| FEA-SLT: A Gloss-Free End-to-End Framework for Facial-Expression-Aware Sign Language Translation added by Syntology |
2026-01 (from id) |
google-research/bleurt/bleurt/lib/bert_tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| AfriHG: News headline generation for African Languages |
28 Dec 2024 |
dadelani/AfriHG/Multi-Lang-Rouge/tokenizers.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| AnyText2: Visual Text Generation and Editing With Customizable Attributes |
22 Nov 2024 |
tyxsspa/anytext2/bert_tokenizer.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| AnyText: Multilingual Visual Text Generation And Editing |
6 Nov 2023 |
tyxsspa/anytext/bert_tokenizer.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| TLM: Token-Level Masking for Transformers |
28 Oct 2023 |
blcuicall/CCL2022-CLTC/baselines/track2/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Sudden Drops in the Loss: Syntax Acquisition, Phase Transitions, and Simplicity Bias in MLMs |
13 Sep 2023 |
angie-chen55/sudden-drops-in-the-loss/extract_attention_pt.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| ReCOGS: How Incidental Details of a Logical Form Overshadow an Evaluation of Semantic Interpretation |
24 Mar 2023 |
frankaging/recogs/utils/cogs_utils.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| NapSS: Paragraph-level Medical Text Simplification via Narrative Prompting and Sentence-matching Summarization |
11 Feb 2023 |
google-research/bert/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese |
2 Nov 2022 |
ofa-sys/chinese-clip/cn_clip/clip/bert_tokenizer.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| EVA2.0: Investigating Open-Domain Chinese Dialogue Systems with Large-Scale Pre-Training |
17 Mar 2022 |
thu-coai/EVA/src/tokenization_eva.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Graph Based Network with Contextualized Representations of Turns in Dialogue |
9 Sep 2021 |
blacknoodle/tucore-gcn/models/BERT/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| MathBERT: A Pre-trained Language Model for General NLP Tasks in Mathematics Education |
2 Jun 2021 |
tbs17/MathBERT/mathbert/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT recorded; this copy not marked cleared · pointer only |
| TUTA: Tree-based Transformers for Generally Structured Table Pre-training |
21 Oct 2020 |
microsoft/TUTA_table_understanding/tuta/tokenizer.py 01922c511206a241 |
unverified |
MIT (permissive) |
| An Empirical Study on Large-Scale Multi-Label Text Classification Including Few and Zero-Shot Labels |
4 Oct 2020 |
iliaschalkidis/lmtc-eurlex57k/neural_networks/layers/bert_tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval |
1 Jul 2020 |
microsoft/ANCE/model/SEED_Encoder/tokenization_seed_encoder.py e8a98cc72991f258 |
unverified |
MIT (permissive) |
| DocVQA: A Dataset for VQA on Document Images |
1 Jul 2020 |
anisha2102/docvqa/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Exploring Cross-sentence Contexts for Named Entity Recognition with BERT |
2 Jun 2020 |
jouniluoma/bert-ner-cmv/bert_tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Common Sense or World Knowledge? Investigating Adapter-Based Knowledge Injection into Pretrained Transformers |
24 May 2020 |
wluper/retrograph/retrograph/modeling/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Named Entity Recognition as Dependency Parsing |
14 May 2020 |
juntaoy/biaffine-ner/extract_bert_features/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Probabilistically Masked Language Model Capable of Autoregressive Generation in Arbitrary Word Order |
24 Apr 2020 |
huawei-noah/Pretrained-Language-Model/PMLM/interactive_conditional_samples_sincos_acrostic.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Invariant Rationalization |
22 Mar 2020 |
code-terminator/invariant_rationalization/imdb.py 5049398bdcfb5fa7 |
unverified |
MIT (permissive) |
| Do Multi-hop Readers Dream of Reasoning Chains? |
31 Oct 2019 |
helloeve/bert-co-matching/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Structured Pruning of Large Language Models |
10 Oct 2019 |
Holldean/BERT-Pruning/bert/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| ALBERT: A Lite BERT for Self-supervised Learning of Language Representations |
26 Sep 2019 |
kpe/bert-for-tf2/bert/tokenization/bert_tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Well-Read Students Learn Better: On the Importance of Pre-training Compact Models |
23 Aug 2019 |
Arthurizijar/Bert_Airport/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Real-Time Open-Domain Question Answering with Dense-Sparse Phrase Index |
13 Jun 2019 |
uwnlp/denspi/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Deeper Text Understanding for IR with Contextual Neural Language Modeling |
22 May 2019 |
AdeDZY/SIGIR19-BERT-IR/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
BSD-3-Clause (permissive) |
| ERNIE: Enhanced Representation through Knowledge Integration |
19 Apr 2019 |
lyqcom/emotect/src/reader.py 01f2dd1d4afbd7c9 |
unverified |
Apache-2.0 (permissive) |
| Utilizing BERT for Aspect-Based Sentiment Analysis via Constructing Auxiliary Sentence |
22 Mar 2019 |
HSLCY/ABSA-BERT-pair/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| BERT for Joint Intent Classification and Slot Filling |
28 Feb 2019 |
mangushev/intent_slot/assistant/functions/intent_slot/preprocessor/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Extracting Multiple-Relations in One-Pass with Pre-Trained Transformers |
4 Feb 2019 |
helloeve/mre-in-one-pass/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| BioBERT: a pre-trained biomedical language representation model for biomedical text mining |
25 Jan 2019 |
dmis-lab/bern/biobert_ner/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
BSD-2-Clause (permissive) |
| Passage Re-ranking with BERT |
13 Jan 2019 |
nyu-dl/dl4marco-bert/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
BSD-3-Clause (permissive) |
| BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding |
11 Oct 2018 |
1wy/bert/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding |
11 Oct 2018 |
IBM/MAX-Question-Answering/core/tokenization.py ef47b227c080ca7a |
unverified |
Apache-2.0 (permissive) |
| BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding |
11 Oct 2018 |
MaZhiyuanBUAA/bert-tf1.4.0/tokenization.py 4874ae2953f3aeab |
unverified |
Apache-2.0 (permissive) |
| BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding |
11 Oct 2018 |
guoyaohua/BERT-Chinese-Annotation/tokenization.py 0953802ea4dc83a7 |
unverified |
Apache-2.0 (permissive) |
| BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding |
11 Oct 2018 |
meizi1114/bert/tokenization.py e2bfbb357751d986 |
unverified |
Apache-2.0 (permissive) |
| BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding |
11 Oct 2018 |
tyxr/bert/tokenization.py 595d45bc07678f87 |
unverified |
Apache-2.0 (permissive) |
| Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation |
26 Sep 2016 |
microsoft/BlingFire/ldbsrc/bert_base_tok/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| arXiv:aaai_17659 |
|
frankaging/Quasi-Attention-ABSA/code/util/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| arXiv:2025.findings-emnlp.1385 |
|
bowen-upenn/ControlText/bert_tokenizer.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| arXiv:2024.findings-acl.532 |
|
kamigaito/SLAHAN/bert/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| arXiv:2024.findings-acl.148 |
|
HillZhang1999/MuCGEC/models/seq2edit-based-CGEC/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| arXiv:2022.acl-long.56 |
|
juntaoy/dali-bridging/extract_bert_features/tokenization.py 1923fc05163d207d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |