| VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation |
18 Nov 2024 |
JT-Sun/Filtering-WoRA/models/xroberta.py 38f46ef4e03fee1c |
ran
|
Apache-2.0 (permissive) |
| Controllable Context Sensitivity and the Knob Behind It |
11 Nov 2024 |
kdu4108/context-vs-prior-finetuning/model_utils/mi_utils.py 640aef5d8b077b96 |
unverified |
no licence file found · pointer only |
| The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models |
9 Oct 2024 |
EIT-NLP/AccuracyParadox-RLHF/reward_modeling/my_longformer.py ce0ed06559b93e81 |
ran
fingerprinted |
MIT (permissive) |
| Propulsion: Steering LLM with Tiny Fine-Tuning |
17 Sep 2024 |
Kowsher/Propulsion/Src/LM/roberta.py 38f46ef4e03fee1c |
ran
|
no licence file found · pointer only |
| CLIBE: Detecting Dynamic Backdoors in Transformer-based NLP Models |
2 Sep 2024 |
raytsang123/clibe/discriminative_backdoors/detection/modeling_roberta.py 38f46ef4e03fee1c |
ran
|
Apache-2.0 (permissive) |
| ParGo: Bridging Vision-Language with Partial and Global Views |
23 Aug 2024 |
bytedance/pargo/pargo/backbone/language/xlm_roberta.py 38f46ef4e03fee1c |
ran
|
BSD-3-Clause (permissive) |
| Zero-Shot Cross-Lingual NER Using Phonemic Representations for Low-Resource Languages |
23 Jun 2024 |
Gabriel819/zeroshot_ner/model/xphonebert.py d4903874fdec3763 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Thinking Forward: Memory-Efficient Federated Finetuning of Language Models |
24 May 2024 |
Astuary/Spry/models/mezo_modeling_roberta.py 38f46ef4e03fee1c |
ran
|
no licence file found · pointer only |
| ALaRM: Align Language Models via Hierarchical Rewards Modeling |
11 Mar 2024 |
halfrot/ALaRM/long-form-QA/my_longformer.py ce0ed06559b93e81 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval |
24 Jan 2024 |
wusiwei0410/scimmir/src/LLM_models/My_kosmos2.py 38f46ef4e03fee1c |
ran
|
no licence file found · pointer only |
| Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives |
30 Nov 2023 |
facebookresearch/EgoVLPv2/EgoVLPv2/model/roberta.py da1c54dffed5d60c |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| MLP Fusion: Towards Efficient Fine-tuning of Dense and Mixture-of-Experts Language Models |
18 Jul 2023 |
weitianxin/MLP_Fusion/models/roberta/modeling_roberta.py 38f46ef4e03fee1c |
ran
|
no licence file found · pointer only |
| MLP Fusion: Towards Efficient Fine-tuning of Dense and Mixture-of-Experts Language Models |
18 Jul 2023 |
weitianxin/MLP_Fusion/models/roberta/modeling_flax_roberta.py 6697ea7953dcb4c6 |
ran
|
no licence file found · pointer only |
| LongCoder: A Long-Range Pre-trained Language Model for Code Completion |
26 Jun 2023 |
microsoft/codebert/LongCoder/longcoder.py da1c54dffed5d60c |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| IndicTrans2: Towards High-Quality and Accessible Machine Translation Models for all 22 Scheduled Indian Languages |
25 May 2023 |
ai4bharat/indictrans2/huggingface_interface/modeling_indictrans.py 336749dd7d699ef6 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| A Causal View of Entity Bias in (Large) Language Models |
24 May 2023 |
luka-group/causal-view-of-entity-bias/roberta.py 2e2109ff964bfb9c |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Text Is All You Need: Learning Language Representations for Sequential Recommendation |
23 May 2023 |
aaronheee/recformer/recformer/models.py 6e4b78022b53cbdb |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Reliable Gradient-free and Likelihood-free Prompt Tuning |
30 Apr 2023 |
maohaos2/SBI_LLM/models/modeling_roberta.py bbfc47d976ee7225 |
unverified |
MIT (permissive) |
| WavCaps: A ChatGPT-Assisted Weakly-Labelled Audio Captioning Dataset for Audio-Language Multimodal Research |
30 Mar 2023 |
gzhu06/cacophony/src/caco/text_models/roberta_text_model.py 12a1bb1ff6ec7d06 |
unverified |
MIT (permissive) |
| DiffusionBERT: Improving Generative Masked Language Models with Diffusion Models |
28 Nov 2022 |
hzfinfdu/diffusion-bert/models/modeling_roberta.py 38f46ef4e03fee1c |
ran
|
Apache-2.0 (permissive) |
| VoLTA: Vision-Language Transformer with Weakly-Supervised Local-Feature Alignment |
9 Oct 2022 |
ShramanPramanick/VoLTA/Pre-training/roberta.py da1c54dffed5d60c |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Causal Proxy Models for Concept-Based Model Explanations |
28 Sep 2022 |
frankaging/causal-proxy-model/models/modelings_roberta.py 38f46ef4e03fee1c |
ran
|
MIT (permissive) |
| TIE: Topological Information Enhanced Structural Reading Comprehension on Web Pages |
13 May 2022 |
x-lance/tie/markuplmft/models/markuplm/modeling_markuplm.py 38f46ef4e03fee1c |
ran
|
MIT (permissive) |
| DiffCSE: Difference-based Contrastive Learning for Sentence Embeddings |
21 Apr 2022 |
voidism/diffcse/modeling_roberta.py d4903874fdec3763 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| InCoder: A Generative Model for Code Infilling and Synthesis |
12 Apr 2022 |
eth-sri/sven/sven/model.py 336749dd7d699ef6 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| LiLT: A Simple yet Effective Language-Independent Layout Transformer for Structured Document Understanding |
28 Feb 2022 |
jpWang/LiLT/LiLTfinetune/models/LiLTRobertaLike/modeling_LiLTRobertaLike.py da1c54dffed5d60c |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Should You Mask 15% in Masked Language Modeling? |
16 Feb 2022 |
princeton-nlp/dinkytrain/huggingface/modeling_roberta_prelayernorm.py 38f46ef4e03fee1c |
ran
|
MIT recorded; this copy not marked cleared · pointer only |
| Revisiting Parameter-Efficient Tuning: Are We Really There Yet? |
16 Feb 2022 |
guanzhchen/petuning/model/roberta/modeling_roberta.py 38f46ef4e03fee1c |
ran
|
MIT (permissive) |
| Revisiting Parameter-Efficient Tuning: Are We Really There Yet? |
16 Feb 2022 |
guanzhchen/petuning/model/roberta/modeling_flax_roberta.py 6697ea7953dcb4c6 |
ran
|
MIT (permissive) |
| IGLUE: A Benchmark for Transfer Learning across Modalities, Tasks, and Languages |
27 Jan 2022 |
e-bug/volta/volta/embeddings.py d4903874fdec3763 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Black-Box Tuning for Language-Model-as-a-Service |
10 Jan 2022 |
jubgjf/plmtuningcompetition/models/deep_modeling_roberta.py bbfc47d976ee7225 |
unverified |
Apache-2.0 (permissive) |
| ProtoTransformer: A Meta-Learning Approach to Providing Student Feedback |
23 Jul 2021 |
mhw32/prototransformer-public/src/models/monkeypatch.py 3976dab6c53b60b0 |
unverified |
MIT (permissive) |
| Learned Token Pruning for Transformers |
2 Jul 2021 |
kssteven418/ltp/src/transformers/models/ltp/modeling_ltp.py 83995d570ba89e6d |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| What to Pre-Train on? Efficient Intermediate Task Selection |
16 Apr 2021 |
adapter-hub/efficient-task-transfer/task_selection/modeling/modeling_roberta.py bbfc47d976ee7225 |
unverified |
MIT (permissive) |
| A Simple Language Model for Task-Oriented Dialogue |
2 May 2020 |
salesforce/simpletod/models/modeling_utils.py d4f3546f2ecff417 |
unverified |
BSD-3-Clause (permissive) |
| arXiv:aaai_33874 |
|
txsun1997/Black-Box-Tuning/models/deep_modeling_roberta.py bbfc47d976ee7225 |
unverified |
MIT (permissive) |
| arXiv:aaai_29850 |
|
ozyyshr/FocalReasoner/modeling_roberta_svo_graph.py da1c54dffed5d60c |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| arXiv:2024.findings-emnlp.286 |
|
kgarg8/Stanceformer/roberta/modeling_roberta.py 38f46ef4e03fee1c |
ran
|
MIT (permissive) |
| arXiv:2024.findings-acl.164 |
|
potter-Zhang/Selective-Prefix-Tuning/model/roberta_select.py 38f46ef4e03fee1c |
ran
|
Apache-2.0 (permissive) |
| arXiv:2023.findings-emnlp.107 |
|
chenxn2020/GOSE/GOSEfinetune/models/LiLTRobertaLike/modeling_LiLTRobertaLike.py da1c54dffed5d60c |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| arXiv:2023.acl-long.264 |
|
DAMO-NLP-SG/MVCR/src/modeling_vaeroberta.py 5ba2b30d6672d2b0 |
unverified |
MIT recorded; this copy not marked cleared · pointer only |
| arXiv:2023.acl-long.182 |
|
Yuanhy1997/HyPe/hype_modeling_roberta.py 38f46ef4e03fee1c |
ran
|
MIT (permissive) |
| arXiv:2021.emnlp-main.154 |
|
Hazelsuko07/TextHide/transformers_hide/modeling_roberta.py 5b2390deea2200c9 |
unverified |
MIT (permissive) |