| Read-ME: Refactorizing LLMs as Router-Decoupled Mixture of Experts with System Co-Design |
24 Oct 2024 |
VITA-Group/READ-ME/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
no licence file found · pointer only |
| Model Balancing Helps Low-data Training and Fine-tuning |
16 Oct 2024 |
zihanghliu/modelbalancing/transformers/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
no licence file found · pointer only |
| Label Confidence Weighted Learning for Target-level Sentence Simplification |
8 Oct 2024 |
astro-jon/LCWL/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Looped Transformers for Length Generalization |
24 Sep 2024 |
UW-Madison-Lee-Lab/looped-tf/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
no licence file found · pointer only |
| Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps |
9 Jul 2024 |
voidism/lookback-lens/transformers-4.32.0/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
no licence file found · pointer only |
| Increasing Model Capacity for Free: A Simple Strategy for Parameter Efficient Fine-tuning |
1 Jul 2024 |
LINs-lab/CapaBoost/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 (permissive) |
| ESCoT: Towards Interpretable Emotional Support Dialogue Systems |
16 Jun 2024 |
thu-coai/Emotional-Support-Conversation/codes/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| VB-LoRA: Extreme Parameter Efficient Fine-Tuning with Vector Banks |
24 May 2024 |
leo-yangli/VB-LoRA/NLU/NLU/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Learning Syntax Without Planting Trees: Understanding When and Why Transformers Generalize Hierarchically |
25 Apr 2024 |
kabirahuja2431/transformers-hg/transformers/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
no licence file found · pointer only |
| ScanTalk: 3D Talking Heads from Unregistered Scans |
16 Mar 2024 |
miccunifi/ScanTalk/src/hubert/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Improving Open-Ended Text Generation via Adaptive Decoding |
28 Feb 2024 |
zwhong714/adaptive_decoding/transformers-main/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
no licence file found · pointer only |
| Self-Guided Masked Autoencoders for Domain-Agnostic Self-Supervised Learning |
22 Feb 2024 |
johnathan-xie/sma/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 (permissive) |
| Riemannian Preconditioned LoRA for Fine-Tuning Foundation Models |
4 Feb 2024 |
identical code first harvested elsewhere a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| Generative Dense Retrieval: Memory Can Be a Burden |
19 Jan 2024 |
ypw0102/gdr/GDR_model/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Mask Grounding for Referring Image Segmentation |
19 Dec 2023 |
yxchng/mask-grounding/bert/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
AGPL-3.0 (copyleft) · pointer only |
| An Analysis and Mitigation of the Reversal Curse |
13 Nov 2023 |
trestad/mitigating-reversal-curse/transformers/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
no licence file found · pointer only |
| Learning Knowledge-Enhanced Contextual Language Representations for Domain Natural Language Understanding |
12 Nov 2023 |
alibaba/EasyNLP/easynlp/modelzoo/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| NOLA: Compressing LoRA using Linear Combination of Random Basis |
4 Oct 2023 |
UCDvision/NOLA/gpt/examples/NLG/src/model_nola.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Learning to Predict Concept Ordering for Common Sense Generation |
12 Sep 2023 |
tianhuizhang/concept_ordering/bart/src/model/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Contrastive Grouping with Transformer for Referring Image Segmentation |
2 Sep 2023 |
toneyaya/cgformer/bert/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Token-wise Decomposition of Autoregressive Language Model Hidden States for Analyzing Model Predictions |
17 May 2023 |
byungdoh/llm_decomposition/huggingface/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 (permissive) |
| AutoPEFT: Automatic Configuration Search for Parameter-Efficient Fine-Tuning |
28 Jan 2023 |
cambridgeltl/autopeft/adapter-transformers-adapters3.1.0/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Entropy- and Distance-Based Predictors From GPT-2 Attention Patterns Predict Reading Times Over and Above GPT-2 Surprisal |
21 Dec 2022 |
byungdoh/attn_dist/huggingface/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 (permissive) |
| AGRO: Adversarial Discovery of Error-prone groups for Robust Optimization |
2 Dec 2022 |
bhargaviparanjape/robust-transformers/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| EvEntS ReaLM: Event Reasoning of Entity States via Language Models |
10 Nov 2022 |
spilioeve/eventsrealm/transformers-single-all-attribute-prompt-experiments/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
MIT (permissive) |
| Locally Typical Sampling |
1 Feb 2022 |
cimeister/typical-sampling/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Locally Typical Sampling |
1 Feb 2022 |
cimeister/typical-sampling/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Attention Approximates Sparse Distributed Memory |
10 Nov 2021 |
trentbrick/attention-approximates-sdm/HugFace/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| DSEE: Dually Sparsity-embedded Efficient Tuning of Pre-trained Language Models |
30 Oct 2021 |
vita-group/dsee/non-GPT-2/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| DSEE: Dually Sparsity-embedded Efficient Tuning of Pre-trained Language Models |
30 Oct 2021 |
vita-group/dsee/non-GPT-2/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
MIT (permissive) |
| Towards a Unified View of Parameter-Efficient Transfer Learning |
8 Oct 2021 |
jxhe/unify-parameter-efficient-tuning/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Towards a Unified View of Parameter-Efficient Transfer Learning |
8 Oct 2021 |
jxhe/unify-parameter-efficient-tuning/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Smelting Gold and Silver for Improved Multilingual AMR-to-Text Generation |
8 Sep 2021 |
UKPLab/m-AMR2Text/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Smelting Gold and Silver for Improved Multilingual AMR-to-Text Generation |
8 Sep 2021 |
UKPLab/m-AMR2Text/transformers/activations_tf.py ac1bc5cb23b3f2ec |
unverified |
Apache-2.0 (permissive) |
| Learned Token Pruning for Transformers |
2 Jul 2021 |
kssteven418/ltp/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Learned Token Pruning for Transformers |
2 Jul 2021 |
kssteven418/ltp/src/transformers/activations_tf.py ac1bc5cb23b3f2ec |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| X-FACT: A New Benchmark Dataset for Multilingual Fact Checking |
17 Jun 2021 |
utahnlp/x-fact/transformers/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Dialogue-oriented Pre-training |
1 Jun 2021 |
xyease/Dialog-PrLM/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Structure-Aware Abstractive Conversation Summarization via Discourse and Action Graphs |
16 Apr 2021 |
GT-SALT/Structure-Aware-BART/transformers/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Effect of Visual Extensions on Natural Language Understanding in Vision-and-Language Models |
16 Apr 2021 |
alab-nii/eval_vl_glue/eval_vl_glue/transformers_volta/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Transformers for Modeling Physical Systems |
4 Oct 2020 |
zabaras/transformer-physx/trphysx/transformer/utils.py b253e96edee87a52 |
unverified |
MIT (permissive) |
| AdapterHub: A Framework for Adapting Transformers |
15 Jul 2020 |
adapter-hub/adapter-transformers-legacy/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| arXiv:2024.findings-naacl.117 |
|
qqplot/dcpmi/transformers/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
MIT (permissive) |
| arXiv:2022.naacl-main.130 |
|
parovicm/BADX/src/transformers/activations.py a4475703ff58ecf9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| arXiv:2022.naacl-main.130 |
|
parovicm/BADX/src/transformers/activations_tf.py 37a5eed2dbd663ca |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |