| UHR-BAT: Budget-Aware Token Compression Vision-Language model for Ultra-High-Resolution Remote Sensing added by Syntology |
2026-04 (from id) |
Yunkaidang/UHR/with-SAM/longva/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| VidLaDA: Bidirectional Diffusion Large Language Models for Efficient Video Understanding added by Syntology |
2026-01 (from id) |
ziHoHe/VidLaDA/train/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| Predict the Retrieval! Test time adaptation for Retrieval Augmented Generation added by Syntology |
2026-01 (from id) |
sunxin000/TTARAG/models/rag_llama_baseline.py bb3da7fc875ebf5d |
unverified |
no licence file found · pointer only |
| Cross-Layer Injection for Deep Vision-Language Fusion added by Syntology |
2026-01 (from id) |
codefuse-ai/CLI/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| InertialAR: Autoregressive 3D Molecule Generation with Inertial Frames added by Syntology |
2025-10 (from id) |
HaoruiLi46/InertialAR/InertialAR/generate_b3lyp.py 71a5b039296e865e |
unverified |
no licence file found · pointer only |
| Phi: Preference Hijacking in Multi-modal Large Language Models at Inference Time added by Syntology |
2025-09 (from id) |
Yifan-Lan/Phi/trl/core.py 072cc4ae40638ba1 |
unverified |
MIT (permissive) |
| Continuously Steering LLMs Sensitivity to Contextual Knowledge with Proxy Models added by Syntology |
2025-08 (from id) |
OliveJuiceLin/CSKS/proxy_model/dexpert.py f24ceca3d1604ba8 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism |
4 Jun 2025 |
weizhepei/adadecode/self_speculation/self_speculation_generator.py b0aaaed3efc56072 |
ran
|
no licence file found · pointer only |
| Learning from Videos for 3D World: Enhancing MLLMs with 3D Vision Geometry Priors |
30 May 2025 |
LaVi-Lab/Video-3D-LLM/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| D-AR: Diffusion via Autoregressive Models |
29 May 2025 |
showlab/d-ar/autoregressive/models/generate.py f2fea48028c7ae80 |
unverified |
MIT (permissive) |
| LaViDa: A Large Diffusion Language Model for Multimodal Understanding |
22 May 2025 |
jacklishufan/lavida/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| Training-Free Watermarking for Autoregressive Image Generation |
20 May 2025 |
maifoundations/indexmark/autoregressive/models/generate.py f2fea48028c7ae80 |
unverified |
MIT (permissive) |
| Projecting Assumptions: The Duality Between Sparse Autoencoders and Concept Geometry |
3 Mar 2025 |
EkdeepSLubana/spadeFormalGrammars/model/model.py 4088666d0047518a |
unverified |
MIT (permissive) |
| DriveMM: All-in-One Large Multimodal Model for Autonomous Driving |
10 Dec 2024 |
zhijian11/DriveMM/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| Moto: Latent Motion Token as the Bridging Language for Learning Robot Manipulation from Videos |
5 Dec 2024 |
tencentarc/moto/moto_gpt/src/models/moto_gpt.py c05a2ee543965828 |
unverified |
licence not identified · pointer only |
| ZipAR: Accelerating Auto-regressive Image Generation through Spatial Locality |
5 Dec 2024 |
ThisisBillhe/ZipAR/LlamaGen-ZipAR/autoregressive/models/generate.py f2fea48028c7ae80 |
unverified |
no licence file found · pointer only |
| AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning |
4 Dec 2024 |
lavi-lab/aim/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| PanoLlama: Generating Endless and Coherent Panoramas with Next-Token-Prediction LLMs |
24 Nov 2024 |
0606zt/panollama/token_generator/generate.py f2fea48028c7ae80 |
unverified |
no licence file found · pointer only |
| DyCoke: Dynamic Compression of Tokens for Fast Video Large Language Models |
22 Nov 2024 |
kd-tao/dycoke/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation |
18 Nov 2024 |
JT-Sun/Filtering-WoRA/models/bert.py 9e8f81565e828ebf |
ran
|
Apache-2.0 (permissive) |
| SymDPO: Boosting In-Context Learning of Large Multimodal Models with Symbol Demonstration Direct Preference Optimization |
17 Nov 2024 |
APiaoG/SymDPO/trl/core.py 2c3da386578af970 |
unverified |
no licence file found · pointer only |
| CCExpert: Advancing MLLM Capability in Remote Sensing Change Captioning with Difference-Aware Integration and a Foundational Dataset |
18 Nov 2024 |
meize0729/ccexpert/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation |
15 Oct 2024 |
zjunlp/Deco/transformers/generation/utils.py 39da6a0fe0e67626 |
unverified |
MIT (permissive) |
| Analyzing (In)Abilities of SAEs via Formal Languages |
15 Oct 2024 |
Abhinav271828/pcfg-sae-causal-arr-oct24/model/model.py 4088666d0047518a |
unverified |
no licence file found · pointer only |
| Coevolving with the Other You: Fine-Tuning LLM with Sequential Cooperative Multi-Agent Reinforcement Learning |
8 Oct 2024 |
Harry67Hu/CORY/trl/core.py 5bcf8e086423cf6b |
ran
|
MIT (permissive) |
| ControlAR: Controllable Image Generation with Autoregressive Models |
3 Oct 2024 |
hustvl/ControlAR/autoregressive/models/generate.py f2fea48028c7ae80 |
unverified |
Apache-2.0 (permissive) |
| Style-Specific Neurons for Steering LLMs in Text Style Transfer |
1 Oct 2024 |
wenlai-lavine/sNeuron-TST/Our/utils_new.py eec474cc8cda034e |
unverified |
no licence file found · pointer only |
| Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval |
30 Sep 2024 |
lijiabei-7/leccr/LECCR/models/xbert.py 9e8f81565e828ebf |
ran
|
no licence file found · pointer only |
| SSR-Speech: Towards Stable, Safe and Robust Zero-shot Text-based Speech Editing and Synthesis |
11 Sep 2024 |
WangHelin1997/SSR-Speech/models/ssr.py 99a311d99448f881 |
ran
|
MIT (permissive) |
| IndicVoices-R: Unlocking a Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS |
9 Sep 2024 |
ai4bharat/indicvoices-r/VoiceCraft/models/voicecraft.py 99a311d99448f881 |
ran
|
CC-BY-4.0 · pointer only |
| A Percolation Model of Emergence: Analyzing Transformers Trained on a Formal Language |
22 Aug 2024 |
ekdeepslubana/conceptpercolation/model/model.py 4088666d0047518a |
unverified |
no licence file found · pointer only |
| Suri: Multi-constraint Instruction Following for Long-form Text Generation |
27 Jun 2024 |
chtmp223/suri/ft/lib/trl_mod/core.py 5bcf8e086423cf6b |
ran
|
no licence file found · pointer only |
| Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation |
10 Jun 2024 |
foundationvision/llamagen/autoregressive/models/generate.py f2fea48028c7ae80 |
unverified |
MIT (permissive) |
| PrE-Text: Training Language Models on Private Federated Data in the Age of LLMs |
5 Jun 2024 |
houcharlie/pre-text/variation.py 13409c48abcde4ce |
unverified |
no licence file found · pointer only |
| Keep It Private: Unsupervised Privatization of Online Text |
16 May 2024 |
csbao/kip-privatization/src/utils_sampling.py acaf70fe7b4941ca |
ran
|
no licence file found · pointer only |
| On the Content Bias in Fréchet Video Distance |
18 Apr 2024 |
songweige/tats/tats/modules/gpt.py 49f3df5f44c69b6f |
ran
|
MIT (permissive) |
| VoiceCraft: Zero-Shot Speech Editing and Text-to-Speech in the Wild |
25 Mar 2024 |
jasonppy/voicecraft/models/voicecraft.py 99a311d99448f881 |
ran
|
no licence file found · pointer only |
| How do Large Language Models Handle Multilingualism? |
29 Feb 2024 |
damo-nlp-sg/multilingual_analysis/layers/transformers/generation/utils.py eec474cc8cda034e |
unverified |
no licence file found · pointer only |
| Query Augmentation by Decoding Semantics from Brain Signals |
24 Feb 2024 |
yeziyi1998/brain-query-augmentation/help/utils.py eec474cc8cda034e |
unverified |
MIT (permissive) |
| GeReA: Question-Aware Prompt Captions for Knowledge-based Visual Question Answering |
4 Feb 2024 |
upper9527/gerea/src/generation_utils.py 0dc219ab6716a127 |
ran
|
no licence file found · pointer only |
| OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation |
29 Nov 2023 |
shikiw/opera/transformers-4.29.2/src/transformers/generation/utils.py eec474cc8cda034e |
unverified |
MIT (permissive) |
| Zero-shot audio captioning with audio-language model guidance and audio context keywords |
14 Nov 2023 |
explainableml/zeraucap/audio_captioning/language_model/utlis.py 96370a572229a5a9 |
ran
|
no licence file found · pointer only |
| Neuroformer: Multimodal and Multitask Generative Pretraining for Brain Data |
31 Oct 2023 |
a-antoniades/neuroformer/neuroformer/simulation.py 120a5300accb00db |
unverified |
MIT (permissive) |
| ODEFormer: Symbolic Regression of Dynamical Systems with Transformers |
9 Oct 2023 |
sdascoli/odeformer/odeformer/model/transformer.py 9ef29a7930f6956c |
ran
|
MIT (permissive) |
| Contrastive Grouping with Transformer for Referring Image Segmentation |
2 Sep 2023 |
toneyaya/cgformer/bert/generation_utils.py 0dc219ab6716a127 |
ran
|
MIT (permissive) |
| SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models |
31 Aug 2023 |
0nutation/uslm/models/valle.py 99a311d99448f881 |
ran
|
no licence file found · pointer only |
| Large Multilingual Models Pivot Zero-Shot Multimodal Learning across Languages |
23 Aug 2023 |
openbmb/viscpm/VisCPM/generation/generation_utils.py d9a981c1eb0f46fc |
ran
|
no licence file found · pointer only |
| PoET: A generative model of protein families as sequences-of-sequences |
9 Jun 2023 |
OpenProteinAI/PoET/poet/models/poet.py 733f20ae4d66d5ae |
ran · fixture could not drive it
|
MIT (permissive) |
| Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks |
7 Jun 2023 |
x-plug/youku-mplug/models/predictor_mplug.py 63e2cc5fb0616ee9 |
unverified |
Apache-2.0 (permissive) |
| Towards Unified Text-based Person Retrieval: A Large-scale Multi-Attribute and Language Search Benchmark |
5 Jun 2023 |
Shuyu-XJTU/APTM/models/bert.py 9e8f81565e828ebf |
ran
|
MIT (permissive) |
| Direct Preference Optimization: Your Language Model is Secretly a Reward Model |
29 May 2023 |
padlex/trl/trl/core.py 5bcf8e086423cf6b |
ran
|
Apache-2.0 (permissive) |
| Pointwise Mutual Information Based Metric and Decoding Strategy for Faithful Generation in Document Grounded Dialogs |
20 May 2023 |
ynandwan/pmi-faith/faithful-decode/models/generation_utils_pmi.py eec474cc8cda034e |
unverified |
Apache-2.0 (permissive) |
| Outline, Then Details: Syntactically Guided Coarse-To-Fine Code Generation |
28 Apr 2023 |
vita-group/chaincoder/model/transformer.py 9ef29a7930f6956c |
ran
|
MIT (permissive) |
| Transformer-based Planning for Symbolic Regression |
13 Mar 2023 |
deep-symbolic-mathematics/TPSR/symbolicregression/model/transformer.py 02e71d97ed55e54d |
ran · fixture could not drive it
|
MIT (permissive) |
| Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling |
7 Mar 2023 |
plachtaa/vall-e-x/models/vallex.py 99a311d99448f881 |
ran
|
MIT (permissive) |
| mPLUG-2: A Modularized Multi-modal Foundation Model Across Text, Image and Video |
1 Feb 2023 |
X-PLUG/mPLUG-2/models/predictor.py 63e2cc5fb0616ee9 |
unverified |
Apache-2.0 (permissive) |
| Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers |
5 Jan 2023 |
lifeiteng/vall-e/valle/models/valle.py 99a311d99448f881 |
ran
|
Apache-2.0 (permissive) |
| EfficientVLM: Fast and Accurate Vision-Language Models via Knowledge Distillation and Modal-adaptive Pruning |
14 Oct 2022 |
swaggy-tn/efficientvlm/efficient_models/eff_bert.py 9e8f81565e828ebf |
ran
|
BSD-3-Clause recorded; this copy not marked cleared · pointer only |
| Predictive Querying for Autoregressive Neural Sequence Models |
12 Oct 2022 |
ajboyd2/prob_seq_queries/seq_queries/sample.py d111355f8bf308f3 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Fengshenbang 1.0: Being the Foundation of Chinese Cognitive Intelligence |
7 Sep 2022 |
idea-ccnl/fengshenbang-lm/fengshen/models/DAVAE/run_latent_generation.py c6f7b6ef9dd8586b |
unverified |
Apache-2.0 (permissive) |
| Can Language Models Make Fun? A Case Study in Chinese Comical Crosstalk |
2 Jul 2022 |
anonNo2/crosstalk-generation/src/cpm/generate_eval_data.py 6f5ddea9dd9cffbb |
unverified |
Apache-2.0 (permissive) |
| Write and Paint: Generative Vision-Language Models are Unified Modal Learners |
15 Jun 2022 |
shizhediao/davinci/models/davinci_pretrain.py 8164123ea94d7bb5 |
unverified |
BSD-3-Clause recorded; this copy not marked cleared · pointer only |
| LAVENDER: Unifying Video-Language Understanding as Masked Language Modeling |
14 Jun 2022 |
microsoft/lavender/model_for_captioning.py 2eab23b94336a71e |
unverified |
MIT (permissive) |
| End-to-end symbolic regression with transformers |
22 Apr 2022 |
deep-symbolic-mathematics/TPSR/symbolicregression/model/transformer.py f434565e2011ab22 |
ran · fixture could not drive it
|
MIT (permissive) |
| Perspective-taking and Pragmatics for Generating Empathetic Responses Focused on Emotion Causes |
18 Sep 2021 |
skywalker023/focused-empathy/from_epitome/modeling_utils.py 8e873fc150477bd1 |
unverified |
MIT (permissive) |
| A Three-Stage Learning Framework for Low-Resource Knowledge-Grounded Dialogue Generation |
9 Sep 2021 |
neukg/kat-tslf/kat/gen_utils.py 685c7abbaa62f39d |
unverified |
MIT (permissive) |
| Image Retrieval on Real-life Images with Pre-trained Vision-and-Language Models |
9 Aug 2021 |
Cuberick-Orion/CIRPLANT/model/OSCAR/modeling/modeling_utils.py 8164123ea94d7bb5 |
unverified |
MIT (permissive) |
| Controlled Text Generation as Continuous Optimization with Multiple Constraints |
4 Aug 2021 |
sachin19/mucoco/mucoco/utils/targets.py cb4a7e224fe59578 |
unverified |
MIT (permissive) |
| SymbolicGPT: A Generative Transformer Model for Symbolic Regression |
27 Jun 2021 |
mojivalipour/symbolicgpt/utils.py 5427a29e4bb4bf46 |
unverified |
MIT (permissive) |
| Towards Understanding and Mitigating Social Biases in Language Models |
24 Jun 2021 |
pliang279/LM_bias/src/global_bias/generate_full_sentence.py cf42150195c933b4 |
unverified |
MIT (permissive) |
| Towards Long-Form Video Understanding |
21 Jun 2021 |
chaoyuaw/lvu/src/models/modeling_utils.py 8e873fc150477bd1 |
unverified |
MIT (permissive) |
| Straight to the Gradient: Learning to Use Novel Tokens for Neural Text Generation |
14 Jun 2021 |
shawnlimn/scalegrad/custom/gpt2/run_gpt2.py d415ae4787477f20 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Enriching Transformers with Structured Tensor-Product Representations for Abstractive Summarization |
2 Jun 2021 |
jiangycTarheel/TPT-Summ/models/modeling_utils.py 8e873fc150477bd1 |
unverified |
MIT (permissive) |
| DExperts: Decoding-Time Controlled Text Generation with Experts and Anti-Experts |
7 May 2021 |
alisawuffles/DExperts/generation/dexperts_generation.py 8c347cb668f95910 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Effect of Visual Extensions on Natural Language Understanding in Vision-and-Language Models |
16 Apr 2021 |
alab-nii/eval_vl_glue/eval_vl_glue/transformers_volta/generation_utils.py 9c750e2920328942 |
unverified |
Apache-2.0 (permissive) |
| Plug-and-Blend: A Framework for Controllable Story Generation with Blended Control Codes |
23 Mar 2021 |
xxbidiao/plug-and-blend/gedi_helpers/modeling_utils.py 49f3df5f44c69b6f |
ran
|
MIT (permissive) |
| Large Pre-trained Language Models Contain Human-like Biases of What is Right and Wrong to Do |
8 Mar 2021 |
ml-research/MoRT_NMI/MoRT/mcm_textgeneration/mcm_models.py 0b42b3fc90731551 |
unverified |
MIT (permissive) |
| Transformer-based Conditional Variational Autoencoder for Controllable Story Generation |
4 Jan 2021 |
fangleai/TransformerCVAE/generate.py a8f76faa252a6539 |
unverified |
no licence file found · pointer only |
| Directed Beam Search: Plug-and-Play Lexically Constrained Language Generation |
31 Dec 2020 |
dapascual/DirectedBeamSearch/main_DBS.py a2d595fc39ea3025 |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| Stylized Dialogue Response Generation Using Stylized Unpaired Texts |
27 Sep 2020 |
silverriver/Stylized_Dialog/TCFC/bt_beam/model/filtering.py dbc1974da366ebe7 |
unverified |
MIT (permissive) |
| GeDi: Generative Discriminator Guided Sequence Generation |
14 Sep 2020 |
salesforce/GeDi/modeling_utils.py 49f3df5f44c69b6f |
ran
|
BSD-3-Clause (permissive) |
| A Simple Language Model for Task-Oriented Dialogue |
2 May 2020 |
salesforce/simpletod/models/modeling_utils.py 8e873fc150477bd1 |
unverified |
BSD-3-Clause (permissive) |
| POINTER: Constrained Progressive Text Generation via Insertion-based Generative Pre-training |
1 May 2020 |
dreasysnail/POINTER/inference.py ee8b235cc453809b |
unverified |
MIT (permissive) |
| Generative Data Augmentation for Commonsense Reasoning |
24 Apr 2020 |
yangyiben/G-DAUG-c-Generative-Data-Augmentation-for-Commonsense-Reasoning/modeling.py 541fe035a87b91ca |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension |
29 Oct 2019 |
i2r-simmc/i2r-simmc-2020/src/generation_utils.py 0dc219ab6716a127 |
ran
|
MIT (permissive) |
| Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer |
23 Oct 2019 |
abelriboulot/onnxt5/onnxt5/models.py bdeda8fda01d8511 |
unverified |
Apache-2.0 (permissive) |
| ALBERT: A Lite BERT for Self-supervised Learning of Language Representations |
26 Sep 2019 |
Soikonomou/albert_final/src/model/ALBERT/modeling_utils.py 8e873fc150477bd1 |
unverified |
Apache-2.0 (permissive) |
| Abductive Commonsense Reasoning |
15 Aug 2019 |
allenai/abductive-commonsense-reasoning/anlg/run_generation.py ee8b235cc453809b |
unverified |
Apache-2.0 (permissive) |
| TransferTransfo: A Transfer Learning Approach for Neural Network Based Conversational Agents |
23 Jan 2019 |
cerebroai/AskIt/decoder.py 31631ed21a9dd3b3 |
unverified |
MIT (permissive) |
| arXiv:openreview_utRSxIkoSJ |
|
hustyyq/ConceptTok/autoregressive/models/generate.py f2fea48028c7ae80 |
unverified |
Apache-2.0 (permissive) |
| arXiv:Zhong_AIM_Adaptive_Inference_of_Multi-Modal_LLMs_via_Token_Merging_and_ICCV_2025_paper |
|
LaVi-Lab/AIM/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| arXiv:Yang_Beyond_Walking_A_Large-Scale_Image-Text_Benchmark_for_Text-based_Person_Anomaly_ICCV_2025_paper |
|
Shuyu-XJTU/CMP/models/bert.py 9e8f81565e828ebf |
ran
|
MIT (permissive) |
| arXiv:2025.findings-emnlp.70 |
|
passing2961/EmpGPT-3/from_epitome/modeling_utils.py 8e873fc150477bd1 |
unverified |
MIT (permissive) |
| arXiv:2025.findings-acl.359 |
|
G-JWLee/TAMP/trl/core.py 2c3da386578af970 |
unverified |
Apache-2.0 (permissive) |
| arXiv:2025.emnlp-main.1418 |
|
avipartho/Synth-SBDH/dss/gen_utils.py 9c750e2920328942 |
unverified |
Apache-2.0 (permissive) |
| arXiv:2022.emnlp-demos.40 |
|
OpenBMB/ModelCenter/model_center/generation/generation_utils.py d9a981c1eb0f46fc |
ran
|
Apache-2.0 (permissive) |