| Ghosted Layers: Unconstrained Activation Alignment for Recovering Layer-Pruned LLMs added by Syntology |
2026-05 (from id) |
sramshetty/ShortGPT/short_gpt/short_llama.py ab63b7d19f77cc5c |
ran
|
MIT (permissive) |
| EnergyLens: Predictive Energy-Aware Exploration for Multi-GPU LLM Inference Optimization added by Syntology |
2026-05 (from id) |
meta-llama/llama3/llama/generation.py 83cc3c258873fe88 |
ran
|
licence not identified · pointer only |
| Back to Basics: Let Conversational Agents Remember with Just Retrieval and Generation added by Syntology |
2026-04 (from id) |
qingyue2014/Rsum/llama/generation.py e28848558038eee4 |
unverified |
no licence file found · pointer only |
| Latent Speech-Text Transformer added by Syntology |
2025-10 (from id) |
facebookresearch/lst/lst/generate.py 2f2141ad0c4f3dbd |
unverified |
licence not identified · pointer only |
| Byte Latent Transformer: Patches Scale Better Than Tokens |
13 Dec 2024 |
facebookresearch/blt/bytelatent/generate.py 2f2141ad0c4f3dbd |
unverified |
no licence file found · pointer only |
| MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models |
2024-12 (from id) |
shansongliu/MuMu-LLaMA/MuMu-LLaMA/llama/utils.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Large-scale moral machine experiment on large language models |
11 Nov 2024 |
kztakemoto/mmllm/llama/generation.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Can Language Models Perform Robust Reasoning in Chain-of-thought Prompting with Noisy Rationales? |
31 Oct 2024 |
tmlr-group/NoisyRationales/llm_model/llama/generation.py e28848558038eee4 |
unverified |
no licence file found · pointer only |
| Structure Language Models for Protein Conformation Generation |
24 Oct 2024 |
lujiarui/esmdiff/slm/sample_hf.py e28848558038eee4 |
unverified |
no licence file found · pointer only |
| Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities |
15 Oct 2024 |
gpt-omni/mini-omni2/litgpt/generate/base.py a2d9cc3c4221f547 |
ran
fingerprinted |
MIT (permissive) |
| On the Influence of Gender and Race in Romantic Relationship Prediction from Large Language Models |
5 Oct 2024 |
facebookresearch/llama/llama/generation.py e28848558038eee4 |
unverified |
no licence file found · pointer only |
| Counterfactual Token Generation in Large Language Models |
25 Sep 2024 |
networks-learning/counterfactual-llms/src/llama3/llama/generation.py 841e0eb44df294a1 |
ran
|
no licence file found · pointer only |
| Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming |
29 Aug 2024 |
gpt-omni/mini-omni/litgpt/generate/base.py a2d9cc3c4221f547 |
ran
fingerprinted |
MIT (permissive) |
| DiReCT: Diagnostic Reasoning for Clinical Notes via Large Language Models |
4 Aug 2024 |
wbw520/DiReCT/llama/generation.py 83cc3c258873fe88 |
ran
|
no licence file found · pointer only |
| An Investigation of Neuron Activation as a Unified Lens to Explain Chain-of-Thought Eliciting Arithmetic Reasoning of LLMs |
18 Jun 2024 |
dakingrai/neuron-analysis-cot-arithmetic-reasoning/llama/generation.py e28848558038eee4 |
unverified |
no licence file found · pointer only |
| Superposed Decoding: Multiple Generations from a Single Autoregressive Inference Pass |
28 May 2024 |
RAIVNLab/SuperposedDecoding/superposed/llama/generation.py 128a40bb8f1f26ef |
ran
|
licence not identified · pointer only |
| Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length |
12 Apr 2024 |
xuezhemax/megalodon/megalodon/generation.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| OpenBias: Open-set Bias Detection in Text-to-Image Generative Models |
11 Apr 2024 |
picsart-ai-research/openbias/llama/generation.py e28848558038eee4 |
unverified |
no licence file found · pointer only |
| PREGO: online mistake detection in PRocedural EGOcentric videos |
2 Apr 2024 |
aleflabo/PREGO/step_anticipation/llama/generation.py e28848558038eee4 |
unverified |
MIT (permissive) |
| TOD3Cap: Towards 3D Dense Captioning in Outdoor Scenes |
28 Mar 2024 |
jxbbb/tod3cap/tod3cap_camera/llama/utils.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Fast Adversarial Attacks on Language Models In One GPU Minute |
23 Feb 2024 |
vinusankars/beast/arutils.py e26c6aee991a471a |
ran
|
no licence file found · pointer only |
| Entropy-Regularized Token-Level Policy Optimization for Language Agent Reinforcement |
9 Feb 2024 |
morning9393/etpo/etpo/models/codellama/generation.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| E-EVAL: A Comprehensive Chinese K-12 Education Evaluation Benchmark for Large Language Models |
29 Jan 2024 |
ai-edu-lab/e-eval/code/evaluator_series/evaluators/llama.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Pre-trained Large Language Models for Financial Sentiment Analysis |
10 Jan 2024 |
luosting/LLaMA-Financial-sentiment-analysis/llm-sentiment-analysis-main/llama_train/generation.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Pre-trained Large Language Models for Financial Sentiment Analysis |
10 Jan 2024 |
luosting/LLaMA-Financial-sentiment-analysis/llm-sentiment-analysis-main/llama/generation.py e28848558038eee4 |
unverified |
no licence file found · pointer only |
| Vamos: Versatile Action Models for Video Understanding |
22 Nov 2023 |
brown-palm/Vamos/finetune/llama/generation.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| DialogueLLM: Context and Emotion Knowledge-Tuned Large Language Models for Emotion Recognition in Conversations |
17 Oct 2023 |
Dreamyao516/DialogueLLM/llama/generation.py e28848558038eee4 |
unverified |
MIT (permissive) |
| GraphLLM: Boosting Graph Reasoning Ability of Large Language Model |
9 Oct 2023 |
mistyreed63849/graph-llm/llama/model_utils.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Misusing Tools in Large Language Models With Visual Adversarial Examples |
4 Oct 2023 |
ZihanWangKi/VLMToolMisuse/llama_adapter/utils.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following |
1 Sep 2023 |
ziyuguo99/point-bind_point-llm/Point-LLM/llama/utils.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Code Llama: Open Foundation Models for Code |
24 Aug 2023 |
facebookresearch/codellama/llama/generation.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Llama 2: Open Foundation and Fine-Tuned Chat Models |
18 Jul 2023 |
IBM/Dromedary/llama_dromedary/llama_dromedary/generation.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
GPL-3.0 (copyleft) · pointer only |
| ToolkenGPT: Augmenting Frozen Language Models with Massive Tools via Tool Embeddings |
19 May 2023 |
lichenghaobuaa/tokenlearning/llama/model.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| C-Eval: A Multi-Level Multi-Discipline Chinese Evaluation Suite for Foundation Models |
15 May 2023 |
hkust-nlp/ceval/code/evaluator_series/evaluators/llama.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| arXiv:2024.findings-emnlp.687 |
|
IAAR-Shanghai/FastMem/src/train_and_inference.py cc1472eec7c03470 |
unverified |
Apache-2.0 (permissive) |
| arXiv:2023.emnlp-main.534 |
|
PreferredAI/superposed-topics/llama/generation.py 8845976729f4c4ee |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |