| Modelpedia: A Catalog of Model Findings for the Meta-Science of AI added by Syntology |
2026-09 (from id) |
tatsu-lab/stanford_alpaca/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| From Compression to Deployment: Real-Time and Energy-Efficient FastGRNN on Ultra-Constrained Microcontrollers added by Syntology |
2026-06 (from id) |
emre1998/fastgrnn-har/analyze_bootstrap.py 2a1e46172a51de87 |
unverified |
Apache-2.0 (permissive) |
| AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering added by Syntology |
2026-05 (from id) |
tatsu-lab/stanford_alpaca/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning NVIDIA added by Syntology |
2026-04 (from id) |
NVIDIA-NeMo/Skills/nemo_skills/file_utils.py 8c48543062658550 |
unverified |
Apache-2.0 (permissive) |
| SPA: A Simple but Tough-to-Beat Baseline for Knowledge Injection added by Syntology |
2026-03 (from id) |
Tangkexian/SPA/src/utils_tools/io_utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Exhaustive Circuit Mapping of a Single-Cell Foundation Model Reveals Massive Redundancy, Heavy-Tailed Hub Architecture, and Layer-Dependent Differentiation Control added by Syntology |
2026-03 (from id) |
Biodyn-AI/sae-biological-map/src/revision/make_figures.py 8c0c8031ac6f0fc4 |
unverified |
MIT (permissive) |
| AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset |
23 Apr 2025 |
kipok/nemo-skills/nemo_skills/file_utils.py 8c48543062658550 |
unverified |
Apache-2.0 (permissive) |
| S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning |
18 Feb 2025 |
nineabyss/s2r/code/utils.py 17c4e4f3635d3ee1 |
unverified |
MIT (permissive) |
| Virus: Harmful Fine-tuning Attack for Large Language Models Bypassing Guardrail Moderation |
29 Jan 2025 |
git-disl/virus/datasets_visualizer.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models |
16 Dec 2024 |
amourwaltz/ualign/code/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Denial-of-Service Poisoning Attacks against Large Language Models |
14 Oct 2024 |
sail-sg/p-dos/pdos_loss.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation |
13 Oct 2024 |
lslland/t-vaccine/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| SecAlign: Defending Against Prompt Injection with Preference Optimization |
7 Oct 2024 |
facebookresearch/secalign/struq.py 331e906ad5e247b9 |
ran
|
licence not identified · pointer only |
| OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data |
2 Oct 2024 |
NVIDIA/NeMo-Skills/nemo_skills/file_utils.py 8c48543062658550 |
unverified |
Apache-2.0 (permissive) |
| Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey |
26 Sep 2024 |
git-disl/booster/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Synthetic continued pretraining |
11 Sep 2024 |
zitongyang/synthetic_continued_pretraining/utils/io_utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Instruct-SkillMix: A Powerful Pipeline for LLM Instruction Tuning |
27 Aug 2024 |
princeton-pli/Instruct-SkillMix/MAmmoTH/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning |
18 Aug 2024 |
git-disl/lisa/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Analyzing the Generalization and Reliability of Steering Vectors |
17 Jul 2024 |
dtch1997/steering-bench/steering_bench/utils/io.py 856c8b4ec46ae412 |
ran
|
no licence file found · pointer only |
| On Large Language Model Continual Unlearning |
14 Jul 2024 |
GCYZSL/O3-LLM-UNLEARNING/ScienceQA_expriment/src/base_model_utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| LLM Critics Help Catch Bugs in Mathematics: Towards a Better Mathematical Verifier with Natural Language Feedback |
20 Jun 2024 |
kbsdjames/math-minos/train_reward.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| SinkLoRA: Enhanced Efficiency and Chat Capabilities for Long-Context Large Language Models |
9 Jun 2024 |
identical code first harvested elsewhere d07d04439cd1d44f |
ran · our draft was wrong
|
licence of this copy not recorded |
| GrootVL: Tree Topology is All You Need in State Space Model |
4 Jun 2024 |
easonxiao-888/grootvl/GrootL/supervised-fine-tune.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Various Lengths, Constant Speed: Efficient Language Modeling with Lightning Attention |
27 May 2024 |
opennlplab/transnormerllm/fine-tune/utils.py 5c5247729979fd74 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Various Lengths, Constant Speed: Efficient Language Modeling with Lightning Attention |
27 May 2024 |
OpenNLPLab/TransnormerLLM/fine-tune/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Can LLMs Solve longer Math Word Problems Better? |
23 May 2024 |
identical code first harvested elsewhere d07d04439cd1d44f |
ran · our draft was wrong
|
licence of this copy not recorded |
| LIRE: listwise reward enhancement for preference alignment |
22 May 2024 |
stevie1023/LIRE/train_alpaca_prompt.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Exploring the Compositional Deficiency of Large Language Models in Mathematical Reasoning |
5 May 2024 |
tongjingqi/MathTrap/train_math.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Exploring Backdoor Vulnerabilities of Chat Models |
3 Apr 2024 |
hychaochao/chat-models-backdoor-attacking/Instructional_Model_Backdoor/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| LLM2LLM: Boosting LLMs with Novel Iterative Data Enhancement |
22 Mar 2024 |
squeezeailab/llm2llm/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
MIT (permissive) |
| Unveiling the Generalization Power of Fine-Tuned Large Language Models |
14 Mar 2024 |
lhryang/generalization_of_ft-llm/utils.py 220607d53f1814c7 |
ran
|
no licence file found · pointer only |
| How Susceptible are Large Language Models to Ideological Manipulation? |
18 Feb 2024 |
kaichen23/llm_ideo_manipulate/code/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
MIT (permissive) |
| Smaller Language Models are capable of selecting Instruction-Tuning Training Data for Larger Language Models |
16 Feb 2024 |
dheeraj7596/small2large/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| NutePrune: Efficient Progressive Pruning with Numerous Teachers for Large Language Models |
15 Feb 2024 |
lucius-lsr/nuteprune/tasks/alpaca.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Answer is All You Need: Instruction-following Text Embedding via Answering the Question |
15 Feb 2024 |
zhang-yu-wei/inbedder/alpaca_train/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
MIT (permissive) |
| StruQ: Defending Against Prompt Injection with Structured Queries |
2024-02 (from id) |
sizhe-chen/struq/struq.py 331e906ad5e247b9 |
ran
|
no licence file found · pointer only |
| WaveCoder: Widespread And Versatile Enhancement For Code Large Language Models By Instruction Tuning |
20 Dec 2023 |
microsoft/wavecoder/src/train/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
MIT (permissive) |
| Sparse is Enough in Fine-tuning Pre-trained Large Language Models |
19 Dec 2023 |
song-wx/SIFT/exp/instruction_finetuning/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Pragmatic Radiology Report Generation |
28 Nov 2023 |
chicagohai/llm_radiology/utils_finetune.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| MoDS: Model-oriented Data Selection for Instruction Tuning |
27 Nov 2023 |
casia-lm/mods/inference/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection |
17 Oct 2023 |
AkariAsai/self-rag/data_creation/train_special_tokens.py 3ca38f9acc00194b |
ran
|
MIT (permissive) |
| Unlocking Emergent Modularity in Large Language Models |
17 Oct 2023 |
qiuzh20/emoe/EMoE_LLaMA/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
MIT (permissive) |
| Explore-Instruct: Enhancing Domain-Specific Instruction Coverage through Active Exploration |
13 Oct 2023 |
fanqiwan/Explore-Instruct/train/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models |
21 Sep 2023 |
dvlab-research/longlora/supervised-fine-tune-qlora.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models |
21 Sep 2023 |
meta-math/MetaMath/train_math.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Publicly Shareable Clinical Large Language Model Built on Synthetic Clinical Notes |
1 Sep 2023 |
starmpcc/asclepius/src/instruction_ft.py 0c9d19493ee33877 |
ran · our draft was wrong
|
no licence file found · pointer only |
| From Quantity to Quality: Boosting LLM Performance with Self-Guided Data Selection for Instruction Tuning |
23 Aug 2023 |
mingliiii/cherry_llm/training/stanford_alpaca/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Token-Scaled Logit Distillation for Ternary Weight Generative Language Models |
13 Aug 2023 |
aiha-lab/TSLD/utils/alpaca_dataset.py d07d04439cd1d44f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Self-Alignment with Instruction Backtranslation |
11 Aug 2023 |
davidkim205/komt/finetune_with_ds.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Extrapolating Large Language Models to Non-English by Aligning Languages |
9 Aug 2023 |
NJUNLP/x-LLM/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| On the Exploitability of Instruction Tuning |
28 Jun 2023 |
azshue/AutoPoison/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners |
24 May 2023 |
xiaojuantang/icsr/finetune/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Making Language Models Better Tool Learners with Execution Feedback |
22 May 2023 |
identical code first harvested elsewhere d07d04439cd1d44f |
ran · our draft was wrong
|
licence of this copy not recorded |
| Lion: Adversarial Distillation of Proprietary Large Language Models |
22 May 2023 |
yjiangcm/lion/src/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
MIT (permissive) |
| LLaMA: Open and Efficient Foundation Language Models |
27 Feb 2023 |
aethercortex/llama-x/src/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| SBERT studies Meaning Representations: Decomposing Sentence Embeddings into Explainable Semantic Features |
14 Jun 2022 |
flipz357/S3BERT/src/data_helpers.py 1779ed436ac80094 |
unverified |
MIT (permissive) |
| arXiv:aaai_34662 |
|
maxindian/3D-RPE-Long-Contex-Modeling/sft_qlora_tuning.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| arXiv:aaai_29784 |
|
HITsz-TMG/Ext-Sub/training/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
MIT (permissive) |
| arXiv:2024.naacl-long.132 |
|
Kent0n-Li/ChatDoctor/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| arXiv:2024.findings-acl.388 |
|
SqueezeAILab/LLM2LLM/utils.py d07d04439cd1d44f |
ran · our draft was wrong
|
MIT (permissive) |