| ReCache: Efficient KV Cache Reuse and Compression for Tool-Augmented LLM Agents added by Syntology |
2026-08 (from id) |
EIT-NLP/ReCache/capsule/src/chat.py 63d36f1757a0760e |
unverified |
no licence file found · pointer only |
| RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction added by Syntology |
2026-08 (from id) |
allenai/WildBench/src/fastchat_conversation.py 2e17447aecaa3b2e |
unverified |
Apache-2.0 (permissive) |
| Focus When Necessary: Adaptive Routing and Collaborative Grounding for Training-Free Visual Grounding added by Syntology |
2026-06 (from id) |
TencentBAC/LazyMCoT/Internvl/conversation.py 4b4a1446c2b464c6 |
unverified |
Apache-2.0 (permissive) |
| Representation-Aware Advantage Estimation: Your Reward Model Provides More Than A Scalar Output added by Syntology |
2026-06 (from id) |
fanqiwan/FuseAI/FuseChat/train/conversation.py 502fd3534e0509c1 |
unverified |
no licence file found · pointer only |
| Psychological Steering of Large Language Models added by Syntology |
2026-04 (from id) |
kaistAI/FLASK/model_output/conversation.py a566c582d26ceb36 |
unverified |
no licence file found · pointer only |
| DSPA: Dynamic SAE Steering for Data-Efficient Preference Alignment added by Syntology |
2026-03 (from id) |
lm-sys/FastChat/fastchat/conversation.py 0e29f2d151ce2bd7 |
unverified |
Apache-2.0 (permissive) |
| R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation added by Syntology |
2026-02 (from id) |
czh24/R3G/src/pipeline.py 7b9e08d60e09fba0 |
unverified |
no licence file found · pointer only |
| UniX: Unifying Autoregression and Diffusion for Chest X-Ray Understanding and Generation added by Syntology |
2026-01 (from id) |
ZrH42/UniX/modeling/unix_vlm/utils/conversation.py 6af454338dc8d2a5 |
unverified |
MIT (permissive) |
| LiveStar: Live Streaming Assistant for Real-World Online Video Understanding added by Syntology |
2025-11 (from id) |
yzy-bupt/LiveStar/livestar/conversation.py 04b172bb0c32a0de |
unverified |
no licence file found · pointer only |
| BiasBusters: Uncovering and Mitigating Tool Selection Bias in Large Language Models added by Syntology |
2025-10 (from id) |
thierry123454/tool-selection-bias/toolbench/tool_conversation.py eab40e1435b85709 |
unverified |
no licence file found · pointer only |
| Unveiling Chain of Step Reasoning for Vision-Language Models with Fine-grained Rewards added by Syntology |
2025-09 (from id) |
baaivision/CoS/internvl/conversation.py 04b172bb0c32a0de |
unverified |
Apache-2.0 (permissive) |
| arXiv:2507.17539 |
2025-07 (from id) |
MeteorElf/FundusExpert/src/internvl_chat/internvl/conversation.py 04b172bb0c32a0de |
unverified |
Apache-2.0 (permissive) |
| Amulet: Putting Complex Multi-Turn Conversations on the Stand with LLM Juries |
26 May 2025 |
thunlp/ChatEval/FastChat/fastchat/conversation.py a619c795784c4043 |
unverified |
Apache-2.0 (permissive) |
| MP-GUI: Modality Perception with MLLMs for GUI Understanding |
18 Mar 2025 |
BigTaige/MP-GUI/model/internvl/conversation.py 04b172bb0c32a0de |
unverified |
MIT (permissive) |
| LServe: Efficient Long-sequence LLM Serving with Unified Sparse Attention |
20 Feb 2025 |
mit-han-lab/qserve/omniserve/conversation.py 3b545dce3ebb9392 |
unverified |
Apache-2.0 (permissive) |
| LEO: Boosting Mixture of Vision Encoders for Multimodal Large Language Models |
13 Jan 2025 |
mozhgan91/leo/leo_chat/leo/conversation.py 04b172bb0c32a0de |
unverified |
Apache-2.0 (permissive) |
| AutoTrust: Benchmarking Trustworthiness in Large Vision Language Models for Autonomous Driving |
19 Dec 2024 |
taco-group/autotrust/Dolphins/conversation.py dbbd17a289ba66a6 |
unverified |
Apache-2.0 (permissive) |
| DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding |
13 Dec 2024 |
deepseek-ai/deepseek-vl2/deepseek_vl2/models/conversation.py 132894d1028abd5a |
unverified |
MIT (permissive) |
| Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity |
9 Dec 2024 |
pipixin321/holmesvau/internvl_chat/internvl/conversation.py 04b172bb0c32a0de |
unverified |
MIT (permissive) |
| Weighted-Reward Preference Optimization for Implicit Model Fusion |
4 Dec 2024 |
fanqiwan/fuseai/FuseChat/train/conversation.py 502fd3534e0509c1 |
unverified |
no licence file found · pointer only |
| StepTool: A Step-grained Reinforcement Learning Framework for Tool Learning in LLMs |
10 Oct 2024 |
yuyq18/steptool/src/baseline-rft/rft.py a822c97372b609f0 |
unverified |
no licence file found · pointer only |
| Visual Perception in Text Strings |
2 Oct 2024 |
JiaQiSJTU/VisionInText/src/utils/conversations.py 174d0858d8cf33ca |
unverified |
no licence file found · pointer only |
| MemoRAG: Moving towards Next-Gen RAG Via Memory-Inspired Knowledge Discovery |
9 Sep 2024 |
qhjqhj00/memorag/train/src/chat.py 63d36f1757a0760e |
unverified |
Apache-2.0 (permissive) |
| Instruct-SkillMix: A Powerful Pipeline for LLM Instruction Tuning |
27 Aug 2024 |
princeton-pli/Instruct-SkillMix/WildBench/src/fastchat_conversation.py 5cb43a8f85ec84a6 |
unverified |
no licence file found · pointer only |
| Imagen 3 |
13 Aug 2024 |
linzhiqiu/t2v_metrics/t2v_metrics/models/vqascore_models/fastchat_utils.py 4b4a1446c2b464c6 |
unverified |
Apache-2.0 (permissive) |
| EfficientQAT: Efficient Quantization-Aware Training for Large Language Models |
10 Jul 2024 |
opengvlab/efficientqat/deita_dataset/conversation.py a566c582d26ceb36 |
unverified |
MIT (permissive) |
| Exploring Design Choices for Building Language-Specific LLMs |
20 Jun 2024 |
atutej/token-language-adaptation/FastChat/fastchat/conversation.py 506ee6d4bab01412 |
unverified |
no licence file found · pointer only |
| DocGenome: An Open Large-scale Scientific Document Benchmark for Training and Testing Multi-modal Large Language Models |
17 Jun 2024 |
Alpha-Innovator/StructEqTable-Deploy/struct_eqtable/internvl/conversation.py bd0b36d66d5aba83 |
unverified |
Apache-2.0 (permissive) |
| Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement |
17 Jun 2024 |
weiminxiong/ipr/fastchat/conversation.py 803c2e6f53564248 |
unverified |
no licence file found · pointer only |
| Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs |
14 Jun 2024 |
ahans30/goldfish-loss/eval/conversation.py 6db38b4d818a2c2e |
unverified |
Apache-2.0 (permissive) |
| Needle In A Multimodal Haystack |
11 Jun 2024 |
opengvlab/mm-niah/utils/conversation.py d701841310923e61 |
unverified |
no licence file found · pointer only |
| FedLLM-Bench: Realistic Benchmarks for Federated Learning of Large Language Models |
7 Jun 2024 |
rui-ye/fedllm-bench/conversation.py 174d0858d8cf33ca |
unverified |
no licence file found · pointer only |
| FedLLM-Bench: Realistic Benchmarks for Federated Learning of Large Language Models |
7 Jun 2024 |
rui-ye/openfedllm/utils/conversation.py 506ee6d4bab01412 |
unverified |
Apache-2.0 (permissive) |
| WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild |
7 Jun 2024 |
allenai/wildbench/src/fastchat_conversation.py 2e17447aecaa3b2e |
unverified |
Apache-2.0 (permissive) |
| Are We Done with MMLU? |
6 Jun 2024 |
wildeval/zeroeval/src/fastchat_conversation.py 2e17447aecaa3b2e |
unverified |
Apache-2.0 (permissive) |
| BadAgent: Inserting and Activating Backdoor Attacks in LLM Agents |
5 Jun 2024 |
DPamK/BadAgent/utils/llm_dataset.py ec71aaaf79219926 |
ran
|
no licence file found · pointer only |
| BadAgent: Inserting and Activating Backdoor Attacks in LLM Agents |
5 Jun 2024 |
DPamK/BadAgent/utils/conversation.py 25634eda28d2513f |
unverified |
no licence file found · pointer only |
| Is In-Context Learning Sufficient for Instruction Following in LLMs? |
30 May 2024 |
tml-epfl/icl-alignment/src/fastchat_conversation.py 86b180ae170d46ff |
unverified |
Apache-2.0 (permissive) |
| Agent Planning with World Knowledge Model |
23 May 2024 |
zjunlp/wkm/src/eval/fastchat/conversation.py eefaa52fc92c8f55 |
unverified |
Apache-2.0 (permissive) |
| GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration |
23 Apr 2024 |
sunanhe/meddr/src/utils/conversation.py bd0b36d66d5aba83 |
unverified |
MIT (permissive) |
| Exploring Backdoor Vulnerabilities of Chat Models |
3 Apr 2024 |
hychaochao/chat-models-backdoor-attacking/fastchat/conversation.py 8ba434c7007f1b4a |
unverified |
Apache-2.0 (permissive) |
| StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models |
12 Mar 2024 |
openbmb/toolbench/toolbench/tool_conversation.py eab40e1435b85709 |
unverified |
Apache-2.0 (permissive) |
| Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference |
7 Mar 2024 |
BirgerMoell/SwedishLLMBenchmark/fastchat/conversation.py c0d59bb5b224b1a8 |
unverified |
Apache-2.0 (permissive) |
| Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference |
7 Mar 2024 |
Peter-Devine/multilingual_mt_bench/fastchat/conversation.py abb1b6149fb6556f |
unverified |
Apache-2.0 (permissive) |
| Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents |
4 Mar 2024 |
yifan-song793/eto/fastchat/conversation.py eefaa52fc92c8f55 |
unverified |
no licence file found · pointer only |
| NewsBench: A Systematic Evaluation Framework for Assessing Editorial Capabilities of Large Language Models in Chinese Journalism |
29 Feb 2024 |
iaar-shanghai/newsbench/eval_scripts/aquila_predict.py f393883643980cc1 |
unverified |
Apache-2.0 (permissive) |
| Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic |
19 Feb 2024 |
declare-lab/red-instruct/starling_training/fastchat/conversation.py bd62a3e93d9530bd |
unverified |
Apache-2.0 (permissive) |
| Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents |
18 Feb 2024 |
reason-wang/nat/prompts/conversations.py f3dc205319b0e027 |
unverified |
no licence file found · pointer only |
| On the Robustness of Editing Large Language Models |
8 Feb 2024 |
xbmxb/edit_analysis/code/conversation.py af4d8c865cc2df7b |
unverified |
no licence file found · pointer only |
| GITA: Graph to Visual and Textual Integration for Vision-Language Graph Reasoning |
3 Feb 2024 |
WEIYanbin1999/GITA/fastchat/conversation.py 25f1c2de6ce6ae79 |
unverified |
no licence file found · pointer only |
| RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning |
16 Jan 2024 |
Junjie-Ye/RoTBench/Code/tool_conversation.py 6aa619d359276223 |
unverified |
Apache-2.0 (permissive) |
| Intention Analysis Makes LLMs A Good Jailbreak Defender |
12 Jan 2024 |
alphadl/safellm_with_intentionanalysis/demo/conversation.py 9ec48e73fc2bd33f |
unverified |
no licence file found · pointer only |
| What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning |
25 Dec 2023 |
hkust-nlp/deita/src/deita/alignment/conversation.py 1c12e90632ed637e |
unverified |
Apache-2.0 (permissive) |
| Dolphins: Multimodal Language Model for Driving |
1 Dec 2023 |
safolab-wisc/dolphins/conversation.py dbbd17a289ba66a6 |
unverified |
MIT (permissive) |
| VIoTGPT: Learning to Schedule Vision Tools in LLMs towards Intelligent Video Internet of Things |
1 Dec 2023 |
zhongyy/viotgpt/train/tool_conversation.py 9cea3d56fecb304b |
unverified |
no licence file found · pointer only |
| Taiwan LLM: Bridging the Linguistic Divide with a Culturally Aligned Language Model |
29 Nov 2023 |
miulab/taiwan-llm/evaluation/conversation.py 2613e53cccc8a4b4 |
unverified |
Apache-2.0 (permissive) |
| Can LLMs Follow Simple Rules? |
6 Nov 2023 |
normster/llm_rules/llm_rules/fastchat_templates.py 47d9aeb23f9ca568 |
unverified |
Apache-2.0 (permissive) |
| JudgeLM: Fine-tuned Large Language Models are Scalable Judges |
26 Oct 2023 |
hitz-zentroa/eval-MCG-COLING-2025/evaluation/JudgeLM-main/judgelm/conversation.py d3f7074a99b4f5dd |
unverified |
Apache-2.0 (permissive) |
| H2O Open Ecosystem for State-of-the-art Large Language Models |
17 Oct 2023 |
h2oai/h2ogpt/models/predict_aquila.py f393883643980cc1 |
unverified |
Apache-2.0 (permissive) |
| Qilin-Med: Multi-stage Knowledge Injection Advanced Medical Large Language Model |
13 Oct 2023 |
williamliujl/Qilin-Med/scripts/supervised_finetuning.py 6cd4acc9be416d47 |
unverified |
no licence file found · pointer only |
| Code Soliloquies for Accurate Calculations in Large Language Models |
21 Sep 2023 |
luffycodes/tutorbot-spock-phys/fastchat/conversation_inference.py a3f8708160205857 |
unverified |
no licence file found · pointer only |
| ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate |
14 Aug 2023 |
chanchimin/chateval/FastChat/fastchat/conversation.py a619c795784c4043 |
unverified |
Apache-2.0 (permissive) |
| UniversalNER: Targeted Distillation from Large Language Models for Open Named Entity Recognition |
7 Aug 2023 |
universal-ner/universal-ner/src/conversation.py dab937567e76e9e3 |
unverified |
MIT (permissive) |
| LLaMA: Open and Efficient Foundation Language Models |
27 Feb 2023 |
aethercortex/llama-x/src/conversation.py 614f3f063d0ec7f0 |
unverified |
Apache-2.0 (permissive) |
| Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks |
16 Apr 2022 |
naver-ai/rethinking-proxy-reward/conversation.py a3ac8409c5a9506e |
unverified |
Apache-2.0 (permissive) |
| Self-attention Does Not Need $O(n^2)$ Memory |
10 Dec 2021 |
stability-ai/fastchat/fastchat/conversation.py 2613e53cccc8a4b4 |
unverified |
Apache-2.0 (permissive) |
| CommonGen: A Constrained Text Generation Challenge for Generative Commonsense Reasoning |
9 Nov 2019 |
allenai/CommonGen-Eval/fastchat_conversation.py eefaa52fc92c8f55 |
unverified |
Apache-2.0 (permissive) |
| arXiv:openreview_9aU4vrHPKD |
|
Octobrist/CoPE/src/fastchat/conversation.py 0e29f2d151ce2bd7 |
unverified |
Apache-2.0 (permissive) |
| arXiv:Zhang_Holmes-VAU_Towards_Long-term_Video_Anomaly_Understanding_at_Any_Granularity_CVPR_2025_paper |
|
pipixin321/HolmesVAU/internvl_chat/internvl/conversation.py 04b172bb0c32a0de |
unverified |
MIT (permissive) |
| arXiv:2025.emnlp-main.1337 |
|
RazvanDu/CopySpec/FastChat/fastchat/conversation.py 0e29f2d151ce2bd7 |
unverified |
MIT (permissive) |
| arXiv:2025.acl-long.498 |
|
OpenGVLab/EfficientQAT/deita_dataset/conversation.py a566c582d26ceb36 |
unverified |
MIT (permissive) |
| arXiv:2024.findings-emnlp.559 |
|
hitz-zentroa/cn-eval/evaluation/JudgeLM-main/judgelm/conversation.py d3f7074a99b4f5dd |
unverified |
Apache-2.0 (permissive) |
| arXiv:2024.findings-emnlp.504 |
|
Re-Align/URIAL/src/fastchat_conversation.py 86b180ae170d46ff |
unverified |
Apache-2.0 (permissive) |
| arXiv:2024.findings-acl.39 |
|
passing2961/DribeR/conversation.py a566c582d26ceb36 |
unverified |
MIT (permissive) |
| arXiv:2024.findings-acl.259 |
|
JoeYing1019/UltraTool/inference/conversation.py a566c582d26ceb36 |
unverified |
Apache-2.0 (permissive) |