Browse State-of-the-Art › Instruction Following › Papers, page 9
Instruction Following
Papers archive 2025-07-28
archive papers tagged: 1,135 · with a code link: 609 · where Syntology ran a sample: 311 (255 with a run with no instrument failure, 56 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (311 of 1,135 tagged: 255 with a run with no instrument failure, 56 where every run was a failure of Syntology's instrument)
Page 9 of 12: papers 801 to 900 of 1,135, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Optimizing Latent Goal by Learning from Trajectory Preference3 Dec 2024 0 repositories listed
-
T-REG: Preference Optimization with Token-Level Reward Regularization3 Dec 2024 0 repositories listed
-
AlignFormer: Modality Matching Can Achieve Better Zero-shot Instruction-Following Speech-LLM2 Dec 2024 0 repositories listed
-
Enhancing Function-Calling Capabilities in LLMs: Strategies for Prompt Formats, Data Integration, and Multilingual Translation2 Dec 2024 0 repositories listed
-
MiningGPT -- A Domain-Specific Large Language Model for the Mining Industry2 Dec 2024 0 repositories listed
-
VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by Video Spatiotemporal Augmentation1 Dec 2024 0 repositories listed
-
InsightEdit: Towards Better Instruction Following for Image Editing26 Nov 2024 0 repositories listed
-
Gaussian Scenes: Pose-Free Sparse-View Scene Reconstruction using Depth-Enhanced Diffusion Priors24 Nov 2024 0 repositories listed
-
Enhancing Instruction-Following Capability of Visual-Language Models by Reducing Image Redundancy23 Nov 2024 0 repositories listed
-
From MTEB to MTOB: Retrieval-Augmented Classification for Descriptive Grammars23 Nov 2024 0 repositories listed
-
Separable Mixture of Low-Rank Adaptation for Continual Visual Instruction Tuning21 Nov 2024 0 repositories listed
-
Adaptive Decoding via Latent Preference Optimization14 Nov 2024 0 repositories listed
-
Zero-shot Object-Centric Instruction Following: Integrating Foundation Models with Traditional Navigation12 Nov 2024 0 repositories listed
-
MrSteve: Instruction-Following Agents in Minecraft with What-Where-When Memory11 Nov 2024 0 repositories listed
-
Stronger Models are NOT Stronger Teachers for Instruction Tuning11 Nov 2024 0 repositories listed
-
Fox-1 Technical Report8 Nov 2024 0 repositories listed
-
Multi-Reward as Condition for Instruction-based Image Editing6 Nov 2024 0 repositories listed
-
Data Extraction Attacks in Retrieval-Augmented Generation via Backdoors3 Nov 2024 0 repositories listed
-
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models2 Nov 2024 0 repositories listed
-
UFT: Unifying Fine-Tuning of SFT and RLHF/DPO/UNA through a Generalized Implicit Reward Function28 Oct 2024 0 repositories listed
-
SWITCH: Studying with Teacher for Knowledge Distillation of Large Language Models25 Oct 2024 0 repositories listed
-
BioMistral-NLU: Towards More Generalizable Medical Language Understanding through Instruction Tuning24 Oct 2024 0 repositories listed
-
Unbounded: A Generative Infinite Game of Character Life Simulation24 Oct 2024 0 repositories listed
-
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains23 Oct 2024 0 repositories listed
-
Towards Understanding the Fragility of Multilingual LLMs against Fine-Tuning Attacks23 Oct 2024 0 repositories listed
-
Large Language Models for Autonomous Driving (LLM4AD): Concept, Benchmark, Experiments, and Challenges20 Oct 2024 0 repositories listed
-
LLaVA-Ultra: Large Chinese Language and Vision Assistant for Ultrasound19 Oct 2024 0 repositories listed
-
Boosting LLM Translation Skills without General Ability Loss via Rationale Distillation17 Oct 2024 0 repositories listed
-
POROver: Improving Safety and Reducing Overrefusal in Large Language Models with Overgeneration and Preference Optimization16 Oct 2024 0 repositories listed
-
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding15 Oct 2024 0 repositories listed
-
Speculative Knowledge Distillation: Bridging the Teacher-Student Gap Through Interleaved Sampling15 Oct 2024 0 repositories listed
-
Balancing Continuous Pre-Training and Instruction Fine-Tuning: Optimizing Instruction-Following in LLMs14 Oct 2024 0 repositories listed
-
DrivingDojo Dataset: Advancing Interactive and Knowledge-Enriched Driving World Model14 Oct 2024 0 repositories listed
-
ForgeryGPT: Multimodal Large Language Model For Explainable Image Forgery Detection and Localization14 Oct 2024 0 repositories listed
-
Optimizing Instruction Synthesis: Effective Exploration of Evolutionary Space with Tree Search14 Oct 2024 0 repositories listed
-
Thinking LLMs: General Instruction Following with Thought Generation14 Oct 2024 0 repositories listed
-
Conversational Code Generation: a Case Study of Designing a Dialogue System for Generating Driving Scenarios for Testing Autonomous Vehicles13 Oct 2024 0 repositories listed
-
Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models13 Oct 2024 0 repositories listed
-
Are You Human? An Adversarial Benchmark to Expose LLMs12 Oct 2024 0 repositories listed
-
SeRA: Self-Reviewing and Alignment of Large Language Models using Implicit Reward Margins12 Oct 2024 0 repositories listed
-
Nudging: Inference-time Alignment of LLMs via Guided Decoding11 Oct 2024 0 repositories listed
-
Evolutionary Contrastive Distillation for Language Model Alignment10 Oct 2024 0 repositories listed
-
HERM: Benchmarking and Enhancing Multimodal LLMs for Human-Centric Understanding9 Oct 2024 0 repositories listed
-
Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy9 Oct 2024 0 repositories listed
-
Large Language Model Compression with Neural Architecture Search9 Oct 2024 0 repositories listed
-
LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints9 Oct 2024 0 repositories listed
-
Self-Boosting Large Language Models with Synthetic Preference Data9 Oct 2024 0 repositories listed
-
Multimodal Situational Safety8 Oct 2024 0 repositories listed
-
Direct Preference Optimization for LLM-Enhanced Recommendation Systems8 Oct 2024 0 repositories listed
-
TOWER: Tree Organized Weighting for Evaluating Complex Instructions8 Oct 2024 0 repositories listed
-
On Instruction-Finetuning Neural Machine Translation Models7 Oct 2024 0 repositories listed
-
RevisEval: Improving LLM-as-a-Judge via Response-Adapted References7 Oct 2024 0 repositories listed
-
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe7 Oct 2024 0 repositories listed
-
Superficial Safety Alignment Hypothesis7 Oct 2024 0 repositories listed
-
Only-IF:Revealing the Decisive Effect of Instruction Diversity on Generalization7 Oct 2024 0 repositories listed
-
SAG: Style-Aligned Article Generation via Model Collaboration4 Oct 2024 0 repositories listed
-
TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation4 Oct 2024 0 repositories listed
-
Better Instruction-Following Through Minimum Bayes Risk3 Oct 2024 0 repositories listed
-
LLaVA-Critic: Learning to Evaluate Multimodal Models3 Oct 2024 0 repositories listed
-
LoGra-Med: Long Context Multi-Graph Alignment for Medical Vision-Language Model3 Oct 2024 0 repositories listed
-
3 Oct 2024 0 repositories listed
-
The Perfect Blend: Redefining RLHF with Mixture of Judges30 Sep 2024 0 repositories listed
-
Revisiting the Superficial Alignment Hypothesis27 Sep 2024 0 repositories listed
-
Inference-Time Language Model Alignment via Integrated Value Guidance26 Sep 2024 0 repositories listed
-
MMMT-IF: A Challenging Multimodal Multi-Turn Instruction Following Benchmark26 Sep 2024 0 repositories listed
-
EAGLE: Towards Efficient Arbitrary Referring Visual Prompts Comprehension for Multimodal Large Language Models25 Sep 2024 0 repositories listed
-
Eliciting Instruction-tuned Code Language Models' Capabilities to Utilize Auxiliary Function for Code Generation20 Sep 2024 0 repositories listed
-
CamelEval: Advancing Culturally Aligned Arabic Language Models and Benchmarks19 Sep 2024 0 repositories listed
-
Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models17 Sep 2024 0 repositories listed
-
SIFToM: Robust Spoken Instruction Following through Theory of Mind17 Sep 2024 0 repositories listed
-
Emo-DPO: Controllable Emotional Speech Synthesis through Direct Preference Optimization16 Sep 2024 0 repositories listed
-
SFR-RAG: Towards Contextually Faithful LLMs16 Sep 2024 0 repositories listed
-
Keypoints-Integrated Instruction-Following Data Generation for Enhanced Human Pose Understanding in Multimodal Models14 Sep 2024 0 repositories listed
-
StressPrompt: Does Stress Impact Large Language Models and Human Performance Similarly?14 Sep 2024 0 repositories listed
-
RNR: Teaching Large Language Models to Follow Roles and Rules10 Sep 2024 0 repositories listed
-
Leveraging LLMs for Influence Path Planning in Proactive Recommendation7 Sep 2024 0 repositories listed
-
Continual Skill and Task Learning via Dialogue5 Sep 2024 0 repositories listed
-
Prompt Baking4 Sep 2024 0 repositories listed
-
Language Models Benefit from Preparation with Elicited Knowledge2 Sep 2024 0 repositories listed
-
Entropic Distribution Matching in Supervised Fine-tuning of LLMs: Less Overfitting and Better Diversity29 Aug 2024 0 repositories listed
-
M4CXR: Exploring Multi-task Potentials of Multi-modal Large Language Models for Chest X-ray Interpretation29 Aug 2024 0 repositories listed
-
Multi-Modal Instruction-Tuning Small-Scale Language-and-Vision Assistant for Semiconductor Electron Micrograph Analysis27 Aug 2024 0 repositories listed
-
Parameter-Efficient Quantized Mixture-of-Experts Meets Vision-Language Instruction Tuning for Semiconductor Electron Micrograph Analysis27 Aug 2024 0 repositories listed
-
Foundational Model for Electron Micrograph Analysis: Instruction-Tuning Small-Scale Language-and-Vision Assistant for Enterprise Adoption23 Aug 2024 0 repositories listed
-
Preference Consistency Matters: Enhancing Preference Learning in Language Models with Automated Self-Curation of Training Corpora23 Aug 2024 0 repositories listed
-
Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation19 Aug 2024 0 repositories listed
-
Can Large Language Models Understand Symbolic Graphics Programs?15 Aug 2024 0 repositories listed
-
13 Aug 2024 0 repositories listed
-
Space-LLaVA: a Vision-Language Model Adapted to Extraterrestrial Applications12 Aug 2024 0 repositories listed
-
Creating Arabic LLM Prompts at Scale12 Aug 2024 0 repositories listed
-
Empirical Analysis of Large Vision-Language Models against Goal Hijacking via Visual Prompt Injection7 Aug 2024 0 repositories listed
-
EXAONE 3.0 7.8B Instruction Tuned Language Model7 Aug 2024 0 repositories listed
-
A Framework for Fine-Tuning LLMs using Heterogeneous Feedback5 Aug 2024 0 repositories listed
-
Semantic Skill Grounding for Embodied Instruction-Following in Cross-Domain Environments2 Aug 2024 0 repositories listed
-
SaulLM-54B & SaulLM-141B: Scaling Up Domain Adaptation for the Legal Domain28 Jul 2024 0 repositories listed
-
HAPFI: History-Aware Planning based on Fused Information23 Jul 2024 0 repositories listed
-
Failures to Find Transferable Image Jailbreaks Between Vision-Language Models21 Jul 2024 0 repositories listed
-
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities19 Jul 2024 0 repositories listed
-
Situated Instruction Following15 Jul 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.