Browse State-of-the-Art › Instruction Following › Papers, page 10
Instruction Following
Papers archive 2025-07-28
archive papers tagged: 1,135 · with a code link: 609 · where Syntology ran a sample: 311 (255 with a run with no instrument failure, 56 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (311 of 1,135 tagged: 255 with a run with no instrument failure, 56 where every run was a failure of Syntology's instrument)
Page 10 of 12: papers 901 to 1,000 of 1,135, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Beyond Instruction Following: Evaluating Inferential Rule Following of Large Language Models11 Jul 2024 0 repositories listed
-
LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition9 Jul 2024 0 repositories listed
-
Large Language Model as an Assignment Evaluator: Insights, Feedback, and Challenges in a 1000+ Student Course7 Jul 2024 0 repositories listed
-
Diverse and Fine-Grained Instruction-Following Ability Exploration with Synthetic Data4 Jul 2024 0 repositories listed
-
D-Rax: Domain-specific Radiologic assistant leveraging multi-modal data and eXpert model predictions2 Jul 2024 0 repositories listed
-
Pelican: Correcting Hallucination in Vision-LLMs via Claim Decomposition and Program of Thought Verification2 Jul 2024 0 repositories listed
-
Improving Multilingual Instruction Finetuning via Linguistically Natural and Diverse Datasets1 Jul 2024 0 repositories listed
-
Iterative Data Generation with Large Language Models for Aspect-based Sentiment Analysis29 Jun 2024 0 repositories listed
-
ScaleBiO: Scalable Bilevel Optimization for LLM Data Reweighting28 Jun 2024 0 repositories listed
-
DeSTA: Enhancing Speech Language Models through Descriptive Speech-Text Alignment27 Jun 2024 0 repositories listed
-
OmniJARVIS: Unified Vision-Language-Action Tokenization Enables Open-World Instruction Following Agents27 Jun 2024 0 repositories listed
-
Role-Play Zero-Shot Prompting with Large Language Models for Open-Domain Human-Machine Conversation26 Jun 2024 0 repositories listed
-
A Text is Worth Several Tokens: Text Embedding from LLMs Secretly Aligns Well with The Key Tokens25 Jun 2024 0 repositories listed
-
Following Length Constraints in Instructions25 Jun 2024 0 repositories listed
-
Evaluation of Instruction-Following Ability for Large Language Models on Story-Ending Generation24 Jun 2024 0 repositories listed
-
Teach Better or Show Smarter? On Instructions and Exemplars in Automatic Prompt Optimization22 Jun 2024 0 repositories listed
-
DEM: Distribution Edited Model for Training with Mixed Data Distributions21 Jun 2024 0 repositories listed
-
AdaGrad under Anisotropic Smoothness21 Jun 2024 0 repositories listed
-
VLM Agents Generate Their Own Memories: Distilling Experience into Embodied Programs of Thought20 Jun 2024 0 repositories listed
-
Refine Large Language Model Fine-tuning via Instruction Vector18 Jun 2024 0 repositories listed
-
The Comparative Trap: Pairwise Comparisons Amplifies Biased Preferences of LLM Evaluators18 Jun 2024 0 repositories listed
-
Unveiling the Flaws: Exploring Imperfections in Synthetic Data and Mitigation Strategies for Large Language Models18 Jun 2024 0 repositories listed
-
Embodied Instruction Following in Unknown Environments17 Jun 2024 0 repositories listed
-
How Far Can In-Context Alignment Go? Exploring the State of In-Context Alignment17 Jun 2024 0 repositories listed
-
Enhancing and Assessing Instruction-Following with Fine-Grained Instruction Variants17 Jun 2024 0 repositories listed
-
Reminding Multimodal Large Language Models of Object-aware Knowledge with Retrieved Tags16 Jun 2024 0 repositories listed
-
Comparison Visual Instruction Tuning13 Jun 2024 0 repositories listed
-
DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding13 Jun 2024 0 repositories listed
-
Mimicking User Data: On Mitigating Fine-Tuning Risks in Closed Large Language Models12 Jun 2024 0 repositories listed
-
3D-Properties: Identifying Challenges in DPO and Charting a Path Forward11 Jun 2024 0 repositories listed
-
FaceGPT: Self-supervised Learning to Chat about 3D Human Faces11 Jun 2024 0 repositories listed
-
OPTune: Efficient Online Preference Tuning11 Jun 2024 0 repositories listed
-
Attend and Enrich: Enhanced Visual Prompt for Zero-Shot Learning5 Jun 2024 0 repositories listed
-
Scalable Ensembling For Mitigating Reward Overoptimisation3 Jun 2024 0 repositories listed
-
LIDAO: Towards Limited Interventions for Debiasing (Large) Language Models1 Jun 2024 0 repositories listed
-
clembench-2024: A Challenging, Dynamic, Complementary, Multilingual Benchmark and Underlying Flexible Framework for LLMs as Multi-Action Agents31 May 2024 0 repositories listed
-
Improving Reward Models with Synthetic Critiques31 May 2024 0 repositories listed
-
Joint Embeddings for Graph Instruction Tuning31 May 2024 0 repositories listed
-
InstructionCP: A fast approach to transfer Large Language Models into target language30 May 2024 0 repositories listed
-
Nadine: An LLM-driven Intelligent Social Robot with Affective Capabilities and Human-like Memory30 May 2024 0 repositories listed
-
30 May 2024 0 repositories listed
-
TS-Align: A Teacher-Student Collaborative Framework for Scalable Iterative Finetuning of Large Language Models30 May 2024 0 repositories listed
-
BLSP-KD: Bootstrapping Language-Speech Pre-training via Knowledge Distillation29 May 2024 0 repositories listed
-
X-VILA: Cross-Modality Alignment for Large Language Model29 May 2024 0 repositories listed
-
Self-Corrected Multimodal Large Language Model for End-to-End Robot Manipulation27 May 2024 0 repositories listed
-
From Role-Play to Drama-Interaction: An LLM Solution23 May 2024 0 repositories listed
-
RE-Adapt: Reverse Engineered Adaptation of Large Language Models23 May 2024 0 repositories listed
-
Distilling Instruction-following Abilities of Large Language Models with Task-aware Curriculum Planning22 May 2024 0 repositories listed
-
Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning16 May 2024 0 repositories listed
-
SpeechGuard: Exploring the Adversarial Robustness of Multimodal Large Language Models14 May 2024 0 repositories listed
-
SpeechVerse: A Large-scale Generalizable Audio Language Model14 May 2024 0 repositories listed
-
Long Context Alignment with Short Instructions and Synthesized Positions7 May 2024 0 repositories listed
-
Enabling High-Sparsity Foundational Llama Models with Efficient Pretraining and Deployment6 May 2024 0 repositories listed
-
FLAME: Factuality-Aware Alignment for Large Language Models2 May 2024 0 repositories listed
-
LLM-AD: Large Language Model based Audio Description System2 May 2024 0 repositories listed
-
WildChat: 1M ChatGPT Interaction Logs in the Wild2 May 2024 0 repositories listed
-
Visual Fact Checker: Enabling High-Fidelity Detailed Caption Generation30 Apr 2024 0 repositories listed
-
HELPER-X: A Unified Instructable Embodied Agent to Tackle Four Interactive Vision-Language Domains with Memory-Augmented Language Models29 Apr 2024 0 repositories listed
-
From Persona to Personalization: A Survey on Role-Playing Language Agents28 Apr 2024 0 repositories listed
-
Automatic Layout Planning for Visually-Rich Documents with Instruction-Following Models23 Apr 2024 0 repositories listed
-
Socratic Planner: Self-QA-Based Zero-Shot Planning for Embodied Instruction Following21 Apr 2024 0 repositories listed
-
Look Before You Decide: Prompting Active Deduction of MLLMs for Assumptive Reasoning19 Apr 2024 0 repositories listed
-
Closed-Loop Open-Vocabulary Mobile Manipulation with GPT-4V16 Apr 2024 0 repositories listed
-
Unveiling the Misuse Potential of Base Large Language Models via In-Context Learning16 Apr 2024 0 repositories listed
-
Sketch-Plan-Generalize: Learning and Planning with Neuro-Symbolic Programmatic Representations for Inductive Spatial Concepts11 Apr 2024 0 repositories listed
-
CodecLM: Aligning Language Models with Tailored Synthetic Data8 Apr 2024 0 repositories listed
-
Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs8 Apr 2024 0 repositories listed
-
CantTalkAboutThis: Aligning Language Models to Stay on Topic in Dialogues4 Apr 2024 0 repositories listed
-
An Incomplete Loop: Deductive, Inductive, and Abductive Learning in Large Language Models3 Apr 2024 0 repositories listed
-
HyperCLOVA X Technical Report2 Apr 2024 0 repositories listed
-
LLaMA-Excitor: General Instruction Tuning via Indirect Feature Interaction1 Apr 2024 0 repositories listed
-
30 Mar 2024 0 repositories listed
-
28 Mar 2024 0 repositories listed
-
Argument Quality Assessment in the Age of Instruction-Following Large Language Models24 Mar 2024 0 repositories listed
-
Few-shot Dialogue Strategy Learning for Motivational Interviewing via Inductive Reasoning23 Mar 2024 0 repositories listed
-
Improving the Robustness of Large Language Models via Consistency Alignment21 Mar 2024 0 repositories listed
-
VisualCritic: Making LMMs Perceive Visual Quality Like Humans19 Mar 2024 0 repositories listed
-
WoLF: Wide-scope Large Language Model Framework for CXR Understanding19 Mar 2024 0 repositories listed
-
Don't Half-listen: Capturing Key-part Information in Continual Instruction Tuning15 Mar 2024 0 repositories listed
-
Mitigating Dialogue Hallucination for Large Vision Language Models via Adversarial Instruction Tuning15 Mar 2024 0 repositories listed
-
DiffChat: Learning to Chat with Text-to-Image Synthesis Models for Interactive Image Creation8 Mar 2024 0 repositories listed
-
CoTBal: Comprehensive Task Balancing for Multi-Task Visual Instruction Tuning7 Mar 2024 0 repositories listed
-
KIWI: A Dataset of Knowledge-Intensive Writing Instructions for Answering Research Questions6 Mar 2024 0 repositories listed
-
CoGenesis: A Framework Collaborating Large and Small Language Models for Secure Context-Aware Instruction Following5 Mar 2024 0 repositories listed
-
OPEx: A Component-Wise Analysis of LLM-Centric Agents in Embodied Instruction Following5 Mar 2024 0 repositories listed
-
Collaborative decoding of critical tokens for boosting factuality of large language models28 Feb 2024 0 repositories listed
-
Think Big, Generate Quick: LLM-to-SLM for Fast Autoregressive Decoding26 Feb 2024 0 repositories listed
-
NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation24 Feb 2024 0 repositories listed
-
Fine-Tuning Enhances Existing Mechanisms: A Case Study on Entity Tracking22 Feb 2024 0 repositories listed
-
Zero-shot cross-lingual transfer in instruction tuning of large language models22 Feb 2024 0 repositories listed
-
Investigating Multilingual Instruction-Tuning: Do Polyglot Models Demand for Multilingual Instructions?21 Feb 2024 0 repositories listed
-
VL-Trojan: Multimodal Instruction Backdoor Attacks against Autoregressive Visual Language Models21 Feb 2024 0 repositories listed
-
CIF-Bench: A Chinese Instruction-Following Benchmark for Evaluating the Generalizability of Large Language Models20 Feb 2024 0 repositories listed
-
Synthetic Data (Almost) from Scratch: Generalized Instruction Tuning for Language Models20 Feb 2024 0 repositories listed
-
Transformer-based Causal Language Models Perform Clustering19 Feb 2024 0 repositories listed
-
Efficient Prompt Optimization Through the Lens of Best Arm Identification15 Feb 2024 0 repositories listed
-
Multi-Query Focused Disaster Summarization via Instruction-Based Prompting14 Feb 2024 0 repositories listed
-
Investigating the Impact of Data Contamination of Large Language Models in Text-to-SQL Translation12 Feb 2024 0 repositories listed
-
PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs12 Feb 2024 0 repositories listed
-
Nevermind: Instruction Override and Moderation in Large Language Models5 Feb 2024 0 repositories listed