Methods › General › Active Learning › BASE › Papers where code ran, page 1
Balanced Selection
BASE
Papers archive 2025-07-28
archive papers tagged: 5,784 · with a code link: 1,913 · where Syntology ran a sample: 621 (523 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (621 of 5,784 tagged: 523 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 1 of 7: papers 1 to 100 of the 621 tagged papers where Syntology ran at least one harvested sample (523 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CultureCLIP: Empowering CLIP with Cultural Awareness through Synthetic Images and Contextualized Captions 8 Jul 2025 · 1 repository · arXiv:2507.06210Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving 8 Jul 2025 · 1 repository · arXiv:2507.06229Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training 7 Jul 2025 · 0 repositories · arXiv:2507.05386Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling 25 Jun 2025 · 1 repository · arXiv:2506.20512Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 17 harvested samples) · 4 pointer-only (licence)
-
LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning 23 Jun 2025 · 1 repository · arXiv:2506.18841Syntology 12 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
DiffO: Single-step Diffusion for Image Compression at Ultra-Low Bitrates 19 Jun 2025 · 1 repository · arXiv:2506.16572Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph Refinement 18 Jun 2025 · 2 repositories · arXiv:2506.15583Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 1 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
SoundMind: RL-Incentivized Logic Reasoning for Audio-Language Models 15 Jun 2025 · 1 repository · arXiv:2506.12935Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples) · 4 pointer-only (licence)
-
ClimateChat: Designing Data and Methods for Instruction Tuning LLMs to Answer Climate Change Queries 12 Jun 2025 · 1 repository · arXiv:2506.13796Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Vision Transformers Don't Need Trained Registers 9 Jun 2025 · 1 repository · arXiv:2506.08010Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Play to Generalize: Learning to Reason Through Game Play 9 Jun 2025 · 1 repository · arXiv:2506.08011Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples)
-
Homogeneous Keys, Heterogeneous Values: Exploiting Local KV Cache Asymmetry for Long-Context LLMs 4 Jun 2025 · 0 repositories · arXiv:2506.05410Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples)
-
Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning 3 Jun 2025 · 2 repositories · arXiv:2506.03136Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
The Surprising Effectiveness of Negative Reinforcement in LLM Reasoning 2 Jun 2025 · 1 repository · arXiv:2506.01347Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models 30 May 2025 · 1 repository · arXiv:2505.24864Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Verify-in-the-Graph: Entity Disambiguation Enhancement for Complex Claim Verification with Interactive Graph Representation 29 May 2025 · 0 repositories · arXiv:2505.22993Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation 28 May 2025 · 1 repository · arXiv:2505.22647Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Update Your Transformer to the Latest Release: Re-Basin of Task Vectors 28 May 2025 · 1 repository · arXiv:2505.22697Syntology official (archive's flag): 1 ran · 5 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 3 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Concise Reasoning, Big Gains: Pruning Long Reasoning Trace with Difficulty-Aware Prompting 26 May 2025 · 1 repository · arXiv:2505.19716Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 9 pointer-only (licence)
-
FinLoRA: Benchmarking LoRA Methods for Fine-Tuning LLMs on Financial Datasets 26 May 2025 · 2 repositories · arXiv:2505.19819Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Behavior Injection: Preparing Language Models for Reinforcement Learning 25 May 2025 · 1 repository · arXiv:2505.18917Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Diffusion Blend: Inference-Time Multi-Preference Alignment for Diffusion Models 24 May 2025 · 1 repository · arXiv:2505.18547Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
UFT: Unifying Supervised and Reinforcement Fine-Tuning 22 May 2025 · 1 repository · arXiv:2505.16984Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
JanusDNA: A Powerful Bi-directional Hybrid DNA Foundation Model 22 May 2025 · 1 repository · arXiv:2505.17257Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Synthetic Data RL: Task Definition Is All You Need 18 May 2025 · 1 repository · arXiv:2505.17063Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
ADHMR: Aligning Diffusion-based Human Mesh Recovery via Direct Preference Optimization 15 May 2025 · 1 repository · arXiv:2505.10250Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving 12 May 2025 · 2 repositories · arXiv:2505.07773Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ZeroSearch: Incentivize the Search Capability of LLMs without Searching 7 May 2025 · 1 repository · arXiv:2505.04588Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 8 unverified (of 15 harvested samples)
-
Don't be lazy: CompleteP enables compute-efficient deep transformers 2 May 2025 · 1 repository · arXiv:2505.01618Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LLM-Empowered Embodied Agent for Memory-Augmented Task Planning in Household Robotics 30 Apr 2025 · 1 repository · arXiv:2504.21716Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Reinforcement Learning for Reasoning in Large Language Models with One Training Example 29 Apr 2025 · 1 repository · arXiv:2504.20571Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 9 pointer-only (licence)
-
AegisLLM: Scaling Agentic Systems for Self-Reflective Defense in LLM Security 29 Apr 2025 · 1 repository · arXiv:2504.20965Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Tina: Tiny Reasoning Models via LoRA 22 Apr 2025 · 1 repository · arXiv:2504.15777Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Physics-Informed Inference Time Scaling via Simulation-Calibrated Scientific Machine Learning 22 Apr 2025 · 1 repository · arXiv:2504.16172Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluating Judges as Evaluators: The JETTS Benchmark of LLM-as-Judges as Test-Time Scaling Evaluators 21 Apr 2025 · 1 repository · arXiv:2504.15253Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning 21 Apr 2025 · 1 repository · arXiv:2504.15275Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
ToolRL: Reward is All Tool Learning Needs 16 Apr 2025 · 1 repository · arXiv:2504.13958Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Adaptive Decision Boundary for Few-Shot Class-Incremental Learning 15 Apr 2025 · 1 repository · arXiv:2504.10976Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
GRPO-LEAD: A Difficulty-Aware Reinforcement Learning Approach for Concise Mathematical Reasoning in Language Models 13 Apr 2025 · 1 repository · arXiv:2504.09696Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
SPICE: A Synergistic, Precise, Iterative, and Customizable Image Editing Workflow 13 Apr 2025 · 1 repository · arXiv:2504.09697Syntology official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning 10 Apr 2025 · 1 repository · arXiv:2504.07891Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization 8 Apr 2025 · 1 repository · arXiv:2504.05812Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Robustly identifying concepts introduced during chat fine-tuning using crosscoders 3 Apr 2025 · 1 repository · arXiv:2504.02922Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples)
-
Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model 31 Mar 2025 · 1 repository · arXiv:2503.24290Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Exploring the Effect of Reinforcement Learning on Video Understanding: Insights from SEED-Bench-R1 31 Mar 2025 · 1 repository · arXiv:2503.24376Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 3 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning 27 Mar 2025 · 1 repository · arXiv:2503.21620Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 3 pointer-only (licence)
-
Open Deep Search: Democratizing Search with Open-source Reasoning Agents 26 Mar 2025 · 1 repository · arXiv:2503.20201Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Understanding R1-Zero-Like Training: A Critical Perspective 26 Mar 2025 · 1 repository · arXiv:2503.20783Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Exploring CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation 26 Mar 2025 · 1 repository · arXiv:2503.20826Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Cross-Tokenizer Distillation via Approximate Likelihood Matching 25 Mar 2025 · 1 repository · arXiv:2503.20083Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild 24 Mar 2025 · 1 repository · arXiv:2503.18892Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity 20 Mar 2025 · 1 repository · arXiv:2503.16418Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
High Temporal Consistency through Semantic Similarity Propagation in Semi-Supervised Video Semantic Segmentation for Autonomous Flight 19 Mar 2025 · 1 repository · arXiv:2503.15676Syntology official (archive's flag): 25 ran · 25 ran (of which 15 constructed an object rather than computing a result; 20 with no instrument failure: 0 honoured, 0 violated, 20 with no contract checked; 5 where Syntology's instrument failed) · 14 unverified (of 39 harvested samples) · 39 pointer-only (licence)
-
DPC: Dual-Prompt Collaboration for Tuning Vision-Language Models 17 Mar 2025 · 1 repository · arXiv:2503.13443Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
MOS: Modeling Object-Scene Associations in Generalized Category Discovery 15 Mar 2025 · 1 repository · arXiv:2503.12035Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Reflect-DiT: Inference-Time Scaling for Text-to-Image Diffusion Transformers via In-Context Reflection 15 Mar 2025 · 1 repository · arXiv:2503.12271Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash Equilibrium 14 Mar 2025 · 1 repository · arXiv:2503.10990Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Rethinking Few-Shot Adaptation of Vision-Language Models in Two Stages 14 Mar 2025 · 1 repository · arXiv:2503.11609Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 8 unverified (of 12 harvested samples)
-
Towards Interpretable Protein Structure Prediction with Sparse Autoencoders 11 Mar 2025 · 1 repository · arXiv:2503.08764Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Task Vector Quantization for Memory-Efficient Model Merging 10 Mar 2025 · 1 repository · arXiv:2503.06921Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning 10 Mar 2025 · 1 repository · arXiv:2503.07459Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 18 with no instrument failure: 0 honoured, 0 violated, 18 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 23 harvested samples) · 4 pointer-only (licence)
-
R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model 7 Mar 2025 · 1 repository · arXiv:2503.05132Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning 7 Mar 2025 · 5 repositories · arXiv:2503.05592Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Activation Space Interventions Can Be Transferred Between Large Language Models 6 Mar 2025 · 1 repository · arXiv:2503.04429Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think 2 Mar 2025 · 1 repository · arXiv:2503.00948Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 4 pointer-only (licence)
-
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success 27 Feb 2025 · 1 repository · arXiv:2502.19645Syntology 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Reward Shaping to Mitigate Reward Hacking in RLHF 26 Feb 2025 · 1 repository · arXiv:2502.18770Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
AgentRM: Enhancing Agent Generalization with Reward Modeling 25 Feb 2025 · 0 repositories · arXiv:2502.18407Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
JUREX-4E: Juridical Expert-Annotated Four-Element Knowledge Base for Legal Reasoning 24 Feb 2025 · 1 repository · arXiv:2502.17166Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples)
-
Delta Decompression for MoE-based LLMs Compression 24 Feb 2025 · 1 repository · arXiv:2502.17298Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Function-Space Learning Rates 24 Feb 2025 · 1 repository · arXiv:2502.17405Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Moving Beyond Medical Exam Questions: A Clinician-Annotated Dataset of Real-World Tasks and Ambiguity in Mental Healthcare 22 Feb 2025 · 1 repository · arXiv:2502.16051Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples)
-
Hyperspherical Normalization for Scalable Deep Reinforcement Learning 21 Feb 2025 · 0 repositories · arXiv:2502.15280Syntology 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
TESS 2: A Large-Scale Generalist Diffusion Language Model 19 Feb 2025 · 1 repository · arXiv:2502.13917Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
S²R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning 18 Feb 2025 · 1 repository · arXiv:2502.12853Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
GENERator: A Long-Context Generative Genomic Foundation Model 11 Feb 2025 · 1 repository · arXiv:2502.07272Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Demystifying Long Chain-of-Thought Reasoning in LLMs 5 Feb 2025 · 1 repository · arXiv:2502.03373Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
AgentBreeder: Mitigating the AI Safety Impact of Multi-Agent Scaffolds via Self-Improvement 2 Feb 2025 · 1 repository · arXiv:2502.00757Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search 31 Jan 2025 · 1 repository · arXiv:2501.18922Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 8 harvested samples)
-
Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate 29 Jan 2025 · 1 repository · arXiv:2501.17703Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
SAeUron: Interpretable Concept Unlearning in Diffusion Models with Sparse Autoencoders 29 Jan 2025 · 1 repository · arXiv:2501.18052Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
CHiP: Cross-modal Hierarchical Direct Preference Optimization for Multimodal LLMs 28 Jan 2025 · 1 repository · arXiv:2501.16629Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Video Depth Anything: Consistent Depth Estimation for Super-Long Videos 21 Jan 2025 · 1 repository · arXiv:2501.12375Syntology 10 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Self-supervised Transformation Learning for Equivariant Representations 15 Jan 2025 · 1 repository · arXiv:2501.08712Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Vchitect-2.0: Parallel Transformer for Scaling Up Video Diffusion Models 14 Jan 2025 · 1 repository · arXiv:2501.08453Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Enhancing Retrieval-Augmented Generation: A Study of Best Practices 13 Jan 2025 · 1 repository · arXiv:2501.07391Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing Unsupervised Graph Few-shot Learning via Set Functions and Optimal Transport 10 Jan 2025 · 1 repository · arXiv:2501.05635Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
EpiCoder: Encompassing Diversity and Complexity in Code Generation 8 Jan 2025 · 0 repositories · arXiv:2501.04694Syntology 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 14 pointer-only (licence)
-
2 OLMo 2 Furious 31 Dec 2024 · 3 repositories · arXiv:2501.00656Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 21 harvested samples) · 2 pointer-only (licence)
-
HumanEval Pro and MBPP Pro: Evaluating Large Language Models on Self-invoking Code Generation 30 Dec 2024 · 1 repository · arXiv:2412.21199Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
CiteBART: Learning to Generate Citations for Local Citation Recommendation 23 Dec 2024 · 1 repository · arXiv:2412.17534Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples)
-
YuLan-Mini: An Open Data-efficient Language Model 23 Dec 2024 · 2 repositories · arXiv:2412.17743Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Qwen2.5 Technical Report 19 Dec 2024 · 6 repositories · arXiv:2412.15115Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Preference-Oriented Supervised Fine-Tuning: Favoring Target Model Over Aligned Large Language Models 17 Dec 2024 · 1 repository · arXiv:2412.12865Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Vocabulary Expansion of Chat Models with Unlabeled Target Language Data 16 Dec 2024 · 1 repository · arXiv:2412.11704Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Memory Layers at Scale 12 Dec 2024 · 1 repository · arXiv:2412.09764Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
PaliGemma 2: A Family of Versatile VLMs for Transfer 4 Dec 2024 · 1 repository · arXiv:2412.03555Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
MV-Adapter: Multi-view Consistent Image Generation Made Easy 4 Dec 2024 · 1 repository · arXiv:2412.03632Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment 27 Nov 2024 · 1 repository · arXiv:2411.18688Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
GrokFormer: Graph Fourier Kolmogorov-Arnold Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17296Syntology official (archive's flag): 4 ran · 4 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)