Methods › General › Active Learning › BASE › Papers where code ran, page 2
Balanced Selection
BASE
Papers archive 2025-07-28
archive papers tagged: 5,784 · with a code link: 1,913 · where Syntology ran a sample: 621 (523 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (621 of 5,784 tagged: 523 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 2 of 7: papers 101 to 200 of the 621 tagged papers where Syntology ran at least one harvested sample (523 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration 25 Nov 2024 · 1 repository · arXiv:2411.16044Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Preference Optimization for Reasoning with Pseudo Feedback 25 Nov 2024 · 1 repository · arXiv:2411.16345Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Tulu 3: Pushing Frontiers in Open Language Model Post-Training 22 Nov 2024 · 1 repository · arXiv:2411.15124Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models 21 Nov 2024 · 1 repository · arXiv:2411.14432Syntology official (archive's flag): 3 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback 18 Nov 2024 · 1 repository · arXiv:2412.03578Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
LLaVA-CoT: Let Vision Language Models Reason Step-by-Step 15 Nov 2024 · 2 repositories · arXiv:2411.10440Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Dynamic Rewarding with Prompt Optimization Enables Tuning-free Self-Alignment of Language Models 13 Nov 2024 · 1 repository · arXiv:2411.08733Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
The Limited Impact of Medical Adaptation of Large Language and Vision-Language Models 13 Nov 2024 · 1 repository · arXiv:2411.08870Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
DPU: Dynamic Prototype Updating for Multimodal Out-of-Distribution Detection 12 Nov 2024 · 2 repositories · arXiv:2411.08227Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
The Surprising Effectiveness of Test-Time Training for Few-Shot Learning 11 Nov 2024 · 1 repository · arXiv:2411.07279Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)
-
Controllable Context Sensitivity and the Knob Behind It 11 Nov 2024 · 1 repository · arXiv:2411.07404Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Rethinking Bradley-Terry Models in Preference-Based Reward Modeling: Foundations, Theory, and Alternatives 7 Nov 2024 · 1 repository · arXiv:2411.04991Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Both Text and Images Leaked! A Systematic Analysis of Multimodal LLM Data Contamination 6 Nov 2024 · 1 repository · arXiv:2411.03823Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Medical Adaptation of Large Language and Vision-Language Models: Are We Making Progress? 6 Nov 2024 · 1 repository · arXiv:2411.04118Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding 6 Nov 2024 · 1 repository · arXiv:2411.04282Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Inference Optimal VLMs Need Fewer Visual Tokens and More Parameters 5 Nov 2024 · 1 repository · arXiv:2411.03312Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Zipfian Whitening 1 Nov 2024 · 1 repository · arXiv:2411.00680Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Can Language Models Perform Robust Reasoning in Chain-of-thought Prompting with Noisy Rationales? 31 Oct 2024 · 2 repositories · arXiv:2410.23856Syntology official (archive's flag): 7 ran · 10 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
ImOV3D: Learning Open-Vocabulary Point Clouds 3D Object Detection from Only 2D Images 31 Oct 2024 · 1 repository · arXiv:2410.24001Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Hidden Persuaders: LLMs' Political Leaning and Their Influence on Voters 31 Oct 2024 · 1 repository · arXiv:2410.24190Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)
-
SelfCodeAlign: Self-Alignment for Code Generation 31 Oct 2024 · 2 repositories · arXiv:2410.24198Syntology official (archive's flag): 9 ran · 30 ran (of which 3 constructed an object rather than computing a result; 22 with no instrument failure: 1 honoured, 0 violated, 21 with no contract checked; 8 where Syntology's instrument failed) · 7 unverified (of 37 harvested samples)
-
Consistency Diffusion Bridge Models 30 Oct 2024 · 2 repositories · arXiv:2410.22637Syntology 4 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
VPO: Leveraging the Number of Votes in Preference Optimization 30 Oct 2024 · 1 repository · arXiv:2410.22891Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
FlowLLM: Flow Matching for Material Generation with Large Language Models as Base Distributions 30 Oct 2024 · 1 repository · arXiv:2410.23405Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
PK-YOLO: Pretrained Knowledge Guided YOLO for Brain Tumor Detection in Multiplanar MRI Slices 29 Oct 2024 · 1 repository · arXiv:2410.21822Syntology official (archive's flag): 16 ran · 16 ran (of which 1 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 1 violated, 5 with no contract checked; 9 where Syntology's instrument failed) · 18 unverified (of 34 harvested samples) · 34 pointer-only (licence)
-
Retrieval-Retro: Retrieval-based Inorganic Retrosynthesis with Expert Knowledge 28 Oct 2024 · 1 repository · arXiv:2410.21341Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Absorb & Escape: Overcoming Single Model Limitations in Generating Genomic Sequences 28 Oct 2024 · 2 repositories · arXiv:2410.21345Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders 27 Oct 2024 · 1 repository · arXiv:2410.20526Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Practical Bayesian Algorithm Execution via Posterior Sampling 27 Oct 2024 · 1 repository · arXiv:2410.20596Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
OpenWebVoyager: Building Multimodal Web Agents via Iterative Real-World Exploration, Feedback and Optimization 25 Oct 2024 · 1 repository · arXiv:2410.19609Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
3D-Adapter: Geometry-Consistent Multi-View Diffusion for High-Quality 3D Generation 24 Oct 2024 · 1 repository · arXiv:2410.18974Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Conditional diffusions for amortized neural posterior estimation 24 Oct 2024 · 1 repository · arXiv:2410.19105Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Fast Graph Sharpness-Aware Minimization for Enhancing and Accelerating Few-Shot Node Classification 22 Oct 2024 · 1 repository · arXiv:2410.16845Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following 21 Oct 2024 · 1 repository · arXiv:2410.15553Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Gradient Rewiring for Editable Graph Neural Network Training 21 Oct 2024 · 1 repository · arXiv:2410.15556Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 8 pointer-only (licence)
-
IPO: Interpretable Prompt Optimization for Vision-Language Models 20 Oct 2024 · 1 repository · arXiv:2410.15397Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Evaluating Deep Unlearning in Large Language Models 19 Oct 2024 · 1 repository · arXiv:2410.15153Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
On the Role of Attention Heads in Large Language Model Safety 17 Oct 2024 · 1 repository · arXiv:2410.13708Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 8 where Syntology's instrument failed) · 8 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Expand and Compress: Exploring Tuning Principles for Continual Spatio-Temporal Graph Forecasting 16 Oct 2024 · 1 repository · arXiv:2410.12593Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Improving Instruction-Following in Language Models through Activation Steering 15 Oct 2024 · 0 repositories · arXiv:2410.12877Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
LoLCATs: On Low-Rank Linearizing of Large Language Models 14 Oct 2024 · 1 repository · arXiv:2410.10254Syntology official (archive's flag): 23 ran · 23 ran (of which 9 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 10 where Syntology's instrument failed) · 9 unverified (of 32 harvested samples)
-
Locking Down the Finetuned LLMs Safety 14 Oct 2024 · 1 repository · arXiv:2410.10343Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 6 where Syntology's instrument failed) · 9 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Towards Reliable Verification of Unauthorized Data Usage in Personalized Text-to-Image Diffusion Models 14 Oct 2024 · 1 repository · arXiv:2410.10437Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
KBLaM: Knowledge Base augmented Language Model 14 Oct 2024 · 1 repository · arXiv:2410.10450Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples)
-
SAMPa: Sharpness-aware Minimization Parallelized 14 Oct 2024 · 1 repository · arXiv:2410.10683Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Large-Scale 3D Medical Image Pre-training with Geometric Context Priors 13 Oct 2024 · 1 repository · arXiv:2410.09890Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Rethinking Data Selection at Scale: Random Selection is Almost All You Need 12 Oct 2024 · 1 repository · arXiv:2410.09335Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Generative Subgraph Retrieval for Knowledge Graph-Grounded Dialog Generation 12 Oct 2024 · 1 repository · arXiv:2410.09350Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation 12 Oct 2024 · 1 repository · arXiv:2410.09400Syntology official (archive's flag): 9 ran · 13 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 1 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 9 unverified (of 22 harvested samples) · 3 pointer-only (licence)
-
VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment 12 Oct 2024 · 0 repositories · arXiv:2410.09421Syntology 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Context-Aware Adapter Tuning for Few-Shot Relation Learning in Knowledge Graphs 11 Oct 2024 · 1 repository · arXiv:2410.09123Syntology official: no sample here; runs from other or unrecorded repositories · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code 10 Oct 2024 · 1 repository · arXiv:2410.08196Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
IterComp: Iterative Composition-Aware Feedback Learning from Model Gallery for Text-to-Image Generation 9 Oct 2024 · 2 repositories · arXiv:2410.07171Syntology official (archive's flag): 3 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 18 harvested samples) · 3 pointer-only (licence)
-
PAD: Personalized Alignment of LLMs at Decoding-Time 5 Oct 2024 · 0 repositories · arXiv:2410.04070Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
Multimodal Large Language Models for Inverse Molecular Design with Retrosynthetic Planning 5 Oct 2024 · 1 repository · arXiv:2410.04223Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Geometric Representation Condition Improves Equivariant Molecule Generation 4 Oct 2024 · 1 repository · arXiv:2410.03655Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 4 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Oscillatory State-Space Models 4 Oct 2024 · 1 repository · arXiv:2410.03943Syntology 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples)
-
LLM-TOPLA: Efficient LLM Ensemble by Maximising Diversity 4 Oct 2024 · 1 repository · arXiv:2410.03953Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Understanding and Mitigating Miscalibration in Prompt Tuning for Vision-Language Models 3 Oct 2024 · 1 repository · arXiv:2410.02681Syntology 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 6 pointer-only (licence)
-
HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly 3 Oct 2024 · 1 repository · arXiv:2410.02694Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
A Simple but Strong Baseline for Sounding Video Generation: Effective Adaptation of Audio and Video Diffusion Models for Joint Generation 26 Sep 2024 · 1 repository · arXiv:2409.17550Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 3 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Zero-Shot Detection of LLM-Generated Text using Token Cohesiveness 25 Sep 2024 · 1 repository · arXiv:2409.16914Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks 20 Sep 2024 · 1 repository · arXiv:2409.13203Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
A Controlled Study on Long Context Extension and Generalization in LLMs 18 Sep 2024 · 1 repository · arXiv:2409.12181Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
THaMES: An End-to-End Tool for Hallucination Mitigation and Evaluation in Large Language Models 17 Sep 2024 · 1 repository · arXiv:2409.11353Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Kolmogorov-Arnold Transformer 16 Sep 2024 · 1 repository · arXiv:2409.10594Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Operator Learning with Gaussian Processes 6 Sep 2024 · 1 repository · arXiv:2409.04538Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Sparse Rewards Can Self-Train Dialogue Agents 6 Sep 2024 · 1 repository · arXiv:2409.04617Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Language Adaptation on a Tight Academic Compute Budget: Tokenizer Swapping Works and Pure bfloat16 Is Enough 28 Aug 2024 · 1 repository · arXiv:2408.15793Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 19 harvested samples)
-
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options 27 Aug 2024 · 1 repository · arXiv:2409.00113Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AgentMove: Predicting Human Mobility Anywhere Using Large Language Model based Agentic Framework 26 Aug 2024 · 1 repository · arXiv:2408.13986Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Learning Unknowns from Unknowns: Diversified Negative Prototypes Generator for Few-Shot Open-Set Recognition 23 Aug 2024 · 1 repository · arXiv:2408.13373Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Critique-out-Loud Reward Models 21 Aug 2024 · 1 repository · arXiv:2408.11791Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Selective Prompt Anchoring for Code Generation 17 Aug 2024 · 1 repository · arXiv:2408.09121Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
xGen-MM (BLIP-3): A Family of Open Large Multimodal Models 16 Aug 2024 · 1 repository · arXiv:2408.08872Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
I-SHEEP: Self-Alignment of LLM from Scratch through an Iterative Self-Enhancement Paradigm 15 Aug 2024 · 1 repository · arXiv:2408.08072Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
ControlNeXt: Powerful and Efficient Control for Image and Video Generation 12 Aug 2024 · 1 repository · arXiv:2408.06070Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine 6 Aug 2024 · 1 repository · arXiv:2408.02900Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters 6 Aug 2024 · 3 repositories · arXiv:2408.03314Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 8 pointer-only (licence)
-
GPUDrive: Data-driven, multi-agent driving simulation at 1 million FPS 2 Aug 2024 · 1 repository · arXiv:2408.01584Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Do LLMs Really Adapt to Domains? An Ontology Learning Perspective 29 Jul 2024 · 1 repository · arXiv:2407.19998Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Self-Training with Direct Preference Optimization Improves Chain-of-Thought Reasoning 25 Jul 2024 · 1 repository · arXiv:2407.18248Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Large Language Model for Verilog Generation with Code-Structure-Guided Reinforcement Learning 21 Jul 2024 · 1 repository · arXiv:2407.18271Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
System-1.x: Learning to Balance Fast and Slow Planning with Language Models 19 Jul 2024 · 1 repository · arXiv:2407.14414Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Catastrophic Goodhart: regularizing RLHF with KL divergence does not mitigate heavy-tailed reward misspecification 19 Jul 2024 · 1 repository · arXiv:2407.14503Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Evaluating language models as risk scores 19 Jul 2024 · 1 repository · arXiv:2407.14614Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples)
-
Scalable Exploration via Ensemble++ 18 Jul 2024 · 2 repositories · arXiv:2407.13195Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
AgentPoison: Red-teaming LLM Agents via Poisoning Memory or Knowledge Bases 17 Jul 2024 · 1 repository · arXiv:2407.12784Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 1 honoured, 0 violated, 14 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 1 pointer-only (licence)
-
AdaLog: Post-Training Quantization for Vision Transformers with Adaptive Logarithm Quantizer 17 Jul 2024 · 1 repository · arXiv:2407.12951Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction 16 Jul 2024 · 1 repository · arXiv:2407.11335Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference 16 Jul 2024 · 2 repositories · arXiv:2407.11550Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
XEdgeAI: A Human-centered Industrial Inspection Framework with Data-centric Explainable Edge AI Approach 16 Jul 2024 · 1 repository · arXiv:2407.11771Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Enhancing Retrieval and Managing Retrieval: A Four-Module Synergy for Improved Quality and Efficiency in RAG Systems 15 Jul 2024 · 1 repository · arXiv:2407.10670Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
PaliGemma: A versatile 3B VLM for transfer 10 Jul 2024 · 1 repository · arXiv:2407.07726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
EfficientQAT: Efficient Quantization-Aware Training for Large Language Models 10 Jul 2024 · 1 repository · arXiv:2407.11062Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct 8 Jul 2024 · 1 repository · arXiv:2407.05700Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Mamba-FSCIL: Dynamic Adaptation with Selective State Space Model for Few-Shot Class-Incremental Learning 8 Jul 2024 · 1 repository · arXiv:2407.06136Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs 5 Jul 2024 · 1 repository · arXiv:2407.04694Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Are Large Language Models Consistent over Value-laden Questions? 3 Jul 2024 · 1 repository · arXiv:2407.02996Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
HRSAM: Efficient Interactive Segmentation in High-Resolution Images 2 Jul 2024 · 1 repository · arXiv:2407.02109Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)