Methods › Natural Language Processing › Transformers › Focus › Papers where code ran, page 4
Focus
Papers archive 2025-07-28
archive papers tagged: 15,340 · with a code link: 5,193 · where Syntology ran a sample: 1,419 (1,210 with a run with no instrument failure, 209 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,419 of 15,340 tagged: 1,210 with a run with no instrument failure, 209 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 4 of 15: papers 301 to 400 of the 1,419 tagged papers where Syntology ran at least one harvested sample (1,210 with a run with no instrument failure, 209 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ODRL: A Benchmark for Off-Dynamics Reinforcement Learning 28 Oct 2024 · 1 repository · arXiv:2410.20750Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 4 pointer-only (licence)
-
RecFlow: An Industrial Full Flow Recommendation Dataset 28 Oct 2024 · 1 repository · arXiv:2410.20868Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
ProtSCAPE: Mapping the landscape of protein conformations in molecular dynamics 27 Oct 2024 · 1 repository · arXiv:2410.20317Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A Cosmic-Scale Benchmark for Symmetry-Preserving Data Processing 27 Oct 2024 · 1 repository · arXiv:2410.20516Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Bongard in Wonderland: Visual Puzzles that Still Make AI Go Mad? 25 Oct 2024 · 1 repository · arXiv:2410.19546Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
What If the Input is Expanded in OOD Detection? 24 Oct 2024 · 1 repository · arXiv:2410.18472Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 2 violated, 7 with no contract checked; 7 where Syntology's instrument failed) · 4 unverified (of 20 harvested samples) · 20 pointer-only (licence)
-
KVSharer: Efficient Inference via Layer-Wise Dissimilar KV Cache Sharing 24 Oct 2024 · 1 repository · arXiv:2410.18517Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Retrieval-Augmented Diffusion Models for Time Series Forecasting 24 Oct 2024 · 1 repository · arXiv:2410.18712Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Should We Really Edit Language Models? On the Evaluation of Edited Language Models 24 Oct 2024 · 1 repository · arXiv:2410.18785Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models 23 Oct 2024 · 1 repository · arXiv:2410.17578Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Graphusion: A RAG Framework for Knowledge Graph Construction with a Global Perspective 23 Oct 2024 · 1 repository · arXiv:2410.17600Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Blendify -- Python rendering framework for Blender 23 Oct 2024 · 1 repository · arXiv:2410.17858Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Trustworthy Alignment of Retrieval-Augmented Large Language Models via Reinforcement Learning 22 Oct 2024 · 1 repository · arXiv:2410.16843Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
VoiceBench: Benchmarking LLM-Based Voice Assistants 22 Oct 2024 · 1 repository · arXiv:2410.17196Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following 21 Oct 2024 · 1 repository · arXiv:2410.15553Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
TALoS: Enhancing Semantic Scene Completion via Test-time Adaptation on the Line of Sight 21 Oct 2024 · 1 repository · arXiv:2410.15674Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 2 honoured, 1 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
On conditional diffusion models for PDE simulations 21 Oct 2024 · 1 repository · arXiv:2410.16415Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Are LLMs Good Zero-Shot Fallacy Classifiers? 19 Oct 2024 · 1 repository · arXiv:2410.15050Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
MultiChartQA: Benchmarking Vision-Language Models on Multi-Chart Problems 18 Oct 2024 · 1 repository · arXiv:2410.14179Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Toward Generalizing Visual Brain Decoding to Unseen Subjects 18 Oct 2024 · 1 repository · arXiv:2410.14445Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Cross-Lingual Auto Evaluation for Assessing Multilingual LLMs 17 Oct 2024 · 1 repository · arXiv:2410.13394Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
LoLDU: Low-Rank Adaptation via Lower-Diag-Upper Decomposition for Parameter-Efficient Fine-Tuning 17 Oct 2024 · 1 repository · arXiv:2410.13618Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Analyzing Deep Transformer Models for Time Series Forecasting via Manifold Learning 17 Oct 2024 · 1 repository · arXiv:2410.13792Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Mitigating the Backdoor Effect for Multi-Task Model Merging via Safety-Aware Subspace 17 Oct 2024 · 1 repository · arXiv:2410.13910Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
CATCH: Channel-Aware multivariate Time Series Anomaly Detection via Frequency Patching 16 Oct 2024 · 1 repository · arXiv:2410.12261Syntology official (archive's flag): 11 ran · 11 ran (of which 8 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 8 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
DAT: Improving Adversarial Robustness via Generative Amplitude Mix-up in Frequency Domain 16 Oct 2024 · 1 repository · arXiv:2410.12307Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Stylistic Multi-Task Analysis of Ukiyo-e Woodblock Prints 16 Oct 2024 · 1 repository · arXiv:2410.12379Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
JudgeBench: A Benchmark for Evaluating LLM-based Judges 16 Oct 2024 · 1 repository · arXiv:2410.12784Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Dual Prototype Evolving for Test-Time Generalization of Vision-Language Models 16 Oct 2024 · 1 repository · arXiv:2410.12790Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
MSc-SQL: Multi-Sample Critiquing Small Language Models For Text-To-SQL Translation 16 Oct 2024 · 1 repository · arXiv:2410.12916Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
Credal Two-Sample Tests of Epistemic Uncertainty 16 Oct 2024 · 1 repository · arXiv:2410.12921Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples)
-
Hypothesis Testing the Circuit Hypothesis in LLMs 16 Oct 2024 · 1 repository · arXiv:2410.13032Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
A Complete Decomposition of KL Error using Refined Information and Mode Interaction Selection 15 Oct 2024 · 1 repository · arXiv:2410.11964Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
MuseTalk: Real-Time High-Fidelity Video Dubbing via Spatio-Temporal Sampling 14 Oct 2024 · 1 repository · arXiv:2410.10122Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
EffiCoder: Enhancing Code Generation in Large Language Models through Efficiency-Aware Fine-tuning 14 Oct 2024 · 2 repositories · arXiv:2410.10209Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 3 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 15 harvested samples)
-
Evaluating Semantic Variation in Text-to-Image Synthesis: A Causal Perspective 14 Oct 2024 · 1 repository · arXiv:2410.10291Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Will LLMs Replace the Encoder-Only Models in Temporal Relation Classification? 14 Oct 2024 · 1 repository · arXiv:2410.10476Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
A Practical Approach to Causal Inference over Time 14 Oct 2024 · 1 repository · arXiv:2410.10502Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
Get Rid of Isolation: A Continuous Multi-task Spatio-Temporal Learning Framework 14 Oct 2024 · 1 repository · arXiv:2410.10524Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
UniMatch V2: Pushing the Limit of Semi-Supervised Semantic Segmentation 14 Oct 2024 · 1 repository · arXiv:2410.10777Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads 14 Oct 2024 · 2 repositories · arXiv:2410.10819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Varying Shades of Wrong: Aligning LLMs with Wrong Answers Only 14 Oct 2024 · 1 repository · arXiv:2410.11055Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MIRAGE: Evaluating and Explaining Inductive Reasoning Process in Language Models 12 Oct 2024 · 0 repositories · arXiv:2410.09542Syntology 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Reinforcement Learning for Control of Non-Markovian Cellular Population Dynamics 11 Oct 2024 · 1 repository · arXiv:2410.08439Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Refusal-Trained LLMs Are Easily Jailbroken As Browser Agents 11 Oct 2024 · 1 repository · arXiv:2410.13886Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Automatic Curriculum Expert Iteration for Reliable LLM Reasoning 10 Oct 2024 · 1 repository · arXiv:2410.07627Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MACPO: Weak-to-Strong Alignment via Multi-Agent Contrastive Preference Optimization 10 Oct 2024 · 0 repositories · arXiv:2410.07672Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Relational Diffusion Distillation for Efficient Image Generation 10 Oct 2024 · 1 repository · arXiv:2410.07679Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Minority-Focused Text-to-Image Generation via Prompt Optimization 10 Oct 2024 · 1 repository · arXiv:2410.07838Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Benchmarking Agentic Workflow Generation 10 Oct 2024 · 1 repository · arXiv:2410.07869Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
COMPL-AI Framework: A Technical Interpretation and LLM Benchmarking Suite for the EU Artificial Intelligence Act 10 Oct 2024 · 1 repository · arXiv:2410.07959Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Efficiently Learning at Test-Time: Active Fine-Tuning of LLMs 10 Oct 2024 · 1 repository · arXiv:2410.08020Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Reward-Augmented Data Enhances Direct Preference Alignment of LLMs 10 Oct 2024 · 1 repository · arXiv:2410.08067Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Robust AI-Generated Text Detection by Restricted Embeddings 10 Oct 2024 · 1 repository · arXiv:2410.08113Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
VibeCheck: Discover and Quantify Qualitative Differences in Large Language Models 10 Oct 2024 · 1 repository · arXiv:2410.12851Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Enhancing Legal Case Retrieval via Scaling High-quality Synthetic Query-Candidate Pairs 9 Oct 2024 · 1 repository · arXiv:2410.06581Syntology official (archive's flag): 3 ran · 14 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
Task-oriented Time Series Imputation Evaluation via Generalized Representers 9 Oct 2024 · 1 repository · arXiv:2410.06652Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Mitigating the Language Mismatch and Repetition Issues in LLM-based Machine Translation via Model Editing 9 Oct 2024 · 1 repository · arXiv:2410.07054Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Parameter Efficient Fine-tuning via Explained Variance Adaptation 9 Oct 2024 · 2 repositories · arXiv:2410.07170Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts 9 Oct 2024 · 3 repositories · arXiv:2410.07348Syntology official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
LLM Embeddings Improve Test-time Adaptation to Tabular Y|X-Shifts 9 Oct 2024 · 1 repository · arXiv:2410.07395Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
T2V-Turbo-v2: Enhancing Video Generation Model Post-Training through Data, Reward, and Conditional Guidance Design 8 Oct 2024 · 1 repository · arXiv:2410.05677Syntology 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Treat Visual Tokens as Text? But Your MLLM Only Needs Fewer Efforts to See 8 Oct 2024 · 1 repository · arXiv:2410.06169Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Entering Real Social World! Benchmarking the Social Intelligence of Large Language Models from a First-person Perspective 8 Oct 2024 · 1 repository · arXiv:2410.06195Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Neural Fourier Modelling: A Highly Compact Approach to Time-Series Analysis 7 Oct 2024 · 1 repository · arXiv:2410.04703Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Fast Training of Sinusoidal Neural Fields via Scaling Initialization 7 Oct 2024 · 1 repository · arXiv:2410.04779Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Mitigating Modality Prior-Induced Hallucinations in Multimodal Large Language Models via Deciphering Attention Causality 7 Oct 2024 · 1 repository · arXiv:2410.04780Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Unitary convolutions for learning on graphs and groups 7 Oct 2024 · 1 repository · arXiv:2410.05499Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Fine-Grained Prediction of Reading Comprehension from Eye Movements 6 Oct 2024 · 1 repository · arXiv:2410.04484Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Neuron-Level Sequential Editing for Large Language Models 5 Oct 2024 · 2 repositories · arXiv:2410.04045Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
LongGenBench: Long-context Generation Benchmark 5 Oct 2024 · 1 repository · arXiv:2410.04199Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
X-ALMA: Plug & Play Modules and Adaptive Rejection for Quality Translation at Scale 4 Oct 2024 · 0 repositories · arXiv:2410.03115Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 9 harvested samples)
-
GraphCroc: Cross-Correlation Autoencoder for Graph Structural Reconstruction 4 Oct 2024 · 1 repository · arXiv:2410.03396Syntology official (archive's flag): 11 ran · 11 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples)
-
PersonalSum: A User-Subjective Guided Personalized Summarization Dataset for Large Language Models 4 Oct 2024 · 1 repository · arXiv:2410.03905Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Retention-Centric Framework for Continual Learning with Guaranteed Model Developmental Safety 4 Oct 2024 · 1 repository · arXiv:2410.03955Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Correlation and Navigation in the Vocabulary Key Representation Space of Language Models 3 Oct 2024 · 1 repository · arXiv:2410.02284Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation 3 Oct 2024 · 1 repository · arXiv:2410.02369Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Towards Comprehensive Detection of Chinese Harmful Memes 3 Oct 2024 · 1 repository · arXiv:2410.02378Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
DivScene: Benchmarking LVLMs for Object Navigation with Diverse Scenes and Objects 3 Oct 2024 · 1 repository · arXiv:2410.02730Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 15 harvested samples)
-
Unveiling Language Skills via Path-Level Circuit Discovery 2 Oct 2024 · 1 repository · arXiv:2410.01334Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Codev-Bench: How Do LLMs Understand Developer-Centric Code Completion? 2 Oct 2024 · 1 repository · arXiv:2410.01353Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Selective Aggregation for Low-Rank Adaptation in Federated Learning 2 Oct 2024 · 1 repository · arXiv:2410.01463Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Positional Attention: Expressivity and Learnability of Algorithmic Computation 2 Oct 2024 · 1 repository · arXiv:2410.01686Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
An Introduction to Deep Survival Analysis Models for Predicting Time-to-Event Outcomes 1 Oct 2024 · 2 repositories · arXiv:2410.01086Syntology official (archive's flag): 20 ran · 20 ran (of which 0 constructed an object rather than computing a result; 18 with no instrument failure: 4 honoured, 0 violated, 14 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Learning to Obstruct Few-Shot Image Classification over Restricted Classes 28 Sep 2024 · 1 repository · arXiv:2409.19210Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Automated conjecturing in mathematics with TxGraffiti 28 Sep 2024 · 1 repository · arXiv:2409.19379Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
Towards Croppable Implicit Neural Representations 28 Sep 2024 · 1 repository · arXiv:2409.19472Syntology official (archive's flag): 13 ran · 13 ran (of which 1 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 22 harvested samples)
-
Do LLMs suffer from Multi-Party Hangover? A Diagnostic Approach to Addressee Recognition and Response Selection in Conversations 27 Sep 2024 · 1 repository · arXiv:2409.18602Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 7 honoured, 1 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Rethinking the Power of Timestamps for Robust Time Series Forecasting: A Global-Local Fusion Perspective 27 Sep 2024 · 1 repository · arXiv:2409.18696Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
HR-Extreme: A High-Resolution Dataset for Extreme Weather Forecasting 27 Sep 2024 · 1 repository · arXiv:2409.18885Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Advancing Open-Set Domain Generalization Using Evidential Bi-Level Hardest Domain Scheduler 26 Sep 2024 · 1 repository · arXiv:2409.17555Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Extracting Affect Aggregates from Longitudinal Social Media Data with Temporal Adapters for Large Language Models 26 Sep 2024 · 1 repository · arXiv:2409.17990Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Holistic Automated Red Teaming for Large Language Models through Top-Down Test Case Generation and Multi-turn Interaction 25 Sep 2024 · 1 repository · arXiv:2409.16783Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Plurals: A System for Guiding LLMs Via Simulated Social Ensembles 25 Sep 2024 · 1 repository · arXiv:2409.17213Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Beyond Redundancy: Information-aware Unsupervised Multiplex Graph Structure Learning 25 Sep 2024 · 1 repository · arXiv:2409.17386Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Looped Transformers for Length Generalization 24 Sep 2024 · 1 repository · arXiv:2409.15647Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
M²PT: Multimodal Prompt Tuning for Zero-shot Instruction Learning 24 Sep 2024 · 2 repositories · arXiv:2409.15657Syntology official (archive's flag): 16 ran · 16 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 10 unverified (of 26 harvested samples) · 26 pointer-only (licence)
-
MOSS: Enabling Code-Driven Evolution and Context Management for AI Agents 24 Sep 2024 · 1 repository · arXiv:2409.16120Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Towards Representation Learning for Weighting Problems in Design-Based Causal Inference 24 Sep 2024 · 1 repository · arXiv:2409.16407Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
ChemEval: A Comprehensive Multi-Level Chemical Evaluation for Large Language Models 21 Sep 2024 · 1 repository · arXiv:2409.13989Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 0 violated, 17 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 18 harvested samples) · 18 pointer-only (licence)