Methods › General › Sparsity › SET › Papers where code ran, page 6
Sparse Evolutionary Training
SET
Papers archive 2025-07-28
archive papers tagged: 13,419 · with a code link: 4,158 · where Syntology ran a sample: 1,186 (1,038 with a run with no instrument failure, 148 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,186 of 13,419 tagged: 1,038 with a run with no instrument failure, 148 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 6 of 12: papers 501 to 600 of the 1,186 tagged papers where Syntology ran at least one harvested sample (1,038 with a run with no instrument failure, 148 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
One Prompt is not Enough: Automated Construction of a Mixture-of-Expert Prompts 28 Jun 2024 · 1 repository · arXiv:2407.00256Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Correspondence-Free Non-Rigid Point Set Registration Using Unsupervised Clustering Analysis 27 Jun 2024 · 2 repositories · arXiv:2406.18817Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Compositional Image Decomposition with Diffusion Models 27 Jun 2024 · 0 repositories · arXiv:2406.19298Syntology 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 1 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
LiveBench: A Challenging, Contamination-Limited LLM Benchmark 27 Jun 2024 · 1 repository · arXiv:2406.19314Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
DICE: End-to-end Deformation Capture of Hand-Face Interactions from a Single Image 26 Jun 2024 · 0 repositories · arXiv:2406.17988Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Selective Prompting Tuning for Personalized Conversations with LLMs 26 Jun 2024 · 1 repository · arXiv:2406.18187Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
AlphaForge: A Framework to Mine and Dynamically Combine Formulaic Alpha Factors 26 Jun 2024 · 1 repository · arXiv:2406.18394Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Detecting Brittle Decisions for Free: Leveraging Margin Consistency in Deep Robust Classifiers 26 Jun 2024 · 1 repository · arXiv:2406.18451Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs 26 Jun 2024 · 4 repositories · arXiv:2406.18495Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems 26 Jun 2024 · 1 repository · arXiv:2406.18747Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
VarBench: Robust Language Model Benchmarking Through Dynamic Variable Perturbation 25 Jun 2024 · 1 repository · arXiv:2406.17681Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
BioTrove: A Large Curated Image Dataset Enabling AI for Biodiversity 25 Jun 2024 · 2 repositories · arXiv:2406.17720Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Recite, Reconstruct, Recollect: Memorization in LMs as a Multifaceted Phenomenon 25 Jun 2024 · 1 repository · arXiv:2406.17746Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Confidence Regulation Neurons in Language Models 24 Jun 2024 · 1 repository · arXiv:2406.16254Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Token-based Decision Criteria Are Suboptimal in In-context Learning 24 Jun 2024 · 1 repository · arXiv:2406.16535Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models 24 Jun 2024 · 1 repository · arXiv:2406.16562Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Confidence Aware Inverse Constrained Reinforcement Learning 24 Jun 2024 · 1 repository · arXiv:2406.16782Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
MD tree: a model-diagnostic tree grown on loss landscape 24 Jun 2024 · 1 repository · arXiv:2406.16988Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 9 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples)
-
AlleNoise: large-scale text classification benchmark dataset with real-world label noise 24 Jun 2024 · 1 repository · arXiv:2407.10992Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Blind Baselines Beat Membership Inference Attacks for Foundation Models 23 Jun 2024 · 1 repository · arXiv:2406.16201Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs 22 Jun 2024 · 1 repository · arXiv:2406.15927Syntology 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation 21 Jun 2024 · 1 repository · arXiv:2406.14830Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
DiPEx: Dispersing Prompt Expansion for Class-Agnostic Object Detection 21 Jun 2024 · 1 repository · arXiv:2406.14924Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
GOAL: A Generalist Combinatorial Optimization Agent Learning 21 Jun 2024 · 1 repository · arXiv:2406.15079Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
ECLIPSE: Expunging Clean-label Indiscriminate Poisons via Sparse Diffusion Purification 21 Jun 2024 · 1 repository · arXiv:2406.15093Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 3 honoured, 0 violated, 12 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 19 harvested samples) · 10 pointer-only (licence)
-
GeoLRM: Geometry-Aware Large Reconstruction Model for High-Quality 3D Gaussian Generation 21 Jun 2024 · 1 repository · arXiv:2406.15333Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning 21 Jun 2024 · 1 repository · arXiv:2406.15334Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 4 honoured, 2 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
NAVSIM: Data-Driven Non-Reactive Autonomous Vehicle Simulation and Benchmarking 21 Jun 2024 · 2 repositories · arXiv:2406.15349Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
The Elusive Pursuit of Reproducing PATE-GAN: Benchmarking, Auditing, Debugging 20 Jun 2024 · 1 repository · arXiv:2406.13985Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
LLM Critics Help Catch Bugs in Mathematics: Towards a Better Mathematical Verifier with Natural Language Feedback 20 Jun 2024 · 1 repository · arXiv:2406.14024Syntology official (archive's flag): 22 ran · 22 ran (of which 0 constructed an object rather than computing a result; 18 with no instrument failure: 0 honoured, 1 violated, 17 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 23 harvested samples) · 23 pointer-only (licence)
-
Timo: Towards Better Temporal Reasoning for Language Models 20 Jun 2024 · 1 repository · arXiv:2406.14192Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
PostMark: A Robust Blackbox Watermark for Large Language Models 20 Jun 2024 · 1 repository · arXiv:2406.14517Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
PathoLM: Identifying pathogenicity from the DNA sequence through the Genome Foundation Model 19 Jun 2024 · 1 repository · arXiv:2406.13133Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Biomedical Visual Instruction Tuning with Clinician Preference Alignment 19 Jun 2024 · 1 repository · arXiv:2406.13173Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
AdaMoE: Token-Adaptive Routing with Null Experts for Mixture-of-Experts Language Models 19 Jun 2024 · 1 repository · arXiv:2406.13233Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Large-Scale Dataset Pruning in Adversarial Training through Data Importance Extrapolation 19 Jun 2024 · 1 repository · arXiv:2406.13283Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words 19 Jun 2024 · 1 repository · arXiv:2406.13340Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Jogging the Memory of Unlearned LLMs Through Targeted Relearning Attacks 19 Jun 2024 · 1 repository · arXiv:2406.13356Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples) · 4 pointer-only (licence)
-
WATT: Weight Average Test-Time Adaptation of CLIP 19 Jun 2024 · 1 repository · arXiv:2406.13875Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
Mathador-LM: A Dynamic Benchmark for Mathematical Reasoning on Large Language Models 18 Jun 2024 · 1 repository · arXiv:2406.12572Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MAGIC: Generating Self-Correction Guideline for In-Context Text-to-SQL 18 Jun 2024 · 1 repository · arXiv:2406.12692Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI 18 Jun 2024 · 1 repository · arXiv:2406.12753Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools 18 Jun 2024 · 7 repositories · arXiv:2406.12793Syntology official (archive's flag): 4 ran · 21 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 0 honoured, 0 violated, 20 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 29 harvested samples) · 1 pointer-only (licence)
-
Dissecting Adversarial Robustness of Multimodal LM Agents 18 Jun 2024 · 1 repository · arXiv:2406.12814Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 21 harvested samples) · 1 pointer-only (licence)
-
Few-Shot Recognition via Stage-Wise Retrieval-Augmented Finetuning 17 Jun 2024 · 1 repository · arXiv:2406.11148Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models 17 Jun 2024 · 1 repository · arXiv:2406.11230Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
How Can We Effectively Expand the Vocabulary of LLMs with 0.01GB of Target Language Text? 17 Jun 2024 · 2 repositories · arXiv:2406.11477Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
GigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinement 17 Jun 2024 · 2 repositories · arXiv:2406.11546Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling 17 Jun 2024 · 1 repository · arXiv:2406.11617Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Unveiling Encoder-Free Vision-Language Models 17 Jun 2024 · 1 repository · arXiv:2406.11832Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MedCalc-Bench: Evaluating Large Language Models for Medical Calculations 17 Jun 2024 · 1 repository · arXiv:2406.12036Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
SCAR: Efficient Instruction-Tuning for Large Language Models via Style Consistency-Aware Response Ranking 16 Jun 2024 · 1 repository · arXiv:2406.10882Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
RWKU: Benchmarking Real-World Knowledge Unlearning for Large Language Models 16 Jun 2024 · 1 repository · arXiv:2406.10890Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
HAIChart: Human and AI Paired Visualization System 16 Jun 2024 · 2 repositories · arXiv:2406.11033Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Active, anytime-valid risk controlling prediction sets 15 Jun 2024 · 1 repository · arXiv:2406.10490Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Unveiling the Ignorance of MLLMs: Seeing Clearly, Answering Incorrectly 15 Jun 2024 · 1 repository · arXiv:2406.10638Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Evaluating the Generalization Ability of Quantized LLMs: Benchmark, Analysis, and Toolbox 15 Jun 2024 · 1 repository · arXiv:2406.12928Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Towards Scalable and Versatile Weight Space Learning 14 Jun 2024 · 1 repository · arXiv:2406.09997Syntology official (archive's flag): 11 ran · 11 ran (of which 7 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
ProtoS-ViT: Visual foundation models for sparse self-explainable classifications 14 Jun 2024 · 1 repository · arXiv:2406.10025Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Task-aligned Part-aware Panoptic Segmentation through Joint Object-Part Representations 14 Jun 2024 · 1 repository · arXiv:2406.10114Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack 14 Jun 2024 · 4 repositories · arXiv:2406.10149Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
DevBench: A multimodal developmental benchmark for language learning 14 Jun 2024 · 1 repository · arXiv:2406.10215Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Pareto Front-Diverse Batch Multi-Objective Bayesian Optimization 13 Jun 2024 · 1 repository · arXiv:2406.08799Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Potion: Towards Poison Unlearning 13 Jun 2024 · 1 repository · arXiv:2406.09173Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
JailbreakEval: An Integrated Toolkit for Evaluating Jailbreak Attempts Against Large Language Models 13 Jun 2024 · 1 repository · arXiv:2406.09321Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Yo'LLaVA: Your Personalized Language and Vision Assistant 13 Jun 2024 · 1 repository · arXiv:2406.09400Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Interpreting the Weight Space of Customized Diffusion Models 13 Jun 2024 · 1 repository · arXiv:2406.09413Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding 13 Jun 2024 · 1 repository · arXiv:2406.09418Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
CleanDiffuser: An Easy-to-use Modularized Library for Diffusion Models in Decision Making 13 Jun 2024 · 3 repositories · arXiv:2406.09509Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
DrivAerNet++: A Large-Scale Multimodal Car Dataset with Computational Fluid Dynamics Simulations and Deep Learning Benchmarks 13 Jun 2024 · 2 repositories · arXiv:2406.09624Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Understanding and Mitigating Compositional Issues in Text-to-Image Generative Models 12 Jun 2024 · 1 repository · arXiv:2406.07844Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
LVBench: An Extreme Long Video Understanding Benchmark 12 Jun 2024 · 1 repository · arXiv:2406.08035Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
RRLS : Robust Reinforcement Learning Suite 12 Jun 2024 · 1 repository · arXiv:2406.08406Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Fine-Tuned 'Small' LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification 12 Jun 2024 · 1 repository · arXiv:2406.08660Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Reconciling Kaplan and Chinchilla Scaling Laws 12 Jun 2024 · 1 repository · arXiv:2406.12907Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A Probabilistic Framework for LLM Hallucination Detection via Belief Tree Propagation 11 Jun 2024 · 1 repository · arXiv:2406.06950Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs 11 Jun 2024 · 3 repositories · arXiv:2406.07476Syntology community repositories only · 12 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 17 harvested samples) · 8 pointer-only (licence)
-
MAP: Low-compute Model Merging with Amortized Pareto Fronts via Quadratic Approximation 11 Jun 2024 · 1 repository · arXiv:2406.07529Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Hearing Anything Anywhere 11 Jun 2024 · 1 repository · arXiv:2406.07532Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena 11 Jun 2024 · 1 repository · arXiv:2406.07545Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Chain-of-Scrutiny: Detecting Backdoor Attacks for Large Language Models 10 Jun 2024 · 1 repository · arXiv:2406.05948Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Get rich quick: exact solutions reveal how unbalanced initializations promote rapid feature learning 10 Jun 2024 · 1 repository · arXiv:2406.06158Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dereverberation 10 Jun 2024 · 2 repositories · arXiv:2406.06185Syntology community repositories only · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Self-Tuning: Instructing LLMs to Effectively Acquire New Knowledge through Self-Teaching 10 Jun 2024 · 1 repository · arXiv:2406.06326Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Husky: A Unified, Open-Source Language Agent for Multi-Step Reasoning 10 Jun 2024 · 1 repository · arXiv:2406.06469Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
When is Multicalibration Post-Processing Necessary? 10 Jun 2024 · 3 repositories · arXiv:2406.06487Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
Merlin: A Vision Language Foundation Model for 3D Computed Tomography 10 Jun 2024 · 2 repositories · arXiv:2406.06512Syntology 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
Controlling Counterfactual Harm in Decision Support Systems Based on Prediction Sets 10 Jun 2024 · 1 repository · arXiv:2406.06671Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Stable Neighbor Denoising for Source-free Domain Adaptive Segmentation 10 Jun 2024 · 1 repository · arXiv:2406.06813Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Conformal Prediction for Class-wise Coverage via Augmented Label Rank Calibration 10 Jun 2024 · 1 repository · arXiv:2406.06818Syntology official (archive's flag): 21 ran · 21 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 0 honoured, 0 violated, 20 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 24 harvested samples) · 1 pointer-only (licence)
-
An LLM-Assisted Easy-to-Trigger Backdoor Attack on Code Completion Models: Injecting Disguised Vulnerabilities against Strong Detection 10 Jun 2024 · 1 repository · arXiv:2406.06822Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
What Can We Learn from State Space Models for Machine Learning on Graphs? 9 Jun 2024 · 1 repository · arXiv:2406.05815Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Winner-takes-all learners are geometry-aware conditional density estimators 7 Jun 2024 · 1 repository · arXiv:2406.04706Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Probabilistic Perspectives on Error Minimization in Adversarial Reinforcement Learning 7 Jun 2024 · 1 repository · arXiv:2406.04724Syntology official (archive's flag): 2 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
Massively Multiagent Minigames for Training Generalist Agents 7 Jun 2024 · 1 repository · arXiv:2406.05071Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Split-and-Fit: Learning B-Reps via Structure-Aware Voronoi Partitioning 7 Jun 2024 · 1 repository · arXiv:2406.05261Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Quality-Diversity with Limited Resources 6 Jun 2024 · 1 repository · arXiv:2406.03731Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search 6 Jun 2024 · 2 repositories · arXiv:2406.03816Syntology official (archive's flag): 6 ran · 9 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 7 where Syntology's instrument failed) · 13 unverified (of 22 harvested samples) · 13 pointer-only (licence)
-
Vectorized Conditional Neural Fields: A Framework for Solving Time-dependent Parametric Partial Differential Equations 6 Jun 2024 · 1 repository · arXiv:2406.03919Syntology official (archive's flag): 9 ran · 9 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; every one of the 9 samples that ran constructed an object rather than computing a result (of 14 harvested samples)
-
What is Dataset Distillation Learning? 6 Jun 2024 · 1 repository · arXiv:2406.04284Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)