| Physics-Guided Synthetic High-Frequency Ultrasound Generation for Skin Layer Segmentation added by Syntology |
2026-09 (from id) |
Finn-02/synthetic-hfus-skin-layer-segmentation/code/segablation/result_io.py 552dcbc7016953db |
unverified |
no licence file found · pointer only |
| Hardware-Aware FP4 FlashAttention-4 added by Syntology |
2026-09 (from id) |
MrHuff/fp4-fa4/results/fp4_fa4_b300_tuning_20260802/build_summary.py cf9fc9656dfc42dd |
unverified |
Apache-2.0 (permissive) |
| EraseSAE: Surgical Concept Erasure in Text-to-Video Diffusion Models via Sparse Autoencoders added by Syntology |
2026-09 (from id) |
HiDream-ai/EraseSAE/configs/config_loader.py 30180de1929a7196 |
unverified |
no licence file found · pointer only |
| Towards Generalizable Visually Grounded Exploration of Household Devices added by Syntology |
2026-09 (from id) |
BITHLP/VGEBench/bench_close.py a0c7393409fd0be9 |
unverified |
MIT (permissive) |
| SOVER: Formal Certification of Optimization Reformulations via LLM-Assisted SMT Verification added by Syntology |
2026-09 (from id) |
baranwa2/SOVER/code/Sover_solver_final.py b39e419911556004 |
unverified |
licence not identified · pointer only |
| Control-Data Flow Separation: Stable Prompt Optimization in Multi-Agent LLMs added by Syntology |
2026-09 (from id) |
yuntian-group/cdsep/experiments/fill_paper_tables.py 1c44cf41ee922d18 |
unverified |
MIT (permissive) |
| Do General NLP Embeddings Capture Ontological Reasoning? added by Syntology |
2026-09 (from id) |
sciknoworg/AVA/src/utils.py 2fbbe01c3721715d |
unverified |
MIT (permissive) |
| Sparse Competition during Training For the Emergence of Specialized Modules added by Syntology |
2026-08 (from id) |
BabaVegato/Minimal-SCM/scm/figures.py e7bb42db6987f3a4 |
unverified |
no licence file found · pointer only |
| Reactivating Test-Time Scaling for Plane Geometry Problem Solving added by Syntology |
2026-08 (from id) |
Jason8Kang/ReTTS-PGPS/src/retts_pgp/eval/cg_mte.py 81fcd61c6366767c |
unverified |
licence not identified · pointer only |
| InteractBench: Benchmarking LLMs on Competitive Programming under Unrevealed Information added by Syntology |
2026-08 (from id) |
kmsgk0/InteractBench/evaluate.py 4c96c8ad1727f8e9 |
unverified |
MIT (permissive) |
| ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning added by Syntology |
2026-08 (from id) |
fangjuntao/ChorusTIC/src/chorustic/model/loading.py 3dae1e0ffba39de6 |
ran
|
no licence file found · pointer only |
| EnSI-RAG: Entity-Structure-Indexed Retrieval-Augmented Generation for Long-Document Question Answering added by Syntology |
2026-08 (from id) |
RamonMeng/EnSI-RAG/ensi-rag-loong-eval/ensi_loong/core/legal_passages.py 70dc1e864dd30ccc |
ran
|
no licence file found · pointer only |
| Beyond Teacher Likelihood: Group-Calibrated On-Policy Distillation for Long-Context Reasoning added by Syntology |
2026-08 (from id) |
SolereZhang/GC-OPD/evaluation/aggregate_main_table.py 5c2ad610574cbda7 |
ran
|
Apache-2.0 (permissive) |
| Fractional Optimizers Meet Fractal Activation Functions: An Empirical Study of Multi-Scale Optimization in Neural Networks added by Syntology |
2026-08 (from id) |
Raubkatz/FractalAndFractional2026/A_03_analyze_himmelblau.py 4d82a982fa751f10 |
ran
|
CC-BY-4.0 · pointer only |
| No Universal Signal Predicts Sample-Level LLM Regression under Version Updates added by Syntology |
2026-08 (from id) |
jiashengsally/llm-regression-signals/analysis/summarize_results.py a1f8a0f919f48ad5 |
unverified |
no licence file found · pointer only |
| Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalignment added by Syntology |
2026-08 (from id) |
PeiYangLiu/icl-em-format-control/run_role_metadata_control.py c58f6281390f6e07 |
ran
|
no licence file found · pointer only |
| When Correct Solutions Repeat: Rarity-Aware Credit Redistribution for GRPO added by Syntology |
2026-08 (from id) |
CzZ12/When-Correct-Solutions-Repeat-Rarity-Aware-Credit-Redistribution-for-GRPO/analyze_mechanism.py 6d732572d6191ed4 |
ran
|
MIT (permissive) |
| Beyond a Single Judge: The Evidence-Grounded, Social-Weighted Persona Panel for Generative UI Evaluation added by Syntology |
2026-07 (from id) |
Wuzheng02/ESPP/scoring_pipeline/common.py 7bb49539f1127719 |
ran
|
no licence file found · pointer only |
| JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety added by Syntology |
2026-07 (from id) |
xiongyuaay/JANUS/eval_framework_online/annotate_final_outputs.py a145434950c05a50 |
ran
|
no licence file found · pointer only |
| Staypoint Detection from Noisy Trajectory Data [Experiment Paper] added by Syntology |
2026-07 (from id) |
amirih/staypoint/code/utils.py 53d425eb1278ddda |
ran · our draft was wrong
|
no licence file found · pointer only |
| LFM: Leveraging Foundation Models for Source-Free Universal Domain Adaptation added by Syntology |
2026-07 (from id) |
iamjingli/LFM/train_target.py 1069918b276d2855 |
ran · our draft was wrong
|
MIT (permissive) |
| Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents added by Syntology |
2026-07 (from id) |
ericjiang18/MemCon/mas/utils.py fb7d225e5f4b88c5 |
ran
fingerprinted |
no licence file found · pointer only |
| From Geometric Recovery to Causal Validation: A Reproducible Audit of Sparse Autoencoder Features, from Superposition Geometry to Causal Inertness added by Syntology |
2026-07 (from id) |
mohamed-bal/sae-causal-audit/src/sae_causal_audit/report.py fe76107a48f0807c |
ran
|
MIT (permissive) |
| Identifiability of Relational Queries in Multi-View Pretraining added by Syntology |
2026-07 (from id) |
danielhz/query-identifiability/analysis/export_pgfdata.py 29244bb954637e6f |
ran
|
MIT (permissive) |
| Can LLM-as-a-Judge Reliably Verify Rubrics in Agentic Scenarios? added by Syntology |
2026-06 (from id) |
THU-KEG/RuVerBench/code/main_leaderboard/compute_main_leaderboard.py d888aa8caa502ac5 |
ran
|
licence not identified · pointer only |
| All Routes Lead to Collapse added by Syntology |
2026-06 (from id) |
parzi-val/all-routes-lead-to-collapse/experiments/calibrate_pstar.py c4557464874373f6 |
unverified |
MIT (permissive) |
| When AUC Misleads: Polarization-Aware Evaluation of Deepfake Detectors under Domain Shift added by Syntology |
2026-06 (from id) |
mapooon/SelfBlendedImages/src/utils/funcs.py b24e1ddfb59f8b53 |
ran
|
licence not identified · pointer only |
| VeriGraph: Towards Verifiable Data-Analytic Agents added by Syntology |
2026-06 (from id) |
ignorejjj/VeriGraph/src/inference/run_inference.py 860c94939e40e6cf |
ran
|
Apache-2.0 (permissive) |
| KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty added by Syntology |
2026-06 (from id) |
naver-ai/KCSAT-ML/src/evaluator.py 88816b951ca46a4d |
ran
|
AGPL-3.0 (copyleft) · pointer only |
| KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty added by Syntology |
2026-06 (from id) |
naver-ai/KCSAT-ML/src/generator.py c700c0687c7ab7df |
ran
|
AGPL-3.0 (copyleft) · pointer only |
| p-adic Bi-Filtrations for Topological Machine Learning on Genomic Sequences added by Syntology |
2026-06 (from id) |
MAHI-Group/pVR/make_aux_tables.py 388975539a13f8a1 |
unverified |
GPL-3.0 (copyleft) · pointer only |
| When Model Merging Breaks Routing: Training-Free Calibration for MoE added by Syntology |
2026-06 (from id) |
huangcb01/HARC/src/utils.py 4ff854f78fd0c6fc |
ran
|
no licence file found · pointer only |
| Benchmarking Recursive-Collapse Warning Claims Under Matched False-Positive Control added by Syntology |
2026-06 (from id) |
davidmullett/loopzero-paper-public/src/loopzero_paper/benchmarks/recommender/audit_engine_spec.py f7b32649f4ec64b3 |
ran
|
Apache-2.0 (permissive) |
| Benchmarking Recursive-Collapse Warning Claims Under Matched False-Positive Control added by Syntology |
2026-06 (from id) |
davidmullett/loopzero-paper-public/src/loopzero_paper/benchmarks/recommender/bridge_check.py eee5a442416d9381 |
ran
|
Apache-2.0 (permissive) |
| Benchmarking Recursive-Collapse Warning Claims Under Matched False-Positive Control added by Syntology |
2026-06 (from id) |
davidmullett/loopzero-paper-public/src/loopzero_paper/benchmarks/recommender/build_user_episode_manifest.py 842e560cd0b69ed7 |
ran
|
Apache-2.0 (permissive) |
| QASM-Eval: A Dataset to Train and Evaluate LLMs on OpenQASM-3 Beyond Quantum Circuits added by Syntology |
2026-05 (from id) |
fuzhenxiao/QASM-Eval/dataset_factory/build_parquet_dataset.py b60d7699ea0bfa77 |
ran
|
no licence file found · pointer only |
| PrionNER: A Named Entity Recognition Dataset for Prion Disease Biomedical Literature added by Syntology |
2026-05 (from id) |
daotuanan/PrionNER/code/convert_w2ner_output_to_brat.py f5868f251e1785a0 |
ran
|
MIT (permissive) |
| Fine-Tuning Over Architectural Complexity: Broad-Coverage PII Detection on PIIBench with DeBERTa added by Syntology |
2026-05 (from id) |
pritesh-2711/pii-bench/compile_comparative_results.py 0e23ecb91603580c |
ran
|
Apache-2.0 (permissive) |
| Agentic Discovery of Cryomicroneedle Formulations added by Syntology |
2026-05 (from id) |
baitmeister/ML-for-CryoMN/src/06_evaluation_explainability/stage_r2_predicted_vs_actual.py 877863db38b995d5 |
ran
|
MIT (permissive) |
| InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search added by Syntology |
2026-05 (from id) |
hbhalpha/InterLV-Search-Bench/download_interlv_images.py 4a42f5d93681e5ac |
ran
|
no licence file found · pointer only |
| From Soliloquy to Agora: Memory-Enhanced LLM Agents with Decentralized Debate for Optimization Modeling added by Syntology |
2026-04 (from id) |
CHIANGEL/Agora-Opt/code/Agora-Opt/src/debate_memory/augment_memory_from_standalone_runs.py b85a1300a89a5565 |
ran
|
no licence file found · pointer only |
| Reasoning Primitives in Hybrid and Non-Hybrid LLMs: Do Architectural Differences Yield Advantages in State-Tracking and Recall? added by Syntology |
2026-04 (from id) |
ultor1996/reasoning_primitives/src/utils.py dffca3653cbd645f |
ran
|
no licence file found · pointer only |
| TriEx: A Game-based Tri-View Framework for Explaining Internal Reasoning in Multi-Agent LLMs added by Syntology |
2026-04 (from id) |
Einsam1819/TriEx/experiments/exp2c_intervention/build_global_summary.py 7698cf0c526ea351 |
unverified |
MIT (permissive) |
| Drift and selection in LLM text ecosystems added by Syntology |
2026-04 (from id) |
SR123/LLM-text-ecosystems/src/drift_selection/checkpoints.py 84a5e93fb6de5139 |
unverified |
MIT (permissive) |
| PIArena: A Platform for Prompt Injection Evaluation added by Syntology |
2026-04 (from id) |
sleeepeer/PIArena/piarena/utils.py ed4e52fc66e57123 |
unverified |
MIT (permissive) |
| Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models added by Syntology |
2026-04 (from id) |
Pepper66/DMLE/run_new.py 3e6c1ee7ad23dfb9 |
unverified |
no licence file found · pointer only |
| CALRK-Bench: Evaluating Context-Aware Legal Reasoning in Korean Law added by Syntology |
2026-03 (from id) |
jhCOR/CALRKBench/src/eval_utils.py 057a1b8fac449e37 |
unverified |
no licence file found · pointer only |
| A Sobering Look at Tabular Data Generation via Probabilistic Circuits added by Syntology |
2026-03 (from id) |
april-tools/tabpc/src/util.py eb460e0605ba191e |
unverified |
Apache-2.0 (permissive) |
| PRISM: A Dual View of LLM Reasoning through Semantic Flow and Latent Computation added by Syntology |
2026-03 (from id) |
chili-lab/PRISM/prism_code/aggregate_lib.py e90b784033e597b2 |
unverified |
no licence file found · pointer only |
| PRISM: A Dual View of LLM Reasoning through Semantic Flow and Latent Computation added by Syntology |
2026-03 (from id) |
chili-lab/PRISM/prism_code/generate_website_prism.py 4f3c32f5060f7280 |
unverified |
no licence file found · pointer only |
| RouterKGQA: Specialized-General Model Routing for Constraint-Aware Knowledge Graph Question Answering added by Syntology |
2026-03 (from id) |
Oldcircle/RouterKGQA/components/utils.py 7277c744809fb106 |
unverified |
MIT (permissive) |
| Flow Matching Policy Optimization with Mirror Descent and Entropy Constraints added by Syntology |
18 Mar 2026 |
lzqw/FLAME/analysis/plot_three_obstacle_pullback_results.py 5d7dd03a436d96f8 |
unverified |
no licence file found · pointer only |
| The Norm-Separation Delay Law of Grokking: A First-Principles Theory of Delayed Generalization added by Syntology |
2026-03 (from id) |
ClevixLab/grokking-norm-separation/make_figures.py 520f5124d96d2f01 |
unverified |
licence not identified · pointer only |
| The Norm-Separation Delay Law of Grokking: A First-Principles Theory of Delayed Generalization added by Syntology |
2026-03 (from id) |
ClevixLab/grokking-norm-separation/make_figures_supplementary.py bdf9e875347c2f9e |
unverified |
licence not identified · pointer only |
| A Tutorial Review of Bayesian Optimization with Gaussian Processes to Accelerate Stationary Point Searches added by Syntology |
2026-03 (from id) |
lode-org/ChemGP/benchmarks/analysis/summarize.py eb7369993fa6881a |
unverified |
MIT (permissive) |
| LLM as a Meta-Judge: Synthetic Data for NLP Evaluation Metric Validation added by Syntology |
2026-03 (from id) |
eiglerl/meta-judge/src/metrics/meta_correlation.py b09dbc63544053c4 |
unverified |
no licence file found · pointer only |
| SmartBench: Evaluating LLMs in Smart Homes with Anomalous Device States and Behavioral Contexts added by Syntology |
2026-03 (from id) |
horizonsinzqs/SmartBench/experiments/common/public_benchmark.py f10deba8c8561bee |
ran · our draft was wrong
|
no licence file found · pointer only |
| Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution added by Syntology |
2026-03 (from id) |
ncbi-nlp/Med-V1/utils.py d5a6f286ba2e7d66 |
unverified |
licence not identified · pointer only |
| According to Me: Long-Term Personalized Referential Memory QA added by Syntology |
2026-03 (from id) |
JingbiaoMei/ATM-Bench/agent_systems/extract_answer.py e1a33f82f601f5fe |
unverified |
MIT (permissive) |
| According to Me: Long-Term Personalized Referential Memory QA added by Syntology |
2026-03 (from id) |
JingbiaoMei/ATM-Bench/agent_systems/extract_usage.py d0d3faa7cc886a81 |
unverified |
MIT (permissive) |
| Efficient Distribution Learning with Error Bounds in Wasserstein Distance added by Syntology |
2026-02 (from id) |
EduardoFMDCosta/WassersteinDistributionLearning/configs/handlers.py d2160be344f39f25 |
unverified |
no licence file found · pointer only |
| Uncertainty Drives Social Bias Changes in Quantized Large Language Models added by Syntology |
2026-02 (from id) |
stan-hua/PostTrainingBiasBenchmark/src/utils/json_utils.py 51658330a10c0e19 |
unverified |
no licence file found · pointer only |
| Boundary Optimization for Weakly Supervised Video Grounding added by Syntology |
2026-02 (from id) |
sunoh-kim/gbo/gbocnm/utils.py 0d7ac1b9deeffa50 |
unverified |
MIT (permissive) |
| UCPO: Uncertainty-Aware Policy Optimization added by Syntology |
2026-01 (from id) |
xzhouzeng/ucpo/eval/eval_general.py a34eb4a9ee65c5fb |
unverified |
Apache-2.0 (permissive) |
| METALEAD: A Comprehensive Human-Curated Leaderboard Dataset for Transparent Reporting of Machine Learning Experiments added by Syntology |
2026-01 (from id) |
RoelTim/metalead/src/evaluate_tuples.py ba57d1fd220fdb48 |
unverified |
no licence file found · pointer only |
| UniFinEval: Towards Unified Evaluation of Financial Multimodal Models across Text, Images and Videos added by Syntology |
2026-01 (from id) |
aifinlab/UniFinEval/evaluate_py/data_loader.py d068ef448a266fb4 |
unverified |
Apache-2.0 (permissive) |
| The Viscosity of Logic: Phase Transitions and Hysteresis in DPO Alignment added by Syntology |
2026-01 (from id) |
anon-repo-317/experiments/analysis/generate_all_data.py eadae9732385cce9 |
unverified |
no licence file found · pointer only |
| Learning Fine-Grained Correspondence with Cross-Perspective Perception for Open-Vocabulary 6D Object Pose Estimation added by Syntology |
2026-01 (from id) |
zjjqinyu/FiCoP/bop_toolkit_lib/inout.py 556be2ec895c3dea |
unverified |
no licence file found · pointer only |
| PrivLEX: Detecting legal concepts in images through Vision-Language Models added by Syntology |
2026-01 (from id) |
idiap/privlex/utils.py a68d8656b672911d |
unverified |
no licence file found · pointer only |
| GenCtrl -A Formal Controllability Toolkit for Generative Models added by Syntology |
2026-01 (from id) |
apple/ml-genctrl/genctrl/utils/setup_utils.py 654a593d62a67e06 |
unverified |
licence not identified · pointer only |
| Linear Dynamics in the RLVR Training of Large Language Models added by Syntology |
2026-01 (from id) |
Miaow-Lab/RLVR-Linearity/analysis/token_logprob/token_logprob_linearity.py db79e1c6f71f41bb |
unverified |
MIT (permissive) |
| Linear Dynamics in the RLVR Training of Large Language Models added by Syntology |
2026-01 (from id) |
Miaow-Lab/RLVR-Linearity/evaluation/pass_at_k_eval.py 98cb765f6409dd94 |
unverified |
MIT (permissive) |
| Re-Rankers as Relevance Judges added by Syntology |
2026-01 (from id) |
ChuanMeng/reranker-as-judge/cross_evaluation.py e36b276d9d394d5f |
unverified |
no licence file found · pointer only |
| INFINITEWEB: Scalable Web Environment Synthesis for GUI Agent Training added by Syntology |
2026-01 (from id) |
microsoft/FIVE-UI-Evol/InfiniteWeb/src/generate_task_jsons.py 752cc2290bb5485c |
unverified |
MIT (permissive) |
| ICLR 2026 Workshop: Principled Design for Trustworthy AI REDBENCH: A UNIVERSAL DATASET FOR COMPRE-HENSIVE RED TEAMING OF LARGE LANGUAGE MOD-ELS added by Syntology |
2026-01 (from id) |
knoveleng/redeval/redeval/score.py 67e96c33980c36f3 |
unverified |
MIT (permissive) |
| JP-TL-Bench: Anchored Pairwise LLM Evaluation for Bidirectional Japanese-English Translation added by Syntology |
2026-01 (from id) |
lhl/liquid-ai-hackathon-tokyo/eval/tui-viewer.py b19d72f495a1c73e |
unverified |
no licence file found · pointer only |
| VCWorld: A Biological World Model for Virtual Cell Simulation added by Syntology |
2025-12 (from id) |
GENTEL-lab/VCWorld/src/cli_pipeline/stages/prompt.py 581e1a3430c33f5f |
unverified |
no licence file found · pointer only |
| DCcluster-Opt: Benchmarking Dynamic Multi-Objective Optimization for Geo-Distributed Data Center Workloads added by Syntology |
2025-11 (from id) |
HewlettPackard/sustain-cluster/utils/config_loader.py 593e5e4d03d4d4df |
ran · our draft was wrong
|
MIT (permissive) |
| MM-OPERA: Benchmarking Open-ended Association Reasoning for Large Vision-Language Models added by Syntology |
2025-10 (from id) |
MM-OPERA-Bench/MM-OPERA/evaluation/utils.py 2b770451741c0c2c |
unverified |
MIT (permissive) |
| RESCUE: Retrieval Augmented Secure Code Generation added by Syntology |
2025-10 (from id) |
steven1518/RESCUE/src/common/utils.py 70da73eeeea12da5 |
unverified |
MIT (permissive) |
| Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model added by Syntology |
2025-10 (from id) |
RuipingL/Situat3DChange/SCReasoner/common/io_utils.py bace1acf5a748097 |
unverified |
CC-BY-4.0 · pointer only |
| Preserving LLM Capabilities through Calibration Data Curation: From Analysis to Optimization added by Syntology |
2025-10 (from id) |
BokwaiHo/COLA/cola/utils.py 3d492573a8beef65 |
unverified |
no licence file found · pointer only |
| Base Models Know How to Reason, Thinking Models Learn When added by Syntology |
8 Oct 2025 |
cvenhoff/thinking-llms-interp/human_eval/sample.py 03331bee0c7d9d6d |
unverified |
no licence file found · pointer only |
| CAM: A Constructivist View of Agentic Memory for LLM-Based Reading Comprehension added by Syntology |
2025-10 (from id) |
rui9812/CAM/prototype/constructivist_memory.py 3ffef03aeacde643 |
ran · our draft was wrong
|
no licence file found · pointer only |
| BiasBusters: Uncovering and Mitigating Tool Selection Bias in Large Language Models added by Syntology |
2025-10 (from id) |
thierry123454/tool-selection-bias/5_bias_investigation/extract_features/add_selection_rate.py a0c7393409fd0be9 |
unverified |
no licence file found · pointer only |
| From Conversation to Query Execution: Benchmarking User and Tool Interactions for EHR Database Agents added by Syntology |
2025-09 (from id) |
glee4810/EHR-ChatQA/src/utils.py fd3ad6e7caf61e76 |
unverified |
no licence file found · pointer only |
| Unveiling the Merits and Defects of LLMs in Automatic Review Generation for Scientific Papers added by Syntology |
2025-09 (from id) |
RichardLRC/Peer-Review/Code/Semantic_similarity.py a0c7393409fd0be9 |
unverified |
no licence file found · pointer only |
| PruneCD: Contrasting Pruned Self Model to Improve Decoding Factuality added by Syntology |
2025-09 (from id) |
hoeng4/PruneCD/2_benchmark/1_tfqa_llama_rating.py d0bb7013dd302153 |
unverified |
no licence file found · pointer only |
| Pun Unintended: LLMs and the Illusion of Humor Understanding added by Syntology |
2025-09 (from id) |
alezanga/punintended/utils/io.py cff0a4f4e937c72c |
unverified |
no licence file found · pointer only |
| Bridging Graph and State-Space Modeling for Intensive Care Unit Length of Stay Prediction added by Syntology |
2025-08 (from id) |
ShuqiZi1/S2G-Net/src/utils.py 0ca0c1bca741caa0 |
unverified |
MIT (permissive) |
| arXiv:2507.22171 |
2025-07 (from id) |
CjangCjengh/Generic_Persona/trust_utils.py a0c7393409fd0be9 |
unverified |
no licence file found · pointer only |
| arXiv:2507.15501 |
2025-07 (from id) |
apple/ml-aspera/src/aspera/readers.py ef54978796fcb878 |
unverified |
licence not identified · pointer only |
| EvoAgentX: An Automated Framework for Evolving Agentic Workflows |
4 Jul 2025 |
evoagentx/evoagentx/evoagentx/core/module_utils.py 2c12d75b45786664 |
unverified |
licence not identified · pointer only |
| arXiv:2506.19054 |
2025-06 (from id) |
AI-secure/PolyGuard/code/evaluate.py a0c7393409fd0be9 |
unverified |
no licence file found · pointer only |
| Unable to Forget: Proactive lnterference Reveals Working Memory Limits in LLMs Beyond Context Length |
9 Jun 2025 |
zhuangziGiantfish/Unable-to-Forget/automation/analyze_pi_flow_final.py 484ae19889957547 |
unverified |
MIT (permissive) |
| arXiv:2506.03949 |
2025-06 (from id) |
wenge-research/TableEval/utils.py c3b9cc79f7f285ce |
unverified |
Apache-2.0 (permissive) |
| EvaLearn: Quantifying the Learning Capability and Efficiency of LLMs via Sequential Problem Solving |
3 Jun 2025 |
ByteDance-Seed/EvaLearn/Evaluate/evaluate_metric.py 9872647e95bfe2f0 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| VeriThoughts: Enabling Automated Verilog Code Generation using Reasoning and Formal Verification |
16 May 2025 |
wilyub/verithoughts/evaluation_verithoughts/verilog_vllm_multi.py b534fded631201b9 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Surrogate Modeling of 3D Rayleigh-Benard Convection with Equivariant Autoencoders |
19 May 2025 |
fynnfromme/equivariant-rb-forecasting/experiments/evaluate.py 4170a4a36bf2db15 |
unverified |
MIT (permissive) |
| ESC-Judge: A Framework for Comparing Emotional Support Conversational Agents |
18 May 2025 |
navidmdn/ESC-Judge/multidim_judge_from_merged.py dd20d05a85fb7631 |
unverified |
Apache-2.0 (permissive) |
| Data Whisperer: Efficient Data Selection for Task-Specific LLM Fine-Tuning via Few-Shot In-Context Learning |
18 May 2025 |
gszfwsb/Data-Whisperer/pruning/datawhisperer_gsm_pruner.py f33868b25affe264 |
ran · our draft was wrong
|
no licence file found · pointer only |
| One2Any: One-Reference 6D Pose Estimation for Any Object |
7 May 2025 |
lmy1001/One2Any/bop_toolkit_lib/inout.py 556be2ec895c3dea |
unverified |
MIT (permissive) |
| SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image Editing |
5 May 2025 |
bytedance/superedit/eval/eval_instructpix2pix.py 53d425eb1278ddda |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| arXiv:2504.04907 |
2025-04 (from id) |
Video-Bench/Video-Bench/videobench/utils.py 81155ad798ce29cf |
unverified |
Apache-2.0 (permissive) |
| arXiv:2504.04635 |
2025-04 (from id) |
patqdasilva/steering-off-course/DoLa/tfqa_gpt3_rating.py d0bb7013dd302153 |
unverified |
MIT (permissive) |
| Unraveling the Effects of Synthetic Data on End-to-End Autonomous Driving |
23 Mar 2025 |
cancaries/SceneCrafter/data_utils/get_kinemic_hyper.py bff1f7ec64588b7c |
unverified |
MIT (permissive) |
| MathFusion: Enhancing Mathematic Problem-solving of LLM through Instruction Fusion |
20 Mar 2025 |
qizhipei/mathfusion/evaluation/dart_math/utils.py 3c40f841274b3441 |
ran
|
Apache-2.0 (permissive) |