| VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages added by Syntology |
2026-09 (from id) |
Usneek1/VakyArth/src/tasks/mcq.py de9ae4128c68a13e |
unverified |
licence not identified · pointer only |
| VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages added by Syntology |
2026-09 (from id) |
Usneek1/VakyArth/src/tasks/nli.py e8db056ec5c17672 |
unverified |
licence not identified · pointer only |
| Dotting the Eye: An Intent-Driven Image Retouching Agent for Visual Focus Enhancement added by Syntology |
2026-09 (from id) |
DragonisCV/EyeControl/src/inference/infer.py a69c968fe0bfe0bd |
unverified |
no licence file found · pointer only |
| DERELAB: Probing Defeasible Reasoning and Confirmation Bias in LLMs with a Generative Benchmark added by Syntology |
2026-08 (from id) |
Jayanta47/DeReLab/generate_paraphrases.py 395655befcf85fb6 |
unverified |
licence not identified · pointer only |
| SPEAR: Distilling Domain-Adaptive Reasoning Skeletons via Sequential Symbolic Alignment in Reinforcement Learning added by Syntology |
2026-08 (from id) |
zhuochunli/SPEAR/utils.py a383de780ac0aa57 |
ran
fingerprinted |
no licence file found · pointer only |
| DeflectBench: A Benchmark for Evaluating Rhetorical Fallacy Generation in LLMs added by Syntology |
2026-08 (from id) |
ArtKanke/DeflectBench/config.py 8737904f0167a07d |
ran
|
MIT (permissive) |
| Reconstructing the Right Episode: Evaluating Interleaved Conversational Memory Beyond Long Context added by Syntology |
2026-08 (from id) |
LordTARN1SHED/SCALE-QA/tsim_reference/src/eval/run_eval.py ef22962a80a1e411 |
ran
|
licence not identified · pointer only |
| Decorrelation Is Not Complementarity: Skill, Not Lineage, Governs Trusted-Monitor Ensembles added by Syntology |
2026-08 (from id) |
anik-jha/challenger-panels/src/monitor.py 5ca319bba2a6961c |
ran
|
MIT (permissive) |
| SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning added by Syntology |
2026-07 (from id) |
chiefovoavicii/SERPO/eval/eval_gpqa_diamond.py ae8e147c9885175d |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models added by Syntology |
2026-07 (from id) |
js-lee-AI/answer-leakage/answer_leakage/prompts.py 6d43f4e39a356fd7 |
ran
|
MIT (permissive) |
| Do We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters? added by Syntology |
2026-07 (from id) |
GAIR-Lab/Light-MER/train_stage2_mgrpo.py 09028445218f8ba8 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Cross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on Romanian added by Syntology |
2026-06 (from id) |
DS4AI-UPB/crosslingual-romanian-re/infer_classification.py a326b5c18eb8d0da |
ran
fingerprinted |
no licence file found · pointer only |
| Cross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on Romanian added by Syntology |
2026-06 (from id) |
DS4AI-UPB/crosslingual-romanian-re/infer_e2e.py dc60f75e190164fc |
ran
fingerprinted |
no licence file found · pointer only |
| A Good Talk Doesn't Look Like a Summary, It Teaches You! Measuring Takeaways from Paper-to-Video Talks added by Syntology |
2026-06 (from id) |
showlab/Paper2Video/src/evaluation/MetaSim_content.py 5f2503b0d7ac1d38 |
ran
|
MIT (permissive) |
| ReNIO: Reweighting Negative Trajectory Importance for LLM On-Policy Distillation added by Syntology |
2026-06 (from id) |
BDML-lab/ReNIO/eval/evaluate_code.py 736f1be23541cbcd |
ran
|
MIT (permissive) |
| BLUEX v2: Benchmarking LLMs on Open-Ended Questions from Brazilian University Entrance Exams added by Syntology |
2026-06 (from id) |
TropicAI-Research/BLUEXv2/inference/run_inference.py a542c779b01fcdc2 |
ran
fingerprinted |
no licence file found · pointer only |
| TrustMargin: Training-Free Arbitration between Parametric Memory and Retrieved Evidence in Large Language Models added by Syntology |
2026-06 (from id) |
mojixu/TrustMargin/src/trustmargin.py 1d8f494789547560 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Tangram: Unlocking Non-Uniform KV Cache Compression for Efficient Multi-turn LLM Serving added by Syntology |
2026-06 (from id) |
aiha-lab/TANGRAM/benchmarks/tangram/ruler_local.py bcd9d69939662364 |
ran
|
Apache-2.0 (permissive) |
| Enhancing LLM Metacognition via Cognitive Pairwise Training added by Syntology |
2026-06 (from id) |
Tsinghua-dhy/CPT/data_construction/4_build_sft/build_sft_dataset.py 74944fb1c9af1d08 |
ran
|
Apache-2.0 (permissive) |
| Can LLMs Time Travel? Enhancing Temporal Consistency in Legal Agentic Search through Reinforcement Learning added by Syntology |
2026-05 (from id) |
AlexFanw/LegalSearch-R1/user/legalsearch_data_process.py d6beb2151fab8888 |
unverified |
Apache-2.0 (permissive) |
| Generalization or Memorization? Brittleness Testing for Chess-Trained Language Models added by Syntology |
2026-05 (from id) |
ethanjtang/GAMBIT/eval_models_on_puzzles/eval_all_models_base.py e3266e4f554042c3 |
ran
fingerprinted |
licence not identified · pointer only |
| Skill Retrieval Augmentation for Agentic AI added by Syntology |
2026-04 (from id) |
oneal2000/SR-Agents/src/sragents/prompts.py 87fb084b0ab65529 |
unverified |
MIT (permissive) |
| A Systematic Analysis of the Impact of Persona Steering on LLM Capabilities added by Syntology |
2026-04 (from id) |
cjia7/DPR/src/npti/eval/gpt4_score.py 16529904d2d5be9f |
unverified |
no licence file found · pointer only |
| Unified Deployment-Aware Evaluation of Open Reasoning Language Models added by Syntology |
2026-04 (from id) |
mkboch/UDAE/prompts/builder.py 35b0ce59f15f774d |
unverified |
no licence file found · pointer only |
| THIVLVC: Retrieval Augmented Dependency Parsing for Latin added by Syntology |
2026-04 (from id) |
l-pommeret/THIVLVC/src/pipeline/stage1_rag_with_baseline.py c22afcb8a1f118a3 |
ran · our draft was wrong
|
MIT (permissive) |
| RAG or Learning? Understanding the Limits of LLM Adaptation under Continuous Knowledge Drift in the Real World added by Syntology |
2026-04 (from id) |
hbing-l/chronos/baselines/memit.py 22a46266c7fa5cbf |
unverified |
no licence file found · pointer only |
| Same Geometry, Opposite Noise: Transformer Magnitude Representations Lack Scalar Variability added by Syntology |
2026-04 (from id) |
synthiumjp/weber/m3_pilot/m3_rerun_identification.py a192415d50000694 |
unverified |
no licence file found · pointer only |
| Prompts Without Evidence: How Neuroimaging Mentions Shift Clinical Vision-Language Model Predictions added by Syntology |
2026-03 (from id) |
long21wt/scaffold-effect/src/inference_joint.py a1ee925fdd89e7a4 |
unverified |
no licence file found · pointer only |
| CALRK-Bench: Evaluating Context-Aware Legal Reasoning in Korean Law added by Syntology |
2026-03 (from id) |
jhCOR/CALRKBench/src/type2_eval.py 324410cfb2d1081e |
unverified |
no licence file found · pointer only |
| TreeTeaming: Autonomous Red-Teaming of Vision-Language Models via Hierarchical Strategy Exploration added by Syntology |
2026-03 (from id) |
ChunXiaostudy/TreeTeaming/generate/batch_qwen_image_edit.py 05b3c14592dae070 |
unverified |
no licence file found · pointer only |
| PRISM: A Dual View of LLM Reasoning through Semantic Flow and Latent Computation added by Syntology |
2026-03 (from id) |
chili-lab/PRISM/prism_code/classifier.py 6a9d7dbe0ac8d9bf |
unverified |
no licence file found · pointer only |
| VietJobs: A Vietnamese Job Advertisement Dataset added by Syntology |
2026-03 (from id) |
VinNLP/VietJobs/format_prompt_category.py ff09cb63990754ad |
unverified |
MIT (permissive) |
| BLUFF: Benchmarking the Detection of False and Synthetic Content across 58 Low-Resource Languages added by Syntology |
2026-03 (from id) |
jsl5710/BLUFF/code-base/generation_codes/eng_x_f/eng_x_f.py a328dee234881191 |
unverified |
licence not identified · pointer only |
| BLUFF: Benchmarking the Detection of False and Synthetic Content across 58 Low-Resource Languages added by Syntology |
2026-03 (from id) |
jsl5710/BLUFF/code-base/generation_codes/eng_x_r/eng_x_r.py 239c2c0b061fb692 |
unverified |
licence not identified · pointer only |
| BLUFF: Benchmarking the Detection of False and Synthetic Content across 58 Low-Resource Languages added by Syntology |
2026-03 (from id) |
jsl5710/BLUFF/code-base/generation_codes/x_eng_f/x_eng_f.py 42ff7e217bee8614 |
unverified |
licence not identified · pointer only |
| ScrapeGraphAI-100k: Dataset for Schema-Constrained LLM Generation added by Syntology |
2026-02 (from id) |
ScrapeGraphAI/scrapegraph-100k-paper/modelling/convert_awq.py 9cca23f2500b9d75 |
unverified |
no licence file found · pointer only |
| HyperRAG: Reasoning N-ary Facts over Hypergraphs for Retrieval Augmented Generation added by Syntology |
2026-02 (from id) |
Vincent-Lien/HyperRAG/dataset/query2text.py a6e6bdcb966b6c62 |
unverified |
MIT (permissive) |
| Reasoning Beyond Literal: Cross-style Multimodal Reasoning for Figurative Language Understanding added by Syntology |
2026-01 (from id) |
scheshmi/CrossStyle-MMR/infer/vllm_baseline.py 2d4a4f42c50d8de2 |
unverified |
no licence file found · pointer only |
| Domain-Specific Knowledge Graphs in RAG-Enhanced Healthcare LLMs added by Syntology |
2026-01 (from id) |
sydneyanuyah/RAGComparison/artifacts/csv_coref_switcher.py 21a18e971bb86c8c |
unverified |
no licence file found · pointer only |
| Domain-Specific Knowledge Graphs in RAG-Enhanced Healthcare LLMs added by Syntology |
2026-01 (from id) |
sydneyanuyah/RAGComparison/artifacts/complex_simplify_sentences.py 718b35386828443e |
unverified |
no licence file found · pointer only |
| Domain-Specific Knowledge Graphs in RAG-Enhanced Healthcare LLMs added by Syntology |
2026-01 (from id) |
sydneyanuyah/RAGComparison/artifacts/rel_extraction_pipeline.py 36c1eae64901861f |
unverified |
no licence file found · pointer only |
| Domain-Specific Knowledge Graphs in RAG-Enhanced Healthcare LLMs added by Syntology |
2026-01 (from id) |
sydneyanuyah/RAGComparison/artifacts/simplify_by_label.py 1b77216c9a0e9270 |
unverified |
no licence file found · pointer only |
| Steering Language Models Before They Speak: Logit-Level Interventions added by Syntology |
2026-01 (from id) |
hsannn/swai/logit_steering.py 824dcd5664e63215 |
unverified |
Apache-2.0 (permissive) |
| ShortCoder: Knowledge-Augmented Syntax Optimization for Token-Efficient Code Generation added by Syntology |
2026-01 (from id) |
DeepSoftwareAnalytics/ShorterCode/reference_lora.py df9332e0280d4162 |
unverified |
no licence file found · pointer only |
| SEEK: Steering LLM Reasoning for RAG via Internal Reasoning Sketches added by Syntology |
2026-01 (from id) |
OpenBMB/PAGER/src/infer_page.py 508d2df531171de9 |
unverified |
MIT (permissive) |
| FINCARDS: Card-Based Analyst Reranking for Financial Document Question Answering added by Syntology |
2026-01 (from id) |
XanderZhou2022/FINCARDS/pipeline/stage0_query_intent.py e244d9d2aab28bd0 |
unverified |
no licence file found · pointer only |
| HAPS: Hierarchical LLM Routing with Joint Architecture and Parameter Search added by Syntology |
2026-01 (from id) |
zihangtian/HAPS/HotpotQA/joint_rl/rl_batch.py 07274e6bcd20a571 |
ran
|
no licence file found · pointer only |
| Calibrating Generative Models to Distributional Constraints added by Syntology |
2025-10 (from id) |
smithhenryd/cgm/gemma/cgm_gemma.py 083602ed7c768390 |
unverified |
MIT (permissive) |
| Cache-to-Cache: Direct Semantic Communication Between Large Language Models added by Syntology |
2025-10 (from id) |
thu-nics/C2C/rosetta/utils/evaluate.py 472ca2ea3d9cecca |
unverified |
Apache-2.0 (permissive) |
| Automated Knowledge Graph Construction using Large Language Models and Sentence Complexity Modelling added by Syntology |
2025-09 (from id) |
KaushikMahmud/CoDe-KG_EMNLP_2025/CoDe-KG/prompts.py ec50c014e2ffb4f5 |
unverified |
MIT (permissive) |
| Scaling up Multi-Turn Off-Policy RL and Multi-Agent Tree Search for LLM Step-Provers added by Syntology |
8 Sep 2025 |
ByteDance-Seed/BFS-Prover-V2/src/plan/generate.py 7bf0d8bea27ef263 |
unverified |
Apache-2.0 (permissive) |
| MIRAGE: A Benchmark for Multimodal Information-Seeking and Reasoning in Agricultural Expert-Guided Conversations |
25 Jun 2025 |
mirage-benchmark/mirage-benchmark/MMMT/src/prompt_builder.py 73f970619d92b562 |
unverified |
no licence file found · pointer only |
| Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation |
3 Jun 2025 |
CuSO4-Chen/PLI/src/evaluation/gsm8k_eval.py 2c2ea637ced966fa |
unverified |
MIT (permissive) |
| MF-LLM: Simulating Population Decision Dynamics via a Mean-Field Large Language Model Framework |
30 Apr 2025 |
Miracle1207/Mean-Field-LLM/mf_llm/evaluate/evaluate_gpt_batch.py 07ad866b84a13b94 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Valid Text-to-SQL Generation with Unification-based DeepStochLog |
17 Mar 2025 |
identical code first harvested elsewhere abcf216a0f054de7 |
ran · our draft was wrong
|
licence of this copy not recorded |
| LongSafety: Evaluating Long-Context Safety of Large Language Models |
24 Feb 2025 |
thu-coai/LongSafety/src/gen_model_response.py 3e34927af2210615 |
unverified |
MIT (permissive) |
| Primus: A Pioneering Collection of Open-Source Datasets for Cybersecurity LLM Training |
16 Feb 2025 |
huggingface/cosmopedia/prompts/auto_math_text/build_science_prompts.py 35dcf93579591fec |
unverified |
Apache-2.0 (permissive) |
| Leveraging Online Olympiad-Level Math Problems for LLMs Training and Contamination-Resistant Evaluation |
24 Jan 2025 |
dsl-lab/aops/classify_aops.py 403712b231130602 |
unverified |
Apache-2.0 (permissive) |
| SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information |
21 Sep 2024 |
GasolSun36/SURf/eval/pope.py c8c51dd88ad58cb1 |
ran
fingerprinted |
no licence file found · pointer only |
| SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information |
21 Sep 2024 |
GasolSun36/SURf/initial/data_filter.py 48f2a9620f2ac197 |
ran
fingerprinted |
no licence file found · pointer only |
| Improving Factuality in Large Language Models via Decoding-Time Hallucinatory and Truthful Comparators |
22 Aug 2024 |
ydk122024/cdt/src/benchmark_evaluation/alpaca_eval.py b13ae15599bf9354 |
ran
|
MIT (permissive) |
| Improving Factuality in Large Language Models via Decoding-Time Hallucinatory and Truthful Comparators |
22 Aug 2024 |
ydk122024/cdt/src/benchmark_evaluation/knight_eval.py a3be7b801f8fda87 |
ran
|
MIT (permissive) |
| Improving Factuality in Large Language Models via Decoding-Time Hallucinatory and Truthful Comparators |
22 Aug 2024 |
ydk122024/cdt/src/benchmark_evaluation/xsum_eval.py d80c60be83bb2afa |
ran
|
MIT (permissive) |
| MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs |
15 Jul 2024 |
mail-research/metallm-wrapper/benchmark.py 10a73330aef4702b |
ran
|
GPL-3.0 (copyleft) · pointer only |
| ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools |
18 Jun 2024 |
thudm/chatglm-6b/cli_demo.py 20f3c3d6999848e9 |
ran
|
Apache-2.0 (permissive) |
| ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools |
18 Jun 2024 |
thudm/chatglm-6b/cli_demo_vision.py 7ce6ba0a6898058f |
ran
|
Apache-2.0 (permissive) |
| ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools |
18 Jun 2024 |
thudm/chatglm/basic_demo/cli_demo.py 45ac4198df42b5f0 |
unverified |
Apache-2.0 (permissive) |
| Instruct-MusicGen: Unlocking Text-to-Music Editing for Music Language Models via Instruction Tuning |
28 May 2024 |
MS-P3/code5/chatglm3/cli_demo.py 45ac4198df42b5f0 |
unverified |
Apache-2.0 (permissive) |
| Large Language Models-guided Dynamic Adaptation for Temporal Knowledge Graph Reasoning |
23 May 2024 |
jiapuwang/LLM-DA/Iteration_reasoning.py ce7c550a5364d864 |
ran · our draft was wrong
|
no licence file found · pointer only |
| LoongServe: Efficiently Serving Long-Context Large Language Models with Elastic Sequence Parallelism |
15 Apr 2024 |
LoongServe/LoongServe/loongserve/longserve_server/build_prompt.py b710e60c3d04dc58 |
ran
|
Apache-2.0 (permissive) |
| How Much are Large Language Models Contaminated? A Comprehensive Survey and the LLMSanitize Library |
31 Mar 2024 |
ntunlp/llmsanitize/llmsanitize/closed_data_methods/ts_guessing_question_based.py a29d8491ca5349bd |
ran
|
Apache-2.0 (permissive) |
| WikiFactDiff: A Large, Realistic, and Temporally Adaptable Dataset for Atomic Factual Knowledge Update in Causal Language Models |
21 Mar 2024 |
orange-opensource/wikifactdiff/evaluate/baselines/prompt.py 5d99924f6d4fd65d |
ran
|
MIT (permissive) |
| Evolutionary Optimization of Model Merging Recipes |
19 Mar 2024 |
sakanaai/evolutionary-model-merge/evomerge/models/llava.py cd57ef201d9c8c10 |
ran
|
Apache-2.0 (permissive) |
| How Susceptible are Large Language Models to Ideological Manipulation? |
18 Feb 2024 |
kaichen23/llm_ideo_manipulate/code/run_tuned_llama2.py ab98f3588bdf16a7 |
ran
fingerprinted |
MIT (permissive) |
| Limits of Transformer Language Models on Learning to Compose Algorithms |
8 Feb 2024 |
ibm/limitations-lm-algorithmic-compositional-learning/prompting_experiment_utils/prompt_builder_perm.py 7bfe356714c7bb6d |
unverified |
Apache-2.0 (permissive) |
| MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries |
27 Jan 2024 |
yixuantt/MultiHop-RAG/qa_llama.py 0768880b580a1361 |
ran · our draft was wrong
|
no licence file found · pointer only |
| SciInstruct: a Self-Reflective Instruction Annotated Dataset for Training Scientific Language Models |
15 Jan 2024 |
THUDM/SciGLM/inference/cli_demo.py 938a3cf8db709378 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Promptly Predicting Structures: The Return of Inference |
12 Jan 2024 |
utahnlp/prompts-for-structures/src/tasks/srl/qasrl2/prompts.py c050aaa66fbc820f |
ran
|
Apache-2.0 (permissive) |
| Alleviating Hallucinations of Large Language Models through Induced Hallucinations |
25 Dec 2023 |
hillzhang1999/icd/src/benchmark_evaluation/factscore_eval.py db8a8ad434c0dd70 |
ran
|
MIT (permissive) |
| CharacterGLM: Customizing Chinese Conversational AI Characters with Large Language Models |
28 Nov 2023 |
thu-coai/characterglm-6b/basic_demo/cli_demo.py 0b952c9d3758f73e |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Prompts have evil twins |
13 Nov 2023 |
rimon15/propane/evil_twins/prompt_optim.py 9180e51a8ac82897 |
unverified |
GPL-2.0 (copyleft) · pointer only |
| LLM-driven Multimodal Target Volume Contouring in Radiation Oncology |
3 Nov 2023 |
tvseg/mm-llm-ro/utils/data_utils.py 9bd73c5e00265752 |
ran
|
no licence file found · pointer only |
| Calibrating LLM-Based Evaluator |
23 Sep 2023 |
FKIRSTE/emnlp2024-personalized-meeting-sum/baseline/G-all.py 7fc2fb12340dc132 |
unverified |
Apache-2.0 (permissive) |
| Calibrating LLM-Based Evaluator |
23 Sep 2023 |
FKIRSTE/emnlp2024-personalized-meeting-sum/baseline/OUT-P-all.py 101c21be672b0d7b |
unverified |
Apache-2.0 (permissive) |
| Calibrating LLM-Based Evaluator |
23 Sep 2023 |
FKIRSTE/emnlp2024-personalized-meeting-sum/baseline/P-all.py 8d97761d2b2796e5 |
unverified |
Apache-2.0 (permissive) |
| Calibrating LLM-Based Evaluator |
23 Sep 2023 |
FKIRSTE/emnlp2024-personalized-meeting-sum/baseline/P-none.py 228f2ec1f2816b90 |
unverified |
Apache-2.0 (permissive) |
| PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization |
8 Jun 2023 |
weopenml/pandalm/pandalm/run-gradio.py f4bb8eecfe1a4f76 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| A New Dataset and Empirical Study for Sentence Simplification in Chinese |
7 Jun 2023 |
THUDM/ChatGLM-6B/cli_demo.py 20f3c3d6999848e9 |
ran
|
Apache-2.0 (permissive) |
| A New Dataset and Empirical Study for Sentence Simplification in Chinese |
7 Jun 2023 |
THUDM/ChatGLM-6B/cli_demo_vision.py 7ce6ba0a6898058f |
ran
|
Apache-2.0 (permissive) |
| TheoremQA: A Theorem-driven Question Answering dataset |
21 May 2023 |
THUDM/VisualGLM-6B/cli_demo_hf.py b7583b2677118645 |
unverified |
Apache-2.0 (permissive) |
| DeepStochLog: Neural Stochastic Logic Programming |
23 Jun 2021 |
identical code first harvested elsewhere abcf216a0f054de7 |
ran · our draft was wrong
|
licence of this copy not recorded |
| Spider: A Large-Scale Human-Labeled Dataset for Complex and Cross-Domain Semantic Parsing and Text-to-SQL Task |
24 Sep 2018 |
ML-KULeuven/deepstochlog-lm/src/task2/task2.py abcf216a0f054de7 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| arXiv:openreview_XtIRCAEYoJ |
|
WayneTomas/Artemis/infer_artemis.py 5b32b161ba880bb2 |
unverified |
Apache-2.0 (permissive) |
| arXiv:2025.findings-acl.804 |
|
WalterPaci/IMPAQTS-PID/MCG_task.py 1db7ebc69f9ff2cb |
unverified |
MIT (permissive) |