| The Dialect Tax: Dialectal Biases Persist throughout the Language Modeling Pipeline added by Syntology |
2026-08 (from id) |
socialnlp/dialecttax/src/dialecttax/prompts.py efdb63cf50def444 |
ran
fingerprinted |
MIT (permissive) |
| Enhancing LLM Metacognition via Cognitive Pairwise Training added by Syntology |
2026-06 (from id) |
Tsinghua-dhy/CPT/eval/utils/judge_math_answer_gpt.py 04002d844af70af7 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Enhancing LLM Metacognition via Cognitive Pairwise Training added by Syntology |
2026-06 (from id) |
Tsinghua-dhy/CPT/eval/utils/math_equal.py a8eed4c1f2f2831b |
ran
|
Apache-2.0 (permissive) |
| REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations added by Syntology |
2026-05 (from id) |
Buyun-Liang/REALISTA/src/realista.py 4c2b9df401898003 |
ran · our draft was wrong
|
MIT (permissive) |
| A Stability Benchmark of Generative Regularizers for Inverse Problems added by Syntology |
2026-05 (from id) |
alexdenker/GenRegBench/main_flow.py cab1fb2b4d059325 |
ran
fingerprinted |
no licence file found · pointer only |
| REDPARROT: Accelerating NL-to-DSL for Business Analytics via Query Semantic Caching added by Syntology |
2026-04 (from id) |
TommyIsNotHere/RedParrot/hybrid_rewrite/prompt.py 1bd873161388cd83 |
unverified |
no licence file found · pointer only |
| Bootstrapping Post-training Signals for Open-ended Tasks via Rubric-based Self-play on Pre-training Text added by Syntology |
2026-04 (from id) |
HCY123902/POP/synthesis/sample.py fb0f7009b88df9e4 |
unverified |
no licence file found · pointer only |
| CBRS: Cognitive Blood Request System with Bilingual Dataset and Dual-Layer Filtering for Multi-Platform Social Streams added by Syntology |
2026-04 (from id) |
aaniksahaa/CBRS/binary-classifier/dual-layer-filtering/eval-v1.py 62cf1ea9a6a67f98 |
unverified |
no licence file found · pointer only |
| Reinforcement Learning with Conditional Expectation Reward added by Syntology |
2026-03 (from id) |
changyi7231/CER/recipe/cer/src/data_preparation.py a6f637988f82514b |
unverified |
Apache-2.0 (permissive) |
| LLM-as-an-Annotator: Training Lightweight Models with LLM-Annotated Examples for Aspect Sentiment Tuple Prediction added by Syntology |
2026-03 (from id) |
NilsHellwig/LA-ABSA/01_create_annotated_examples/03_llm_eda_few_shot_augmenter.py eb03a5076b44c920 |
unverified |
no licence file found · pointer only |
| Safety Alignment as Continual Learning: Mitigating the Alignment Tax via Orthogonal Gradient Projection added by Syntology |
2026-02 (from id) |
SunGL001/OGPSA/eval/AlpacaEval_pre.py fa528f9eb5d2a2af |
unverified |
MIT (permissive) |
| SEMPIPES -Optimizable Semantic Data Operators for Tabular Machine Learning Pipelines added by Syntology |
2026-02 (from id) |
noahho/CAAFE/caafe/caafe.py 1288c06136735dd7 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| BEAR: Towards Beam-Search-Aware Optimization for Recommendation with Large Language Models added by Syntology |
2026-01 (from id) |
Tiny-Snow/BEAR-SIGIR-2026/utils.py 369f30df2ac7e4ca |
unverified |
MIT (permissive) |
| UniFinEval: Towards Unified Evaluation of Financial Multimodal Models across Text, Images and Videos added by Syntology |
2026-01 (from id) |
aifinlab/UniFinEval/evaluate_py/prompts.py eb48ceb1095ac012 |
unverified |
Apache-2.0 (permissive) |
| LILO: Bayesian Optimization with Natural Language Feedback added by Syntology |
2025-10 (from id) |
facebookresearch/lilo/lilo/human_feedback_simulator.py af08af9d44db5a5c |
unverified |
MIT (permissive) |
| SECA: Semantically Equivalent and Coherent Attacks for Eliciting LLM Hallucinations added by Syntology |
2025-10 (from id) |
Buyun-Liang/SECA/src/seca.py cf0a660abde18864 |
ran · our draft was wrong
|
MIT (permissive) |
| AgentTTS: Large Language Model Agent for Test-time Compute-optimal Scaling Strategy in Complex Tasks added by Syntology |
2025-08 (from id) |
FairyFali/AgentTTS/code_others/archon_taskbench.py 043fc4b1c4483d0c |
unverified |
Apache-2.0 (permissive) |
| arXiv:2507.10302 |
2025-07 (from id) |
ZJHTerry18/DisCo/evaluation/eval_egoschema.py 7321ecad6597e9ff |
unverified |
no licence file found · pointer only |
| Sample Efficient Demonstration Selection for In-Context Learning |
10 Jun 2025 |
kiranpurohit/case/Code/LLM_experiments/CASE_Gsm8K_selection.py f4cfbae1ebeb837e |
ran · our draft was wrong
|
MIT (permissive) |
| ReasonGen-R1: CoT for Autoregressive Image generation models through SFT and RL |
30 May 2025 |
Franklin-Zhang0/ReasonGen-R1/benchmark/generate_inference_dpg.py 0b420baae60fbb33 |
unverified |
Apache-2.0 (permissive) |
| Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation |
29 May 2025 |
mainlp/CoT2EL/Pipeline/generator.py 556513f9f88952ef |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning |
20 May 2025 |
visual-agent/deepeyes/eval/judge_result.py 1b4e5e1a7c7a8eee |
unverified |
Apache-2.0 (permissive) |
| DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving |
2025-05 (from id) |
paparare/driveagent/enviroment.py 8dc3ad180e974ab6 |
unverified |
MIT (permissive) |
| Two Heads are Better Than One: Test-time Scaling of Multi-agent Collaborative Reasoning |
14 Apr 2025 |
jincan333/MAS-TTS/evaluate_dataset.py 3cbafc44dba51a7d |
unverified |
Apache-2.0 (permissive) |
| Cultural Learning-Based Culture Adaptation of Language Models |
3 Apr 2025 |
ukplab/arxiv2025-clca/CLCA/llm_roleplay/common/model_tuning.py ebd71a3edf5b5421 |
unverified |
Apache-2.0 (permissive) |
| FaceBench: A Multi-View Multi-Level Facial Attribute VQA Dataset for Benchmarking Face Perception MLLMs |
27 Mar 2025 |
CVI-SZU/FaceBench/evaluation/evaluation.py 7af521bbe681ec46 |
unverified |
MIT (permissive) |
| OpenVLThinker: An Early Exploration to Complex Vision-Language Reasoning via Iterative Self-Improvement |
21 Mar 2025 |
yihedeng9/openvlthinker/v1/evaluation/verify_mathverse_gpt4.py ccfc28a818189693 |
unverified |
Apache-2.0 (permissive) |
| DPImageBench: A Unified Benchmark for Differentially Private Image Synthesis |
18 Mar 2025 |
2019chengong/dpimagebench/evaluation/evaluator.py 092ceca055502762 |
ran · our draft was wrong
|
MIT (permissive) |
| Rank1: Test-Time Compute for Reranking in Information Retrieval |
25 Feb 2025 |
orionw/rank1/prompts.py b52aabdc316d82de |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Unveiling the Key Factors for Distilling Chain-of-Thought Reasoning |
25 Feb 2025 |
eit-nlp/distilling-cot-reasoning/Evaluation/reasoning_eval/prompt_utils.py 9dcdd19fa07a9553 |
unverified |
Apache-2.0 (permissive) |
| MDCrow: Automating Molecular Dynamics Workflows with Large Language Models |
13 Feb 2025 |
ur-whitelab/MDCrow/notebooks/experiments/Robustness/robustness_prompts.py 5a0080520de90215 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Cached Multi-Lora Composition for Multi-Concept Image Generation |
7 Feb 2025 |
Yqcca/CMLoRA/utils.py bb2bba270e29d9ee |
unverified |
MIT (permissive) |
| STP: Self-play LLM Theorem Provers with Iterative Conjecturing and Proving |
31 Jan 2025 |
kfdong/STP/RL/utils/model_utils.py ba7c25429714f2b5 |
unverified |
MIT (permissive) |
| Multi-Grained Patch Training for Efficient LLM-based Recommendation |
25 Jan 2025 |
ljy0ustc/patchrec/utils.py 7fcc27faf959d335 |
unverified |
Apache-2.0 (permissive) |
| No Preference Left Behind: Group Distributional Preference Optimization |
28 Dec 2024 |
BigBinnie/GDPO/evaluate_BPC.py d420ed106490686d |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| DISCERN: Decoding Systematic Errors in Natural Language for Text Classifiers |
29 Oct 2024 |
rrmenon10/DISCERN/src/discern/refine.py 478e1775f2ce81ae |
ran · our draft was wrong
|
no licence file found · pointer only |
| IntLoRA: Integral Low-rank Adaptation of Quantized Diffusion Models |
29 Oct 2024 |
csguoh/IntLoRA/evaluation.py ed99491f78bd49e1 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Infogent: An Agent-Based Framework for Web Information Aggregation |
24 Oct 2024 |
gangiswag/infogent/direct-api-driven/fanoutqa_answer.py 05e25bd96ace7a5c |
unverified |
Apache-2.0 (permissive) |
| Refusal-Trained LLMs Are Easily Jailbroken As Browser Agents |
11 Oct 2024 |
scaleapi/browser-art/src/agents/OpenDevin/agenthub/browsing_agent/browsing_agent.py dd9d65d0d8999bb0 |
ran
fingerprinted |
licence not identified · pointer only |
| Neuron-based Personality Trait Induction in Large Language Models |
16 Oct 2024 |
rucaibox/npti/NPTI/code/gpt4_score.py 6340f1429608e7df |
ran · our draft was wrong
|
no licence file found · pointer only |
| MMIE: Massive Multimodal Interleaved Comprehension Benchmark for Large Vision-Language Models |
14 Oct 2024 |
Lillianwei-h/MMIE/prompts.py 3007270c1b69f9a2 |
unverified |
MIT (permissive) |
| Regressing the Relative Future: Efficient Policy Optimization for Multi-turn RLHF |
6 Oct 2024 |
zhaolingao/refuel/setting_one/user_generator.py fc0417f911166c83 |
ran
|
Apache-2.0 (permissive) |
| MC-CoT: A Modular Collaborative CoT Framework for Zero-shot Medical-VQA with LLM and MLLM Integration |
6 Oct 2024 |
thomaswei-cn/MC-CoT/method/VisualOnly.py 9c7f5d246b874c50 |
ran
fingerprinted |
no licence file found · pointer only |
| From Reading to Compressing: Exploring the Multi-document Reader for Prompt Compression |
5 Oct 2024 |
eunseongc/r2c/src_comp/compress_utils.py edc9d2c346eeeec0 |
ran
|
no licence file found · pointer only |
| TPP-LLM: Modeling Temporal Point Processes by Efficiently Fine-Tuning Large Language Models |
2 Oct 2024 |
zefang-liu/TPP-LLM/src/tpp_llm/utils.py dd21bdbaacebe685 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| MECD: Unlocking Multi-Event Causal Discovery in Video Reasoning |
26 Sep 2024 |
tychen-sjtu/mecd-benchmark/mecd_vllm_fewshot/VideoChat2/multi_event.py 748d5a2586ca3178 |
ran · our draft was wrong
|
MIT (permissive) |
| Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation |
2 Sep 2024 |
mayuelala/followyourcanvas/inference_outpainting-dir-with-prompt.py f67980fe15864660 |
unverified |
no licence file found · pointer only |
| HERMES: temporal-coHERent long-forM understanding with Episodes and Semantics |
30 Aug 2024 |
joslefaure/HERMES/lavis/tasks/moviecore_eval_prompts.py 50228589040ec9ac |
unverified |
MIT (permissive) |
| Automated Design of Agentic Systems |
15 Aug 2024 |
shengranhu/adas/_arc/arc_prompt.py e303009857c65536 |
unverified |
Apache-2.0 (permissive) |
| Automated Design of Agentic Systems |
15 Aug 2024 |
shengranhu/adas/_drop/drop_prompt.py 7c697edb0437e5f5 |
unverified |
Apache-2.0 (permissive) |
| Automated Design of Agentic Systems |
15 Aug 2024 |
shengranhu/adas/_gpqa/gpqa_prompt.py 505fd0fb8ba0fca7 |
unverified |
Apache-2.0 (permissive) |
| Future Events as Backdoor Triggers: Investigating Temporal Vulnerabilities in LLMs |
4 Jul 2024 |
sbp354/future_triggered_backdoors/future_probing/headline_prompting/prompting_utils.py 2a75fb471a0321a6 |
ran
|
no licence file found · pointer only |
| Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning |
30 Jun 2024 |
mathllm/Step-Controlled_DPO/src/step_controled_dpo_lce/lce_solution_gen_different_negative_gsm8k.py 587a8be3d359400c |
ran
|
no licence file found · pointer only |
| Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning |
30 Jun 2024 |
mathllm/Step-Controlled_DPO/src/step_controled_dpo_lce/lce_solution_gen_different_negative_gsm8k_divided.py a681b908de92655d |
ran
|
no licence file found · pointer only |
| ALiiCE: Evaluating Positional Fine-grained Citation Generation |
19 Jun 2024 |
ylXuu/ALiiCE/generate.py a9aab662c8ac6543 |
ran · our draft was wrong
|
MIT (permissive) |
| Tokenization Falling Short: On Subword Robustness in Large Language Models |
17 Jun 2024 |
floatai/tkeval/src/utils.py 3d2dcc489e259a67 |
ran
|
MIT (permissive) |
| It Takes Two: On the Seamlessness between Reward and Policy Model in RLHF |
12 Jun 2024 |
taiminglu/seamless/code/retrival/gpt/retrival.py 845ccc17a73fe529 |
ran
fingerprinted |
no licence file found · pointer only |
| Failures Are Fated, But Can Be Faded: Characterizing and Mitigating Unwanted Behaviors in Large-Scale Vision and Language Models |
11 Jun 2024 |
somsagar07/FailureShiftRL/Baselines/Generation/config.py 9f09587c18b2d305 |
ran
fingerprinted |
MIT (permissive) |
| Vript: A Video Is Worth Thousands of Words |
10 Jun 2024 |
mutonix/Vript/vript-hard/models/videochat2/videochat2_vriptCAP.py 748d5a2586ca3178 |
ran · our draft was wrong
|
no licence file found · pointer only |
| MLVU: Benchmarking Multi-task Long Video Understanding |
6 Jun 2024 |
identical code first harvested elsewhere 748d5a2586ca3178 |
ran · our draft was wrong
|
licence of this copy not recorded |
| LLMGeo: Benchmarking Large Language Models on Image Geolocation In-the-wild |
30 May 2024 |
yeyimilk/llmgeo/src/fschat.py b691056ddf790e1d |
ran
fingerprinted |
no licence file found · pointer only |
| ATM: Adversarial Tuning Multi-agent System Makes a Robust Retrieval-Augmented Generator |
28 May 2024 |
chuhac/ATM-RAG/atm_train/attacker_build_data/prompting_for_rag.py 3df45990990f233a |
unverified |
no licence file found · pointer only |
| Evaluating and Safeguarding the Adversarial Robustness of Retrieval-Based In-Context Learning |
24 May 2024 |
simonucl/adv-retreival-icl/src/quick_exp.py 8eddb6f291aed485 |
ran
|
no licence file found · pointer only |
| Evaluating and Safeguarding the Adversarial Robustness of Retrieval-Based In-Context Learning |
24 May 2024 |
simonucl/adv-retreival-icl/src/transfer_attack.py d47716bfc331903f |
unverified |
no licence file found · pointer only |
| Constrained Decoding for Secure Code Generation |
30 Apr 2024 |
dynamite321/codeguardplus/correctness_eval.py 718990946e149fe1 |
ran
|
MIT (permissive) |
| LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models |
8 Apr 2024 |
maitrix-org/llm-reasoners/examples/ReasonerAgent-Web/baseline/openhands_browsing_agent.py d59734e9fc209dc9 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models |
3 Apr 2024 |
coniferlm/conifer/utils.py 82360028338c7371 |
ran · our draft was wrong
|
no licence file found · pointer only |
| The Impact of Prompts on Zero-Shot Detection of AI-Generated Text |
29 Mar 2024 |
kaito25atugich/detector/tmp/generate_estimation_prompts.py 9a9e58cbb8b7370b |
ran
fingerprinted |
no licence file found · pointer only |
| How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments |
18 Mar 2024 |
cuhk-arise/gamabench/server.py 73776d01a9a3cb5a |
ran
|
GPL-3.0 (copyleft) · pointer only |
| A Comprehensive Study of Multimodal Large Language Models for Image Quality Assessment |
16 Mar 2024 |
tianhewu/mllms-for-iqa/prompts/gpt4v_prompt.py 0ed0bc77cad17572 |
unverified |
no licence file found · pointer only |
| ERBench: An Entity-Relationship based Automatically Verifiable Hallucination Benchmark for Large Language Models |
8 Mar 2024 |
dilab-kaist/erbench/binary/run_qa.py cd65beff1b6f34f4 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| DS-Agent: Automated Data Science by Empowering Large Language Models with Case-Based Reasoning |
27 Feb 2024 |
guosyjlu/DS-Agent/deployment/prompt.py 15c1e503e527a5bf |
unverified |
no licence file found · pointer only |
| EHRNoteQA: An LLM Benchmark for Real-World Clinical Practice Using Discharge Summaries |
25 Feb 2024 |
ji-youn-kim/ehrnoteqa/src/evaluation/utils.py e80d57247e8567c0 |
ran
|
MIT (permissive) |
| EHRNoteQA: An LLM Benchmark for Real-World Clinical Practice Using Discharge Summaries |
25 Feb 2024 |
ji-youn-kim/ehrnoteqa/src/generation/utils.py 574a354a5ba5d4c1 |
ran
fingerprinted |
MIT (permissive) |
| When Do LLMs Need Retrieval Augmentation? Mitigating LLMs' Overconfidence Helps Retrieval Augmentation |
18 Feb 2024 |
shiyunee/when-to-retrieve/utils/prompt.py 277b2d82e854b570 |
unverified |
no licence file found · pointer only |
| DE-COP: Detecting Copyrighted Content in Language Models Training Data |
15 Feb 2024 |
avduarte333/de-cop_method/2_decop_hf.py abeb23f8b2eef76f |
unverified |
Apache-2.0 (permissive) |
| Policy Improvement using Language Feedback Models |
12 Feb 2024 |
vzhong/language_feedback_models/experiments/prompt_utils.py fe361276102577d5 |
ran
|
MIT (permissive) |
| ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models |
24 Jan 2024 |
rohan598/contextual/eval/response_eval_gpt4.py bf992be8225b649f |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint |
11 Jan 2024 |
rucaibox/rlmec/evaluate/Math/prompt_utils.py ad26ca57156a1cf3 |
ran
|
no licence file found · pointer only |
| Exploring Large Language Model based Intelligent Agents: Definitions, Methods, and Prospects |
7 Jan 2024 |
melih-unsal/demogpt/demogpt/prompt.py 938213d47dbb3323 |
ran
|
MIT (permissive) |
| Supervised Knowledge Makes Large Language Models Better In-context Learners |
26 Dec 2023 |
yanglinyi/supervised-knowledge-makes-large-language-models-better-in-context-learners/NLG/model.py 0dd5a42ce6af7976 |
ran · our draft was wrong
|
no licence file found · pointer only |
| SOLAR 10.7B: Scaling Large Language Models with Simple yet Effective Depth Up-Scaling |
23 Dec 2023 |
jquesnelle/yarn/eval/quality.py 8d8ff31fdde07297 |
ran
|
MIT (permissive) |
| Safety Alignment in NLP Tasks: Weakly Aligned Summarization as an In-Context Attack |
12 Dec 2023 |
fyyfu/safetyalignnlp/multi_prompt_generation.py 873e7b82a3f92f78 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Biomedical knowledge graph-optimized prompt generation for large language models |
29 Nov 2023 |
BaranziniLab/KG_RAG/kg_rag/utility.py a7b09ec8d815a408 |
unverified |
Apache-2.0 (permissive) |
| MVBench: A Comprehensive Multi-modal Video Understanding Benchmark |
28 Nov 2023 |
opengvlab/ask-anything/video_chat/conversation.py 748d5a2586ca3178 |
ran · our draft was wrong
|
MIT (permissive) |
| GENOME: GenerativE Neuro-symbOlic visual reasoning by growing and reusing ModulEs |
8 Nov 2023 |
umass-foundation-model/genome/engine/prompt.py 04269c1b860b83ab |
unverified |
Apache-2.0 (permissive) |
| Can Foundation Models Watch, Talk and Guide You Step by Step to Make a Cake? |
1 Nov 2023 |
sled-group/watch-talk-and-guide/src/util.py 508259e833b104dd |
ran
|
MIT (permissive) |
| How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition |
9 Oct 2023 |
wangrongsheng/caregpt/Gradio/model.py 0dd5a42ce6af7976 |
ran · our draft was wrong
|
MIT recorded; this copy not marked cleared · pointer only |
| Large Language Models Only Pass Primary School Exams in Indonesia: A Comprehensive Test on IndoMMLU |
7 Oct 2023 |
fajri91/indommlu/evaluate.py 3565fb4a9a36a3c8 |
ran
|
MIT (permissive) |
| FinGPT: Instruction Tuning Benchmark for Open-Source Large Language Models in Financial Datasets |
7 Oct 2023 |
AI4Finance-Foundation/FinGPT/fingpt/FinGPT_Benchmark/utils.py 130ab5832e26fbdd |
ran
|
MIT (permissive) |
| Evaluating Hallucinations in Chinese Large Language Models |
5 Oct 2023 |
xiami2019/halluqa/calculate_metrics.py 38fc10020d01d898 |
ran
|
Apache-2.0 (permissive) |
| Beyond Task Performance: Evaluating and Reducing the Flaws of Large Multimodal Models with In-Context Learning |
1 Oct 2023 |
mshukor/EvALign-ICL/open_flamingo/eval/caption_utils.py 48d71ff25acc020d |
ran
|
no licence file found · pointer only |
| Causal Discovery with Language Models as Imperfect Experts |
5 Jul 2023 |
StephLong614/Causal-disco/utils/language_models.py 835e7af85a1f4c96 |
ran
|
no licence file found · pointer only |
| Diffusion Model is an Effective Planner and Data Synthesizer for Multi-Task Reinforcement Learning |
29 May 2023 |
tinnerhrhe/MTDiff/diffuser/models/temporal.py 5321d9d86ff019c1 |
unverified |
MIT (permissive) |
| Conformal Prediction with Large Language Models for Multi-Choice Question Answering |
28 May 2023 |
bhaweshiitk/conformalllm/conformal_llm_scores.py 62b2e921ddd66bd8 |
ran · our draft was wrong
|
MIT (permissive) |
| QLoRA: Efficient Finetuning of Quantized LLMs |
23 May 2023 |
georgesung/llm_qlora/inference.py 8bdd290bac0f9c85 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| InstructAlign: High-and-Low Resource Language Alignment via Continual Crosslingual Instruction Tuning |
23 May 2023 |
hltchkust/instructalign/nlu_prompt.py 9239c0977078e7fe |
unverified |
Apache-2.0 (permissive) |
| Adaptive Chameleon or Stubborn Sloth: Revealing the Behavior of Large Language Models in Knowledge Conflicts |
22 May 2023 |
copenlu/context-utilisation-for-rag/src/get_model_predictions/get_model_predictions.py 525dbdd48bc6f54c |
ran
|
GPL-3.0 (copyleft) · pointer only |
| Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models |
6 May 2023 |
AGI-Edgerunners/Plan-and-Solve-Prompting/prompt.py c90b87435028c310 |
unverified |
no licence file found · pointer only |
| Large Language Models for Automated Data Science: Introducing CAAFE for Context-Aware Automated Feature Engineering |
5 May 2023 |
noahho/caafe/caafe/caafe.py 1288c06136735dd7 |
ran · fixture could not drive it
|
licence not identified · pointer only |
| Differentiable Data Augmentation for Contrastive Sentence Representation Learning |
29 Oct 2022 |
TianduoWang/DiffAug/diffaug/models.py 1694baf70cd8e233 |
ran · our draft was wrong
|
MIT (permissive) |
| DreamBooth: Fine Tuning Text-to-Image Diffusion Models for Subject-Driven Generation |
25 Aug 2022 |
identical code first harvested elsewhere ed99491f78bd49e1 |
ran · our draft was wrong
|
licence of this copy not recorded |
| arXiv:aaai_34547 |
|
Aatrox103/CrAM/utils/re_weighting.py 1472b8313235999a |
unverified |
no licence file found · pointer only |
| arXiv:2024.findings-emnlp.86 |
|
FloatAI/TKEval/src/utils.py 3d2dcc489e259a67 |
ran
|
MIT (permissive) |
| arXiv:2024.findings-acl.326 |
|
dscc-admin-ch/statbot.swiss/src/utility.py 0dd5a42ce6af7976 |
ran · our draft was wrong
|
Unlicense (permissive) |
| arXiv:2024.findings-acl.326 |
|
dscc-admin-ch/statbot.swiss/src/main-mixtral.py 0f6f984fa9637589 |
unverified |
Unlicense (permissive) |