| TAO-Attack: Toward Advanced Optimization-Based Jailbreak Attacks for Large Language Models added by Syntology |
2026-03 (from id) |
ZevineXu/TAO-Attack/llm_attacks/base/attack_manager.py f556dda0785f3c77 |
unverified |
MIT (permissive) |
| Improving Set Function Approximation with Quasi-Arithmetic Neural Networks added by Syntology |
2026-02 (from id) |
tomastokar/Quasi-Arithmetic-Neural-Networks/MNIST_qualitative.py d4aa88c9ae85da11 |
unverified |
no licence file found · pointer only |
| CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval added by Syntology |
2026-01 (from id) |
idirlab/CaseFacts/experiments/check_similarity.py a5a629d94ddef395 |
unverified |
no licence file found · pointer only |
| Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers added by Syntology |
2026-01 (from id) |
glad-lab/MGEO/attack_autodan.py 19c802b4dd21c8b8 |
unverified |
no licence file found · pointer only |
| CLINIC: Evaluating Multilingual Trustworthiness in Language Models for Healthcare added by Syntology |
2025-12 (from id) |
AikyamLab/clinic/evaluation/adverserial/adv_eval.py e51b60b9add827b6 |
unverified |
MIT (permissive) |
| CLINIC: Evaluating Multilingual Trustworthiness in Language Models for Healthcare added by Syntology |
2025-12 (from id) |
AikyamLab/clinic/evaluation/consistency/consistency_eval.py daf8bacd15f6bdea |
unverified |
MIT (permissive) |
| Building Data-Driven Occupation Taxonomies: A Bottom-Up Multi-Stage Approach via Semantic Clustering and Multi-Agent Collaboration added by Syntology |
2025-09 (from id) |
aida-ugent/CLIMB/src/get_embeddings.py 5eb54fb92e768705 |
unverified |
licence not identified · pointer only |
| Breaking Free from MMI: A New Frontier in Rationalization by Probing Input Utilization |
8 Mar 2025 |
jugechengzi/Rationalization-N2R/embedding.py a7fd4428a7a26977 |
ran
|
MIT (permissive) |
| MERGE$^3$: Efficient Evolutionary Merging on Consumer-grade GPUs |
9 Feb 2025 |
tommasomncttn/merge3/src/mergenetic/analysis/representation.py 1286c013573044fc |
unverified |
MIT (permissive) |
| WyckoffDiff -- A Generative Diffusion Model for Crystal Symmetry |
10 Feb 2025 |
httk/wyckoffdiff/wyckoff_generation/evaluation/compute_fwd.py 408fa913477a97c5 |
unverified |
MIT (permissive) |
| LLMs Can Simulate Standardized Patients via Agent Coevolution |
16 Dec 2024 |
zjumai/evopatient/embedding_function/sentence_embedding.py fc512197798fe2a7 |
unverified |
no licence file found · pointer only |
| MC-LLaVA: Multi-Concept Personalized Vision-Language Model |
18 Nov 2024 |
arctanxarc/mc-llava/train/train_joint.py 5db573c804b47ea8 |
ran · our draft was wrong
|
MIT (permissive) |
| Understanding Scaling Laws with Statistical and Approximation Theory for Transformer Neural Networks on Intrinsically Low-dimensional Data |
11 Nov 2024 |
dahoas/transformer_manifolds_learning/embeddings.py a4824be70cdf2b77 |
ran · our draft was wrong
|
no licence file found · pointer only |
| BendVLM: Test-Time Debiasing of Vision-Language Embeddings |
7 Nov 2024 |
waltergerych/bend_vlm/bend_utils.py 050b743339edef0b |
unverified |
no licence file found · pointer only |
| Robust AI-Generated Text Detection by Restricted Embeddings |
10 Oct 2024 |
silversolver/robustatd/fit_eraser_probing_tasks.py e99730bed9668d68 |
unverified |
no licence file found · pointer only |
| Is the MMI Criterion Necessary for Interpretability? Degenerating Non-causal Features to Plain Noise for Self-Rationalization |
8 Oct 2024 |
jugechengzi/Rationalization-MRD/embedding.py a7fd4428a7a26977 |
ran
|
MIT (permissive) |
| Infer Human's Intentions Before Following Natural Language Instructions |
26 Sep 2024 |
simon-wan/fiser/networks/embeddings.py 16b037202af1a18b |
ran
|
Apache-2.0 (permissive) |
| PROSE-FD: A Multimodal PDE Foundation Model for Learning Multiple Operators for Forecasting Fluid Dynamics |
15 Sep 2024 |
felix-lyx/prose/prose_fd/models/attention_utils.py df45f4439f402a46 |
ran
|
MIT (permissive) |
| Unifying Causal Representation Learning with the Invariance Principle |
4 Sep 2024 |
causallearningai/istant/src/model.py c2c6795c770350e2 |
unverified |
MIT (permissive) |
| AgentPoison: Red-teaming LLM Agents via Poisoning Memory or Knowledge Bases |
17 Jul 2024 |
BillChan226/AgentPoison/algo/utils.py 655e5409659cee36 |
unverified |
MIT (permissive) |
| Evidential Concept Embedding Models: Towards Reliable Concept Explanations for Skin Disease Diagnosis |
27 Jun 2024 |
obiyoag/evi-cem/learn_cavs.py 6cd52c2775a843cb |
ran
|
Apache-2.0 (permissive) |
| Improved Few-Shot Jailbreaking Can Circumvent Aligned Language Models and Their Defenses |
3 Jun 2024 |
sail-sg/I-FSJ/llm_attacks/base/attack_manager.py f556dda0785f3c77 |
unverified |
MIT (permissive) |
| Improved Techniques for Optimization-Based Jailbreaking on Large Language Models |
31 May 2024 |
jiaxiaojunqaq/i-gcg/llm_attacks/base/attack_manager.py f556dda0785f3c77 |
unverified |
no licence file found · pointer only |
| Quantitative Certification of Bias in Large Language Models |
29 May 2024 |
uiuc-focal-lab/LLMCert-B/utils.py e00307369c36acc3 |
ran
|
AGPL-3.0 (copyleft) · pointer only |
| Protecting Your LLMs with Information Bottleneck |
22 Apr 2024 |
llm-attacks/llm-attacks/llm_attacks/base/attack_manager.py f556dda0785f3c77 |
unverified |
MIT (permissive) |
| No "Zero-Shot" Without Exponential Data: Pretraining Concept Frequency Determines Multimodal Model Performance |
4 Apr 2024 |
bethgelab/frequency_determines_performance/src/retrieval_eval.py 507a009866a94a6f |
ran
|
MIT (permissive) |
| DiLM: Distilling Dataset into Language Model for Text-level Dataset Distillation |
30 Mar 2024 |
arumaekawa/dilm/src/coreset/coreset_utils.py 499bced10bfc04a9 |
unverified |
MIT (permissive) |
| Automatic and Universal Prompt Injection Attacks against Large Language Models |
7 Mar 2024 |
sheltonliu-n/universal-prompt-injection/utils/opt_utils.py f556dda0785f3c77 |
unverified |
MIT (permissive) |
| Accelerating Greedy Coordinate Gradient and General Prompt Optimization via Probe Sampling |
2 Mar 2024 |
zhaoyiran924/probe-sampling/llm_attacks/base/attack_manager.py db2b1708ea3c49ed |
unverified |
MIT (permissive) |
| GISTEmbed: Guided In-sample Selection of Training Negatives for Text Embedding Fine-tuning |
26 Feb 2024 |
avsolatorio/gistembed/gist_embed/trainer/loss.py 7f91a871ff0d03ed |
ran
|
no licence file found · pointer only |
| Defending LLMs against Jailbreaking Attacks via Backtranslation |
26 Feb 2024 |
yihanwang617/llm-jailbreaking-defense-backtranslation/GCG/llm_attacks/base/attack_manager.py f556dda0785f3c77 |
unverified |
BSD-3-Clause (permissive) |
| Query Augmentation by Decoding Semantics from Brain Signals |
24 Feb 2024 |
yeziyi1998/brain-query-augmentation/ict/model_utils/model_config.py bde4c1a685708c16 |
unverified |
MIT (permissive) |
| Learning to Poison Large Language Models for Downstream Manipulation |
21 Feb 2024 |
rookiezxy/gbtl/llm_attacks/base/attack_manager.py a42389ea4ae4a27a |
unverified |
no licence file found · pointer only |
| Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia |
8 Feb 2024 |
solidshen/ripple_official/src/utils/attack_manager.py 708a052bc12dad60 |
unverified |
no licence file found · pointer only |
| Robust Prompt Optimization for Defending Language Models Against Jailbreaking Attacks |
30 Jan 2024 |
lapisrocks/rpo/rpo/opt_utils.py 9a3d93b72c5ec034 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| GOAt: Explaining Graph Neural Networks via Graph Output Attribution |
26 Jan 2024 |
sluxsr/GOAt/node_classi_plot_pack/nc_plots.py 1ddb0751a34af6db |
ran
|
no licence file found · pointer only |
| Enhancing the Rationale-Input Alignment for Self-explaining Rationalization |
7 Dec 2023 |
jugechengzi/dar/embedding.py a7fd4428a7a26977 |
ran
|
MIT (permissive) |
| Hijacking Large Language Models via Adversarial In-Context Learning |
16 Nov 2023 |
RookieZxy/GGI-attack/GGI-attack/llm_attacks/base/attack_manager.py 33377834cc7d0c84 |
unverified |
no licence file found · pointer only |
| AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models |
23 Oct 2023 |
rotaryhammer/code-autodan/autodan/autodan/attack_manager.py 01abe3bd48ba14fc |
unverified |
MIT (permissive) |
| D-Separation for Causal Self-Explanation |
23 Sep 2023 |
jugechengzi/Rationalization-MCD/embedding.py a7fd4428a7a26977 |
ran
|
MIT (permissive) |
| Adversarial Illusions in Multi-Modal Embeddings |
22 Aug 2023 |
ebagdasa/adversarial_illusions/dataset_utils.py 996d7797acf3500c |
ran
|
MIT (permissive) |
| Still No Lie Detector for Language Models: Probing Empirical and Conceptual Roadblocks |
30 Jun 2023 |
balevinstein/probes/Train_CCSProbe.py 1fd11bf75e70a151 |
ran · our draft was wrong
|
MIT (permissive) |
| NOTABLE: Transferable Backdoor Attacks Against Prompt-based NLP Models |
28 May 2023 |
ru-system-software-and-security/notable/autoprompt/create_trigger.py 20a9293a820f617d |
ran · our draft was wrong
|
no licence file found · pointer only |
| Augmentation-Adapted Retriever Improves Generalization of Language Models as Generic Plug-In |
27 May 2023 |
openai/chatgpt-retrieval-plugin/services/openai.py b009eb83df585f1c |
unverified |
MIT (permissive) |
| Discover and Cure: Concept-aware Mitigation of Spurious Correlation |
1 May 2023 |
Wuyxin/DISC/disc/concept_utils/cav_utils.py 87689c095e68cee7 |
unverified |
MIT (permissive) |
| Investigating the Effectiveness of Task-Agnostic Prefix Prompt for Instruction Following |
28 Feb 2023 |
seonghyeonye/icil/src/run_nearest_demo.py 73680032c6cc1a34 |
unverified |
MIT (permissive) |
| Cluster & Tune: Boost Cold Start Performance in Text Classification |
20 Mar 2022 |
ibm/intermediate-training-using-clustering/run_experiment.py c1fcf89801c4ea9e |
unverified |
Apache-2.0 (permissive) |
| A Neural Network Solves, Explains, and Generates University Math Problems by Program Synthesis and Few-Shot Learning at Human Level |
31 Dec 2021 |
idrori/mathq/code/embedding.py caf8168c846f56d5 |
unverified |
MIT (permissive) |
| Virtual Augmentation Supported Contrastive Learning of Sentence Representations |
16 Oct 2021 |
amazon-science/sentence-representations/DownstreamEval/clustering/clustering_eval.py 13e3886c6d762ec6 |
unverified |
Apache-2.0 (permissive) |
| AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated Prompts |
29 Oct 2020 |
ucinlp/autoprompt/autoprompt/create_trigger.py 20a9293a820f617d |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| ProtTrans: Towards Cracking the Language of Life's Code Through Self-Supervised Deep Learning and High Performance Computing |
13 Jul 2020 |
agemagician/ProtTrans/Embedding/prott5_embedder.py dd492312bc93c287 |
unverified |
MIT (permissive) |
| The Oxford Radar RobotCar Dataset: A Radar Extension to the Oxford RobotCar Dataset |
2019-09 (from id) |
mttgdd/oord-dataset/src/compute_distance_matrix.py adbf95aba4ffd8f5 |
unverified |
MIT (permissive) |
| Show, Attend and Tell: Neural Image Caption Generation with Visual Attention |
10 Feb 2015 |
LinXueyuanStdio/LaTeX_OCR/model/decoder.py 9b24ffeda92bc6e1 |
unverified |
Apache-2.0 (permissive) |
| arXiv:aaai_16580 |
|
modriczhang/HRL-Rec/layer_util.py fb6bc9cda68a32c1 |
unverified |
Apache-2.0 (permissive) |
| arXiv:2024.findings-naacl.294 |
|
balevinstein/Probes/Train_CCSProbe.py 1fd11bf75e70a151 |
ran · our draft was wrong
|
MIT (permissive) |
| arXiv:2024.findings-naacl.294 |
|
balevinstein/Probes/Generate_CCS_predictions.py e4680e27c1f698ff |
unverified |
MIT (permissive) |
| arXiv:2024.findings-naacl.199 |
|
arumaekawa/DiLM/src/coreset/coreset_utils.py 499bced10bfc04a9 |
unverified |
MIT (permissive) |
| arXiv:2023.acl-long.426 |
|
umanlp/babelbert/utils/functions.py 4ebe45034029f9ba |
unverified |
MIT (permissive) |