Home › Code › get_embeddings

get_embeddings

Syntologyentry name in harvested coderead from the graph 2026-09-24

get_embeddings appears in the code Syntology harvested for 56 papers, as 47 distinct code bodies found in 58 places (a place is one code body under one paper). At least one of them ran in 19 of the papers; 0 of the code bodies carry a behaviour fingerprint.

What this page is not. Routines are grouped here by the exact string of their function or class name. Nothing asserts that two samples named get_embeddings do the same thing, share code, or are comparable; the name is a string, not an identity. Behaviour outputs (what a fingerprinted sample returned on the shared battery) are not in this export and are not shown here; the graph at syntology.ai holds them. "Ran" means executed on a synthesized fixture, not that the code is correct or reproduces a paper.

Samples Syntology

Syntology ran 14 of the 47 distinct code bodies named get_embeddings; 33 are unverified. One tile per status, in the site's fixed vocabulary, each code body counted once:

0ran · honoured contract
0ran · violated contract
4ran · our draft was wrong
1ran · fixture could not drive it
9ran
33unverified
0fingerprinted

Licence is a property of each copy, so it is counted per place: 17 of the 58 places are pointer only (Syntology does not serve that copy's text). This site shows no code text for any sample; every row below links to the file in its repository where the record names one.

“Ran” means the sample executed on a synthesized input; it does not mean the output is correct. “Honoured” and “violated” refer to a contract Syntology drafted from the code itself; “our draft was wrong” and “fixture could not drive it” are failures of Syntology's instrument, not of the code, and those samples did run. The ran count above is every status except unverified, the same rule as each paper page.

Papers

56 papers shown of 56, newest first; 58 places in the table. A paper with no recorded date is placed by the month its arXiv id encodes, shown in the Date column as YYYY-MM (from id). One row per place: a paper whose repository defines the name more than once appears more than once, and the same code body held for several papers appears once under each, with the same status. Titles and dates are the archive's archive 2025-07-28 for papers in the archive, and the graph's for 6 papers added by Syntology; 4 papers have no page here and are shown by arXiv id only. Status and fingerprint are Syntology's record of each code body; licence is recorded for each place. The File cell ends with the code body's code_sha256, Syntology's identity for that exact code: an agent fetches the stored sample with Syntology's MCP tool get_code(code_sha256="…") (how to connect); click an identity to copy that call.

PaperDateFileStatus SyntologyLicence
TAO-Attack: Toward Advanced Optimization-Based Jailbreak Attacks for Large Language Models added by Syntology 2026-03 (from id) ZevineXu/TAO-Attack/llm_attacks/base/attack_manager.py f556dda0785f3c77 unverified MIT (permissive)
Improving Set Function Approximation with Quasi-Arithmetic Neural Networks added by Syntology 2026-02 (from id) tomastokar/Quasi-Arithmetic-Neural-Networks/MNIST_qualitative.py d4aa88c9ae85da11 unverified no licence file found · pointer only
CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval added by Syntology 2026-01 (from id) idirlab/CaseFacts/experiments/check_similarity.py a5a629d94ddef395 unverified no licence file found · pointer only
Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers added by Syntology 2026-01 (from id) glad-lab/MGEO/attack_autodan.py 19c802b4dd21c8b8 unverified no licence file found · pointer only
CLINIC: Evaluating Multilingual Trustworthiness in Language Models for Healthcare added by Syntology 2025-12 (from id) AikyamLab/clinic/evaluation/adverserial/adv_eval.py e51b60b9add827b6 unverified MIT (permissive)
CLINIC: Evaluating Multilingual Trustworthiness in Language Models for Healthcare added by Syntology 2025-12 (from id) AikyamLab/clinic/evaluation/consistency/consistency_eval.py daf8bacd15f6bdea unverified MIT (permissive)
Building Data-Driven Occupation Taxonomies: A Bottom-Up Multi-Stage Approach via Semantic Clustering and Multi-Agent Collaboration added by Syntology 2025-09 (from id) aida-ugent/CLIMB/src/get_embeddings.py 5eb54fb92e768705 unverified licence not identified · pointer only
Breaking Free from MMI: A New Frontier in Rationalization by Probing Input Utilization 8 Mar 2025 jugechengzi/Rationalization-N2R/embedding.py a7fd4428a7a26977 ran MIT (permissive)
MERGE$^3$: Efficient Evolutionary Merging on Consumer-grade GPUs 9 Feb 2025 tommasomncttn/merge3/src/mergenetic/analysis/representation.py 1286c013573044fc unverified MIT (permissive)
WyckoffDiff -- A Generative Diffusion Model for Crystal Symmetry 10 Feb 2025 httk/wyckoffdiff/wyckoff_generation/evaluation/compute_fwd.py 408fa913477a97c5 unverified MIT (permissive)
LLMs Can Simulate Standardized Patients via Agent Coevolution 16 Dec 2024 zjumai/evopatient/embedding_function/sentence_embedding.py fc512197798fe2a7 unverified no licence file found · pointer only
MC-LLaVA: Multi-Concept Personalized Vision-Language Model 18 Nov 2024 arctanxarc/mc-llava/train/train_joint.py 5db573c804b47ea8 ran · our draft was wrong MIT (permissive)
Understanding Scaling Laws with Statistical and Approximation Theory for Transformer Neural Networks on Intrinsically Low-dimensional Data 11 Nov 2024 dahoas/transformer_manifolds_learning/embeddings.py a4824be70cdf2b77 ran · our draft was wrong no licence file found · pointer only
BendVLM: Test-Time Debiasing of Vision-Language Embeddings 7 Nov 2024 waltergerych/bend_vlm/bend_utils.py 050b743339edef0b unverified no licence file found · pointer only
Robust AI-Generated Text Detection by Restricted Embeddings 10 Oct 2024 silversolver/robustatd/fit_eraser_probing_tasks.py e99730bed9668d68 unverified no licence file found · pointer only
Is the MMI Criterion Necessary for Interpretability? Degenerating Non-causal Features to Plain Noise for Self-Rationalization 8 Oct 2024 jugechengzi/Rationalization-MRD/embedding.py a7fd4428a7a26977 ran MIT (permissive)
Infer Human's Intentions Before Following Natural Language Instructions 26 Sep 2024 simon-wan/fiser/networks/embeddings.py 16b037202af1a18b ran Apache-2.0 (permissive)
PROSE-FD: A Multimodal PDE Foundation Model for Learning Multiple Operators for Forecasting Fluid Dynamics 15 Sep 2024 felix-lyx/prose/prose_fd/models/attention_utils.py df45f4439f402a46 ran MIT (permissive)
Unifying Causal Representation Learning with the Invariance Principle 4 Sep 2024 causallearningai/istant/src/model.py c2c6795c770350e2 unverified MIT (permissive)
AgentPoison: Red-teaming LLM Agents via Poisoning Memory or Knowledge Bases 17 Jul 2024 BillChan226/AgentPoison/algo/utils.py 655e5409659cee36 unverified MIT (permissive)
Evidential Concept Embedding Models: Towards Reliable Concept Explanations for Skin Disease Diagnosis 27 Jun 2024 obiyoag/evi-cem/learn_cavs.py 6cd52c2775a843cb ran Apache-2.0 (permissive)
Improved Few-Shot Jailbreaking Can Circumvent Aligned Language Models and Their Defenses 3 Jun 2024 sail-sg/I-FSJ/llm_attacks/base/attack_manager.py f556dda0785f3c77 unverified MIT (permissive)
Improved Techniques for Optimization-Based Jailbreaking on Large Language Models 31 May 2024 jiaxiaojunqaq/i-gcg/llm_attacks/base/attack_manager.py f556dda0785f3c77 unverified no licence file found · pointer only
Quantitative Certification of Bias in Large Language Models 29 May 2024 uiuc-focal-lab/LLMCert-B/utils.py e00307369c36acc3 ran AGPL-3.0 (copyleft) · pointer only
Protecting Your LLMs with Information Bottleneck 22 Apr 2024 llm-attacks/llm-attacks/llm_attacks/base/attack_manager.py f556dda0785f3c77 unverified MIT (permissive)
No "Zero-Shot" Without Exponential Data: Pretraining Concept Frequency Determines Multimodal Model Performance 4 Apr 2024 bethgelab/frequency_determines_performance/src/retrieval_eval.py 507a009866a94a6f ran MIT (permissive)
DiLM: Distilling Dataset into Language Model for Text-level Dataset Distillation 30 Mar 2024 arumaekawa/dilm/src/coreset/coreset_utils.py 499bced10bfc04a9 unverified MIT (permissive)
Automatic and Universal Prompt Injection Attacks against Large Language Models 7 Mar 2024 sheltonliu-n/universal-prompt-injection/utils/opt_utils.py f556dda0785f3c77 unverified MIT (permissive)
Accelerating Greedy Coordinate Gradient and General Prompt Optimization via Probe Sampling 2 Mar 2024 zhaoyiran924/probe-sampling/llm_attacks/base/attack_manager.py db2b1708ea3c49ed unverified MIT (permissive)
GISTEmbed: Guided In-sample Selection of Training Negatives for Text Embedding Fine-tuning 26 Feb 2024 avsolatorio/gistembed/gist_embed/trainer/loss.py 7f91a871ff0d03ed ran no licence file found · pointer only
Defending LLMs against Jailbreaking Attacks via Backtranslation 26 Feb 2024 yihanwang617/llm-jailbreaking-defense-backtranslation/GCG/llm_attacks/base/attack_manager.py f556dda0785f3c77 unverified BSD-3-Clause (permissive)
Query Augmentation by Decoding Semantics from Brain Signals 24 Feb 2024 yeziyi1998/brain-query-augmentation/ict/model_utils/model_config.py bde4c1a685708c16 unverified MIT (permissive)
Learning to Poison Large Language Models for Downstream Manipulation 21 Feb 2024 rookiezxy/gbtl/llm_attacks/base/attack_manager.py a42389ea4ae4a27a unverified no licence file found · pointer only
Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia 8 Feb 2024 solidshen/ripple_official/src/utils/attack_manager.py 708a052bc12dad60 unverified no licence file found · pointer only
Robust Prompt Optimization for Defending Language Models Against Jailbreaking Attacks 30 Jan 2024 lapisrocks/rpo/rpo/opt_utils.py 9a3d93b72c5ec034 ran · fixture could not drive it no licence file found · pointer only
GOAt: Explaining Graph Neural Networks via Graph Output Attribution 26 Jan 2024 sluxsr/GOAt/node_classi_plot_pack/nc_plots.py 1ddb0751a34af6db ran no licence file found · pointer only
Enhancing the Rationale-Input Alignment for Self-explaining Rationalization 7 Dec 2023 jugechengzi/dar/embedding.py a7fd4428a7a26977 ran MIT (permissive)
Hijacking Large Language Models via Adversarial In-Context Learning 16 Nov 2023 RookieZxy/GGI-attack/GGI-attack/llm_attacks/base/attack_manager.py 33377834cc7d0c84 unverified no licence file found · pointer only
AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models 23 Oct 2023 rotaryhammer/code-autodan/autodan/autodan/attack_manager.py 01abe3bd48ba14fc unverified MIT (permissive)
D-Separation for Causal Self-Explanation 23 Sep 2023 jugechengzi/Rationalization-MCD/embedding.py a7fd4428a7a26977 ran MIT (permissive)
Adversarial Illusions in Multi-Modal Embeddings 22 Aug 2023 ebagdasa/adversarial_illusions/dataset_utils.py 996d7797acf3500c ran MIT (permissive)
Still No Lie Detector for Language Models: Probing Empirical and Conceptual Roadblocks 30 Jun 2023 balevinstein/probes/Train_CCSProbe.py 1fd11bf75e70a151 ran · our draft was wrong MIT (permissive)
NOTABLE: Transferable Backdoor Attacks Against Prompt-based NLP Models 28 May 2023 ru-system-software-and-security/notable/autoprompt/create_trigger.py 20a9293a820f617d ran · our draft was wrong no licence file found · pointer only
Augmentation-Adapted Retriever Improves Generalization of Language Models as Generic Plug-In 27 May 2023 openai/chatgpt-retrieval-plugin/services/openai.py b009eb83df585f1c unverified MIT (permissive)
Discover and Cure: Concept-aware Mitigation of Spurious Correlation 1 May 2023 Wuyxin/DISC/disc/concept_utils/cav_utils.py 87689c095e68cee7 unverified MIT (permissive)
Investigating the Effectiveness of Task-Agnostic Prefix Prompt for Instruction Following 28 Feb 2023 seonghyeonye/icil/src/run_nearest_demo.py 73680032c6cc1a34 unverified MIT (permissive)
Cluster & Tune: Boost Cold Start Performance in Text Classification 20 Mar 2022 ibm/intermediate-training-using-clustering/run_experiment.py c1fcf89801c4ea9e unverified Apache-2.0 (permissive)
A Neural Network Solves, Explains, and Generates University Math Problems by Program Synthesis and Few-Shot Learning at Human Level 31 Dec 2021 idrori/mathq/code/embedding.py caf8168c846f56d5 unverified MIT (permissive)
Virtual Augmentation Supported Contrastive Learning of Sentence Representations 16 Oct 2021 amazon-science/sentence-representations/DownstreamEval/clustering/clustering_eval.py 13e3886c6d762ec6 unverified Apache-2.0 (permissive)
AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated Prompts 29 Oct 2020 ucinlp/autoprompt/autoprompt/create_trigger.py 20a9293a820f617d ran · our draft was wrong Apache-2.0 (permissive)
ProtTrans: Towards Cracking the Language of Life's Code Through Self-Supervised Deep Learning and High Performance Computing 13 Jul 2020 agemagician/ProtTrans/Embedding/prott5_embedder.py dd492312bc93c287 unverified MIT (permissive)
The Oxford Radar RobotCar Dataset: A Radar Extension to the Oxford RobotCar Dataset 2019-09 (from id) mttgdd/oord-dataset/src/compute_distance_matrix.py adbf95aba4ffd8f5 unverified MIT (permissive)
Show, Attend and Tell: Neural Image Caption Generation with Visual Attention 10 Feb 2015 LinXueyuanStdio/LaTeX_OCR/model/decoder.py 9b24ffeda92bc6e1 unverified Apache-2.0 (permissive)
arXiv:aaai_16580 modriczhang/HRL-Rec/layer_util.py fb6bc9cda68a32c1 unverified Apache-2.0 (permissive)
arXiv:2024.findings-naacl.294 balevinstein/Probes/Train_CCSProbe.py 1fd11bf75e70a151 ran · our draft was wrong MIT (permissive)
arXiv:2024.findings-naacl.294 balevinstein/Probes/Generate_CCS_predictions.py e4680e27c1f698ff unverified MIT (permissive)
arXiv:2024.findings-naacl.199 arumaekawa/DiLM/src/coreset/coreset_utils.py 499bced10bfc04a9 unverified MIT (permissive)
arXiv:2023.acl-long.426 umanlp/babelbert/utils/functions.py 4ebe45034029f9ba unverified MIT (permissive)

This site shows no code text; each File cell links to the file on GitHub at the repository's current default branch, which may have changed since the harvest. "Pointer only" means Syntology does not serve that copy's text, for one of four reasons: no licence file was found; the licence was not identified; the licence is recorded as permissive but that copy's record is not marked cleared; or the licence is outside the permissive list Syntology serves text under (MIT, Apache-2.0, BSD and similar). Some licences outside that list permit redistribution, such as WTFPL, and GPL-3.0 under its conditions; they are simply not on the list. Hover a licence cell for the reason. Per-sample records for a paper are on its paper page under "Code Syntology ran".

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections