Browse State-of-the-Art › Semantic Textual Similarity › Papers, page 8
Semantic Textual Similarity
Papers archive 2025-07-28
archive papers tagged: 2,381 · with a code link: 693 · where Syntology ran a sample: 144 (118 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (144 of 2,381 tagged: 118 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument)
Page 8 of 24: papers 701 to 800 of 2,381, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
PrivacyXray: Detecting Privacy Breaches in LLMs through Semantic Consistency and Probability Certainty24 Jun 2025 0 repositories listed
-
Semantic similarity estimation for domain specific data using BERT and other techniques23 Jun 2025 0 repositories listed
-
GrFormer: A Novel Transformer on Grassmann Manifold for Infrared and Visible Image Fusion17 Jun 2025 0 repositories listed
-
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking17 Jun 2025 0 repositories listed
-
Similarity = Value? Consultation Value Assessment and Alignment for Personalized Search17 Jun 2025 0 repositories listed
-
FindMeIfYouCan: Bringing Open Set metrics to near, far and farther Out-of-Distribution Object Detection16 Jun 2025 0 repositories listed
-
Conservative Bias in Large Language Models: Measuring Relation Predictions9 Jun 2025 0 repositories listed
-
Hierarchical Scoring with 3D Gaussian Splatting for Instance Image-Goal Navigation9 Jun 2025 0 repositories listed
-
Statistical Hypothesis Testing for Auditing Robustness in Language Models9 Jun 2025 0 repositories listed
-
Denoising Programming Knowledge Tracing with a Code Graph-based Tuning Adaptor7 Jun 2025 0 repositories listed
-
Plugging Schema Graph into Multi-Table QA: A Human-Guided Framework for Reducing LLM Reliance4 Jun 2025 0 repositories listed
-
MCP-Zero: Active Tool Discovery for Autonomous LLM Agents1 Jun 2025 0 repositories listed
-
GATE: General Arabic Text Embedding for Enhanced Semantic Textual Similarity with Matryoshka Representation Learning and Hybrid Loss Training30 May 2025 0 repositories listed
-
VUDG: A Dataset for Video Understanding Domain Generalization30 May 2025 0 repositories listed
-
Document Valuation in LLM Summaries: A Cluster Shapley Approach28 May 2025 0 repositories listed
-
Improving Brain-to-Image Reconstruction via Fine-Grained Text Bridging28 May 2025 0 repositories listed
-
LLMs as Better Recommenders with Natural Language Collaborative Signals: A Self-Assessing Retrieval Approach26 May 2025 0 repositories listed
-
CrosGrpsABS: Cross-Attention over Syntactic and Semantic Graphs for Aspect-Based Sentiment Analysis in a Low-Resource Language25 May 2025 0 repositories listed
-
Sparse VideoGen2: Accelerate Video Generation with Sparse Attention via Semantic-Aware Permutation24 May 2025 0 repositories listed
-
Automated Feedback Loops to Protect Text Simplification with Generative AI from Information Loss22 May 2025 0 repositories listed
-
Omni TM-AE: A Scalable and Interpretable Embedding Model Using the Full Tsetlin Machine State Space22 May 2025 0 repositories listed
-
EcomScriptBench: A Multi-task Benchmark for E-commerce Script Planning via Step-wise Intention-Driven Product Association21 May 2025 0 repositories listed
-
Language Specific Knowledge: Do Models Know Better in X than in English?21 May 2025 0 repositories listed
-
Community Search in Time-dependent Road-social Attributed Networks18 May 2025 0 repositories listed
-
Fine-Grained ECG-Text Contrastive Learning via Waveform Understanding Enhancement17 May 2025 0 repositories listed
-
AI-enhanced semantic feature norms for 786 concepts15 May 2025 0 repositories listed
-
Evaluations at Work: Measuring the Capabilities of GenAI in Use15 May 2025 0 repositories listed
-
FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation15 May 2025 0 repositories listed
-
Towards Automated Situation Awareness: A RAG-Based Framework for Peacebuilding Reports14 May 2025 0 repositories listed
-
A 2D Semantic-Aware Position Encoding for Vision Transformers14 May 2025 0 repositories listed
-
TrialMatchAI: An End-to-End AI-powered Clinical Trial Recommendation System to Streamline Patient-to-Trial Matching13 May 2025 0 repositories listed
-
Hypernym Mercury: Token Optimization Through Semantic Field Constriction And Reconstruction From Hypernyms. A New Text Compression Method12 May 2025 0 repositories listed
-
Jailbreaking the Text-to-Video Generative Models10 May 2025 0 repositories listed
-
Estimating Quality in Therapeutic Conversations: A Multi-Dimensional Natural Language Processing Framework9 May 2025 0 repositories listed
-
Sparse Attention Remapping with Clustering for Efficient LLM Decoding on PIM9 May 2025 0 repositories listed
-
Stealthy LLM-Driven Data Poisoning Attacks Against Embedding-Based Retrieval-Augmented Recommender Systems8 May 2025 0 repositories listed
-
R&B: Domain Regrouping and Data Mixture Balancing for Efficient Foundation Model Training1 May 2025 0 repositories listed
-
Homa at SemEval-2025 Task 5: Aligning Librarian Records with OntoAligner for Subject Tagging30 Apr 2025 0 repositories listed
-
Retrieval-Enhanced Few-Shot Prompting for Speech Event Extraction30 Apr 2025 0 repositories listed
-
ReCellTy: Domain-specific knowledge graph retrieval-augmented LLMs workflow for single-cell annotation24 Apr 2025 0 repositories listed
-
Cyc3D: Fine-grained Controllable 3D Generation via Cycle Consistency Regularization21 Apr 2025 0 repositories listed
-
Stay Hungry, Stay Foolish: On the Extended Reading Articles Generation with LLMs21 Apr 2025 0 repositories listed
-
Exploring Language Patterns of Prompts in Text-to-Image Generation and Their Impact on Visual Diversity19 Apr 2025 0 repositories listed
-
Semantic Similarity-Informed Bayesian Borrowing for Quantitative Signal Detection of Adverse Events16 Apr 2025 0 repositories listed
-
Self-Controlled Dynamic Expansion Model for Continual Learning14 Apr 2025 0 repositories listed
-
HD-RAG: Retrieval-Augmented Generation for Hybrid Documents Containing Text and Hierarchical Tables13 Apr 2025 0 repositories listed
-
Embodied Image Captioning: Self-supervised Learning Agents for Spatially Coherent Image Descriptions11 Apr 2025 0 repositories listed
-
Evaluating Retrieval Augmented Generative Models for Document Queries in Transportation Safety9 Apr 2025 0 repositories listed
-
Balancing Complexity and Informativeness in LLM-Based Clustering: Finding the Goldilocks Zone6 Apr 2025 0 repositories listed
-
Horizon Scans can be accelerated using novel information retrieval and artificial intelligence tools2 Apr 2025 0 repositories listed
-
ProtoGuard-guided PROPEL: Class-Aware Prototype Enhancement and Progressive Labeling for Incremental 3D Point Cloud Segmentation2 Apr 2025 0 repositories listed
-
Context-Aware Human Behavior Prediction Using Multimodal Large Language Models: Challenges and Insights1 Apr 2025 0 repositories listed
-
SentenceKV: Efficient LLM Inference via Sentence-Level Semantic KV Caching1 Apr 2025 0 repositories listed
-
Beyond Detection: Designing AI-Resilient Assessments with Automated Feedback Tool to Foster Critical Thinking30 Mar 2025 0 repositories listed
-
Discovering Knowledge Deficiencies of Language Models on Massive Knowledge Base30 Mar 2025 0 repositories listed
-
A Quantitative Approach to Evaluating Open-Source EHR Systems for Indian Healthcare27 Mar 2025 0 repositories listed
-
HyperFree: A Channel-adaptive and Tuning-free Foundation Model for Hyperspectral Remote Sensing Imagery27 Mar 2025 0 repositories listed
-
BeLightRec: A lightweight recommender system enhanced with BERT26 Mar 2025 0 repositories listed
-
CausalRAG: Integrating Causal Graphs into Retrieval-Augmented Generation25 Mar 2025 0 repositories listed
-
SeLIP: Similarity Enhanced Contrastive Language Image Pretraining for Multi-modal Head MRI25 Mar 2025 0 repositories listed
-
Unleashing the power of text for credit default prediction: Comparing human-written and generative AI-refined texts23 Mar 2025 0 repositories listed
-
CASE -- Condition-Aware Sentence Embeddings for Conditional Semantic Textual Similarity Measurement21 Mar 2025 0 repositories listed
-
Vision Transformer Based Semantic Communications for Next Generation Wireless Networks21 Mar 2025 0 repositories listed
-
KVShare: An LLM Service System with Efficient and Effective Multi-Tenant KV Cache Reuse17 Mar 2025 0 repositories listed
-
A General Close-loop Predictive Coding Framework for Auditory Working Memory16 Mar 2025 0 repositories listed
-
Measuring Similarity in Causal Graphs: A Framework for Semantic and Structural Analysis14 Mar 2025 0 repositories listed
-
Are We Truly Forgetting? A Critical Re-examination of Machine Unlearning Evaluation Protocols10 Mar 2025 0 repositories listed
-
AuthorMist: Evading AI Text Detectors with Reinforcement Learning10 Mar 2025 0 repositories listed
-
MIGA: Mutual Information-Guided Attack on Denoising Models for Semantic Manipulation10 Mar 2025 0 repositories listed
-
AutoTestForge: A Multidimensional Automated Testing Framework for Natural Language Processing Models7 Mar 2025 0 repositories listed
-
Improving RAG Retrieval via Propositional Content Extraction: a Speech Act Theory Approach7 Mar 2025 0 repositories listed
-
SEOE: A Scalable and Reliable Semantic Evaluation Framework for Open Domain Event Detection5 Mar 2025 0 repositories listed
-
Token-Level Privacy in Large Language Models5 Mar 2025 0 repositories listed
-
Language-agnostic, automated assessment of listeners' speech recall using large language models2 Mar 2025 0 repositories listed
-
Statistical Mechanics of Semantic Compression1 Mar 2025 0 repositories listed
-
TempRetriever: Fusion-based Temporal Dense Passage Retrieval for Time-Sensitive Questions28 Feb 2025 0 repositories listed
-
Towards Label-Only Membership Inference Attack against Pre-trained Large Language Models26 Feb 2025 0 repositories listed
-
EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models25 Feb 2025 0 repositories listed
-
How Vital is the Jurisprudential Relevance: Law Article Intervened Legal Case Retrieval and Matching25 Feb 2025 0 repositories listed
-
ATEB: Evaluating and Improving Advanced NLP Tasks for Text Embedding Models24 Feb 2025 0 repositories listed
-
Uncertainty Quantification of Large Language Models through Multi-Dimensional Responses24 Feb 2025 0 repositories listed
-
Constructing a Norm for Children's Scientific Drawing: Distribution Features Based on Semantic Similarity of Large Language Models21 Feb 2025 0 repositories listed
-
A Meta-Evaluation of Style and Attribute Transfer Metrics20 Feb 2025 0 repositories listed
-
DeepRTL: Bridging Verilog Understanding and Generation with a Unified Representation Model20 Feb 2025 0 repositories listed
-
Evolutionary Algorithms Approach For Search Based On Semantic Document Similarity20 Feb 2025 0 repositories listed
-
Breaking the Clusters: Uniformity-Optimization for Text-Based Sequential Recommendation19 Feb 2025 0 repositories listed
-
Event Segmentation Applications in Large Language Model Enabled Automated Recall Assessments19 Feb 2025 0 repositories listed
-
HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation18 Feb 2025 0 repositories listed
-
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models17 Feb 2025 0 repositories listed
-
FaMTEB: Massive Text Embedding Benchmark in Persian Language17 Feb 2025 0 repositories listed
-
PropNet: a White-Box and Human-Like Network for Sentence Representation15 Feb 2025 0 repositories listed
-
Examining Multilingual Embedding Models Cross-Lingually Through LLM-Generated Adversarial Examples12 Feb 2025 0 repositories listed
-
PDV: Prompt Directional Vectors for Zero-shot Composed Image Retrieval11 Feb 2025 0 repositories listed
-
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering10 Feb 2025 0 repositories listed
-
Enhancing Health Information Retrieval with RAG by Prioritizing Topical Relevance and Factual Accuracy7 Feb 2025 0 repositories listed
-
Detecting Backdoor Attacks via Similarity in Semantic Communication Systems6 Feb 2025 0 repositories listed
-
How does a Multilingual LM Handle Multiple Languages?6 Feb 2025 0 repositories listed
-
How do Humans and Language Models Reason About Creativity? A Comparative Analysis5 Feb 2025 0 repositories listed
-
HSI: A Holistic Style Injector for Arbitrary Style Transfer5 Feb 2025 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.