Methods › Natural Language Processing › Subword Segmentation › WordPiece › Papers, page 7
WordPiece
Papers archive 2025-07-28
archive papers tagged: 7,063 · with a code link: 2,910 · where Syntology ran a sample: 650 (529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,063 tagged: 529 with a run with no instrument failure, 121 where every run was a failure of Syntology's instrument)
Page 7 of 71: papers 601 to 700 of 7,063, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
FoQA: A Faroese Question-Answering Dataset 11 Feb 2025 · 0 repositories · arXiv:2502.07642
-
Graph RAG-Tool Fusion 11 Feb 2025 · 1 repository · arXiv:2502.07223
-
Making Language Models Robust Against Negation 11 Feb 2025 · 0 repositories · arXiv:2502.07717
-
C-3PO: Compact Plug-and-Play Proxy Optimization to Achieve Human-like Retrieval-Augmented Generation 10 Feb 2025 · 0 repositories · arXiv:2502.06205Syntology 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
ConMeC: A Dataset for Metonymy Resolution with Common Nouns 10 Feb 2025 · 1 repository · arXiv:2502.06087
-
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights 10 Feb 2025 · 0 repositories · arXiv:2502.07049
-
Multimodal Task Representation Memory Bank vs. Catastrophic Forgetting in Anomaly Detection 10 Feb 2025 · 0 repositories · arXiv:2502.06194
-
Optimizing Knowledge Integration in Retrieval-Augmented Generation with Self-Selection 10 Feb 2025 · 0 repositories · arXiv:2502.06148
-
RALLRec: Improving Retrieval Augmented Large Language Model Recommendation with Representation Learning 10 Feb 2025 · 1 repository · arXiv:2502.06101
-
Towards Copyright Protection for Knowledge Bases of Retrieval-augmented Language Models via Reasoning 10 Feb 2025 · 0 repositories · arXiv:2502.10440
-
Enhancing Financial Time-Series Forecasting with Retrieval-Augmented Large Language Models 9 Feb 2025 · 0 repositories · arXiv:2502.05878
-
APE: Faster and Longer Context-Augmented Generation via Adaptive Parallel Encoding 8 Feb 2025 · 1 repository · arXiv:2502.05431Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Knowledge Graph-Guided Retrieval Augmented Generation 8 Feb 2025 · 1 repository · arXiv:2502.06864
-
On the Effectiveness of Large Language Models in Automating Categorization of Scientific Texts 8 Feb 2025 · 0 repositories · arXiv:2502.15745
-
Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey 8 Feb 2025 · 1 repository · arXiv:2502.06872
-
Efficient Knowledge Feeding to Language Models: A Novel Integrated Encoder-Decoder Architecture 7 Feb 2025 · 0 repositories · arXiv:2502.05233
-
Probing Internal Representations of Multi-Word Verbs in Large Language Models 7 Feb 2025 · 0 repositories · arXiv:2502.04789
-
A Classification System Approach in Predicting Chinese Censorship 6 Feb 2025 · 0 repositories · arXiv:2502.04234
-
Experiments with Large Language Models on Retrieval-Augmented Generation for Closed-Source Simulation Software 6 Feb 2025 · 0 repositories · arXiv:2502.03916
-
LLMs to Support a Domain Specific Knowledge Assistant 6 Feb 2025 · 0 repositories · arXiv:2502.04095
-
MD-BERT: Action Recognition in Dark Videos via Dynamic Multi-Stream Fusion and Temporal Modeling 6 Feb 2025 · 1 repository · arXiv:2502.03724
-
MedRAG: Enhancing Retrieval-augmented Generation with Knowledge Graph-Elicited Reasoning for Healthcare Copilot 6 Feb 2025 · 1 repository · arXiv:2502.04413
-
MRAMG-Bench: A Comprehensive Benchmark for Advancing Multimodal Retrieval-Augmented Multimodal Generation 6 Feb 2025 · 1 repository · arXiv:2502.04176
-
Cache-Craft: Managing Chunk-Caches for Efficient Retrieval-Augmented Generation 5 Feb 2025 · 0 repositories · arXiv:2502.15734
-
MARAGE: Transferable Multi-Model Adversarial Attack for Retrieval-Augmented Generation Data Extraction 5 Feb 2025 · 0 repositories · arXiv:2502.04360
-
OPTIC: Optimizing Patient-Provider Triaging & Improving Communications in Clinical Operations using GPT-4 Data Labeling and Model Distillation 5 Feb 2025 · 0 repositories · arXiv:2503.05701
-
Path Planning for Masked Diffusion Model Sampling 5 Feb 2025 · 0 repositories · arXiv:2502.03540
-
Open Foundation Models in Healthcare: Challenges, Paradoxes, and Opportunities with GenAI Driven Personalized Prescription 4 Feb 2025 · 0 repositories · arXiv:2502.04356
-
OverThink: Slowdown Attacks on Reasoning LLMs 4 Feb 2025 · 1 repository · arXiv:2502.02542
-
Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation 4 Feb 2025 · 1 repository · arXiv:2502.02464
-
Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions 4 Feb 2025 · 0 repositories · arXiv:2502.18470
-
Topic Modeling in Marathi 4 Feb 2025 · 0 repositories · arXiv:2502.02100
-
BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation 3 Feb 2025 · 0 repositories · arXiv:2502.01697
-
GFM-RAG: Graph Foundation Model for Retrieval Augmented Generation 3 Feb 2025 · 1 repository · arXiv:2502.01113Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
Meursault as a Data Point 3 Feb 2025 · 0 repositories · arXiv:2502.01364
-
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models 3 Feb 2025 · 0 repositories · arXiv:2502.01386
-
VideoRAG: Retrieval-Augmented Generation with Extreme Long-Context Videos 3 Feb 2025 · 1 repository · arXiv:2502.01549
-
Explainability in Practice: A Survey of Explainable NLP Across Various Domains 2 Feb 2025 · 0 repositories · arXiv:2502.00837
-
LIBRA: Measuring Bias of Large Language Model from a Local Context 2 Feb 2025 · 0 repositories · arXiv:2502.01679
-
Riddle Me This! Stealthy Membership Inference for Retrieval-Augmented Generation 1 Feb 2025 · 1 repository · arXiv:2502.00306
-
Through the Looking Glass: LLM-Based Analysis of AR/VR Android Applications Privacy Policies 31 Jan 2025 · 0 repositories · arXiv:2501.19223
-
AlphaAdam:Asynchronous Masked Optimization with Dynamic Alpha for Selective Updates 30 Jan 2025 · 0 repositories · arXiv:2501.18094
-
Can we Retrieve Everything All at Once? ARM: An Alignment-Oriented LLM-based Retrieval Method 30 Jan 2025 · 0 repositories · arXiv:2501.18539
-
General Embedding vs. Task-Specific Embedding: A Comparative Approach to Enhancing NLP Performance 30 Jan 2025 · 0 repositories
-
Israel-Hamas war through Telegram, Reddit and Twitter 30 Jan 2025 · 0 repositories · arXiv:2502.00060
-
Leveraging LLM Agents for Automated Optimization Modeling for SASP Problems: A Graph-RAG based Approach 30 Jan 2025 · 0 repositories · arXiv:2501.18320
-
RbFT: Robust Fine-tuning for Retrieval-Augmented Generation against Retrieval Defects 30 Jan 2025 · 1 repository · arXiv:2501.18365
-
Retrieval Augmented Generation Based LLM Evaluation For Protocol State Machine Inference With Chain-of-Thought Reasoning 30 Jan 2025 · 0 repositories · arXiv:2502.15727
-
Scalable and Cost-Efficient ML Inference: Parallel Batch Processing with Serverless Functions 30 Jan 2025 · 0 repositories · arXiv:2502.12017
-
Leveraging In-Context Learning and Retrieval-Augmented Generation for Automatic Question Generation in Educational Domains 29 Jan 2025 · 0 repositories · arXiv:2501.17397
-
Attribution analysis of legal language as used by LLM 28 Jan 2025 · 0 repositories · arXiv:2501.17330
-
Balancing Content Size in RAG-Text2SQL System 28 Jan 2025 · 0 repositories · arXiv:2502.15723
-
Detecting harassment and defamation in cyberbullying with emotion-adaptive training 28 Jan 2025 · 1 repository · arXiv:2501.16925
-
Multiple Abstraction Level Retrieve Augment Generation 28 Jan 2025 · 0 repositories · arXiv:2501.16952
-
SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model 28 Jan 2025 · 1 repository · arXiv:2501.18636
-
A Comprehensive Study on Fine-Tuning Large Language Models for Medical Question Answering Using Classification Models and Comparative Analysis 27 Jan 2025 · 0 repositories · arXiv:2501.17190
-
Enhancing and Exploring Mild Cognitive Impairment Detection with W2V-BERT-2.0 27 Jan 2025 · 0 repositories · arXiv:2501.16201
-
LemmaHead: RAG Assisted Proof Generation Using Large Language Models 27 Jan 2025 · 0 repositories · arXiv:2501.15797
-
Parametric Retrieval Augmented Generation 27 Jan 2025 · 1 repository · arXiv:2501.15915
-
Provence: efficient and robust context pruning for retrieval-augmented generation 27 Jan 2025 · 0 repositories · arXiv:2501.16214
-
RelCAT: Advancing Extraction of Clinical Inter-Entity Relationships from Unstructured Electronic Health Records 27 Jan 2025 · 1 repository · arXiv:2501.16077
-
URAG: Implementing a Unified Hybrid RAG for Precise Answers in University Admission Chatbots -- A Case Study at HCMUT 27 Jan 2025 · 0 repositories · arXiv:2501.16276
-
Decentralized Low-Rank Fine-Tuning of Large Language Models 26 Jan 2025 · 0 repositories · arXiv:2501.15361
-
Advanced Real-Time Fraud Detection Using RAG-Based LLMs 25 Jan 2025 · 0 repositories · arXiv:2501.15290
-
ASRank: Zero-Shot Re-Ranking with Answer Scent for Document Retrieval 25 Jan 2025 · 0 repositories · arXiv:2501.15245
-
CG-RAG: Research Question Answering by Citation Graph Retrieval-Augmented LLMs 25 Jan 2025 · 0 repositories · arXiv:2501.15067
-
Improving Retrieval-Augmented Generation through Multi-Agent Reinforcement Learning 25 Jan 2025 · 1 repository · arXiv:2501.15228Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
TrustDataFilter:Leveraging Trusted Knowledge Base Data for More Effective Filtering of Unknown Information 25 Jan 2025 · 0 repositories · arXiv:2502.15714
-
Causal Graphs Meet Thoughts: Enhancing Complex Reasoning in Graph-Augmented LLMs 24 Jan 2025 · 1 repository · arXiv:2501.14892
-
Chain-of-Retrieval Augmented Generation 24 Jan 2025 · 0 repositories · arXiv:2501.14342
-
Fast Think-on-Graph: Wider, Deeper and Faster Reasoning of Large Language Model on Knowledge Graph 24 Jan 2025 · 1 repository · arXiv:2501.14300
-
GraPPI: A Retrieve-Divide-Solve GraphRAG Framework for Large-scale Protein-protein Interaction Exploration 24 Jan 2025 · 1 repository · arXiv:2501.16382
-
Idiom Detection in Sorani Kurdish Texts 24 Jan 2025 · 0 repositories · arXiv:2501.14528
-
A Transformer-based Autoregressive Decoder Architecture for Hierarchical Text Classification 23 Jan 2025 · 1 repository · arXiv:2501.13598
-
CAPRAG: A Large Language Model Solution for Customer Service and Automatic Reporting using Vector and Graph Retrieval-Augmented Generation 23 Jan 2025 · 0 repositories · arXiv:2501.13993
-
GraphRAG under Fire 23 Jan 2025 · 0 repositories · arXiv:2501.14050
-
Multi-Level Attention and Contrastive Learning for Enhanced Text Classification with an Optimized Transformer 23 Jan 2025 · 0 repositories · arXiv:2501.13467
-
Retrievals Can Be Detrimental: A Contrastive Backdoor Attack Paradigm on Retrieval-Augmented Diffusion Models 23 Jan 2025 · 0 repositories · arXiv:2501.13340
-
RPO: Retrieval Preference Optimization for Robust Retrieval-Augmented Generation 23 Jan 2025 · 0 repositories · arXiv:2501.13726
-
StreamingRAG: Real-time Contextual Retrieval and Generation Framework 23 Jan 2025 · 0 repositories · arXiv:2501.14101
-
Adaptive Retrieval Without Self-Knowledge? Bringing Uncertainty Back Home 22 Jan 2025 · 0 repositories · arXiv:2501.12835Syntology 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 5 pointer-only (licence)
-
EvidenceMap: Learning Evidence Analysis to Unleash the Power of Small Language Models for Biomedical Question Answering 22 Jan 2025 · 0 repositories · arXiv:2501.12746
-
Generating Diverse Q&A Benchmarks for RAG Evaluation with DataMorgana 22 Jan 2025 · 0 repositories · arXiv:2501.12789
-
RAG-Reward: Optimizing RAG with Reward Modeling and RLHF 22 Jan 2025 · 0 repositories · arXiv:2501.13264
-
A Survey of Graph Retrieval-Augmented Generation for Customized Large Language Models 21 Jan 2025 · 1 repository · arXiv:2501.13958
-
Academic Case Reports Lack Diversity: Assessing the Presence and Diversity of Sociodemographic and Behavioral Factors related to Post COVID-19 Condition 21 Jan 2025 · 0 repositories · arXiv:2501.12538
-
ALoFTRAG: Automatic Local Fine Tuning for Retrieval Augmented Generation 21 Jan 2025 · 1 repository · arXiv:2501.11929
-
Assisting Mathematical Formalization with A Learning-based Premise Retriever 21 Jan 2025 · 1 repository · arXiv:2501.13959
-
Comparative Approaches to Sentiment Analysis Using Datasets in Major European and Arabic Languages 21 Jan 2025 · 0 repositories · arXiv:2501.12540
-
Med-R²: Crafting Trustworthy LLM Physicians via Retrieval and Reasoning of Evidence-Based Medicine 21 Jan 2025 · 1 repository · arXiv:2501.11885
-
Network-informed Prompt Engineering against Organized Astroturf Campaigns under Extreme Class Imbalance 21 Jan 2025 · 1 repository · arXiv:2501.11849
-
Explainable Lane Change Prediction for Near-Crash Scenarios Using Knowledge Graph Embeddings and Retrieval Augmented Generation 20 Jan 2025 · 0 repositories · arXiv:2501.11560
-
KEIR @ ECIR 2025: The Second Workshop on Knowledge-Enhanced Information Retrieval 20 Jan 2025 · 0 repositories · arXiv:2501.11499
-
PIKE-RAG: sPecIalized KnowledgE and Rationale Augmented Generation 20 Jan 2025 · 1 repository · arXiv:2501.11551Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Poison-RAG: Adversarial Data Poisoning Attacks on Retrieval-Augmented Generation in Recommender Systems 20 Jan 2025 · 1 repository · arXiv:2501.11759
-
Trustformer: A Trusted Federated Transformer 20 Jan 2025 · 0 repositories · arXiv:2501.11706
-
TutorLLM: Customizing Learning Recommendations with Knowledge Tracing and Retrieval-Augmented Generation 20 Jan 2025 · 0 repositories · arXiv:2502.15709
-
GEC-RAG: Improving Generative Error Correction via Retrieval-Augmented Generation for Automatic Speech Recognition Systems 18 Jan 2025 · 0 repositories · arXiv:2501.10734
-
Visual RAG: Expanding MLLM visual knowledge without fine-tuning 18 Jan 2025 · 0 repositories · arXiv:2501.10834
-
4bit-Quantization in Vector-Embedding for RAG 17 Jan 2025 · 1 repository · arXiv:2501.10534