Methods › Natural Language Processing › Language Models › LLaMA › Papers, page 4
LLaMA
Papers archive 2025-07-28
archive papers tagged: 1,062 · with a code link: 423 · where Syntology ran a sample: 143 (127 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (143 of 1,062 tagged: 127 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 4 of 11: papers 301 to 400 of 1,062, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TUMLU: A Unified and Native Language Understanding Benchmark for Turkic Languages 16 Feb 2025 · 1 repository · arXiv:2502.11020
-
Multilingual Encoder Knows more than You Realize: Shared Weights Pretraining for Extremely Low-Resource Languages 15 Feb 2025 · 1 repository · arXiv:2502.10852
-
Benchmarking the rationality of AI decision making using the transitivity axiom 14 Feb 2025 · 0 repositories · arXiv:2502.10554
-
Leveraging large language models for structured information extraction from pathology reports 14 Feb 2025 · 1 repository · arXiv:2502.12183
-
RoSTE: An Efficient Quantization-Aware Supervised Fine-Tuning Approach for Large Language Models 13 Feb 2025 · 0 repositories · arXiv:2502.09003Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
SQuARE: Sequential Question Answering Reasoning Engine for Enhanced Chain-of-Thought in Large Language Models 13 Feb 2025 · 1 repository · arXiv:2502.09390
-
The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Safety Analysis 13 Feb 2025 · 3 repositories · arXiv:2502.09674
-
Cancer Vaccine Adjuvant Name Recognition from Biomedical Literature using Large Language Models 12 Feb 2025 · 1 repository · arXiv:2502.09659
-
Inference-time sparse attention with asymmetric indexing 12 Feb 2025 · 0 repositories · arXiv:2502.08246
-
Automated Capability Discovery via Model Self-Exploration 11 Feb 2025 · 2 repositories · arXiv:2502.07577
-
FinRL-DeepSeek: LLM-Infused Risk-Sensitive Reinforcement Learning for Trading Agents 11 Feb 2025 · 2 repositories · arXiv:2502.07393
-
Forget What You Know about LLMs Evaluations - LLMs are Like a Chameleon 11 Feb 2025 · 1 repository · arXiv:2502.07445
-
Improving Adaptive Moment Optimization via Preconditioner Diagonalization 11 Feb 2025 · 0 repositories · arXiv:2502.07488
-
NanoVLMs: How small can we go and still make coherent Vision Language Models? 11 Feb 2025 · 0 repositories · arXiv:2502.07838
-
Towards Efficient Optimizer Design for LLM via Structured Fisher Approximation with a Low-Rank Extension 11 Feb 2025 · 0 repositories · arXiv:2502.07752
-
TransMLA: Multi-Head Latent Attention Is All You Need 11 Feb 2025 · 1 repository · arXiv:2502.07864
-
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring 10 Feb 2025 · 1 repository · arXiv:2502.07087
-
HSI: Head-Specific Intervention Can Induce Misaligned AI Coordination in Large Language Models 9 Feb 2025 · 1 repository · arXiv:2502.05945
-
On the Effectiveness of Large Language Models in Automating Categorization of Scientific Texts 8 Feb 2025 · 0 repositories · arXiv:2502.15745
-
Refining Positive and Toxic Samples for Dual Safety Self-Alignment of LLMs with Minimal Human Interventions 8 Feb 2025 · 0 repositories · arXiv:2502.08657
-
Can Large Language Models Understand Intermediate Representations? 7 Feb 2025 · 0 repositories · arXiv:2502.06854
-
Mitigating Unintended Memorization with LoRA in Federated Learning for LLMs 7 Feb 2025 · 1 repository · arXiv:2502.05087
-
Decoding AI Judgment: How LLMs Assess News Credibility and Bias 6 Feb 2025 · 0 repositories · arXiv:2502.04426
-
Identify Critical KV Cache in LLM Inference from an Output Perturbation Perspective 6 Feb 2025 · 2 repositories · arXiv:2502.03805Syntology official (archive's flag): 1 ran · 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 11 harvested samples)
-
Mediator: Memory-efficient LLM Merging with Less Parameter Conflicts and Uncertainty Based Routing 6 Feb 2025 · 0 repositories · arXiv:2502.04411
-
Probing a Vision-Language-Action Model for Symbolic States and Integration into a Cognitive Architecture 6 Feb 2025 · 0 repositories · arXiv:2502.04558
-
Entropy Adaptive Decoding: Dynamic Model Switching for Efficient Inference 5 Feb 2025 · 0 repositories · arXiv:2502.06833
-
On Zero-Initialized Attention: Optimal Prompt and Gating Factor Estimation 5 Feb 2025 · 0 repositories · arXiv:2502.03029
-
Evaluating the Effectiveness of LLMs in Fixing Maintainability Issues in Real-World Projects 4 Feb 2025 · 0 repositories · arXiv:2502.02368
-
LV-XAttn: Distributed Cross-Attention for Long Visual Inputs in Multimodal Large Language Models 4 Feb 2025 · 0 repositories · arXiv:2502.02406
-
An Agentic AI Workflow for Detecting Cognitive Concerns in Real-world Data 3 Feb 2025 · 0 repositories · arXiv:2502.01789
-
BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation 3 Feb 2025 · 0 repositories · arXiv:2502.01697
-
Discovering Chunks in Neural Embeddings for Interpretability 3 Feb 2025 · 0 repositories · arXiv:2502.01803
-
Evaluation of Large Language Models via Coupled Token Generation 3 Feb 2025 · 1 repository · arXiv:2502.01754
-
SelfCheckAgent: Zero-Resource Hallucination Detection in Generative Large Language Models 3 Feb 2025 · 0 repositories · arXiv:2502.01812
-
Agent-Based Uncertainty Awareness Improves Automated Radiology Report Labeling with an Open-Source Large Language Model 2 Feb 2025 · 0 repositories · arXiv:2502.01691
-
How Effective Is Constitutional AI in Small LLMs? A Study on DeepSeek-R1 and Its Peers 1 Feb 2025 · 0 repositories · arXiv:2503.17365
-
Pause-Tuning for Long-Context Comprehension: A Lightweight Approach to LLM Attention Recalibration 1 Feb 2025 · 0 repositories · arXiv:2502.20405
-
Cache Me If You Must: Adaptive Key-Value Quantization for Large Language Models 31 Jan 2025 · 1 repository · arXiv:2501.19392
-
Can AI Solve the Peer Review Crisis? A Large Scale Cross Model Experiment of LLMs' Performance and Biases in Evaluating over 1000 Economics Papers 31 Jan 2025 · 0 repositories · arXiv:2502.00070
-
DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights 30 Jan 2025 · 0 repositories · arXiv:2501.18596
-
GENIE: Generative Note Information Extraction model for structuring EHR data 30 Jan 2025 · 0 repositories · arXiv:2501.18435
-
GuardReasoner: Towards Reasoning-based LLM Safeguards 30 Jan 2025 · 1 repository · arXiv:2501.18492Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
High-Accuracy ECG Image Interpretation using Parameter-Efficient LoRA Fine-Tuning with Multimodal LLaMA 3.2 30 Jan 2025 · 0 repositories · arXiv:2501.18670
-
R.I.P.: Better Models by Survival of the Fittest Prompts 30 Jan 2025 · 0 repositories · arXiv:2501.18578
-
LLM Assistance for Pediatric Depression 29 Jan 2025 · 0 repositories · arXiv:2501.17510
-
COS(M+O)S: Curiosity and RL-Enhanced MCTS for Exploring Story Space via Language Models 28 Jan 2025 · 0 repositories · arXiv:2501.17104
-
Sparse Autoencoders Trained on the Same Data Learn Different Features 28 Jan 2025 · 0 repositories · arXiv:2501.16615
-
Evaluating The Performance of Using Large Language Models to Automate Summarization of CT Simulation Orders in Radiation Oncology 27 Jan 2025 · 0 repositories · arXiv:2501.16309
-
RelCAT: Advancing Extraction of Clinical Inter-Entity Relationships from Unstructured Electronic Health Records 27 Jan 2025 · 1 repository · arXiv:2501.16077
-
Targeting Alignment: Extracting Safety Classifiers of Aligned LLMs 27 Jan 2025 · 0 repositories · arXiv:2501.16534
-
The TIP of the Iceberg: Revealing a Hidden Class of Task-in-Prompt Adversarial Attacks on LLMs 27 Jan 2025 · 0 repositories · arXiv:2501.18626
-
Decentralized Low-Rank Fine-Tuning of Large Language Models 26 Jan 2025 · 0 repositories · arXiv:2501.15361
-
Improving Estonian Text Simplification through Pretrained Language Models and Custom Datasets 26 Jan 2025 · 0 repositories · arXiv:2501.15624
-
PatentLMM: Large Multimodal Model for Generating Descriptions for Patent Figures 25 Jan 2025 · 1 repository · arXiv:2501.15074
-
Mitigating Forgetting in LLM Fine-Tuning via Low-Perplexity Token Learning 24 Jan 2025 · 0 repositories · arXiv:2501.14315
-
Rethinking Table Instruction Tuning 24 Jan 2025 · 1 repository · arXiv:2501.14693
-
MedSlice: Fine-Tuned Large Language Models for Secure Clinical Note Sectioning 23 Jan 2025 · 1 repository · arXiv:2501.14105
-
RAMQA: A Unified Framework for Retrieval-Augmented Multi-Modal Question Answering 23 Jan 2025 · 1 repository · arXiv:2501.13297
-
RECALL: Library-Like Behavior In Language Models is Enhanced by Self-Referencing Causal Cycles 23 Jan 2025 · 1 repository · arXiv:2501.13491
-
The Breeze 2 Herd of Models: Traditional Chinese LLMs Based on Llama with Vision-Aware and Function-Calling Capabilities 23 Jan 2025 · 1 repository · arXiv:2501.13921
-
Benchmarking Generative AI for Scoring Medical Student Interviews in Objective Structured Clinical Examinations (OSCEs) 21 Jan 2025 · 0 repositories · arXiv:2501.13957
-
Can open source large language models be used for tumor documentation in Germany? -- An evaluation on urological doctors' notes 21 Jan 2025 · 1 repository · arXiv:2501.12106
-
Episodic Memories Generation and Evaluation Benchmark for Large Language Models 21 Jan 2025 · 1 repository · arXiv:2501.13121Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Leveraging Large Language Models to Enhance Machine Learning Interpretability and Predictive Performance: A Case Study on Emergency Department Returns for Mental Health Patients 21 Jan 2025 · 0 repositories · arXiv:2502.00025
-
Pushing the Limits of BFP on Narrow Precision LLM Inference 21 Jan 2025 · 0 repositories · arXiv:2502.00026
-
GREEN-CODE: Learning to Optimize Energy Efficiency in LLM-based Code Generation 19 Jan 2025 · 1 repository · arXiv:2501.11006
-
DNA 1.0 Technical Report 18 Jan 2025 · 0 repositories · arXiv:2501.10648
-
Dialogue Benchmark Generation from Knowledge Graphs with Cost-Effective Retrieval-Augmented LLMs 17 Jan 2025 · 1 repository · arXiv:2501.09928
-
Confidence Estimation for Error Detection in Text-to-SQL Systems 16 Jan 2025 · 1 repository · arXiv:2501.09527
-
Domain Adaptation of Foundation LLMs for e-Commerce 16 Jan 2025 · 0 repositories · arXiv:2501.09706
-
FASP: Fast and Accurate Structured Pruning of Large Language Models 16 Jan 2025 · 0 repositories · arXiv:2501.09412
-
Analyzing the Ethical Logic of Six Large Language Models 15 Jan 2025 · 0 repositories · arXiv:2501.08951
-
Development and Validation of the Provider Documentation Summarization Quality Instrument for Large Language Models 15 Jan 2025 · 0 repositories · arXiv:2501.08977
-
IDEA: Image Description Enhanced CLIP-Adapter 15 Jan 2025 · 1 repository · arXiv:2501.08816
-
Consistency of Responses and Continuations Generated by Large Language Models on Social Media 14 Jan 2025 · 0 repositories · arXiv:2501.08102
-
PokerBench: Training Large Language Models to become Professional Poker Players 14 Jan 2025 · 1 repository · arXiv:2501.08328
-
Potential and Perils of Large Language Models as Judges of Unstructured Textual Data 14 Jan 2025 · 0 repositories · arXiv:2501.08167
-
LLM-Net: Democratizing LLMs-as-a-Service through Blockchain-based Expert Networks 13 Jan 2025 · 0 repositories · arXiv:2501.07288
-
Language Fusion for Parameter-Efficient Cross-lingual Transfer 12 Jan 2025 · 1 repository · arXiv:2501.06892
-
Guided Code Generation with LLMs: A Multi-Agent Framework for Complex Code Tasks 11 Jan 2025 · 0 repositories · arXiv:2501.06625
-
Exploring Large Language Models for Translating Romanian Computational Problems into English 9 Jan 2025 · 0 repositories · arXiv:2501.05601
-
Using LLMs to Infer Non-Binary COVID-19 Sentiments of Chinese Micro-bloggers 9 Jan 2025 · 0 repositories · arXiv:2501.05423
-
Language and Planning in Robotic Navigation: A Multilingual Evaluation of State-of-the-Art Models 7 Jan 2025 · 0 repositories · arXiv:2501.05478
-
Leveraging Explainable AI for LLM Text Attribution: Differentiating Human-Written and Multiple LLMs-Generated Text 6 Jan 2025 · 0 repositories · arXiv:2501.03212
-
CHAIR -- Classifier of Hallucination as Improver 5 Jan 2025 · 1 repository · arXiv:2501.02518
-
Decoding specialised feature neurons in LLMs with the final projection layer 5 Jan 2025 · 0 repositories · arXiv:2501.02688
-
A Survey on Large Language Models with some Insights on their Capabilities and Limitations 3 Jan 2025 · 0 repositories · arXiv:2501.04040
-
PersonaAI: Leveraging Retrieval-Augmented Generation and Personalized Context for AI-Driven Digital Avatars 3 Jan 2025 · 0 repositories · arXiv:2503.15489
-
Large Language Models for Mental Health Diagnostic Assessments: Exploring The Potential of Large Language Models for Assisting with Mental Health Diagnostic Assessments -- The Depression and Anxiety Case 2 Jan 2025 · 0 repositories · arXiv:2501.01305
-
Toward Inclusive Educational AI: Auditing Frontier LLMs through a Multiplexity Lens 2 Jan 2025 · 0 repositories · arXiv:2501.03259
-
IGC: Integrating a Gated Calculator into an LLM to Solve Arithmetic Tasks Reliably and Efficiently 1 Jan 2025 · 0 repositories · arXiv:2501.00684
-
2 OLMo 2 Furious 31 Dec 2024 · 3 repositories · arXiv:2501.00656Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 21 harvested samples) · 2 pointer-only (licence)
-
An Empirical Evaluation of Large Language Models on Consumer Health Questions 31 Dec 2024 · 0 repositories · arXiv:2501.00208
-
EQUATOR: A Deterministic Framework for Evaluating LLM Reasoning with Open-Ended Questions. # v1.0.0-beta 31 Dec 2024 · 0 repositories · arXiv:2501.00257
-
Towards Sustainable Large Language Model Serving 31 Dec 2024 · 0 repositories · arXiv:2501.01990
-
Zero-Shot Strategies for Length-Controllable Summarization 31 Dec 2024 · 0 repositories · arXiv:2501.00233
-
Adaptive Batch Size Schedules for Distributed Training of Language Models with Data and Model Parallelism 30 Dec 2024 · 0 repositories · arXiv:2412.21124
-
Distilling Desired Comments for Enhanced Code Review with Large Language Models 29 Dec 2024 · 0 repositories · arXiv:2412.20340
-
Gradient Weight-normalized Low-rank Projection for Efficient LLM Training 27 Dec 2024 · 1 repository · arXiv:2412.19616