Methods › Natural Language Processing › Language Models › LLaMA › Papers, page 3
LLaMA
Papers archive 2025-07-28
archive papers tagged: 1,062 · with a code link: 423 · where Syntology ran a sample: 143 (127 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (143 of 1,062 tagged: 127 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 3 of 11: papers 201 to 300 of 1,062, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LLM Benchmarking with LLaMA2: Evaluating Code Development Performance Across Multiple Programming Languages 24 Mar 2025 · 1 repository · arXiv:2503.19217
-
Investigating Recent Large Language Models for Vietnamese Machine Reading Comprehension 23 Mar 2025 · 0 repositories · arXiv:2503.18062
-
Language-specific Neurons Do Not Facilitate Cross-Lingual Transfer 21 Mar 2025 · 0 repositories · arXiv:2503.17456
-
SaudiCulture: A Benchmark for Evaluating Large Language Models Cultural Competence within Saudi Arabia 21 Mar 2025 · 0 repositories · arXiv:2503.17485
-
Text2Model: Generating dynamic chemical reactor models using large language models (LLMs) 21 Mar 2025 · 0 repositories · arXiv:2503.17004
-
V-Seek: Accelerating LLM Reasoning on Open-hardware Server-class RISC-V Platforms 21 Mar 2025 · 0 repositories · arXiv:2503.17422
-
Variance Control via Weight Rescaling in LLM Pre-training 21 Mar 2025 · 1 repository · arXiv:2503.17500
-
Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models 20 Mar 2025 · 0 repositories · arXiv:2503.16148
-
Poly-FEVER: A Multilingual Fact Verification Benchmark for Hallucination Detection in Large Language Models 19 Mar 2025 · 0 repositories · arXiv:2503.16541
-
Reinforcement Learning Environment with LLM-Controlled Adversary in D&D 5th Edition Combat 19 Mar 2025 · 0 repositories · arXiv:2503.15726
-
Empowering Smaller Models: Tuning LLaMA and Gemma with Chain-of-Thought for Ukrainian Exam Tasks 18 Mar 2025 · 1 repository · arXiv:2503.13988
-
ROMA: a Read-Only-Memory-based Accelerator for QLoRA-based On-Device LLM 17 Mar 2025 · 0 repositories · arXiv:2503.12988
-
VeriLeaky: Navigating IP Protection vs Utility in Fine-Tuning for LLM-Driven Verilog Coding 17 Mar 2025 · 0 repositories · arXiv:2503.13116
-
LLM-Driven Multi-step Translation from C to Rust using Static Analysis 16 Mar 2025 · 1 repository · arXiv:2503.12511
-
Integrating Chain-of-Thought and Retrieval Augmented Generation Enhances Rare Disease Diagnosis from Clinical Notes 15 Mar 2025 · 0 repositories · arXiv:2503.12286
-
Optimizing Large Language Models for Detecting Symptoms of Comorbid Depression or Anxiety in Chronic Diseases: Insights from Patient Messages 14 Mar 2025 · 0 repositories · arXiv:2503.11384
-
Prompt Sentiment: The Catalyst for LLM Change 14 Mar 2025 · 0 repositories · arXiv:2503.13510
-
Towards Extreme Pruning of LLMs with Plug-and-Play Mixed Sparsity 14 Mar 2025 · 0 repositories · arXiv:2503.11164
-
Scalable Evaluation of Online Facilitation Strategies via Synthetic Simulation of Discussions 13 Mar 2025 · 1 repository · arXiv:2503.16505
-
Aligning to What? Limits to RLHF Based Alignment 12 Mar 2025 · 1 repository · arXiv:2503.09025
-
Battling Misinformation: An Empirical Study on Adversarial Factuality in Open-Source Large Language Models 12 Mar 2025 · 0 repositories · arXiv:2503.10690
-
Can A Society of Generative Agents Simulate Human Behavior and Inform Public Health Policy? A Case Study on Vaccine Hesitancy 12 Mar 2025 · 0 repositories · arXiv:2503.09639
-
CyberLLMInstruct: A New Dataset for Analysing Safety of Fine-Tuned LLMs Using Cyber Security Data 12 Mar 2025 · 1 repository · arXiv:2503.09334
-
Enhancing High-Quality Code Generation in Large Language Models with Comparative Prefix-Tuning 12 Mar 2025 · 1 repository · arXiv:2503.09020
-
Improving the Reusability of Conversational Search Test Collections 12 Mar 2025 · 1 repository · arXiv:2503.09899
-
SurgicalVLM-Agent: Towards an Interactive AI Co-Pilot for Pituitary Surgery 12 Mar 2025 · 0 repositories · arXiv:2503.09474
-
Enhancing Large Language Models for Hardware Verification: A Novel SystemVerilog Assertion Dataset 11 Mar 2025 · 1 repository · arXiv:2503.08923
-
Fact-checking with Generative AI: A Systematic Cross-Topic Examination of LLMs Capacity to Detect Veracity of Political Information 11 Mar 2025 · 0 repositories · arXiv:2503.08404
-
Llms, Virtual Users, and Bias: Predicting Any Survey Question Without Human Data 11 Mar 2025 · 0 repositories · arXiv:2503.16498
-
Evaluating LLaMA 3.2 for Software Vulnerability Detection 10 Mar 2025 · 0 repositories · arXiv:2503.07770
-
Fully Autonomous Programming using Iterative Multi-Agent Debugging with Large Language Models 10 Mar 2025 · 0 repositories · arXiv:2503.07693
-
Identifying Non-Replicable Social Science Studies with Language Models 10 Mar 2025 · 0 repositories · arXiv:2503.10671
-
Roamify: Designing and Evaluating an LLM Based Google Chrome Extension for Personalised Itinerary Planning 10 Mar 2025 · 1 repository · arXiv:2504.10489
-
Sometimes the Model doth Preach: Quantifying Religious Bias in Open LLMs through Demographic Analysis in Asian Nations 10 Mar 2025 · 1 repository · arXiv:2503.07510
-
Towards Superior Quantization Accuracy: A Layer-sensitive Approach 9 Mar 2025 · 0 repositories · arXiv:2503.06518
-
Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues 9 Mar 2025 · 1 repository · arXiv:2503.06424
-
Critical Foreign Policy Decisions (CFPD)-Benchmark: Measuring Diplomatic Preferences in Large Language Models 8 Mar 2025 · 0 repositories · arXiv:2503.06263
-
Reinforced Diffuser for Red Teaming Large Vision-Language Models 8 Mar 2025 · 0 repositories · arXiv:2503.06223
-
This Is Your Doge, If It Please You: Exploring Deception and Robustness in Mixture of LLMs 7 Mar 2025 · 1 repository · arXiv:2503.05856
-
HelpSteer3: Human-Annotated Feedback and Edit Data to Empower Inference-Time Scaling in Open-Ended General-Domain Tasks 6 Mar 2025 · 0 repositories · arXiv:2503.04378
-
Memory Is All You Need: Testing How Model Memory Affects LLM Performance in Annotation Tasks 6 Mar 2025 · 0 repositories · arXiv:2503.04874
-
PokéChamp: an Expert-level Minimax Language Agent 6 Mar 2025 · 0 repositories · arXiv:2503.04094
-
Wanda++: Pruning Large Language Models via Regional Gradients 6 Mar 2025 · 1 repository · arXiv:2503.04992
-
Large language models in finance : what is financial sentiment? 5 Mar 2025 · 0 repositories · arXiv:2503.03612
-
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders 4 Mar 2025 · 0 repositories · arXiv:2503.02993
-
LLMInit: A Free Lunch from Large Language Models for Selective Initialization of Recommendation 3 Mar 2025 · 0 repositories · arXiv:2503.01814
-
Cancer Type, Stage and Prognosis Assessment from Pathology Reports using LLMs 3 Mar 2025 · 1 repository · arXiv:2503.01194
-
Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs 3 Mar 2025 · 1 repository · arXiv:2503.01307Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Evidence of conceptual mastery in the application of rules by Large Language Models 2 Mar 2025 · 0 repositories · arXiv:2503.00992
-
LADDER: Self-Improving LLMs Through Recursive Problem Decomposition 2 Mar 2025 · 0 repositories · arXiv:2503.00735
-
Collective Reasoning Among LLMs A Framework for Answer Validation Without Ground Truth 28 Feb 2025 · 0 repositories · arXiv:2502.20758
-
NutriGen: Personalized Meal Plan Generator Leveraging Large Language Models to Enhance Dietary and Nutritional Adherence 28 Feb 2025 · 1 repository · arXiv:2502.20601
-
Optimizing Large Language Models for ESG Activity Detection in Financial Texts 28 Feb 2025 · 1 repository · arXiv:2502.21112
-
Passage Query Methods for Retrieval and Reranking in Conversational Agents 28 Feb 2025 · 0 repositories · arXiv:2503.00238
-
RuCCoD: Towards Automated ICD Coding in Russian 28 Feb 2025 · 1 repository · arXiv:2502.21263
-
An exploration of features to improve the generalisability of fake news detection models 27 Feb 2025 · 0 repositories · arXiv:2502.20299
-
HaLoRA: Hardware-aware Low-Rank Adaptation for Large Language Models Based on Hybrid Compute-in-Memory Architecture 27 Feb 2025 · 0 repositories · arXiv:2502.19747
-
SkipPipe: Partial and Reordered Pipelining Framework for Training LLMs in Heterogeneous Networks 27 Feb 2025 · 1 repository · arXiv:2502.19913
-
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation 26 Feb 2025 · 0 repositories · arXiv:2502.18726
-
NeoBERT: A Next-Generation BERT 26 Feb 2025 · 1 repository · arXiv:2502.19587
-
Nexus: An Omni-Perceptive And -Interactive Model for Language, Audio, And Vision 26 Feb 2025 · 0 repositories · arXiv:2503.01879
-
Starjob: Dataset for LLM-Driven Job Shop Scheduling 26 Feb 2025 · 1 repository · arXiv:2503.01877
-
The Sharpness Disparity Principle in Transformers for Accelerating Language Model Pre-Training 26 Feb 2025 · 0 repositories · arXiv:2502.19002
-
AMPO: Active Multi-Preference Optimization 25 Feb 2025 · 0 repositories · arXiv:2502.18293
-
FRIDA to the Rescue! Analyzing Synthetic Data Effectiveness in Object-Based Common Sense Reasoning for Disaster Response 25 Feb 2025 · 0 repositories · arXiv:2502.18452
-
NusaAksara: A Multimodal and Multilingual Benchmark for Preserving Indonesian Indigenous Scripts 25 Feb 2025 · 0 repositories · arXiv:2502.18148
-
Single- vs. Dual-Prompt Dialogue Generation with LLMs for Job Interviews in Human Resources 25 Feb 2025 · 0 repositories · arXiv:2502.18650
-
SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution 25 Feb 2025 · 0 repositories · arXiv:2502.18449
-
Are Large Language Models Good Data Preprocessors? 24 Feb 2025 · 0 repositories · arXiv:2502.16790
-
Correlating and Predicting Human Evaluations of Language Models from Natural Language Processing Benchmarks 24 Feb 2025 · 0 repositories · arXiv:2502.18339
-
CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought 24 Feb 2025 · 1 repository · arXiv:2502.17214
-
StatLLM: A Dataset for Evaluating the Performance of Large Language Models in Statistical Analysis 24 Feb 2025 · 1 repository · arXiv:2502.17657
-
Dynamic LLM Routing and Selection based on User Preferences: Balancing Performance, Cost, and Ethics 23 Feb 2025 · 0 repositories · arXiv:2502.16696
-
A Multi-Agent Framework for Automated Vulnerability Detection and Repair in Solidity and Move Smart Contracts 22 Feb 2025 · 0 repositories · arXiv:2502.18515
-
Fine-Tuning Qwen 2.5 3B for Realistic Movie Dialogue Generation 22 Feb 2025 · 0 repositories · arXiv:2502.16274
-
Toward a Flexible Framework for Linear Representation Hypothesis Using Maximum Likelihood Estimation 22 Feb 2025 · 0 repositories · arXiv:2502.16385
-
AutoMedPrompt: A New Framework for Optimizing LLM Medical Prompts Using Textual Gradients 21 Feb 2025 · 0 repositories · arXiv:2502.15944
-
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models 21 Feb 2025 · 0 repositories · arXiv:2502.18505
-
MMRAG: Multi-Mode Retrieval-Augmented Generation with Large Language Models for Biomedical In-Context Learning 21 Feb 2025 · 0 repositories · arXiv:2502.15954
-
Do LLMs Consider Security? An Empirical Study on Responses to Programming Questions 20 Feb 2025 · 0 repositories · arXiv:2502.14202
-
English Please: Evaluating Machine Translation with Large Language Models for Multilingual Bug Reports 20 Feb 2025 · 1 repository · arXiv:2502.14338
-
Enhancing Conversational Agents with Theory of Mind: Aligning Beliefs, Desires, and Intentions for Human-Like Interaction 20 Feb 2025 · 1 repository · arXiv:2502.14171
-
Less is More: Improving LLM Alignment via Preference Data Selection 20 Feb 2025 · 0 repositories · arXiv:2502.14560
-
Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation 20 Feb 2025 · 0 repositories · arXiv:2502.14846
-
FairKV: Balancing Per-Head KV Cache for Fast Multi-GPU Inference 19 Feb 2025 · 0 repositories · arXiv:2502.15804
-
Giving AI Personalities Leads to More Human-Like Reasoning 19 Feb 2025 · 0 repositories · arXiv:2502.14155
-
Theoretical Physics Benchmark (TPBench) -- a Dataset and Study of AI Reasoning Capabilities in Theoretical Physics 19 Feb 2025 · 0 repositories · arXiv:2502.15815
-
ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails 19 Feb 2025 · 1 repository · arXiv:2502.13458
-
DSMoE: Matrix-Partitioned Experts with Dynamic Routing for Computation-Efficient Dense LLMs 18 Feb 2025 · 0 repositories · arXiv:2502.12455
-
Investigating the Impact of Quantization Methods on the Safety and Reliability of Large Language Models 18 Feb 2025 · 1 repository · arXiv:2502.15799
-
K-Paths: Reasoning over Graph Paths for Drug Repurposing and Drug Interaction Prediction 18 Feb 2025 · 1 repository · arXiv:2502.13344
-
OCCULT: Evaluating Large Language Models for Offensive Cyber Operation Capabilities 18 Feb 2025 · 0 repositories · arXiv:2502.15797
-
Revisiting Privacy, Utility, and Efficiency Trade-offs when Fine-Tuning Large Language Models 18 Feb 2025 · 0 repositories · arXiv:2502.13313
-
Testing Prompt Engineering Methods for Knowledge Extraction from Text 18 Feb 2025 · 1 repository
-
Independence Tests for Language Models 17 Feb 2025 · 0 repositories · arXiv:2502.12292
-
LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws 17 Feb 2025 · 0 repositories · arXiv:2502.12120
-
Logic.py: Bridging the Gap between LLMs and Constraint Solvers 17 Feb 2025 · 1 repository · arXiv:2502.15776
-
SmartLLM: Smart Contract Auditing using Custom Generative AI 17 Feb 2025 · 0 repositories · arXiv:2502.13167
-
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction 17 Feb 2025 · 1 repository · arXiv:2502.11946
-
CoLA: Compute-Efficient Pre-Training of LLMs via Low-Rank Activation 16 Feb 2025 · 1 repository · arXiv:2502.10940