Methods › Natural Language Processing › Language Models › LLaMA › Papers, page 5
LLaMA
Papers archive 2025-07-28
archive papers tagged: 1,062 · with a code link: 423 · where Syntology ran a sample: 143 (127 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (143 of 1,062 tagged: 127 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 5 of 11: papers 401 to 500 of 1,062, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MLLM-SUL: Multimodal Large Language Model for Semantic Scene Understanding and Localization in Traffic Scenarios 27 Dec 2024 · 1 repository · arXiv:2412.19406
-
Dynamic Skill Adaptation for Large Language Models 26 Dec 2024 · 0 repositories · arXiv:2412.19361
-
Whose Morality Do They Speak? Unraveling Cultural Bias in Multilingual Language Models 25 Dec 2024 · 0 repositories · arXiv:2412.18863
-
GenAI Content Detection Task 2: AI vs. Human -- Academic Essay Authenticity Challenge 24 Dec 2024 · 0 repositories · arXiv:2412.18274
-
Segment-Based Attention Masking for GPTs 24 Dec 2024 · 1 repository · arXiv:2412.18487
-
SlimGPT: Layer-wise Structured Pruning for Large Language Models 24 Dec 2024 · 0 repositories · arXiv:2412.18110
-
GQSA: Group Quantization and Sparsity for Accelerating Large Language Model Inference 23 Dec 2024 · 0 repositories · arXiv:2412.17560
-
Highly Optimized Kernels and Fine-Grained Codebooks for LLM Inference on Arm CPUs 23 Dec 2024 · 1 repository · arXiv:2501.00032
-
LMV-RPA: Large Model Voting-based Robotic Process Automation 23 Dec 2024 · 1 repository · arXiv:2412.17965
-
Evaluating LLM Reasoning in the Operations Research Domain with ORQA 22 Dec 2024 · 2 repositories · arXiv:2412.17874
-
Joint Knowledge Editing for Information Enrichment and Probability Promotion 22 Dec 2024 · 1 repository · arXiv:2412.17872
-
PsychAdapter: Adapting LLM Transformers to Reflect Traits, Personality and Mental Health 22 Dec 2024 · 1 repository · arXiv:2412.16882
-
Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification 21 Dec 2024 · 0 repositories · arXiv:2412.16486
-
InfoTech Assistant : A Multimodal Conversational Agent for InfoTechnology Web Portal Queries 21 Dec 2024 · 0 repositories · arXiv:2412.16412
-
SilVar: Speech Driven Multimodal Model for Reasoning Visual Question Answering and Object Localization 21 Dec 2024 · 1 repository · arXiv:2412.16771
-
A Machine Learning Approach for Emergency Detection in Medical Scenarios Using Large Language Models 20 Dec 2024 · 0 repositories · arXiv:2412.16341
-
Can LLMs Obfuscate Code? A Systematic Analysis of Large Language Models into Assembly Code Obfuscation 20 Dec 2024 · 0 repositories · arXiv:2412.16135
-
ConfliBERT: A Language Model for Political Conflict 19 Dec 2024 · 1 repository · arXiv:2412.15060
-
DirectorLLM for Human-Centric Video Generation 19 Dec 2024 · 0 repositories · arXiv:2412.14484
-
MixLLM: LLM Quantization with Global Mixed-precision between Output-features and Highly-efficient System Design 19 Dec 2024 · 0 repositories · arXiv:2412.14590
-
Understanding the Dark Side of LLMs' Intrinsic Self-Correction 19 Dec 2024 · 0 repositories · arXiv:2412.14959
-
Domain-adaptative Continual Learning for Low-resource Tasks: Evaluation on Nepali 18 Dec 2024 · 0 repositories · arXiv:2412.13860
-
Enhancing Knowledge Distillation for LLMs with Response-Priming Prompting 18 Dec 2024 · 1 repository · arXiv:2412.17846
-
LIFT: Improving Long Context Understanding Through Long Input Fine-Tuning 18 Dec 2024 · 0 repositories · arXiv:2412.13626
-
LLM-SEM: A Sentiment-Based Student Engagement Metric Using LLMS for E-Learning Platforms 18 Dec 2024 · 0 repositories · arXiv:2412.13765
-
Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN 18 Dec 2024 · 1 repository · arXiv:2412.13795
-
ResQ: Mixed-Precision Quantization of Large Language Models with Low-Rank Residuals 18 Dec 2024 · 1 repository · arXiv:2412.14363
-
Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models 18 Dec 2024 · 1 repository · arXiv:2412.13702
-
Algorithmic Fidelity of Large Language Models in Generating Synthetic German Public Opinions: A Case Study 17 Dec 2024 · 1 repository · arXiv:2412.13169
-
Extending LLMs to New Languages: A Case Study of Llama and Persian Adaptation 17 Dec 2024 · 1 repository · arXiv:2412.13375
-
LLMCL-GEC: Advancing Grammatical Error Correction with LLM-Driven Curriculum Learning 17 Dec 2024 · 0 repositories · arXiv:2412.12541
-
LLMs are Also Effective Embedding Models: An In-depth Overview 17 Dec 2024 · 0 repositories · arXiv:2412.12591
-
SWAN: SGD with Normalization and Whitening Enables Stateless LLM Training 17 Dec 2024 · 0 repositories · arXiv:2412.13148
-
Causal Diffusion Transformers for Generative Modeling 16 Dec 2024 · 1 repository · arXiv:2412.12095Syntology official (archive's flag): 7 ran · 8 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Codenames as a Benchmark for Large Language Models 16 Dec 2024 · 0 repositories · arXiv:2412.11373
-
Combining Large Language Models with Tutoring System Intelligence: A Case Study in Caregiver Homework Support 16 Dec 2024 · 1 repository · arXiv:2412.11995
-
The Open Source Advantage in Large Language Models (LLMs) 16 Dec 2024 · 0 repositories · arXiv:2412.12004
-
Beyond Discrete Personas: Personality Modeling Through Journal Intensive Conversations 15 Dec 2024 · 1 repository · arXiv:2412.11250
-
Sequence-Level Leakage Risk of Training Data in Large Language Models 15 Dec 2024 · 0 repositories · arXiv:2412.11302
-
A recent evaluation on the performance of LLMs on radiation oncology physics using questions of randomly shuffled options 14 Dec 2024 · 0 repositories · arXiv:2412.10622
-
Advancing Vehicle Plate Recognition: Multitasking Visual Language Models with VehiclePaliGemma 14 Dec 2024 · 0 repositories · arXiv:2412.14197
-
Do large language vision models understand 3D shapes? 14 Dec 2024 · 1 repository · arXiv:2412.10908
-
Llama 3 Meets MoE: Efficient Upcycling 13 Dec 2024 · 1 repository · arXiv:2412.09952
-
MST-R: Multi-Stage Tuning for Retrieval Systems and Metric Evaluation 13 Dec 2024 · 1 repository · arXiv:2412.10313
-
Targeted Angular Reversal of Weights (TARS) for Knowledge Removal in Large Language Models 13 Dec 2024 · 0 repositories · arXiv:2412.10257
-
Benchmarking LLMs for Mimicking Child-Caregiver Language in Interaction 12 Dec 2024 · 0 repositories · arXiv:2412.09318
-
Foundational Large Language Models for Materials Research 12 Dec 2024 · 1 repository · arXiv:2412.09560Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Lexico: Extreme KV Cache Compression via Sparse Coding over Universal Dictionaries 12 Dec 2024 · 1 repository · arXiv:2412.08890
-
Assessing Personalized AI Mentoring with Large Language Models in the Computing Field 11 Dec 2024 · 0 repositories · arXiv:2412.08430
-
NyayaAnumana & INLegalLlama: The Largest Indian Legal Judgment Prediction Dataset and Specialized Language Model for Enhanced Decision Analysis 11 Dec 2024 · 1 repository · arXiv:2412.08385
-
Frame Representation Hypothesis: Multi-Token LLM Interpretability and Concept-Guided Text Generation 10 Dec 2024 · 1 repository · arXiv:2412.07334
-
Generating Knowledge Graphs from Large Language Models: A Comparative Study of GPT-4, LLaMA 2, and BERT 10 Dec 2024 · 0 repositories · arXiv:2412.07412
-
Identifying and Manipulating Personality Traits in LLMs Through Activation Engineering 10 Dec 2024 · 1 repository · arXiv:2412.10427
-
MemHunter: Automated and Verifiable Memorization Detection at Dataset-scale in LLMs 10 Dec 2024 · 0 repositories · arXiv:2412.07261
-
TrojanWhisper: Evaluating Pre-trained LLMs to Detect and Localize Hardware Trojans 10 Dec 2024 · 0 repositories · arXiv:2412.07636
-
Zero-Shot ATC Coding with Large Language Models for Clinical Assessments 10 Dec 2024 · 0 repositories · arXiv:2412.07743
-
SUPERMERGE: An Approach For Gradient-Based Model Merging 9 Dec 2024 · 0 repositories · arXiv:2412.10416
-
Domain-Specific Translation with Open-Source Large Language Models: Resource-Oriented Analysis 8 Dec 2024 · 0 repositories · arXiv:2412.05862
-
Fully Open Source Moxin-7B Technical Report 8 Dec 2024 · 1 repository · arXiv:2412.06845
-
Taming Sensitive Weights : Noise Perturbation Fine-tuning for Robust LLM Quantization 8 Dec 2024 · 0 repositories · arXiv:2412.06858
-
Comprehensive Evaluation of Multimodal AI Models in Medical Imaging Diagnosis: From Data Augmentation to Preference-Based Comparison 7 Dec 2024 · 0 repositories · arXiv:2412.05536
-
ChatNVD: Advancing Cybersecurity Vulnerability Assessment with Large Language Models 6 Dec 2024 · 0 repositories · arXiv:2412.04756
-
Enhancing Cross-Language Code Translation via Task-Specific Embedding Alignment in Retrieval-Augmented Generation 6 Dec 2024 · 0 repositories · arXiv:2412.05159
-
Frontier Models are Capable of In-context Scheming 6 Dec 2024 · 0 repositories · arXiv:2412.04984
-
Aya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier 5 Dec 2024 · 0 repositories · arXiv:2412.04261
-
Extractive Structures Learned in Pretraining Enable Generalization on Finetuned Facts 5 Dec 2024 · 1 repository · arXiv:2412.04614Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Florence-VL: Enhancing Vision-Language Models with Generative Vision Encoder and Depth-Breadth Fusion 5 Dec 2024 · 1 repository · arXiv:2412.04424
-
SKIM: Any-bit Quantization Pushing The Limits of Post-Training Quantization 5 Dec 2024 · 0 repositories · arXiv:2412.04180
-
Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.03537
-
From Language Models over Tokens to Language Models over Characters 4 Dec 2024 · 0 repositories · arXiv:2412.03719Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Enhancing CLIP Conceptual Embedding through Knowledge Distillation 4 Dec 2024 · 0 repositories · arXiv:2412.03513
-
CEGI: Measuring the trade-off between efficiency and carbon emissions for SLMs and VLMs 3 Dec 2024 · 0 repositories · arXiv:2412.02602
-
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity 3 Dec 2024 · 0 repositories · arXiv:2412.02252
-
Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset 3 Dec 2024 · 0 repositories · arXiv:2412.02595
-
RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models 3 Dec 2024 · 1 repository · arXiv:2412.02830
-
Early Exit Is a Natural Capability in Transformer-based Models: An Empirical Study on Early Exit without Joint Optimization 2 Dec 2024 · 0 repositories · arXiv:2412.01455
-
MALT: Improving Reasoning with Multi-Agent LLM Training 2 Dec 2024 · 0 repositories · arXiv:2412.01928
-
Multimodal Medical Disease Classification with LLaMA II 2 Dec 2024 · 0 repositories · arXiv:2412.01306
-
Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages 1 Dec 2024 · 0 repositories · arXiv:2412.00948
-
LLaMA-Gene: A General-purpose Gene Task Large Language Model Based on Instruction Fine-tuning 30 Nov 2024 · 0 repositories · arXiv:2412.00471
-
On Explaining Recommendations with Large Language Models: A Review 29 Nov 2024 · 0 repositories · arXiv:2411.19576
-
Fine Tuning Large Language Models to Deliver CBT for Depression 29 Nov 2024 · 1 repository · arXiv:2412.00251
-
Sensitive Content Classification in Social Media: A Holistic Resource and Evaluation 29 Nov 2024 · 0 repositories · arXiv:2411.19832
-
An Extensive Evaluation of Factual Consistency in Large Language Models for Data-to-Text Generation 28 Nov 2024 · 0 repositories · arXiv:2411.19203
-
DENIAHL: In-Context Features Influence LLM Needle-In-A-Haystack Abilities 28 Nov 2024 · 1 repository · arXiv:2411.19360
-
DIESEL -- Dynamic Inference-Guidance via Evasion of Semantic Embeddings in LLMs 28 Nov 2024 · 0 repositories · arXiv:2411.19038
-
Sneaking Syntax into Transformer Language Models with Tree Regularization 28 Nov 2024 · 0 repositories · arXiv:2411.18885
-
Emergence of Self-Identity in AI: A Mathematical Framework and Empirical Study with Generative Large Language Models 27 Nov 2024 · 1 repository · arXiv:2411.18530
-
Training and Evaluating Language Models with Template-based Data Generation 27 Nov 2024 · 1 repository · arXiv:2411.18104Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Advancing Content Moderation: Evaluating Large Language Models for Detecting Sensitive Content Across Text, Images, and Videos 26 Nov 2024 · 0 repositories · arXiv:2411.17123
-
On Limitations of LLM as Annotator for Low Resource Languages 26 Nov 2024 · 0 repositories · arXiv:2411.17637
-
BayLing 2: A Multilingual Large Language Model with Efficient Language Alignment 25 Nov 2024 · 1 repository · arXiv:2411.16300
-
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring 25 Nov 2024 · 0 repositories · arXiv:2411.16337
-
Cautious Optimizers: Improving Training with One Line of Code 25 Nov 2024 · 3 repositories · arXiv:2411.16085Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16991
-
Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16797
-
The Two-Hop Curse: LLMs trained on A→B, B→C fail to learn A→C 25 Nov 2024 · 0 repositories · arXiv:2411.16353
-
Anda: Unlocking Efficient LLM Inference with a Variable-Length Grouped Activation Data Format 24 Nov 2024 · 0 repositories · arXiv:2411.15982
-
LLaMA-MoE v2: Exploring Sparsity of LLaMA from Perspective of Mixture-of-Experts with Post-Training 24 Nov 2024 · 1 repository · arXiv:2411.15708Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
"Moralized" Multi-Step Jailbreak Prompts: Black-Box Testing of Guardrails in Large Language Models for Verbal Attacks 23 Nov 2024 · 1 repository · arXiv:2411.16730