Methods › Natural Language Processing › Language Models › LLaMA › Papers, page 11
LLaMA
Papers archive 2025-07-28
archive papers tagged: 1,062 · with a code link: 423 · where Syntology ran a sample: 143 (127 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (143 of 1,062 tagged: 127 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 11 of 11: papers 1,001 to 1,062 of 1,062, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Building Accurate Translation-Tailored LLMs with Language Aware Instruction Tuning 21 Mar 2024 · 1 repository · arXiv:2403.14399
-
ChainLM: Empowering Large Language Models with Improved Chain-of-Thought Prompting 21 Mar 2024 · 1 repository · arXiv:2403.14312Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
From Representational Harms to Quality-of-Service Harms: A Case Study on Llama 2 Safety Safeguards 20 Mar 2024 · 1 repository · arXiv:2403.13213
-
Llama meets EU: Investigating the European Political Spectrum through the Lens of LLMs 20 Mar 2024 · 1 repository · arXiv:2403.13592Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
VL-Mamba: Exploring State Space Models for Multimodal Learning 20 Mar 2024 · 0 repositories · arXiv:2403.13600
-
DEE: Dual-stage Explainable Evaluation Method for Text Generation 18 Mar 2024 · 0 repositories · arXiv:2403.11509
-
Enhancing Taiwanese Hokkien Dual Translation by Exploring and Standardizing of Four Writing Systems 18 Mar 2024 · 1 repository · arXiv:2403.12024
-
FinLlama: Financial Sentiment Classification for Algorithmic Trading Applications 18 Mar 2024 · 0 repositories · arXiv:2403.12285
-
Metaphor Understanding Challenge Dataset for LLMs 18 Mar 2024 · 0 repositories · arXiv:2403.11810
-
Can Large Language Models abstract Medical Coded Language? 16 Mar 2024 · 0 repositories · arXiv:2403.10822
-
Empirical Studies of Parameter Efficient Methods for Large Language Models of Code and Knowledge Transfer to R 16 Mar 2024 · 1 repository · arXiv:2405.01553
-
Neural Erosion: Emulating Controlled Neurodegeneration and Aging in AI Systems 15 Mar 2024 · 0 repositories · arXiv:2403.10596
-
Can Large Language Models Identify Authorship? 13 Mar 2024 · 1 repository · arXiv:2403.08213
-
A Novel Nuanced Conversation Evaluation Framework for Large Language Models in Mental Health 8 Mar 2024 · 0 repositories · arXiv:2403.09705
-
MeanCache: User-Centric Semantic Caching for LLM Web Services 5 Mar 2024 · 0 repositories · arXiv:2403.02694
-
Information Flow Routes: Automatically Interpreting Language Models at Scale 27 Feb 2024 · 1 repository · arXiv:2403.00824
-
TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space 27 Feb 2024 · 1 repository · arXiv:2402.17811Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Pandora's White-Box: Precise Training Data Detection and Extraction in Large Language Models 26 Feb 2024 · 1 repository · arXiv:2402.17012Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts 26 Feb 2024 · 0 repositories · arXiv:2402.16822
-
How Do Humans Write Code? Large Models Do It the Same Way Too 24 Feb 2024 · 1 repository · arXiv:2402.15729
-
ProSparse: Introducing and Enhancing Intrinsic Activation Sparsity within Large Language Models 21 Feb 2024 · 1 repository · arXiv:2402.13516Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Test-Driven Development for Code Generation 21 Feb 2024 · 0 repositories · arXiv:2402.13521
-
OneBit: Towards Extremely Low-bit Large Language Models 17 Feb 2024 · 1 repository · arXiv:2402.11295Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 7 pointer-only (licence)
-
Network Formation and Dynamics Among Multi-LLMs 16 Feb 2024 · 1 repository · arXiv:2402.10659
-
Chain-of-Planned-Behaviour Workflow Elicits Few-Shot Mobility Generation in LLMs 15 Feb 2024 · 0 repositories · arXiv:2402.09836
-
UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction 10 Feb 2024 · 1 repository · arXiv:2402.06861Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Pedagogical Alignment of Large Language Models 7 Feb 2024 · 1 repository · arXiv:2402.05000
-
Beyond the Black Box: A Statistical Model for LLM Reasoning and Inference 5 Feb 2024 · 0 repositories · arXiv:2402.03175
-
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models 4 Feb 2024 · 1 repository · arXiv:2402.02370
-
LLM4Vuln: A Unified Evaluation Framework for Decoupling and Enhancing LLMs' Vulnerability Reasoning 29 Jan 2024 · 0 repositories · arXiv:2401.16185
-
Learning from Implicit User Feedback, Emotions and Demographic Information in Task-Oriented and Document-Grounded Dialogues 17 Jan 2024 · 1 repository · arXiv:2401.09248
-
Low-Rank Approximation for Sparse Attention in Multi-Modal LLMs 1 Jan 2024 · 0 repositories
-
Do LLM Agents Exhibit Social Behavior? 23 Dec 2023 · 0 repositories · arXiv:2312.15198
-
Training With "Paraphrasing the Original Text" Improves Long-Context Performance 18 Dec 2023 · 0 repositories · arXiv:2312.11193
-
Jellyfish: A Large Language Model for Data Preprocessing 4 Dec 2023 · 0 repositories · arXiv:2312.01678
-
Generative Parameter-Efficient Fine-Tuning 1 Dec 2023 · 1 repository · arXiv:2312.00700Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples) · 10 pointer-only (licence)
-
Mark My Words: Analyzing and Evaluating Language Model Watermarks 1 Dec 2023 · 1 repository · arXiv:2312.00273Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Positioning Political Texts with Large Language Models by Asking and Averaging 28 Nov 2023 · 0 repositories · arXiv:2311.16639
-
MEDITRON-70B: Scaling Medical Pretraining for Large Language Models 27 Nov 2023 · 1 repository · arXiv:2311.16079Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples)
-
OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning 16 Nov 2023 · 1 repository · arXiv:2311.09724Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Extrinsically-Focused Evaluation of Omissions in Medical Summarization 14 Nov 2023 · 1 repository · arXiv:2311.08303
-
Fair Abstractive Summarization of Diverse Perspectives 14 Nov 2023 · 1 repository · arXiv:2311.07884
-
LLMs and Finetuning: Benchmarking cross-domain performance for hate speech detection 29 Oct 2023 · 0 repositories · arXiv:2310.18964
-
Large Language Models can Share Images, Too! 23 Oct 2023 · 2 repositories · arXiv:2310.14804
-
Entity Matching using Large Language Models 17 Oct 2023 · 1 repository · arXiv:2310.11244
-
Alexpaca: Learning Factual Clarification Question Generation Without Examples 17 Oct 2023 · 0 repositories · arXiv:2310.11571
-
Assessing and Enhancing the Robustness of Large Language Models with Task Structure Variations for Logical Reasoning 13 Oct 2023 · 1 repository · arXiv:2310.09430Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Typing to Listen at the Cocktail Party: Text-Guided Target Speaker Extraction 11 Oct 2023 · 1 repository · arXiv:2310.07284
-
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning 9 Oct 2023 · 1 repository · arXiv:2310.05506
-
Compresso: Structured Pruning with Collaborative Prompting Learns Compact Large Language Models 8 Oct 2023 · 1 repository · arXiv:2310.05015Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
FaceGemma: Enhancing Image Captioning with Facial Attributes for Portrait Images 24 Sep 2023 · 0 repositories · arXiv:2309.13601
-
CPLLM: Clinical Prediction with Large Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11295
-
AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification 18 Sep 2023 · 1 repository · arXiv:2309.09800
-
Empowering In-Browser Deep Learning Inference on Edge Devices with Just-in-Time Kernel Optimizations 16 Sep 2023 · 0 repositories · arXiv:2309.08978
-
Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions 25 Aug 2023 · 1 repository · arXiv:2309.12342
-
Do Multilingual Language Models Think Better in English? 2 Aug 2023 · 1 repository · arXiv:2308.01223Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Scaling Sentence Embeddings with Large Language Models 31 Jul 2023 · 1 repository · arXiv:2307.16645Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs 22 Jun 2023 · 1 repository · arXiv:2306.13063Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
In-Context Learning User Simulators for Task-Oriented Dialog Systems 1 Jun 2023 · 2 repositories · arXiv:2306.00774
-
CoEdIT: Text Editing by Task-Specific Instruction Tuning 17 May 2023 · 1 repository · arXiv:2305.09857
-
NeuroComparatives: Neuro-Symbolic Distillation of Comparative Knowledge 8 May 2023 · 1 repository · arXiv:2305.04978
-
LLaMA: Open and Efficient Foundation Language Models 27 Feb 2023 · 57 repositories · arXiv:2302.13971Syntology official: no sample here; runs from other or unrecorded repositories · 37 ran (of which 9 constructed an object rather than computing a result; 25 with no instrument failure: 3 honoured, 0 violated, 22 with no contract checked; 12 where Syntology's instrument failed) · 21 unverified (of 58 harvested samples) · 4 pointer-only (licence)