Methods › Natural Language Processing › Language Models › GPT-4 › Papers, page 4
GPT-4
Papers archive 2025-07-28
archive papers tagged: 2,870 · with a code link: 1,244 · where Syntology ran a sample: 526 (417 with a run with no instrument failure, 109 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (526 of 2,870 tagged: 417 with a run with no instrument failure, 109 where every run was a failure of Syntology's instrument)
Page 4 of 29: papers 301 to 400 of 2,870, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Decentralized Low-Rank Fine-Tuning of Large Language Models 26 Jan 2025 · 0 repositories · arXiv:2501.15361
-
Improving Estonian Text Simplification through Pretrained Language Models and Custom Datasets 26 Jan 2025 · 0 repositories · arXiv:2501.15624
-
SedarEval: Automated Evaluation using Self-Adaptive Rubrics 26 Jan 2025 · 1 repository · arXiv:2501.15595Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
An AI-Driven Live Systematic Reviews in the Brain-Heart Interconnectome: Minimizing Research Waste and Advancing Evidence Synthesis 25 Jan 2025 · 1 repository · arXiv:2501.17181
-
Knowledge Hierarchy Guided Biological-Medical Dataset Distillation for Domain LLM Training 25 Jan 2025 · 0 repositories · arXiv:2501.15108
-
LLM Evaluation Based on Aerospace Manufacturing Expertise: Automated Generation and Multi-Model Question Answering 25 Jan 2025 · 0 repositories · arXiv:2501.17183
-
Using Large Language Models for education managements in Vietnamese with low resources 25 Jan 2025 · 0 repositories · arXiv:2501.15022
-
Rethinking Table Instruction Tuning 24 Jan 2025 · 1 repository · arXiv:2501.14693
-
Test-Time Code-Switching for Cross-lingual Aspect Sentiment Triplet Extraction 24 Jan 2025 · 0 repositories · arXiv:2501.14144
-
Enhancing Biomedical Relation Extraction with Directionality 23 Jan 2025 · 1 repository · arXiv:2501.14079
-
LLMs are Vulnerable to Malicious Prompts Disguised as Scientific Language 23 Jan 2025 · 0 repositories · arXiv:2501.14073
-
LLMs Can Plan Only If We Tell Them 23 Jan 2025 · 0 repositories · arXiv:2501.13545
-
Question Answering on Patient Medical Records with Private Fine-Tuned LLMs 23 Jan 2025 · 0 repositories · arXiv:2501.13687
-
Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models 23 Jan 2025 · 0 repositories · arXiv:2501.13629
-
Automatic Labelling with Open-source LLMs using Dynamic Label Schema Integration 21 Jan 2025 · 0 repositories · arXiv:2501.12332
-
Episodic Memories Generation and Evaluation Benchmark for Large Language Models 21 Jan 2025 · 1 repository · arXiv:2501.13121Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
KEIR @ ECIR 2025: The Second Workshop on Knowledge-Enhanced Information Retrieval 20 Jan 2025 · 0 repositories · arXiv:2501.11499
-
Chain-of-Reasoning: Towards Unified Mathematical Reasoning in Large Language Models via a Multi-Paradigm Perspective 19 Jan 2025 · 0 repositories · arXiv:2501.11110
-
PaSa: An LLM Agent for Comprehensive Academic Paper Search 17 Jan 2025 · 1 repository · arXiv:2501.10120Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Perspective Transition of Large Language Models for Solving Subjective Tasks 16 Jan 2025 · 0 repositories · arXiv:2501.09265
-
Enhanced Large Language Models for Effective Screening of Depression and Anxiety 15 Jan 2025 · 0 repositories · arXiv:2501.08769
-
Exploring Narrative Clustering in Large Language Models: A Layerwise Analysis of BERT 14 Jan 2025 · 0 repositories · arXiv:2501.08053
-
Large Language Models For Text Classification: Case Study And Comprehensive Review 14 Jan 2025 · 0 repositories · arXiv:2501.08457
-
PokerBench: Training Large Language Models to become Professional Poker Players 14 Jan 2025 · 1 repository · arXiv:2501.08328
-
Generative Artificial Intelligence-Supported Pentesting: A Comparison between Claude Opus, GPT-4, and Copilot 12 Jan 2025 · 0 repositories · arXiv:2501.06963
-
ZNO-Eval: Benchmarking reasoning capabilities of large language models in Ukrainian 12 Jan 2025 · 1 repository · arXiv:2501.06715
-
Iconicity in Large Language Models 10 Jan 2025 · 0 repositories · arXiv:2501.05643
-
Model Inversion in Split Learning for Personalized LLMs: New Insights from Information Bottleneck Theory 10 Jan 2025 · 0 repositories · arXiv:2501.05965
-
Large language models streamline automated systematic review: A preliminary study 9 Jan 2025 · 0 repositories · arXiv:2502.15702
-
LongViTU: Instruction Tuning for Long-Form Video Understanding 9 Jan 2025 · 0 repositories · arXiv:2501.05037
-
OpenAI ChatGPT interprets Radiological Images: GPT-4 as a Medical Doctor for a Fast Check-Up 9 Jan 2025 · 0 repositories · arXiv:2501.06269
-
The dynamics of meaning through time: Assessment of Large Language Models 9 Jan 2025 · 0 repositories · arXiv:2501.05552
-
Language and Planning in Robotic Navigation: A Multilingual Evaluation of State-of-the-Art Models 7 Jan 2025 · 0 repositories · arXiv:2501.05478
-
Developing an Artificial Intelligence Tool for Personalized Breast Cancer Treatment Plans based on the NCCN Guidelines 6 Jan 2025 · 0 repositories · arXiv:2502.15698
-
VicSim: Enhancing Victim Simulation with Emotional and Linguistic Fidelity 6 Jan 2025 · 0 repositories · arXiv:2501.03139
-
Evaluating Large Language Models Against Human Annotators in Latent Content Analysis: Sentiment, Political Leaning, Emotional Intensity, and Sarcasm 5 Jan 2025 · 0 repositories · arXiv:2501.02532
-
HonkaiChat: Companions from Anime that feel alive! 5 Jan 2025 · 0 repositories · arXiv:2501.03277
-
Towards New Benchmark for AI Alignment & Sentiment Analysis in Socially Important Issues: A Comparative Study of Human and LLMs in the Context of AGI 5 Jan 2025 · 0 repositories · arXiv:2501.02531
-
Examining the Robustness of Homogeneity Bias to Hyperparameter Adjustments in GPT-4 4 Jan 2025 · 0 repositories · arXiv:2501.02211
-
Exploring the Capabilities and Limitations of Large Language Models for Radiation Oncology Decision Support 4 Jan 2025 · 0 repositories · arXiv:2501.02346
-
The Application of Large Language Models in Recommendation Systems 4 Jan 2025 · 0 repositories · arXiv:2501.02178
-
Classifier-Guided Captioning Across Modalities 3 Jan 2025 · 0 repositories · arXiv:2501.03183
-
LLMs & Legal Aid: Understanding Legal Needs Exhibited Through User Queries 3 Jan 2025 · 0 repositories · arXiv:2501.01711
-
MIRAGE: Exploring How Large Language Models Perform in Complex Social Interactive Environments 3 Jan 2025 · 1 repository · arXiv:2501.01652
-
Turning Logic Against Itself : Probing Model Defenses Through Contrastive Questions 3 Jan 2025 · 1 repository · arXiv:2501.01872Syntology official: no sample here; runs from other or unrecorded repositories · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Toward Inclusive Educational AI: Auditing Frontier LLMs through a Multiplexity Lens 2 Jan 2025 · 0 repositories · arXiv:2501.03259
-
Column Property Annotation using Large Language Models 1 Jan 2025 · 1 repository
-
R2C: Mapping Room to Chessboard to Unlock LLM As Low-Level Action Planner 1 Jan 2025 · 0 repositories
-
Separation of Powers: On Segregating Knowledge from Observation in LLM-enabled Knowledge-based Visual Question Answering 1 Jan 2025 · 0 repositories
-
Yo'Chameleon: Personalized Vision and Language Generation 1 Jan 2025 · 0 repositories
-
Echoes in AI: Quantifying Lack of Plot Diversity in LLM Outputs 31 Dec 2024 · 0 repositories · arXiv:2501.00273
-
GPT-4 on Clinic Depression Assessment: An LLM-Based Pilot Study 31 Dec 2024 · 0 repositories · arXiv:2501.00199
-
Probing Visual Language Priors in VLMs 31 Dec 2024 · 0 repositories · arXiv:2501.00569
-
An Unsupervised Anomaly Detection in Electricity Consumption Using Reinforcement Learning and Time Series Forest Based Framework 30 Dec 2024 · 0 repositories · arXiv:2501.00107
-
CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions 30 Dec 2024 · 0 repositories · arXiv:2501.00097
-
Facilitating large language model Russian adaptation with Learned Embedding Propagation 30 Dec 2024 · 1 repository · arXiv:2412.21140
-
NLP-based Regulatory Compliance -- Using GPT 4.0 to Decode Regulatory Documents 29 Dec 2024 · 0 repositories · arXiv:2412.20602
-
DDD-GenDT: Dynamic Data-driven Generative Digital Twin Framework 28 Dec 2024 · 0 repositories · arXiv:2501.00051
-
Efficient Multi-Agent Collaboration with Tool Use for Online Planning in Complex Table Question Answering 28 Dec 2024 · 0 repositories · arXiv:2412.20145
-
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models 27 Dec 2024 · 0 repositories · arXiv:2412.19449
-
Toward Adaptive Reasoning in Large Language Models with Thought Rollback 27 Dec 2024 · 1 repository · arXiv:2412.19707Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MEDEC: A Benchmark for Medical Error Detection and Correction in Clinical Notes 26 Dec 2024 · 1 repository · arXiv:2412.19260
-
Reversed in Time: A Novel Temporal-Emphasized Benchmark for Cross-Modal Video-Text Retrieval 26 Dec 2024 · 1 repository · arXiv:2412.19178
-
Optimizing Large Language Models with an Enhanced LoRA Fine-Tuning Algorithm for Efficiency and Robustness in NLP Tasks 25 Dec 2024 · 0 repositories · arXiv:2412.18729
-
Using Large Language Models for Automated Grading of Student Writing about Science 25 Dec 2024 · 0 repositories · arXiv:2412.18719
-
Decentralized Intelligence in GameFi: Embodied AI Agents and the Convergence of DeFi and Virtual Ecosystems 24 Dec 2024 · 1 repository · arXiv:2412.18601
-
EvoPat: A Multi-LLM-based Patents Summarization and Analysis Agent 24 Dec 2024 · 0 repositories · arXiv:2412.18100
-
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English 24 Dec 2024 · 1 repository · arXiv:2412.18415
-
TimelyLLM: Segmented LLM Serving System for Time-sensitive Robotic Applications 24 Dec 2024 · 0 repositories · arXiv:2412.18695
-
Multimodal Preference Data Synthetic Alignment with Reward Model 23 Dec 2024 · 1 repository · arXiv:2412.17417
-
On Fusing ChatGPT and Ensemble Learning in Discon-tinuous Named Entity Recognition in Health Corpora 22 Dec 2024 · 0 repositories · arXiv:2412.16976
-
SubstationAI: Multimodal Large Model-Based Approaches for Analyzing Substation Equipment Faults 22 Dec 2024 · 0 repositories · arXiv:2412.17077
-
Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline 20 Dec 2024 · 0 repositories · arXiv:2412.15660
-
Benchmarking LLMs and SLMs for patient reported outcomes 20 Dec 2024 · 0 repositories · arXiv:2412.16291
-
Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring 20 Dec 2024 · 0 repositories · arXiv:2412.16108
-
Humanlike Cognitive Patterns as Emergent Phenomena in Large Language Models 20 Dec 2024 · 0 repositories · arXiv:2412.15501
-
Linguistic Features Extracted by GPT-4 Improve Alzheimer's Disease Detection based on Spontaneous Speech 20 Dec 2024 · 1 repository · arXiv:2412.15772
-
PromptOptMe: Error-Aware Prompt Compression for LLM-based MT Evaluation Metrics 20 Dec 2024 · 0 repositories · arXiv:2412.16120
-
How good is GPT at writing political speeches for the White House? 19 Dec 2024 · 0 repositories · arXiv:2412.14617
-
Systematic Evaluation of Long-Context LLMs on Financial Concepts 19 Dec 2024 · 0 repositories · arXiv:2412.15386
-
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data 18 Dec 2024 · 1 repository · arXiv:2412.14276
-
PsyDT: Using LLMs to Construct the Digital Twin of Psychological Counselor with Personalized Counseling Style for Psychological Counseling 18 Dec 2024 · 1 repository · arXiv:2412.13660
-
Reinforcement Learning from Automatic Feedback for High-Quality Unit Test Generation 18 Dec 2024 · 0 repositories · arXiv:2412.14308
-
JudgeBlender: Ensembling Judgments for Automatic Relevance Assessment 17 Dec 2024 · 1 repository · arXiv:2412.13268
-
Can Language Models Rival Mathematics Students? Evaluating Mathematical Reasoning through Textual Manipulation and Human Experiments 16 Dec 2024 · 0 repositories · arXiv:2412.11908
-
OpenReviewer: A Specialized Large Language Model for Generating Critical Scientific Paper Reviews 16 Dec 2024 · 0 repositories · arXiv:2412.11948
-
Second Language (Arabic) Acquisition of LLMs via Progressive Vocabulary Expansion 16 Dec 2024 · 0 repositories · arXiv:2412.12310
-
The Impact of AI Assistance on Radiology Reporting: A Pilot Study Using Simulated AI Draft Reports 16 Dec 2024 · 0 repositories · arXiv:2412.12042
-
The Open Source Advantage in Large Language Models (LLMs) 16 Dec 2024 · 0 repositories · arXiv:2412.12004
-
Smaller Language Models Are Better Instruction Evolvers 15 Dec 2024 · 1 repository · arXiv:2412.11231
-
MedG-KRP: Medical Graph Knowledge Representation Probing 14 Dec 2024 · 1 repository · arXiv:2412.10982
-
SusGen-GPT: A Data-Centric LLM for Financial NLP and Sustainability Report Generation 14 Dec 2024 · 1 repository · arXiv:2412.10906
-
WHAT-IF: Exploring Branching Narratives by Meta-Prompting Large Language Models 13 Dec 2024 · 0 repositories · arXiv:2412.10582
-
Adversarial Vulnerabilities in Large Language Models for Time Series Forecasting 11 Dec 2024 · 1 repository · arXiv:2412.08099
-
Assessing Personalized AI Mentoring with Large Language Models in the Computing Field 11 Dec 2024 · 0 repositories · arXiv:2412.08430
-
Automatic Item Generation for Personality Situational Judgment Tests with Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.12144
-
BiMediX2: Bio-Medical EXpert LMM for Diverse Medical Modalities 10 Dec 2024 · 1 repository · arXiv:2412.07769
-
ConceptSearch: Towards Efficient Program Search Using LLMs for Abstraction and Reasoning Corpus (ARC) 10 Dec 2024 · 1 repository · arXiv:2412.07322
-
Generating Knowledge Graphs from Large Language Models: A Comparative Study of GPT-4, LLaMA 2, and BERT 10 Dec 2024 · 0 repositories · arXiv:2412.07412
-
Ontology-driven Prompt Tuning for LLM-based Task and Motion Planning 10 Dec 2024 · 0 repositories · arXiv:2412.07493