Methods › Natural Language Processing › Language Models › GPT-4
GPT-4
Introduced by OpenAI et al. in GPT-4 Technical Report
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
GPT-4 is a transformer based model pre-trained to predict the next token in a document.
Papers archive 2025-07-28
30 shown of 2,870, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Foundation Models for Logistics: Toward Certifiable, Conversational Planning Interfaces 15 Jul 2025 · 0 repositories · arXiv:2507.11352
-
An Empirical Evaluation of AI-Powered Non-Player Characters' Perceived Realism and Performance in Virtual Reality Environments 14 Jul 2025 · 0 repositories · arXiv:2507.10469
-
Learning from Synthetic Labs: Language Models as Auction Participants 12 Jul 2025 · 0 repositories · arXiv:2507.09083
-
Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving 8 Jul 2025 · 1 repository · arXiv:2507.06229Syntology ran 1 of 1 samples · 0 unverified
-
Large Language Models Don't Make Sense of Word Problems. A Scoping Review from a Mathematics Education Perspective 30 Jun 2025 · 0 repositories · arXiv:2506.24006
-
Probing AI Safety with Source Code 25 Jun 2025 · 1 repository · arXiv:2506.20471
-
Lost in Translation? Converting RegExes for Log Parsing into Dynatrace Pattern Language 24 Jun 2025 · 0 repositories · arXiv:2506.19539
-
Security Assessment of DeepSeek and GPT Series Models against Jailbreak Attacks 23 Jun 2025 · 0 repositories · arXiv:2506.18543
-
SWE-SQL: Illuminating LLM Pathways to Solve User SQL Issues in Real-World Applications 23 Jun 2025 · 0 repositories · arXiv:2506.18951
-
Leveraging LLMs to Assess Tutor Moves in Real-Life Dialogues: A Feasibility Study 20 Jun 2025 · 0 repositories · arXiv:2506.17410
-
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making 20 Jun 2025 · 1 repository · arXiv:2506.17419
-
DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph Refinement 18 Jun 2025 · 2 repositories · arXiv:2506.15583Syntology ran 1 of 19 samples · 18 unverified · 19 pointer-only (licence)
-
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution 18 Jun 2025 · 0 repositories · arXiv:2506.17323
-
ImpliRet: Benchmarking the Implicit Fact Retrieval Challenge 17 Jun 2025 · 1 repository · arXiv:2506.14407Syntology ran 0 of 2 samples · 2 unverified · 2 pointer-only (licence)
-
Scaling Intelligence: Designing Data Centers for Next-Gen Language Models 17 Jun 2025 · 0 repositories · arXiv:2506.15006
-
Alphabet Index Mapping: Jailbreaking LLMs through Semantic Dissimilarity 15 Jun 2025 · 0 repositories · arXiv:2506.12685
-
Evaluating Cell Type Inference in Vision Language Models Under Varying Visual Context 15 Jun 2025 · 1 repository · arXiv:2506.12683
-
Language Models Enable Data-Augmented Synthesis Planning for Inorganic Materials 14 Jun 2025 · 0 repositories · arXiv:2506.12557
-
Identifying Helpful Context for LLM-based Vulnerability Repair: A Preliminary Study 13 Jun 2025 · 0 repositories · arXiv:2506.11561
-
Leveraging GPT-4 for Vulnerability-Witnessing Unit Test Generation 13 Jun 2025 · 0 repositories · arXiv:2506.11559
-
"Check My Work?": Measuring Sycophancy in a Simulated Educational Context 12 Jun 2025 · 1 repository · arXiv:2506.10297
-
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs 12 Jun 2025 · 0 repositories · arXiv:2506.10527
-
SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarks 12 Jun 2025 · 1 repository · arXiv:2506.10954
-
Think before You Simulate: Symbolic Reasoning to Orchestrate Neural Computation for Counterfactual Question Answering 12 Jun 2025 · 1 repository · arXiv:2506.10753
-
Can LLMs Generate Good Stories? Insights and Challenges from a Narrative Planning Perspective 11 Jun 2025 · 0 repositories · arXiv:2506.10161
-
Large Language Models for Toxic Language Detection in Low-Resource Balkan Languages 11 Jun 2025 · 1 repository · arXiv:2506.09992
-
Latent Multi-Head Attention for Small Language Models 11 Jun 2025 · 0 repositories · arXiv:2506.09342
-
Mutual-Supervised Learning for Sequential-to-Parallel Code Translation 11 Jun 2025 · 1 repository · arXiv:2506.11153
-
CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmark of Large Language Models in Mental Health Counseling 10 Jun 2025 · 1 repository · arXiv:2506.08584
-
TACTIC: Translation Agents with Cognitive-Theoretic Interactive Collaboration 10 Jun 2025 · 1 repository · arXiv:2506.08403
Tasks archive 2025-07-28
20 shown of 734 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
| Task | Papers |
|---|---|
| Language Modelling | 447 |
| Language Modeling | 322 |
| Large Language Model | 298 |
| Question Answering | 251 |
| Retrieval | 167 |
| Decision Making | 122 |
| In-Context Learning | 121 |
| Benchmarking | 120 |
| RAG | 111 |
| Code Generation | 110 |
| Retrieval-augmented Generation | 109 |
| Prompt Engineering | 105 |
| Text Generation | 99 |
| Math | 94 |
| Hallucination | 87 |
| Multiple-choice | 79 |
| Sentence | 69 |
| Instruction Following | 68 |
| Mathematical Reasoning | 58 |
| Chatbot | 57 |
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections