Methods › Natural Language Processing › Transformers › GPT-3
GPT-3
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
GPT-3 is an autoregressive transformer model with 175 billion parameters. It uses the same architecture/model as GPT-2, including the modified initialization, pre-normalization, and reversible tokenization, with the exception that GPT-3 uses alternating dense and locally banded sparse attention patterns in the layers of the transformer, similar to the Sparse Transformer.
Papers archive 2025-07-28
30 shown of 1,906, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Augmenting Large Language Models with Static Code Analysis for Automated Code Quality Improvements 12 Jun 2025 · 0 repositories · arXiv:2506.10330
-
Think before You Simulate: Symbolic Reasoning to Orchestrate Neural Computation for Counterfactual Question Answering 12 Jun 2025 · 1 repository · arXiv:2506.10753
-
Multilingual Hate Speech Detection in Social Media Using Translation-Based Approaches with Large Language Models 9 Jun 2025 · 0 repositories · arXiv:2506.08147
-
Direct Behavior Optimization: Unlocking the Potential of Lightweight LLMs 6 Jun 2025 · 0 repositories · arXiv:2506.06401
-
Benchmarking Large Language Models on Homework Assessment in Circuit Analysis 5 Jun 2025 · 0 repositories · arXiv:2506.06390
-
Multiple-Choice Question Generation Using Large Language Models: Methodology and Educator Insights 5 Jun 2025 · 0 repositories · arXiv:2506.04851
-
Facts are Harder Than Opinions -- A Multilingual, Comparative Analysis of LLM-Based Fact-Checking Reliability 4 Jun 2025 · 0 repositories · arXiv:2506.03655
-
FinBERT2: A Specialized Bidirectional Encoder for Bridging the Gap in Finance-Specific Deployment of Large Language Models 31 May 2025 · 0 repositories · arXiv:2506.06335
-
Critical Batch Size Revisited: A Simple Empirical Approach to Large-Batch Language Model Training 29 May 2025 · 0 repositories · arXiv:2505.23971
-
Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach 29 May 2025 · 0 repositories · arXiv:2505.23953
-
Say What You Mean: Natural Language Access Control with Large Language Models for Internet of Things 28 May 2025 · 0 repositories · arXiv:2505.23835
-
Conversational Lexicography: Querying Lexicographic Data on Knowledge Graphs with SPARQL through Natural Language 26 May 2025 · 0 repositories · arXiv:2505.19971
-
Generative AI and Creativity: A Systematic Literature Review and Meta-Analysis 22 May 2025 · 1 repository · arXiv:2505.17241
-
Adversarial Testing in LLMs: Insights into Decision-Making Vulnerabilities 19 May 2025 · 0 repositories · arXiv:2505.13195
-
Are Large Language Models Good at Detecting Propaganda? 19 May 2025 · 0 repositories · arXiv:2505.13706
-
EVALOOP: Assessing LLM Robustness in Programming from a Self-consistency Perspective 18 May 2025 · 0 repositories · arXiv:2505.12185
-
Let the Trial Begin: A Mock-Court Approach to Vulnerability Detection using LLM-Based Agents 16 May 2025 · 0 repositories · arXiv:2505.10961
-
Comparing LLM Text Annotation Skills: A Study on Human Rights Violations in Social Media Data 15 May 2025 · 1 repository · arXiv:2505.10260
-
Achieving Scalable Robot Autonomy via neurosymbolic planning using lightweight local LLM 13 May 2025 · 1 repository · arXiv:2505.08492
-
Evaluating the Effectiveness of Black-Box Prompt Optimization as the Scale of LLMs Continues to Grow 13 May 2025 · 0 repositories · arXiv:2505.08303
-
HealthBench: Evaluating Large Language Models Towards Improved Human Health 13 May 2025 · 1 repository · arXiv:2505.08775Syntology ran 3 of 3 samples · 0 unverified
-
GRADA: Graph-based Reranker against Adversarial Documents Attack 12 May 2025 · 1 repository · arXiv:2505.07546
-
REFINE-AF: A Task-Agnostic Framework to Align Language Models via Self-Generated Instructions using Reinforcement Learning from Automated Feedback 10 May 2025 · 0 repositories · arXiv:2505.06548
-
An empathic GPT-based chatbot to talk about mental disorders with Spanish teenagers 9 May 2025 · 0 repositories · arXiv:2505.05828
-
What Is Next for LLMs? Next-Generation AI Computing Hardware Using Photonic Chips 9 May 2025 · 0 repositories · arXiv:2505.05794
-
Performance Evaluation of Large Language Models in Bangla Consumer Health Query Summarization 8 May 2025 · 0 repositories · arXiv:2505.05070
-
Bringing legal knowledge to the public by constructing a legal question bank using large-scale pre-trained language model 7 May 2025 · 0 repositories · arXiv:2505.04132
-
A Domain Adaptation of Large Language Models for Classifying Mechanical Assembly Components 2 May 2025 · 0 repositories · arXiv:2505.01627
-
JaccDiv: A Metric and Benchmark for Quantifying Diversity of Generated Marketing Text in the Music Industry 29 Apr 2025 · 0 repositories · arXiv:2504.20849
-
VeriDebug: A Unified LLM for Verilog Debugging via Contrastive Embedding and Guided Correction 27 Apr 2025 · 0 repositories · arXiv:2504.19099
Tasks archive 2025-07-28
20 shown of 665 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections