Methods › Natural Language Processing › Transformers › ERNIE
ERNIE
Introduced by Yu Sun et al. in ERNIE: Enhanced Representation through Knowledge Integration
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
ERNIE is a transformer-based model consisting of two stacked modules: 1) textual encoder and 2) knowledgeable encoder, which is responsible to integrate extra token-oriented knowledge information into textual information. This layer consists of stacked aggregators, designed for encoding both tokens and entities as well as fusing their heterogeneous features. To integrate this layer of enhancing representations via knowledge, a special pre-training task is adopted for ERNIE - it involves randomly masking token-entity alignments and training the model to predict all corresponding entities based on aligned tokens (aka denoising entity auto-encoder).
Papers archive 2025-07-28
30 shown of 54, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Evaluating Moral Beliefs across LLMs through a Pluralistic Framework 6 Nov 2024 · 1 repository · arXiv:2411.03665
-
The Potential of LLMs in Medical Education: Generating Questions and Answers for Qualification Exams 31 Oct 2024 · 0 repositories · arXiv:2410.23769
-
Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles 24 Jul 2024 · 0 repositories · arXiv:2407.17211
-
Unlocking the Potential: Benchmarking Large Language Models in Water Engineering and Research 22 Jul 2024 · 0 repositories · arXiv:2407.21045
-
The Solution for the AIGC Inference Performance Optimization Competition 6 Jul 2024 · 0 repositories · arXiv:2407.04991
-
Chumor 1.0: A Truly Funny and Challenging Chinese Humor Understanding Dataset from Ruo Zhi Ba 18 Jun 2024 · 0 repositories · arXiv:2406.12754
-
NewsBench: A Systematic Evaluation Framework for Assessing Editorial Capabilities of Large Language Models in Chinese Journalism 29 Feb 2024 · 1 repository · arXiv:2403.00862Syntology ran 0 of 3 samples · 3 unverified
-
Research about the Ability of LLM in the Tamper-Detection Area 24 Jan 2024 · 0 repositories · arXiv:2401.13504
-
How Robust is Google's Bard to Adversarial Image Attacks? 21 Sep 2023 · 1 repository · arXiv:2309.11751Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)
-
Robust Multi-Agent Reinforcement Learning via Adversarial Regularization: Theoretical Foundation and Stable Algorithms 21 Sep 2023 · 1 repository
-
The Impact of Artificial Intelligence on the Evolution of Digital Education: A Comparative Study of OpenAI Text Generation Tools including ChatGPT, Bing Chat, Bard, and Ernie 5 Sep 2023 · 0 repositories · arXiv:2309.02029
-
An APT Event Extraction Method Based on BERT-BiGRU-CRF for APT Attack Detection 4 Aug 2023 · 0 repositories
-
Towards Effective Ancient Chinese Translation: Dataset, Model, and Evaluation 1 Aug 2023 · 1 repository · arXiv:2308.00240
-
HouYi: An open-source large language model specially designed for renewable energy and carbon neutrality field 31 Jul 2023 · 0 repositories · arXiv:2308.01414
-
MedGPTEval: A Dataset and Benchmark to Evaluate Responses of Large Language Models in Medicine 12 May 2023 · 0 repositories · arXiv:2305.07340
-
Technical Report: Impact of Position Bias on Language Models in Token Classification 26 Apr 2023 · 2 repositories · arXiv:2304.13567
-
Co-Driven Recognition of Semantic Consistency via the Fusion of Transformer and HowNet Sememes Knowledge 21 Feb 2023 · 1 repository · arXiv:2302.10570
-
ERNIE 3.0 Tiny: Frustratingly Simple Method to Improve Task-Agnostic Distillation Generalization 9 Jan 2023 · 1 repository · arXiv:2301.03416Syntology ran 0 of 3 samples · 3 unverified
-
Vote'n'Rank: Revision of Benchmarking with Social Choice Theory 11 Oct 2022 · 1 repository · arXiv:2210.05769Syntology ran 0 of 17 samples · 17 unverified
-
GLM-130B: An Open Bilingual Pre-trained Model 5 Oct 2022 · 9 repositories · arXiv:2210.02414Syntology ran 5 of 21 samples · 16 unverified
-
ARGUABLY@SMM4H’22: Classification of Health Related Tweets using Ensemble, Zero-Shot and Fine-Tuned Language Model 1 Oct 2022 · 0 repositories
-
Exploiting Word Semantics to Enrich Character Representations of Chinese Pre-trained Models 13 Jul 2022 · 1 repository · arXiv:2207.05928
-
Argumentative Text Generation in Economic Domain 18 Jun 2022 · 1 repository · arXiv:2206.09251
-
Automatic Summarization of Russian Texts: Comparison of Extractive and Abstractive Methods 18 Jun 2022 · 0 repositories · arXiv:2206.09253
-
Training on Lexical Resources 1 Jun 2022 · 1 repository
-
Enhancing Chinese Pre-trained Language Model via Heterogeneous Linguistics Graph 1 May 2022 · 3 repositories
-
IDIAP Submission@LT-EDI-ACL2022 : Hope Speech Detection for Equality, Diversity and Inclusion 1 May 2022 · 1 repository
-
Research on Dual Channel News Headline Classification Based on ERNIE Pre-training Model 14 Feb 2022 · 0 repositories · arXiv:2202.06600
-
What Has Been Enhanced in my Knowledge-Enhanced Language Model? 2 Feb 2022 · 1 repository · arXiv:2202.00964
-
ERNIE 3.0 Titan: Exploring Larger-scale Knowledge Enhanced Pre-training for Language Understanding and Generation 23 Dec 2021 · 3 repositories · arXiv:2112.12731
Tasks archive 2025-07-28
20 shown of 94 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections