Methods › Natural Language Processing › Transformers › ALBERT › Papers, page 2
ALBERT
Papers archive 2025-07-28
archive papers tagged: 172 · with a code link: 68 · where Syntology ran a sample: 14 (9 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (14 of 172 tagged: 9 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument)
Page 2 of 2: papers 101 to 172 of 172, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
NeuralLog: Natural Language Inference with Joint Neural and Logical Reasoning 29 May 2021 · 1 repository · arXiv:2105.14167Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Using Adversarial Attacks to Reveal the Statistical Bias in Machine Reading Comprehension Models 24 May 2021 · 0 repositories · arXiv:2105.11136
-
Fine-Tuned Transformers Show Clusters of Similar Representations Across Layers 16 May 2021 · 0 repositories
-
Which transformer architecture fits my data? A vocabulary bottleneck in self-attention 9 May 2021 · 0 repositories · arXiv:2105.03928
-
When to Fold'em: How to answer Unanswerable questions 1 May 2021 · 1 repository · arXiv:2105.00328
-
Optimizing small BERTs trained for German NER 23 Apr 2021 · 2 repositories · arXiv:2104.11559
-
Transformers: "The End of History" for NLP? 9 Apr 2021 · 0 repositories · arXiv:2105.00813
-
Layer Reduction: Accelerating Conformer-Based Self-Supervised Model via Layer Consistency 8 Apr 2021 · 0 repositories · arXiv:2105.00812
-
MCL@IITK at SemEval-2021 Task 2: Multilingual and Cross-lingual Word-in-Context Disambiguation using Augmented Data, Signals, and Transformers 4 Apr 2021 · 0 repositories · arXiv:2104.01567
-
ReCAM@IITK at SemEval-2021 Task 4: BERT and ALBERT based Ensemble for Abstract Word Prediction 4 Apr 2021 · 1 repository · arXiv:2104.01563
-
Czert -- Czech BERT-like Model for Language Representation 24 Mar 2021 · 1 repository · arXiv:2103.13031
-
Text Mining of Stocktwits Data for Predicting Stock Prices 13 Mar 2021 · 0 repositories · arXiv:2103.16388
-
Hopeful_Men@LT-EDI-EACL2021: Hope Speech Detection Using Indic Transliteration and Transformers 24 Feb 2021 · 0 repositories · arXiv:2102.12082
-
Towards Emotion Recognition in Hindi-English Code-Mixed Data: A Transformer Based Approach 19 Feb 2021 · 1 repository · arXiv:2102.09943
-
Improved Customer Transaction Classification using Semi-Supervised Knowledge Distillation 15 Feb 2021 · 0 repositories · arXiv:2102.07635
-
Scaling Federated Learning for Fine-tuning of Large Language Models 1 Feb 2021 · 0 repositories · arXiv:2102.00875
-
The characteristic equation of the exceptional Jordan algebra: its eigenvalues, and their possible connection with the mass ratios of quarks and leptons 30 Jan 2021 · 0 repositories
-
Consequence of Enterprise Resource Planning in the Environs of Pedagogical Organization. 28 Jan 2021 · 0 repositories
-
Consequence of Enterprise Resource Planning in the Environs of Pedagogical Organization. 28 Jan 2021 · 0 repositories
-
KoreALBERT: Pretraining a Lite BERT Model for Korean Language Understanding 27 Jan 2021 · 0 repositories · arXiv:2101.11363
-
Evaluation of BERT and ALBERT Sentence Embedding Performance on Downstream NLP Tasks 26 Jan 2021 · 0 repositories · arXiv:2101.10642
-
Transformer-Based Models for Question Answering on COVID19 16 Jan 2021 · 0 repositories · arXiv:2101.11432
-
Grid Search Hyperparameter Benchmarking of BERT, ALBERT, and LongFormer on DuoRC 15 Jan 2021 · 0 repositories · arXiv:2101.06326
-
Transformer based Automatic COVID-19 Fake News Detection System 1 Jan 2021 · 2 repositories · arXiv:2101.00180
-
A Graph Reasoning Network for Multi-turn Response Selection via Customized Pre-training 21 Dec 2020 · 0 repositories · arXiv:2012.11099
-
Detecting Insincere Questions from Text: A Transfer Learning Approach 7 Dec 2020 · 1 repository · arXiv:2012.07587
-
Deep Learning Brasil - NLP at SemEval-2020 Task 9: Sentiment Analysis of Code-Mixed Tweets Using Ensemble of Language Models 1 Dec 2020 · 0 repositories
-
Lee at SemEval-2020 Task 5: ALBERT Model Based on the Maximum Ensemble Strategy and Different Data Sampling Methods for Detecting Counterfactual Statements 1 Dec 2020 · 0 repositories
-
Lijunyi at SemEval-2020 Task 4: An ALBERT Model Based Maximum Ensemble with Different Training Sizes and Depths for Commonsense Validation and Explanation 1 Dec 2020 · 0 repositories
-
Robust Machine Reading Comprehension by Learning Soft labels 1 Dec 2020 · 0 repositories
-
TeamJUST at SemEval-2020 Task 4: Commonsense Validation and Explanation Using Ensembling Techniques 1 Dec 2020 · 0 repositories
-
Warren at SemEval-2020 Task 4: ALBERT and Multi-Task Learning for Commonsense Validation 1 Dec 2020 · 0 repositories
-
EdgeBERT: Sentence-Level Energy Optimizations for Latency-Aware Multi-Task NLP Inference 28 Nov 2020 · 0 repositories · arXiv:2011.14203
-
Two Stage Transformer Model for COVID-19 Fake News Detection and Fact Checking 26 Nov 2020 · 1 repository · arXiv:2011.13253
-
IndicNLPSuite: Monolingual Corpora, Evaluation Benchmarks and Pre-trained Multilingual Language Models for Indian Languages 8 Nov 2020 · 1 repository
-
A Transformer Based Pitch Sequence Autoencoder with MIDI Augmentation 15 Oct 2020 · 0 repositories · arXiv:2010.07758
-
An Investigation on Different Underlying Quantization Schemes for Pre-trained Language Models 14 Oct 2020 · 0 repositories · arXiv:2010.07109
-
Infusing Disease Knowledge into BERT for Health Question Answering, Medical Inference and Disease Name Recognition 8 Oct 2020 · 1 repository · arXiv:2010.03746
-
On the Interplay Between Fine-tuning and Sentence-level Probing for Linguistic Knowledge in Pre-trained Transformers 6 Oct 2020 · 0 repositories · arXiv:2010.02616
-
Pretrained Language Model Embryology: The Birth of ALBERT 6 Oct 2020 · 1 repository · arXiv:2010.02480
-
A Technical Question Answering System with Transfer Learning 1 Oct 2020 · 1 repository
-
Interpretable Machine Learning for COVID-19: An Empirical Study on Severity Prediction Task 30 Sep 2020 · 1 repository · arXiv:2010.02006
-
BET: A Backtranslation Approach for Easy Data Augmentation in Transformer-based Paraphrase Identification Context 25 Sep 2020 · 1 repository · arXiv:2009.12452
-
BioALBERT: A Simple and Effective Pre-trained Language Model for Biomedical Named Entity Recognition 19 Sep 2020 · 0 repositories · arXiv:2009.09223
-
Learning Universal Representations from Word to Sentence 10 Sep 2020 · 0 repositories · arXiv:2009.04656
-
Comparative Study of Language Models on Cross-Domain Data with Model Agnostic Explainability 9 Sep 2020 · 0 repositories · arXiv:2009.04095
-
ERNIE at SemEval-2020 Task 10: Learning Word Emphasis Selection by Pre-trained Language Model 8 Sep 2020 · 0 repositories · arXiv:2009.03706
-
UPB at SemEval-2020 Task 8: Joint Textual and Visual Modeling in a Multi-Task Learning Architecture for Memotion Analysis 6 Sep 2020 · 0 repositories · arXiv:2009.02779
-
Deep Learning Brasil -- NLP at SemEval-2020 Task 9: Overview of Sentiment Analysis of Code-Mixed Tweets 28 Jul 2020 · 0 repositories · arXiv:2008.01544
-
Variants of BERT, Random Forests and SVM approach for Multimodal Emotion-Target Sub-challenge 28 Jul 2020 · 0 repositories · arXiv:2007.13928
-
Mono vs Multilingual Transformer-based Models: a Comparison across Several Language Tasks 19 Jul 2020 · 1 repository · arXiv:2007.09757
-
LMVE at SemEval-2020 Task 4: Commonsense Validation and Explanation using Pretraining Language Model 6 Jul 2020 · 0 repositories · arXiv:2007.02540
-
A Transformer Approach to Contextual Sarcasm Detection in Twitter 1 Jul 2020 · 0 repositories
-
BERTology Meets Biology: Interpreting Attention in Protein Language Models 26 Jun 2020 · 2 repositories · arXiv:2006.15222Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
New Vietnamese Corpus for Machine Reading Comprehension of Health News Articles 19 Jun 2020 · 0 repositories · arXiv:2006.11138
-
On the Stability of Fine-tuning BERT: Misconceptions, Explanations, and Strong Baselines 8 Jun 2020 · 2 repositories · arXiv:2006.04884Syntology official (archive's flag): 8 ran · 13 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 0 violated, 6 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 20 harvested samples) · 3 pointer-only (licence)
-
BERT Loses Patience: Fast and Robust Inference with Early Exit 7 Jun 2020 · 1 repository · arXiv:2006.04152
-
BERT-based Ensembles for Modeling Disclosure and Support in Conversational Social Media Text 1 Jun 2020 · 0 repositories · arXiv:2006.01222
-
Language Representation Models for Fine-Grained Sentiment Classification 27 May 2020 · 1 repository · arXiv:2005.13619
-
Audio ALBERT: A Lite BERT for Self-supervised Learning of Audio Representation 18 May 2020 · 4 repositories · arXiv:2005.08575
-
Value-at-Risk substitute for non-ruin capital is fallacious and redundant 11 May 2020 · 0 repositories · arXiv:2005.05428
-
ImpactCite: An XLNet-based method for Citation Impact Analysis 5 May 2020 · 1 repository · arXiv:2005.06611
-
TAVAT: Token-Aware Virtual Adversarial Training for Language Understanding 30 Apr 2020 · 1 repository · arXiv:2004.14543
-
Recall and Learn: Fine-tuning Deep Pretrained Language Models with Less Forgetting 27 Apr 2020 · 1 repository · arXiv:2004.12651
-
UHH-LT at SemEval-2020 Task 12: Fine-Tuning of Pre-Trained Transformer Networks for Offensive Language Detection 23 Apr 2020 · 0 repositories · arXiv:2004.11493
-
Investigating the Effectiveness of Representations Based on Pretrained Transformer-based Language Models in Active Learning for Labelling Text Datasets 21 Apr 2020 · 0 repositories · arXiv:2004.13138
-
Gestalt: a Stacking Ensemble for SQuAD2.0 2 Apr 2020 · 0 repositories · arXiv:2004.07067
-
Deep Entity Matching with Pre-Trained Language Models 1 Apr 2020 · 1 repository · arXiv:2004.00584Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Retrospective Reader for Machine Reading Comprehension 27 Jan 2020 · 2 repositories · arXiv:2001.09694
-
PoWER-BERT: Accelerating BERT Inference via Progressive Word-vector Elimination 24 Jan 2020 · 1 repository · arXiv:2001.08950Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Perceiving the arrow of time in autoregressive motion 1 Dec 2019 · 0 repositories
-
ALBERT: A Lite BERT for Self-supervised Learning of Language Representations 26 Sep 2019 · 48 repositories · arXiv:1909.11942Syntology official (archive's flag): 6 ran · 81 ran (of which 17 constructed an object rather than computing a result; 59 with no instrument failure: 4 honoured, 0 violated, 55 with no contract checked; 22 where Syntology's instrument failed) · 45 unverified (of 126 harvested samples) · 28 pointer-only (licence)