Methods › General › Regularization › Weight Decay › Papers, page 66
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 66 of 108: papers 6,501 to 6,600 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BiTimeBERT: Extending Pre-Trained Language Representations with Bi-Temporal Information 27 Apr 2022 · 0 repositories · arXiv:2204.13032
-
MILES: Visual BERT Pre-training with Injected Language Semantics for Video-text Retrieval 26 Apr 2022 · 1 repository · arXiv:2204.12408
-
PLOD: An Abbreviation Detection Dataset for Scientific Documents 26 Apr 2022 · 1 repository · arXiv:2204.12061
-
Pretraining Chinese BERT for Detecting Word Insertion and Deletion Errors 26 Apr 2022 · 0 repositories · arXiv:2204.12052
-
Evaluating Interpolation and Extrapolation Performance of Neural Retrieval Models 25 Apr 2022 · 1 repository · arXiv:2204.11447
-
Groupwise Query Performance Prediction with BERT 25 Apr 2022 · 1 repository · arXiv:2204.11489
-
Faster Learned Sparse Retrieval with Guided Traversal 24 Apr 2022 · 1 repository · arXiv:2204.11314
-
Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training 24 Apr 2022 · 1 repository · arXiv:2204.11218
-
Fine-Tuning BERT Models to Classify Misinformation on Garlic and COVID-19 on Twitter 22 Apr 2022 · 0 repositories
-
Taygete at SemEval-2022 Task 4: RoBERTa based models for detecting Patronising and Condescending Language 22 Apr 2022 · 0 repositories · arXiv:2204.10519
-
WaBERT: A Low-resource End-to-end Model for Spoken Language Understanding and Speech-to-BERT Alignment 22 Apr 2022 · 0 repositories · arXiv:2204.10461
-
Measuring artificial intelligence: a systematic assessment and implications for governance 21 Apr 2022 · 0 repositories · arXiv:2204.10304
-
Is Neural Topic Modelling Better than Clustering? An Empirical Study on Clustering with Contextual Embeddings for Topics 21 Apr 2022 · 1 repository · arXiv:2204.09874Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Revisiting Gaussian mixture critics in off-policy reinforcement learning: a sample-based approach 21 Apr 2022 · 1 repository · arXiv:2204.10256
-
Deep Learning meets Nonparametric Regression: Are Weight-Decayed DNNs Locally Adaptive? 20 Apr 2022 · 0 repositories · arXiv:2204.09664
-
Is BERT Robust to Label Noise? A Study on Learning with Noisy Labels in Text Classification 20 Apr 2022 · 1 repository · arXiv:2204.09371Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
ALBETO and DistilBETO: Lightweight Spanish Language Models 19 Apr 2022 · 2 repositories · arXiv:2204.09145
-
CodexDB: Generating Code for Processing SQL Queries using GPT-3 Codex 19 Apr 2022 · 0 repositories · arXiv:2204.08941
-
DecBERT: Enhancing the Language Understanding of BERT with Causal Attention Masks 19 Apr 2022 · 0 repositories · arXiv:2204.08688
-
Impact of Tokenization on Language Models: An Analysis for Turkish 19 Apr 2022 · 0 repositories · arXiv:2204.08832
-
LitMC-BERT: transformer-based multi-label classification of biomedical literature with an application on COVID-19 literature curation 19 Apr 2022 · 1 repository · arXiv:2204.08649
-
Mono vs Multilingual BERT for Hate Speech Detection and Text Classification: A Case Study in Marathi 19 Apr 2022 · 0 repositories · arXiv:2204.08669
-
Multimodal Hate Speech Detection from Bengali Memes and Texts 19 Apr 2022 · 1 repository · arXiv:2204.10196
-
Probing for the Usage of Grammatical Number 19 Apr 2022 · 0 repositories · arXiv:2204.08831
-
Ingredient Extraction from Text in the Recipe Domain 18 Apr 2022 · 1 repository · arXiv:2204.08137
-
L3Cube-HingCorpus and HingBERT: A Code Mixed Hindi-English Dataset and BERT Language Models 18 Apr 2022 · 1 repository · arXiv:2204.08398
-
UTNLP at SemEval-2022 Task 6: A Comparative Analysis of Sarcasm Detection Using Generative-based and Mutation-based Data Augmentation 18 Apr 2022 · 2 repositories · arXiv:2204.08198
-
Zero-shot Entity and Tweet Characterization with Designed Conditional Prompts and Contexts 18 Apr 2022 · 0 repositories · arXiv:2204.08405
-
Knowledgeable Salient Span Mask for Enhancing Language Models as Knowledge Base 17 Apr 2022 · 0 repositories · arXiv:2204.07994
-
Pathologies of Pre-trained Language Models in Few-shot Fine-tuning 17 Apr 2022 · 0 repositories · arXiv:2204.08039
-
Probing Script Knowledge from Pre-Trained Models 16 Apr 2022 · 0 repositories · arXiv:2204.10176
-
SimpleBERT: A Pre-trained Model That Learns to Generate Simple Words 16 Apr 2022 · 0 repositories · arXiv:2204.07779
-
Improving Pre-trained Language Models with Syntactic Dependency Prediction Task for Chinese Semantic Error Recognition 15 Apr 2022 · 0 repositories · arXiv:2204.07464
-
mGPT: Few-Shot Learners Go Multilingual 15 Apr 2022 · 1 repository · arXiv:2204.07580
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 15 Apr 2022 · 1 repository · arXiv:2204.07483
-
Resource-Aware Distributed Submodular Maximization: A Paradigm for Multi-Robot Decision-Making 15 Apr 2022 · 0 repositories · arXiv:2204.07520
-
Analysing similarities between legal court documents using natural language processing approaches based on Transformers 14 Apr 2022 · 0 repositories · arXiv:2204.07182
-
CalBERT - Code-mixed Adaptive Language representations using BERT 14 Apr 2022 · 1 repository
-
Does BERT really agree ? Fine-grained Analysis of Lexical Dependence on a Syntactic Task 14 Apr 2022 · 0 repositories · arXiv:2204.06889
-
GPT-NeoX-20B: An Open-Source Autoregressive Language Model 14 Apr 2022 · 11 repositories · arXiv:2204.06745Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Latent Aspect Detection from Online Unsolicited Customer Reviews 14 Apr 2022 · 1 repository · arXiv:2204.06964
-
Multi-label topic classification for COVID-19 literature with Bioformer 14 Apr 2022 · 0 repositories · arXiv:2204.06758
-
Rows from Many Sources: Enriching row completions from Wikidata with a pre-trained Language Model 14 Apr 2022 · 0 repositories · arXiv:2204.07014
-
HuBERT-EE: Early Exiting HuBERT for Efficient Speech Recognition 13 Apr 2022 · 1 repository · arXiv:2204.06328
-
IIITDWD-ShankarB@ Dravidian-CodeMixi-HASOC2021: mBERT based model for identification of offensive content in south Indian languages 13 Apr 2022 · 0 repositories · arXiv:2204.10195
-
METRO: Efficient Denoising Pretraining of Large Scale Autoencoding Language Models with Model Generated Signals 13 Apr 2022 · 0 repositories · arXiv:2204.06644
-
Probing for Constituency Structure in Neural Language Models 13 Apr 2022 · 1 repository · arXiv:2204.06201Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TangoBERT: Reducing Inference Cost by using Cascaded Architecture 13 Apr 2022 · 0 repositories · arXiv:2204.06271
-
Team ÚFAL at CMCL 2022 Shared Task: Figuring out the correct recipe for predicting Eye-Tracking features using Pretrained Language Models 11 Apr 2022 · 0 repositories · arXiv:2204.04998
-
Tokenwise Contrastive Pretraining for Finer Speech-to-BERT Alignment in End-to-End Speech-to-Intent Systems 11 Apr 2022 · 0 repositories · arXiv:2204.05188
-
Towards Generalizable Semantic Product Search by Text Similarity Pre-training on Search Click Logs 11 Apr 2022 · 0 repositories · arXiv:2204.05231
-
Uniform Complexity for Text Generation 11 Apr 2022 · 1 repository · arXiv:2204.05185
-
Confidence Estimation Transformer for Long-term Renewable Energy Forecasting in Reinforcement Learning-based Power Grid Dispatching 10 Apr 2022 · 1 repository · arXiv:2204.04612
-
Fake news detection using parallel BERT deep neural networks 10 Apr 2022 · 0 repositories · arXiv:2204.04793
-
Few-Shot Cross-lingual Transfer for Coarse-grained De-identification of Code-Mixed Clinical Texts 10 Apr 2022 · 1 repository · arXiv:2204.04775
-
MA-Dreamer: Coordination and communication through shared imagination 10 Apr 2022 · 0 repositories · arXiv:2204.04687
-
Pushing on Personality Detection from Verbal Behavior: A Transformer Meets Text Contours of Psycholinguistic Features 10 Apr 2022 · 0 repositories · arXiv:2204.04629
-
FoundationLayerNorm: Scaling BERT and GPT to 1,000 Layers 9 Apr 2022 · 0 repositories · arXiv:2204.04477
-
Modeling Multi-Granularity Hierarchical Features for Relation Extraction 9 Apr 2022 · 1 repository · arXiv:2204.04437Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Are We Really Making Much Progress in Text Classification? A Comparative Review 8 Apr 2022 · 1 repository · arXiv:2204.03954
-
Contextual Representation Learning beyond Masked Language Modeling 8 Apr 2022 · 1 repository · arXiv:2204.04163
-
Infusing Knowledge from Wikipedia to Enhance Stance Detection 8 Apr 2022 · 2 repositories · arXiv:2204.03839
-
Towards Understanding Large-Scale Discourse Structures in Pre-Trained and Fine-Tuned Language Models 8 Apr 2022 · 0 repositories · arXiv:2204.04289
-
Accelerating Attention through Gradient-Based Learned Runtime Pruning 7 Apr 2022 · 0 repositories · arXiv:2204.03227
-
Autoencoding Language Model Based Ensemble Learning for Commonsense Validation and Explanation 7 Apr 2022 · 0 repositories · arXiv:2204.03324
-
BERTuit: Understanding Spanish language in Twitter through a native transformer 7 Apr 2022 · 0 repositories · arXiv:2204.03465
-
CoCoSoDa: Effective Contrastive Learning for Code Search 7 Apr 2022 · 0 repositories · arXiv:2204.03293
-
PALBERT: Teaching ALBERT to Ponder 7 Apr 2022 · 1 repository · arXiv:2204.03276Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Pretraining Text Encoders with Adversarial Mixture of Training Signal Generators 7 Apr 2022 · 1 repository · arXiv:2204.03243Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Testing the limits of natural language models for predicting human language judgments 7 Apr 2022 · 1 repository · arXiv:2204.03592
-
The Effects of Regularization and Data Augmentation are Class Dependent 7 Apr 2022 · 0 repositories · arXiv:2204.03632
-
drsphelps at SemEval-2022 Task 2: Learning idiom representations using BERTRAM 6 Apr 2022 · 0 repositories · arXiv:2204.02821
-
Fusing finetuned models for better pretraining 6 Apr 2022 · 2 repositories · arXiv:2204.03044
-
Knowledge Infused Decoding 6 Apr 2022 · 1 repository · arXiv:2204.03084Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Paying More Attention to Self-attention: Improving Pre-trained Language Models via Attention Guiding 6 Apr 2022 · 0 repositories · arXiv:2204.02922
-
Using Synthetic Data for Conversational Response Generation in Low-resource Settings 6 Apr 2022 · 0 repositories · arXiv:2204.02653
-
Abstractive summarization of hospitalisation histories with transformer networks 5 Apr 2022 · 0 repositories · arXiv:2204.02208
-
An Exploratory Study on Code Attention in BERT 5 Apr 2022 · 0 repositories · arXiv:2204.10200
-
Data Augmentation for Intent Classification with Off-the-shelf Large Language Models 5 Apr 2022 · 1 repository · arXiv:2204.01959Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
How Different are Pre-trained Transformers for Text Ranking? 5 Apr 2022 · 1 repository · arXiv:2204.07233
-
Multilinguals at SemEval-2022 Task 11: Transformer Based Architecture for Complex NER 5 Apr 2022 · 1 repository · arXiv:2204.02173
-
POS-BERT: Point Cloud One-Stage BERT Pre-Training 3 Apr 2022 · 1 repository · arXiv:2204.00989
-
BERT-Assisted Semantic Annotation Correction for Emotion-Related Questions 2 Apr 2022 · 1 repository · arXiv:2204.00916
-
Efficient comparison of sentence embeddings 2 Apr 2022 · 0 repositories · arXiv:2204.00820
-
Analyzing how BERT performs entity matching 1 Apr 2022 · 1 repository
-
CharacterBERT and Self-Teaching for Improving the Robustness of Dense Retrievers on Queries with Typos 1 Apr 2022 · 1 repository · arXiv:2204.00716
-
Cyberbullying detection across social media platforms via platform-aware adversarial encoding 1 Apr 2022 · 0 repositories · arXiv:2204.00334
-
Effect and Analysis of Large-scale Language Model Rescoring on Competitive ASR Systems 1 Apr 2022 · 0 repositories · arXiv:2204.00212
-
Feature Structure Distillation with Centered Kernel Alignment in BERT Transferring 1 Apr 2022 · 1 repository · arXiv:2204.08922
-
Monarch: Expressive Structured Matrices for Efficient and Accurate Training 1 Apr 2022 · 2 repositories · arXiv:2204.00595Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 1 violated, 13 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 15 harvested samples)
-
Syntax-informed Question Answering with Heterogeneous Graph Transformer 1 Apr 2022 · 0 repositories · arXiv:2204.09655
-
A Baseline Readability Model for Cebuano 31 Mar 2022 · 1 repository · arXiv:2203.17225
-
A Character-level Span-based Model for Mandarin Prosodic Structure Prediction 31 Mar 2022 · 1 repository · arXiv:2203.16922
-
CatIss: An Intelligent Tool for Categorizing Issues Reports using Transformers 31 Mar 2022 · 1 repository · arXiv:2203.17196
-
CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation 31 Mar 2022 · 0 repositories · arXiv:2203.16763
-
ESGBERT: Language Model to Help with Classification Tasks Related to Companies Environmental, Social, and Governance Practices 31 Mar 2022 · 0 repositories · arXiv:2203.16788
-
Generative Pre-Trained Transformers for Biologically Inspired Design 31 Mar 2022 · 0 repositories · arXiv:2204.09714
-
Leveraging pre-trained language models for conversational information seeking from text 31 Mar 2022 · 0 repositories · arXiv:2204.03542
-
Mixed-Phoneme BERT: Improving BERT with Mixed Phoneme and Sup-Phoneme Representations for Text to Speech 31 Mar 2022 · 0 repositories · arXiv:2203.17190
-
Incorporating Dynamic Semantics into Pre-Trained Language Model for Aspect-based Sentiment Analysis 30 Mar 2022 · 0 repositories · arXiv:2203.16369