Methods › General › Activation Functions › Gated Linear Unit › Papers, page 6
Gated Linear Unit
Papers archive 2025-07-28
archive papers tagged: 798 · with a code link: 400 · where Syntology ran a sample: 115 (99 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (115 of 798 tagged: 99 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 6 of 8: papers 501 to 600 of 798, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Scaling Instruction-Finetuned Language Models 20 Oct 2022 · 9 repositories · arXiv:2210.11416Syntology 8 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 7 where Syntology's instrument failed) · 9 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
Self-supervised Graph Masking Pre-training for Graph-to-Text Generation 19 Oct 2022 · 1 repository · arXiv:2210.10599
-
Swinv2-Imagen: Hierarchical Vision Transformer Diffusion Models for Text-to-Image Generation 18 Oct 2022 · 0 repositories · arXiv:2210.09549
-
"John is 50 years old, can his son be 65?" Evaluating NLP Models' Understanding of Feasibility 14 Oct 2022 · 1 repository · arXiv:2210.07471
-
Self-Repetition in Abstractive Neural Summarizers 14 Oct 2022 · 0 repositories · arXiv:2210.08145
-
Tone prediction and orthographic conversion for Basaa 13 Oct 2022 · 0 repositories · arXiv:2210.06986
-
Entity Tracking via Effective Use of Multi-Task Learning Model and Mention-guided Decoding 12 Oct 2022 · 2 repositories · arXiv:2210.06444Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Instruction Tuning for Few-Shot Aspect-Based Sentiment Analysis 12 Oct 2022 · 1 repository · arXiv:2210.06629
-
RankT5: Fine-Tuning T5 for Text Ranking with Ranking Losses 12 Oct 2022 · 0 repositories · arXiv:2210.10634
-
Are Pretrained Multilingual Models Equally Fair Across Languages? 11 Oct 2022 · 1 repository · arXiv:2210.05457
-
Reflection of Thought: Inversely Eliciting Numerical Reasoning in Language Models via Solving Linear Systems 11 Oct 2022 · 0 repositories · arXiv:2210.05075
-
T5 for Hate Speech, Augmented Data and Ensemble 11 Oct 2022 · 1 repository · arXiv:2210.05480
-
ASDOT: Any-Shot Data-to-Text Generation with Pretrained Language Models 9 Oct 2022 · 1 repository · arXiv:2210.04325
-
CHARD: Clinical Health-Aware Reasoning Across Dimensions for Text Generation Models 9 Oct 2022 · 1 repository · arXiv:2210.04191
-
Better Pre-Training by Reducing Representation Confusion 9 Oct 2022 · 0 repositories · arXiv:2210.04246
-
Generating Quizzes to Support Training on Quality Management and Assurance in Space Science and Engineering 7 Oct 2022 · 0 repositories · arXiv:2210.03427
-
How Large Language Models are Transforming Machine-Paraphrased Plagiarism 7 Oct 2022 · 3 repositories · arXiv:2210.03568
-
LLMEffiChecker: Understanding and Testing Efficiency Degradation of Large Language Models 7 Oct 2022 · 1 repository · arXiv:2210.03696
-
LambdaKG: A Library for Pre-trained Language Model-Based Knowledge Graph Embeddings 1 Oct 2022 · 2 repositories · arXiv:2210.00305
-
Bidirectional Language Models Are Also Few-shot Learners 29 Sep 2022 · 0 repositories · arXiv:2209.14500
-
WikiDes: A Wikipedia-Based Dataset for Generating Short Descriptions from Paragraphs 27 Sep 2022 · 1 repository · arXiv:2209.13101
-
Application of Deep Learning in Generating Structured Radiology Reports: A Transformer-Based Technique 25 Sep 2022 · 1 repository · arXiv:2209.12177
-
ET5: A Novel End-to-end Framework for Conversational Machine Reading Comprehension 23 Sep 2022 · 1 repository · arXiv:2209.11484
-
On Efficient Reinforcement Learning for Full-length Game of StarCraft II 23 Sep 2022 · 2 repositories · arXiv:2209.11553Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
XF2T: Cross-lingual Fact-to-Text Generation for Low-Resource Languages 22 Sep 2022 · 0 repositories · arXiv:2209.11252
-
T5QL: Taming language models for SQL generation 21 Sep 2022 · 0 repositories · arXiv:2209.10254
-
Chain of Explanation: New Prompting Method to Generate Higher Quality Natural Language Explanation for Implicit Hate Speech 11 Sep 2022 · 0 repositories · arXiv:2209.04889
-
Simple and Effective Gradient-Based Tuning of Sequence-to-Sequence Models 10 Sep 2022 · 0 repositories · arXiv:2209.04683
-
IDIAPers @ Causal News Corpus 2022: Extracting Cause-Effect-Signal Triplets via Pre-trained Autoregressive Language Model 8 Sep 2022 · 1 repository · arXiv:2209.03891
-
Interpreting Black-box Machine Learning Models for High Dimensional Datasets 29 Aug 2022 · 0 repositories · arXiv:2208.13405
-
MDIA: A Benchmark for Multilingual Dialogue Generation in 46 Languages 27 Aug 2022 · 1 repository · arXiv:2208.13078
-
AutoQGS: Auto-Prompt for Low-Resource Knowledge-based Question Generation from SPARQL 26 Aug 2022 · 1 repository · arXiv:2208.12461
-
Training a T5 Using Lab-sized Resources 25 Aug 2022 · 0 repositories · arXiv:2208.12097
-
Diverse Title Generation for Stack Overflow Posts with Multiple Sampling Enhanced Transformer 24 Aug 2022 · 1 repository · arXiv:2208.11523
-
MulZDG: Multilingual Code-Switching Framework for Zero-shot Dialogue Generation 18 Aug 2022 · 1 repository · arXiv:2208.08629
-
Summarizing Patients Problems from Hospital Progress Notes Using Pre-trained Sequence-to-Sequence Models 17 Aug 2022 · 0 repositories · arXiv:2208.08408
-
Continuous Active Learning Using Pretrained Transformers 15 Aug 2022 · 0 repositories · arXiv:2208.06955
-
A Boring-yet-effective Approach for the Product Ranking Task of the Amazon KDD Cup 2022 9 Aug 2022 · 0 repositories · arXiv:2208.06264
-
Neural Knowledge Bank for Pretrained Transformers 31 Jul 2022 · 0 repositories · arXiv:2208.00399
-
HorNet: Efficient High-Order Spatial Interactions with Recursive Gated Convolutions 28 Jul 2022 · 8 repositories · arXiv:2207.14284Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Sequence to sequence pretraining for a less-resourced Slovenian language 28 Jul 2022 · 1 repository · arXiv:2207.13988
-
No More Fine-Tuning? An Experimental Evaluation of Prompt Tuning in Code Intelligence 24 Jul 2022 · 1 repository · arXiv:2207.11680Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Context based lemmatizer for Polish language 23 Jul 2022 · 0 repositories · arXiv:2207.11565
-
Effectiveness of French Language Models on Abstractive Dialogue Summarization Task 17 Jul 2022 · 0 repositories · arXiv:2207.08305
-
DocPrompting: Generating Code by Retrieving the Docs 13 Jul 2022 · 2 repositories · arXiv:2207.05987Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples)
-
Re2G: Retrieve, Rerank, Generate 13 Jul 2022 · 1 repository · arXiv:2207.06300Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Learning to Diversify for Product Question Generation 6 Jul 2022 · 0 repositories · arXiv:2207.02534
-
CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning 5 Jul 2022 · 2 repositories · arXiv:2207.01780Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Generating Repetitions with Appropriate Repeated Words 3 Jul 2022 · 1 repository · arXiv:2207.00929
-
Is neural language acquisition similar to natural? A chronological probing study 1 Jul 2022 · 1 repository · arXiv:2207.00560
-
Argumentative Text Generation in Economic Domain 18 Jun 2022 · 1 repository · arXiv:2206.09251
-
Automatic Summarization of Russian Texts: Comparison of Extractive and Abstractive Methods 18 Jun 2022 · 0 repositories · arXiv:2206.09253
-
Alexa Teacher Model: Pretraining and Distilling Multi-Billion-Parameter Encoders for Natural Language Understanding Systems 15 Jun 2022 · 0 repositories · arXiv:2206.07808
-
NatGen: Generative pre-training by "Naturalizing" source code 15 Jun 2022 · 2 repositories · arXiv:2206.07585Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LST: Ladder Side-Tuning for Parameter and Memory Efficient Transfer Learning 13 Jun 2022 · 2 repositories · arXiv:2206.06522Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
GASP: Gated Attention For Saliency Prediction 9 Jun 2022 · 1 repository · arXiv:2206.04590
-
CodeAttack: Code-Based Adversarial Attacks for Pre-trained Programming Language Models 31 May 2022 · 2 repositories · arXiv:2206.00052
-
E2S2: Encoding-Enhanced Sequence-to-Sequence Pretraining for Language Understanding and Generation 30 May 2022 · 1 repository · arXiv:2205.14912
-
Conditional set generation using Seq2seq models 25 May 2022 · 0 repositories · arXiv:2205.12485
-
RobustLR: Evaluating Robustness to Logical Perturbation in Deductive Reasoning 25 May 2022 · 1 repository · arXiv:2205.12598Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
TaCube: Pre-computing Data Cubes for Answering Numerical-Reasoning Questions over Tabular Data 25 May 2022 · 1 repository · arXiv:2205.12682
-
EdiT5: Semi-Autoregressive Text-Editing with T5 Warm-Start 24 May 2022 · 0 repositories · arXiv:2205.12209
-
FLUTE: Figurative Language Understanding through Textual Explanations 24 May 2022 · 1 repository · arXiv:2205.12404
-
GeoMLAMA: Geo-Diverse Commonsense Probing on Multilingual Pre-Trained Language Models 24 May 2022 · 1 repository · arXiv:2205.12247
-
Medical Scientific Table-to-Text Generation with Human-in-the-Loop under the Data Sparsity Constraint 24 May 2022 · 0 repositories · arXiv:2205.12368
-
Workflow Discovery from Dialogues in the Low Data Regime 24 May 2022 · 1 repository · arXiv:2205.11690
-
A Question-Answer Driven Approach to Reveal Affirmative Interpretations from Verbal Negations 23 May 2022 · 1 repository · arXiv:2205.11467
-
BanglaNLG and BanglaT5: Benchmarks and Resources for Evaluating Low-Resource Natural Language Generation in Bangla 23 May 2022 · 2 repositories · arXiv:2205.11081
-
Life after BERT: What do Other Muppets Understand about Language? 21 May 2022 · 1 repository · arXiv:2205.10696
-
Evaluation of Transfer Learning for Polish with a Text-to-Text Model 18 May 2022 · 0 repositories · arXiv:2205.08808
-
PASH at TREC 2021 Deep Learning Track: Generative Enhanced Model for Multi-stage Ranking 18 May 2022 · 0 repositories · arXiv:2205.11245
-
M6-Rec: Generative Pretrained Language Models are Open-Ended Recommender Systems 17 May 2022 · 0 repositories · arXiv:2205.08084
-
SKILL: Structured Knowledge Infusion for Large Language Models 17 May 2022 · 0 repositories · arXiv:2205.08184
-
Downstream Transformer Generation of Question-Answer Pairs with Preprocessing and Postprocessing Pipelines 15 May 2022 · 1 repository · arXiv:2205.07387
-
RASAT: Integrating Relational Structures into Pretrained Seq2Seq Model for Text-to-SQL 14 May 2022 · 1 repository · arXiv:2205.06983Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
UL2: Unifying Language Learning Paradigms 10 May 2022 · 2 repositories · arXiv:2205.05131Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 16 harvested samples)
-
Vector Representations of Idioms in Conversational Systems 7 May 2022 · 0 repositories · arXiv:2205.03666
-
Two New Datasets for Italian-Language Abstractive Text Summarization 29 Apr 2022 · 1 repository
-
Modern Baselines for SPARQL Semantic Parsing 27 Apr 2022 · 1 repository · arXiv:2204.12793
-
Towards Arabic Sentence Simplification via Classification and Generative Approaches 20 Apr 2022 · 0 repositories · arXiv:2204.09292
-
MASSIVE: A 1M-Example Multilingual Natural Language Understanding Dataset with 51 Typologically-Diverse Languages 18 Apr 2022 · 6 repositories · arXiv:2204.08582Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
WordAlchemy: A transformer-based Reverse Dictionary 16 Apr 2022 · 0 repositories · arXiv:2204.10181
-
ML_LTU at SemEval-2022 Task 4: T5 Towards Identifying Patronizing and Condescending Language 15 Apr 2022 · 0 repositories · arXiv:2204.07432
-
Building Markovian Generative Architectures over Pretrained LM Backbones for Efficient Task-Oriented Dialog Systems 13 Apr 2022 · 2 repositories · arXiv:2204.06452
-
Enhance Incomplete Utterance Restoration by Joint Learning Token Extraction and Text Generation 8 Apr 2022 · 1 repository · arXiv:2204.03958
-
MMTAfrica: Multilingual Machine Translation for African Languages 8 Apr 2022 · 1 repository · arXiv:2204.04306
-
ByT5 model for massively multilingual grapheme-to-phoneme conversion 6 Apr 2022 · 1 repository · arXiv:2204.03067
-
Contextual Attention Mechanism, SRGAN Based Inpainting System for Eliminating Interruptions from Images 6 Apr 2022 · 0 repositories · arXiv:2204.02591
-
Example-based Hypernetworks for Out-of-Distribution Generalization 27 Mar 2022 · 1 repository · arXiv:2203.14276
-
Ensembling and Knowledge Distilling of Large Sequence Taggers for Grammatical Error Correction 24 Mar 2022 · 1 repository · arXiv:2203.13064
-
Recommendation as Language Processing (RLP): A Unified Pretrain, Personalized Prompt & Predict Paradigm (P5) 24 Mar 2022 · 2 repositories · arXiv:2203.13366
-
AraBART: a Pretrained Arabic Sequence-to-Sequence Model for Abstractive Summarization 21 Mar 2022 · 0 repositories · arXiv:2203.10945
-
DQ-BART: Efficient Sequence-to-Sequence Model via Joint Distillation and Quantization 21 Mar 2022 · 2 repositories · arXiv:2203.11239Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Coloring the Blank Slate: Pre-training Imparts a Hierarchical Inductive Bias to Sequence-to-sequence Models 17 Mar 2022 · 1 repository · arXiv:2203.09397Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
IT5: Text-to-text Pretraining for Italian Language Understanding and Generation 7 Mar 2022 · 3 repositories · arXiv:2203.03759
-
Parameter-Efficient Mixture-of-Experts Architecture for Pre-trained Language Models 2 Mar 2022 · 2 repositories · arXiv:2203.01104Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
E-LANG: Energy-Based Joint Inferencing of Super and Swift Language Models 1 Mar 2022 · 0 repositories · arXiv:2203.00748
-
HyperPrompt: Prompt-based Task-Conditioning of Transformers 1 Mar 2022 · 0 repositories · arXiv:2203.00759
-
Harmonic gated compensation network plus for ICASSP 2022 DNS CHALLENGE 25 Feb 2022 · 0 repositories · arXiv:2202.12643
-
Mixture-of-Experts with Expert Choice Routing 18 Feb 2022 · 0 repositories · arXiv:2202.09368