Methods › General › Activation Functions › Gated Linear Unit › Papers, page 4
Gated Linear Unit
Papers archive 2025-07-28
archive papers tagged: 798 · with a code link: 400 · where Syntology ran a sample: 115 (99 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (115 of 798 tagged: 99 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 4 of 8: papers 301 to 400 of 798, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Benchmarking Procedural Language Understanding for Low-Resource Languages: A Case Study on Turkish 13 Sep 2023 · 1 repository · arXiv:2309.06698
-
RAP-Gen: Retrieval-Augmented Patch Generation with CodeT5 for Automatic Program Repair 12 Sep 2023 · 0 repositories · arXiv:2309.06057
-
Detecting Natural Language Biases with Prompt-based Learning 11 Sep 2023 · 0 repositories · arXiv:2309.05227
-
Can NLP Models 'Identify', 'Distinguish', and 'Justify' Questions that Don't have a Definitive Answer? 8 Sep 2023 · 0 repositories · arXiv:2309.04635
-
nanoT5: A PyTorch Framework for Pre-training and Fine-tuning T5-style Models with Limited Resources 5 Sep 2023 · 1 repository · arXiv:2309.02373Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
SP³: Enhancing Structured Pruning via PCA Projection 31 Aug 2023 · 1 repository · arXiv:2308.16475Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Multi-party Goal Tracking with LLMs: Comparing Pre-training, Fine-tuning, and Prompt Engineering 29 Aug 2023 · 1 repository · arXiv:2308.15231
-
FIRE: Food Image to REcipe generation 28 Aug 2023 · 1 repository · arXiv:2308.14391Syntology official (archive's flag): 12 ran · 12 ran (of which 2 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
UniPT: Universal Parallel Tuning for Transfer Learning with Efficient Parameter and Memory 28 Aug 2023 · 1 repository · arXiv:2308.14316
-
Improving Knowledge Distillation for BERT Models: Loss Functions, Mapping Methods, and Weight Tuning 26 Aug 2023 · 0 repositories · arXiv:2308.13958
-
Interpretable Graph Neural Networks for Tabular Data 17 Aug 2023 · 2 repositories · arXiv:2308.08945
-
Mitigating the Exposure Bias in Sentence-Level Grapheme-to-Phoneme (G2P) Transduction 16 Aug 2023 · 0 repositories · arXiv:2308.08442
-
Domain Adaptation for Code Model-based Unit Test Case Generation 15 Aug 2023 · 0 repositories · arXiv:2308.08033
-
"Beware of deception": Detecting Half-Truth and Debunking it through Controlled Claim Editing 15 Aug 2023 · 0 repositories · arXiv:2308.07973
-
EasyEdit: An Easy-to-use Knowledge Editing Framework for Large Language Models 14 Aug 2023 · 2 repositories · arXiv:2308.07269Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Thinking Like an Expert:Multimodal Hypergraph-of-Thought (HoT) Reasoning to boost Foundation Modals 11 Aug 2023 · 0 repositories · arXiv:2308.06207
-
You Only Prompt Once: On the Capabilities of Prompt Learning on Large Language Models to Tackle Toxic Content 10 Aug 2023 · 1 repository · arXiv:2308.05596Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning 7 Aug 2023 · 1 repository · arXiv:2308.03526Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
LoRA-FA: Memory-efficient Low-rank Adaptation for Large Language Models Fine-tuning 7 Aug 2023 · 0 repositories · arXiv:2308.03303
-
Towards Controllable Natural Language Inference through Lexical Inference Types 7 Aug 2023 · 0 repositories · arXiv:2308.03581
-
Machine learning methods for the search for L&T brown dwarfs in the data of modern sky surveys 6 Aug 2023 · 1 repository · arXiv:2308.03045
-
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias 1 Aug 2023 · 1 repository · arXiv:2308.00225Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
No that's not what I meant: Handling Third Position Repair in Conversational Question Answering 31 Jul 2023 · 1 repository · arXiv:2307.16689
-
Evaluating Generative Models for Graph-to-Text Generation 27 Jul 2023 · 1 repository · arXiv:2307.14712
-
Speed Reading Tool Powered by Artificial Intelligence for Students with ADHD, Dyslexia, or Short Attention Span 26 Jul 2023 · 0 repositories · arXiv:2307.14544
-
An End-to-End Workflow using Topic Segmentation and Text Summarisation Methods for Improved Podcast Comprehension 25 Jul 2023 · 0 repositories · arXiv:2307.13394
-
Learning minimal representations of stochastic processes with variational autoencoders 21 Jul 2023 · 1 repository · arXiv:2307.11608
-
Controlling Equational Reasoning in Large Language Models with Prompt Interventions 19 Jul 2023 · 0 repositories · arXiv:2307.09998
-
Large Language Models Understand and Can be Enhanced by Emotional Stimuli 14 Jul 2023 · 0 repositories · arXiv:2307.11760
-
Agreement Tracking for Multi-Issue Negotiation Dialogues 13 Jul 2023 · 0 repositories · arXiv:2307.06524
-
No Train No Gain: Revisiting Efficient Training Algorithms For Transformer-based Language Models 12 Jul 2023 · 1 repository · arXiv:2307.06440Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 3 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Edge-Aware Mirror Network for Camouflaged Object Detection 8 Jul 2023 · 1 repository · arXiv:2307.03932
-
Text Simplification of Scientific Texts for Non-Expert Readers 7 Jul 2023 · 0 repositories · arXiv:2307.03569
-
Exploring Continual Learning for Code Generation Models 5 Jul 2023 · 0 repositories · arXiv:2307.02435
-
Leveraging Denoised Abstract Meaning Representation for Grammatical Error Correction 5 Jul 2023 · 0 repositories · arXiv:2307.02127
-
Multilingual Controllable Transformer-Based Lexical Simplification 5 Jul 2023 · 1 repository · arXiv:2307.02120
-
Interpretability and Transparency-Driven Detection and Transformation of Textual Adversarial Examples (IT-DT) 3 Jul 2023 · 0 repositories · arXiv:2307.01225
-
How far is Language Model from 100% Few-shot Named Entity Recognition in Medical Domain 1 Jul 2023 · 1 repository · arXiv:2307.00186
-
Abstractive Text Summarization for Resumes With Cutting Edge NLP Transformers and LSTM 23 Jun 2023 · 0 repositories · arXiv:2306.13315
-
On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes 23 Jun 2023 · 0 repositories · arXiv:2306.13649
-
Incorporating Graph Information in Transformer-based AMR Parsing 23 Jun 2023 · 1 repository · arXiv:2306.13467
-
Investigating Pre-trained Language Models on Cross-Domain Datasets, a Step Closer to General AI 21 Jun 2023 · 0 repositories · arXiv:2306.12205
-
A Novel Counterfactual Data Augmentation Method for Aspect-Based Sentiment Analysis 20 Jun 2023 · 0 repositories · arXiv:2306.11260
-
Deep Fusion: Efficient Network Training via Pre-trained Initializations 20 Jun 2023 · 0 repositories · arXiv:2306.11903
-
Fine-Tuning Language Models for Scientific Writing Support 19 Jun 2023 · 1 repository · arXiv:2306.10974
-
Summarization from Leaderboards to Practice: Choosing A Representation Backbone and Ensuring Robustness 18 Jun 2023 · 0 repositories · arXiv:2306.10555
-
UniPoll: A Unified Social Media Poll Generation Framework via Multi-Objective Optimization 12 Jun 2023 · 1 repository · arXiv:2306.06851
-
CoTran: An LLM-based Code Translator using Reinforcement Learning with Feedback from Compiler and Symbolic Execution 11 Jun 2023 · 1 repository · arXiv:2306.06755
-
A Unified Generative Approach to Product Attribute-Value Identification 9 Jun 2023 · 0 repositories · arXiv:2306.05605
-
Triggering Multi-Hop Reasoning for Question Answering in Language Models using Soft Prompts and Random Walks 6 Jun 2023 · 0 repositories · arXiv:2306.04009
-
Detector Guidance for Multi-Object Text-to-Image Generation 4 Jun 2023 · 1 repository · arXiv:2306.02236
-
Modular Transformers: Compressing Transformers into Modularized Layers for Flexible Efficient Inference 4 Jun 2023 · 0 repositories · arXiv:2306.02379
-
5IDER: Unified Query Rewriting for Steering, Intent Carryover, Disfluencies, Entity Carryover and Repair 2 Jun 2023 · 0 repositories · arXiv:2306.01855
-
Adapting Pre-trained Language Models to Vision-Language Tasks via Dynamic Visual Prompting 1 Jun 2023 · 1 repository · arXiv:2306.00409
-
PreQuant: A Task-agnostic Quantization Approach for Pre-trained Language Models 30 May 2023 · 0 repositories · arXiv:2306.00014
-
Coeditor: Leveraging Contextual Changes for Multi-round Code Auto-editing 29 May 2023 · 0 repositories · arXiv:2305.18584
-
How Effective Are Neural Networks for Fixing Security Vulnerabilities 29 May 2023 · 1 repository · arXiv:2305.18607Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Knowledge-Augmented Reasoning Distillation for Small Language Models in Knowledge-Intensive Tasks 28 May 2023 · 1 repository · arXiv:2305.18395Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Towards Explainable Conversational Recommender Systems 27 May 2023 · 1 repository · arXiv:2305.18363
-
Beyond Chain-of-Thought, Effective Graph-of-Thought Reasoning in Language Models 26 May 2023 · 2 repositories · arXiv:2305.16582Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Learning to Imagine: Visually-Augmented Natural Language Generation 26 May 2023 · 1 repository · arXiv:2305.16944Syntology official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Pre-training Meets Clustering: A Hybrid Extractive Multi-document Summarization Model 25 May 2023 · 1 repository
-
Neural Summarization of Electronic Health Records 24 May 2023 · 0 repositories · arXiv:2305.15222
-
Segmented Recurrent Transformer: An Efficient Sequence-to-Sequence Model 24 May 2023 · 1 repository · arXiv:2305.16340
-
Exploring Large Language Models for Classical Philology 23 May 2023 · 1 repository · arXiv:2305.13698
-
GrACE: Generation using Associated Code Edits 23 May 2023 · 0 repositories · arXiv:2305.14129
-
mmT5: Modular Multilingual Pre-Training Solves Source Language Hallucinations 23 May 2023 · 0 repositories · arXiv:2305.14224
-
NAIL: Lexical Retrieval Indices with Efficient Non-Autoregressive Decoders 23 May 2023 · 0 repositories · arXiv:2305.14499
-
On Robustness of Finetuned Transformer-based NLP Models 23 May 2023 · 1 repository · arXiv:2305.14453
-
ReadMe++: Benchmarking Multilingual Language Models for Multi-Domain Readability Assessment 23 May 2023 · 1 repository · arXiv:2305.14463
-
GNCformer Enhanced Self-attention for Automatic Speech Recognition 22 May 2023 · 0 repositories · arXiv:2305.12755
-
SPARSEFIT: Few-shot Prompting with Sparse Fine-tuning for Jointly Generating Predictions and Natural Language Explanations 22 May 2023 · 1 repository · arXiv:2305.13235
-
Model-Generated Pretraining Signals Improves Zero-Shot Generalization of Text-to-Text Transformers 21 May 2023 · 1 repository · arXiv:2305.12567
-
Can NLP Models Correctly Reason Over Contexts that Break the Common Assumptions? 20 May 2023 · 0 repositories · arXiv:2305.12096
-
Do Models Really Learn to Follow Instructions? An Empirical Study of Instruction Tuning 19 May 2023 · 0 repositories · arXiv:2305.11383
-
Generalized Multiple Intent Conditioned Slot Filling 18 May 2023 · 0 repositories · arXiv:2305.11023
-
mLongT5: A Multilingual and Efficient Text-To-Text Transformer for Longer Sequences 18 May 2023 · 1 repository · arXiv:2305.11129
-
Instruction Tuned Models are Quick Learners 17 May 2023 · 1 repository · arXiv:2306.05539Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Document Understanding Dataset and Evaluation (DUDE) 15 May 2023 · 1 repository · arXiv:2305.08455Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Sensitivity and Robustness of Large Language Models to Prompt Template in Japanese Text Classification Tasks 15 May 2023 · 0 repositories · arXiv:2305.08714
-
NL2TL: Transforming Natural Languages to Temporal Logics using Large Language Models 12 May 2023 · 3 repositories · arXiv:2305.07766Syntology official (archive's flag): 3 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 8 pointer-only (licence)
-
A Black-Box Attack on Code Models via Representation Nearest Neighbor Search 10 May 2023 · 0 repositories · arXiv:2305.05896
-
An Exploration of Encoder-Decoder Approaches to Multi-Label Classification for Legal and Biomedical Text 9 May 2023 · 1 repository · arXiv:2305.05627
-
Multi-Task End-to-End Training Improves Conversational Recommendation 8 May 2023 · 0 repositories · arXiv:2305.06218
-
Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model Sizes 3 May 2023 · 1 repository · arXiv:2305.02301Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Entity Tracking in Language Models 3 May 2023 · 1 repository · arXiv:2305.02363
-
Neural Keyphrase Generation: Analysis and Evaluation 27 Apr 2023 · 0 repositories · arXiv:2304.13883
-
Introducing MBIB -- the first Media Bias Identification Benchmark Task and Dataset Collection 25 Apr 2023 · 1 repository · arXiv:2304.13148
-
Text-to-Audio Generation using Instruction-Tuned LLM and Latent Diffusion Model 24 Apr 2023 · 1 repository · arXiv:2304.13731
-
LongForm: Effective Instruction Tuning with Reverse Instructions 17 Apr 2023 · 2 repositories · arXiv:2304.08460
-
The MiniPile Challenge for Data-Efficient Language Models 17 Apr 2023 · 1 repository · arXiv:2304.08442
-
Enhancing Automated Program Repair through Fine-tuning and Prompt Engineering 16 Apr 2023 · 0 repositories · arXiv:2304.07840
-
Exploring the Use of Foundation Models for Named Entity Recognition and Lemmatization Tasks in Slavic Languages 11 Apr 2023 · 0 repositories · arXiv:2304.05336
-
Pump It Up: Predict Water Pump Status using Attentive Tabular Learning 8 Apr 2023 · 0 repositories · arXiv:2304.03969
-
ChartReader: A Unified Framework for Chart Derendering and Comprehension without Heuristic Rules 5 Apr 2023 · 1 repository · arXiv:2304.02173Syntology official (archive's flag): 16 ran · 16 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 10 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Generating Natural Language from Logic Expressions with Structural Representation 4 Apr 2023 · 1 repository
-
PEACH: Pre-Training Sequence-to-Sequence Multilingual Models for Translation with Semi-Supervised Pseudo-Parallel Document Generation 3 Apr 2023 · 1 repository · arXiv:2304.01282
-
The StatCan Dialogue Dataset: Retrieving Data Tables through Conversations with Genuine Intents 3 Apr 2023 · 1 repository · arXiv:2304.01412
-
Does Human Collaboration Enhance the Accuracy of Identifying LLM-Generated Deepfake Texts? 3 Apr 2023 · 2 repositories · arXiv:2304.01002Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Better Language Models of Code through Self-Improvement 2 Apr 2023 · 1 repository · arXiv:2304.01228