Methods › General › Activation Functions › Gated Linear Unit
Gated Linear Unit
Introduced by Yann N. Dauphin et al. in Language Modeling with Gated Convolutional Networks
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
A Gated Linear Unit, or GLU computes:
GLU(a, b) = a ⊗σ(b)
It is used in natural language processing architectures, for example the Gated CNN, because here σ(b) is the gate that control what information from a is passed up to the following layer. Intuitively, for a language modeling task, the gating mechanism allows selection of words or features that are important for predicting the next word. The GLU also has non-linear capabilities, but has a linear path for the gradient so diminishes the vanishing gradient problem.
Papers archive 2025-07-28
30 shown of 798, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
LiLM-RDB-SFC: Lightweight Language Model with Relational Database-Guided DRL for Optimized SFC Provisioning 15 Jul 2025 · 0 repositories · arXiv:2507.10903
-
Chat-Ghosting: A Comparative Study of Methods for Auto-Completion in Dialog Systems 8 Jul 2025 · 0 repositories · arXiv:2507.05940
-
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution 18 Jun 2025 · 0 repositories · arXiv:2506.17323
-
Fretting-Transformer: Encoder-Decoder Model for MIDI to Tablature Transcription 17 Jun 2025 · 0 repositories · arXiv:2506.14223
-
A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation 9 Jun 2025 · 0 repositories · arXiv:2506.08210
-
The Impact of Feature Scaling In Machine Learning: Effects on Regression and Classification Tasks 9 Jun 2025 · 0 repositories · arXiv:2506.08274
-
A Multi-Dataset Evaluation of Models for Automated Vulnerability Repair 5 Jun 2025 · 0 repositories · arXiv:2506.04987
-
DuAL-Net: A Hybrid Framework for Alzheimer's Disease Prediction from Whole-Genome Sequencing via Local SNP Windows and Global Annotations 31 May 2025 · 0 repositories · arXiv:2506.00673
-
Decom-Renorm-Merge: Model Merging on the Right Space Improves Multitasking 29 May 2025 · 0 repositories · arXiv:2505.23117
-
ShIOEnv: A CLI Behavior-Capturing Environment Enabling Grammar-Guided Command Synthesis for Dataset Curation 23 May 2025 · 1 repository · arXiv:2505.18374
-
Fusion of Foundation and Vision Transformer Model Features for Dermatoscopic Image Classification 22 May 2025 · 0 repositories · arXiv:2505.16338
-
LogiCase: Effective Test Case Generation from Logical Description in Competitive Programming 21 May 2025 · 0 repositories · arXiv:2505.15039Syntology ran 11 of 19 samples · 8 unverified · 19 pointer-only (licence)
-
EEG-to-Text Translation: A Model for Deciphering Human Brain Activity 20 May 2025 · 1 repository · arXiv:2505.13936
-
Masking in Multi-hop QA: An Analysis of How Language Models Perform with Context Permutation 16 May 2025 · 1 repository · arXiv:2505.11754
-
Multilingual Machine Translation with Quantum Encoder Decoder Attention-based Convolutional Variational Circuits 14 May 2025 · 0 repositories · arXiv:2505.09407
-
Performance Evaluation of Large Language Models in Bangla Consumer Health Query Summarization 8 May 2025 · 0 repositories · arXiv:2505.05070
-
Benchmarking Traditional Machine Learning and Deep Learning Models for Fault Detection in Power Transformers 7 May 2025 · 1 repository · arXiv:2505.06295
-
GASCADE: Grouped Summarization of Adverse Drug Event for Enhanced Cancer Pharmacovigilance 7 May 2025 · 1 repository · arXiv:2505.04284
-
A review of DNA restriction-free overlapping sequence cloning techniques for synthetic biology 6 May 2025 · 0 repositories · arXiv:2505.03681
-
JaccDiv: A Metric and Benchmark for Quantifying Diversity of Generated Marketing Text in the Music Industry 29 Apr 2025 · 0 repositories · arXiv:2504.20849
-
Large Language Models are Qualified Benchmark Builders: Rebuilding Pre-Training Datasets for Advancing Code Intelligence Tasks 28 Apr 2025 · 0 repositories · arXiv:2504.19444
-
An Efficient Aerial Image Detection with Variable Receptive Fields 21 Apr 2025 · 0 repositories · arXiv:2504.15165
-
The Geometry of Self-Verification in a Task-Specific Reasoning Model 19 Apr 2025 · 0 repositories · arXiv:2504.14379
-
Can Moran Eigenvectors Improve Machine Learning of Spatial Data? Insights from Synthetic Data Validation 16 Apr 2025 · 0 repositories · arXiv:2504.12450
-
Enhancing Metabolic Syndrome Prediction with Hybrid Data Balancing and Counterfactuals 9 Apr 2025 · 1 repository · arXiv:2504.06987
-
Sigma: A dataset for text-to-code semantic parsing with statistical analysis 5 Apr 2025 · 1 repository · arXiv:2504.04301
-
Advancing Sentiment Analysis in Tamil-English Code-Mixed Texts: Challenges and Transformer-Based Solutions 30 Mar 2025 · 0 repositories · arXiv:2503.23295
-
Enhancing Knowledge Graph Completion with Entity Neighborhood and Relation Context 29 Mar 2025 · 0 repositories · arXiv:2503.23205
-
Scaling Down Text Encoders of Text-to-Image Diffusion Models 25 Mar 2025 · 1 repository · arXiv:2503.19897
-
Exploring Training and Inference Scaling Laws in Generative Retrieval 24 Mar 2025 · 1 repository · arXiv:2503.18941
Tasks archive 2025-07-28
20 shown of 540 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections