Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 65
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 65 of 139: papers 6,401 to 6,500 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Building Real-World Meeting Summarization Systems using Large Language Models: A Practical Perspective 30 Oct 2023 · 0 repositories · arXiv:2310.19233
-
Constituency Parsing using LLMs 30 Oct 2023 · 0 repositories · arXiv:2310.19462
-
Dynamics of Instruction Tuning: Each Ability of Large Language Models Has Its Own Growth Pace 30 Oct 2023 · 1 repository · arXiv:2310.19651
-
AViTMP: A Tracking-Specific Transformer for Single-Branch Visual Tracking 30 Oct 2023 · 1 repository · arXiv:2310.19542
-
Fusing Temporal Graphs into Transformers for Time-Sensitive Question Answering 30 Oct 2023 · 0 repositories · arXiv:2310.19292
-
Bridging Lottery Ticket and Grokking: Understanding Grokking from Inner Structure of Networks 30 Oct 2023 · 1 repository · arXiv:2310.19470Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Interpretable-by-Design Text Understanding with Iteratively Generated Concept Bottleneck 30 Oct 2023 · 1 repository · arXiv:2310.19660
-
Large Trajectory Models are Scalable Motion Predictors and Planners 30 Oct 2023 · 1 repository · arXiv:2310.19620Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MIST: Medical Image Segmentation Transformer with Convolutional Attention Mixing (CAM) Decoder 30 Oct 2023 · 1 repository · arXiv:2310.19898
-
One-for-All: Bridge the Gap Between Heterogeneous Architectures in Knowledge Distillation 30 Oct 2023 · 1 repository · arXiv:2310.19444Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
SolarFormer: Multi-scale Transformer for Solar PV Profiling 30 Oct 2023 · 0 repositories · arXiv:2310.20057
-
Towards Few-Annotation Learning for Object Detection: Are Transformer-based Models More Efficient ? 30 Oct 2023 · 1 repository · arXiv:2310.19936
-
Multimodal ChatGPT for Medical Applications: an Experimental Study of GPT-4V 29 Oct 2023 · 1 repository · arXiv:2310.19061
-
Pushdown Layers: Encoding Recursive Structure in Transformer Language Models 29 Oct 2023 · 1 repository · arXiv:2310.19089Syntology official (archive's flag): 14 ran · 14 ran (of which 8 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 23 harvested samples) · 23 pointer-only (licence)
-
ASTormer: An AST Structure-aware Transformer Decoder for Text-to-SQL 28 Oct 2023 · 0 repositories · arXiv:2310.18662
-
Integration of persistent Laplacian and pre-trained transformer for protein solubility changes upon mutation 28 Oct 2023 · 1 repository · arXiv:2310.18760
-
Patch-Wise Self-Supervised Visual Representation Learning: A Fine-Grained Approach 28 Oct 2023 · 1 repository · arXiv:2310.18651
-
MultiScale Spectral-Spatial Convolutional Transformer for Hyperspectral Image Classification 28 Oct 2023 · 0 repositories · arXiv:2310.18550
-
Using Large Language Models to Support Thematic Analysis in Empirical Legal Studies 28 Oct 2023 · 0 repositories · arXiv:2310.18729
-
Boosting Data Analytics With Synthetic Volume Expansion 27 Oct 2023 · 1 repository · arXiv:2310.17848
-
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory 27 Oct 2023 · 1 repository · arXiv:2310.17884
-
FaultSeg Swin-UNETR: Transformer-Based Self-Supervised Pretraining Model for Fault Recognition 27 Oct 2023 · 0 repositories · arXiv:2310.17974
-
FP8-LM: Training FP8 Large Language Models 27 Oct 2023 · 1 repository · arXiv:2310.18313
-
GPT-4 Vision on Medical Image Classification -- A Case Study on COVID-19 Dataset 27 Oct 2023 · 0 repositories · arXiv:2310.18498
-
Knowing What LLMs DO NOT Know: A Simple Yet Effective Self-Detection Method 27 Oct 2023 · 0 repositories · arXiv:2310.17918
-
Large language models for aspect-based sentiment analysis 27 Oct 2023 · 1 repository · arXiv:2310.18025Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multi-label Emotion Analysis in Conversation via Multimodal Knowledge Distillation 27 Oct 2023 · 0 repositories
-
Qilin-Med-VL: Towards Chinese Large Vision-Language Model for General Healthcare 27 Oct 2023 · 1 repository · arXiv:2310.17956Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Revising with a Backward Glance: Regressions and Skips during Reading as Cognitive Signals for Revision Policies in Incremental Processing 27 Oct 2023 · 1 repository · arXiv:2310.18229
-
Siamese-DETR for Generic Multi-Object Tracking 27 Oct 2023 · 1 repository · arXiv:2310.17875
-
SOUL: Towards Sentiment and Opinion Understanding of Language 27 Oct 2023 · 1 repository · arXiv:2310.17924Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
SQLformer: Deep Auto-Regressive Query Graph Generation for Text-to-SQL Translation 27 Oct 2023 · 1 repository · arXiv:2310.18376
-
Transformers as Graph-to-Graph Models 27 Oct 2023 · 1 repository · arXiv:2310.17936Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
What You See Is What You Detect: Towards better Object Densification in 3D detection 27 Oct 2023 · 1 repository · arXiv:2310.17842
-
A Framework for Automated Measurement of Responsible AI Harms in Generative AI Applications 26 Oct 2023 · 0 repositories · arXiv:2310.17750
-
BERT-PIN: A BERT-based Framework for Recovering Missing Data Segments in Time-series Load Profiles 26 Oct 2023 · 0 repositories · arXiv:2310.17742
-
Can large language models replace humans in the systematic review process? Evaluating GPT-4's efficacy in screening and extracting data from peer-reviewed and grey literature in multiple languages 26 Oct 2023 · 0 repositories · arXiv:2310.17526
-
Can LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset 26 Oct 2023 · 0 repositories · arXiv:2310.18373
-
Codebook Features: Sparse and Discrete Interpretability for Neural Networks 26 Oct 2023 · 1 repository · arXiv:2310.17230Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
CompeteAI: Understanding the Competition Dynamics in Large Language Model-based Agents 26 Oct 2023 · 1 repository · arXiv:2310.17512Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples)
-
Cultural Adaptation of Recipes 26 Oct 2023 · 0 repositories · arXiv:2310.17353
-
Is Explanation the Cure? Misinformation Mitigation in the Short Term and Long Term 26 Oct 2023 · 0 repositories · arXiv:2310.17711
-
Learning to Abstract with Nonparametric Variational Information Bottleneck 26 Oct 2023 · 2 repositories · arXiv:2310.17284Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
LightLM: A Lightweight Deep and Narrow Language Model for Generative Recommendation 26 Oct 2023 · 1 repository · arXiv:2310.17488Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
DecoderTracker: Decoder-Only Method for Multiple-Object Tracking 26 Oct 2023 · 0 repositories · arXiv:2310.17170
-
Privacy-preserving Representation Learning for Speech Understanding 26 Oct 2023 · 0 repositories · arXiv:2310.17194
-
Skill-Mix: a Flexible and Expandable Family of Evaluations for AI models 26 Oct 2023 · 0 repositories · arXiv:2310.17567
-
Sliceformer: Make Multi-head Attention as Simple as Sorting in Discriminative Tasks 26 Oct 2023 · 1 repository · arXiv:2310.17683
-
The Expressive Power of Low-Rank Adaptation 26 Oct 2023 · 1 repository · arXiv:2310.17513Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Transformers Learn to Achieve Second-Order Convergence Rates for In-Context Linear Regression 26 Oct 2023 · 1 repository · arXiv:2310.17086Syntology official: no sample here; runs from other or unrecorded repositories · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
"You Are An Expert Linguistic Annotator": Limits of LLMs as Analyzers of Abstract Meaning Representation 26 Oct 2023 · 0 repositories · arXiv:2310.17793
-
Evaluating, Understanding, and Improving Constrained Text Generation for Large Language Models 25 Oct 2023 · 0 repositories · arXiv:2310.16343
-
A No-Reference Quality Assessment Method for Digital Human Head 25 Oct 2023 · 0 repositories · arXiv:2310.16732
-
An Early Evaluation of GPT-4V(ision) 25 Oct 2023 · 1 repository · arXiv:2310.16534
-
Back Transcription as a Method for Evaluating Robustness of Natural Language Understanding Models to Speech Recognition Errors 25 Oct 2023 · 2 repositories · arXiv:2310.16609
-
Can GPT models Follow Human Summarization Guidelines? Evaluating ChatGPT and GPT-4 for Dialogue Summarization 25 Oct 2023 · 0 repositories · arXiv:2310.16810
-
CLEX: Continuous Length Extrapolation for Large Language Models 25 Oct 2023 · 1 repository · arXiv:2310.16450Syntology official (archive's flag): 9 ran · 9 ran (of which 2 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
GraFT: Gradual Fusion Transformer for Multimodal Re-Identification 25 Oct 2023 · 0 repositories · arXiv:2310.16856
-
Is ChatGPT a Good Multi-Party Conversation Solver? 25 Oct 2023 · 1 repository · arXiv:2310.16301
-
LLM-FP4: 4-Bit Floating-Point Quantized Transformers 25 Oct 2023 · 1 repository · arXiv:2310.16836Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
LLM Performance Predictors are good initializers for Architecture Search 25 Oct 2023 · 1 repository · arXiv:2310.16712Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Modality-Agnostic Self-Supervised Learning with Meta-Learned Masked Auto-Encoder 25 Oct 2023 · 1 repository · arXiv:2310.16318Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
netFound: Foundation Model for Network Security 25 Oct 2023 · 1 repository · arXiv:2310.17025
-
OccuQuest: Mitigating Occupational Bias for Inclusive Large Language Models 25 Oct 2023 · 1 repository · arXiv:2310.16517
-
Prompt-Driven Building Footprint Extraction in Aerial Images with Offset-Building Model 25 Oct 2023 · 0 repositories · arXiv:2310.16717
-
SMURF-THP: Score Matching-based UnceRtainty quantiFication for Transformer Hawkes Process 25 Oct 2023 · 1 repository · arXiv:2310.16336
-
SuperHF: Supervised Iterative Learning from Human Feedback 25 Oct 2023 · 1 repository · arXiv:2310.16763Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer 25 Oct 2023 · 0 repositories · arXiv:2310.16279
-
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring 25 Oct 2023 · 0 repositories · arXiv:2310.18365
-
Confounder Balancing in Adversarial Domain Adaptation for Pre-Trained Large Models Fine-Tuning 24 Oct 2023 · 0 repositories · arXiv:2310.16062
-
Decoupled DETR: Spatially Disentangling Localization and Classification for Improved End-to-End Object Detection 24 Oct 2023 · 0 repositories · arXiv:2310.15955
-
Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models 24 Oct 2023 · 1 repository · arXiv:2310.15648
-
Learning-based Scheduling for Information Accuracy and Freshness in Wireless Networks 24 Oct 2023 · 0 repositories · arXiv:2310.15705
-
Mixture of Tokens: Continuous MoE through Cross-Example Aggregation 24 Oct 2023 · 1 repository · arXiv:2310.15961
-
MuSR: Testing the Limits of Chain-of-thought with Multistep Soft Reasoning 24 Oct 2023 · 3 repositories · arXiv:2310.16049Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
NoteChat: A Dataset of Synthetic Doctor-Patient Conversations Conditioned on Clinical Notes 24 Oct 2023 · 1 repository · arXiv:2310.15959
-
Octopus: A Multitask Model and Toolkit for Arabic Natural Language Generation 24 Oct 2023 · 0 repositories · arXiv:2310.16127
-
Practical Computational Power of Linear Transformers and Their Recurrent and Self-Referential Extensions 24 Oct 2023 · 1 repository · arXiv:2310.16076Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
AGaLiTe: Approximate Gated Linear Transformers for Online Reinforcement Learning 24 Oct 2023 · 2 repositories · arXiv:2310.15719Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
TRAMS: Training-free Memory Selection for Long-range Language Modeling 24 Oct 2023 · 1 repository · arXiv:2310.15494Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
UI Layout Generation with LLMs Guided by UI Grammar 24 Oct 2023 · 0 repositories · arXiv:2310.15455
-
What Algorithms can Transformers Learn? A Study in Length Generalization 24 Oct 2023 · 0 repositories · arXiv:2310.16028
-
AlpaCare:Instruction-tuned Large Language Models for Medical Application 23 Oct 2023 · 1 repository · arXiv:2310.14558Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Analyzing Multilingual Competency of LLMs in Multi-Turn Instruction Following: A Case Study of Arabic 23 Oct 2023 · 0 repositories · arXiv:2310.14819
-
Branch-Solve-Merge Improves Large Language Model Evaluation and Generation 23 Oct 2023 · 0 repositories · arXiv:2310.15123
-
Calibration of Time-Series Forecasting: Detecting and Adapting Context-Driven Distribution Shift 23 Oct 2023 · 2 repositories · arXiv:2310.14838Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Causal Inference Using LLM-Guided Discovery 23 Oct 2023 · 0 repositories · arXiv:2310.15117
-
Evaluating Spatial Understanding of Large Language Models 23 Oct 2023 · 1 repository · arXiv:2310.14540Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Evaluating the Knowledge Base Completion Potential of GPT 23 Oct 2023 · 0 repositories · arXiv:2310.14771
-
Exploring the Boundaries of GPT-4 in Radiology 23 Oct 2023 · 0 repositories · arXiv:2310.14573
-
Generative Pre-trained Transformer for Vietnamese Community-based COVID-19 Question Answering 23 Oct 2023 · 0 repositories · arXiv:2310.14602
-
GPT-4 as an Effective Zero-Shot Evaluator for Scientific Figure Captions 23 Oct 2023 · 0 repositories · arXiv:2310.15405
-
Hallucination Detection for Grounded Instruction Generation 23 Oct 2023 · 0 repositories · arXiv:2310.15319
-
InstructExcel: A Benchmark for Natural Language Instruction in Excel 23 Oct 2023 · 0 repositories · arXiv:2310.14495
-
Large Language Models can Share Images, Too! 23 Oct 2023 · 2 repositories · arXiv:2310.14804
-
LINC: A Neurosymbolic Approach for Logical Reasoning by Combining Language Models with First-Order Logic Provers 23 Oct 2023 · 1 repository · arXiv:2310.15164Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Location-Aware Visual Question Generation with Lightweight Models 23 Oct 2023 · 1 repository · arXiv:2310.15129
-
Non-autoregressive Streaming Transformer for Simultaneous Translation 23 Oct 2023 · 1 repository · arXiv:2310.14883Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
P2AT: Pyramid Pooling Axial Transformer for Real-time Semantic Segmentation 23 Oct 2023 · 1 repository · arXiv:2310.15025
-
PartialFormer: Modeling Part Instead of Whole for Machine Translation 23 Oct 2023 · 1 repository · arXiv:2310.14921