Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 45
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 45 of 139: papers 4,401 to 4,500 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Mumpy: Multilateral Temporal-view Pyramid Transformer for Video Inpainting Detection 17 Apr 2024 · 0 repositories · arXiv:2404.11054
-
Octopus v3: Technical Report for On-device Sub-billion Multimodal AI Agent 17 Apr 2024 · 0 repositories · arXiv:2404.11459
-
Pretraining Billion-scale Geospatial Foundational Models on Frontier 17 Apr 2024 · 0 repositories · arXiv:2404.11706
-
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding 17 Apr 2024 · 0 repositories · arXiv:2404.11589
-
Towards Data-Centric Automatic R&D 17 Apr 2024 · 1 repository · arXiv:2404.11276
-
Revisiting Noise Resilience Strategies in Gesture Recognition: Short-Term Enhancement in Surface Electromyographic Signal Analysis 17 Apr 2024 · 0 repositories · arXiv:2404.11213
-
Self-adaptive PSRO: Towards an Automatic Population-based Game Solver 17 Apr 2024 · 0 repositories · arXiv:2404.11144
-
Supervised Contrastive Vision Transformer for Breast Histopathological Image Classification 17 Apr 2024 · 0 repositories · arXiv:2404.11052
-
Towards Coarse-to-Fine Evaluation of Inference Efficiency for Large Language Models 17 Apr 2024 · 1 repository · arXiv:2404.11502
-
Training Transformer Models by Wavelet Losses Improves Quantitative and Visual Performance in Single Image Super-Resolution 17 Apr 2024 · 1 repository · arXiv:2404.11273Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 7 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 7 pointer-only (licence)
-
AGHINT: Attribute-Guided Representation Learning on Heterogeneous Information Networks with Transformer 16 Apr 2024 · 0 repositories · arXiv:2404.10443
-
Anomaly Correction of Business Processes Using Transformer Autoencoder 16 Apr 2024 · 0 repositories · arXiv:2404.10211
-
Can Language Models Solve Olympiad Programming? 16 Apr 2024 · 1 repository · arXiv:2404.10952Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
CoTAR: Chain-of-Thought Attribution Reasoning with Multi-level Granularity 16 Apr 2024 · 0 repositories · arXiv:2404.10513
-
Deep Learning and LLM-based Methods Applied to Stellar Lightcurve Classification 16 Apr 2024 · 1 repository · arXiv:2404.10757
-
Gasformer: A Transformer-based Architecture for Segmenting Methane Emissions from Livestock in Optical Gas Imaging 16 Apr 2024 · 1 repository · arXiv:2404.10841
-
ClashEval: Quantifying the tug-of-war between an LLM's internal prior and external evidence 16 Apr 2024 · 1 repository · arXiv:2404.10198
-
Incubating Text Classifiers Following User Instruction with Nothing but LLM 16 Apr 2024 · 1 repository · arXiv:2404.10877
-
MathWriting: A Dataset For Handwritten Mathematical Expression Recognition 16 Apr 2024 · 0 repositories · arXiv:2404.10690
-
MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents 16 Apr 2024 · 2 repositories · arXiv:2404.10774Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 8 harvested samples)
-
Neuromorphic Vision-based Motion Segmentation with Graph Transformer Neural Network 16 Apr 2024 · 0 repositories · arXiv:2404.10940
-
Grounded Language Agent for Product Search via Intelligent Web Interactions 16 Apr 2024 · 1 repository · arXiv:2404.10887Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Self-Supervised Visual Preference Alignment 16 Apr 2024 · 1 repository · arXiv:2404.10501Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback 16 Apr 2024 · 0 repositories · arXiv:2404.10271
-
TC-OCR: TableCraft OCR for Efficient Detection & Recognition of Table Structure & Content 16 Apr 2024 · 0 repositories · arXiv:2404.10305
-
Threat Behavior Textual Search by Attention Graph Isomorphism 16 Apr 2024 · 1 repository · arXiv:2404.10944
-
AIGeN: An Adversarial Approach for Instruction Generation in VLN 15 Apr 2024 · 0 repositories · arXiv:2404.10054
-
EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided Reinforcement Learning 15 Apr 2024 · 0 repositories · arXiv:2404.10163
-
Learn Your Reference Model for Real Good Alignment 15 Apr 2024 · 0 repositories · arXiv:2404.09656
-
LLM Evaluators Recognize and Favor Their Own Generations 15 Apr 2024 · 0 repositories · arXiv:2404.13076
-
LoRAP: Transformer Sub-Layers Deserve Differentiated Structured Compression for Large Language Models 15 Apr 2024 · 0 repositories · arXiv:2404.09695
-
Are Medium-Sized Transformers Models still Relevant for Medical Records Processing? 15 Apr 2024 · 0 repositories · arXiv:2404.10171
-
ODFormer: Semantic Fundus Image Segmentation Using Transformer for Optic Nerve Head Detection 15 Apr 2024 · 0 repositories · arXiv:2405.09552
-
SegFormer3D: an Efficient Transformer for 3D Medical Image Segmentation 15 Apr 2024 · 2 repositories · arXiv:2404.10156Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
State Space Model for New-Generation Network Alternative to Transformers: A Survey 15 Apr 2024 · 1 repository · arXiv:2404.09516
-
Unveiling Imitation Learning: Exploring the Impact of Data Falsity to Large Language Model 15 Apr 2024 · 0 repositories · arXiv:2404.09717
-
WiTUnet: A U-Shaped Architecture Integrating CNN and Transformer for Improved Feature Alignment and Local Information Fusion 15 Apr 2024 · 1 repository · arXiv:2404.09533
-
Zero-shot Building Age Classification from Facade Image Using GPT-4 15 Apr 2024 · 1 repository · arXiv:2404.09921
-
Arena: A Patch-of-Interest ViT Inference Acceleration System for Edge-Assisted Video Analytics 14 Apr 2024 · 0 repositories · arXiv:2404.09245
-
MAX-AST: COMBINING CONVOLUTION, LOCAL AND GLOBAL SELF-ATTENTIONS FOR AUDIO EVENT CLASSIFICATION 14 Apr 2024 · 1 repository
-
RF-Diffusion: Radio Signal Generation via Time-Frequency Diffusion 14 Apr 2024 · 1 repository · arXiv:2404.09140Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TransformerFAM: Feedback attention is working memory 14 Apr 2024 · 0 repositories · arXiv:2404.09173
-
Rethinking Low-Rank Adaptation in Vision: Exploring Head-Level Responsiveness across Diverse Tasks 13 Apr 2024 · 0 repositories · arXiv:2404.08894
-
NeurIT: Pushing the Limit of Neural Inertial Tracking for Indoor Robotic IoT 13 Apr 2024 · 1 repository · arXiv:2404.08939
-
OOVs in the Spotlight: How to Inflect them? 13 Apr 2024 · 1 repository · arXiv:2404.08974
-
A Novel Vision Transformer based Load Profile Analysis using Load Images as Inputs 12 Apr 2024 · 0 repositories · arXiv:2404.08175
-
Calibration & Reconstruction: Deep Integrated Language for Referring Image Segmentation 12 Apr 2024 · 0 repositories · arXiv:2404.08281
-
Constrained C-Test Generation via Mixed-Integer Programming 12 Apr 2024 · 1 repository · arXiv:2404.08821
-
Dataset Reset Policy Optimization for RLHF 12 Apr 2024 · 1 repository · arXiv:2404.08495Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
"Don't forget to put the milk back!" Dataset for Enabling Embodied Agents to Detect Anomalous Situations 12 Apr 2024 · 0 repositories · arXiv:2404.08827
-
IFViT: Interpretable Fixed-Length Representation for Fingerprint Matching via Vision Transformer 12 Apr 2024 · 0 repositories · arXiv:2404.08237
-
Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length 12 Apr 2024 · 1 repository · arXiv:2404.08801Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
MSSTNet: A Multi-Scale Spatio-Temporal CNN-Transformer Network for Dynamic Facial Expression Recognition 12 Apr 2024 · 0 repositories · arXiv:2404.08433
-
Scalability in Building Component Data Annotation: Enhancing Facade Material Classification with Synthetic Data 12 Apr 2024 · 0 repositories · arXiv:2404.08557
-
Single-image driven 3d viewpoint training data augmentation for effective wine label recognition 12 Apr 2024 · 0 repositories · arXiv:2404.08820
-
Small Models Are (Still) Effective Cross-Domain Argument Extractors 12 Apr 2024 · 1 repository · arXiv:2404.08579
-
Automatic Generation and Evaluation of Reading Comprehension Test Items with Large Language Models 11 Apr 2024 · 2 repositories · arXiv:2404.07720
-
Comments as Natural Logic Pivots: Improve Code Generation via Comment Perspective 11 Apr 2024 · 1 repository · arXiv:2404.07549
-
DesignQA: A Multimodal Benchmark for Evaluating Large Language Models' Understanding of Engineering Documentation 11 Apr 2024 · 1 repository · arXiv:2404.07917Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Event-Enhanced Snapshot Compressive Videography at 10K FPS 11 Apr 2024 · 0 repositories · arXiv:2404.07551
-
From Words to Numbers: Your Large Language Model Is Secretly A Capable Regressor When Given In-Context Examples 11 Apr 2024 · 1 repository · arXiv:2404.07544Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Graph Integrated Language Transformers for Next Action Prediction in Complex Phone Calls 11 Apr 2024 · 0 repositories · arXiv:2404.08155
-
HGRN2: Gated Linear RNNs with State Expansion 11 Apr 2024 · 4 repositories · arXiv:2404.07904Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Human Latency Conversational Turns for Spoken Avatar Systems 11 Apr 2024 · 0 repositories · arXiv:2404.16053
-
LATTE: Low-Precision Approximate Attention with Head-wise Trainable Threshold for Efficient Transformer 11 Apr 2024 · 0 repositories · arXiv:2404.07519
-
LLM Agents can Autonomously Exploit One-day Vulnerabilities 11 Apr 2024 · 0 repositories · arXiv:2404.08144
-
LUCF-Net: Lightweight U-shaped Cascade Fusion Network for Medical Image Segmentation 11 Apr 2024 · 1 repository · arXiv:2404.07473
-
MM-PhyQA: Multimodal Physics Question-Answering With Multi-Image CoT Prompting 11 Apr 2024 · 0 repositories · arXiv:2404.08704
-
Post-hurricane building damage assessment using street-view imagery and structured data: A multi-modal deep learning approach 11 Apr 2024 · 0 repositories · arXiv:2404.07399
-
Remembering Transformer for Continual Learning 11 Apr 2024 · 0 repositories · arXiv:2404.07518
-
Structure-aware Fine-tuning for Code Pre-trained Models 11 Apr 2024 · 0 repositories · arXiv:2404.07471
-
Token Space: A Category Theory Framework for AI Computations 11 Apr 2024 · 0 repositories · arXiv:2404.11624
-
ViM-UNet: Vision Mamba for Biomedical Segmentation 11 Apr 2024 · 1 repository · arXiv:2404.07705
-
Control-DAG: Constrained Decoding for Non-Autoregressive Directed Acyclic T5 using Weighted Finite State Automata 10 Apr 2024 · 1 repository · arXiv:2404.06854
-
Dynamic Generation of Personalities with Large Language Models 10 Apr 2024 · 1 repository · arXiv:2404.07084
-
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention 10 Apr 2024 · 5 repositories · arXiv:2404.07143Syntology 15 ran (of which 4 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 2 violated, 5 with no contract checked; 8 where Syntology's instrument failed) · 1 unverified (of 16 harvested samples) · 6 pointer-only (licence)
-
NFARec: A Negative Feedback-Aware Recommender Model 10 Apr 2024 · 1 repository · arXiv:2404.06900
-
Characterizing Multimodal Long-form Summarization: A Case Study on Financial Reports 9 Apr 2024 · 0 repositories · arXiv:2404.06162
-
Comparing Two Model Designs for Clinical Note Generation; Is an LLM a Useful Evaluator of Consistency? 9 Apr 2024 · 0 repositories · arXiv:2404.06503
-
Generative Pre-Trained Transformer for Symbolic Regression Base In-Context Reinforcement Learning 9 Apr 2024 · 0 repositories · arXiv:2404.06330
-
InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD 9 Apr 2024 · 2 repositories · arXiv:2404.06512
-
LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders 9 Apr 2024 · 1 repository · arXiv:2404.05961
-
LLMs' Reading Comprehension Is Affected by Parametric Knowledge and Struggles with Hypothetical Statements 9 Apr 2024 · 0 repositories · arXiv:2404.06283
-
Efficient Concertormer for Image Deblurring and Beyond 9 Apr 2024 · 0 repositories · arXiv:2404.06135
-
PGTNet: A Process Graph Transformer Network for Remaining Time Prediction of Business Process Instances 9 Apr 2024 · 1 repository · arXiv:2404.06267
-
Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs 9 Apr 2024 · 0 repositories · arXiv:2404.07242
-
scRDiT: Generating single-cell RNA-seq data by diffusion transformers and accelerating sampling 9 Apr 2024 · 1 repository · arXiv:2404.06153
-
WebCode2M: A Real-World Dataset for Code Generation from Webpage Designs 9 Apr 2024 · 0 repositories · arXiv:2404.06369
-
Decision Transformers for Wireless Communications: A New Paradigm of Resource Management 8 Apr 2024 · 0 repositories · arXiv:2404.05199
-
Deep Optics for Video Snapshot Compressive Imaging 8 Apr 2024 · 1 repository · arXiv:2404.05274Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Enhancing Lip Reading with Multi-Scale Video and Multi-Encoder 8 Apr 2024 · 0 repositories · arXiv:2404.05466
-
Evaluation of an LLM in Identifying Logical Fallacies: A Call for Rigor When Adopting LLMs in HCI Research 8 Apr 2024 · 0 repositories · arXiv:2404.05213
-
Fighting crime with Transformers: Empirical analysis of address parsing methods in payment data 8 Apr 2024 · 1 repository · arXiv:2404.05632
-
HSViT: Horizontally Scalable Vision Transformer 8 Apr 2024 · 1 repository · arXiv:2404.05196
-
LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models 8 Apr 2024 · 1 repository · arXiv:2404.05221Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MLP Can Be A Good Transformer Learner 8 Apr 2024 · 1 repository · arXiv:2404.05657Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-head Attention-based Deep Multiple Instance Learning 8 Apr 2024 · 1 repository · arXiv:2404.05362
-
Relation Extraction Using Large Language Models: A Case Study on Acupuncture Point Locations 8 Apr 2024 · 0 repositories · arXiv:2404.05415
-
Use of a Structured Knowledge Base Enhances Metadata Curation by Large Language Models 8 Apr 2024 · 1 repository · arXiv:2404.05893
-
Xiwu: A Basis Flexible and Learnable LLM for High Energy Physics 8 Apr 2024 · 1 repository · arXiv:2404.08001