Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 8
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 8 of 139: papers 701 to 800 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TransiT: Transient Transformer for Non-line-of-sight Videography 14 Mar 2025 · 0 repositories · arXiv:2503.11328
-
TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing 14 Mar 2025 · 1 repository · arXiv:2503.11629Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 6 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
When Do Transformers Outperform Feedforward and Recurrent Networks? A Statistical Perspective 14 Mar 2025 · 1 repository · arXiv:2503.11272
-
A Frustratingly Simple Yet Highly Effective Attack Baseline: Over 90% Success Rate Against the Strong Black-box Models of GPT-4.5/4o/o1 13 Mar 2025 · 1 repository · arXiv:2503.10635Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Hybrid Architecture with Efficient Fine Tuning for Abstractive Patent Document Summarization 13 Mar 2025 · 0 repositories · arXiv:2503.10354
-
Advanced Tool Learning and Selection System (ATLASS): A Closed-Loop Framework Using LLM 13 Mar 2025 · 0 repositories · arXiv:2503.10071
-
AudioX: Diffusion Transformer for Anything-to-Audio Generation 13 Mar 2025 · 0 repositories · arXiv:2503.10522
-
ChatGPT Encounters Morphing Attack Detection: Zero-Shot MAD with Multi-Modal Large Language Models and General Vision Models 13 Mar 2025 · 0 repositories · arXiv:2503.10937
-
CoCMT: Communication-Efficient Cross-Modal Transformer for Collaborative Perception 13 Mar 2025 · 0 repositories · arXiv:2503.13504
-
Cosh-DiT: Co-Speech Gesture Video Synthesis via Hybrid Audio-Visual Diffusion Transformers 13 Mar 2025 · 0 repositories · arXiv:2503.09942
-
CountPath: Automating Fragment Counting in Digital Pathology 13 Mar 2025 · 0 repositories · arXiv:2503.10520
-
Do I look like a `cat.n.01` to you? A Taxonomy Image Generation Benchmark 13 Mar 2025 · 0 repositories · arXiv:2503.10357
-
Emotion Recognition with CLIP and Sequential Learning 13 Mar 2025 · 0 repositories · arXiv:2503.09929
-
Fixed-Point RNNs: From Diagonal to Dense in a Few Iterations 13 Mar 2025 · 0 repositories · arXiv:2503.10799
-
Gumiho: A Hybrid Architecture to Prioritize Early Tokens in Speculative Decoding 13 Mar 2025 · 0 repositories · arXiv:2503.10135Syntology 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 8 pointer-only (licence)
-
KV-Distill: Nearly Lossless Learnable Context Compression for LLMs 13 Mar 2025 · 0 repositories · arXiv:2503.10337
-
Radar: Fast Long-Context Decoding for Any Transformer 13 Mar 2025 · 1 repository · arXiv:2503.10571Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Robustness Tokens: Towards Adversarial Robustness of Transformers 13 Mar 2025 · 1 repository · arXiv:2503.10191
-
Tempest: Autonomous Multi-Turn Jailbreaking of Large Language Models with Tree Search 13 Mar 2025 · 0 repositories · arXiv:2503.10619
-
TacticExpert: Spatial-Temporal Graph Language Model for Basketball Tactics 13 Mar 2025 · 0 repositories · arXiv:2503.10722
-
TGP: Two-modal occupancy prediction with 3D Gaussian and sparse points for 3D Environment Awareness 13 Mar 2025 · 0 repositories · arXiv:2503.09941
-
Towards Efficient Large Scale Spatial-Temporal Time Series Forecasting via Improved Inverted Transformers 13 Mar 2025 · 0 repositories · arXiv:2503.10858
-
Why Does Your CoT Prompt (Not) Work? Theoretical Analysis of Prompt Space Complexity, its Interaction with Answer Space During CoT Reasoning with LLMs: A Recurrent Perspective 13 Mar 2025 · 0 repositories · arXiv:2503.10084
-
4D-ACFNet: A 4D Attention Mechanism-Based Prognostic Framework for Colorectal Cancer Liver Metastasis Integrating Multimodal Spatiotemporal Features 12 Mar 2025 · 0 repositories · arXiv:2503.09652
-
AgentDAM: Privacy Leakage Evaluation for Autonomous Web Agents 12 Mar 2025 · 1 repository · arXiv:2503.09780Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
An Evaluation of LLMs for Detecting Harmful Computing Terms 12 Mar 2025 · 0 repositories · arXiv:2503.09341
-
Discovering Influential Neuron Path in Vision Transformers 12 Mar 2025 · 0 repositories · arXiv:2503.09046
-
Finding the Muses: Identifying Coresets through Loss Trajectories 12 Mar 2025 · 0 repositories · arXiv:2503.09721
-
How to Protect Yourself from 5G Radiation? Investigating LLM Responses to Implicit Misinformation 12 Mar 2025 · 1 repository · arXiv:2503.09598
-
Minimal Time Series Transformer 12 Mar 2025 · 1 repository · arXiv:2503.09791
-
Language-Enhanced Representation Learning for Single-Cell Transcriptomics 12 Mar 2025 · 1 repository · arXiv:2503.09427Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
NAMI: Efficient Image Generation via Progressive Rectified Flow Transformers 12 Mar 2025 · 0 repositories · arXiv:2503.09242
-
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latant Space 12 Mar 2025 · 0 repositories · arXiv:2503.09215
-
Post-interactive Multimodal Trajectory Prediction for Autonomous Driving 12 Mar 2025 · 0 repositories · arXiv:2503.09366
-
Robust Multimodal Survival Prediction with the Latent Differentiation Conditional Variational AutoEncoder 12 Mar 2025 · 1 repository · arXiv:2503.09496Syntology official (archive's flag): 5 ran · 6 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
SE(3)-Equivariant Robot Learning and Control: A Tutorial Survey 12 Mar 2025 · 0 repositories · arXiv:2503.09829
-
TA-V2A: Textually Assisted Video-to-Audio Generation 12 Mar 2025 · 0 repositories · arXiv:2503.10700
-
Un-Straightening Generative AI: How Queer Artists Surface and Challenge the Normativity of Generative AI Models 12 Mar 2025 · 0 repositories · arXiv:2503.09805
-
Unified Locomotion Transformer with Simultaneous Sim-to-Real Transfer for Quadrupeds 12 Mar 2025 · 0 repositories · arXiv:2503.08997
-
Who Are You Behind the Screen? Implicit MBTI and Gender Detection Using Artificial Intelligence 12 Mar 2025 · 0 repositories · arXiv:2503.09853
-
Accurate INT8 Training Through Dynamic Block-Level Fallback 11 Mar 2025 · 0 repositories · arXiv:2503.08040
-
EFPC: Towards Efficient and Flexible Prompt Compression 11 Mar 2025 · 0 repositories · arXiv:2503.07956
-
External Knowledge Injection for CLIP-Based Class-Incremental Learning 11 Mar 2025 · 3 repositories · arXiv:2503.08510
-
GPT-PPG: A GPT-based Foundation Model for Photoplethysmography Signals 11 Mar 2025 · 0 repositories · arXiv:2503.08015
-
HOTFormerLoc: Hierarchical Octree Transformer for Versatile Lidar Place Recognition Across Ground and Aerial Views 11 Mar 2025 · 0 repositories · arXiv:2503.08140
-
KAN-Mixers: a new deep learning architecture for image classification 11 Mar 2025 · 0 repositories · arXiv:2503.08939
-
QUIET-SR: Quantum Image Enhancement Transformer for Single Image Super-Resolution 11 Mar 2025 · 0 repositories · arXiv:2503.08759
-
Seeing What's Not There: Spurious Correlation in Multimodal LLMs 11 Mar 2025 · 0 repositories · arXiv:2503.08884
-
TransECG: Leveraging Transformers for Explainable ECG Re-identification Risk Analysis 11 Mar 2025 · 0 repositories · arXiv:2503.13495
-
A LSTM-Transformer Model for pulsation control of pVADs 10 Mar 2025 · 0 repositories · arXiv:2503.07110
-
Bot Wars Evolved: Orchestrating Competing LLMs in a Counterstrike Against Phone Scams 10 Mar 2025 · 0 repositories · arXiv:2503.07036
-
Enhancing Time Series Forecasting via Logic-Inspired Regularization 10 Mar 2025 · 0 repositories · arXiv:2503.06867
-
Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings 10 Mar 2025 · 0 repositories · arXiv:2503.06980
-
Exposure Bias Reduction for Enhancing Diffusion Transformer Feature Caching 10 Mar 2025 · 1 repository · arXiv:2503.07120
-
From Idea to Implementation: Evaluating the Influence of Large Language Models in Software Development -- An Opinion Paper 10 Mar 2025 · 0 repositories · arXiv:2503.07450
-
Large model enhanced computational ghost imaging 10 Mar 2025 · 1 repository · arXiv:2503.08710
-
MambaFlow: A Mamba-Centric Architecture for End-to-End Optical Flow Estimation 10 Mar 2025 · 0 repositories · arXiv:2503.07046
-
Post-Training Quantization for Diffusion Transformer via Hierarchical Timestep Grouping 10 Mar 2025 · 0 repositories · arXiv:2503.06930
-
ResMoE: Space-efficient Compression of Mixture of Experts LLMs via Residual Restoration 10 Mar 2025 · 1 repository · arXiv:2503.06881Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
VACE: All-in-One Video Creation and Editing 10 Mar 2025 · 2 repositories · arXiv:2503.07598Syntology 11 ran (of which 5 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 8 unverified (of 19 harvested samples)
-
Beyond Decoder-only: Large Language Models Can be Good Encoders for Machine Translation 9 Mar 2025 · 1 repository · arXiv:2503.06594
-
Global-Aware Monocular Semantic Scene Completion with State Space Models 9 Mar 2025 · 0 repositories · arXiv:2503.06569
-
GroMo: Plant Growth Modeling with Multiview Images 9 Mar 2025 · 1 repository · arXiv:2503.06608
-
LSA: Latent Style Augmentation Towards Stain-Agnostic Cervical Cancer Screening 9 Mar 2025 · 0 repositories · arXiv:2503.06563
-
SKG-LLM: Developing a Mathematical Model for Stroke Knowledge Graph Construction Using Large Language Models 9 Mar 2025 · 0 repositories · arXiv:2503.06475
-
A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment 8 Mar 2025 · 1 repository · arXiv:2503.06241
-
End-to-End Action Segmentation Transformer 8 Mar 2025 · 0 repositories · arXiv:2503.06316
-
End-to-End HOI Reconstruction Transformer with Graph-based Encoding 8 Mar 2025 · 0 repositories · arXiv:2503.06012
-
Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision 8 Mar 2025 · 0 repositories · arXiv:2503.06089
-
Lightweight Software Kernels and Hardware Extensions for Efficient Sparse Deep Neural Networks on Microcontrollers 8 Mar 2025 · 0 repositories · arXiv:2503.06183
-
Optimizing Generative AI's Accuracy and Transparency in Inductive Thematic Analysis: A Human-AI Comparison 8 Mar 2025 · 0 repositories · arXiv:2503.16485
-
X2I: Seamless Integration of Multimodal Understanding into Diffusion Transformer via Attention Distillation 8 Mar 2025 · 1 repository · arXiv:2503.06134Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Hybrid Model/Data-Driven Solution to Channel, Position and Orientation Tracking in mmWave Vehicular Systems 7 Mar 2025 · 0 repositories · arXiv:2503.05091
-
A Real-time Multimodal Transformer Neural Network-powered Wildfire Forecasting System 7 Mar 2025 · 0 repositories · arXiv:2503.05971
-
FastMap: Fast Queries Initialization Based Vectorized HD Map Reconstruction Framework 7 Mar 2025 · 1 repository · arXiv:2503.05492
-
FMCHS: Advancing Traditional Chinese Medicine Herb Recommendation with Fusion of Multiscale Correlations of Herbs and Symptoms 7 Mar 2025 · 0 repositories · arXiv:2503.05167
-
FMT:A Multimodal Pneumonia Detection Model Based on Stacking MOE Framework 7 Mar 2025 · 0 repositories · arXiv:2503.05626
-
Language modelling techniques for analysing the impact of human genetic variation 7 Mar 2025 · 0 repositories · arXiv:2503.10655
-
MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice 7 Mar 2025 · 0 repositories · arXiv:2503.05978
-
A Generalist Cross-Domain Molecular Learning Framework for Structure-Based Drug Discovery 6 Mar 2025 · 0 repositories · arXiv:2503.04362
-
BicliqueEncoder: An Efficient Method for Link Prediction in Bipartite Networks using Formal Concept Analysis and Transformer Encoder 6 Mar 2025 · 0 repositories · arXiv:2503.07645
-
Can We Optimize Deep RL Policy Weights as Trajectory Modeling? 6 Mar 2025 · 0 repositories · arXiv:2503.04074
-
DB-Explore: Automated Database Exploration and Instruction Synthesis for Text-to-SQL 6 Mar 2025 · 0 repositories · arXiv:2503.04959
-
GBT-SAM: Adapting a Foundational Deep Learning Model for Generalizable Brain Tumor Segmentation via Efficient Integration of Multi-Parametric MRI Data 6 Mar 2025 · 1 repository · arXiv:2503.04325
-
Hedging with Sparse Reward Reinforcement Learning 6 Mar 2025 · 0 repositories · arXiv:2503.04218
-
High-Precision Transformer-Based Visual Servoing for Humanoid Robots in Aligning Tiny Objects 6 Mar 2025 · 0 repositories · arXiv:2503.04862
-
Incentivizing Multi-Tenant Split Federated Learning for Foundation Models at the Network Edge 6 Mar 2025 · 0 repositories · arXiv:2503.04971
-
Learning Transformer-based World Models with Contrastive Predictive Coding 6 Mar 2025 · 0 repositories · arXiv:2503.04416
-
Leveraging Large Language Models to Address Data Scarcity in Machine Learning: Applications in Graphene Synthesis 6 Mar 2025 · 1 repository · arXiv:2503.04870
-
Toward Lightweight and Fast Decoders for Diffusion Models in Image and Video Generation 6 Mar 2025 · 1 repository · arXiv:2503.04871
-
Towards Autonomous Reinforcement Learning for Real-World Robotic Manipulation with Large Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04280
-
A Multimodal Framework for Topic Propagation Classification in Social Networks 5 Mar 2025 · 0 repositories · arXiv:2503.03112
-
All-atom Diffusion Transformers: Unified generative modelling of molecules and materials 5 Mar 2025 · 1 repository · arXiv:2503.03965Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
DTU-Net: A Multi-Scale Dilated Transformer Network for Nonlinear Hyperspectral Unmixing 5 Mar 2025 · 0 repositories · arXiv:2503.03465
-
Large language models in finance : what is financial sentiment? 5 Mar 2025 · 0 repositories · arXiv:2503.03612
-
MA-LoT: Multi-Agent Lean-based Long Chain-of-Thought Reasoning enhances Formal Theorem Proving 5 Mar 2025 · 1 repository · arXiv:2503.03205
-
PathRWKV: Enabling Whole Slide Prediction with Recurrent-Transformer 5 Mar 2025 · 0 repositories · arXiv:2503.03199
-
Pretrained LLMs as Real-Time Controllers for Robot Operated Serial Production Line 5 Mar 2025 · 0 repositories · arXiv:2503.03889
-
RiskAgent: Autonomous Medical AI Copilot for Generalist Risk Prediction 5 Mar 2025 · 0 repositories · arXiv:2503.03802
-
ScaleFusionNet: Transformer-Guided Multi-Scale Feature Fusion for Skin Lesion Segmentation 5 Mar 2025 · 1 repository · arXiv:2503.03327