Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 13
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 13 of 139: papers 1,201 to 1,300 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Quantized Spike-driven Transformer 23 Jan 2025 · 1 repository · arXiv:2501.13492Syntology official (archive's flag): 11 ran · 13 ran (of which 9 constructed an object rather than computing a result; 13 with no instrument failure: 2 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 12 unverified (of 25 harvested samples) · 25 pointer-only (licence)
-
Question Answering on Patient Medical Records with Private Fine-Tuned LLMs 23 Jan 2025 · 0 repositories · arXiv:2501.13687
-
Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models 23 Jan 2025 · 0 repositories · arXiv:2501.13629
-
Text-driven Online Action Detection 23 Jan 2025 · 1 repository · arXiv:2501.13518
-
Utilizing Evolution Strategies to Train Transformers in Reinforcement Learning 23 Jan 2025 · 1 repository · arXiv:2501.13883Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Ehrenfeucht-Haussler Rank and Chain of Thought 22 Jan 2025 · 0 repositories · arXiv:2501.12997
-
EmoFormer: A Text-Independent Speech Emotion Recognition using a Hybrid Transformer-CNN model 22 Jan 2025 · 0 repositories · arXiv:2501.12682
-
LiT: Delving into a Simplified Linear Diffusion Transformer for Image Generation 22 Jan 2025 · 0 repositories · arXiv:2501.12976
-
Multimodal AI on Wound Images and Clinical Notes for Home Patient Referral 22 Jan 2025 · 0 repositories · arXiv:2501.13247
-
SRMT: Shared Memory for Multi-agent Lifelong Pathfinding 22 Jan 2025 · 1 repository · arXiv:2501.13200
-
T-Graphormer: Using Transformers for Spatiotemporal Forecasting 22 Jan 2025 · 1 repository · arXiv:2501.13274
-
Automatic Labelling with Open-source LLMs using Dynamic Label Schema Integration 21 Jan 2025 · 0 repositories · arXiv:2501.12332
-
Continuous 3D Perception Model with Persistent State 21 Jan 2025 · 0 repositories · arXiv:2501.12387
-
DLEN: Dual Branch of Transformer for Low-Light Image Enhancement in Dual Domains 21 Jan 2025 · 0 repositories · arXiv:2501.12235
-
Episodic Memories Generation and Evaluation Benchmark for Large Language Models 21 Jan 2025 · 1 repository · arXiv:2501.13121Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Panoramic Interests: Stylistic-Content Aware Personalized Headline Generation 21 Jan 2025 · 1 repository · arXiv:2501.11900
-
Towards Accurate Unified Anomaly Segmentation 21 Jan 2025 · 1 repository · arXiv:2501.12295
-
Vision-Language Models for Automated Chest X-ray Interpretation: Leveraging ViT and GPT-2 21 Jan 2025 · 0 repositories · arXiv:2501.12356
-
Adaptive parameters identification for nonlinear dynamics using deep permutation invariant networks 20 Jan 2025 · 0 repositories · arXiv:2501.11350
-
DLinear-based Prediction of Remaining Useful Life of Lithium-Ion Batteries: Feature Engineering through Explainable Artificial Intelligence 20 Jan 2025 · 0 repositories · arXiv:2501.11542
-
Early evidence of how LLMs outperform traditional systems on OCR/HTR tasks for historical records 20 Jan 2025 · 1 repository · arXiv:2501.11623
-
Generative AI-enabled Blockage Prediction for Robust Dual-Band mmWave Communication 20 Jan 2025 · 0 repositories · arXiv:2501.11763
-
Glinthawk: A Two-Tiered Architecture for Offline LLM Inference 20 Jan 2025 · 1 repository · arXiv:2501.11779
-
KEIR @ ECIR 2025: The Second Workshop on Knowledge-Enhanced Information Retrieval 20 Jan 2025 · 0 repositories · arXiv:2501.11499
-
Trustformer: A Trusted Federated Transformer 20 Jan 2025 · 0 repositories · arXiv:2501.11706
-
Chain-of-Reasoning: Towards Unified Mathematical Reasoning in Large Language Models via a Multi-Paradigm Perspective 19 Jan 2025 · 0 repositories · arXiv:2501.11110
-
A CNN-Transformer for Classification of Longitudinal 3D MRI Images -- A Case Study on Hepatocellular Carcinoma Prediction 18 Jan 2025 · 1 repository · arXiv:2501.10733
-
Dynamic Trend Fusion Module for Traffic Flow Prediction 18 Jan 2025 · 1 repository · arXiv:2501.10796
-
LD-DETR: Loop Decoder DEtection TRansformer for Video Moment Retrieval and Highlight Detection 18 Jan 2025 · 1 repository · arXiv:2501.10787
-
Neural Algorithmic Reasoning for Hypergraphs with Looped Transformers 18 Jan 2025 · 0 repositories · arXiv:2501.10688
-
Enhancing the Reliability in Machine Learning for Gravitational Wave Parameter Estimation with Attention-Based Models 17 Jan 2025 · 0 repositories · arXiv:2501.10486
-
PaSa: An LLM Agent for Comprehensive Academic Paper Search 17 Jan 2025 · 1 repository · arXiv:2501.10120Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Self-Clustering Graph Transformer Approach to Model Resting-State Functional Brain Activity 17 Jan 2025 · 0 repositories · arXiv:2501.16345
-
A Simple Aerial Detection Baseline of Multimodal Language Models 16 Jan 2025 · 1 repository · arXiv:2501.09720
-
Generalized Single-Image-Based Morphing Attack Detection Using Deep Representations from Vision Transformer 16 Jan 2025 · 0 repositories · arXiv:2501.09817
-
HSPFormer: Hierarchical Spatial Perception Transformer for Semantic Segmentation 16 Jan 2025 · 1 repository
-
Learnings from Scaling Visual Tokenizers for Reconstruction and Generation 16 Jan 2025 · 0 repositories · arXiv:2501.09755
-
Perspective Transition of Large Language Models for Solving Subjective Tasks 16 Jan 2025 · 0 repositories · arXiv:2501.09265
-
Practical Continual Forgetting for Pre-trained Vision Models 16 Jan 2025 · 1 repository · arXiv:2501.09705
-
Towards Robust and Realistic Human Pose Estimation via WiFi Signals 16 Jan 2025 · 1 repository · arXiv:2501.09411
-
Unified Face Matching and Physical-Digital Spoofing Attack Detection 16 Jan 2025 · 0 repositories · arXiv:2501.09635
-
Attention is All You Need Until You Need Retention 15 Jan 2025 · 0 repositories · arXiv:2501.09166
-
Beyond Speaker Identity: Text Guided Target Speech Extraction 15 Jan 2025 · 1 repository · arXiv:2501.09169
-
BRIGHT-VO: Brightness-Guided Hybrid Transformer for Visual Odometry with Multi-modality Refinement Module 15 Jan 2025 · 1 repository · arXiv:2501.08659
-
Cancer-Net PCa-Seg: Benchmarking Deep Learning Models for Prostate Cancer Segmentation Using Synthetic Correlated Diffusion Imaging 15 Jan 2025 · 0 repositories · arXiv:2501.09185
-
CT-PatchTST: Channel-Time Patch Time-Series Transformer for Long-Term Renewable Energy Forecasting 15 Jan 2025 · 0 repositories · arXiv:2501.08620
-
Enhanced Large Language Models for Effective Screening of Depression and Anxiety 15 Jan 2025 · 0 repositories · arXiv:2501.08769
-
MIAFEx: An Attention-based Feature Extraction Method for Medical Image Classification 15 Jan 2025 · 0 repositories · arXiv:2501.08562
-
Multi-View Transformers for Airway-To-Lung Ratio Inference on Cardiac CT Scans: The C4R Study 15 Jan 2025 · 0 repositories · arXiv:2501.08902
-
Multimodal Fake News Video Explanation: Dataset, Analysis and Evaluation 15 Jan 2025 · 0 repositories · arXiv:2501.08514
-
SuperSAM: Crafting a SAM Supernetwork via Structured Pruning and Unstructured Parameter Prioritization 15 Jan 2025 · 1 repository · arXiv:2501.08504
-
SwinTExCo: Exemplar-based video colorization using Swin Transformer 15 Jan 2025 · 1 repository
-
Active Sampling for Node Attribute Completion on Graphs 14 Jan 2025 · 0 repositories · arXiv:2501.08450
-
Decision Transformers for RIS-Assisted Systems with Diffusion Model-Based Channel Acquisition 14 Jan 2025 · 0 repositories · arXiv:2501.08007
-
Decoding Interpretable Logic Rules from Neural Networks 14 Jan 2025 · 0 repositories · arXiv:2501.08281
-
Efficient Deep Learning-based Forward Solvers for Brain Tumor Growth Models 14 Jan 2025 · 1 repository · arXiv:2501.08226
-
EmoNeXt: an Adapted ConvNeXt for Facial Emotion Recognition 14 Jan 2025 · 1 repository · arXiv:2501.08199
-
Exploring Narrative Clustering in Large Language Models: A Layerwise Analysis of BERT 14 Jan 2025 · 0 repositories · arXiv:2501.08053
-
Large Language Models For Text Classification: Case Study And Comprehensive Review 14 Jan 2025 · 0 repositories · arXiv:2501.08457
-
Optimizing Language Models for Grammatical Acceptability: A Comparative Study of Fine-Tuning Techniques 14 Jan 2025 · 0 repositories · arXiv:2501.07853
-
PokerBench: Training Large Language Models to become Professional Poker Players 14 Jan 2025 · 1 repository · arXiv:2501.08328
-
PSReg: Prior-guided Sparse Mixture of Experts for Point Cloud Registration 14 Jan 2025 · 0 repositories · arXiv:2501.07762
-
Towards Lightweight Time Series Forecasting: a Patch-wise Transformer with Weak Data Enriching 14 Jan 2025 · 0 repositories · arXiv:2501.10448
-
Transforming Indoor Localization: Advanced Transformer Architecture for NLOS Dominated Wireless Environments with Distributed Sensors 14 Jan 2025 · 0 repositories · arXiv:2501.07774
-
UFGraphFR: An attempt at a federated recommendation system based on user text characteristics 14 Jan 2025 · 1 repository · arXiv:2501.08044
-
Comparative analysis of optical character recognition methods for Sámi texts from the National Library of Norway 13 Jan 2025 · 2 repositories · arXiv:2501.07300
-
D3MES: Diffusion Transformer with multihead equivariant self-attention for 3D molecule generation 13 Jan 2025 · 1 repository · arXiv:2501.07077
-
EdgeTAM: On-Device Track Anything Model 13 Jan 2025 · 1 repository · arXiv:2501.07256Syntology official: no sample here; runs from other or unrecorded repositories · 7 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Estimating Musical Surprisal in Audio 13 Jan 2025 · 1 repository · arXiv:2501.07474
-
Future-Conditioned Recommendations with Multi-Objective Controllable Decision Transformer 13 Jan 2025 · 0 repositories · arXiv:2501.07212
-
Scaling Up ESM2 Architectures for Long Protein Sequences Analysis: Long and Quantized Approaches 13 Jan 2025 · 0 repositories · arXiv:2501.07747
-
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing 13 Jan 2025 · 1 repository · arXiv:2501.07554
-
UNetVL: Enhancing 3D Medical Image Segmentation with Chebyshev KAN Powered Vision-LSTM 13 Jan 2025 · 1 repository · arXiv:2501.07017
-
Better Prompt Compression Without Multi-Layer Perceptrons 12 Jan 2025 · 0 repositories · arXiv:2501.06730
-
DRDT3: Diffusion-Refined Decision Test-Time Training Model 12 Jan 2025 · 0 repositories · arXiv:2501.06718
-
Generative Artificial Intelligence-Supported Pentesting: A Comparison between Claude Opus, GPT-4, and Copilot 12 Jan 2025 · 0 repositories · arXiv:2501.06963
-
MFConvTr: Multi-Frequency Convolutional Transformer for Fetal Arrhythmia Detection in Non-Invasive fECG 12 Jan 2025 · 0 repositories · arXiv:2501.16649
-
Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous Learning 12 Jan 2025 · 1 repository · arXiv:2501.06884Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
ZNO-Eval: Benchmarking reasoning capabilities of large language models in Ukrainian 12 Jan 2025 · 1 repository · arXiv:2501.06715
-
A Comparative Performance Analysis of Classification and Segmentation Models on Bangladeshi Pothole Dataset 11 Jan 2025 · 0 repositories · arXiv:2501.06602
-
CeViT: Copula-Enhanced Vision Transformer in multi-task learning and bi-group image covariates with an application to myopia screening 11 Jan 2025 · 1 repository · arXiv:2501.06540
-
Flash Window Attention: speedup the attention computation for Swin Transformer 11 Jan 2025 · 2 repositories · arXiv:2501.06480
-
FocusDD: Real-World Scene Infusion for Robust Dataset Distillation 11 Jan 2025 · 0 repositories · arXiv:2501.06405
-
Ladder-residual: parallelism-aware architecture for accelerating large model inference with communication overlapping 11 Jan 2025 · 1 repository · arXiv:2501.06589Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Tensor Product Attention Is All You Need 11 Jan 2025 · 1 repository · arXiv:2501.06425Syntology official (archive's flag): 5 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization 11 Jan 2025 · 0 repositories · arXiv:2501.06663
-
A Holistically Point-guided Text Framework for Weakly-Supervised Camouflaged Object Detection 10 Jan 2025 · 0 repositories · arXiv:2501.06038
-
An Attention-Guided Deep Learning Approach for Classifying 39 Skin Lesion Types 10 Jan 2025 · 1 repository · arXiv:2501.05991
-
Analyzing Spatio-Temporal Dynamics of Dissolved Oxygen for the River Thames using Superstatistical Methods and Machine Learning 10 Jan 2025 · 0 repositories · arXiv:2501.07599
-
Binary Event-Driven Spiking Transformer 10 Jan 2025 · 0 repositories · arXiv:2501.05904
-
Iconicity in Large Language Models 10 Jan 2025 · 0 repositories · arXiv:2501.05643
-
Merging Feed-Forward Sublayers for Compressed Transformers 10 Jan 2025 · 1 repository · arXiv:2501.06126
-
Mix-QViT: Mixed-Precision Vision Transformer Quantization Driven by Layer Importance and Quantization Sensitivity 10 Jan 2025 · 0 repositories · arXiv:2501.06357
-
Model Inversion in Split Learning for Personalized LLMs: New Insights from Information Bottleneck Theory 10 Jan 2025 · 0 repositories · arXiv:2501.05965
-
MSCViT: A Small-size ViT architecture with Multi-Scale Self-Attention Mechanism for Tiny Datasets 10 Jan 2025 · 0 repositories · arXiv:2501.06040
-
Multi-subject Open-set Personalization in Video Generation 10 Jan 2025 · 0 repositories · arXiv:2501.06187
-
Swin-X2S: Reconstructing 3D Shape from 2D Biplanar X-ray with Swin Transformers 10 Jan 2025 · 0 repositories · arXiv:2501.05961
-
TTS-Transducer: End-to-End Speech Synthesis with Neural Transducer 10 Jan 2025 · 0 repositories · arXiv:2501.06320
-
Weakly Supervised Segmentation of Hyper-Reflective Foci with Compact Convolutional Transformers and SAM2 10 Jan 2025 · 0 repositories · arXiv:2501.05933
-
Large language models streamline automated systematic review: A preliminary study 9 Jan 2025 · 0 repositories · arXiv:2502.15702