Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 27
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 27 of 139: papers 2,601 to 2,700 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
VMAS: Video-to-Music Generation via Semantic Alignment in Web Music Videos 11 Sep 2024 · 0 repositories · arXiv:2409.07450
-
Weather-Informed Probabilistic Forecasting and Scenario Generation in Power Systems 11 Sep 2024 · 0 repositories · arXiv:2409.07637
-
Mapping Biomedical Ontology Terms to IDs: Effect of Domain Prevalence on Prediction Accuracy 11 Sep 2024 · 0 repositories · arXiv:2409.13746
-
A Dataset for Evaluating LLM-based Evaluation Functions for Research Question Extraction Task 10 Sep 2024 · 0 repositories · arXiv:2409.06883
-
A Practical Gated Recurrent Transformer Network Incorporating Multiple Fusions for Video Denoising 10 Sep 2024 · 0 repositories · arXiv:2409.06603
-
Adaptive Transformer Modelling of Density Function for Nonparametric Survival Analysis 10 Sep 2024 · 1 repository · arXiv:2409.06209Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
AgileIR: Memory-Efficient Group Shifted Windows Attention for Agile Image Restoration 10 Sep 2024 · 0 repositories · arXiv:2409.06206
-
An End-to-End Approach for Chord-Conditioned Song Generation 10 Sep 2024 · 0 repositories · arXiv:2409.06307
-
Can Large Language Models Unlock Novel Scientific Research Ideas? 10 Sep 2024 · 1 repository · arXiv:2409.06185Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)
-
DA-MoE: Towards Dynamic Expert Allocation for Mixture-of-Experts Models 10 Sep 2024 · 0 repositories · arXiv:2409.06669
-
GroUSE: A Benchmark to Evaluate Evaluators in Grounded Question Answering 10 Sep 2024 · 1 repository · arXiv:2409.06595
-
Knowledge Distillation via Query Selection for Detection Transformer 10 Sep 2024 · 0 repositories · arXiv:2409.06443
-
Lightweight single-image super-resolution network based on dual paths 10 Sep 2024 · 0 repositories · arXiv:2409.06590
-
Static for Dynamic: Towards a Deeper Understanding of Dynamic Facial Expressions Using Static Expression Data 10 Sep 2024 · 1 repository · arXiv:2409.06154
-
What is the Role of Small Models in the LLM Era: A Survey 10 Sep 2024 · 1 repository · arXiv:2409.06857
-
Classification performance and reproducibility of GPT-4 omni for information extraction from veterinary electronic health records 9 Sep 2024 · 1 repository · arXiv:2409.13727
-
Rule Extrapolation in Language Models: A Study of Compositional Generalization on OOD Prompts 9 Sep 2024 · 1 repository · arXiv:2409.13728
-
AbGPT: De Novo Antibody Design via Generative Language Modeling 9 Sep 2024 · 1 repository · arXiv:2409.06090
-
Deep Generative Model for Mechanical System Configuration Design 9 Sep 2024 · 0 repositories · arXiv:2409.06016
-
DriveScape: Towards High-Resolution Controllable Multi-View Driving Video Generation 9 Sep 2024 · 0 repositories · arXiv:2409.05463
-
DSDFormer: An Innovative Transformer-Mamba Framework for Robust High-Precision Driver Distraction Identification 9 Sep 2024 · 0 repositories · arXiv:2409.05587
-
Exploring Rich Subjective Quality Information for Image Quality Assessment in the Wild 9 Sep 2024 · 0 repositories · arXiv:2409.05540
-
FairHome: A Fair Housing and Fair Lending Dataset 9 Sep 2024 · 0 repositories · arXiv:2409.05990
-
Identifying the sources of ideological bias in GPT models through linguistic variation in output 9 Sep 2024 · 0 repositories · arXiv:2409.06043
-
Retrofitting Temporal Graph Neural Networks with Transformer 9 Sep 2024 · 1 repository · arXiv:2409.05477
-
RotCAtt-TransUNet++: Novel Deep Neural Network for Sophisticated Cardiac Segmentation 9 Sep 2024 · 1 repository · arXiv:2409.05280
-
Towards Building a Robust Knowledge Intensive Question Answering Model with Large Language Models 9 Sep 2024 · 0 repositories · arXiv:2409.05385
-
Zero-shot Outlier Detection via Prior-data Fitted Networks: Model Selection Bygone! 9 Sep 2024 · 0 repositories · arXiv:2409.05672
-
Audio-Guided Fusion Techniques for Multimodal Emotion Analysis 8 Sep 2024 · 0 repositories · arXiv:2409.05007
-
Lung-DETR: Deformable Detection Transformer for Sparse Lung Nodule Anomaly Detection 8 Sep 2024 · 0 repositories · arXiv:2409.05200
-
Activation Function Optimization Scheme for Image Classification 7 Sep 2024 · 1 repository · arXiv:2409.04915
-
Towards Weather-Robust 3D Human Body Reconstruction: Millimeter-Wave Radar-Based Dataset, Benchmark, and Multi-Modal Fusion 7 Sep 2024 · 0 repositories · arXiv:2409.04851
-
Cross-attention Inspired Selective State Space Models for Target Sound Extraction 7 Sep 2024 · 1 repository · arXiv:2409.04803Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Efficient Training of Transformers for Molecule Property Prediction on Small-scale Datasets 7 Sep 2024 · 0 repositories · arXiv:2409.04909
-
MuAP: Multi-step Adaptive Prompt Learning for Vision-Language Model with Missing Modality 7 Sep 2024 · 0 repositories · arXiv:2409.04693
-
NapTune: Efficient Model Tuning for Mood Classification using Previous Night's Sleep Measures along with Wearable Time-series 7 Sep 2024 · 0 repositories · arXiv:2409.04723
-
Swin Transformer for Robust Differentiation of Real and Synthetic Images: Intra- and Inter-Dataset Analysis 7 Sep 2024 · 0 repositories · arXiv:2409.04734
-
VidLPRO: A Video-Language Pre-training Framework for Robotic and Laparoscopic Surgery 7 Sep 2024 · 0 repositories · arXiv:2409.04732
-
ActionFlow: Equivariant, Accurate, and Efficient Policies with Spatially Symmetric Flow Matching 6 Sep 2024 · 0 repositories · arXiv:2409.04576
-
Advancing SEM Based Nano-Scale Defect Analysis in Semiconductor Manufacturing for Advanced IC Nodes 6 Sep 2024 · 0 repositories · arXiv:2409.04310
-
AnyMatch -- Efficient Zero-Shot Entity Matching with a Small Language Model 6 Sep 2024 · 1 repository · arXiv:2409.04073
-
Combining LLMs and Knowledge Graphs to Reduce Hallucinations in Question Answering 6 Sep 2024 · 0 repositories · arXiv:2409.04181
-
GALLa: Graph Aligned Large Language Models for Improved Source Code Understanding 6 Sep 2024 · 0 repositories · arXiv:2409.04183
-
Qihoo-T2X: An Efficient Proxy-Tokenized Diffusion Transformer for Text-to-Any-Task 6 Sep 2024 · 1 repository · arXiv:2409.04005
-
Retrieval Augmented Generation-Based Incident Resolution Recommendation System for IT Support 6 Sep 2024 · 0 repositories · arXiv:2409.13707
-
UI-JEPA: Towards Active Perception of User Intent through Onscreen User Activity 6 Sep 2024 · 0 repositories · arXiv:2409.04081
-
CACER: Clinical Concept Annotations for Cancer Events and Relations 5 Sep 2024 · 1 repository · arXiv:2409.03905
-
Characterizing Massive Activations of Attention Mechanism in Graph Neural Networks 5 Sep 2024 · 1 repository · arXiv:2409.03463
-
LMLT: Low-to-high Multi-Level Vision Transformer for Image Super-Resolution 5 Sep 2024 · 1 repository · arXiv:2409.03516
-
MaterialBENCH: Evaluating College-Level Materials Science Problem-Solving Abilities of Large Language Models 5 Sep 2024 · 0 repositories · arXiv:2409.03161
-
MVTN: A Multiscale Video Transformer Network for Hand Gesture Recognition 5 Sep 2024 · 1 repository · arXiv:2409.03890
-
Onboard Satellite Image Classification for Earth Observation: A Comparative Study of ViT Models 5 Sep 2024 · 1 repository · arXiv:2409.03901
-
Why mamba is effective? Exploit Linear Transformer-Mamba Network for Multi-Modality Image Fusion 5 Sep 2024 · 0 repositories · arXiv:2409.03223
-
xLAM: A Family of Large Action Models to Empower AI Agent Systems 5 Sep 2024 · 1 repository · arXiv:2409.03215
-
Causality-Aware Transformer Networks for Robotic Navigation 4 Sep 2024 · 0 repositories · arXiv:2409.02669
-
Detecting Calls to Action in Multimodal Content: Analysis of the 2021 German Federal Election Campaign on Instagram 4 Sep 2024 · 0 repositories · arXiv:2409.02690
-
Historical German Text Normalization Using Type- and Token-Based Language Modeling 4 Sep 2024 · 0 repositories · arXiv:2409.02841
-
How DREAMS are made: Emulating Satellite Galaxy and Subhalo Populations with Diffusion Models and Point Clouds 4 Sep 2024 · 1 repository · arXiv:2409.02980Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples)
-
How Privacy-Savvy Are Large Language Models? A Case Study on Compliance and Privacy Technical Review 4 Sep 2024 · 0 repositories · arXiv:2409.02375
-
Hypothesizing Missing Causal Variables with LLMs 4 Sep 2024 · 1 repository · arXiv:2409.02604
-
iConFormer: Dynamic Parameter-Efficient Tuning with Input-Conditioned Adaptation 4 Sep 2024 · 0 repositories · arXiv:2409.02838
-
Incorporating Like-Minded Peers to Overcome Friend Data Sparsity in Session-Based Social Recommendations 4 Sep 2024 · 0 repositories · arXiv:2409.02702
-
Irrelevant Alternatives Bias Large Language Model Hiring Decisions 4 Sep 2024 · 0 repositories · arXiv:2409.15299
-
Leveraging Interpretability in the Transformer to Automate the Proactive Scaling of Cloud Resources 4 Sep 2024 · 0 repositories · arXiv:2409.03103
-
LongLLaVA: Scaling Multi-modal LLMs to 1000 Images Efficiently via a Hybrid Architecture 4 Sep 2024 · 1 repository · arXiv:2409.02889Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
MobileUNETR: A Lightweight End-To-End Hybrid Vision Transformer For Efficient Medical Image Segmentation 4 Sep 2024 · 1 repository · arXiv:2409.03062
-
MOSMOS: Multi-organ segmentation facilitated by medical report supervision 4 Sep 2024 · 0 repositories · arXiv:2409.02418
-
Robust Text-to-Cypher Using Combination of BERT, GraphSAGE, and Transformer (CoBGT) Model 4 Sep 2024 · 0 repositories
-
Towards Data-Centric Face Anti-Spoofing: Improving Cross-domain Generalization via Physics-based Data Synthesis 4 Sep 2024 · 0 repositories · arXiv:2409.03501
-
Training Universal Vocoders with Feature Smoothing-Based Augmentation Methods for High-Quality TTS Systems 4 Sep 2024 · 0 repositories · arXiv:2409.02517
-
Wavelet GPT: Wavelet Inspired Large Language Models 4 Sep 2024 · 0 repositories · arXiv:2409.12924
-
1DCNNTrans: BISINDO Sign Language Interpreters in Improving the Inclusiveness of Public Services 3 Sep 2024 · 0 repositories · arXiv:2409.01975
-
Comparative Analysis of Learning-Based Methods for Transient Stability Assessment 3 Sep 2024 · 0 repositories · arXiv:2409.02336
-
F2former: When Fractional Fourier Meets Deep Wiener Deconvolution and Selective Frequency Transformer for Image Deblurring 3 Sep 2024 · 0 repositories · arXiv:2409.02056
-
Frequency-Spatial Entanglement Learning for Camouflaged Object Detection 3 Sep 2024 · 1 repository · arXiv:2409.01686Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Leveraging Large Language Models for Solving Rare MIP Challenges 3 Sep 2024 · 0 repositories · arXiv:2409.04464
-
On the Design Space Between Transformers and Recursive Neural Nets 3 Sep 2024 · 0 repositories · arXiv:2409.01531
-
PMT-MAE: Dual-Branch Self-Supervised Learning with Distillation for Efficient Point Cloud Classification 3 Sep 2024 · 0 repositories · arXiv:2409.02007
-
Self-Instructed Derived Prompt Generation Meets In-Context Learning: Unlocking New Potential of Black-Box LLMs 3 Sep 2024 · 0 repositories · arXiv:2409.01552
-
SPiKE: 3D Human Pose from Point Cloud Sequences 3 Sep 2024 · 1 repository · arXiv:2409.01879
-
The USTC-NERCSLIP Systems for the CHiME-8 NOTSOFAR-1 Challenge 3 Sep 2024 · 0 repositories · arXiv:2409.02041
-
TimeDiT: General-purpose Diffusion Transformers for Time Series Foundation Model 3 Sep 2024 · 0 repositories · arXiv:2409.02322
-
CLIBE: Detecting Dynamic Backdoors in Transformer-based NLP Models 2 Sep 2024 · 1 repository · arXiv:2409.01193Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 19 harvested samples) · 3 pointer-only (licence)
-
ESP-PCT: Enhanced VR Semantic Performance through Efficient Compression of Temporal and Spatial Redundancies in Point Cloud Transformers 2 Sep 2024 · 1 repository · arXiv:2409.01216
-
Evidential Transformers for Improved Image Retrieval 2 Sep 2024 · 0 repositories · arXiv:2409.01082
-
Large Language Models versus Classical Machine Learning: Performance in COVID-19 Mortality Prediction Using High-Dimensional Tabular Data 2 Sep 2024 · 2 repositories · arXiv:2409.02136
-
Multi-scale Temporal Fusion Transformer for Incomplete Vehicle Trajectory Prediction 2 Sep 2024 · 0 repositories · arXiv:2409.00904
-
Self-Judge: Selective Instruction Following with Alignment Self-Evaluation 2 Sep 2024 · 1 repository · arXiv:2409.00935Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Role of Transformer Models in Advancing Blockchain Technology: A Systematic Survey 2 Sep 2024 · 0 repositories · arXiv:2409.02139
-
ToolACE: Winning the Points of LLM Function Calling 2 Sep 2024 · 0 repositories · arXiv:2409.00920
-
Assessing UHD Image Quality from Aesthetics, Distortions, and Saliency 1 Sep 2024 · 1 repository · arXiv:2409.00749
-
Attention-Guided Multi-scale Interaction Network for Face Super-Resolution 1 Sep 2024 · 0 repositories · arXiv:2409.00591
-
Deep Knowledge-Infusion For Explainable Depression Detection 1 Sep 2024 · 0 repositories · arXiv:2409.02122
-
MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer 1 Sep 2024 · 1 repository · arXiv:2409.00750
-
ProteinRPN: Towards Accurate Protein Function Prediction with Graph-Based Region Proposals 1 Sep 2024 · 0 repositories · arXiv:2409.00610
-
Sample-Efficient Diffusion for Text-To-Speech Synthesis 1 Sep 2024 · 1 repository · arXiv:2409.03717
-
A Hybrid Transformer-Mamba Network for Single Image Deraining 31 Aug 2024 · 1 repository · arXiv:2409.00410
-
An Empirical Study on Information Extraction using Large Language Models 31 Aug 2024 · 0 repositories · arXiv:2409.00369
-
Chatting Up Attachment: Using LLMs to Predict Adult Bonds 31 Aug 2024 · 0 repositories · arXiv:2409.00347
-
Evaluating the Performance of Large Language Models in Competitive Programming: A Multi-Year, Multi-Grade Analysis 31 Aug 2024 · 0 repositories · arXiv:2409.09054