Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 39
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 39 of 139: papers 3,801 to 3,900 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CYCLO: Cyclic Graph Transformer Approach to Multi-Object Relationship Modeling in Aerial Videos 3 Jun 2024 · 0 repositories · arXiv:2406.01029
-
Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models 3 Jun 2024 · 1 repository · arXiv:2406.01698
-
Dimba: Transformer-Mamba Diffusion Models 3 Jun 2024 · 0 repositories · arXiv:2406.01159
-
LLEMamba: Low-Light Enhancement via Relighting-Guided Mamba with Deep Unfolding Network 3 Jun 2024 · 0 repositories · arXiv:2406.01028
-
Predicting Drug-Gene Relations via Analogy Tasks with Word Embeddings 3 Jun 2024 · 1 repository · arXiv:2406.00984
-
Progressive Inference: Explaining Decoder-Only Sequence Classification Models Using Intermediate Predictions 3 Jun 2024 · 0 repositories · arXiv:2406.02625
-
Prototypical Transformer as Unified Motion Learners 3 Jun 2024 · 0 repositories · arXiv:2406.01559
-
Seeing the Forest through the Trees: Data Leakage from Partial Transformer Gradients 3 Jun 2024 · 1 repository · arXiv:2406.00999
-
Sparse Focus Network for Multi-Source Remote Sensing Data Classification 3 Jun 2024 · 0 repositories · arXiv:2406.01245
-
Superhuman performance in urology board questions by an explainable large language model enabled for context integration of the European Association of Urology guidelines: the UroBot study 3 Jun 2024 · 0 repositories · arXiv:2406.01428
-
TimeCMA: Towards LLM-Empowered Multivariate Time Series Forecasting via Cross-Modality Alignment 3 Jun 2024 · 3 repositories · arXiv:2406.01638Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation 3 Jun 2024 · 2 repositories · arXiv:2406.01188
-
Utilizing Large Language Models for Automating Technical Customer Support 3 Jun 2024 · 0 repositories · arXiv:2406.01407
-
An Early Investigation into the Utility of Multimodal Large Language Models in Medical Imaging 2 Jun 2024 · 0 repositories · arXiv:2406.00667
-
Correlation Matching Transformation Transformers for UHD Image Restoration 2 Jun 2024 · 1 repository · arXiv:2406.00629Syntology official (archive's flag): 16 ran · 17 ran (of which 2 constructed an object rather than computing a result; 12 with no instrument failure: 2 honoured, 0 violated, 10 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Diffusion Tuning: Transferring Diffusion Models via Chain of Forgetting 2 Jun 2024 · 0 repositories · arXiv:2406.00773
-
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction 2 Jun 2024 · 1 repository · arXiv:2406.00755Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Exploiting Frequency Correlation for Hyperspectral Image Reconstruction 2 Jun 2024 · 0 repositories · arXiv:2406.00683
-
GLADformer: A Mixed Perspective for Graph-level Anomaly Detection 2 Jun 2024 · 0 repositories · arXiv:2406.00734
-
Kolmogorov-Arnold Network for Satellite Image Classification in Remote Sensing 2 Jun 2024 · 1 repository · arXiv:2406.00600
-
Learning to Play 7 Wonders Duel Without Human Supervision 2 Jun 2024 · 0 repositories · arXiv:2406.00741
-
Maximum-Entropy Regularized Decision Transformer with Reward Relabelling for Dynamic Recommendation 2 Jun 2024 · 0 repositories · arXiv:2406.00725
-
MGI: Multimodal Contrastive pre-training of Genomic and Medical Imaging 2 Jun 2024 · 0 repositories · arXiv:2406.00631
-
Presence or Absence: Are Unknown Word Usages in Dictionaries? 2 Jun 2024 · 1 repository · arXiv:2406.00656
-
An Evaluation Benchmark for Autoformalization in Lean4 1 Jun 2024 · 0 repositories · arXiv:2406.06555
-
Beyond Metrics: Evaluating LLMs' Effectiveness in Culturally Nuanced, Low-Resource Real-World Scenarios 1 Jun 2024 · 0 repositories · arXiv:2406.00343
-
Causal Contrastive Learning for Counterfactual Regression Over Time 1 Jun 2024 · 0 repositories · arXiv:2406.00535
-
Cross-Table Pretraining towards a Universal Function Space for Heterogeneous Tabular Data 1 Jun 2024 · 0 repositories · arXiv:2406.00281
-
Image Captioning via Dynamic Path Customization 1 Jun 2024 · 1 repository · arXiv:2406.00334
-
Phased Instruction Fine-Tuning for Large Language Models 1 Jun 2024 · 1 repository · arXiv:2406.04371Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
RoBERTa-BiLSTM: A Context-Aware Hybrid Model for Sentiment Analysis 1 Jun 2024 · 1 repository · arXiv:2406.00367
-
A Deep Learning Model for Coronary Artery Segmentation and Quantitative Stenosis Detection in Angiographic Images 1 Jun 2024 · 1 repository · arXiv:2406.00492
-
You Only Need Less Attention at Each Stage in Vision Transformers 1 Jun 2024 · 0 repositories · arXiv:2406.00427
-
A Comparative Study of CNN, ResNet, and Vision Transformers for Multi-Classification of Chest Diseases 31 May 2024 · 1 repository · arXiv:2406.00237
-
A Survey of Deep Learning Audio Generation Methods 31 May 2024 · 0 repositories · arXiv:2406.00146
-
Arbitrary-Length Generalization for Addition in a Tiny Transformer 31 May 2024 · 1 repository · arXiv:2406.00075
-
Decision Mamba: Reinforcement Learning via Hybrid Selective Sequence Modeling 31 May 2024 · 0 repositories · arXiv:2406.00079
-
Graph External Attention Enhanced Transformer 31 May 2024 · 1 repository · arXiv:2405.21061Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 2 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought 31 May 2024 · 1 repository · arXiv:2405.20692Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Knowledge Enhanced Multi-intent Transformer Network for Recommendation 31 May 2024 · 1 repository · arXiv:2405.20565
-
Leveraging Large Language Models for Entity Matching 31 May 2024 · 0 repositories · arXiv:2405.20624
-
Mamba State-Space Models Are Lyapunov-Stable Learners 31 May 2024 · 0 repositories · arXiv:2406.00209
-
MVAD: A Multiple Visual Artifact Detector for Video Streaming 31 May 2024 · 0 repositories · arXiv:2406.00212
-
Position Coupling: Improving Length Generalization of Arithmetic Transformers Using Task Structure 31 May 2024 · 1 repository · arXiv:2405.20671Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Query2CAD: Generating CAD models using natural language queries 31 May 2024 · 1 repository · arXiv:2406.00144Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Rough Transformers: Lightweight and Continuous Time Series Modelling through Signature Patching 31 May 2024 · 1 repository · arXiv:2405.20799Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Superlatives in Context: Modeling the Implicit Semantics of Superlatives 31 May 2024 · 1 repository · arXiv:2405.20967
-
Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis 31 May 2024 · 1 repository · arXiv:2405.21075Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
A Structure-Aware Lane Graph Transformer Model for Vehicle Trajectory Prediction 30 May 2024 · 0 repositories · arXiv:2405.20121
-
An Automatic Question Usability Evaluation Toolkit 30 May 2024 · 1 repository · arXiv:2405.20529
-
ANAH: Analytical Annotation of Hallucinations in Large Language Models 30 May 2024 · 1 repository · arXiv:2405.20315Syntology official: no sample here; runs from other or unrecorded repositories · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
AutoBreach: Universal and Adaptive Jailbreaking with Efficient Wordplay-Guided Optimization 30 May 2024 · 0 repositories · arXiv:2405.19668
-
Automated Generation and Tagging of Knowledge Components from Multiple-Choice Questions 30 May 2024 · 1 repository · arXiv:2405.20526
-
Automatic Graph Topology-Aware Transformer 30 May 2024 · 1 repository · arXiv:2405.19779
-
CLAY: A Controllable Large-scale Generative Model for Creating High-quality 3D Assets 30 May 2024 · 2 repositories · arXiv:2406.13897Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 3 pointer-only (licence)
-
Contextual Counting: A Mechanistic Study of Transformers on a Quantitative Task 30 May 2024 · 0 repositories · arXiv:2406.02585
-
Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation 30 May 2024 · 0 repositories · arXiv:2405.20092
-
Encoding and Controlling Global Semantics for Long-form Video Question Answering 30 May 2024 · 1 repository · arXiv:2405.19723
-
Fourier Controller Networks for Real-Time Decision-Making in Embodied Learning 30 May 2024 · 0 repositories · arXiv:2405.19885
-
GNN-RAG: Graph Neural Retrieval for Large Language Model Reasoning 30 May 2024 · 1 repository · arXiv:2405.20139Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Improving Object Detector Training on Synthetic Data by Starting With a Strong Baseline Methodology 30 May 2024 · 0 repositories · arXiv:2405.19822
-
LLaMEA: A Large Language Model Evolutionary Algorithm for Automatically Generating Metaheuristics 30 May 2024 · 2 repositories · arXiv:2405.20132Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
PATIENT-Ψ: Using Large Language Models to Simulate Patients for Training Mental Health Professionals 30 May 2024 · 1 repository · arXiv:2405.19660Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations 30 May 2024 · 1 repository · arXiv:2405.19740Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Phantom: General Trigger Attacks on Retrieval Augmented Language Generation 30 May 2024 · 0 repositories · arXiv:2405.20485
-
Preference Alignment with Flow Matching 30 May 2024 · 1 repository · arXiv:2405.19806Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
QClusformer: A Quantum Transformer-based Framework for Unsupervised Visual Clustering 30 May 2024 · 0 repositories · arXiv:2405.19722
-
Robust Image Semantic Coding with Learnable CSI Fusion Masking over MIMO Fading Channels 30 May 2024 · 0 repositories · arXiv:2406.07389
-
Sharing Key Semantics in Transformer Makes Efficient Image Restoration 30 May 2024 · 1 repository · arXiv:2405.20008
-
TAIA: Large Language Models are Out-of-Distribution Data Learners 30 May 2024 · 1 repository · arXiv:2405.20192Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Towards RGB-NIR Cross-modality Image Registration and Beyond 30 May 2024 · 0 repositories · arXiv:2405.19914
-
Transformers and Slot Encoding for Sample Efficient Physical World Modelling 30 May 2024 · 1 repository · arXiv:2405.20180
-
Use of a Multiscale Vision Transformer to predict Nursing Activities Score from Low Resolution Thermal Videos in an Intensive Care Unit 30 May 2024 · 0 repositories · arXiv:2406.04364
-
WebUOT-1M: Advancing Deep Underwater Object Tracking with A Million-Scale Benchmark 30 May 2024 · 1 repository · arXiv:2405.19818
-
YotoR-You Only Transform One Representation 30 May 2024 · 0 repositories · arXiv:2405.19629
-
Are You Sure? Rank Them Again: Repeated Ranking For Better Preference Datasets 29 May 2024 · 0 repositories · arXiv:2405.18952
-
Beyond Agreement: Diagnosing the Rationale Alignment of Automated Essay Scoring Methods based on Linguistically-informed Counterfactuals 29 May 2024 · 1 repository · arXiv:2405.19433
-
Can We Enhance the Quality of Mobile Crowdsensing Data Without Ground Truth? 29 May 2024 · 1 repository · arXiv:2405.18725
-
Contextual Position Encoding: Learning to Count What's Important 29 May 2024 · 0 repositories · arXiv:2405.18719
-
Deep Grokking: Would Deep Neural Networks Generalize Better? 29 May 2024 · 0 repositories · arXiv:2405.19454
-
Does learning the right latent variables necessarily improve in-context learning? 29 May 2024 · 1 repository · arXiv:2405.19162Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Enhancing Vision-Language Model with Unmasked Token Alignment 29 May 2024 · 1 repository · arXiv:2405.19009
-
LLM-based Hierarchical Concept Decomposition for Interpretable Fine-Grained Image Classification 29 May 2024 · 0 repositories · arXiv:2405.18672
-
LLMs achieve adult human performance on higher-order theory of mind tasks 29 May 2024 · 0 repositories · arXiv:2405.18870
-
Matryoshka Query Transformer for Large Vision-Language Models 29 May 2024 · 1 repository · arXiv:2405.19315Syntology official (archive's flag): 8 ran · 9 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
MDS-ViTNet: Improving saliency prediction for Eye-Tracking with Vision Transformer 29 May 2024 · 1 repository · arXiv:2405.19501
-
MindSemantix: Deciphering Brain Visual Experiences with a Brain-Language Model 29 May 2024 · 0 repositories · arXiv:2405.18812
-
Multi-Channel Multi-Step Spectrum Prediction Using Transformer and Stacked Bi-LSTM 29 May 2024 · 0 repositories · arXiv:2405.19138
-
PediatricsGPT: Large Language Models as Chinese Medical Assistants for Pediatric Applications 29 May 2024 · 1 repository · arXiv:2405.19266Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Reverse Image Retrieval Cues Parametric Memory in Multimodal LLMs 29 May 2024 · 1 repository · arXiv:2405.18740
-
Transcending Fusion: A Multi-Scale Alignment Method for Remote Sensing Image-Text Retrieval 29 May 2024 · 1 repository · arXiv:2405.18959
-
Two-Layer Retrieval-Augmented Generation Framework for Low-Resource Medical Question Answering Using Reddit Data: Proof-of-Concept Study 29 May 2024 · 0 repositories · arXiv:2405.19519
-
Aligning to Thousands of Preferences via System Message Generalization 28 May 2024 · 1 repository · arXiv:2405.17977Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
An Empirical Analysis on Large Language Models in Debate Evaluation 28 May 2024 · 1 repository · arXiv:2406.00050Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Benchmarks Underestimate the Readiness of Multi-lingual Dialogue Agents 28 May 2024 · 0 repositories · arXiv:2405.17840
-
Delving into Differentially Private Transformer 28 May 2024 · 0 repositories · arXiv:2405.18194
-
Dual-Path Multi-Scale Transformer for High-Quality Image Deraining 28 May 2024 · 0 repositories · arXiv:2405.18124
-
Edinburgh Clinical NLP at MEDIQA-CORR 2024: Guiding Large Language Models with Hints 28 May 2024 · 0 repositories · arXiv:2405.18028
-
FASTopic: Pretrained Transformer is a Fast, Adaptive, Stable, and Transferable Topic Model 28 May 2024 · 2 repositories · arXiv:2405.17978Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
ForecastGrapher: Redefining Multivariate Time Series Forecasting with Graph Neural Networks 28 May 2024 · 0 repositories · arXiv:2405.18036