Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 17
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 17 of 139: papers 1,601 to 1,700 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Rethinking Emotion Annotations in the Era of Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.07906
-
STIV: Scalable Text and Image Conditioned Video Generation 10 Dec 2024 · 0 repositories · arXiv:2412.07730
-
Towards Automated Cross-domain Exploratory Data Analysis through Large Language Models 10 Dec 2024 · 2 repositories · arXiv:2412.07214
-
Anchoring Bias in Large Language Models: An Experimental Study 9 Dec 2024 · 0 repositories · arXiv:2412.06593
-
Bridging the Divide: Reconsidering Softmax and Linear Attention 9 Dec 2024 · 1 repository · arXiv:2412.06590Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 22 harvested samples) · 22 pointer-only (licence)
-
Efficient user history modeling with amortized inference for deep learning recommendation models 9 Dec 2024 · 0 repositories · arXiv:2412.06924
-
EMOv2: Pushing 5M Vision Model Frontier 9 Dec 2024 · 1 repository · arXiv:2412.06674
-
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit 9 Dec 2024 · 0 repositories · arXiv:2412.06370
-
Inverting Transformer-based Vision Models 9 Dec 2024 · 2 repositories · arXiv:2412.06534
-
Normalizing Flows are Capable Generative Models 9 Dec 2024 · 3 repositories · arXiv:2412.06329Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06249
-
S²FT: Efficient, Scalable and Generalizable LLM Fine-tuning by Structured Sparsity 9 Dec 2024 · 0 repositories · arXiv:2412.06289
-
The Computational Limits of State-Space Models and Mamba via the Lens of Circuit Complexity 9 Dec 2024 · 0 repositories · arXiv:2412.06148
-
Toward Non-Invasive Diagnosis of Bankart Lesions with Deep Learning 9 Dec 2024 · 1 repository · arXiv:2412.06717
-
Enhanced Computationally Efficient Long LoRA Inspired Perceiver Architectures for Auto-Regressive Language Modeling 8 Dec 2024 · 0 repositories · arXiv:2412.06106
-
Evaluating Robustness of LLMs on Crisis-Related Microblogs across Events, Information Types, and Linguistic Features 8 Dec 2024 · 0 repositories · arXiv:2412.10413
-
Fully Open Source Moxin-7B Technical Report 8 Dec 2024 · 1 repository · arXiv:2412.06845
-
KITE-DDI: A Knowledge graph Integrated Transformer Model for accurately predicting Drug-Drug Interaction Events from Drug SMILES and Biomedical Knowledge Graph 8 Dec 2024 · 0 repositories · arXiv:2412.05770
-
Language-Guided Image Tokenization for Generation 8 Dec 2024 · 0 repositories · arXiv:2412.05796
-
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor 8 Dec 2024 · 1 repository · arXiv:2412.07801
-
M³-20M: A Large-Scale Multi-Modal Molecule Dataset for AI-driven Drug Design and Discovery 8 Dec 2024 · 1 repository · arXiv:2412.06847
-
Paddy Disease Detection and Classification Using Computer Vision Techniques: A Mobile Application to Detect Paddy Disease 8 Dec 2024 · 0 repositories · arXiv:2412.05996
-
A Comparative Study on Code Generation with Transformers 7 Dec 2024 · 0 repositories · arXiv:2412.05749
-
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking 7 Dec 2024 · 0 repositories · arXiv:2412.05561
-
Innovative Sentiment Analysis and Prediction of Stock Price Using FinBERT, GPT-4 and Logistic Regression: A Data-Driven Approach 7 Dec 2024 · 0 repositories · arXiv:2412.06837
-
M³PC: Test-time Model Predictive Control for Pretrained Masked Trajectory Model 7 Dec 2024 · 1 repository · arXiv:2412.05675Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
RefSAM3D: Adapting SAM with Cross-modal Reference for 3D Medical Image Segmentation 7 Dec 2024 · 0 repositories · arXiv:2412.05605
-
Shifting NER into High Gear: The Auto-AdvER Approach 7 Dec 2024 · 0 repositories · arXiv:2412.05655
-
Towards Learning to Reason: Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning 7 Dec 2024 · 2 repositories · arXiv:2412.05586Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo 6 Dec 2024 · 0 repositories · arXiv:2412.05223
-
Are Frontier Large Language Models Suitable for Q&A in Science Centres? 6 Dec 2024 · 0 repositories · arXiv:2412.05200
-
BEExformer: A Fast Inferencing Transformer Architecture via Binarization with Multiple Early Exits 6 Dec 2024 · 0 repositories · arXiv:2412.05225
-
DHIL-GT: Scalable Graph Transformer with Decoupled Hierarchy Labeling 6 Dec 2024 · 0 repositories · arXiv:2412.04738
-
Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System 6 Dec 2024 · 0 repositories · arXiv:2412.06828
-
Feature Group Tabular Transformer: A Novel Approach to Traffic Crash Modeling and Causality Analysis 6 Dec 2024 · 0 repositories · arXiv:2412.06825
-
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens 6 Dec 2024 · 1 repository · arXiv:2412.04680Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Addressing Hallucinations with RAG and NMISS in Italian Healthcare LLM Chatbots 5 Dec 2024 · 0 repositories · arXiv:2412.04235
-
ARTeFACT: Benchmarking Segmentation Models on Diverse Analogue Media Damage 5 Dec 2024 · 0 repositories · arXiv:2412.04580
-
Cubify Anything: Scaling Indoor 3D Object Detection 5 Dec 2024 · 1 repository · arXiv:2412.04458
-
DEIM: DETR with Improved Matching for Fast Convergence 5 Dec 2024 · 1 repository · arXiv:2412.04234Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Dynamic Graph Representation with Contrastive Learning for Financial Market Prediction: Integrating Temporal Evolution and Static Relations 5 Dec 2024 · 1 repository · arXiv:2412.04034
-
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments? 5 Dec 2024 · 0 repositories · arXiv:2412.03856
-
TransAdapter: Vision Transformer for Feature-Centric Unsupervised Domain Adaptation 5 Dec 2024 · 1 repository · arXiv:2412.04073
-
A Water Efficiency Dataset for African Data Centers 4 Dec 2024 · 0 repositories · arXiv:2412.03716
-
Advanced Risk Prediction and Stability Assessment of Banks Using Time Series Transformer Models 4 Dec 2024 · 0 repositories · arXiv:2412.03606
-
AntLM: Bridging Causal and Masked Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.03275
-
Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts? 4 Dec 2024 · 0 repositories · arXiv:2412.03235
-
EMPATH: MediaPipe-Aided Ensemble Learning with Attention-Based Transformers for Accurate Recognition of Bangla Word-Level Sign Language 4 Dec 2024 · 1 repository
-
Interpreting Transformers for Jet Tagging 4 Dec 2024 · 1 repository · arXiv:2412.03673Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MaterialPicker: Multi-Modal Material Generation with Diffusion Transformers 4 Dec 2024 · 0 repositories · arXiv:2412.03225
-
Multi-Branch Mutual-Distillation Transformer for EEG-Based Seizure Subtype Classification 4 Dec 2024 · 0 repositories · arXiv:2412.15224
-
Navigation World Models 4 Dec 2024 · 1 repository · arXiv:2412.03572Syntology 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Seeing Beyond Views: Multi-View Driving Scene Video Generation with Holistic Attention 4 Dec 2024 · 0 repositories · arXiv:2412.03520
-
Theoretical limitations of multi-layer Transformer 4 Dec 2024 · 1 repository · arXiv:2412.02975
-
FCL-ViT: Task-Aware Attention Tuning for Continual Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02509
-
GQWformer: A Quantum-based Transformer for Graph Representation Learning 3 Dec 2024 · 0 repositories · arXiv:2412.02285
-
Leveraging Large Language Models for Comparative Literature Summarization with Reflective Incremental Mechanisms 3 Dec 2024 · 0 repositories · arXiv:2412.02149
-
MAGMA: Manifold Regularization for MAEs 3 Dec 2024 · 1 repository · arXiv:2412.02871
-
Optimization of Transformer heart disease prediction model based on particle swarm optimization algorithm 3 Dec 2024 · 0 repositories · arXiv:2412.02801
-
Patent-CR: A Dataset for Patent Claim Revision 3 Dec 2024 · 0 repositories · arXiv:2412.02549
-
RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models 3 Dec 2024 · 1 repository · arXiv:2412.02830
-
Revisiting the Initial Steps in Adaptive Gradient Descent Optimization 3 Dec 2024 · 0 repositories · arXiv:2412.02153
-
Transformer-Based Auxiliary Loss for Face Recognition Across Age Variations 3 Dec 2024 · 0 repositories · arXiv:2412.02198
-
Automated Extraction of Acronym-Expansion Pairs from Scientific Papers 2 Dec 2024 · 0 repositories · arXiv:2412.01093
-
Automated Toll Management System Using RFID and Image Processing 2 Dec 2024 · 0 repositories · arXiv:2412.01728
-
Convolutional Transformer Neural Collaborative Filtering 2 Dec 2024 · 0 repositories · arXiv:2412.01376
-
CPA: Camera-pose-awareness Diffusion Transformer for Video Generation 2 Dec 2024 · 0 repositories · arXiv:2412.01429
-
Enhancing Crop Segmentation in Satellite Image Time Series with Transformer Networks 2 Dec 2024 · 0 repositories · arXiv:2412.01944
-
FGATT: A Robust Framework for Wireless Data Imputation Using Fuzzy Graph Attention Networks and Transformer Encoders 2 Dec 2024 · 0 repositories · arXiv:2412.01979
-
GETAE: Graph information Enhanced deep neural NeTwork ensemble ArchitecturE for fake news detection 2 Dec 2024 · 1 repository · arXiv:2412.01825
-
Global Average Feature Augmentation for Robust Semantic Segmentation with Transformers 2 Dec 2024 · 0 repositories · arXiv:2412.01941
-
High-Throughput Detection of Risk Factors to Sudden Cardiac Arrest in Youth Athletes: A Smartwatch-Based Screening Platform 2 Dec 2024 · 0 repositories · arXiv:2412.12118
-
Identifying Reliable Predictions in Detection Transformers 2 Dec 2024 · 0 repositories · arXiv:2412.01782
-
Mutli-View 3D Reconstruction using Knowledge Distillation 2 Dec 2024 · 1 repository · arXiv:2412.02039
-
NYT-Connections: A Deceptively Simple Text Classification Task that Stumps System-1 Thinkers 2 Dec 2024 · 0 repositories · arXiv:2412.01621
-
PKRD-CoT: A Unified Chain-of-thought Prompting for Multi-Modal Large Language Models in Autonomous Driving 2 Dec 2024 · 0 repositories · arXiv:2412.02025
-
R-Bot: An LLM-based Query Rewrite System 2 Dec 2024 · 0 repositories · arXiv:2412.01661
-
ReHub: Linear Complexity Graph Transformers with Adaptive Hub-Spoke Reassignment 2 Dec 2024 · 0 repositories · arXiv:2412.01519
-
The Promise and Peril of Generative AI: Evidence from GPT-4 as Sell-Side Analysts 2 Dec 2024 · 0 repositories · arXiv:2412.01069
-
AniMer: Animal Pose and Shape Estimation Using Family Aware Transformer 1 Dec 2024 · 0 repositories · arXiv:2412.00837
-
Categorical Keypoint Positional Embedding for Robust Animal Re-Identification 1 Dec 2024 · 0 repositories · arXiv:2412.00818
-
Decision Transformer vs. Decision Mamba: Analysing the Complexity of Sequential Decision Making in Atari Games 1 Dec 2024 · 1 repository · arXiv:2412.00725
-
DSSRNN: Decomposition-Enhanced State-Space Recurrent Neural Network for Time-Series Analysis 1 Dec 2024 · 1 repository · arXiv:2412.00994
-
EDTformer: An Efficient Decoder Transformer for Visual Place Recognition 1 Dec 2024 · 1 repository · arXiv:2412.00784Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
MIMIC: Multimodal Islamophobic Meme Identification and Classification 1 Dec 2024 · 1 repository · arXiv:2412.00681
-
Precise Facial Landmark Detection by Dynamic Semantic Aggregation Transformer 1 Dec 2024 · 1 repository · arXiv:2412.00740
-
TGTOD: A Global Temporal Graph Transformer for Outlier Detection at Scale 1 Dec 2024 · 1 repository · arXiv:2412.00984
-
Cognitive Biases in Large Language Models: A Survey and Mitigation Experiments 30 Nov 2024 · 0 repositories · arXiv:2412.00323
-
Dynamic Token Selection for Aerial-Ground Person Re-Identification 30 Nov 2024 · 0 repositories · arXiv:2412.00433
-
Multi-scale Feature Enhancement in Multi-task Learning for Medical Image Analysis 30 Nov 2024 · 0 repositories · arXiv:2412.00351
-
Dynamic ETF Portfolio Optimization Using enhanced Transformer-Based Models for Covariance and Semi-Covariance Prediction(Work in Progress) 29 Nov 2024 · 0 repositories · arXiv:2411.19649
-
Excretion Detection in Pigsties Using Convolutional and Transformerbased Deep Neural Networks 29 Nov 2024 · 0 repositories · arXiv:2412.00256
-
Graph Neural Networks for Heart Failure Prediction on an EHR-Based Patient Similarity Graph 29 Nov 2024 · 1 repository · arXiv:2411.19742
-
HVAC-DPT: A Decision Pretrained Transformer for HVAC Control 29 Nov 2024 · 0 repositories · arXiv:2411.19746
-
LLM Teacher-Student Framework for Text Classification With No Manually Annotated Data: A Case Study in IPTC News Topic Classification 29 Nov 2024 · 1 repository · arXiv:2411.19638
-
Multi-task CNN Behavioral Embedding Model For Transaction Fraud Detection 29 Nov 2024 · 0 repositories · arXiv:2411.19457
-
On Domain-Specific Post-Training for Multimodal Large Language Models 29 Nov 2024 · 0 repositories · arXiv:2411.19930
-
RL-MILP Solver: A Reinforcement Learning Approach for Solving Mixed-Integer Linear Programs with Graph Neural Networks 29 Nov 2024 · 0 repositories · arXiv:2411.19517
-
SAT-HMR: Real-Time Multi-Person 3D Mesh Estimation via Scale-Adaptive Tokens 29 Nov 2024 · 0 repositories · arXiv:2411.19824
-
Training Agents with Weakly Supervised Feedback from Large Language Models 29 Nov 2024 · 0 repositories · arXiv:2411.19547