Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 25
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 25 of 139: papers 2,401 to 2,500 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
GTransPDM: A Graph-embedded Transformer with Positional Decoupling for Pedestrian Crossing Intention Prediction 30 Sep 2024 · 0 repositories · arXiv:2409.20223
-
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation 30 Sep 2024 · 0 repositories · arXiv:2409.19937
-
Exploring Social Media Image Categorization Using Large Models with Different Adaptation Methods: A Case Study on Cultural Nature's Contributions to People 30 Sep 2024 · 0 repositories · arXiv:2410.00275
-
On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability 30 Sep 2024 · 2 repositories · arXiv:2409.19924
-
Adversarial Examples for DNA Classification 29 Sep 2024 · 0 repositories · arXiv:2409.19788
-
Can Models Learn Skill Composition from Examples? 29 Sep 2024 · 0 repositories · arXiv:2409.19808
-
GenTel-Safe: A Unified Benchmark and Shielding Framework for Defending Against Prompt Injection Attacks 29 Sep 2024 · 0 repositories · arXiv:2409.19521
-
DATransNet: Dynamic Attention Transformer Network for Infrared Small Target Detection 29 Sep 2024 · 1 repository · arXiv:2409.19599
-
MedHalu: Hallucinations in Responses to Healthcare Queries by Large Language Models 29 Sep 2024 · 0 repositories · arXiv:2409.19492
-
See then Tell: Enhancing Key Information Extraction with Vision Grounding 29 Sep 2024 · 0 repositories · arXiv:2409.19573
-
Spiking Transformer with Spatial-Temporal Attention 29 Sep 2024 · 1 repository · arXiv:2409.19764
-
DENEB: A Hallucination-Robust Automatic Evaluation Metric for Image Captioning 28 Sep 2024 · 0 repositories · arXiv:2409.19255
-
Multi-Atlas Brain Network Classification through Consistency Distillation and Complementary Information Fusion 28 Sep 2024 · 0 repositories · arXiv:2410.08228
-
Unveil Benign Overfitting for Transformer in Vision: Training Dynamics, Convergence, and Generalization 28 Sep 2024 · 0 repositories · arXiv:2409.19345
-
How Effective is Pre-training of Large Masked Autoencoders for Downstream Earth Observation Tasks? 27 Sep 2024 · 0 repositories · arXiv:2409.18536
-
LML-DAP: Language Model Learning a Dataset for Data-Augmented Prediction 27 Sep 2024 · 1 repository · arXiv:2409.18957
-
Not the Silver Bullet: LLM-enhanced Programming Error Messages are Ineffective in Practice 27 Sep 2024 · 0 repositories · arXiv:2409.18661
-
On the Power of Decision Trees in Auto-Regressive Language Modeling 27 Sep 2024 · 0 repositories · arXiv:2409.19150
-
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs 27 Sep 2024 · 0 repositories · arXiv:2409.18794
-
Pruning then Reweighting: Towards Data-Efficient Training of Diffusion Models 27 Sep 2024 · 1 repository · arXiv:2409.19128
-
Query matching for spatio-temporal action detection with query-based object detector 27 Sep 2024 · 0 repositories · arXiv:2409.18408
-
Speech-Mamba: Long-Context Speech Recognition with Selective State Spaces Models 27 Sep 2024 · 0 repositories · arXiv:2409.18654
-
A Fuzzy-based Approach to Predict Human Interaction by Functional Near-Infrared Spectroscopy 26 Sep 2024 · 0 repositories · arXiv:2409.17661
-
AgMTR: Agent Mining Transformer for Few-shot Segmentation in Remote Sensing 26 Sep 2024 · 1 repository · arXiv:2409.17453
-
CASPFormer: Trajectory Prediction from BEV Images with Deformable Attention 26 Sep 2024 · 0 repositories · arXiv:2409.17790
-
DARE: Diverse Visual Question Answering with Robustness Evaluation 26 Sep 2024 · 0 repositories · arXiv:2409.18023
-
Developing a Dual-Stage Vision Transformer Model for Lung Disease Classification 26 Sep 2024 · 0 repositories · arXiv:2409.18257
-
Dynamic Subframe Splitting and Spatio-Temporal Motion Entangled Sparse Attention for RGB-E Tracking 26 Sep 2024 · 0 repositories · arXiv:2409.17560
-
EM-Net: Efficient Channel and Frequency Learning with Mamba for 3D Medical Image Segmentation 26 Sep 2024 · 1 repository · arXiv:2409.17675Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Just Say What You Want: Only-prompting Self-rewarding Online Preference Optimization 26 Sep 2024 · 0 repositories · arXiv:2409.17534
-
NeuroPath: A Neural Pathway Transformer for Joining the Dots of Human Connectomes 26 Sep 2024 · 1 repository · arXiv:2409.17510Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Ophthalmic Biomarker Detection with Parallel Prediction of Transformer and Convolutional Architecture 26 Sep 2024 · 0 repositories · arXiv:2409.17788
-
PEDRO: Parameter-Efficient Fine-tuning with Prompt DEpenDent Representation MOdification 26 Sep 2024 · 0 repositories · arXiv:2409.17834
-
Predicting Anchored Text from Translation Memories for Machine Translation Using Deep Learning Methods 26 Sep 2024 · 0 repositories · arXiv:2409.17939
-
Retrospective Comparative Analysis of Prostate Cancer In-Basket Messages: Responses from Closed-Domain LLM vs. Clinical Teams 26 Sep 2024 · 1 repository · arXiv:2409.18290
-
Self-supervised Monocular Depth Estimation with Large Kernel Attention 26 Sep 2024 · 0 repositories · arXiv:2409.17895
-
The application of GPT-4 in grading design university students' assignment and providing feedback: An exploratory study 26 Sep 2024 · 0 repositories · arXiv:2409.17698
-
Unifying Dimensions: A Linear Adaptive Approach to Lightweight Image Super-Resolution 26 Sep 2024 · 1 repository · arXiv:2409.17597
-
Beyond Turing Test: Can GPT-4 Sway Experts' Decisions? 25 Sep 2024 · 0 repositories · arXiv:2409.16710
-
CodeInsight: A Curated Dataset of Practical Coding Solutions from Stack Overflow 25 Sep 2024 · 1 repository · arXiv:2409.16819
-
Enhancing Automatic Keyphrase Labelling with Text-to-Text Transfer Transformer (T5) Architecture: A Framework for Keyphrase Generation and Filtering 25 Sep 2024 · 0 repositories · arXiv:2409.16760
-
Going Beyond U-Net: Assessing Vision Transformers for Semantic Segmentation in Microscopy Image Analysis 25 Sep 2024 · 0 repositories · arXiv:2409.16940
-
Gradient Boosting Decision Trees on Medical Diagnosis over Tabular Data 25 Sep 2024 · 1 repository · arXiv:2410.03705
-
HVT: A Comprehensive Vision Framework for Learning in Non-Euclidean Space 25 Sep 2024 · 1 repository · arXiv:2409.16897
-
Investigating OCR-Sensitive Neurons to Improve Entity Recognition in Historical Documents 25 Sep 2024 · 1 repository · arXiv:2409.16934
-
Post-hoc Reward Calibration: A Case Study on Length Bias 25 Sep 2024 · 1 repository · arXiv:2409.17407Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Pre-trained Graphformer-based Ranking at Web-scale Search (Extended Abstract) 25 Sep 2024 · 0 repositories · arXiv:2409.16590
-
Quantum-Classical Sentiment Analysis 25 Sep 2024 · 0 repositories · arXiv:2409.16928
-
The Credibility Transformer 25 Sep 2024 · 0 repositories · arXiv:2409.16653
-
Trading through Earnings Seasons using Self-Supervised Contrastive Representation Learning 25 Sep 2024 · 0 repositories · arXiv:2409.17392
-
A Comprehensive Evaluation of Large Language Models on Mental Illnesses 24 Sep 2024 · 0 repositories · arXiv:2409.15687
-
AI Can Be Cognitively Biased: An Exploratory Study on Threshold Priming in LLM-Based Batch Relevance Assessment 24 Sep 2024 · 0 repositories · arXiv:2409.16022
-
dnaGrinder: a lightweight and high-capacity genomic foundation model 24 Sep 2024 · 0 repositories · arXiv:2409.15697
-
Double-Path Adaptive-correlation Spatial-Temporal Inverted Transformer for Stock Time Series Forecasting 24 Sep 2024 · 0 repositories · arXiv:2409.15662
-
Language-based Audio Moment Retrieval 24 Sep 2024 · 1 repository · arXiv:2409.15672
-
Looped Transformers for Length Generalization 24 Sep 2024 · 1 repository · arXiv:2409.15647Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
MaskBit: Embedding-free Image Generation via Bit Tokens 24 Sep 2024 · 1 repository · arXiv:2409.16211Syntology official (archive's flag): 8 ran · 8 ran (of which 1 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
MonoFormer: One Transformer for Both Diffusion and Autoregression 24 Sep 2024 · 1 repository · arXiv:2409.16280Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Predicting Deterioration in Mild Cognitive Impairment with Survival Transformers, Extreme Gradient Boosting and Cox Proportional Hazard Modelling 24 Sep 2024 · 0 repositories · arXiv:2409.16231
-
Task-oriented Prompt Enhancement via Script Generation 24 Sep 2024 · 0 repositories · arXiv:2409.16418
-
TiM4Rec: An Efficient Sequential Recommendation Model Based on Time-Aware Structured State Space Duality Model 24 Sep 2024 · 1 repository · arXiv:2409.16182
-
A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor? 23 Sep 2024 · 0 repositories · arXiv:2409.15277
-
DepthART: Monocular Depth Estimation as Autoregressive Refinement Task 23 Sep 2024 · 0 repositories · arXiv:2409.15010
-
Designing Pre-training Datasets from Unlabeled Data for EEG Classification with Transformers 23 Sep 2024 · 0 repositories · arXiv:2410.07190
-
Diffusion-based RGB-D Semantic Segmentation with Deformable Attention Transformer 23 Sep 2024 · 0 repositories · arXiv:2409.15117
-
Dual Stream Graph Transformer Fusion Networks for Enhanced Brain Decoding 23 Sep 2024 · 0 repositories · arXiv:2410.07189
-
EDGE-Rec: Efficient and Data-Guided Edge Diffusion For Recommender Systems Graphs 23 Sep 2024 · 0 repositories · arXiv:2409.14689
-
PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs 23 Sep 2024 · 1 repository · arXiv:2409.14866Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Generalizing monocular colonoscopy image depth estimation by uncertainty-based global and local fusion network 23 Sep 2024 · 0 repositories · arXiv:2409.15006
-
HydroVision: LiDAR-Guided Hydrometric Prediction with Vision Transformers and Hybrid Graph Learning 23 Sep 2024 · 0 repositories · arXiv:2409.15213
-
Kriformer: A Novel Spatiotemporal Kriging Approach Based on Graph Transformers 23 Sep 2024 · 0 repositories · arXiv:2409.14906
-
M2OST: Many-to-one Regression for Predicting Spatial Transcriptomics from Digital Pathology Images 23 Sep 2024 · 1 repository · arXiv:2409.15092Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
MemeCLIP: Leveraging CLIP Representations for Multimodal Meme Classification 23 Sep 2024 · 1 repository · arXiv:2409.14703
-
Micrometer: Micromechanics Transformer for Predicting Mechanical Responses of Heterogeneous Materials 23 Sep 2024 · 0 repositories · arXiv:2410.05281
-
PALLM: Evaluating and Enhancing PALLiative Care Conversations with Large Language Models 23 Sep 2024 · 1 repository · arXiv:2409.15188
-
RACER: Rich Language-Guided Failure Recovery Policies for Imitation Learning 23 Sep 2024 · 0 repositories · arXiv:2409.14674
-
RoWSFormer: A Robust Watermarking Framework with Swin Transformer for Enhanced Geometric Attack Resilience 23 Sep 2024 · 0 repositories · arXiv:2409.14829
-
Scaling Laws of Decoder-Only Models on the Multilingual Machine Translation Task 23 Sep 2024 · 0 repositories · arXiv:2409.15051
-
TransUKAN:Computing-Efficient Hybrid KAN-Transformer for Enhanced Medical Image Segmentation 23 Sep 2024 · 0 repositories · arXiv:2409.14676
-
Beyond Words: Evaluating Large Language Models in Transportation Planning 22 Sep 2024 · 0 repositories · arXiv:2409.14516
-
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks 22 Sep 2024 · 0 repositories · arXiv:2409.14488
-
Evaluating the Quality of Code Comments Generated by Large Language Models for Novice Programmers 22 Sep 2024 · 0 repositories · arXiv:2409.14368
-
Large Model Based Agents: State-of-the-Art, Cooperation Paradigms, Security and Privacy, and Future Trends 22 Sep 2024 · 0 repositories · arXiv:2409.14457
-
More Effective LLM Compressed Tokens with Uniformly Spread Position Identifiers and Compression Loss 22 Sep 2024 · 0 repositories · arXiv:2409.14364
-
Patch Ranking: Efficient CLIP by Learning to Rank Local Patches 22 Sep 2024 · 1 repository · arXiv:2409.14607
-
Sparse Low-Ranked Self-Attention Transformer for Remaining Useful Lifetime Prediction of Optical Fiber Amplifiers 22 Sep 2024 · 0 repositories · arXiv:2409.14378
-
ChemEval: A Comprehensive Multi-Level Chemical Evaluation for Large Language Models 21 Sep 2024 · 1 repository · arXiv:2409.13989Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 0 violated, 17 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators 21 Sep 2024 · 1 repository · arXiv:2409.14037
-
AI Assistants for Spaceflight Procedures: Combining Generative Pre-Trained Transformer and Retrieval-Augmented Generation on Knowledge Graphs With Augmented Reality Cues 21 Sep 2024 · 0 repositories · arXiv:2409.14206
-
Developing a Thailand solar irradiance map using Himawari-8 satellite imageries and deep learning models 21 Sep 2024 · 1 repository · arXiv:2409.16320
-
Window-based Channel Attention for Wavelet-enhanced Learned Image Compression 21 Sep 2024 · 0 repositories · arXiv:2409.14090
-
A Personalised 3D+t Mesh Generative Model for Unveiling Normal Heart Dynamics 20 Sep 2024 · 1 repository · arXiv:2409.13825
-
Aligning Language Models Using Follow-up Likelihood as Reward Signal 20 Sep 2024 · 1 repository · arXiv:2409.13948
-
AVG-LLaVA: A Large Multimodal Model with Adaptive Visual Granularity 20 Sep 2024 · 1 repository · arXiv:2410.02745Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Localized Gaussians as Self-Attention Weights for Point Clouds Correspondence 20 Sep 2024 · 0 repositories · arXiv:2409.13291
-
Prompting Large Language Models for Supporting the Differential Diagnosis of Anemia 20 Sep 2024 · 0 repositories · arXiv:2409.15377
-
ShizishanGPT: An Agricultural Large Language Model Integrating Tools and Resources 20 Sep 2024 · 1 repository · arXiv:2409.13537
-
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions 20 Sep 2024 · 1 repository · arXiv:2409.13843
-
Tackling fluffy clouds: field boundaries detection using time series of S2 and/or S1 imagery 20 Sep 2024 · 1 repository · arXiv:2409.13568
-
ViTGuard: Attention-aware Detection against Adversarial Examples for Vision Transformer 20 Sep 2024 · 0 repositories · arXiv:2409.13828