Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 30
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 30 of 139: papers 2,901 to 3,000 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
PAFormer: Part Aware Transformer for Person Re-identification 12 Aug 2024 · 0 repositories · arXiv:2408.05918
-
PhaGO: Protein function annotation for bacteriophages by integrating the genomic context 12 Aug 2024 · 0 repositories · arXiv:2408.06402
-
Spacetime E(n)-Transformer: Equivariant Attention for Spatio-temporal Graphs 12 Aug 2024 · 1 repository · arXiv:2408.06039
-
The Language of Trauma: Modeling Traumatic Event Descriptions Across Domains with Explainable AI 12 Aug 2024 · 0 repositories · arXiv:2408.05977
-
Utilize Transformers for translating Wikipedia category names 12 Aug 2024 · 0 repositories · arXiv:2408.06124
-
GPT-4 Emulates Average-Human Emotional Cognition from a Third-Person Perspective 11 Aug 2024 · 0 repositories · arXiv:2408.13718
-
HySparK: Hybrid Sparse Masking for Large Scale Medical Image Pre-Training 11 Aug 2024 · 1 repository · arXiv:2408.05815
-
Sampling Foundational Transformer: A Theoretical Perspective 11 Aug 2024 · 0 repositories · arXiv:2408.05822
-
U-DECN: End-to-End Underwater Object Detection ConvNet with Improved DeNoising Training 11 Aug 2024 · 1 repository · arXiv:2408.05780
-
BeyondCT: A deep learning model for predicting pulmonary function from chest CT scans 10 Aug 2024 · 0 repositories · arXiv:2408.05645
-
Chain of Condition: Construct, Verify and Solve Conditions for Conditional Question Answering 10 Aug 2024 · 0 repositories · arXiv:2408.05442
-
Modeling Multi-Step Scientific Processes with Graph Transformer Networks 10 Aug 2024 · 0 repositories · arXiv:2408.05425
-
PersonViT: Large-scale Self-supervised Vision Transformer for Person Re-Identification 10 Aug 2024 · 1 repository · arXiv:2408.05398
-
PointMT: Efficient Point Cloud Analysis with Hybrid MLP-Transformer Architecture 10 Aug 2024 · 0 repositories · arXiv:2408.05508
-
SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning 10 Aug 2024 · 3 repositories · arXiv:2408.05517Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ChatGPT Meets Iris Biometrics 9 Aug 2024 · 0 repositories · arXiv:2408.04868
-
DeepInteraction++: Multi-Modality Interaction for Autonomous Driving 9 Aug 2024 · 1 repository · arXiv:2408.05075Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Evaluating the capability of large language models to personalize science texts for diverse middle-school-age learners 9 Aug 2024 · 0 repositories · arXiv:2408.05204
-
From Text to Insight: Leveraging Large Language Models for Performance Evaluation in Management 9 Aug 2024 · 0 repositories · arXiv:2408.05328
-
Large Language Models and Thematic Analysis: Human-AI Synergy in Researching Hate Speech on Social Media 9 Aug 2024 · 0 repositories · arXiv:2408.05126
-
LLMJudge: LLMs for Relevance Judgments 9 Aug 2024 · 1 repository · arXiv:2408.08896
-
MIDI-to-Tab: Guitar Tablature Inference via Masked Language Modeling 9 Aug 2024 · 0 repositories · arXiv:2408.05024
-
Attention Mechanism and Context Modeling System for Text Mining Machine Translation 8 Aug 2024 · 0 repositories · arXiv:2408.04216
-
Can GPT-4 Models Detect Misleading Visualizations? 8 Aug 2024 · 0 repositories · arXiv:2408.12617
-
M2EF-NNs: Multimodal Multi-instance Evidence Fusion Neural Networks for Cancer Survival Prediction 8 Aug 2024 · 0 repositories · arXiv:2408.04170
-
Multi-Turn Context Jailbreak Attack on Large Language Models From First Principles 8 Aug 2024 · 0 repositories · arXiv:2408.04686
-
Scalable Transformer for High Dimensional Multivariate Time Series Forecasting 8 Aug 2024 · 1 repository · arXiv:2408.04245
-
SCENE: Evaluating Explainable AI Techniques Using Soft Counterfactuals 8 Aug 2024 · 0 repositories · arXiv:2408.04575
-
Survey: Transformer-based Models in Data Modality Conversion 8 Aug 2024 · 0 repositories · arXiv:2408.04723
-
Towards Explainable Network Intrusion Detection using Large Language Models 8 Aug 2024 · 0 repositories · arXiv:2408.04342
-
Towards Resilient and Efficient LLMs: A Comparative Study of Efficiency, Performance, and Adversarial Robustness 8 Aug 2024 · 0 repositories · arXiv:2408.04585
-
Transformer Explainer: Interactive Learning of Text-Generative Models 8 Aug 2024 · 1 repository · arXiv:2408.04619
-
UHNet: An Ultra-Lightweight and High-Speed Edge Detection Network 8 Aug 2024 · 0 repositories · arXiv:2408.04258
-
A Comparison of LLM Finetuning Methods & Evaluation Metrics with Travel Chatbot Use Case 7 Aug 2024 · 0 repositories · arXiv:2408.03562
-
Bi-Level Spatial and Channel-aware Transformer for Learned Image Compression 7 Aug 2024 · 0 repositories · arXiv:2408.03842
-
Can Rule-Based Insights Enhance LLMs for Radiology Report Classification? Introducing the RadPrompt Methodology 7 Aug 2024 · 0 repositories · arXiv:2408.04121
-
Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants 7 Aug 2024 · 0 repositories · arXiv:2408.11841
-
Early Prediction of Causes (not Effects) in Healthcare by Long-Term Clinical Time Series Forecasting 7 Aug 2024 · 1 repository · arXiv:2408.03816
-
FMiFood: Multi-modal Contrastive Learning for Food Image Classification 7 Aug 2024 · 0 repositories · arXiv:2408.03922
-
No-Reference Image Quality Assessment with Global-Local Progressive Integration and Semantic-Aligned Quality Transfer 7 Aug 2024 · 1 repository · arXiv:2408.03885
-
Image-to-LaTeX Converter for Mathematical Formulas and Text 7 Aug 2024 · 1 repository · arXiv:2408.04015
-
Inter-Series Transformer: Attending to Products in Time Series Forecasting 7 Aug 2024 · 0 repositories · arXiv:2408.03872
-
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling 7 Aug 2024 · 0 repositories · arXiv:2408.03612
-
PackMamba: Efficient Processing of Variable-Length Sequences in Mamba training 7 Aug 2024 · 0 repositories · arXiv:2408.03865
-
PaveCap: The First Multimodal Framework for Comprehensive Pavement Condition Assessment with Dense Captioning and PCI Estimation 7 Aug 2024 · 1 repository · arXiv:2408.04110
-
RailTrack-DaViT: A Vision Transformer-Based Approach for Automated Railway Track Defect Detection 7 Aug 2024 · 1 repository
-
Surgformer: Surgical Transformer with Hierarchical Temporal Attention for Surgical Phase Recognition 7 Aug 2024 · 1 repository · arXiv:2408.03867
-
SwinShadow: Shifted Window for Ambiguous Adjacent Shadow Detection 7 Aug 2024 · 1 repository · arXiv:2408.03521
-
Intermediate direct preference optimization 6 Aug 2024 · 0 repositories · arXiv:2408.02923
-
Analysis of Argument Structure Constructions in a Deep Recurrent Language Model 6 Aug 2024 · 0 repositories · arXiv:2408.03062
-
ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer 6 Aug 2024 · 0 repositories · arXiv:2408.03284
-
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation 6 Aug 2024 · 0 repositories · arXiv:2408.03312
-
Advancing EEG-Based Gaze Prediction Using Depthwise Separable Convolution and Enhanced Pre-Processing 6 Aug 2024 · 1 repository · arXiv:2408.03480
-
Can LLMs Serve As Time Series Anomaly Detectors? 6 Aug 2024 · 0 repositories · arXiv:2408.03475
-
LLM-Aided Compilation for Tensor Accelerators 6 Aug 2024 · 0 repositories · arXiv:2408.03408
-
LLM-based MOFs Synthesis Condition Extraction using Few-Shot Demonstrations 6 Aug 2024 · 0 repositories · arXiv:2408.04665
-
Set2Seq Transformer: Learning Permutation Aware Set Representations of Artistic Sequences 6 Aug 2024 · 0 repositories · arXiv:2408.03404
-
TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement 6 Aug 2024 · 1 repository · arXiv:2408.03440
-
AssemAI: Interpretable Image-Based Anomaly Detection for Manufacturing Pipelines 5 Aug 2024 · 1 repository · arXiv:2408.02181
-
Is Large Language Model Good at Database Knob Tuning? A Comprehensive Experimental Evaluation 5 Aug 2024 · 0 repositories · arXiv:2408.02213
-
Cross-modulated Attention Transformer for RGBT Tracking 5 Aug 2024 · 0 repositories · arXiv:2408.02222
-
Do Large Language Models Speak All Languages Equally? A Comparative Study in Low-Resource Settings 5 Aug 2024 · 0 repositories · arXiv:2408.02237
-
DRFormer: Multi-Scale Transformer Utilizing Diverse Receptive Fields for Long Time-Series Forecasting 5 Aug 2024 · 1 repository · arXiv:2408.02279
-
The NPU-ASLP System Description for Visual Speech Recognition in CNVSRC 2024 5 Aug 2024 · 1 repository · arXiv:2408.02369
-
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02416
-
On Using Quasirandom Sequences in Machine Learning for Model Weight Initialization 5 Aug 2024 · 1 repository · arXiv:2408.02654
-
Dimensionality Reduction and Nearest Neighbors for Improving Out-of-Distribution Detection in Medical Image Segmentation 5 Aug 2024 · 1 repository · arXiv:2408.02761
-
SEAS: Self-Evolving Adversarial Safety Optimization for Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02632
-
Self-Taught Evaluators 5 Aug 2024 · 0 repositories · arXiv:2408.02666
-
Toward Attention-based TinyML: A Heterogeneous Accelerated Architecture and Automated Deployment Flow 5 Aug 2024 · 1 repository · arXiv:2408.02473
-
DiReCT: Diagnostic Reasoning for Clinical Notes via Large Language Models 4 Aug 2024 · 1 repository · arXiv:2408.01933Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
MedSyn: LLM-based Synthetic Medical Text Generation Framework 4 Aug 2024 · 1 repository · arXiv:2408.02056
-
KAN-RCBEVDepth: A multi-modal fusion algorithm in object detection for autonomous driving 4 Aug 2024 · 0 repositories · arXiv:2408.02088
-
Advancing Mental Health Pre-Screening: A New Custom GPT for Psychological Distress Assessment 3 Aug 2024 · 0 repositories · arXiv:2408.01614
-
Self-Emotion Blended Dialogue Generation in Social Simulation Agents 3 Aug 2024 · 0 repositories · arXiv:2408.01633
-
Stimulating Imagination: Towards General-purpose Object Rearrangement 3 Aug 2024 · 0 repositories · arXiv:2408.01655
-
A Novel Evaluation Framework for Image2Text Generation 3 Aug 2024 · 0 repositories · arXiv:2408.01723
-
LAM3D: Leveraging Attention for Monocular 3D Object Detection 3 Aug 2024 · 0 repositories · arXiv:2408.01739
-
GLDiTalker: Speech-Driven 3D Facial Animation with Graph Latent Diffusion Transformer 3 Aug 2024 · 0 repositories · arXiv:2408.01826
-
MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance 3 Aug 2024 · 1 repository · arXiv:2408.01869
-
Distinguishing Chatbot from Human 3 Aug 2024 · 0 repositories · arXiv:2408.04647
-
JambaTalk: Speech-Driven 3D Talking Head Generation Based on Hybrid Transformer-Mamba Language Model 3 Aug 2024 · 0 repositories · arXiv:2408.01627
-
POA: Pre-training Once for Models of All Sizes 2 Aug 2024 · 1 repository · arXiv:2408.01031
-
MambaST: A Plug-and-Play Cross-Spectral Spatial-Temporal Fuser for Efficient Pedestrian Detection 2 Aug 2024 · 1 repository · arXiv:2408.01037
-
LLM as Runtime Error Handler: A Promising Pathway to Adaptive Self-Healing of Software Systems 2 Aug 2024 · 0 repositories · arXiv:2408.01055
-
Leveraging Encoder-only Large Language Models for Mobile App Review Feature Extraction 2 Aug 2024 · 1 repository · arXiv:2408.01063
-
An Efficient and Effective Transformer Decoder-Based Framework for Multi-Task Visual Grounding 2 Aug 2024 · 1 repository · arXiv:2408.01120Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
A Survey of Mamba 2 Aug 2024 · 0 repositories · arXiv:2408.01129
-
Rethinking Pre-Trained Feature Extractor Selection in Multiple Instance Learning for Whole Slide Image Classification 2 Aug 2024 · 1 repository · arXiv:2408.01167
-
Nested Music Transformer: Sequentially Decoding Compound Tokens in Symbolic Music and Audio Generation 2 Aug 2024 · 1 repository · arXiv:2408.01180
-
High-Throughput Phenotyping of Clinical Text Using Large Language Models 2 Aug 2024 · 0 repositories · arXiv:2408.01214
-
Multi-head Spatial-Spectral Mamba for Hyperspectral Image Classification 2 Aug 2024 · 1 repository · arXiv:2408.01224
-
HeteroMorpheus: Universal Control Based on Morphological Heterogeneity Modeling 2 Aug 2024 · 1 repository · arXiv:2408.01230
-
WaveMamba: Spatial-Spectral Wavelet Mamba for Hyperspectral Image Classification 2 Aug 2024 · 0 repositories · arXiv:2408.01231
-
Underwater Object Detection Enhancement via Channel Stabilization 2 Aug 2024 · 1 repository · arXiv:2408.01293
-
Spatial and Spatial-Spectral Morphological Mamba for Hyperspectral Image Classification 2 Aug 2024 · 2 repositories · arXiv:2408.01372
-
NOLO: Navigate Only Look Once 2 Aug 2024 · 0 repositories · arXiv:2408.01384
-
Pre-trained Language Models Improve the Few-shot Prompt Ability of Decision Transformer 2 Aug 2024 · 0 repositories · arXiv:2408.01402
-
THOR2: Topological Analysis for 3D Shape and Color-Based Human-Inspired Object Recognition in Unseen Environments 2 Aug 2024 · 1 repository · arXiv:2408.01579
-
Enhanced Structured State Space Models via Grouped FIR Filtering and Attention Sink Mechanisms 1 Aug 2024 · 0 repositories · arXiv:2408.00244