Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 60
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 60 of 139: papers 5,901 to 6,000 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Research on Multilingual Natural Scene Text Detection Algorithm 18 Dec 2023 · 0 repositories · arXiv:2312.11153
-
Stronger Graph Transformer with Regularized Attention Scores 18 Dec 2023 · 1 repository · arXiv:2312.11730
-
Time-Transformer: Integrating Local and Global Features for Better Time Series Generation 18 Dec 2023 · 1 repository · arXiv:2312.11714
-
Unleashing the Power of CNN and Transformer for Balanced RGB-Event Video Recognition 18 Dec 2023 · 1 repository · arXiv:2312.11128
-
An Evaluation of GPT-4V and Gemini in Online VQA 17 Dec 2023 · 0 repositories · arXiv:2312.10637
-
AutoVisual Fusion Suite: A Comprehensive Evaluation of Image Segmentation and Voice Conversion Tools on HuggingFace Platform 17 Dec 2023 · 1 repository · arXiv:2401.05379
-
CEIR: Concept-based Explainable Image Representation Learning 17 Dec 2023 · 0 repositories · arXiv:2312.10747
-
DER-GCN: Dialogue and Event Relation-Aware Graph Convolutional Neural Network for Multimodal Dialogue Emotion Recognition 17 Dec 2023 · 0 repositories · arXiv:2312.10579
-
Multi-level Reasoning for Robotic Assembly: From Sequence Inference to Contact Selection 17 Dec 2023 · 0 repositories · arXiv:2312.10571
-
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion 17 Dec 2023 · 2 repositories · arXiv:2312.10692
-
T2M-HiFiGPT: Generating High Quality Human Motion from Textual Descriptions with Residual Discrete Representations 17 Dec 2023 · 0 repositories · arXiv:2312.10628
-
Towards Compact 3D Representations via Point Feature Enhancement Masked Autoencoders 17 Dec 2023 · 1 repository · arXiv:2312.10726
-
A Comparative Analysis of Large Language Models for Code Documentation Generation 16 Dec 2023 · 0 repositories · arXiv:2312.10349
-
An Attentive Inductive Bias for Sequential Recommendation beyond the Self-Attention 16 Dec 2023 · 2 repositories · arXiv:2312.10325Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
DeepArt: A Benchmark to Advance Fidelity Research in AI-Generated Content 16 Dec 2023 · 0 repositories · arXiv:2312.10407
-
RecPrompt: A Self-tuning Prompting Framework for News Recommendation Using Large Language Models 16 Dec 2023 · 1 repository · arXiv:2312.10463
-
ResoNet: Robust and Explainable ENSO Forecasts with Hybrid Convolution and Transformer Networks 16 Dec 2023 · 0 repositories · arXiv:2312.10429
-
Self-Supervised Disentangled Representation Learning for Robust Target Speech Extraction 16 Dec 2023 · 0 repositories · arXiv:2312.10305
-
SPT: Fine-Tuning Transformer-based Language Models Efficiently with Sparsification 16 Dec 2023 · 1 repository · arXiv:2312.10365
-
A Case Study of Image Enhancement Algorithms' Effectiveness of Improving Neural Networks' Performance on Adverse Images 15 Dec 2023 · 0 repositories · arXiv:2312.09509
-
Accelerating Neural Network Training: A Brief Review 15 Dec 2023 · 1 repository · arXiv:2312.10024
-
Beyond Empirical Windowing: An Attention-Based Approach for Trust Prediction in Autonomous Vehicles 15 Dec 2023 · 0 repositories · arXiv:2312.10209
-
Binary Code Summarization: Benchmarking ChatGPT/GPT-4 and Other Large Language Models 15 Dec 2023 · 1 repository · arXiv:2312.09601
-
Distilling Large Language Models for Matching Patients to Clinical Trials 15 Dec 2023 · 0 repositories · arXiv:2312.09958
-
Lever LM: Configuring In-Context Sequence to Lever Large Vision Language Models 15 Dec 2023 · 2 repositories · arXiv:2312.10104Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Integrating AI and Learning Analytics for Data-Driven Pedagogical Decisions and Personalized Interventions in Education 15 Dec 2023 · 0 repositories · arXiv:2312.09548
-
OpenMedCalc: Augmentation of ChatGPT with Clinician-Informed Tools Improves Performance on Medical Calculation Tasks 15 Dec 2023 · 1 repository
-
Part Representation Learning with Teacher-Student Decoder for Occluded Person Re-identification 15 Dec 2023 · 1 repository · arXiv:2312.09797
-
Point Transformer V3: Simpler, Faster, Stronger 15 Dec 2023 · 3 repositories · arXiv:2312.10035Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
Acoustic models of Brazilian Portuguese Speech based on Neural Transformers 14 Dec 2023 · 0 repositories · arXiv:2312.09265
-
Auto-Prox: Training-Free Vision Transformer Architecture Search via Automatic Proxy Discovery 14 Dec 2023 · 1 repository · arXiv:2312.09059Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 4 pointer-only (licence)
-
BiPFT: Binary Pre-trained Foundation Transformer with Low-rank Estimation of Binarization Residual Polynomials 14 Dec 2023 · 1 repository · arXiv:2312.08937
-
Heterogeneous Graph Neural Architecture Search with GPT-4 14 Dec 2023 · 1 repository · arXiv:2312.08680
-
Holodeck: Language Guided Generation of 3D Embodied AI Environments 14 Dec 2023 · 1 repository · arXiv:2312.09067Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Learning Long Sequences in Spiking Neural Networks 14 Dec 2023 · 0 repositories · arXiv:2401.00955
-
Modeling Complex Mathematical Reasoning via Large Language Model based MathAgent 14 Dec 2023 · 1 repository · arXiv:2312.08926
-
NeXt-TDNN: Modernizing Multi-Scale Temporal Convolution Backbone for Speaker Verification 14 Dec 2023 · 1 repository · arXiv:2312.08603
-
Polyper: Boundary Sensitive Polyp Segmentation 14 Dec 2023 · 1 repository · arXiv:2312.08735
-
VL-GPT: A Generative Pre-trained Transformer for Vision and Language Understanding and Generation 14 Dec 2023 · 1 repository · arXiv:2312.09251
-
VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning 14 Dec 2023 · 1 repository · arXiv:2312.08774
-
Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision 14 Dec 2023 · 0 repositories · arXiv:2312.09390
-
Zebra: Extending Context Window with Layerwise Grouped Local-Global Attention 14 Dec 2023 · 0 repositories · arXiv:2312.08618
-
Assessing GPT4-V on Structured Reasoning Tasks 13 Dec 2023 · 0 repositories · arXiv:2312.11524
-
Beyond English: Evaluating LLMs for Arabic Grammatical Error Correction 13 Dec 2023 · 0 repositories · arXiv:2312.08400
-
Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers 13 Dec 2023 · 2 repositories · arXiv:2312.08168Syntology official (archive's flag): 6 ran · 11 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 5 pointer-only (licence)
-
CoIE: Chain-of-Instruct Editing for Multi-Attribute Face Manipulation 13 Dec 2023 · 0 repositories · arXiv:2312.07879
-
Enhancing CT Image synthesis from multi-modal MRI data based on a multi-task neural network framework 13 Dec 2023 · 0 repositories · arXiv:2312.08343
-
High-throughput Biomedical Relation Extraction for Semi-Structured Web Articles Empowered by Large Language Models 13 Dec 2023 · 0 repositories · arXiv:2312.08274
-
Fine-grained Graph Rationalization 13 Dec 2023 · 0 repositories · arXiv:2312.07859
-
Large Language Models are Complex Table Parsers 13 Dec 2023 · 0 repositories · arXiv:2312.11521
-
Mono3DVG: 3D Visual Grounding in Monocular Images 13 Dec 2023 · 1 repository · arXiv:2312.08022Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
N-Gram Unsupervised Compoundation and Feature Injection for Better Symbolic Music Understanding 13 Dec 2023 · 2 repositories · arXiv:2312.08931
-
Native Language Identification with Large Language Models 13 Dec 2023 · 0 repositories · arXiv:2312.07819
-
Prompt Engineering-assisted Malware Dynamic Analysis Using GPT-4 13 Dec 2023 · 1 repository · arXiv:2312.08317
-
SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention 13 Dec 2023 · 2 repositories · arXiv:2312.07987Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Prompted Contextual Transformer for Incomplete-View CT Reconstruction 13 Dec 2023 · 1 repository · arXiv:2312.07846
-
AI Control: Improving Safety Despite Intentional Subversion 12 Dec 2023 · 1 repository · arXiv:2312.06942Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Benchmarking Deep Learning Classifiers for SAR Automatic Target Recognition 12 Dec 2023 · 0 repositories · arXiv:2312.06940
-
Can a Transformer Represent a Kalman Filter? 12 Dec 2023 · 0 repositories · arXiv:2312.06937
-
Context Matters: Data-Efficient Augmentation of Large Language Models for Scientific Applications 12 Dec 2023 · 2 repositories · arXiv:2312.07069
-
Exploring Large Language Models to Facilitate Variable Autonomy for Human-Robot Teaming 12 Dec 2023 · 0 repositories · arXiv:2312.07214
-
Exploring Plain ViT Reconstruction for Multi-class Unsupervised Anomaly Detection 12 Dec 2023 · 1 repository · arXiv:2312.07495
-
Hyper-Restormer: A General Hyperspectral Image Restoration Transformer for Remote Sensing Imaging 12 Dec 2023 · 0 repositories · arXiv:2312.07016
-
Language-Guided Transformer for Federated Multi-Label Classification 12 Dec 2023 · 1 repository · arXiv:2312.07165
-
Large Foundation Models for Power Systems 12 Dec 2023 · 1 repository · arXiv:2312.07044
-
LLMEval: A Preliminary Study on How to Evaluate Large Language Models 12 Dec 2023 · 0 repositories · arXiv:2312.07398
-
Neural Machine Translation of Clinical Text: An Empirical Investigation into Multilingual Pre-Trained Language Models and Transfer-Learning 12 Dec 2023 · 1 repository · arXiv:2312.07250
-
One-Step Diffusion Distillation via Deep Equilibrium Models 12 Dec 2023 · 1 repository · arXiv:2401.08639Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Building Universal Foundation Models for Medical Image Analysis with Spatially Adaptive Networks 12 Dec 2023 · 1 repository · arXiv:2312.07630
-
Safety Alignment in NLP Tasks: Weakly Aligned Summarization as an In-Context Attack 12 Dec 2023 · 1 repository · arXiv:2312.06924Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Towards Equipping Transformer with the Ability of Systematic Compositionality 12 Dec 2023 · 1 repository · arXiv:2312.07280
-
Transformer-based No-Reference Image Quality Assessment via Supervised Contrastive Learning 12 Dec 2023 · 1 repository · arXiv:2312.06995
-
X4D-SceneFormer: Enhanced Scene Understanding on 4D Point Cloud Videos through Cross-modal Knowledge Transfer 12 Dec 2023 · 1 repository · arXiv:2312.07378Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
4M: Massively Multimodal Masked Modeling 11 Dec 2023 · 1 repository · arXiv:2312.06647Syntology official (archive's flag): 11 ran · 11 ran (of which 7 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples)
-
Audio-Visual LLM for Video Understanding 11 Dec 2023 · 0 repositories · arXiv:2312.06720
-
BACTrack: Building Appearance Collection for Aerial Tracking 11 Dec 2023 · 0 repositories · arXiv:2312.06136
-
Gated Linear Attention Transformers with Hardware-Efficient Training 11 Dec 2023 · 6 repositories · arXiv:2312.06635Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Generative Large Language Models Are All-purpose Text Analytics Engines: Text-to-text Learning Is All Your Need 11 Dec 2023 · 0 repositories · arXiv:2312.06099
-
Genixer: Empowering Multimodal Large Language Models as a Powerful Data Generator 11 Dec 2023 · 1 repository · arXiv:2312.06731Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
GPTBIAS: A Comprehensive Framework for Evaluating Bias in Large Language Models 11 Dec 2023 · 0 repositories · arXiv:2312.06315
-
GTA: Gated Toxicity Avoidance for LM Performance Preservation 11 Dec 2023 · 1 repository · arXiv:2312.06122
-
HyPE-GT: where Graph Transformers meet Hyperbolic Positional Encodings 11 Dec 2023 · 1 repository · arXiv:2312.06576
-
Interactive Planning Using Large Language Models for Partially Observable Robotics Tasks 11 Dec 2023 · 0 repositories · arXiv:2312.06876
-
KnowGPT: Knowledge Graph based Prompting for Large Language Models 11 Dec 2023 · 0 repositories · arXiv:2312.06185
-
TabMT: Generating tabular data with masked transformers 11 Dec 2023 · 0 repositories · arXiv:2312.06089
-
Transformers Implement Functional Gradient Descent to Learn Non-Linear Functions In Context 11 Dec 2023 · 0 repositories · arXiv:2312.06528
-
U-MixFormer: UNet-like Transformer with Mix-Attention for Efficient Semantic Segmentation 11 Dec 2023 · 1 repository · arXiv:2312.06272
-
VisionTraj: A Noise-Robust Trajectory Recovery Framework based on Large-scale Camera Network 11 Dec 2023 · 1 repository · arXiv:2312.06428Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Why "classic" Transformers are shallow and how to make them go deep 11 Dec 2023 · 0 repositories · arXiv:2312.06182
-
SIFU: Side-view Conditioned Implicit Function for Real-world Usable Clothed Human Reconstruction 10 Dec 2023 · 1 repository · arXiv:2312.06704Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 18 harvested samples) · 4 pointer-only (licence)
-
Take an Irregular Route: Enhance the Decoder of Time-Series Forecasting Transformer 10 Dec 2023 · 1 repository · arXiv:2312.05792
-
Context Tuning for Retrieval Augmented Generation 9 Dec 2023 · 0 repositories · arXiv:2312.05708
-
From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Recognition in Videos 9 Dec 2023 · 2 repositories · arXiv:2312.05447Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 4 pointer-only (licence)
-
GPT-4 and Safety Case Generation: An Exploratory Analysis 9 Dec 2023 · 0 repositories · arXiv:2312.05696
-
Identifying and Mitigating Model Failures through Few-shot CLIP-aided Diffusion Generation 9 Dec 2023 · 0 repositories · arXiv:2312.05464
-
Labrador: Exploring the Limits of Masked Language Modeling for Laboratory Data 9 Dec 2023 · 1 repository · arXiv:2312.11502Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Sim-GPT: Text Similarity via GPT Annotated Data 9 Dec 2023 · 1 repository · arXiv:2312.05603
-
The Counterattack of CNNs in Self-Supervised Learning: Larger Kernel Size might be All You Need 9 Dec 2023 · 0 repositories · arXiv:2312.05695
-
Transformer as Linear Expansion of Learngene 9 Dec 2023 · 1 repository · arXiv:2312.05614
-
Learning to Break: Knowledge-Enhanced Reasoning in Multi-Agent Debate System 8 Dec 2023 · 2 repositories · arXiv:2312.04854