Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 70
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 70 of 139: papers 6,901 to 7,000 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TrTr: A Versatile Pre-Trained Large Traffic Model based on Transformer for Capturing Trajectory Diversity in Vehicle Population 22 Sep 2023 · 0 repositories · arXiv:2309.12677
-
Vision Transformers for Computer Go 22 Sep 2023 · 0 repositories · arXiv:2309.12675
-
A Robust and Opponent-Aware League Training Method for StarCraft II 21 Sep 2023 · 0 repositories
-
AceGPT, Localizing Large Language Models in Arabic 21 Sep 2023 · 1 repository · arXiv:2309.12053
-
Adaptive Input-image Normalization for Solving the Mode Collapse Problem in GAN-based X-ray Images 21 Sep 2023 · 0 repositories · arXiv:2309.12245
-
BadTrack: A Poison-Only Backdoor Attack on Visual Object Tracking 21 Sep 2023 · 0 repositories
-
BIOT: Biosignal Transformer for Cross-data Learning in the Wild 21 Sep 2023 · 1 repository
-
Blockwise Parallel Transformers for Large Context Models 21 Sep 2023 · 1 repository
-
Boolformer: Symbolic Regression of Logic Functions with Transformers 21 Sep 2023 · 1 repository · arXiv:2309.12207
-
Can LLMs Augment Low-Resource Reading Comprehension Datasets? Opportunities and Challenges 21 Sep 2023 · 0 repositories · arXiv:2309.12426
-
Cheaply Estimating Inference Efficiency Metrics for Autoregressive Transformer Models 21 Sep 2023 · 1 repository
-
ClusterFomer: Clustering As A Universal Visual Learner 21 Sep 2023 · 1 repository
-
Code Soliloquies for Accurate Calculations in Large Language Models 21 Sep 2023 · 1 repository · arXiv:2309.12161Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Conformal PID Control for Time Series Prediction 21 Sep 2023 · 1 repository
-
Convolution and Attention Mixer for Synthetic Aperture Radar Image Change Detection 21 Sep 2023 · 1 repository · arXiv:2309.12010Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
DAC-DETR: Divide the Attention Layers and Conquer 21 Sep 2023 · 1 repository
-
DEYOv3: DETR with YOLO for Real-time Object Detection 21 Sep 2023 · 0 repositories · arXiv:2309.11851
-
Fully Transformer-Equipped Architecture for End-to-End Referring Video Object Segmentation 21 Sep 2023 · 0 repositories · arXiv:2309.11933
-
Geometric Transformer with Interatomic Positional Encoding 21 Sep 2023 · 1 repository
-
H3T: Efficient Integration of Memory Optimization and Parallelism for Large-scale Transformer Training 21 Sep 2023 · 1 repository
-
LART: Neural Correspondence Learning with Latent Regularization Transformer for 3D Motion Transfer 21 Sep 2023 · 1 repository
-
LLMR: Real-time Prompting of Interactive Worlds using Large Language Models 21 Sep 2023 · 0 repositories · arXiv:2309.12276
-
LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation Dataset 21 Sep 2023 · 5 repositories · arXiv:2309.11998
-
MG-ViT: A Multi-Granularity Method for Compact and Efficient Vision Transformers 21 Sep 2023 · 0 repositories
-
MiChao-HuaFen 1.0: A Specialized Pre-trained Corpus Dataset for Domain-specific Large Models 21 Sep 2023 · 0 repositories · arXiv:2309.13079
-
Mitigating Over-smoothing in Transformers via Regularized Nonlocal Functionals 21 Sep 2023 · 0 repositories
-
Multimodal Deep Learning for Scientific Imaging Interpretation 21 Sep 2023 · 0 repositories · arXiv:2309.12460
-
Neural Data Transformer 2: Multi-context Pretraining for Neural Spiking Activity 21 Sep 2023 · 1 repository
-
One Fits All: Power General Time Series Analysis by Pretrained LM 21 Sep 2023 · 2 repositories
-
OSNet & MNetO: Two Types of General Reconstruction Architectures for Linear Computed Tomography in Multi-Scenarios 21 Sep 2023 · 0 repositories · arXiv:2309.11858
-
PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation 21 Sep 2023 · 1 repository · arXiv:2309.12303
-
Patch n’ Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution 21 Sep 2023 · 0 repositories
-
QuantSR: Accurate Low-bit Quantization for Efficient Image Super-Resolution 21 Sep 2023 · 1 repository
-
Random-Access Infinite Context Length for Transformers 21 Sep 2023 · 1 repository
-
[Re] On the Reproducibility of CartoonX 21 Sep 2023 · 0 repositories
-
Spatial-Temporal Transformer based Video Compression Framework 21 Sep 2023 · 0 repositories · arXiv:2309.11913
-
SPRING: Studying Papers and Reasoning to play Games 21 Sep 2023 · 0 repositories
-
The Cambridge Law Corpus: A Dataset for Legal AI Research 21 Sep 2023 · 0 repositories · arXiv:2309.12269
-
The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A" 21 Sep 2023 · 2 repositories · arXiv:2309.12288Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Toward Re-Identifying Any Animal 21 Sep 2023 · 0 repositories
-
Unified 3D Segmenter As Prototypical Classifiers 21 Sep 2023 · 1 repository
-
Automatic Bat Call Classification using Transformer Networks 20 Sep 2023 · 0 repositories · arXiv:2309.11218
-
Embed-Search-Align: DNA Sequence Alignment using Transformer Models 20 Sep 2023 · 0 repositories · arXiv:2309.11087
-
Generalized Face Forgery Detection via Adaptive Learning for Pre-trained Vision Transformer 20 Sep 2023 · 1 repository · arXiv:2309.11092
-
Generative AI in Mafia-like Game Simulation 20 Sep 2023 · 0 repositories · arXiv:2309.11672
-
Generative Pre-Training of Time-Series Data for Unsupervised Fault Detection in Semiconductor Manufacturing 20 Sep 2023 · 0 repositories · arXiv:2309.11427
-
GPT-MolBERTa: GPT Molecular Features Language Model for molecular property prediction 20 Sep 2023 · 0 repositories · arXiv:2310.03030
-
Is GPT4 a Good Trader? 20 Sep 2023 · 0 repositories · arXiv:2309.10982
-
KOSMOS-2.5: A Multimodal Literate Model 20 Sep 2023 · 0 repositories · arXiv:2309.11419
-
Multi-image Tranformer for Multi-focus Image Fusion 20 Sep 2023 · 1 repository
-
PRAT: PRofiling Adversarial aTtacks 20 Sep 2023 · 0 repositories · arXiv:2309.11111
-
Rating Prediction in Conversational Task Assistants with Behavioral and Conversational-Flow Features 20 Sep 2023 · 1 repository · arXiv:2309.11307
-
RMT: Retentive Networks Meet Vision Transformers 20 Sep 2023 · 1 repository · arXiv:2309.11523Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
SkeleTR: Towrads Skeleton-based Action Recognition in the Wild 20 Sep 2023 · 0 repositories · arXiv:2309.11445
-
Transformers versus LSTMs for electronic trading 20 Sep 2023 · 1 repository · arXiv:2309.11400
-
A Family of Pretrained Transformer Language Models for Russian 19 Sep 2023 · 0 repositories · arXiv:2309.10931
-
An Evaluation of GPT-4 on the ETHICS Dataset 19 Sep 2023 · 0 repositories · arXiv:2309.10492
-
Audio signal based danger detection using signal processing and deep learning 19 Sep 2023 · 1 repository
-
CFGPT: Chinese Financial Assistant with Large Language Model 19 Sep 2023 · 1 repository · arXiv:2309.10654
-
Context-Aware Neural Video Compression on Solar Dynamics Observatory 19 Sep 2023 · 0 repositories · arXiv:2309.10784
-
Exploring Iterative Enhancement for Improving Learnersourced Multiple-Choice Question Explanations with Large Language Models 19 Sep 2023 · 1 repository · arXiv:2309.10444Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
FoleyGen: Visually-Guided Audio Generation 19 Sep 2023 · 0 repositories · arXiv:2309.10537
-
Generative AI vs. AGI: The Cognitive Strengths and Weaknesses of Modern LLMs 19 Sep 2023 · 0 repositories · arXiv:2309.10371
-
Interpret Vision Transformers as ConvNets with Dynamic Convolutions 19 Sep 2023 · 0 repositories · arXiv:2309.10713
-
Learning Dynamic MRI Reconstruction with Convolutional Network Assisted Reconstruction Swin Transformer 19 Sep 2023 · 0 repositories · arXiv:2309.10227
-
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition 19 Sep 2023 · 0 repositories · arXiv:2309.10294
-
LineMarkNet: Line Landmark Detection for Valet Parking 19 Sep 2023 · 0 repositories · arXiv:2309.10475
-
MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback 19 Sep 2023 · 1 repository · arXiv:2309.10691
-
PolicyGPT: Automated Analysis of Privacy Policies with Large Language Models 19 Sep 2023 · 0 repositories · arXiv:2309.10238
-
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing 19 Sep 2023 · 0 repositories · arXiv:2309.10356
-
Deep Prompt Tuning for Graph Transformers 18 Sep 2023 · 0 repositories · arXiv:2309.10131
-
Discovering Sounding Objects by Audio Queries for Audio Visual Segmentation 18 Sep 2023 · 0 repositories · arXiv:2309.09501
-
Distilling HuBERT with LSTMs via Decoupled Knowledge Distillation 18 Sep 2023 · 0 repositories · arXiv:2309.09920
-
EGFE: End-to-end Grouping of Fragmented Elements in UI Designs with Multimodal Learning 18 Sep 2023 · 1 repository · arXiv:2309.09867
-
Facilitating NSFW Text Detection in Open-Domain Dialogue Systems via Knowledge Distillation 18 Sep 2023 · 1 repository · arXiv:2309.09749
-
Harnessing Collective Intelligence Under a Lack of Cultural Consensus 18 Sep 2023 · 0 repositories · arXiv:2309.09787
-
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling 18 Sep 2023 · 0 repositories · arXiv:2309.09571
-
Integration of Swin UNETR and statistical shape modeling for a semi-automated segmentation of the knee and biomechanical modeling of articular cartilage 18 Sep 2023 · 0 repositories · arXiv:2312.00169
-
Q-Transformer: Scalable Offline Reinforcement Learning via Autoregressive Q-Functions 18 Sep 2023 · 0 repositories · arXiv:2309.10150
-
Target-aware Bi-Transformer for Few-shot Segmentation 18 Sep 2023 · 0 repositories · arXiv:2309.09492
-
Deep Neighbor Layer Aggregation for Lightweight Self-Supervised Monocular Depth Estimation 17 Sep 2023 · 1 repository · arXiv:2309.09272
-
Embrace Divergence for Richer Insights: A Multi-document Summarization Benchmark and a Case Study on Summarizing Diverse Information from News Articles 17 Sep 2023 · 1 repository · arXiv:2309.09369
-
MVP: Meta Visual Prompt Tuning for Few-Shot Remote Sensing Image Scene Classification 17 Sep 2023 · 0 repositories · arXiv:2309.09276
-
Performance of the Pre-Trained Large Language Model GPT-4 on Automated Short Answer Grading 17 Sep 2023 · 0 repositories · arXiv:2309.09338
-
Examining the Influence of Varied Levels of Domain Knowledge Base Inclusion in GPT-based Intelligent Tutors 16 Sep 2023 · 1 repository · arXiv:2309.12367
-
MMST-ViT: Climate Change-aware Crop Yield Prediction via Multi-Modal Spatial-Temporal Vision Transformer 16 Sep 2023 · 1 repository · arXiv:2309.09067Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
RingMo-lite: A Remote Sensing Multi-task Lightweight Network with CNN-Transformer Hybrid Framework 16 Sep 2023 · 0 repositories · arXiv:2309.09003
-
Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 16 Sep 2023 · 1 repository · arXiv:2309.08963
-
Cross-Modal Synthesis of Structural MRI and Functional Connectivity Networks via Conditional ViT-GANs 15 Sep 2023 · 0 repositories · arXiv:2309.08160
-
CoCA: Fusing Position Embedding with Collinear Constrained Attention in Transformers for Long Context Window Extending 15 Sep 2023 · 1 repository · arXiv:2309.08646
-
Differentiable Resolution Compression and Alignment for Efficient Video Classification and Retrieval 15 Sep 2023 · 1 repository · arXiv:2309.08167
-
GPT-Lab: Next Generation Of Optimal Chemistry Discovery By GPT Driven Robotic Lab 15 Sep 2023 · 0 repositories · arXiv:2309.16721
-
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer 15 Sep 2023 · 1 repository · arXiv:2309.08583
-
InvestLM: A Large Language Model for Investment using Financial Domain Instruction Tuning 15 Sep 2023 · 1 repository · arXiv:2309.13064Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
M³Net: Multilevel, Mixed and Multistage Attention Network for Salient Object Detection 15 Sep 2023 · 1 repository · arXiv:2309.08365
-
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax? 15 Sep 2023 · 0 repositories · arXiv:2309.09992
-
SculptBot: Pre-Trained Models for 3D Deformable Object Manipulation 15 Sep 2023 · 0 repositories · arXiv:2309.08728
-
Structural Self-Supervised Objectives for Transformers 15 Sep 2023 · 1 repository · arXiv:2309.08272
-
TransMUSIC: A Transformer-Aided Subspace Method for DOA Estimation with Low-Resolution ADCs 15 Sep 2023 · 1 repository · arXiv:2309.08174
-
UniST: Towards Unifying Saliency Transformer for Video Saliency Prediction and Detection 15 Sep 2023 · 0 repositories · arXiv:2309.08220