Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 77
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 77 of 139: papers 7,601 to 7,700 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ACDNet: Attention-guided Collaborative Decision Network for Effective Medication Recommendation 6 Jul 2023 · 0 repositories · arXiv:2307.03332
-
Art Authentication with Vision Transformers 6 Jul 2023 · 0 repositories · arXiv:2307.03039
-
Contrast Is All You Need 6 Jul 2023 · 0 repositories · arXiv:2307.02882
-
Cross-Spatial Pixel Integration and Cross-Stage Feature Fusion Based Transformer Network for Remote Sensing Image Super-Resolution 6 Jul 2023 · 0 repositories · arXiv:2307.02974
-
Focused Transformer: Contrastive Training for Context Scaling 6 Jul 2023 · 1 repository · arXiv:2307.03170Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples)
-
Large Language Models Empowered Autonomous Edge AI for Connected Intelligence 6 Jul 2023 · 0 repositories · arXiv:2307.02779
-
LEA: Improving Sentence Similarity Robustness to Typos Using Lexical Attention Bias 6 Jul 2023 · 1 repository · arXiv:2307.02912
-
Structure Guided Multi-modal Pre-trained Transformer for Knowledge Graph Reasoning 6 Jul 2023 · 0 repositories · arXiv:2307.03591
-
Track Mix Generation on Music Streaming Services using Transformers 6 Jul 2023 · 0 repositories · arXiv:2307.03045
-
UIT-Saviors at MEDVQA-GI 2023: Improving Multimodal Learning with Image Enhancement for Gastrointestinal Visual Question Answering 6 Jul 2023 · 0 repositories · arXiv:2307.02783
-
Vision Language Transformers: A Survey 6 Jul 2023 · 0 repositories · arXiv:2307.03254
-
Building Cooperative Embodied Agents Modularly with Large Language Models 5 Jul 2023 · 2 repositories · arXiv:2307.02485
-
Comparative Analysis of GPT-4 and Human Graders in Evaluating Praise Given to Students in Synthetic Dialogues 5 Jul 2023 · 0 repositories · arXiv:2307.02018
-
Elastic Decision Transformer 5 Jul 2023 · 0 repositories · arXiv:2307.02484
-
External Reasoning: Towards Multi-Large-Language-Models Interchangeable Assistance with Human Feedback 5 Jul 2023 · 1 repository · arXiv:2307.12057
-
Hoodwinked: Deception and Cooperation in a Text-Based Game for Language Models 5 Jul 2023 · 1 repository · arXiv:2308.01404
-
Improving Automatic Parallel Training via Balanced Memory Workload Optimization 5 Jul 2023 · 1 repository · arXiv:2307.02031
-
Jailbroken: How Does LLM Safety Training Fail? 5 Jul 2023 · 1 repository · arXiv:2307.02483
-
LongNet: Scaling Transformers to 1,000,000,000 Tokens 5 Jul 2023 · 3 repositories · arXiv:2307.02486
-
MAE-DFER: Efficient Masked Autoencoder for Self-supervised Dynamic Facial Expression Recognition 5 Jul 2023 · 1 repository · arXiv:2307.02227Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 7 pointer-only (licence)
-
Multi-Scale Prototypical Transformer for Whole Slide Image Classification 5 Jul 2023 · 0 repositories · arXiv:2307.02308
-
Deductive Additivity for Planning of Natural Language Proofs 5 Jul 2023 · 1 repository · arXiv:2307.02472
-
Open-Source LLMs for Text Annotation: A Practical Guide for Model Setting and Fine-Tuning 5 Jul 2023 · 0 repositories · arXiv:2307.02179
-
Sumformer: Universal Approximation for Efficient Transformers 5 Jul 2023 · 0 repositories · arXiv:2307.02301
-
Task-Specific Alignment and Multiple Level Transformer for Few-Shot Action Recognition 5 Jul 2023 · 1 repository · arXiv:2307.01985
-
Deep Attention Q-Network for Personalized Treatment Recommendation 4 Jul 2023 · 1 repository · arXiv:2307.01519
-
Deep Features for Contactless Fingerprint Presentation Attack Detection: Can They Be Generalized? 4 Jul 2023 · 0 repositories · arXiv:2307.01845
-
DiT-3D: Exploring Plain Diffusion Transformers for 3D Shape Generation 4 Jul 2023 · 1 repository · arXiv:2307.01831Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
EdgeFace: Efficient Face Recognition Model for Edge Devices 4 Jul 2023 · 3 repositories · arXiv:2307.01838Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Exploring Transformers for On-Line Handwritten Signature Verification 4 Jul 2023 · 0 repositories · arXiv:2307.01663
-
H-DenseFormer: An Efficient Hybrid Densely Connected Transformer for Multimodal Tumor Segmentation 4 Jul 2023 · 1 repository · arXiv:2307.01486
-
Knowledge Graph for NLG in the context of conversational agents 4 Jul 2023 · 0 repositories · arXiv:2307.01548
-
Last layer state space model for representation learning and uncertainty quantification 4 Jul 2023 · 0 repositories · arXiv:2307.01566
-
Pretraining is All You Need: A Multi-Atlas Enhanced Transformer Framework for Autism Spectrum Disorder Classification 4 Jul 2023 · 1 repository · arXiv:2307.01759
-
SageFormer: Series-Aware Framework for Long-term Multivariate Time Series Forecasting 4 Jul 2023 · 1 repository · arXiv:2307.01616Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
SelfFed: Self-supervised Federated Learning for Data Heterogeneity and Label Scarcity in IoMT 4 Jul 2023 · 0 repositories · arXiv:2307.01514
-
Spike-driven Transformer 4 Jul 2023 · 1 repository · arXiv:2307.01694Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Transformed Protoform Reconstruction 4 Jul 2023 · 1 repository · arXiv:2307.01896
-
Beyond the Snapshot: Brain Tokenized Graph Transformer for Longitudinal Brain Functional Connectome Embedding 3 Jul 2023 · 1 repository · arXiv:2307.00858
-
End-To-End Prediction of Knee Osteoarthritis Progression With Multi-Modal Transformers 3 Jul 2023 · 1 repository · arXiv:2307.00873
-
Evaluating Shutdown Avoidance of Language Models in Textual Scenarios 3 Jul 2023 · 1 repository · arXiv:2307.00787
-
Guided Patch-Grouping Wavelet Transformer with Spatial Congruence for Ultra-High Resolution Segmentation 3 Jul 2023 · 0 repositories · arXiv:2307.00711
-
Implicit Memory Transformer for Computationally Efficient Simultaneous Speech Translation 3 Jul 2023 · 1 repository · arXiv:2307.01381
-
Population Age Group Sensitivity for COVID-19 Infections with Deep Learning 3 Jul 2023 · 0 repositories · arXiv:2307.00751
-
Shiftable Context: Addressing Training-Inference Context Mismatch in Simultaneous Speech Translation 3 Jul 2023 · 1 repository · arXiv:2307.01377
-
Trainable Transformer in Transformer 3 Jul 2023 · 1 repository · arXiv:2307.01189Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
VOLTA: Improving Generative Diversity by Variational Mutual Information Maximizing Autoencoder 3 Jul 2023 · 0 repositories · arXiv:2307.00852
-
MedCPT: Contrastive Pre-trained Transformers with Large-scale PubMed Search Logs for Zero-shot Biomedical Information Retrieval 2 Jul 2023 · 2 repositories · arXiv:2307.00589
-
ClipSitu: Effectively Leveraging CLIP for Conditional Predictions in Situation Recognition 2 Jul 2023 · 1 repository · arXiv:2307.00586
-
Conformer LLMs -- Convolution Augmented Large Language Models 2 Jul 2023 · 0 repositories · arXiv:2307.00461
-
Bidirectional Correlation-Driven Inter-Frame Interaction Transformer for Referring Video Object Segmentation 2 Jul 2023 · 0 repositories · arXiv:2307.00536
-
AutoST: Training-free Neural Architecture Search for Spiking Transformers 1 Jul 2023 · 1 repository · arXiv:2307.00293
-
Effective Matching of Patients to Clinical Trials using Entity Extraction and Neural Re-ranking 1 Jul 2023 · 0 repositories · arXiv:2307.00381
-
Rearrangement Planning for General Part Assembly 1 Jul 2023 · 0 repositories · arXiv:2307.00206
-
Learning Content-enhanced Mask Transformer for Domain Generalized Urban-Scene Segmentation 1 Jul 2023 · 1 repository · arXiv:2307.00371
-
More for Less: Compact Convolutional Transformers Enable Robust Medical Image Classification with Limited Data 1 Jul 2023 · 0 repositories · arXiv:2307.00213
-
PM-DETR: Domain Adaptive Prompt Memory for Object Detection with Transformers 1 Jul 2023 · 0 repositories · arXiv:2307.00313
-
Spatial-Temporal Graph Enhanced DETR Towards Multi-Frame 3D Object Detection 1 Jul 2023 · 1 repository · arXiv:2307.00347Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation 30 Jun 2023 · 2 repositories · arXiv:2306.17817
-
Harnessing LLMs in Curricular Design: Using GPT-4 to Support Authoring of Learning Objectives 30 Jun 2023 · 0 repositories · arXiv:2306.17459
-
HVTSurv: Hierarchical Vision Transformer for Patient-Level Survival Prediction from Whole Slide Image 30 Jun 2023 · 1 repository · arXiv:2306.17373Syntology official (archive's flag): 4 ran · 4 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting 30 Jun 2023 · 0 repositories · arXiv:2306.17563
-
Learning to Localize with Attention: from sparse mmWave channel estimates from a single BS to high accuracy 3D location 30 Jun 2023 · 0 repositories · arXiv:2307.00167
-
Preference Ranking Optimization for Human Alignment 30 Jun 2023 · 1 repository · arXiv:2306.17492
-
SpATr: MoCap 3D Human Action Recognition based on Spiral Auto-encoder and Transformer Network 30 Jun 2023 · 1 repository · arXiv:2306.17574
-
SummQA at MEDIQA-Chat 2023:In-Context Learning with GPT-4 for Medical Summarization 30 Jun 2023 · 1 repository · arXiv:2306.17384
-
The Shaped Transformer: Attention Models in the Infinite Depth-and-Width Limit 30 Jun 2023 · 0 repositories · arXiv:2306.17759
-
Towards Improving the Performance of Pre-Trained Speech Models for Low-Resource Languages Through Lateral Inhibition 30 Jun 2023 · 0 repositories · arXiv:2306.17792
-
Transformers in Healthcare: A Survey 30 Jun 2023 · 0 repositories · arXiv:2307.00067
-
A negation detection assessment of GPTs: analysis with the xNot360 dataset 29 Jun 2023 · 0 repositories · arXiv:2306.16638
-
CMATH: Can Your Language Model Pass Chinese Elementary School Math Test? 29 Jun 2023 · 0 repositories · arXiv:2306.16636
-
Generative AI for Programming Education: Benchmarking ChatGPT, GPT-4, and Human Tutors 29 Jun 2023 · 0 repositories · arXiv:2306.17156
-
LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding 29 Jun 2023 · 2 repositories · arXiv:2306.17107Syntology official (archive's flag): 1 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT 29 Jun 2023 · 1 repository · arXiv:2306.17103
-
MNISQ: A Large-Scale Quantum Circuit Dataset for Machine Learning on/for Quantum Computers in the NISQ era 29 Jun 2023 · 1 repository · arXiv:2306.16627
-
UMASS_BioNLP at MEDIQA-Chat 2023: Can LLMs generate high-quality synthetic note-oriented doctor-patient conversations? 29 Jun 2023 · 1 repository · arXiv:2306.16931Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Pareto Optimal Learning for Estimating Large Language Model Errors 28 Jun 2023 · 0 repositories · arXiv:2306.16564
-
Chatlaw: A Multi-Agent Collaborative Legal Assistant with Knowledge Graph Enhanced Mixture-of-Experts Large Language Model 28 Jun 2023 · 1 repository · arXiv:2306.16092
-
Is ChatGPT a Biomedical Expert? -- Exploring the Zero-Shot Performance of Current GPT Models in Biomedical Tasks 28 Jun 2023 · 1 repository · arXiv:2306.16108
-
Leveraging GPT-4 for Food Effect Summarization to Enhance Product-Specific Guidance Development via Iterative Prompting 28 Jun 2023 · 0 repositories · arXiv:2306.16275
-
Mass Spectra Prediction with Structural Motif-based Graph Neural Networks 28 Jun 2023 · 0 repositories · arXiv:2306.16085
-
C²Former: Calibrated and Complementary Transformer for RGB-Infrared Object Detection 28 Jun 2023 · 2 repositories · arXiv:2306.16175
-
SkillNet-X: A Multilingual Multitask Model with Sparsely Activated Skills 28 Jun 2023 · 0 repositories · arXiv:2306.16176
-
Taqyim: Evaluating Arabic NLP Tasks Using ChatGPT Models 28 Jun 2023 · 1 repository · arXiv:2306.16322
-
The 2nd Place Solution for 2023 Waymo Open Sim Agents Challenge 28 Jun 2023 · 0 repositories · arXiv:2306.15914
-
CellViT: Vision Transformers for Precise Cell Segmentation and Classification 27 Jun 2023 · 3 repositories · arXiv:2306.15350Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Evaluating GPT-3.5 and GPT-4 on Grammatical Error Correction for Brazilian Portuguese 27 Jun 2023 · 0 repositories · arXiv:2306.15788
-
FedET: A Communication-Efficient Federated Class-Incremental Learning Framework Based on Enhanced Transformer 27 Jun 2023 · 0 repositories · arXiv:2306.15347
-
HyenaDNA: Long-Range Genomic Sequence Modeling at Single Nucleotide Resolution 27 Jun 2023 · 4 repositories · arXiv:2306.15794Syntology official (archive's flag): 5 ran · 17 ran (of which 11 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 4 where Syntology's instrument failed) · 11 unverified (of 28 harvested samples) · 1 pointer-only (licence)
-
LeanDojo: Theorem Proving with Retrieval-Augmented Language Models 27 Jun 2023 · 3 repositories · arXiv:2306.15626Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Novel Hybrid-Learning Algorithms for Improved Millimeter-Wave Imaging Systems 27 Jun 2023 · 1 repository · arXiv:2306.15341
-
Style-transfer based Speech and Audio-visual Scene Understanding for Robot Action Sequence Acquisition from Videos 27 Jun 2023 · 0 repositories · arXiv:2306.15644
-
Taming Detection Transformers for Medical Object Detection 27 Jun 2023 · 0 repositories · arXiv:2306.15472
-
Towards predicting Pedestrian Evacuation Time and Density from Floorplans using a Vision Transformer 27 Jun 2023 · 1 repository · arXiv:2306.15318
-
Variational latent discrete representation for time series modelling 27 Jun 2023 · 0 repositories · arXiv:2306.15282
-
CST-YOLO: A Novel Method for Blood Cell Detection Based on Improved YOLOv7 and CNN-Swin Transformer 26 Jun 2023 · 1 repository · arXiv:2306.14590
-
DNABERT-2: Efficient Foundation Model and Benchmark For Multi-Species Genome 26 Jun 2023 · 6 repositories · arXiv:2306.15006Syntology community repositories only · 13 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 10 unverified (of 23 harvested samples) · 1 pointer-only (licence)
-
FeSViBS: Federated Split Learning of Vision Transformer with Block Sampling 26 Jun 2023 · 1 repository · arXiv:2306.14638
-
Large Multimodal Models: Notes on CVPR 2023 Tutorial 26 Jun 2023 · 0 repositories · arXiv:2306.14895
-
LM4HPC: Towards Effective Language Model Application in High-Performance Computing 26 Jun 2023 · 0 repositories · arXiv:2306.14979