Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 35
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 35 of 139: papers 3,401 to 3,500 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Large Language Model Enhanced Knowledge Representation Learning: A Survey 1 Jul 2024 · 0 repositories · arXiv:2407.00936
-
Multi-branch CNN and grouping cascade attention for medical image classification 1 Jul 2024 · 0 repositories
-
Multi-State-Action Tokenisation in Decision Transformers for Multi-Discrete Action Spaces 1 Jul 2024 · 0 repositories · arXiv:2407.01310
-
Papez: Resource-Efficient Speech Separation with Auditory Working Memory 1 Jul 2024 · 1 repository · arXiv:2407.00888
-
Pictures Of MIDI: Controlled Music Generation via Graphical Prompts for Image-Based Diffusion Inpainting 1 Jul 2024 · 0 repositories · arXiv:2407.01499
-
Pron vs Prompt: Can Large Language Models already Challenge a World-Class Fiction Author at Creative Text Writing? 1 Jul 2024 · 0 repositories · arXiv:2407.01119
-
Roleplay-doh: Enabling Domain-Experts to Create LLM-simulated Patients via Eliciting and Adhering to Principles 1 Jul 2024 · 0 repositories · arXiv:2407.00870
-
The Solution for Temporal Sound Localisation Task of ICCV 1st Perception Test Challenge 2023 1 Jul 2024 · 0 repositories · arXiv:2407.02318
-
Uni-DVPS: Unified Model for Depth-Aware Video Panoptic Segmentation 1 Jul 2024 · 1 repository
-
Dynamic Universal Approximation Theory: The Basic Theory for Transformer-based Large Language Models 1 Jul 2024 · 0 repositories · arXiv:2407.00958
-
Evaluation of Bias Towards Medical Professionals in Large Language Models 30 Jun 2024 · 0 repositories · arXiv:2407.12031
-
Exploring a Physics-Informed Decision Transformer for Distribution System Restoration: Methodology and Performance Analysis 30 Jun 2024 · 0 repositories · arXiv:2407.00808
-
Instruct-IPT: All-in-One Image Processing Transformer via Weight Modulation 30 Jun 2024 · 1 repository · arXiv:2407.00676
-
LegalTurk Optimized BERT for Multi-Label Text Classification and NER 30 Jun 2024 · 0 repositories · arXiv:2407.00648
-
NAIST Simultaneous Speech Translation System for IWSLT 2024 30 Jun 2024 · 0 repositories · arXiv:2407.00826
-
Interpreting Pretrained Speech Models for Automatic Speech Assessment of Voice Disorders 29 Jun 2024 · 0 repositories · arXiv:2407.00531
-
LLM-Generated Natural Language Meets Scaling Laws: New Explorations and Data Augmentation Methods 29 Jun 2024 · 0 repositories · arXiv:2407.00322
-
Too Late to Train, Too Early To Use? A Study on Necessity and Viability of Low-Resource Bengali LLMs 29 Jun 2024 · 0 repositories · arXiv:2407.00416
-
Towards Universal Mesh Movement Networks 29 Jun 2024 · 1 repository · arXiv:2407.00382Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Urban Visual Appeal According to ChatGPT: Contrasting AI and Human Insights 29 Jun 2024 · 0 repositories · arXiv:2407.14268
-
AnomaLLMy -- Detecting anomalous tokens in black-box LLMs through low-confidence single-token predictions 28 Jun 2024 · 1 repository · arXiv:2406.19840
-
Attention Meets UAVs: A Comprehensive Evaluation of DDoS Detection in Low-Cost UAVs 28 Jun 2024 · 0 repositories · arXiv:2406.19881
-
Can GPT-4 Help Detect Quit Vaping Intentions? An Exploration of Automatic Data Annotation Approach 28 Jun 2024 · 0 repositories · arXiv:2407.00167
-
Covert Malicious Finetuning: Challenges in Safeguarding LLM Adaptation 28 Jun 2024 · 0 repositories · arXiv:2406.20053
-
Generative Iris Prior Embedded Transformer for Iris Restoration 28 Jun 2024 · 1 repository · arXiv:2407.00261
-
InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management 28 Jun 2024 · 1 repository · arXiv:2406.19707Syntology 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Multimodal Prototyping for cancer survival prediction 28 Jun 2024 · 1 repository · arXiv:2407.00224Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Can Large Language Models Generate High-quality Patent Claims? 27 Jun 2024 · 1 repository · arXiv:2406.19465
-
Diminishing Stereotype Bias in Image Generation Model using Reinforcemenlent Learning Feedback 27 Jun 2024 · 0 repositories · arXiv:2407.09551
-
Enhancing Video-Language Representations with Structural Spatio-Temporal Alignment 27 Jun 2024 · 0 repositories · arXiv:2406.19255
-
Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads 27 Jun 2024 · 1 repository · arXiv:2406.19391
-
Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions 27 Jun 2024 · 1 repository · arXiv:2406.19236Syntology official (archive's flag): 7 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Leveraging Contrastive Learning for Enhanced Node Representations in Tokenized Graph Transformers 27 Jun 2024 · 0 repositories · arXiv:2406.19258
-
NTFormer: A Composite Node Tokenized Graph Transformer for Node Classification 27 Jun 2024 · 0 repositories · arXiv:2406.19249
-
Retain, Blend, and Exchange: A Quality-aware Spatial-Stereo Fusion Approach for Event Stream Recognition 27 Jun 2024 · 1 repository · arXiv:2406.18845
-
Sonnet or Not, Bot? Poetry Evaluation for Large Models and Datasets 27 Jun 2024 · 1 repository · arXiv:2406.18906Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Structural Attention: Rethinking Transformer for Unpaired Medical Image Synthesis 27 Jun 2024 · 1 repository · arXiv:2406.18967
-
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models 27 Jun 2024 · 0 repositories · arXiv:2406.19358
-
UniGen: A Unified Framework for Textual Dataset Generation Using Large Language Models 27 Jun 2024 · 1 repository · arXiv:2406.18966Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
YZS-model: A Predictive Model for Organic Drug Solubility Based on Graph Convolutional Networks and Transformer-Attention 27 Jun 2024 · 1 repository · arXiv:2406.19136
-
3D-MVP: 3D Multiview Pretraining for Robotic Manipulation 26 Jun 2024 · 0 repositories · arXiv:2406.18158
-
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems 26 Jun 2024 · 1 repository · arXiv:2406.18747Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Adversarial Search Engine Optimization for Large Language Models 26 Jun 2024 · 0 repositories · arXiv:2406.18382
-
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets 26 Jun 2024 · 0 repositories · arXiv:2406.18518
-
BADGE: BADminton report Generation and Evaluation with LLM 26 Jun 2024 · 1 repository · arXiv:2406.18116
-
Geometric Features Enhanced Human-Object Interaction Detection 26 Jun 2024 · 1 repository · arXiv:2406.18691
-
Human-Free Automated Prompting for Vision-Language Anomaly Detection: Prompt Optimization with Meta-guiding Prompt Scheme 26 Jun 2024 · 0 repositories · arXiv:2406.18197
-
Improving Entity Recognition Using Ensembles of Deep Learning and Fine-tuned Large Language Models: A Case Study on Adverse Event Extraction from Multiple Sources 26 Jun 2024 · 0 repositories · arXiv:2406.18049
-
Jailbreaking LLMs with Arabic Transliteration and Arabizi 26 Jun 2024 · 1 repository · arXiv:2406.18725Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MFDNet: Multi-Frequency Deflare Network for Efficient Nighttime Flare Removal 26 Jun 2024 · 1 repository · arXiv:2406.18079Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Octo-planner: On-device Language Model for Planner-Action Agents 26 Jun 2024 · 0 repositories · arXiv:2406.18082
-
Re-Ranking Step by Step: Investigating Pre-Filtering for Re-Ranking with Large Language Models 26 Jun 2024 · 0 repositories · arXiv:2406.18740
-
Repeat and Concatenate: 2D to 3D Image Translation with 3D to 3D Generative Modeling 26 Jun 2024 · 1 repository · arXiv:2406.18422
-
Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs 26 Jun 2024 · 1 repository · arXiv:2406.18629Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Themis: A Reference-free NLG Evaluation Language Model with Flexibility and Interpretability 26 Jun 2024 · 1 repository · arXiv:2406.18365Syntology official (archive's flag): 5 ran · 5 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Unveiling and Controlling Anomalous Attention Distribution in Transformers 26 Jun 2024 · 0 repositories · arXiv:2407.01601
-
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs 26 Jun 2024 · 4 repositories · arXiv:2406.18495Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Accelerating Clinical Evidence Synthesis with Large Language Models 25 Jun 2024 · 0 repositories · arXiv:2406.17755
-
ARES: Alternating Reinforcement Learning and Supervised Fine-Tuning for Enhanced Multi-Modal Chain-of-Thought Reasoning Through Diverse AI Feedback 25 Jun 2024 · 1 repository · arXiv:2407.00087
-
Autonomous Prompt Engineering in Large Language Models 25 Jun 2024 · 0 repositories · arXiv:2407.11000
-
Cross-Modal Spherical Aggregation for Weakly Supervised Remote Sensing Shadow Removal 25 Jun 2024 · 1 repository · arXiv:2406.17469
-
Dark Transformer: A Video Transformer for Action Recognition in the Dark 25 Jun 2024 · 0 repositories · arXiv:2407.12805
-
Director3D: Real-world Camera Trajectory and 3D Scene Generation from Text 25 Jun 2024 · 1 repository · arXiv:2406.17601Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 3 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Discrete Diffusion Language Model for Long Text Summarization 25 Jun 2024 · 0 repositories · arXiv:2407.10998
-
ET tu, CLIP? Addressing Common Object Errors for Unseen Environments 25 Jun 2024 · 0 repositories · arXiv:2406.17876
-
Knowledge Distillation in Automated Annotation: Supervised Text Classification with LLM-Generated Training Labels 25 Jun 2024 · 0 repositories · arXiv:2406.17633
-
LongIns: A Challenging Long-context Instruction-based Exam for LLMs 25 Jun 2024 · 0 repositories · arXiv:2406.17588
-
Point Tree Transformer for Point Cloud Registration 25 Jun 2024 · 0 repositories · arXiv:2406.17530
-
Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers 25 Jun 2024 · 1 repository · arXiv:2406.17343Syntology official (archive's flag): 16 ran · 17 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 4 honoured, 0 violated, 7 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Revitalizing Convolutional Network for Image Restoration 25 Jun 2024 · 1 repository
-
Semi-supervised classification of dental conditions in panoramic radiographs using large language model and instance segmentation: A real-world dataset evaluation 25 Jun 2024 · 0 repositories · arXiv:2406.17915
-
Sound Tagging in Infant-centric Home Soundscapes 25 Jun 2024 · 0 repositories · arXiv:2406.17190
-
Structured Unrestricted-Rank Matrices for Parameter Efficient Fine-tuning 25 Jun 2024 · 1 repository · arXiv:2406.17740Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Task-Agnostic Federated Learning 25 Jun 2024 · 0 repositories · arXiv:2406.17235
-
Temporal-Channel Modeling in Multi-head Self-Attention for Synthetic Speech Detection 25 Jun 2024 · 1 repository · arXiv:2406.17376
-
Towards Optimal Trade-offs in Knowledge Distillation for CNNs and Vision Transformers at the Edge 25 Jun 2024 · 0 repositories · arXiv:2407.12808
-
Improving ovarian cancer segmentation accuracy with transformers through AI-guided labeling 25 Jun 2024 · 0 repositories · arXiv:2406.17666
-
Univariate Skeleton Prediction in Multivariate Systems Using Transformers 25 Jun 2024 · 1 repository · arXiv:2406.17834
-
Anomaly Detection of Tabular Data Using LLMs 24 Jun 2024 · 0 repositories · arXiv:2406.16308
-
Building on Efficient Foundations: Effectively Training LLMs with Structured Feedforward Layers 24 Jun 2024 · 1 repository · arXiv:2406.16450
-
Classification of Geological Borehole Descriptions Using a Domain Adapted Large Language Model 24 Jun 2024 · 0 repositories · arXiv:2407.10991
-
Diff3Dformer: Leveraging Slice Sequence Diffusion for Enhanced 3D CT Classification with Transformer Networks 24 Jun 2024 · 0 repositories · arXiv:2406.17173
-
Evaluation of Language Models in the Medical Context Under Resource-Constrained Settings 24 Jun 2024 · 1 repository · arXiv:2406.16611
-
Exploring Factual Entailment with NLI: A News Media Study 24 Jun 2024 · 0 repositories · arXiv:2406.16842
-
Feature Fusion for Human Activity Recognition using Parameter-Optimized Multi-Stage Graph Convolutional Network and Transformer Models 24 Jun 2024 · 0 repositories · arXiv:2406.16638
-
GeoMFormer: A General Architecture for Geometric Molecular Representation Learning 24 Jun 2024 · 1 repository · arXiv:2406.16853
-
GMT: Guided Mask Transformer for Leaf Instance Segmentation 24 Jun 2024 · 1 repository · arXiv:2406.17109
-
Large Language Models in Student Assessment: Comparing ChatGPT and Human Graders 24 Jun 2024 · 0 repositories · arXiv:2406.16510
-
Make Graph Neural Networks Great Again: A Generic Integration Paradigm of Topology-Free Patterns for Traffic Speed Prediction 24 Jun 2024 · 1 repository · arXiv:2406.16992
-
METRIK: Measurement-Efficient Randomized Controlled Trials using Transformers with Input Masking 24 Jun 2024 · 0 repositories · arXiv:2406.16351
-
Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models 24 Jun 2024 · 1 repository · arXiv:2406.17169
-
Multi-Modal Vision Transformers for Crop Mapping from Satellite Image Time Series 24 Jun 2024 · 0 repositories · arXiv:2406.16513
-
OTCE: Hybrid SSM and Attention with Cross Domain Mixture of Experts to construct Observer-Thinker-Conceiver-Expresser 24 Jun 2024 · 1 repository · arXiv:2406.16495
-
PlagBench: Exploring the Duality of Large Language Models in Plagiarism Generation and Detection 24 Jun 2024 · 0 repositories · arXiv:2406.16288
-
MixTex: Unambiguous Recognition Should Not Rely Solely on Real Data 24 Jun 2024 · 1 repository · arXiv:2406.17148
-
UNO Arena for Evaluating Sequential Decision-Making Capability of Large Language Models 24 Jun 2024 · 0 repositories · arXiv:2406.16382
-
USDC: A Dataset of User Stance and Dogmatism in Long Conversations 24 Jun 2024 · 0 repositories · arXiv:2406.16833
-
Venturing into Uncharted Waters: The Navigation Compass from Transformer to Mamba 24 Jun 2024 · 0 repositories · arXiv:2406.16722
-
Breaking the Frame: Visual Place Recognition by Overlap Prediction 23 Jun 2024 · 1 repository · arXiv:2406.16204Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
EditFollower: Tunable Car Following Models for Customizable Adaptive Cruise Control Systems 23 Jun 2024 · 0 repositories · arXiv:2407.02516