Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 29
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 29 of 139: papers 2,801 to 2,900 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Leveraging Fine-Tuned Retrieval-Augmented Generation with Long-Context Support: For 3GPP Standards 21 Aug 2024 · 1 repository · arXiv:2408.11775
-
Macformer: Transformer with Random Maclaurin Feature Attention 21 Aug 2024 · 0 repositories · arXiv:2408.11656
-
OAPT: Offset-Aware Partition Transformer for Double JPEG Artifacts Removal 21 Aug 2024 · 1 repository · arXiv:2408.11480Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Positional Prompt Tuning for Efficient 3D Representation Learning 21 Aug 2024 · 1 repository · arXiv:2408.11567
-
SarcasmBench: Towards Evaluating Large Language Models on Sarcasm Understanding 21 Aug 2024 · 0 repositories · arXiv:2408.11319
-
Classification of Endoscopy and Video Capsule Images using CNN-Transformer Model 20 Aug 2024 · 0 repositories · arXiv:2408.10733
-
Crafting Tomorrow's Headlines: Neural News Generation and Detection in English, Turkish, Hungarian, and Persian 20 Aug 2024 · 0 repositories · arXiv:2408.10724
-
Dr.Academy: A Benchmark for Evaluating Questioning Capability in Education for Large Language Models 20 Aug 2024 · 0 repositories · arXiv:2408.10947
-
EdgeNAT: Transformer for Efficient Edge Detection 20 Aug 2024 · 1 repository · arXiv:2408.10527
-
How Well Do Large Language Models Serve as End-to-End Secure Code Agents for Python? 20 Aug 2024 · 0 repositories · arXiv:2408.10495
-
Integrating Multi-Modal Input Token Mixer Into Mamba-Based Decision Models: Decision MetaMamba 20 Aug 2024 · 0 repositories · arXiv:2408.10517
-
MambaEVT: Event Stream based Visual Object Tracking using State Space Model 20 Aug 2024 · 1 repository · arXiv:2408.10487Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
Navigating Spatio-Temporal Heterogeneity: A Graph Transformer Approach for Traffic Forecasting 20 Aug 2024 · 1 repository · arXiv:2408.10822
-
Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications 20 Aug 2024 · 0 repositories · arXiv:2408.11878
-
Out-of-Distribution Detection with Attention Head Masking for Multimodal Document Classification 20 Aug 2024 · 1 repository · arXiv:2408.11237
-
PRformer: Pyramidal Recurrent Transformer for Multivariate Time Series Forecasting 20 Aug 2024 · 1 repository · arXiv:2408.10483Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Quantum Inverse Contextual Vision Transformers (Q-ICVT): A New Frontier in 3D Object Detection for AVs 20 Aug 2024 · 1 repository · arXiv:2408.11207
-
Revisiting VerilogEval: A Year of Improvements in Large-Language Models for Hardware Code Generation 20 Aug 2024 · 1 repository · arXiv:2408.11053Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Soda-Eval: Open-Domain Dialogue Evaluation in the age of LLMs 20 Aug 2024 · 1 repository · arXiv:2408.10902
-
GACL: Graph Attention Collaborative Learning for Temporal QoS Prediction 20 Aug 2024 · 0 repositories · arXiv:2408.10555
-
UIE-UnFold: Deep Unfolding Network with Color Priors and Vision Transformer for Underwater Image Enhancement 20 Aug 2024 · 1 repository · arXiv:2408.10653
-
Large Language Models for Classical Chinese Poetry Translation: Benchmarking, Evaluating, and Improving 19 Aug 2024 · 0 repositories · arXiv:2408.09945
-
Edge-Cloud Collaborative Motion Planning for Autonomous Driving with Large Language Models 19 Aug 2024 · 0 repositories · arXiv:2408.09972
-
Goldfish: Monolingual Language Models for 350 Languages 19 Aug 2024 · 1 repository · arXiv:2408.10441Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation 19 Aug 2024 · 0 repositories · arXiv:2408.10123
-
LightWeather: Harnessing Absolute Positional Encoding to Efficient and Scalable Global Weather Forecasting 19 Aug 2024 · 0 repositories · arXiv:2408.09695
-
MoDeGPT: Modular Decomposition for Large Language Model Compression 19 Aug 2024 · 0 repositories · arXiv:2408.09632
-
Multi-Scale Representation Learning for Image Restoration with State-Space Model 19 Aug 2024 · 0 repositories · arXiv:2408.10145
-
Pedestrian Attribute Recognition: A New Benchmark Dataset and A Large Language Model Augmented Framework 19 Aug 2024 · 2 repositories · arXiv:2408.09720
-
Propagating the prior from shallow to deep with a pre-trained velocity-model Generative Transformer network 19 Aug 2024 · 0 repositories · arXiv:2408.09767
-
R2GenCSR: Retrieving Context Samples for Large Language Model based X-ray Medical Report Generation 19 Aug 2024 · 1 repository · arXiv:2408.09743
-
SAM-UNet:Enhancing Zero-Shot Segmentation of SAM for Universal Medical Images 19 Aug 2024 · 1 repository · arXiv:2408.09886
-
Self-Directed Turing Test for Large Language Models 19 Aug 2024 · 0 repositories · arXiv:2408.09853
-
Toward Large-scale Spiking Neural Networks: A Comprehensive Survey and Future Directions 19 Aug 2024 · 0 repositories · arXiv:2409.02111
-
Transformers to SSMs: Distilling Quadratic Knowledge to Subquadratic Models 19 Aug 2024 · 1 repository · arXiv:2408.10189
-
A Unified Framework for Interpretable Transformers Using PDEs and Information Theory 18 Aug 2024 · 0 repositories · arXiv:2408.09523
-
FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model 18 Aug 2024 · 0 repositories · arXiv:2408.09384
-
Out-of-distribution generalization via composition: a lens through induction heads in Transformers 18 Aug 2024 · 1 repository · arXiv:2408.09503Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Cross-Species Data Integration for Enhanced Layer Segmentation in Kidney Pathology 17 Aug 2024 · 1 repository · arXiv:2408.09278
-
HybridOcc: NeRF Enhanced Transformer-based Multi-Camera 3D Occupancy Prediction 17 Aug 2024 · 0 repositories · arXiv:2408.09104
-
Linear Attention is Enough in Spatial-Temporal Forecasting 17 Aug 2024 · 1 repository · arXiv:2408.09158
-
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation 17 Aug 2024 · 0 repositories · arXiv:2408.09122
-
Sentiment analysis of preservice teachers' reflections using a large language model 17 Aug 2024 · 0 repositories · arXiv:2408.11862
-
TableBench: A Comprehensive and Complex Benchmark for Table Question Answering 17 Aug 2024 · 0 repositories · arXiv:2408.09174
-
Unraveling Text Generation in LLMs: A Stochastic Differential Equation Approach 17 Aug 2024 · 0 repositories · arXiv:2408.11863
-
A Novel Approach to Classify Power Quality Signals Using Vision Transformers 16 Aug 2024 · 0 repositories · arXiv:2409.00025
-
Blockchain-Enabled Accountability in Data Supply Chain: A Data Bill of Materials Approach 16 Aug 2024 · 0 repositories · arXiv:2408.08536
-
Can Large Language Models Improve the Adversarial Robustness of Graph Neural Networks? 16 Aug 2024 · 1 repository · arXiv:2408.08685
-
GeoTransformer: Enhancing Urban Forecasting with Dependency Retrieval and Geospatial Attention 16 Aug 2024 · 0 repositories · arXiv:2408.08852
-
HyCoT: A Transformer-Based Autoencoder for Hyperspectral Image Compression 16 Aug 2024 · 0 repositories · arXiv:2408.08700
-
MAT-SED: A Masked Audio Transformer with Masked-Reconstruction Based Pre-training for Sound Event Detection 16 Aug 2024 · 1 repository · arXiv:2408.08673
-
OpenCity: Open Spatio-Temporal Foundation Models for Traffic Prediction 16 Aug 2024 · 1 repository · arXiv:2408.10269Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Persona is a Double-edged Sword: Mitigating the Negative Impact of Role-playing Prompts in Zero-shot Reasoning Tasks 16 Aug 2024 · 0 repositories · arXiv:2408.08631
-
Quantifying the Effectiveness of Student Organization Activities using Natural Language Processing 16 Aug 2024 · 0 repositories · arXiv:2408.08694
-
Research on Personalized Compression Algorithm for Pre-trained Models Based on Homomorphic Entropy Increase 16 Aug 2024 · 0 repositories · arXiv:2408.08684
-
Self-Explainable Graph Transformer for Link Sign Prediction 16 Aug 2024 · 1 repository · arXiv:2408.08754Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
See What LLMs Cannot Answer: A Self-Challenge Framework for Uncovering LLM Weaknesses 16 Aug 2024 · 1 repository · arXiv:2408.08978
-
TAMER: Tree-Aware Transformer for Handwritten Mathematical Expression Recognition 16 Aug 2024 · 1 repository · arXiv:2408.08578Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Task-Aware Dynamic Transformer for Efficient Arbitrary-Scale Image Super-Resolution 16 Aug 2024 · 1 repository · arXiv:2408.08736
-
ArabLegalEval: A Multitask Benchmark for Assessing Arabic Legal Knowledge in Large Language Models 15 Aug 2024 · 1 repository · arXiv:2408.07983
-
Benchmarking the Capabilities of Large Language Models in Transportation System Engineering: Accuracy, Consistency, and Reasoning Behaviors 15 Aug 2024 · 0 repositories · arXiv:2408.08302
-
Beyond Uniform Query Distribution: Key-Driven Grouped Query Attention 15 Aug 2024 · 1 repository · arXiv:2408.08454
-
Computer Vision Model Compression Techniques for Embedded Systems: A Survey 15 Aug 2024 · 1 repository · arXiv:2408.08250
-
Distributional Drift Detection in Medical Imaging with Sketching and Fine-Tuned Transformer 15 Aug 2024 · 0 repositories · arXiv:2408.08456
-
Evaluating the Validity of Word-level Adversarial Attacks with Large Language Models 15 Aug 2024 · 1 repository
-
Leveraging Web-Crawled Data for High-Quality Fine-Tuning 15 Aug 2024 · 1 repository · arXiv:2408.08003
-
MAG-SQL: Multi-Agent Generative Approach with Soft Schema Linking and Iterative Sub-SQL Refinement for Text-to-SQL 15 Aug 2024 · 1 repository · arXiv:2408.07930Syntology official (archive's flag): 21 ran · 21 ran (of which 0 constructed an object rather than computing a result; 19 with no instrument failure: 1 honoured, 3 violated, 15 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 25 harvested samples) · 4 pointer-only (licence)
-
MambaVT: Spatio-Temporal Contextual Modeling for robust RGB-T Tracking 15 Aug 2024 · 1 repository · arXiv:2408.07889
-
Polaris: Open-ended Interactive Robotic Manipulation via Syn2Real Visual Grounding and Large Language Models 15 Aug 2024 · 0 repositories · arXiv:2408.07975
-
PQV-Mobile: A Combined Pruning and Quantization Toolkit to Optimize Vision Transformers for Mobile Applications 15 Aug 2024 · 1 repository · arXiv:2408.08437
-
Unsupervised Part Discovery via Dual Representation Alignment 15 Aug 2024 · 1 repository · arXiv:2408.08108
-
Your Turn: At Home Turning Angle Estimation for Parkinson's Disease Severity Assessment 15 Aug 2024 · 0 repositories · arXiv:2408.08182
-
A Spitting Image: Modular Superpixel Tokenization in Vision Transformers 14 Aug 2024 · 1 repository · arXiv:2408.07680Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples)
-
CodeMirage: Hallucinations in Code Generated by Large Language Models 14 Aug 2024 · 0 repositories · arXiv:2408.08333
-
Cross-aware Early Fusion with Stage-divided Vision and Language Transformer Encoders for Referring Image Segmentation 14 Aug 2024 · 0 repositories · arXiv:2408.07539
-
End-to-end Semantic-centric Video-based Multimodal Affective Computing 14 Aug 2024 · 0 repositories · arXiv:2408.07694
-
G²V²former: Graph Guided Video Vision Transformer for Face Anti-Spoofing 14 Aug 2024 · 0 repositories · arXiv:2408.07675
-
Improved 3D Whole Heart Geometry from Sparse CMR Slices 14 Aug 2024 · 1 repository · arXiv:2408.07532
-
Kraken: Inherently Parallel Transformers For Efficient Multi-Device Inference 14 Aug 2024 · 0 repositories · arXiv:2408.07802
-
MetaSeg: MetaFormer-based Global Contexts-aware Network for Efficient Semantic Segmentation 14 Aug 2024 · 1 repository · arXiv:2408.07576
-
Multi-periodicity dependency Transformer based on spectrum offset for radio frequency fingerprint identification 14 Aug 2024 · 0 repositories · arXiv:2408.07592
-
UAHOI: Uncertainty-aware Robust Interaction Learning for HOI Detection 14 Aug 2024 · 0 repositories · arXiv:2408.07430
-
A Perspective on Large Language Models, Intelligent Machines, and Knowledge Acquisition 13 Aug 2024 · 0 repositories · arXiv:2408.06598
-
Cross-View Geolocalization and Disaster Mapping with Street-View and VHR Satellite Imagery: A Case Study of Hurricane IAN 13 Aug 2024 · 1 repository · arXiv:2408.06761
-
Divide and Conquer: Improving Multi-Camera 3D Perception with 2D Semantic-Depth Priors and Input-Dependent Queries 13 Aug 2024 · 0 repositories · arXiv:2408.06901
-
FlatFusion: Delving into Details of Sparse Transformer-based Camera-LiDAR Fusion for Autonomous Driving 13 Aug 2024 · 0 repositories · arXiv:2408.06832
-
Generative AI for automatic topic labelling 13 Aug 2024 · 0 repositories · arXiv:2408.07003
-
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach 13 Aug 2024 · 0 repositories · arXiv:2408.06634
-
Leveraging Language Models for Emotion and Behavior Analysis in Education 13 Aug 2024 · 0 repositories · arXiv:2408.06874
-
Optimal Preprocessing for Joint Detection and Classification of Wireless Communication Signals in Congested Spectrum Using Computer Vision Methods 13 Aug 2024 · 0 repositories · arXiv:2408.06545
-
Spectrum Prediction With Deep 3D Pyramid Vision Transformer Learning 13 Aug 2024 · 1 repository · arXiv:2408.06870
-
Unlocking Efficiency: Adaptive Masking for Gene Transformer Models 13 Aug 2024 · 1 repository · arXiv:2408.07180
-
Using Advanced LLMs to Enhance Smaller LLMs: An Interpretable Knowledge Distillation Approach 13 Aug 2024 · 0 repositories · arXiv:2408.07238
-
Advanced Vision Transformers and Open-Set Learning for Robust Mosquito Classification: A Novel Approach to Entomological Studies 12 Aug 2024 · 0 repositories · arXiv:2408.06457
-
Body Transformer: Leveraging Robot Embodiment for Policy Learning 12 Aug 2024 · 0 repositories · arXiv:2408.06316
-
Cross-Lingual Conversational Speech Summarization with Large Language Models 12 Aug 2024 · 0 repositories · arXiv:2408.06484
-
DPDETR: Decoupled Position Detection Transformer for Infrared-Visible Object Detection 12 Aug 2024 · 0 repositories · arXiv:2408.06123
-
Enhancing 3D Transformer Segmentation Model for Medical Image with Token-level Representation Learning 12 Aug 2024 · 1 repository · arXiv:2408.05889
-
HAT: History-Augmented Anchor Transformer for Online Temporal Action Localization 12 Aug 2024 · 1 repository · arXiv:2408.06437Syntology official (archive's flag): 13 ran · 13 ran (of which 3 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Med42-v2: A Suite of Clinical LLMs 12 Aug 2024 · 0 repositories · arXiv:2408.06142