Methods › General › Feedforward Networks › Position-Wise Feed-Forward Layer › Papers, page 41
Position-Wise Feed-Forward Layer
Papers archive 2025-07-28
archive papers tagged: 13,895 · with a code link: 6,514 · where Syntology ran a sample: 2,229 (1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,229 of 13,895 tagged: 1,902 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 41 of 139: papers 4,001 to 4,100 of 13,895, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data 23 May 2024 · 0 repositories · arXiv:2405.14333
-
Dinomaly: The Less Is More Philosophy in Multi-Class Unsupervised Anomaly Detection 23 May 2024 · 2 repositories · arXiv:2405.14325Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 5 pointer-only (licence)
-
Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer 23 May 2024 · 1 repository · arXiv:2405.14832Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Efficient Medical Question Answering with Knowledge-Augmented Question Generation 23 May 2024 · 1 repository · arXiv:2405.14654
-
Efficient Point Transformer with Dynamic Token Aggregating for LiDAR Point Cloud Processing 23 May 2024 · 0 repositories · arXiv:2405.15827
-
Evaluating Large Language Models for Public Health Classification and Extraction Tasks 23 May 2024 · 0 repositories · arXiv:2405.14766
-
Exploring the use of a Large Language Model for data extraction in systematic reviews: a rapid feasibility study 23 May 2024 · 0 repositories · arXiv:2405.14445
-
Impact of Non-Standard Unicode Characters on Security and Comprehension in Large Language Models 23 May 2024 · 1 repository · arXiv:2405.14490
-
Improving Gloss-free Sign Language Translation by Reducing Representation Density 23 May 2024 · 1 repository · arXiv:2405.14312Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Improving Language Models Trained on Translated Data with Continual Pre-Training and Dictionary Learning Analysis 23 May 2024 · 0 repositories · arXiv:2405.14277
-
JiuZhang3.0: Efficiently Improving Mathematical Reasoning by Training Small Data Synthesis Models 23 May 2024 · 1 repository · arXiv:2405.14365Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Leveraging Semantic Segmentation Masks with Embeddings for Fine-Grained Form Classification 23 May 2024 · 0 repositories · arXiv:2405.14162
-
Linking In-context Learning in Transformers to Human Episodic Memory 23 May 2024 · 1 repository · arXiv:2405.14992Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Lorentz-Equivariant Geometric Algebra Transformers for High-Energy Physics 23 May 2024 · 1 repository · arXiv:2405.14806Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Magnetic Resonance Image Processing Transformer for General Accelerated Image Reconstruction 23 May 2024 · 0 repositories · arXiv:2405.15098
-
Optimizing example selection for retrieval-augmented machine translation with translation memories 23 May 2024 · 0 repositories · arXiv:2405.15070
-
Perception of Knowledge Boundary for Large Language Models through Semi-open-ended Question Answering 23 May 2024 · 0 repositories · arXiv:2405.14383
-
PrivCirNet: Efficient Private Inference via Block Circulant Transformation 23 May 2024 · 1 repository · arXiv:2405.14569Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
PuTR: A Pure Transformer for Decoupled and Online Multi-Object Tracking 23 May 2024 · 1 repository · arXiv:2405.14119
-
Scalable Visual State Space Model with Fractal Scanning 23 May 2024 · 0 repositories · arXiv:2405.14480
-
ShapeFormer: Shapelet Transformer for Multivariate Time Series Classification 23 May 2024 · 0 repositories · arXiv:2405.14608
-
Sparse-Tuning: Adapting Vision Transformers with Efficient Fine-tuning and Inference 23 May 2024 · 1 repository · arXiv:2405.14700Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Transformers for Image-Goal Navigation 23 May 2024 · 0 repositories · arXiv:2405.14128
-
A Transformer variant for multi-step forecasting of water level and hydrometeorological sensitivity analysis based on explainable artificial intelligence technology 22 May 2024 · 0 repositories · arXiv:2405.13646
-
A General Graph Spectral Wavelet Convolution via Chebyshev Order Decomposition 22 May 2024 · 1 repository · arXiv:2405.13806Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Affine-based Deformable Attention and Selective Fusion for Semi-dense Matching 22 May 2024 · 0 repositories · arXiv:2405.13874
-
CViT: Continuous Vision Transformer for Operator Learning 22 May 2024 · 2 repositories · arXiv:2405.13998Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Comparative Analysis of Hyperspectral Image Reconstruction Using Deep Learning for Agricultural and Biological Applications 22 May 2024 · 0 repositories · arXiv:2405.13331
-
Discrete Cosine Transform Based Decorrelated Attention for Vision Transformers 22 May 2024 · 0 repositories · arXiv:2405.13901
-
Evaluating Large Language Models with Human Feedback: Establishing a Swedish Benchmark 22 May 2024 · 1 repository · arXiv:2405.14006
-
From CNNs to Transformers in Multimodal Human Action Recognition: A Survey 22 May 2024 · 0 repositories · arXiv:2405.15813
-
High Performance P300 Spellers Using GPT2 Word Prediction With Cross-Subject Training 22 May 2024 · 0 repositories · arXiv:2405.13329
-
Leveraging 2D Information for Long-term Time Series Forecasting with Vanilla Transformers 22 May 2024 · 1 repository · arXiv:2405.13810
-
Unlocking the Power of Patch: Patch-Based MLP for Long-Term Time Series Forecasting 22 May 2024 · 0 repositories · arXiv:2405.13575
-
Semantic Equitable Clustering: A Simple and Effective Strategy for Clustering Vision Tokens 22 May 2024 · 0 repositories · arXiv:2405.13337
-
Task-agnostic Decision Transformer for Multi-type Agent Control with Federated Split Training 22 May 2024 · 0 repositories · arXiv:2405.13445
-
Unsupervised Pre-training with Language-Vision Prompts for Low-Data Instance Segmentation 22 May 2024 · 1 repository · arXiv:2405.13388
-
Why Not Transform Chat Large Language Models to Non-English? 22 May 2024 · 1 repository · arXiv:2405.13923
-
WordGame: Efficient & Effective LLM Jailbreak via Simultaneous Obfuscation in Query and Response 22 May 2024 · 0 repositories · arXiv:2405.14023
-
A Masked Semi-Supervised Learning Approach for Otago Micro Labels Recognition 21 May 2024 · 0 repositories · arXiv:2405.12711
-
BIMM: Brain Inspired Masked Modeling for Video Representation Learning 21 May 2024 · 1 repository · arXiv:2405.12757
-
BiomedParse: a biomedical foundation model for image parsing of everything everywhere all at once 21 May 2024 · 0 repositories · arXiv:2405.12971
-
Enhancing Transformer-based models for Long Sequence Time Series Forecasting via Structured Matrix 21 May 2024 · 1 repository · arXiv:2405.12462
-
Computational Tradeoffs in Image Synthesis: Diffusion, Masked-Token, and Next-Token Prediction 21 May 2024 · 0 repositories · arXiv:2405.13218
-
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities 21 May 2024 · 0 repositories · arXiv:2405.12750
-
Global-Local Detail Guided Transformer for Sea Ice Recognition in Optical Remote Sensing Images 21 May 2024 · 0 repositories · arXiv:2405.13197
-
GPT-4 Jailbreaks Itself with Near-Perfect Success Using Self-Explanation 21 May 2024 · 0 repositories · arXiv:2405.13077
-
Is Dataset Quality Still a Concern in Diagnosis Using Large Foundation Model? 21 May 2024 · 0 repositories · arXiv:2405.12584
-
Mamba in Speech: Towards an Alternative to Self-Attention 21 May 2024 · 1 repository · arXiv:2405.12609
-
Mitigating Overconfidence in Out-of-Distribution Detection by Capturing Extreme Activations 21 May 2024 · 1 repository · arXiv:2405.12658
-
Global-local Fourier Neural Operator for Accelerating Coronal Magnetic Field Model 21 May 2024 · 1 repository · arXiv:2405.12754Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4 21 May 2024 · 0 repositories · arXiv:2405.12450
-
Pseudo Channel: Time Embedding for Motor Imagery Decoding 21 May 2024 · 0 repositories · arXiv:2405.15812
-
Self-Supervised Modality-Agnostic Pre-Training of Swin Transformers 21 May 2024 · 1 repository · arXiv:2405.12781
-
System Safety Monitoring of Learned Components Using Temporal Metric Forecasting 21 May 2024 · 0 repositories · arXiv:2405.13254
-
Transformer in Touch: A Survey 21 May 2024 · 0 repositories · arXiv:2405.12779
-
Asymptotic theory of in-context learning by linear attention 20 May 2024 · 1 repository · arXiv:2405.11751Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Can AI Relate: Testing Large Language Model Response for Mental Health Support 20 May 2024 · 1 repository · arXiv:2405.12021Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples)
-
CT-Eval: Benchmarking Chinese Text-to-Table Performance in Large Language Models 20 May 2024 · 0 repositories · arXiv:2405.12174
-
Efficiency optimization of large-scale language models based on deep learning in natural language processing tasks 20 May 2024 · 0 repositories · arXiv:2405.11704
-
Fennec: Fine-grained Language Model Evaluation and Correction Extended through Branching and Bridging 20 May 2024 · 1 repository · arXiv:2405.12163
-
Is Mamba Compatible with Trajectory Optimization in Offline Reinforcement Learning? 20 May 2024 · 1 repository · arXiv:2405.12094Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Large-Scale Multi-Center CT and MRI Segmentation of Pancreas with Deep Learning 20 May 2024 · 1 repository · arXiv:2405.12367
-
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving 20 May 2024 · 0 repositories · arXiv:2405.12205
-
SSAMBA: Self-Supervised Audio Representation Learning with Mamba State Space Model 20 May 2024 · 1 repository · arXiv:2405.11831Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
A Multi-Perspective Analysis of Memorization in Large Language Models 19 May 2024 · 0 repositories · arXiv:2405.11577
-
ColorFoil: Investigating Color Blindness in Large Vision and Language Models 19 May 2024 · 1 repository · arXiv:2405.11685
-
Du-IN: Discrete units-guided mask modeling for decoding speech from Intracranial Neural signals 19 May 2024 · 1 repository · arXiv:2405.11459
-
Hummer: Towards Limited Competitive Preference Dataset 19 May 2024 · 0 repositories · arXiv:2405.11647
-
Hybrid CNN-Transformer Architecture for Efficient Large-Scale Video Snapshot Compressive Imaging 19 May 2024 · 1 repository
-
Large Language Models Can Infer Personality from Free-Form User Interactions 19 May 2024 · 0 repositories · arXiv:2405.13052
-
MHPP: Exploring the Capabilities and Limitations of Language Models Beyond Basic Code Generation 19 May 2024 · 1 repository · arXiv:2405.11430
-
NetMamba: Efficient Network Traffic Classification via Pre-training Unidirectional Mamba 19 May 2024 · 1 repository · arXiv:2405.11449
-
Review of deep learning models for crypto price prediction: implementation and evaluation 19 May 2024 · 2 repositories · arXiv:2405.11431
-
VCformer: Variable Correlation Transformer with Inherent Lagged Correlation for Multivariate Time Series Forecasting 19 May 2024 · 1 repository · arXiv:2405.11470Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A Dual Power Grid Cascading Failure Model for the Vulnerability Analysis 18 May 2024 · 0 repositories · arXiv:2405.11311
-
Automating PTSD Diagnostics in Clinical Interviews: Leveraging Large Language Models for Trauma Assessments 18 May 2024 · 0 repositories · arXiv:2405.11178
-
Can Public LLMs be used for Self-Diagnosis of Medical Conditions ? 18 May 2024 · 0 repositories · arXiv:2405.11407
-
ActiveLLM: Large Language Model-based Active Learning for Textual Few-Shot Scenarios 17 May 2024 · 0 repositories · arXiv:2405.10808
-
Are Large Language Models Moral Hypocrites? A Study Based on Moral Foundations 17 May 2024 · 0 repositories · arXiv:2405.11100
-
Benchmarking Large Language Models on CFLUE -- A Chinese Financial Language Understanding Evaluation Dataset 17 May 2024 · 2 repositories · arXiv:2405.10542
-
Enhancing Dialogue State Tracking Models through LLM-backed User-Agents Simulation 17 May 2024 · 0 repositories · arXiv:2405.13037
-
Enhancing the analysis of murine neonatal ultrasonic vocalizations: Development, evaluation, and application of different mathematical models 17 May 2024 · 1 repository · arXiv:2405.12957
-
Evaluation of large language model performance on the Biomedical Language Understanding and Reasoning Benchmark 17 May 2024 · 0 repositories
-
Hi-GMAE: Hierarchical Graph Masked Autoencoders 17 May 2024 · 1 repository · arXiv:2405.10642
-
Know in AdVance: Linear-Complexity Forecasting of Ad Campaign Performance with Evolving User Interest 17 May 2024 · 1 repository · arXiv:2405.10681
-
Language Models can Evaluate Themselves via Probability Discrepancy 17 May 2024 · 1 repository · arXiv:2405.10516
-
Language Models can Exploit Cross-Task In-context Learning for Data-Scarce Novel Tasks 17 May 2024 · 1 repository · arXiv:2405.10548Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Large Language Models in Wireless Application Design: In-Context Learning-enhanced Automatic Network Intrusion Detection 17 May 2024 · 0 repositories · arXiv:2405.11002
-
Observational Scaling Laws and the Predictability of Language Model Performance 17 May 2024 · 1 repository · arXiv:2405.10938Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Persian Pronoun Resolution: Leveraging Neural Networks and Language Models 17 May 2024 · 0 repositories · arXiv:2405.10714
-
Simultaneous Deep Learning of Myocardium Segmentation and T2 Quantification for Acute Myocardial Infarction MRI 17 May 2024 · 0 repositories · arXiv:2405.10570
-
Uncertainty Distribution Assessment of Jiles-Atherton Parameter Estimation for Inrush Current Studies 17 May 2024 · 0 repositories · arXiv:2405.11011
-
A Tale of Two Languages: Large-Vocabulary Continuous Sign Language Recognition from Spoken Language Supervision 16 May 2024 · 0 repositories · arXiv:2405.10266
-
Distilling Implicit Multimodal Knowledge into Large Language Models for Zero-Resource Dialogue Generation 16 May 2024 · 1 repository · arXiv:2405.10121
-
Dynamic In-context Learning with Conversational Models for Data Extraction and Materials Property Prediction 16 May 2024 · 1 repository · arXiv:2405.10448
-
FinTextQA: A Dataset for Long-form Financial Question Answering 16 May 2024 · 0 repositories · arXiv:2405.09980
-
GPT Store Mining and Analysis 16 May 2024 · 0 repositories · arXiv:2405.10210
-
ImgAdaPoinTr: Improving Point Cloud Completion via Images and Segmentation 16 May 2024 · 1 repository
-
Infrared Adversarial Car Stickers 16 May 2024 · 0 repositories · arXiv:2405.09924