Methods › General › Feedforward Networks › Linear Layer › Papers, page 150
Linear Layer
Papers archive 2025-07-28
archive papers tagged: 25,421 · with a code link: 11,479 · where Syntology ran a sample: 3,523 (2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,523 of 25,421 tagged: 2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument)
Page 150 of 255: papers 14,901 to 15,000 of 25,421, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
GPS++: Reviving the Art of Message Passing for Molecular Property Prediction 6 Feb 2023 · 1 repository · arXiv:2302.02947
-
V1T: large-scale mouse V1 response prediction using a Vision Transformer 6 Feb 2023 · 1 repository · arXiv:2302.03023Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 17 harvested samples)
-
cross-modal fusion techniques for utterance-level emotion recognition from text and speech 5 Feb 2023 · 0 repositories · arXiv:2302.02447
-
KDEformer: Accelerating Transformers via Kernel Density Estimation 5 Feb 2023 · 1 repository · arXiv:2302.02451Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 7 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples)
-
Nationality Bias in Text Generation 5 Feb 2023 · 0 repositories · arXiv:2302.02463
-
Quantized Distributed Training of Large Models with Convergence Guarantees 5 Feb 2023 · 0 repositories · arXiv:2302.02390
-
VuLASTE: Long Sequence Model with Abstract Syntax Tree Embedding for vulnerability Detection 5 Feb 2023 · 0 repositories · arXiv:2302.02345
-
A New cross-domain strategy based XAI models for fake news detection 4 Feb 2023 · 0 repositories · arXiv:2302.02122
-
Knowledge Graph Completion Method Combined With Adaptive Enhanced Semantic Information 4 Feb 2023 · 0 repositories · arXiv:2302.02116
-
Learning to Agree on Vision Attention for Visual Commonsense Reasoning 4 Feb 2023 · 0 repositories · arXiv:2302.02117
-
REaLTabFormer: Generating Realistic Relational and Tabular Data using Transformers 4 Feb 2023 · 3 repositories · arXiv:2302.02041Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Revisiting Image Deblurring with an Efficient ConvNet 4 Feb 2023 · 1 repository · arXiv:2302.02234
-
Evaluating Large Language Models in Theory of Mind Tasks 4 Feb 2023 · 0 repositories · arXiv:2302.02083
-
Greedy Ordering of Layer Weight Matrices in Transformers Improves Translation 4 Feb 2023 · 0 repositories · arXiv:2302.02123
-
ANTM: An Aligned Neural Topic Model for Exploring Evolving Topics 3 Feb 2023 · 1 repository · arXiv:2302.01501
-
Bioformer: an efficient transformer language model for biomedical text mining 3 Feb 2023 · 1 repository · arXiv:2302.01588
-
CFFT-GAN: Cross-domain Feature Fusion Transformer for Exemplar-based Image Translation 3 Feb 2023 · 0 repositories · arXiv:2302.01608
-
Coinductive guide to inductive transformer heads 3 Feb 2023 · 0 repositories · arXiv:2302.01834
-
Detecting Reddit Users with Depression Using a Hybrid Neural Network SBERT-CNN 3 Feb 2023 · 0 repositories · arXiv:2302.02759
-
DEVICE: DEpth and VIsual ConcEpts Aware Transformer for TextCaps 3 Feb 2023 · 0 repositories · arXiv:2302.01540
-
DilateFormer: Multi-Scale Dilated Transformer for Visual Recognition 3 Feb 2023 · 1 repository · arXiv:2302.01791
-
HDFormer: High-order Directed Transformer for 3D Human Pose Estimation 3 Feb 2023 · 1 repository · arXiv:2302.01825Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Revisiting Intermediate Layer Distillation for Compressing Language Models: An Overfitting Perspective 3 Feb 2023 · 1 repository · arXiv:2302.01530
-
Towards Few-Shot Identification of Morality Frames using In-Context Learning 3 Feb 2023 · 0 repositories · arXiv:2302.02029
-
A Survey on Efficient Training of Transformers 2 Feb 2023 · 0 repositories · arXiv:2302.01107
-
Boosting Low-Data Instance Segmentation by Unsupervised Pre-training with Saliency Prompt 2 Feb 2023 · 0 repositories · arXiv:2302.01171
-
Creating a Large Language Model of a Philosopher 2 Feb 2023 · 0 repositories · arXiv:2302.01339
-
Curriculum-Guided Abstractive Summarization 2 Feb 2023 · 0 repositories · arXiv:2302.01342
-
Dual PatchNorm 2 Feb 2023 · 7 repositories · arXiv:2302.01327Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 4 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
FCB-SwinV2 Transformer for Polyp Segmentation 2 Feb 2023 · 0 repositories · arXiv:2302.01027
-
History-Aware Hierarchical Transformer for Multi-session Open-domain Dialogue System 2 Feb 2023 · 0 repositories · arXiv:2302.00907
-
idT5: Indonesian Version of Multilingual T5 Transformer 2 Feb 2023 · 0 repositories · arXiv:2302.00856
-
Language Quantized AutoEncoders: Towards Unsupervised Text-Image Alignment 2 Feb 2023 · 1 repository · arXiv:2302.00902
-
Longformer: Longitudinal Transformer for Alzheimer's Disease Classification with Structural MRIs 2 Feb 2023 · 1 repository · arXiv:2302.00901
-
Mnemosyne: Learning to Train Transformers with Transformers 2 Feb 2023 · 0 repositories · arXiv:2302.01128
-
Molecular Geometry-aware Transformer for accurate 3D Atomic System modeling 2 Feb 2023 · 0 repositories · arXiv:2302.00855
-
Resilient Binary Neural Network 2 Feb 2023 · 1 repository · arXiv:2302.00956Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Semantic Coherence Markers for the Early Diagnosis of the Alzheimer Disease 2 Feb 2023 · 1 repository · arXiv:2302.01025
-
Vision Transformer-based Feature Extraction for Generalized Zero-Shot Learning 2 Feb 2023 · 0 repositories · arXiv:2302.00875
-
Large language models predict human sensory judgments across six modalities 2 Feb 2023 · 0 repositories · arXiv:2302.01308
-
An Empirical Study on the Transferability of Transformer Modules in Parameter-Efficient Fine-Tuning 1 Feb 2023 · 0 repositories · arXiv:2302.00378
-
Analyzing Leakage of Personally Identifiable Information in Language Models 1 Feb 2023 · 1 repository · arXiv:2302.00539
-
Clinical Decision Transformer: Intended Treatment Recommendation through Goal Prompting 1 Feb 2023 · 0 repositories · arXiv:2302.00612
-
Co-Writing with Opinionated Language Models Affects Users' Views 1 Feb 2023 · 0 repositories · arXiv:2302.00560
-
Analyzing Feed-Forward Blocks in Transformers through the Lens of Attention Maps 1 Feb 2023 · 1 repository · arXiv:2302.00456Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
HunSum-1: an Abstractive Summarization Dataset for Hungarian 1 Feb 2023 · 1 repository · arXiv:2302.00455
-
Image-Based Vehicle Classification by Synergizing Features from Supervised and Self-Supervised Learning Paradigms 1 Feb 2023 · 0 repositories · arXiv:2302.00648
-
Improving Few-Shot Generalization by Exploring and Exploiting Auxiliary Data 1 Feb 2023 · 1 repository · arXiv:2302.00674
-
MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balanced Optimization 1 Feb 2023 · 1 repository · arXiv:2302.00290
-
Turning the Curse of Heterogeneity in Federated Learning into a Blessing for Out-of-Distribution Detection 1 Feb 2023 · 1 repository
-
An Comparative Analysis of Different Pitch and Metrical Grid Encoding Methods in the Task of Sequential Music Generation 31 Jan 2023 · 0 repositories · arXiv:2301.13383
-
Continuous Spatiotemporal Transformers 31 Jan 2023 · 1 repository · arXiv:2301.13338Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Design and Implementation of A Soccer Ball Detection System with Multiple Cameras 31 Jan 2023 · 0 repositories · arXiv:2302.00123
-
Fairness-aware Vision Transformer via Debiased Self-Attention 31 Jan 2023 · 1 repository · arXiv:2301.13803
-
FLAME: A small language model for spreadsheet formulas 31 Jan 2023 · 0 repositories · arXiv:2301.13779
-
Numeracy from Literacy: Data Science as an Emergent Skill from Large Language Models 31 Jan 2023 · 0 repositories · arXiv:2301.13382
-
Skill Decision Transformer 31 Jan 2023 · 0 repositories · arXiv:2301.13573
-
The Flan Collection: Designing Data and Methods for Effective Instruction Tuning 31 Jan 2023 · 1 repository · arXiv:2301.13688
-
The Touché23-ValueEval Dataset for Identifying Human Values behind Arguments 31 Jan 2023 · 1 repository · arXiv:2301.13771
-
UPop: Unified and Progressive Pruning for Compressing Vision-Language Transformers 31 Jan 2023 · 2 repositories · arXiv:2301.13741Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Adaptive Machine Translation with Large Language Models 30 Jan 2023 · 1 repository · arXiv:2301.13294
-
BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models 30 Jan 2023 · 17 repositories · arXiv:2301.12597Syntology community repositories only · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 8 harvested samples) · 1 pointer-only (licence)
-
CSDR-BERT: a pre-trained scientific dataset match model for Chinese Scientific Dataset Retrieval 30 Jan 2023 · 0 repositories · arXiv:2301.12700
-
DepGraph: Towards Any Structural Pruning 30 Jan 2023 · 1 repository · arXiv:2301.12900Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Guiding Online Reinforcement Learning with Action-Free Offline Pretraining 30 Jan 2023 · 1 repository · arXiv:2301.12876Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multi-modal Large Language Model Enhanced Pseudo 3D Perception Framework for Visual Commonsense Reasoning 30 Jan 2023 · 0 repositories · arXiv:2301.13335
-
REPLUG: Retrieval-Augmented Black-Box Language Models 30 Jan 2023 · 3 repositories · arXiv:2301.12652Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Representation biases in sentence transformers 30 Jan 2023 · 0 repositories · arXiv:2301.13039
-
Specializing Smaller Language Models towards Multi-Step Reasoning 30 Jan 2023 · 2 repositories · arXiv:2301.12726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Discerning Several Thousand Judgments: GPT-3 Rates the Article + Adjective + Numeral + Noun Construction 29 Jan 2023 · 0 repositories · arXiv:2301.12564
-
BERT-based Authorship Attribution on the Romanian Dataset called ROST 29 Jan 2023 · 0 repositories · arXiv:2301.12500
-
Exploring Attention Map Reuse for Efficient Transformer Neural Networks 29 Jan 2023 · 0 repositories · arXiv:2301.12444
-
Global Flood Prediction: a Multimodal Machine Learning Approach 29 Jan 2023 · 0 repositories · arXiv:2301.12548
-
Graph Mixer Networks 29 Jan 2023 · 1 repository · arXiv:2301.12493
-
PhaVIP: Phage VIrion Protein classification based on chaos game representation and Vision Transformer 29 Jan 2023 · 1 repository · arXiv:2301.12422
-
Progressive Prompts: Continual Learning for Language Models 29 Jan 2023 · 2 repositories · arXiv:2301.12314Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples)
-
Schema-Guided Semantic Accuracy: Faithfulness in Task-Oriented Dialogue Response Generation 29 Jan 2023 · 1 repository · arXiv:2301.12568
-
Semantics-enhanced Temporal Graph Networks for Content Popularity Prediction 29 Jan 2023 · 0 repositories · arXiv:2301.12355
-
Towards Vision Transformer Unrolling Fixed-Point Algorithm: a Case Study on Image Restoration 29 Jan 2023 · 0 repositories · arXiv:2301.12332
-
Aerial Image Object Detection With Vision Transformer Detector (ViTDet) 28 Jan 2023 · 1 repository · arXiv:2301.12058
-
Bipol: Multi-axes Evaluation of Bias with Explainability in Benchmark Datasets 28 Jan 2023 · 2 repositories · arXiv:2301.12139
-
Multilingual Sentence Transformer as A Multilingual Word Aligner 28 Jan 2023 · 1 repository · arXiv:2301.12140
-
Physics-guided Residual Learning for Probabilistic Power Flow Analysis 28 Jan 2023 · 0 repositories · arXiv:2301.12062
-
Predicting Visit Cost of Obstructive Sleep Apnea using Electronic Healthcare Records with Transformer 28 Jan 2023 · 1 repository · arXiv:2301.12289
-
Semantic Tagging with LSTM-CRF 28 Jan 2023 · 0 repositories · arXiv:2301.12206
-
CancerUniT: Towards a Single Unified Model for Effective Detection, Segmentation, and Diagnosis of Eight Major Cancers Using a Large Collection of CT Scans 28 Jan 2023 · 0 repositories · arXiv:2301.12291
-
Towards Equitable Representation in Text-to-Image Synthesis Models with the Cross-Cultural Understanding Benchmark (CCUB) Dataset 28 Jan 2023 · 1 repository · arXiv:2301.12073
-
A Comparative Study of Pretrained Language Models for Long Clinical Text 27 Jan 2023 · 1 repository · arXiv:2301.11847
-
A Multi-View Joint Learning Framework for Embedding Clinical Codes and Text Using Graph Neural Networks 27 Jan 2023 · 0 repositories · arXiv:2301.11608
-
Can We Use Probing to Better Understand Fine-tuning and Knowledge Distillation of the BERT NLU? 27 Jan 2023 · 0 repositories · arXiv:2301.11688
-
Context Matters: A Strategy to Pre-train Language Model for Science Education 27 Jan 2023 · 0 repositories · arXiv:2301.12031
-
Cross-Architectural Positive Pairs improve the effectiveness of Self-Supervised Learning 27 Jan 2023 · 1 repository · arXiv:2301.12025
-
Predicting Sentence-Level Factuality of News and Bias of Media Outlets 27 Jan 2023 · 1 repository · arXiv:2301.11850
-
The Exploration of Knowledge-Preserving Prompts for Document Summarisation 27 Jan 2023 · 0 repositories · arXiv:2301.11719
-
Large Language Models Are Latent Variable Models: Explaining and Finding Good Demonstrations for In-Context Learning 27 Jan 2023 · 1 repository · arXiv:2301.11916Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Large-Scale Traffic Data Imputation with Spatiotemporal Semantic Understanding 27 Jan 2023 · 0 repositories · arXiv:2301.11691
-
On the Connection Between MPNN and Graph Transformer 27 Jan 2023 · 1 repository · arXiv:2301.11956
-
Pre-training for Speech Translation: CTC Meets Optimal Transport 27 Jan 2023 · 1 repository · arXiv:2301.11716
-
Skeleton-based Action Recognition through Contrasting Two-Stream Spatial-Temporal Networks 27 Jan 2023 · 0 repositories · arXiv:2301.11495
-
SWARM Parallelism: Training Large Models Can Be Surprisingly Communication-Efficient 27 Jan 2023 · 2 repositories · arXiv:2301.11913