Methods › General › Feedforward Networks › Linear Layer › Papers, page 135
Linear Layer
Papers archive 2025-07-28
archive papers tagged: 25,421 · with a code link: 11,479 · where Syntology ran a sample: 3,523 (2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,523 of 25,421 tagged: 2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument)
Page 135 of 255: papers 13,401 to 13,500 of 25,421, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
InterFormer: Interactive Local and Global Features Fusion for Automatic Speech Recognition 24 May 2023 · 0 repositories · arXiv:2305.16342
-
Is GPT-4 a Good Data Analyst? 24 May 2023 · 1 repository · arXiv:2305.15038
-
Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback 24 May 2023 · 1 repository · arXiv:2305.14975
-
KNN-LM Does Not Improve Open-ended Text Generation 24 May 2023 · 0 repositories · arXiv:2305.14625
-
Investigating Table-to-Text Generation Capabilities of LLMs in Real-World Information Seeking Scenarios 24 May 2023 · 2 repositories · arXiv:2305.14987
-
Leveraging GPT-4 for Automatic Translation Post-Editing 24 May 2023 · 0 repositories · arXiv:2305.14878
-
Enabling and Analyzing How to Efficiently Extract Information from Hybrid Long Documents with LLMs 24 May 2023 · 0 repositories · arXiv:2305.16344
-
Leveraging Pre-trained Large Language Models to Construct and Utilize World Models for Model-based Task Planning 24 May 2023 · 0 repositories · arXiv:2305.14909
-
LLMDet: A Third Party Large Language Models Generated Text Detection Tool 24 May 2023 · 1 repository · arXiv:2305.15004Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Mastering the ABCDs of Complex Questions: Answer-Based Claim Decomposition for Fine-grained Self-Evaluation 24 May 2023 · 0 repositories · arXiv:2305.14750
-
Multi-Modal Mutual Attention and Iterative Interaction for Referring Image Segmentation 24 May 2023 · 0 repositories · arXiv:2305.15302
-
Multiresolution Feature Guidance Based Transformer for Anomaly Detection 24 May 2023 · 0 repositories · arXiv:2305.14880
-
Neural Summarization of Electronic Health Records 24 May 2023 · 0 repositories · arXiv:2305.15222
-
P-vectors: A Parallel-Coupled TDNN/Transformer Network for Speaker Verification 24 May 2023 · 0 repositories · arXiv:2305.14778
-
Peek Across: Improving Multi-Document Modeling via Cross-Document Question-Answering 24 May 2023 · 1 repository · arXiv:2305.15387Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 18 harvested samples)
-
Pre-RMSNorm and Pre-CRMSNorm Transformers: Equivalent and Efficient Pre-LN Transformers 24 May 2023 · 1 repository · arXiv:2305.14858Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Predicting Token Impact Towards Efficient Vision Transformer 24 May 2023 · 0 repositories · arXiv:2305.14840
-
AutoPlan: Automatic Planning of Interactive Decision-Making Tasks With Large Language Models 24 May 2023 · 1 repository · arXiv:2305.15064Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Reasoning with Language Model is Planning with World Model 24 May 2023 · 3 repositories · arXiv:2305.14992Syntology 4 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
RefGPT: Dialogue Generation of GPT, by GPT, and for GPT 24 May 2023 · 1 repository · arXiv:2305.14994
-
Learning UI-to-Code Reverse Generator Using Visual Critic Without Rendering 24 May 2023 · 0 repositories · arXiv:2305.14637
-
Revisiting Token Dropping Strategy in Efficient BERT Pretraining 24 May 2023 · 1 repository · arXiv:2305.15273
-
Segmented Recurrent Transformer: An Efficient Sequence-to-Sequence Model 24 May 2023 · 1 repository · arXiv:2305.16340
-
Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models 24 May 2023 · 1 repository · arXiv:2305.14623Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Semi-Supervised and Long-Tailed Object Detection with CascadeMatch 24 May 2023 · 1 repository · arXiv:2305.14813
-
SPRING: Studying the Paper and Reasoning to Play Games 24 May 2023 · 1 repository · arXiv:2305.15486
-
Testing Causal Models of Word Meaning in GPT-3 and -4 24 May 2023 · 1 repository · arXiv:2305.14630
-
ToMChallenges: A Principle-Guided Dataset and Diverse Evaluation Tasks for Exploring Theory of Mind 24 May 2023 · 1 repository · arXiv:2305.15068
-
Towards Adaptive Prefix Tuning for Parameter-Efficient Language Model Fine-tuning 24 May 2023 · 0 repositories · arXiv:2305.15212
-
Towards Reliable Misinformation Mitigation: Generalization, Uncertainty, and GPT-4 24 May 2023 · 1 repository · arXiv:2305.14928Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Tricking LLMs into Disobedience: Formalizing, Analyzing, and Detecting Jailbreaks 24 May 2023 · 1 repository · arXiv:2305.14965Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Trusting Your Evidence: Hallucinate Less with Context-aware Decoding 24 May 2023 · 3 repositories · arXiv:2305.14739Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Few-shot Adaptation to Distribution Shifts By Mixing Source and Target Embeddings 23 May 2023 · 0 repositories · arXiv:2305.14521
-
When should we prefer Decision Transformers for Offline Reinforcement Learning? 23 May 2023 · 1 repository · arXiv:2305.14550Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
When your Cousin has the Right Connections: Unsupervised Bilingual Lexicon Induction for Related Data-Imbalanced Languages 23 May 2023 · 1 repository · arXiv:2305.14012
-
A Trip Towards Fairness: Bias and De-Biasing in Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.13862
-
Active Learning Principles for In-Context Learning with Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.14264
-
Aligning Large Language Models through Synthetic Feedback 23 May 2023 · 1 repository · arXiv:2305.13735Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
All Roads Lead to Rome? Exploring the Invariance of Transformers' Representations 23 May 2023 · 1 repository · arXiv:2305.14555Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Assessing Linguistic Generalisation in Language Models: A Dataset for Brazilian Portuguese 23 May 2023 · 0 repositories · arXiv:2305.14070
-
Automatic Model Selection with Large Language Models for Reasoning 23 May 2023 · 1 repository · arXiv:2305.14333Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AxomiyaBERTa: A Phonologically-aware Transformer Model for Assamese 23 May 2023 · 1 repository · arXiv:2305.13641
-
Causal Intervention for Abstractive Related Work Generation 23 May 2023 · 0 repositories · arXiv:2305.13685
-
CGCE: A Chinese Generative Chat Evaluation Benchmark for General and Financial Domains 23 May 2023 · 1 repository · arXiv:2305.14471
-
Fine-tuned LLMs Know More, Hallucinate Less with Few-Shot Sequence-to-Sequence Semantic Parsing over Wikidata 23 May 2023 · 1 repository · arXiv:2305.14202
-
Condensing Multilingual Knowledge with Lightweight Language-Specific Modules 23 May 2023 · 1 repository · arXiv:2305.13993Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Connecting the Dots: What Graph-Based Text Representations Work Best for Text Classification Using Graph Neural Networks? 23 May 2023 · 1 repository · arXiv:2305.14578
-
Cross-Attention is Not Enough: Incongruity-Aware Dynamic Hierarchical Fusion for Multimodal Affect Recognition 23 May 2023 · 1 repository · arXiv:2305.13583
-
Dancing Between Success and Failure: Edit-level Simplification Evaluation using SALSA 23 May 2023 · 0 repositories · arXiv:2305.14458
-
Deduction under Perturbed Evidence: Probing Student Simulation Capabilities of Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.14507
-
Detecting automatically the layout of clinical documents to enhance the performances of downstream natural language processing 23 May 2023 · 0 repositories · arXiv:2305.13817
-
Dynosaur: A Dynamic Growth Paradigm for Instruction-Tuning Data Curation 23 May 2023 · 1 repository · arXiv:2305.14327Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
Eeg2vec: Self-Supervised Electroencephalographic Representation Learning 23 May 2023 · 0 repositories · arXiv:2305.13957
-
Few-Shot Data Synthesis for Open Domain Multi-Hop Question Answering 23 May 2023 · 0 repositories · arXiv:2305.13691
-
Benchmarking Machine Translation with Cultural Awareness 23 May 2023 · 1 repository · arXiv:2305.14328
-
Enhancing Black-Box Few-Shot Text Classification with Prompt-Based Data Augmentation 23 May 2023 · 0 repositories · arXiv:2305.13785
-
Advancing Precise Outline-Conditioned Text Generation with Task Duality and Explicit Outline Control 23 May 2023 · 0 repositories · arXiv:2305.14459
-
Evaluating Factual Consistency of Summaries with Large Language Models 23 May 2023 · 2 repositories · arXiv:2305.14069
-
Exploring Large Language Models for Classical Philology 23 May 2023 · 1 repository · arXiv:2305.13698
-
FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation 23 May 2023 · 4 repositories · arXiv:2305.14251Syntology official (archive's flag): 10 ran · 11 ran (of which 1 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 15 harvested samples) · 6 pointer-only (licence)
-
From Characters to Words: Hierarchical Pre-trained Language Model for Open-vocabulary Language Understanding 23 May 2023 · 0 repositories · arXiv:2305.14571
-
GenSpectrum Chat: Data Exploration in Public Health Using Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.13821
-
Goat: Fine-tuned LLaMA Outperforms GPT-4 on Arithmetic Tasks 23 May 2023 · 1 repository · arXiv:2305.14201
-
GrACE: Generation using Associated Code Edits 23 May 2023 · 0 repositories · arXiv:2305.14129
-
Handling Realistic Label Noise in BERT Text Classification 23 May 2023 · 0 repositories · arXiv:2305.16337
-
HumBEL: A Human-in-the-Loop Approach for Evaluating Demographic Factors of Language Models in Human-Machine Conversations 23 May 2023 · 1 repository · arXiv:2305.14195
-
IfQA: A Dataset for Open-domain Question Answering under Counterfactual Presuppositions 23 May 2023 · 0 repositories · arXiv:2305.14010
-
Images in Language Space: Exploring the Suitability of Large Language Models for Vision & Language Tasks 23 May 2023 · 1 repository · arXiv:2305.13782
-
INSTRUCTSCORE: Explainable Text Generation Evaluation with Finegrained Feedback 23 May 2023 · 2 repositories · arXiv:2305.14282
-
SciMON: Scientific Inspiration Machines Optimized for Novelty 23 May 2023 · 1 repository · arXiv:2305.14259
-
Let's Think Frame by Frame with VIP: A Video Infilling and Prediction Dataset for Evaluating Video Chain-of-Thought 23 May 2023 · 1 repository · arXiv:2305.13903
-
LLM-powered Data Augmentation for Enhanced Cross-lingual Performance 23 May 2023 · 1 repository · arXiv:2305.14288
-
LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond 23 May 2023 · 1 repository · arXiv:2305.14540
-
MathDial: A Dialogue Tutoring Dataset with Rich Pedagogical Properties Grounded in Math Reasoning Problems 23 May 2023 · 1 repository · arXiv:2305.14536Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
mmT5: Modular Multilingual Pre-Training Solves Source Language Hallucinations 23 May 2023 · 0 repositories · arXiv:2305.14224
-
NAIL: Lexical Retrieval Indices with Efficient Non-Autoregressive Decoders 23 May 2023 · 0 repositories · arXiv:2305.14499
-
NarrativeXL: A Large-scale Dataset For Long-Term Memory Models 23 May 2023 · 1 repository · arXiv:2305.13877
-
NORM: Knowledge Distillation via N-to-One Representation Matching 23 May 2023 · 1 repository · arXiv:2305.13803Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified; the one sample that ran constructed an object rather than computing a result (of 8 harvested samples)
-
On Robustness of Finetuned Transformer-based NLP Models 23 May 2023 · 1 repository · arXiv:2305.14453
-
On Structural Expressive Power of Graph Transformers 23 May 2023 · 0 repositories · arXiv:2305.13987
-
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification 23 May 2023 · 1 repository · arXiv:2305.14032Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Physics of Language Models: Part 1, Learning Hierarchical Language Structures 23 May 2023 · 0 repositories · arXiv:2305.13673
-
Pre-training Multi-task Contrastive Learning Models for Scientific Literature Understanding 23 May 2023 · 0 repositories · arXiv:2305.14232
-
Probing Brain Context-Sensitivity with Masked-Attention Generation 23 May 2023 · 0 repositories · arXiv:2305.13863
-
QLoRA: Efficient Finetuning of Quantized LLMs 23 May 2023 · 20 repositories · arXiv:2305.14314Syntology official (archive's flag): 3 ran · 18 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 2 violated, 2 with no contract checked; 12 where Syntology's instrument failed) · 8 unverified (of 26 harvested samples) · 17 pointer-only (licence)
-
Rethinking Speech Recognition with A Multimodal Perspective via Acoustic and Semantic Cooperative Decoding 23 May 2023 · 0 repositories · arXiv:2305.14049
-
SMAP: A Novel Heterogeneous Information Framework for Scenario-based Optimal Model Assignment 23 May 2023 · 0 repositories · arXiv:2305.13634
-
Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training 23 May 2023 · 7 repositories · arXiv:2305.14342Syntology 13 ran (of which 4 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 19 harvested samples)
-
Source-Free Domain Adaptation for RGB-D Semantic Segmentation with Vision Transformers 23 May 2023 · 0 repositories · arXiv:2305.14269
-
Sources of Hallucination by Large Language Models on Inference Tasks 23 May 2023 · 1 repository · arXiv:2305.14552
-
Text Is All You Need: Learning Language Representations for Sequential Recommendation 23 May 2023 · 1 repository · arXiv:2305.13731Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards A Unified View of Sparse Feed-Forward Network in Pretraining Large Language Model 23 May 2023 · 0 repositories · arXiv:2305.13999
-
ReadMe++: Benchmarking Multilingual Language Models for Multi-Domain Readability Assessment 23 May 2023 · 1 repository · arXiv:2305.14463
-
Training Transitive and Commutative Multimodal Transformers with LoReTTa 23 May 2023 · 0 repositories · arXiv:2305.14243
-
Two Failures of Self-Consistency in the Multi-Step Reasoning of LLMs 23 May 2023 · 0 repositories · arXiv:2305.14279
-
VDD: Varied Drone Dataset for Semantic Segmentation 23 May 2023 · 1 repository · arXiv:2305.13608
-
Weakly Supervised 3D Open-vocabulary Segmentation 23 May 2023 · 1 repository · arXiv:2305.14093Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 10 pointer-only (licence)
-
WikiChat: Stopping the Hallucination of Large Language Model Chatbots by Few-Shot Grounding on Wikipedia 23 May 2023 · 1 repository · arXiv:2305.14292Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
ZeroSCROLLS: A Zero-Shot Benchmark for Long Text Understanding 23 May 2023 · 1 repository · arXiv:2305.14196Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Let GPT be a Math Tutor: Teaching Math Word Problem Solvers with Customized Exercise Generation 22 May 2023 · 0 repositories · arXiv:2305.14386