Methods › General › Feedforward Networks › Linear Layer › Papers, page 19
Linear Layer
Papers archive 2025-07-28
archive papers tagged: 25,421 · with a code link: 11,479 · where Syntology ran a sample: 3,523 (2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,523 of 25,421 tagged: 2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument)
Page 19 of 255: papers 1,801 to 1,900 of 25,421, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Is Relevance Propagated from Retriever to Generator in RAG? 20 Feb 2025 · 0 repositories · arXiv:2502.15025
-
KITAB-Bench: A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understanding 20 Feb 2025 · 0 repositories · arXiv:2502.14949
-
Mechanistic Understanding of Language Models in Syntactic Code Completion 20 Feb 2025 · 0 repositories · arXiv:2502.18499
-
Multiscale Byte Language Models -- A Hierarchical Architecture for Causal Million-Length Sequence Modeling 20 Feb 2025 · 1 repository · arXiv:2502.14553
-
On the Influence of Context Size and Model Choice in Retrieval-Augmented Generation Systems 20 Feb 2025 · 1 repository · arXiv:2502.14759
-
PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant 20 Feb 2025 · 0 repositories · arXiv:2502.14271
-
Predicting Fetal Birthweight from High Dimensional Data using Advanced Machine Learning 20 Feb 2025 · 0 repositories · arXiv:2502.14270
-
QUAD-LLM-MLTC: Large Language Models Ensemble Learning for Healthcare Text Multi-Label Classification 20 Feb 2025 · 0 repositories · arXiv:2502.14189
-
RelaCtrl: Relevance-Guided Efficient Control for Diffusion Transformers 20 Feb 2025 · 0 repositories · arXiv:2502.14377
-
Tabular Embeddings for Tables with Bi-Dimensional Hierarchical Metadata and Nesting 20 Feb 2025 · 0 repositories · arXiv:2502.15819
-
Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-based LLMs 20 Feb 2025 · 1 repository · arXiv:2502.14837Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples)
-
WavRAG: Audio-Integrated Retrieval Augmented Generation for Spoken Dialogue Models 20 Feb 2025 · 0 repositories · arXiv:2502.14727
-
Adapting Large Language Models for Time Series Modeling via a Novel Parameter-efficient Adaptation Method 19 Feb 2025 · 0 repositories · arXiv:2502.13725
-
Are Large Language Models In-Context Graph Learners? 19 Feb 2025 · 0 repositories · arXiv:2502.13562
-
Building Age Estimation: A New Multi-Modal Benchmark Dataset and Community Challenge 19 Feb 2025 · 1 repository · arXiv:2502.13818
-
Capturing Rich Behavior Representations: A Dynamic Action Semantic-Aware Graph Transformer for Video Captioning 19 Feb 2025 · 0 repositories · arXiv:2502.13754
-
DH-RAG: A Dynamic Historical Context-Powered Retrieval-Augmented Generation Method for Multi-Turn Dialogue 19 Feb 2025 · 0 repositories · arXiv:2502.13847
-
Extracting Social Connections from Finnish Karelian Refugee Interviews Using LLMs 19 Feb 2025 · 0 repositories · arXiv:2502.13566
-
FairKV: Balancing Per-Head KV Cache for Fast Multi-GPU Inference 19 Feb 2025 · 0 repositories · arXiv:2502.15804
-
FlexTok: Resampling Images into 1D Token Sequences of Flexible Length 19 Feb 2025 · 0 repositories · arXiv:2502.13967
-
From Correctness to Comprehension: AI Agents for Personalized Error Diagnosis in Education 19 Feb 2025 · 0 repositories · arXiv:2502.13789
-
Giving AI Personalities Leads to More Human-Like Reasoning 19 Feb 2025 · 0 repositories · arXiv:2502.14155
-
HawkBench: Investigating Resilience of RAG Methods on Stratified Information-Seeking Tasks 19 Feb 2025 · 0 repositories · arXiv:2502.13465
-
Hidden Darkness in LLM-Generated Designs: Exploring Dark Patterns in Ecommerce Web Components Generated by LLMs 19 Feb 2025 · 0 repositories · arXiv:2502.13499
-
In-Place Updates of a Graph Index for Streaming Approximate Nearest Neighbor Search 19 Feb 2025 · 0 repositories · arXiv:2502.13826
-
Inner Thinking Transformer: Leveraging Dynamic Depth Scaling to Foster Adaptive Internal Thinking 19 Feb 2025 · 0 repositories · arXiv:2502.13842
-
Learning Novel Transformer Architecture for Time-series Forecasting 19 Feb 2025 · 0 repositories · arXiv:2502.13721
-
Medical Image Classification with KAN-Integrated Transformers and Dilated Neighborhood Attention 19 Feb 2025 · 1 repository · arXiv:2502.13693
-
MoM: Linear Sequence Modeling with Mixture-of-Memories 19 Feb 2025 · 2 repositories · arXiv:2502.13685Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
PitVQA++: Vector Matrix-Low-Rank Adaptation for Open-Ended Visual Question Answering in Pituitary Surgery 19 Feb 2025 · 1 repository · arXiv:2502.14149
-
Qwen2.5-VL Technical Report 19 Feb 2025 · 4 repositories · arXiv:2502.13923Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
RAG-Gym: Optimizing Reasoning and Search Agents with Process Supervision 19 Feb 2025 · 0 repositories · arXiv:2502.13957
-
RAPTOR: Refined Approach for Product Table Object Recognition 19 Feb 2025 · 0 repositories · arXiv:2502.14918
-
RGAR: Recurrence Generation-augmented Retrieval for Factual-aware Medical Question Answering 19 Feb 2025 · 0 repositories · arXiv:2502.13361
-
Spiking Point Transformer for Point Cloud Classification 19 Feb 2025 · 1 repository · arXiv:2502.15811
-
STaR-SQL: Self-Taught Reasoner for Text-to-SQL 19 Feb 2025 · 0 repositories · arXiv:2502.13550
-
Token Adaptation via Side Graph Convolution for Temporally and Spatially Efficient Fine-tuning of 3D Point Cloud Transformers 19 Feb 2025 · 1 repository · arXiv:2502.14142
-
Universal Semantic Embeddings of Chemical Elements for Enhanced Materials Inference and Discovery 19 Feb 2025 · 0 repositories · arXiv:2502.14912
-
What are Models Thinking about? Understanding Large Language Model Hallucinations "Psychology" through Model Inner State Analysis 19 Feb 2025 · 0 repositories · arXiv:2502.13490
-
An LLM-Powered Agent for Physiological Data Analysis: A Case Study on PPG-based Heart Rate Estimation 18 Feb 2025 · 0 repositories · arXiv:2502.12836
-
DeepResonance: Enhancing Multimodal Music Understanding via Music-centric Multi-way Instruction Tuning 18 Feb 2025 · 0 repositories · arXiv:2502.12623Syntology 7 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation 18 Feb 2025 · 0 repositories · arXiv:2502.12442
-
Improving Clinical Question Answering with Multi-Task Learning: A Joint Approach for Answer Extraction and Medical Categorization 18 Feb 2025 · 0 repositories · arXiv:2502.13108
-
Language Barriers: Evaluating Cross-Lingual Performance of CNN and Transformer Architectures for Speech Quality Estimation 18 Feb 2025 · 0 repositories · arXiv:2502.13004
-
Language Models are Few-Shot Graders 18 Feb 2025 · 0 repositories · arXiv:2502.13337
-
LMN: A Tool for Generating Machine Enforceable Policies from Natural Language Access Control Rules using LLMs 18 Feb 2025 · 0 repositories · arXiv:2502.12460
-
MatterChat: A Multi-Modal LLM for Material Science 18 Feb 2025 · 0 repositories · arXiv:2502.13107
-
MVCNet: Multi-View Contrastive Network for Motor Imagery Classification 18 Feb 2025 · 1 repository · arXiv:2502.17482
-
Multimodal Mamba: Decoder-only Multimodal State Space Model via Quadratic to Linear Distillation 18 Feb 2025 · 1 repository · arXiv:2502.13145
-
Multimodal Sleep Stage and Sleep Apnea Classification Using Vision Transformer: A Multitask Explainable Learning Approach 18 Feb 2025 · 0 repositories · arXiv:2502.17486
-
Myna: Masking-Based Contrastive Learning of Musical Representations 18 Feb 2025 · 1 repository · arXiv:2502.12511
-
Oreo: A Plug-in Context Reconstructor to Enhance Retrieval-Augmented Generation 18 Feb 2025 · 0 repositories · arXiv:2502.13019
-
PathRAG: Pruning Graph-based Retrieval Augmented Generation with Relational Paths 18 Feb 2025 · 1 repository · arXiv:2502.14902Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
RingFormer: Rethinking Recurrent Transformer with Adaptive Level Signals 18 Feb 2025 · 0 repositories · arXiv:2502.13181
-
Self-Supervised Transformers as Iterative Solution Improvers for Constraint Satisfaction 18 Feb 2025 · 0 repositories · arXiv:2502.15794
-
Towards an automated workflow in materials science for combining multi-modal simulative and experimental information using data mining and large language models 18 Feb 2025 · 0 repositories · arXiv:2502.14904
-
When Segmentation Meets Hyperspectral Image: New Paradigm for Hyperspectral Image Classification 18 Feb 2025 · 1 repository · arXiv:2502.12541
-
A Survey on Bridging EEG Signals and Generative AI: From Image and Text to Beyond 17 Feb 2025 · 0 repositories · arXiv:2502.12048
-
AdaSplash: Adaptive Sparse Flash Attention 17 Feb 2025 · 1 repository · arXiv:2502.12082Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
AI-generated Text Detection with a GLTR-based Approach 17 Feb 2025 · 0 repositories · arXiv:2502.12064
-
Biases in Edge Language Models: Detection, Analysis, and Mitigation 17 Feb 2025 · 0 repositories · arXiv:2502.11349
-
Can LLMs Simulate Social Media Engagement? A Study on Action-Guided Response Generation 17 Feb 2025 · 0 repositories · arXiv:2502.12073
-
CMQCIC-Bench: A Chinese Benchmark for Evaluating Large Language Models in Medical Quality Control Indicator Calculation 17 Feb 2025 · 0 repositories · arXiv:2502.11703
-
Deep Spatio-Temporal Neural Network for Air Quality Reanalysis 17 Feb 2025 · 1 repository · arXiv:2502.11941
-
DiSCo: Device-Server Collaborative LLM-Based Text Streaming Services 17 Feb 2025 · 0 repositories · arXiv:2502.11417
-
Does RAG Really Perform Bad For Long-Context Processing? 17 Feb 2025 · 0 repositories · arXiv:2502.11444
-
Fast or Better? Balancing Accuracy and Cost in Retrieval-Augmented Generation with Flexible User Control 17 Feb 2025 · 1 repository · arXiv:2502.12145
-
FineFilter: A Fine-grained Noise Filtering Mechanism for Retrieval-Augmented Large Language Models 17 Feb 2025 · 0 repositories · arXiv:2502.11811
-
GLTW: Joint Improved Graph Transformer and LLM via Three-Word Language for Knowledge Graph Completion 17 Feb 2025 · 0 repositories · arXiv:2502.11471
-
Hierarchical Graph Topic Modeling with Topic Tree-based Transformer 17 Feb 2025 · 0 repositories · arXiv:2502.11345
-
Hyperspherical Energy Transformer with Recurrent Depth 17 Feb 2025 · 0 repositories · arXiv:2502.11646
-
If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation? 17 Feb 2025 · 0 repositories · arXiv:2502.11469Syntology 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction 17 Feb 2025 · 1 repository · arXiv:2502.11663Syntology official (archive's flag): 3 ran · 4 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training 17 Feb 2025 · 1 repository · arXiv:2502.11541
-
OCT Data is All You Need: How Vision Transformers with and without Pre-training Benefit Imaging 17 Feb 2025 · 0 repositories · arXiv:2502.12379
-
RAG vs. GraphRAG: A Systematic Evaluation and Key Insights 17 Feb 2025 · 0 repositories · arXiv:2502.11371
-
REAL-MM-RAG: A Real-World Multi-Modal Retrieval Benchmark 17 Feb 2025 · 0 repositories · arXiv:2502.12342
-
Market-Derived Financial Sentiment Analysis: Context-Aware Language Models for Crypto Forecasting 17 Feb 2025 · 1 repository · arXiv:2502.14897
-
Revisiting Robust RAG: Do We Still Need Complex Robust Training in the Era of Powerful LLMs? 17 Feb 2025 · 0 repositories · arXiv:2502.11400
-
S2TX: Cross-Attention Multi-Scale State-Space Transformer for Time Series Forecasting 17 Feb 2025 · 0 repositories · arXiv:2502.11340
-
SmartLLM: Smart Contract Auditing using Custom Generative AI 17 Feb 2025 · 0 repositories · arXiv:2502.13167
-
The geometry of BERT 17 Feb 2025 · 0 repositories · arXiv:2502.12033
-
Towards Efficient Pre-training: Exploring FP4 Precision in Large Language Models 17 Feb 2025 · 0 repositories · arXiv:2502.11458
-
Towards Mechanistic Interpretability of Graph Transformers via Attention Graphs 17 Feb 2025 · 1 repository · arXiv:2502.12352
-
X-IL: Exploring the Design Space of Imitation Learning Policies 17 Feb 2025 · 1 repository · arXiv:2502.12330Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Zero Token-Driven Deep Thinking in LLMs: Unlocking the Full Potential of Existing Parameters via Cyclic Refinement 17 Feb 2025 · 0 repositories · arXiv:2502.12214
-
A recurrent vision transformer shows signatures of primate visual attention 16 Feb 2025 · 0 repositories · arXiv:2502.10955
-
AnyRefill: A Unified, Data-Efficient Framework for Left-Prompt-Guided Vision Tasks 16 Feb 2025 · 0 repositories · arXiv:2502.11158
-
AudioSpa: Spatializing Sound Events with Text 16 Feb 2025 · 0 repositories · arXiv:2502.11219
-
Bridging the Gap: Enabling Natural Language Queries for NoSQL Databases through Text-to-NoSQL Translation 16 Feb 2025 · 0 repositories · arXiv:2502.11201
-
DA-Mamba: Domain Adaptive Hybrid Mamba-Transformer Based One-Stage Object Detection 16 Feb 2025 · 2 repositories · arXiv:2502.11178
-
Empirical evaluation of LLMs in predicting fixes of Configuration bugs in Smart Home System 16 Feb 2025 · 0 repositories · arXiv:2502.10953
-
Exposing Numeracy Gaps: A Benchmark to Evaluate Fundamental Numerical Abilities in Large Language Models 16 Feb 2025 · 1 repository · arXiv:2502.11075
-
Integrating Language Models for Enhanced Network State Monitoring in DRL-Based SFC Provisioning 16 Feb 2025 · 0 repositories · arXiv:2502.11298
-
Investigating Language Preference of Multilingual RAG Systems 16 Feb 2025 · 0 repositories · arXiv:2502.11175
-
Knowing Your Target: Target-Aware Transformer Makes Better Spatio-Temporal Video Grounding 16 Feb 2025 · 1 repository · arXiv:2502.11168
-
Leveraging Conditional Mutual Information to Improve Large Language Model Fine-Tuning For Classification 16 Feb 2025 · 0 repositories · arXiv:2502.11258
-
MultiTEND: A Multilingual Benchmark for Natural Language to NoSQL Query Translation 16 Feb 2025 · 0 repositories · arXiv:2502.11022
-
Performance Review on LLM for solving leetcode problems 16 Feb 2025 · 0 repositories · arXiv:2502.15770
-
QuOTE: Question-Oriented Text Embeddings 16 Feb 2025 · 0 repositories · arXiv:2502.10976