Methods › General › Feedforward Networks › Linear Layer › Papers, page 31
Linear Layer
Papers archive 2025-07-28
archive papers tagged: 25,421 · with a code link: 11,479 · where Syntology ran a sample: 3,523 (2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,523 of 25,421 tagged: 2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument)
Page 31 of 255: papers 3,001 to 3,100 of 25,421, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Demystifying Workload Imbalances in Large Transformer Model Training over Variable-length Sequences 10 Dec 2024 · 0 repositories · arXiv:2412.07894
-
Enhancing radioisotope identification in gamma spectra via supervised domain adaptation 10 Dec 2024 · 0 repositories · arXiv:2412.07069
-
Generating Knowledge Graphs from Large Language Models: A Comparative Study of GPT-4, LLaMA 2, and BERT 10 Dec 2024 · 0 repositories · arXiv:2412.07412
-
GPT-2 Through the Lens of Vector Symbolic Architectures 10 Dec 2024 · 0 repositories · arXiv:2412.07947
-
HARP: Hesitation-Aware Reframing in Transformer Inference Pass 10 Dec 2024 · 1 repository · arXiv:2412.07282
-
IntellectSeeker: A Personalized Literature Management System with the Probabilistic Model and Large Language Model 10 Dec 2024 · 1 repository · arXiv:2412.07213
-
Ontology-driven Prompt Tuning for LLM-based Task and Motion Planning 10 Dec 2024 · 0 repositories · arXiv:2412.07493
-
Post-Training Statistical Calibration for Higher Activation Sparsity 10 Dec 2024 · 1 repository · arXiv:2412.07174Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
RADIO Amplified: Improved Baselines for Agglomerative Vision Foundation Models 10 Dec 2024 · 1 repository · arXiv:2412.07679
-
RAG-based Question Answering over Heterogeneous Data and Text 10 Dec 2024 · 0 repositories · arXiv:2412.07420
-
Rethinking Emotion Annotations in the Era of Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.07906
-
STIV: Scalable Text and Image Conditioned Video Generation 10 Dec 2024 · 0 repositories · arXiv:2412.07730
-
Superficial Consciousness Hypothesis for Autoregressive Transformers 10 Dec 2024 · 1 repository · arXiv:2412.07278
-
Towards Automated Cross-domain Exploratory Data Analysis through Large Language Models 10 Dec 2024 · 2 repositories · arXiv:2412.07214
-
Towards Predictive Communication with Brain-Computer Interfaces integrating Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.07355
-
Anchoring Bias in Large Language Models: An Experimental Study 9 Dec 2024 · 0 repositories · arXiv:2412.06593
-
BatchTopK Sparse Autoencoders 9 Dec 2024 · 2 repositories · arXiv:2412.06410Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Bridging the Divide: Reconsidering Softmax and Linear Attention 9 Dec 2024 · 1 repository · arXiv:2412.06590Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 22 harvested samples) · 22 pointer-only (licence)
-
Efficient user history modeling with amortized inference for deep learning recommendation models 9 Dec 2024 · 0 repositories · arXiv:2412.06924
-
EMOv2: Pushing 5M Vision Model Frontier 9 Dec 2024 · 1 repository · arXiv:2412.06674
-
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit 9 Dec 2024 · 0 repositories · arXiv:2412.06370
-
Inverting Transformer-based Vision Models 9 Dec 2024 · 2 repositories · arXiv:2412.06534
-
Knowledge Transfer and Domain Adaptation for Fine-Grained Remote Sensing Image Segmentation 9 Dec 2024 · 1 repository · arXiv:2412.06664
-
LLM as HPC Expert: Extending RAG Architecture for HPC Data 9 Dec 2024 · 0 repositories · arXiv:2501.14733
-
Normalizing Flows are Capable Generative Models 9 Dec 2024 · 3 repositories · arXiv:2412.06329Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Open-Vocabulary High-Resolution 3D (OVHR3D) Data Segmentation and Annotation Framework 9 Dec 2024 · 0 repositories · arXiv:2412.06268
-
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06249
-
S²FT: Efficient, Scalable and Generalizable LLM Fine-tuning by Structured Sparsity 9 Dec 2024 · 0 repositories · arXiv:2412.06289
-
SiReRAG: Indexing Similar and Related Information for Multihop Reasoning 9 Dec 2024 · 0 repositories · arXiv:2412.06206Syntology 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
The Computational Limits of State-Space Models and Mamba via the Lens of Circuit Complexity 9 Dec 2024 · 0 repositories · arXiv:2412.06148
-
The Rosetta Paradox: Domain-Specific Performance Inversions in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.17821
-
Toward Non-Invasive Diagnosis of Bankart Lesions with Deep Learning 9 Dec 2024 · 1 repository · arXiv:2412.06717
-
Unseen Attack Detection in Software-Defined Networking Using a BERT-Based Large Language Model 9 Dec 2024 · 0 repositories · arXiv:2412.06239
-
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06292
-
A Collaborative Multi-Agent Approach to Retrieval-Augmented Generation Across Diverse Data 8 Dec 2024 · 0 repositories · arXiv:2412.05838
-
Are Clinical T5 Models Better for Clinical Text? 8 Dec 2024 · 1 repository · arXiv:2412.05845
-
Enhanced Computationally Efficient Long LoRA Inspired Perceiver Architectures for Auto-Regressive Language Modeling 8 Dec 2024 · 0 repositories · arXiv:2412.06106
-
Enhancing Content Representation for AR Image Quality Assessment Using Knowledge Distillation 8 Dec 2024 · 0 repositories · arXiv:2412.06003
-
Evaluating Robustness of LLMs on Crisis-Related Microblogs across Events, Information Types, and Linguistic Features 8 Dec 2024 · 0 repositories · arXiv:2412.10413
-
Fully Open Source Moxin-7B Technical Report 8 Dec 2024 · 1 repository · arXiv:2412.06845
-
KITE-DDI: A Knowledge graph Integrated Transformer Model for accurately predicting Drug-Drug Interaction Events from Drug SMILES and Biomedical Knowledge Graph 8 Dec 2024 · 0 repositories · arXiv:2412.05770
-
Language-Guided Image Tokenization for Generation 8 Dec 2024 · 0 repositories · arXiv:2412.05796
-
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor 8 Dec 2024 · 1 repository · arXiv:2412.07801
-
M³-20M: A Large-Scale Multi-Modal Molecule Dataset for AI-driven Drug Design and Discovery 8 Dec 2024 · 1 repository · arXiv:2412.06847
-
Mixture-of-PageRanks: Replacing Long-Context with Real-Time, Sparse GraphRAG 8 Dec 2024 · 0 repositories · arXiv:2412.06078
-
Paddy Disease Detection and Classification Using Computer Vision Techniques: A Mobile Application to Detect Paddy Disease 8 Dec 2024 · 0 repositories · arXiv:2412.05996
-
Vision Transformer-based Semantic Communications With Importance-Aware Quantization 8 Dec 2024 · 0 repositories · arXiv:2412.06038
-
A Comparative Study on Code Generation with Transformers 7 Dec 2024 · 0 repositories · arXiv:2412.05749
-
BERTCaps: BERT Capsule for Persian Multi-Domain Sentiment Analysis 7 Dec 2024 · 0 repositories · arXiv:2412.05591
-
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds 7 Dec 2024 · 1 repository · arXiv:2412.05631Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking 7 Dec 2024 · 0 repositories · arXiv:2412.05561
-
Innovative Sentiment Analysis and Prediction of Stock Price Using FinBERT, GPT-4 and Logistic Regression: A Data-Driven Approach 7 Dec 2024 · 0 repositories · arXiv:2412.06837
-
KG-Retriever: Efficient Knowledge Indexing for Retrieval-Augmented Large Language Models 7 Dec 2024 · 1 repository · arXiv:2412.05547Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
M³PC: Test-time Model Predictive Control for Pretrained Masked Trajectory Model 7 Dec 2024 · 1 repository · arXiv:2412.05675Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
PrivAgent: Agentic-based Red-teaming for LLM Privacy Leakage 7 Dec 2024 · 1 repository · arXiv:2412.05734
-
RefSAM3D: Adapting SAM with Cross-modal Reference for 3D Medical Image Segmentation 7 Dec 2024 · 0 repositories · arXiv:2412.05605
-
Shifting NER into High Gear: The Auto-AdvER Approach 7 Dec 2024 · 0 repositories · arXiv:2412.05655
-
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering 7 Dec 2024 · 0 repositories · arXiv:2412.06832
-
STONet: A novel neural operator for modeling solute transport in micro-cracked reservoirs 7 Dec 2024 · 1 repository · arXiv:2412.05576
-
Towards 3D Acceleration for low-power Mixture-of-Experts and Multi-Head Attention Spiking Transformers 7 Dec 2024 · 0 repositories · arXiv:2412.05540
-
Towards Learning to Reason: Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning 7 Dec 2024 · 2 repositories · arXiv:2412.05586Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo 6 Dec 2024 · 0 repositories · arXiv:2412.05223
-
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG 6 Dec 2024 · 0 repositories · arXiv:2412.05447
-
Are Frontier Large Language Models Suitable for Q&A in Science Centres? 6 Dec 2024 · 0 repositories · arXiv:2412.05200
-
BEExformer: A Fast Inferencing Transformer Architecture via Binarization with Multiple Early Exits 6 Dec 2024 · 0 repositories · arXiv:2412.05225
-
DHIL-GT: Scalable Graph Transformer with Decoupled Hierarchy Labeling 6 Dec 2024 · 0 repositories · arXiv:2412.04738
-
Enhancing Cross-Language Code Translation via Task-Specific Embedding Alignment in Retrieval-Augmented Generation 6 Dec 2024 · 0 repositories · arXiv:2412.05159
-
Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System 6 Dec 2024 · 0 repositories · arXiv:2412.06828
-
Feature Group Tabular Transformer: A Novel Approach to Traffic Crash Modeling and Causality Analysis 6 Dec 2024 · 0 repositories · arXiv:2412.06825
-
IterL2Norm: Fast Iterative L2-Normalization 6 Dec 2024 · 0 repositories · arXiv:2412.04778
-
NLP-ADBench: NLP Anomaly Detection Benchmark 6 Dec 2024 · 1 repository · arXiv:2412.04784
-
PCTreeS: 3D Point Cloud Tree Species Classification Using Airborne LiDAR Images 6 Dec 2024 · 0 repositories · arXiv:2412.04714
-
Privacy-Preserving Retrieval-Augmented Generation with Differential Privacy 6 Dec 2024 · 0 repositories · arXiv:2412.04697
-
QueEn: A Large Language Model for Quechua-English Translation 6 Dec 2024 · 0 repositories · arXiv:2412.05184
-
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens 6 Dec 2024 · 1 repository · arXiv:2412.04680Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Addressing Hallucinations with RAG and NMISS in Italian Healthcare LLM Chatbots 5 Dec 2024 · 0 repositories · arXiv:2412.04235
-
ARTeFACT: Benchmarking Segmentation Models on Diverse Analogue Media Damage 5 Dec 2024 · 0 repositories · arXiv:2412.04580
-
Automated LaTeX Code Generation from Handwritten Math Expressions Using Vision Transformer 5 Dec 2024 · 0 repositories · arXiv:2412.03853
-
Comprehensive Audio Query Handling System with Integrated Expert Models and Contextual Understanding 5 Dec 2024 · 0 repositories · arXiv:2412.03980
-
Cubify Anything: Scaling Indoor 3D Object Detection 5 Dec 2024 · 1 repository · arXiv:2412.04458
-
DEIM: DETR with Improved Matching for Fast Convergence 5 Dec 2024 · 1 repository · arXiv:2412.04234Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Dynamic Graph Representation with Contrastive Learning for Financial Market Prediction: Integrating Temporal Evolution and Static Relations 5 Dec 2024 · 1 repository · arXiv:2412.04034
-
Exploring AI Text Generation, Retrieval-Augmented Generation, and Detection Technologies: a Comprehensive Overview 5 Dec 2024 · 0 repositories · arXiv:2412.03933
-
Florence-VL: Enhancing Vision-Language Models with Generative Vision Encoder and Depth-Breadth Fusion 5 Dec 2024 · 1 repository · arXiv:2412.04424
-
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning 5 Dec 2024 · 1 repository · arXiv:2412.04661
-
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments? 5 Dec 2024 · 0 repositories · arXiv:2412.03856
-
TransAdapter: Vision Transformer for Feature-Centric Unsupervised Domain Adaptation 5 Dec 2024 · 1 repository · arXiv:2412.04073
-
Uniform Discretized Integrated Gradients: An effective attribution based method for explaining large language models 5 Dec 2024 · 0 repositories · arXiv:2412.03886
-
A Water Efficiency Dataset for African Data Centers 4 Dec 2024 · 0 repositories · arXiv:2412.03716
-
Advanced Risk Prediction and Stability Assessment of Banks Using Time Series Transformer Models 4 Dec 2024 · 0 repositories · arXiv:2412.03606
-
Advancing Conversational Psychotherapy: Integrating Privacy, Dual-Memory, and Domain Expertise with Large Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.02987
-
AntLM: Bridging Causal and Masked Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.03275
-
Controlling the Mutation in Large Language Models for the Efficient Evolution of Algorithms 4 Dec 2024 · 0 repositories · arXiv:2412.03250
-
DIVE: Taming DINO for Subject-Driven Video Editing 4 Dec 2024 · 0 repositories · arXiv:2412.03347
-
Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts? 4 Dec 2024 · 0 repositories · arXiv:2412.03235
-
EMPATH: MediaPipe-Aided Ensemble Learning with Attention-Based Transformers for Accurate Recognition of Bangla Word-Level Sign Language 4 Dec 2024 · 1 repository
-
FANAL -- Financial Activity News Alerting Language Modeling Framework 4 Dec 2024 · 0 repositories · arXiv:2412.03527
-
GraPix: Exploring Graph Modularity Optimization for Unsupervised Pixel Clustering 4 Dec 2024 · 1 repository
-
HIIF: Hierarchical Encoding based Implicit Image Function for Continuous Super-resolution 4 Dec 2024 · 0 repositories · arXiv:2412.03748
-
Interpreting Transformers for Jet Tagging 4 Dec 2024 · 1 repository · arXiv:2412.03673Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)