Methods › General › Feedforward Networks › Linear Layer › Papers, page 66
Linear Layer
Papers archive 2025-07-28
archive papers tagged: 25,421 · with a code link: 11,479 · where Syntology ran a sample: 3,523 (2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,523 of 25,421 tagged: 2,976 with a run with no instrument failure, 547 where every run was a failure of Syntology's instrument)
Page 66 of 255: papers 6,501 to 6,600 of 25,421, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Scoreformer: A Surrogate Model For Large-Scale Prediction of Docking Scores 13 Jun 2024 · 0 repositories · arXiv:2406.09346
-
Separations in the Representational Capabilities of Transformers and Recurrent Architectures 13 Jun 2024 · 0 repositories · arXiv:2406.09347
-
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models 13 Jun 2024 · 0 repositories · arXiv:2406.09519
-
Transformers meet Neural Algorithmic Reasoners 13 Jun 2024 · 0 repositories · arXiv:2406.09308
-
Vertical LoRA: Dense Expectation-Maximization Interpretation of Transformers 13 Jun 2024 · 1 repository · arXiv:2406.09315
-
A Robust Pipeline for Classification and Detection of Bleeding Frames in Wireless Capsule Endoscopy using Swin Transformer and RT-DETR 12 Jun 2024 · 0 repositories · arXiv:2406.08046
-
A Sociotechnical Lens for Evaluating Computer Vision Models: A Case Study on Detecting and Reasoning about Gender and Emotion 12 Jun 2024 · 0 repositories · arXiv:2406.08222
-
Ad Auctions for LLMs via Retrieval Augmented Generation 12 Jun 2024 · 0 repositories · arXiv:2406.09459
-
AdaNCA: Neural Cellular Automata As Adaptors For More Robust Vision Transformer 12 Jun 2024 · 0 repositories · arXiv:2406.08298
-
An Empirical Study of Mamba-based Language Models 12 Jun 2024 · 1 repository · arXiv:2406.07887
-
Automated Information Extraction from Thyroid Operation Narrative: A Comparative Study of GPT-4 and Fine-tuned KoELECTRA 12 Jun 2024 · 0 repositories · arXiv:2406.07922
-
Bilateral Interaction for Local-Global Collaborative Perception in Low-Light Image Enhancement 12 Jun 2024 · 1 repository
-
CLDTA: Contrastive Learning based on Diagonal Transformer Autoencoder for Cross-Dataset EEG Emotion Recognition 12 Jun 2024 · 0 repositories · arXiv:2406.08081
-
ConceptHash: Interpretable Fine-Grained Hashing via Concept Discovery 12 Jun 2024 · 1 repository · arXiv:2406.08457
-
CT3D++: Improving 3D Object Detection with Keypoint-induced Channel-wise Transformer 12 Jun 2024 · 1 repository · arXiv:2406.08152
-
DafnyBench: A Benchmark for Formal Software Verification 12 Jun 2024 · 1 repository · arXiv:2406.08467
-
Exploring Fact Memorization and Style Imitation in LLMs Using QLoRA: An Experimental Study and Quality Assessment Methods 12 Jun 2024 · 0 repositories · arXiv:2406.08582
-
FaithFill: Faithful Inpainting for Object Completion Using a Single Reference Image 12 Jun 2024 · 0 repositories · arXiv:2406.07865
-
Fine-Tuned 'Small' LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification 12 Jun 2024 · 1 repository · arXiv:2406.08660Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Fully Few-shot Class-incremental Audio Classification Using Expandable Dual-embedding Extractor 12 Jun 2024 · 1 repository · arXiv:2406.08122
-
HelpSteer2: Open-source dataset for training top-performing reward models 12 Jun 2024 · 1 repository · arXiv:2406.08673
-
How well it works: Benchmarking performance of GPT models on medical natural language processing tasks 12 Jun 2024 · 0 repositories
-
I Don't Know You, But I Can Catch You: Real-Time Defense against Diverse Adversarial Patches for Object Detectors 12 Jun 2024 · 0 repositories · arXiv:2406.10285
-
ICE-G: Image Conditional Editing of 3D Gaussian Splats 12 Jun 2024 · 0 repositories · arXiv:2406.08488
-
Making Task-Oriented Dialogue Datasets More Natural by Synthetically Generating Indirect User Requests 12 Jun 2024 · 0 repositories · arXiv:2406.07794
-
Judging the Judges: A Systematic Study of Position Bias in LLM-as-a-Judge 12 Jun 2024 · 1 repository · arXiv:2406.07791Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
Label-aware Hard Negative Sampling Strategies with Momentum Contrastive Learning for Implicit Hate Speech Detection 12 Jun 2024 · 1 repository · arXiv:2406.07886
-
Leveraging Large Language Models for Web Scraping 12 Jun 2024 · 0 repositories · arXiv:2406.08246
-
MaIL: Improving Imitation Learning with Mamba 12 Jun 2024 · 1 repository · arXiv:2406.08234Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Mistral-C2F: Coarse to Fine Actor for Analytical and Reasoning Enhancement in RLHF and Effective-Merged LLMs 12 Jun 2024 · 0 repositories · arXiv:2406.08657
-
Multimodal Representation Loss Between Timed Text and Audio for Regularized Speech Separation 12 Jun 2024 · 0 repositories · arXiv:2406.08328
-
On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models 12 Jun 2024 · 1 repository · arXiv:2406.08486
-
Scaling Manipulation Learning with Visual Kinematic Chain Prediction 12 Jun 2024 · 1 repository · arXiv:2406.07837
-
Supportiveness-based Knowledge Rewriting for Retrieval-augmented Language Modeling 12 Jun 2024 · 0 repositories · arXiv:2406.08116
-
Tailoring Generative AI Chatbots for Multiethnic Communities in Disaster Preparedness Communication: Extending the CASA Paradigm 12 Jun 2024 · 1 repository · arXiv:2406.08411
-
Learning positional encodings in transformers depends on initialization 12 Jun 2024 · 0 repositories · arXiv:2406.08272
-
Transformer-based Model for ASR N-Best Rescoring and Rewriting 12 Jun 2024 · 0 repositories · arXiv:2406.08207
-
VeraCT Scan: Retrieval-Augmented Fake News Detection with Justifiable Reasoning 12 Jun 2024 · 0 repositories · arXiv:2406.10289
-
What If We Recaption Billions of Web Images with LLaMA-3? 12 Jun 2024 · 0 repositories · arXiv:2406.08478
-
Agent-SiMT: Agent-assisted Simultaneous Machine Translation with Large Language Models 11 Jun 2024 · 1 repository · arXiv:2406.06910
-
AI Sandbagging: Language Models can Strategically Underperform on Evaluations 11 Jun 2024 · 1 repository · arXiv:2406.07358Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Beyond Words: On Large Language Models Actionability in Mission-Critical Risk Analysis 11 Jun 2024 · 0 repositories · arXiv:2406.10273
-
Bilingual Sexism Classification: Fine-Tuned XLM-RoBERTa and GPT-3.5 Few-Shot Learning 11 Jun 2024 · 0 repositories · arXiv:2406.07287
-
COVID-19 Twitter Sentiment Classification Using Hybrid Deep Learning Model Based on Grid Search Methodology 11 Jun 2024 · 0 repositories · arXiv:2406.10266
-
DARA: Decomposition-Alignment-Reasoning Autonomous Language Agent for Question Answering over Knowledge Graphs 11 Jun 2024 · 1 repository · arXiv:2406.07080Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
DR-RAG: Applying Dynamic Document Relevance to Retrieval-Augmented Generation for Question-Answering 11 Jun 2024 · 0 repositories · arXiv:2406.07348
-
Effectively Compress KV Heads for LLM 11 Jun 2024 · 0 repositories · arXiv:2406.07056
-
Entropy-Reinforced Planning with Large Language Models for Drug Discovery 11 Jun 2024 · 1 repository · arXiv:2406.07025
-
Evolving Subnetwork Training for Large Language Models 11 Jun 2024 · 0 repositories · arXiv:2406.06962
-
ExHuBERT: Enhancing HuBERT Through Block Extension and Fine-Tuning on 37 Emotion Datasets 11 Jun 2024 · 1 repository
-
FastAST: Accelerating Audio Spectrogram Transformer via Token Merging and Cross-Model Knowledge Distillation 11 Jun 2024 · 1 repository · arXiv:2406.07676
-
Flextron: Many-in-One Flexible Large Language Model 11 Jun 2024 · 0 repositories · arXiv:2406.10260
-
Grapevine Disease Prediction Using Climate Variables from Multi-Sensor Remote Sensing Imagery via a Transformer Model 11 Jun 2024 · 0 repositories · arXiv:2406.07094
-
GridPE: Unifying Positional Encoding in Transformers with a Grid Cell-Inspired Framework 11 Jun 2024 · 0 repositories · arXiv:2406.07049
-
MLLMGuard: A Multi-dimensional Safety Evaluation Suite for Multimodal Large Language Models 11 Jun 2024 · 1 repository · arXiv:2406.07594Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Multi-objective Reinforcement learning from AI Feedback 11 Jun 2024 · 1 repository · arXiv:2406.07295
-
Multimodal Belief Prediction 11 Jun 2024 · 1 repository · arXiv:2406.07466
-
Noise-robust Speech Separation with Fast Generative Correction 11 Jun 2024 · 1 repository · arXiv:2406.07461
-
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language 11 Jun 2024 · 0 repositories · arXiv:2406.08519
-
Towards Generalized Hydrological Forecasting using Transformer Models for 120-Hour Streamflow Prediction 11 Jun 2024 · 0 repositories · arXiv:2406.07484
-
Unused information in token probability distribution of generative LLM: improving LLM reading comprehension through calculation of expected values 11 Jun 2024 · 1 repository · arXiv:2406.10267
-
UVIS: Unsupervised Video Instance Segmentation 11 Jun 2024 · 0 repositories · arXiv:2406.06908
-
Validating LLM-Generated Programs with Metamorphic Prompt Testing 11 Jun 2024 · 0 repositories · arXiv:2406.06864
-
A Comparative Survey of Vision Transformers for Feature Extraction in Texture Analysis 10 Jun 2024 · 0 repositories · arXiv:2406.06136
-
AGB-DE: A Corpus for the Automated Legal Assessment of Clauses in German Consumer Contracts 10 Jun 2024 · 1 repository · arXiv:2406.06809
-
Annotation alignment: Comparing LLM and human annotations of conversational safety 10 Jun 2024 · 0 repositories · arXiv:2406.06369
-
Can Language Models Serve as Text-Based World Simulators? 10 Jun 2024 · 0 repositories · arXiv:2406.06485
-
Compute Better Spent: Replacing Dense Layers with Structured Matrices 10 Jun 2024 · 1 repository · arXiv:2406.06248Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Data-Efficient Learning with Neural Programs 10 Jun 2024 · 1 repository · arXiv:2406.06246Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Diving into Underwater: Segment Anything Model Guided Underwater Salient Instance Segmentation and A Large-scale Dataset 10 Jun 2024 · 1 repository · arXiv:2406.06039
-
Emotion-Aware Speech Self-Supervised Representation Learning with Intensity Knowledge 10 Jun 2024 · 0 repositories · arXiv:2406.06646
-
Husky: A Unified, Open-Source Language Agent for Multi-Step Reasoning 10 Jun 2024 · 1 repository · arXiv:2406.06469Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
In-Context Learning and Fine-Tuning GPT for Argument Mining 10 Jun 2024 · 1 repository · arXiv:2406.06699
-
Learning Physical Simulation with Message Passing Transformer 10 Jun 2024 · 0 repositories · arXiv:2406.06060
-
Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing 10 Jun 2024 · 0 repositories · arXiv:2406.06723
-
LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching 10 Jun 2024 · 0 repositories · arXiv:2406.06799
-
PointABM:Integrating Bidirectional State Space Model with Multi-Head Self-Attention for Point Cloud Analysis 10 Jun 2024 · 0 repositories · arXiv:2406.06069
-
SecureNet: A Comparative Study of DeBERTa and Large Language Models for Phishing Detection 10 Jun 2024 · 0 repositories · arXiv:2406.06663
-
Separate and Reconstruct: Asymmetric Encoder-Decoder for Speech Separation 10 Jun 2024 · 1 repository · arXiv:2406.05983
-
Symmetric Dot-Product Attention for Efficient Training of BERT Language Models 10 Jun 2024 · 0 repositories · arXiv:2406.06366
-
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs 10 Jun 2024 · 0 repositories · arXiv:2406.10251
-
UMBRELA: UMbrela is the (Open-Source Reproduction of the) Bing RELevance Assessor 10 Jun 2024 · 1 repository · arXiv:2406.06519
-
A Knowledge-Component-Based Methodology for Evaluating AI Assistants 9 Jun 2024 · 0 repositories · arXiv:2406.05603
-
Are Large Language Models Actually Good at Text Style Transfer? 9 Jun 2024 · 1 repository · arXiv:2406.05885
-
Attention as a Hypernetwork 9 Jun 2024 · 1 repository · arXiv:2406.05816Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CAMS: Convolution and Attention-Free Mamba-based Cardiac Image Segmentation 9 Jun 2024 · 1 repository · arXiv:2406.05786
-
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation 9 Jun 2024 · 2 repositories · arXiv:2406.05654Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Exploring the Efficacy of Large Language Models (GPT-4) in Binary Reverse Engineering 9 Jun 2024 · 0 repositories · arXiv:2406.06637
-
GCtx-UNet: Efficient Network for Medical Image Segmentation 9 Jun 2024 · 1 repository · arXiv:2406.05891
-
Hidden Holes: topological aspects of language models 9 Jun 2024 · 0 repositories · arXiv:2406.05798
-
Large Language Models Memorize Sensor Datasets! Implications on Human Activity Recognition Research 9 Jun 2024 · 0 repositories · arXiv:2406.05900
-
Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents 9 Jun 2024 · 0 repositories · arXiv:2406.05870
-
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering 9 Jun 2024 · 0 repositories · arXiv:2406.05845
-
OD-DETR: Online Distillation for Stabilizing Training of Detection Transformer 9 Jun 2024 · 0 repositories · arXiv:2406.05791
-
RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation 9 Jun 2024 · 1 repository · arXiv:2406.05794
-
SinkLoRA: Enhanced Efficiency and Chat Capabilities for Long-Context Large Language Models 9 Jun 2024 · 1 repository · arXiv:2406.05678Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Smiles2Dock: an open large-scale multi-task dataset for ML-based molecular docking 9 Jun 2024 · 1 repository · arXiv:2406.05738
-
Text2VP: Generative AI for Visual Programming and Parametric Modeling 9 Jun 2024 · 0 repositories · arXiv:2407.07732
-
Vision Mamba: Cutting-Edge Classification of Alzheimer's Disease with 3D MRI Scans 9 Jun 2024 · 0 repositories · arXiv:2406.05757
-
1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation 8 Jun 2024 · 0 repositories · arXiv:2406.05352