Methods › General › Output Functions › Softmax › Papers, page 42
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 42 of 375: papers 4,101 to 4,200 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Detecting Knowledge Boundary of Vision Large Language Models by Sampling-Based Inference 25 Feb 2025 · 0 repositories · arXiv:2502.18023
-
Dual Classification Head Self-training Network for Cross-scene Hyperspectral Image Classification 25 Feb 2025 · 0 repositories · arXiv:2502.17879
-
Enhancing Speech Quality through the Integration of BGRU and Transformer Architectures 25 Feb 2025 · 0 repositories · arXiv:2502.17911
-
Enhancing Text Classification with a Novel Multi-Agent Collaboration Framework Leveraging BERT 25 Feb 2025 · 0 repositories · arXiv:2502.18653
-
Examining the Threat Landscape: Foundation Models and Model Stealing 25 Feb 2025 · 0 repositories · arXiv:2502.18077
-
Faster, Cheaper, Better: Multi-Objective Hyperparameter Optimization for LLM and RAG Systems 25 Feb 2025 · 0 repositories · arXiv:2502.18635
-
FoREST: Frame of Reference Evaluation in Spatial Reasoning Tasks 25 Feb 2025 · 1 repository · arXiv:2502.17775
-
FwNet-ECA: Facilitating Window Attention with Global Receptive Fields through Fourier Filtering Operations 25 Feb 2025 · 1 repository · arXiv:2502.18094
-
GHOST 2.0: generative high-fidelity one shot transfer of heads 25 Feb 2025 · 0 repositories · arXiv:2502.18417
-
Graph Inference with Effective Resistance Queries 25 Feb 2025 · 0 repositories · arXiv:2502.18350
-
H-FLTN: A Privacy-Preserving Hierarchical Framework for Electric Vehicle Spatio-Temporal Charge Prediction 25 Feb 2025 · 0 repositories · arXiv:2502.18697
-
Harnessing Multiple Large Language Models: A Survey on LLM Ensemble 25 Feb 2025 · 1 repository · arXiv:2502.18036
-
How Vital is the Jurisprudential Relevance: Law Article Intervened Legal Case Retrieval and Matching 25 Feb 2025 · 0 repositories · arXiv:2502.18292
-
Improved YOLOv7x-Based Defect Detection Algorithm for Power Equipment 25 Feb 2025 · 0 repositories · arXiv:2502.17961
-
Independent Mobility GPT (IDM-GPT): A Self-Supervised Multi-Agent Large Language Model Framework for Customized Traffic Mobility Analysis Using Machine Learning Models 25 Feb 2025 · 0 repositories · arXiv:2502.18652
-
InVDriver: Intra-Instance Aware Vectorized Query-Based Autonomous Driving Transformer 25 Feb 2025 · 0 repositories · arXiv:2502.17949
-
K-LoRA: Unlocking Training-Free Fusion of Any Subject and Style LoRAs 25 Feb 2025 · 0 repositories · arXiv:2502.18461
-
LAM: Large Avatar Model for One-shot Animatable Gaussian Head 25 Feb 2025 · 0 repositories · arXiv:2502.17796
-
LDGen: Enhancing Text-to-Image Synthesis via Large Language Model-Driven Language Representation 25 Feb 2025 · 0 repositories · arXiv:2502.18302
-
Learning Structure-Supporting Dependencies via Keypoint Interactive Transformer for General Mammal Pose Estimation 25 Feb 2025 · 1 repository · arXiv:2502.18214
-
LevelRAG: Enhancing Retrieval-Augmented Generation with Multi-hop Logic Planning over Rewriting Augmented Searchers 25 Feb 2025 · 1 repository · arXiv:2502.18139
-
LightFC-X: Lightweight Convolutional Tracker for RGB-X Tracking 25 Feb 2025 · 1 repository · arXiv:2502.18143
-
MAGE: Multi-Head Attention Guided Embeddings for Low Resource Sentiment Classification 25 Feb 2025 · 0 repositories · arXiv:2502.17987
-
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks 25 Feb 2025 · 1 repository · arXiv:2502.17832
-
MuCoS: Efficient Drug-Target Prediction through Multi-Context-Aware Sampling 25 Feb 2025 · 0 repositories · arXiv:2502.17784
-
Neural Network Graph Similarity Computation Based on Graph Fusion 25 Feb 2025 · 1 repository · arXiv:2502.18291
-
Opus: A Workflow Intention Framework for Complex Workflow Generation 25 Feb 2025 · 0 repositories · arXiv:2502.19532
-
RefuteBench 2.0 -- Agentic Benchmark for Dynamic Evaluation of LLM Responses to Refutation Instruction 25 Feb 2025 · 0 repositories · arXiv:2502.18308
-
Say Less, Mean More: Leveraging Pragmatics in Retrieval-Augmented Generation 25 Feb 2025 · 0 repositories · arXiv:2502.17839
-
Scaling LLM Pre-training with Vocabulary Curriculum 25 Feb 2025 · 0 repositories · arXiv:2502.17910
-
Self-Adjust Softmax 25 Feb 2025 · 0 repositories · arXiv:2502.18277
-
Software implemented fault diagnosis of natural gas pumping unit based on feedforward neural network 25 Feb 2025 · 0 repositories · arXiv:2502.18233
-
SpargeAttention: Accurate and Training-free Sparse Attention Accelerating Any Model Inference 25 Feb 2025 · 1 repository · arXiv:2502.18137
-
Stackelberg Game Preference Optimization for Data-Efficient Alignment of Language Models 25 Feb 2025 · 0 repositories · arXiv:2502.18099
-
Synthesizing Consistent Novel Views via 3D Epipolar Attention without Re-Training 25 Feb 2025 · 0 repositories · arXiv:2502.18219
-
Systems and Algorithms for Convolutional Multi-Hybrid Language Models at Scale 25 Feb 2025 · 0 repositories · arXiv:2503.01868
-
SPECTRE: An FFT-Based Efficient Drop-In Replacement to Self-Attention for Long Contexts 25 Feb 2025 · 2 repositories · arXiv:2502.18394
-
VesselSAM: Leveraging SAM for Aortic Vessel Segmentation with AtrousLoRA 25 Feb 2025 · 1 repository · arXiv:2502.18185
-
ViDoRAG: Visual Document Retrieval-Augmented Generation via Dynamic Iterative Reasoning Agents 25 Feb 2025 · 1 repository · arXiv:2502.18017Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Weakly Supervised Pixel-Level Annotation with Visual Interpretability 25 Feb 2025 · 0 repositories · arXiv:2502.17824
-
A Concise Lyapunov Analysis of Nesterov's Accelerated Gradient Method 24 Feb 2025 · 0 repositories · arXiv:2502.17373
-
A novel approach to navigate the taxonomic hierarchy to address the Open-World Scenarios in Medicinal Plant Classification 24 Feb 2025 · 0 repositories · arXiv:2502.17289
-
A Pragmatic Note on Evaluating Generative Models with Fréchet Inception Distance for Retinal Image Synthesis 24 Feb 2025 · 0 repositories · arXiv:2502.17160
-
A Transformer-in-Transformer Network Utilizing Knowledge Distillation for Image Recognition 24 Feb 2025 · 0 repositories · arXiv:2502.16762
-
"Actionable Help" in Crises: A Novel Dataset and Resource-Efficient Models for Identifying Request and Offer Social Media Posts 24 Feb 2025 · 0 repositories · arXiv:2502.16839
-
Adversarial Training for Defense Against Label Poisoning Attacks 24 Feb 2025 · 1 repository · arXiv:2502.17121Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
AnyTop: Character Animation Diffusion with Any Topology 24 Feb 2025 · 1 repository · arXiv:2502.17327
-
Applying LLMs to Active Learning: Towards Cost-Efficient Cross-Task Text Classification without Manually Labeled Data 24 Feb 2025 · 0 repositories · arXiv:2502.16892
-
Are Large Language Models Good Data Preprocessors? 24 Feb 2025 · 0 repositories · arXiv:2502.16790
-
Atten-Transformer: A Deep Learning Framework for User App Usage Prediction 24 Feb 2025 · 0 repositories · arXiv:2502.16957
-
Benchmarking Retrieval-Augmented Generation in Multi-Modal Contexts 24 Feb 2025 · 2 repositories · arXiv:2502.17297
-
CalibRefine: Deep Learning-Based Online Automatic Targetless LiDAR-Camera Calibration with Iterative and Attention-Driven Post-Refinement 24 Feb 2025 · 1 repository · arXiv:2502.17648
-
Child vs. machine language learning: Can the logical structure of human language unleash LLMs? 24 Feb 2025 · 0 repositories · arXiv:2502.17304
-
CipherPrune: Efficient and Scalable Private Transformer Inference 24 Feb 2025 · 1 repository · arXiv:2502.16782Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
CLEP-GAN: An Innovative Approach to Subject-Independent ECG Reconstruction from PPG Signals 24 Feb 2025 · 0 repositories · arXiv:2502.17536
-
DBudgetKV: Dynamic Budget in KV Cache Compression for Ensuring Optimal Performance 24 Feb 2025 · 0 repositories · arXiv:2502.16886
-
Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation 24 Feb 2025 · 0 repositories · arXiv:2502.17198
-
Disentangling Visual Transformers: Patch-level Interpretability for Image Classification 24 Feb 2025 · 0 repositories · arXiv:2502.17196
-
ENACT-Heart -- ENsemble-based Assessment Using CNN and Transformer on Heart Sounds 24 Feb 2025 · 0 repositories · arXiv:2502.16914
-
Enhancing Image Matting in Real-World Scenes with Mask-Guided Iterative Refinement 24 Feb 2025 · 0 repositories · arXiv:2502.17093
-
Erwin: A Tree-based Hierarchical Transformer for Large-scale Physical Systems 24 Feb 2025 · 2 repositories · arXiv:2502.17019Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Evaluating the Effect of Retrieval Augmentation on Social Biases 24 Feb 2025 · 0 repositories · arXiv:2502.17611
-
Functional BART with Shape Priors: A Bayesian Tree Approach to Constrained Functional Regression 24 Feb 2025 · 0 repositories · arXiv:2502.16888
-
Gabor-Enhanced Physics-Informed Neural Networks for Fast Simulations of Acoustic Wavefields 24 Feb 2025 · 1 repository · arXiv:2502.17134
-
GaussianFlowOcc: Sparse and Weakly Supervised Occupancy Estimation using Gaussian Splatting and Temporal Flow 24 Feb 2025 · 0 repositories · arXiv:2502.17288
-
GuidedBench: Equipping Jailbreak Evaluation with Guidelines 24 Feb 2025 · 0 repositories · arXiv:2502.16903
-
Hallucination Detection in LLMs Using Spectral Features of Attention Maps 24 Feb 2025 · 1 repository · arXiv:2502.17598
-
Improving the Transferability of Adversarial Examples by Inverse Knowledge Distillation 24 Feb 2025 · 0 repositories · arXiv:2502.17003
-
LettuceDetect: A Hallucination Detection Framework for RAG Applications 24 Feb 2025 · 2 repositories · arXiv:2502.17125
-
LLM Inference Acceleration via Efficient Operation Fusion 24 Feb 2025 · 0 repositories · arXiv:2502.17728
-
Logic Haystacks: Probing LLMs Long-Context Logical Reasoning (Without Easily Identifiable Unrelated Padding) 24 Feb 2025 · 0 repositories · arXiv:2502.17169
-
LongSafety: Evaluating Long-Context Safety of Large Language Models 24 Feb 2025 · 1 repository · arXiv:2502.16971Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
LongSpec: Long-Context Speculative Decoding with Efficient Drafting and Verification 24 Feb 2025 · 1 repository · arXiv:2502.17421Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
MambaFlow: A Novel and Flow-guided State Space Model for Scene Flow Estimation 24 Feb 2025 · 1 repository · arXiv:2502.16907Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MaxGlaViT: A novel lightweight vision transformer-based approach for early diagnosis of glaucoma stages from fundus images 24 Feb 2025 · 1 repository · arXiv:2502.17154
-
MDN: Mamba-Driven Dualstream Network For Medical Hyperspectral Image Segmentation 24 Feb 2025 · 0 repositories · arXiv:2502.17255
-
MEDA: Dynamic KV Cache Allocation for Efficient Multimodal Long-Context Inference 24 Feb 2025 · 1 repository · arXiv:2502.17599Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
MEMERAG: A Multilingual End-to-End Meta-Evaluation Benchmark for Retrieval Augmented Generation 24 Feb 2025 · 1 repository · arXiv:2502.17163
-
Mitigating Bias in RAG: Controlling the Embedder 24 Feb 2025 · 1 repository · arXiv:2502.17390
-
Mitigating Hallucinations in Diffusion Models through Adaptive Attention Modulation 24 Feb 2025 · 0 repositories · arXiv:2502.16872
-
MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs 24 Feb 2025 · 1 repository · arXiv:2502.17422Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Mutual Reinforcement of LLM Dialogue Synthesis and Summarization Capabilities for Few-Shot Dialogue Summarization 24 Feb 2025 · 0 repositories · arXiv:2502.17328
-
Neural Attention: A Novel Mechanism for Enhanced Expressive Power in Transformer Models 24 Feb 2025 · 0 repositories · arXiv:2502.17206
-
NUTSHELL: A Dataset for Abstract Generation from Scientific Talks 24 Feb 2025 · 0 repositories · arXiv:2502.16942
-
Optimal Recovery Meets Minimax Estimation 24 Feb 2025 · 0 repositories · arXiv:2502.17671
-
Order Matters: Investigate the Position Bias in Multi-constraint Instruction Following 24 Feb 2025 · 1 repository · arXiv:2502.17204
-
Quantifying Logical Consistency in Transformers via Query-Key Alignment 24 Feb 2025 · 0 repositories · arXiv:2502.17017
-
Shakti-VLMs: Scalable Vision-Language Models for Enterprise AI 24 Feb 2025 · 0 repositories · arXiv:2502.17092
-
TabulaTime: A Novel Multimodal Deep Learning Framework for Advancing Acute Coronary Syndrome Prediction through Environmental and Clinical Data Integration 24 Feb 2025 · 0 repositories · arXiv:2502.17049
-
The Lottery LLM Hypothesis, Rethinking What Abilities Should LLM Compression Preserve? 24 Feb 2025 · 0 repositories · arXiv:2502.17535
-
The Role of Sparsity for Length Generalization in Transformers 24 Feb 2025 · 0 repositories · arXiv:2502.16792
-
Towards Typologically Aware Rescoring to Mitigate Unfaithfulness in Lower-Resource Languages 24 Feb 2025 · 0 repositories · arXiv:2502.17664
-
Unraveling the geometry of visual relational reasoning 24 Feb 2025 · 1 repository · arXiv:2502.17382
-
VGFL-SA: Vertical Graph Federated Learning Structure Attack Based on Contrastive Learning 24 Feb 2025 · 0 repositories · arXiv:2502.16793
-
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing 24 Feb 2025 · 0 repositories · arXiv:2502.17258
-
VR-Pipe: Streamlining Hardware Graphics Pipeline for Volume Rendering 24 Feb 2025 · 0 repositories · arXiv:2502.17078
-
X-Dancer: Expressive Music to Human Dance Video Generation 24 Feb 2025 · 0 repositories · arXiv:2502.17414
-
A Fine-Tuning Approach for T5 Using Knowledge Graphs to Address Complex Tasks 23 Feb 2025 · 0 repositories · arXiv:2502.16484
-
A Reverse Mamba Attention Network for Pathological Liver Segmentation 23 Feb 2025 · 1 repository · arXiv:2502.18232
-
A Split-Window Transformer for Multi-Model Sequence Spammer Detection using Multi-Model Variational Autoencoder 23 Feb 2025 · 0 repositories · arXiv:2502.16483