Methods › General › Output Functions › Softmax › Papers, page 44
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 44 of 375: papers 4,301 to 4,400 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition 20 Feb 2025 · 1 repository · arXiv:2502.14195
-
Cardiac Evidence Backtracking for Eating Behavior Monitoring using Collocative Electrocardiogram Imagining 20 Feb 2025 · 0 repositories · arXiv:2502.14430
-
DeepRTL: Bridging Verilog Understanding and Generation with a Unified Representation Model 20 Feb 2025 · 0 repositories · arXiv:2502.15832
-
Designing Parameter and Compute Efficient Diffusion Transformers using Distillation 20 Feb 2025 · 0 repositories · arXiv:2502.14226
-
Do LLMs Consider Security? An Empirical Study on Responses to Programming Questions 20 Feb 2025 · 0 repositories · arXiv:2502.14202
-
Does Time Have Its Place? Temporal Heads: Where Language Models Recall Time-specific Information 20 Feb 2025 · 1 repository · arXiv:2502.14258
-
Earlier Tokens Contribute More: Learning Direct Preference Optimization From Temporal Decay Perspective 20 Feb 2025 · 1 repository · arXiv:2502.14340Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Effects of Prompt Length on Domain-specific Tasks for Large Language Models 20 Feb 2025 · 0 repositories · arXiv:2502.14255
-
Entropy-UID: A Method for Optimizing Information Density 20 Feb 2025 · 0 repositories · arXiv:2502.14366
-
Exploring RWKV for Sentence Embeddings: Layer-wise Analysis and Baseline Comparison for Semantic Similarity 20 Feb 2025 · 1 repository · arXiv:2502.14620
-
FIND: Fine-grained Information Density Guided Adaptive Retrieval-Augmented Generation for Disease Diagnosis 20 Feb 2025 · 0 repositories · arXiv:2502.14614
-
Forecasting Local Ionospheric Parameters Using Transformers 20 Feb 2025 · 1 repository · arXiv:2502.15093
-
From Knowledge Generation to Knowledge Verification: Examining the BioMedical Generative Capabilities of ChatGPT 20 Feb 2025 · 0 repositories · arXiv:2502.14714
-
From RAG to Memory: Non-Parametric Continual Learning for Large Language Models 20 Feb 2025 · 1 repository · arXiv:2502.14802Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Model Inversion Attack against Federated Unlearning 20 Feb 2025 · 0 repositories · arXiv:2502.14558
-
Full-Step-DPO: Self-Supervised Preference Optimization with Step-wise Rewards for Mathematical Reasoning 20 Feb 2025 · 0 repositories · arXiv:2502.14356
-
H3DE-Net: Efficient and Accurate 3D Landmark Detection in Medical Imaging 20 Feb 2025 · 1 repository · arXiv:2502.14221
-
Hallucination Detection in Large Language Models with Metamorphic Relations 20 Feb 2025 · 0 repositories · arXiv:2502.15844
-
Hardware-Friendly Static Quantization Method for Video Diffusion Transformers 20 Feb 2025 · 0 repositories · arXiv:2502.15077
-
How Far are LLMs from Being Our Digital Twins? A Benchmark for Persona-Based Behavior Chain Simulation 20 Feb 2025 · 1 repository · arXiv:2502.14642
-
Is Relevance Propagated from Retriever to Generator in RAG? 20 Feb 2025 · 0 repositories · arXiv:2502.15025
-
KITAB-Bench: A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understanding 20 Feb 2025 · 0 repositories · arXiv:2502.14949
-
LIFT: Improving Long Context Understanding of Large Language Models through Long Input Fine-Tuning 20 Feb 2025 · 0 repositories · arXiv:2502.14644
-
LServe: Efficient Long-sequence LLM Serving with Unified Sparse Attention 20 Feb 2025 · 2 repositories · arXiv:2502.14866Syntology official: harvested, nothing ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Mechanistic Understanding of Language Models in Syntactic Code Completion 20 Feb 2025 · 0 repositories · arXiv:2502.18499
-
Multiscale Byte Language Models -- A Hierarchical Architecture for Causal Million-Length Sequence Modeling 20 Feb 2025 · 1 repository · arXiv:2502.14553
-
NeRF-3DTalker: Neural Radiance Field with 3D Prior Aided Audio Disentanglement for Talking Head Synthesis 20 Feb 2025 · 0 repositories · arXiv:2502.14178
-
On the Influence of Context Size and Model Choice in Retrieval-Augmented Generation Systems 20 Feb 2025 · 1 repository · arXiv:2502.14759
-
PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant 20 Feb 2025 · 0 repositories · arXiv:2502.14271
-
ParallelComp: Parallel Long-Context Compressor for Length Extrapolation 20 Feb 2025 · 0 repositories · arXiv:2502.14317
-
PLPHP: Per-Layer Per-Head Vision Token Pruning for Efficient Large Vision-Language Models 20 Feb 2025 · 0 repositories · arXiv:2502.14504
-
Predicting Fetal Birthweight from High Dimensional Data using Advanced Machine Learning 20 Feb 2025 · 0 repositories · arXiv:2502.14270
-
QUAD-LLM-MLTC: Large Language Models Ensemble Learning for Healthcare Text Multi-Label Classification 20 Feb 2025 · 0 repositories · arXiv:2502.14189
-
Reducing false positives in strong lens detection through effective augmentation and ensemble learning 20 Feb 2025 · 0 repositories · arXiv:2502.14936
-
Reinforcement Learning with Graph Attention for Routing and Wavelength Assignment with Lightpath Reuse 20 Feb 2025 · 0 repositories · arXiv:2502.14741
-
RelaCtrl: Relevance-Guided Efficient Control for Diffusion Transformers 20 Feb 2025 · 0 repositories · arXiv:2502.14377
-
RendBEV: Semantic Novel View Synthesis for Self-Supervised Bird's Eye View Segmentation 20 Feb 2025 · 0 repositories · arXiv:2502.14792
-
Revealing and Mitigating Over-Attention in Knowledge Editing 20 Feb 2025 · 1 repository · arXiv:2502.14838
-
Role of the Pretraining and the Adaptation data sizes for low-resource real-time MRI video segmentation 20 Feb 2025 · 0 repositories · arXiv:2502.14418
-
Tabular Embeddings for Tables with Bi-Dimensional Hierarchical Metadata and Nesting 20 Feb 2025 · 0 repositories · arXiv:2502.15819
-
Textured 3D Regenerative Morphing with 3D Diffusion Prior 20 Feb 2025 · 0 repositories · arXiv:2502.14316
-
Topology-Aware Wavelet Mamba for Airway Structure Segmentation in Postoperative Recurrent Nasopharyngeal Carcinoma CT Scans 20 Feb 2025 · 0 repositories · arXiv:2502.14363
-
Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-based LLMs 20 Feb 2025 · 1 repository · arXiv:2502.14837Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples)
-
Unshackling Context Length: An Efficient Selective Attention Approach through Query-Key Compression 20 Feb 2025 · 0 repositories · arXiv:2502.14477
-
WavRAG: Audio-Integrated Retrieval Augmented Generation for Spoken Dialogue Models 20 Feb 2025 · 0 repositories · arXiv:2502.14727
-
A consensus set for the aggregation of partial rankings: the case of the Optimal Set of Bucket Orders Problem 19 Feb 2025 · 0 repositories · arXiv:2502.13769
-
Activation-aware Probe-Query: Effective Key-Value Retrieval for Long-Context LLMs Inference 19 Feb 2025 · 0 repositories · arXiv:2502.13542
-
Adapting Large Language Models for Time Series Modeling via a Novel Parameter-efficient Adaptation Method 19 Feb 2025 · 0 repositories · arXiv:2502.13725
-
Are Large Language Models In-Context Graph Learners? 19 Feb 2025 · 0 repositories · arXiv:2502.13562
-
Building Age Estimation: A New Multi-Modal Benchmark Dataset and Community Challenge 19 Feb 2025 · 1 repository · arXiv:2502.13818
-
Capturing Rich Behavior Representations: A Dynamic Action Semantic-Aware Graph Transformer for Video Captioning 19 Feb 2025 · 0 repositories · arXiv:2502.13754
-
CARE: Confidence-Aware Regression Estimation of building density fine-tuning EO Foundation Models 19 Feb 2025 · 0 repositories · arXiv:2502.13734
-
Contrastive Learning-Based privacy metrics in Tabular Synthetic Datasets 19 Feb 2025 · 1 repository · arXiv:2502.13833
-
Conveniently Identify Coils in Inductive Power Transfer System Using Machine Learning 19 Feb 2025 · 0 repositories · arXiv:2502.13915
-
DH-RAG: A Dynamic Historical Context-Powered Retrieval-Augmented Generation Method for Multi-Turn Dialogue 19 Feb 2025 · 0 repositories · arXiv:2502.13847
-
Diffusion Model Agnostic Social Influence Maximization in Hyperbolic Space 19 Feb 2025 · 0 repositories · arXiv:2502.13571
-
Extracting Social Connections from Finnish Karelian Refugee Interviews Using LLMs 19 Feb 2025 · 0 repositories · arXiv:2502.13566
-
FairKV: Balancing Per-Head KV Cache for Fast Multi-GPU Inference 19 Feb 2025 · 0 repositories · arXiv:2502.15804
-
FlexTok: Resampling Images into 1D Token Sequences of Flexible Length 19 Feb 2025 · 0 repositories · arXiv:2502.13967
-
From Correctness to Comprehension: AI Agents for Personalized Error Diagnosis in Education 19 Feb 2025 · 0 repositories · arXiv:2502.13789
-
Generative Detail Enhancement for Physically Based Materials 19 Feb 2025 · 0 repositories · arXiv:2502.13994
-
GIMMICK -- Globally Inclusive Multimodal Multitask Cultural Knowledge Benchmarking 19 Feb 2025 · 0 repositories · arXiv:2502.13766
-
Giving AI Personalities Leads to More Human-Like Reasoning 19 Feb 2025 · 0 repositories · arXiv:2502.14155
-
HawkBench: Investigating Resilience of RAG Methods on Stratified Information-Seeking Tasks 19 Feb 2025 · 0 repositories · arXiv:2502.13465
-
Helix-mRNA: A Hybrid Foundation Model For Full Sequence mRNA Therapeutics 19 Feb 2025 · 1 repository · arXiv:2502.13785
-
Hidden Darkness in LLM-Generated Designs: Exploring Dark Patterns in Ecommerce Web Components Generated by LLMs 19 Feb 2025 · 0 repositories · arXiv:2502.13499
-
In-Place Updates of a Graph Index for Streaming Approximate Nearest Neighbor Search 19 Feb 2025 · 0 repositories · arXiv:2502.13826
-
Inner Thinking Transformer: Leveraging Dynamic Depth Scaling to Foster Adaptive Internal Thinking 19 Feb 2025 · 0 repositories · arXiv:2502.13842
-
Integration of Agentic AI with 6G Networks for Mission-Critical Applications: Use-case and Challenges 19 Feb 2025 · 0 repositories · arXiv:2502.13476
-
Learning Novel Transformer Architecture for Time-series Forecasting 19 Feb 2025 · 0 repositories · arXiv:2502.13721
-
MambaLiteSR: Image Super-Resolution with Low-Rank Mamba using Knowledge Distillation 19 Feb 2025 · 0 repositories · arXiv:2502.14090
-
MaskPrune: Mask-based LLM Pruning for Layer-wise Uniform Structures 19 Feb 2025 · 0 repositories · arXiv:2502.14008
-
Medical Image Classification with KAN-Integrated Transformers and Dilated Neighborhood Attention 19 Feb 2025 · 1 repository · arXiv:2502.13693
-
Modeling Behavior Change for Multi-model At-Risk Students Early Prediction (extended version) 19 Feb 2025 · 0 repositories · arXiv:2503.05734
-
ModSkill: Physical Character Skill Modularization 19 Feb 2025 · 0 repositories · arXiv:2502.14140
-
MoM: Linear Sequence Modeling with Mixture-of-Memories 19 Feb 2025 · 2 repositories · arXiv:2502.13685Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
MuDAF: Long-Context Multi-Document Attention Focusing through Contrastive Learning on Attention Heads 19 Feb 2025 · 1 repository · arXiv:2502.13963Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples) · 7 pointer-only (licence)
-
PitVQA++: Vector Matrix-Low-Rank Adaptation for Open-Ended Visual Question Answering in Pituitary Surgery 19 Feb 2025 · 1 repository · arXiv:2502.14149
-
PLDR-LLMs Learn A Generalizable Tensor Operator That Can Replace Its Own Deep Neural Net At Inference 19 Feb 2025 · 1 repository · arXiv:2502.13502
-
Qwen2.5-VL Technical Report 19 Feb 2025 · 4 repositories · arXiv:2502.13923Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
RAG-Gym: Optimizing Reasoning and Search Agents with Process Supervision 19 Feb 2025 · 0 repositories · arXiv:2502.13957
-
RAPTOR: Refined Approach for Product Table Object Recognition 19 Feb 2025 · 0 repositories · arXiv:2502.14918
-
Rectified Lagrangian for Out-of-Distribution Detection in Modern Hopfield Networks 19 Feb 2025 · 0 repositories · arXiv:2502.14003
-
Reproducing NevIR: Negation in Neural Information Retrieval 19 Feb 2025 · 2 repositories · arXiv:2502.13506
-
RGAR: Recurrence Generation-augmented Retrieval for Factual-aware Medical Question Answering 19 Feb 2025 · 0 repositories · arXiv:2502.13361
-
RocketKV: Accelerating Long-Context LLM Inference via Two-Stage KV Cache Compression 19 Feb 2025 · 0 repositories · arXiv:2502.14051Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Spiking Point Transformer for Point Cloud Classification 19 Feb 2025 · 1 repository · arXiv:2502.15811
-
STaR-SQL: Self-Taught Reasoner for Text-to-SQL 19 Feb 2025 · 0 repositories · arXiv:2502.13550
-
The Risk-Neutral Equivalent Pricing of Model-Uncertainty 19 Feb 2025 · 0 repositories · arXiv:2502.13744
-
Token Adaptation via Side Graph Convolution for Temporally and Spatially Efficient Fine-tuning of 3D Point Cloud Transformers 19 Feb 2025 · 1 repository · arXiv:2502.14142
-
Toward Robust Non-Transferable Learning: A Survey and Benchmark 19 Feb 2025 · 1 repository · arXiv:2502.13593
-
Towards Adaptive Memory-Based Optimization for Enhanced Retrieval-Augmented Generation 19 Feb 2025 · 0 repositories · arXiv:2504.05312
-
TrustRAG: An Information Assistant with Retrieval Augmented Generation 19 Feb 2025 · 1 repository · arXiv:2502.13719
-
UNGT: Ultrasound Nasogastric Tube Dataset for Medical Image Analysis 19 Feb 2025 · 0 repositories · arXiv:2502.14915
-
Universal Semantic Embeddings of Chemical Elements for Enhanced Materials Inference and Discovery 19 Feb 2025 · 0 repositories · arXiv:2502.14912
-
What are Models Thinking about? Understanding Large Language Model Hallucinations "Psychology" through Model Inner State Analysis 19 Feb 2025 · 0 repositories · arXiv:2502.13490
-
Where's the Bug? Attention Probing for Scalable Fault Localization 19 Feb 2025 · 0 repositories · arXiv:2502.13966
-
A²ATS: Retrieval-Based KV Cache Reduction via Windowed Rotary Position Embedding and Query-Aware Vector Quantization 18 Feb 2025 · 0 repositories · arXiv:2502.12665
-
A Survey of Sim-to-Real Methods in RL: Progress, Prospects and Challenges with Foundation Models 18 Feb 2025 · 0 repositories · arXiv:2502.13187
-
An Attention-Assisted Multi-Modal Data Fusion Model for Real-Time Estimation of Underwater Sound Velocity 18 Feb 2025 · 0 repositories · arXiv:2502.12817