Methods › General › Output Functions › Softmax › Papers, page 78
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 78 of 375: papers 7,701 to 7,800 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
DMQR-RAG: Diverse Multi-Query Rewriting for RAG 20 Nov 2024 · 0 repositories · arXiv:2411.13154
-
DrugGen: Advancing Drug Discovery with Large Language Models and Reinforcement Learning Feedback 20 Nov 2024 · 4 repositories · arXiv:2411.14157
-
Exploring Large Language Models for Climate Forecasting 20 Nov 2024 · 0 repositories · arXiv:2411.13724
-
Human Age and Gender Prediction Management system project report 20 Nov 2024 · 0 repositories
-
Hymba: A Hybrid-head Architecture for Small Language Models 20 Nov 2024 · 0 repositories · arXiv:2411.13676
-
Learning to Reason Iteratively and Parallelly for Complex Visual Reasoning Scenarios 20 Nov 2024 · 0 repositories · arXiv:2411.13754
-
LLMSteer: Improving Long-Context LLM Inference by Steering Attention on Reused Contexts 20 Nov 2024 · 0 repositories · arXiv:2411.13009
-
M2oE: Multimodal Collaborative Expert Peptide Model 20 Nov 2024 · 1 repository · arXiv:2411.15208
-
MAS-Attention: Memory-Aware Stream Processing for Attention Acceleration on Resource-Constrained Edge Devices 20 Nov 2024 · 0 repositories · arXiv:2411.17720
-
MemoryFormer: Minimize Transformer Computation by Removing Fully-Connected Layers 20 Nov 2024 · 0 repositories · arXiv:2411.12992Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Multimodal large language model for wheat breeding: a new exploration of smart breeding 20 Nov 2024 · 0 repositories · arXiv:2411.15203
-
Multipath Mitigation Technology-integrated GNSS Direct Position Estimation Plug-in Module 20 Nov 2024 · 0 repositories · arXiv:2411.13339
-
Limitations of Automatic Relevance Assessments with Large Language Models for Fair and Reliable Retrieval Evaluation 20 Nov 2024 · 0 repositories · arXiv:2411.13212
-
On the Way to LLM Personalization: Learning to Remember User Conversations 20 Nov 2024 · 0 repositories · arXiv:2411.13405
-
Paying more attention to local contrast: improving infrared small target detection performance via prior knowledge 20 Nov 2024 · 0 repositories · arXiv:2411.13260
-
Practical Compact Deep Compressed Sensing 20 Nov 2024 · 1 repository · arXiv:2411.13081
-
Quantum Attention for Vision Transformers in High Energy Physics 20 Nov 2024 · 0 repositories · arXiv:2411.13520
-
Retrieval-Augmented Generation for Domain-Specific Question Answering: A Case Study on Pittsburgh and CMU 20 Nov 2024 · 0 repositories · arXiv:2411.13691
-
RobustFormer: Noise-Robust Pre-training for images and videos 20 Nov 2024 · 0 repositories · arXiv:2411.13040
-
Scaling Laws for Online Advertisement Retrieval 20 Nov 2024 · 0 repositories · arXiv:2411.13322
-
The Impossible Test: A 2024 Unsolvable Dataset and A Chance for an AGI Quiz 20 Nov 2024 · 0 repositories · arXiv:2411.14486
-
Transformers with Sparse Attention for Granger Causality 20 Nov 2024 · 0 repositories · arXiv:2411.13264
-
Unlocking Historical Clinical Trial Data with ALIGN: A Compositional Large Language Model System for Medical Coding 20 Nov 2024 · 0 repositories · arXiv:2411.13163
-
Verifying Machine Unlearning with Explainable AI 20 Nov 2024 · 1 repository · arXiv:2411.13332
-
When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training 20 Nov 2024 · 1 repository · arXiv:2411.13476
-
A Combined Encoder and Transformer Approach for Coherent and High-Quality Text Generation 19 Nov 2024 · 0 repositories · arXiv:2411.12157
-
A Full-History Network Dataset for BTC Asset Decentralization Profiling 19 Nov 2024 · 0 repositories · arXiv:2411.13603
-
Action-Attentive Deep Reinforcement Learning for Autonomous Alignment of Beamlines 19 Nov 2024 · 1 repository · arXiv:2411.12183
-
Adaptively Controllable Diffusion Model for Efficient Conditional Image Generation 19 Nov 2024 · 0 repositories · arXiv:2411.15199
-
Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction 19 Nov 2024 · 1 repository · arXiv:2411.17835
-
Benchmarking Positional Encodings for GNNs and Graph Transformers 19 Nov 2024 · 1 repository · arXiv:2411.12732Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Can ChatGPT Overcome Behavioral Biases in the Financial Sector? Classify-and-Rethink: Multi-Step Zero-Shot Reasoning in the Gold Investment 19 Nov 2024 · 0 repositories · arXiv:2411.13599
-
CCIS-Diff: A Generative Model with Stable Diffusion Prior for Controlled Colonoscopy Image Synthesis 19 Nov 2024 · 0 repositories · arXiv:2411.12198
-
Classification of Geographical Land Structure Using Convolution Neural Network and Transfer Learning 19 Nov 2024 · 0 repositories · arXiv:2411.12415
-
Comparing Prior and Learned Time Representations in Transformer Models of Timeseries 19 Nov 2024 · 0 repositories · arXiv:2411.12476
-
Cross-Layer Encrypted Semantic Communication Framework for Panoramic Video Transmission 19 Nov 2024 · 0 repositories · arXiv:2411.12776
-
Deep Learning-Based Classification of Hyperkinetic Movement Disorders in Children 19 Nov 2024 · 0 repositories · arXiv:2411.15200
-
DLBacktrace: A Model Agnostic Explainability for any Deep Learning Models 19 Nov 2024 · 1 repository · arXiv:2411.12643
-
Enhancing Low Dose Computed Tomography Images Using Consistency Training Techniques 19 Nov 2024 · 0 repositories · arXiv:2411.12181
-
Enhancing Multi-Class Disease Classification: Neoplasms, Cardiovascular, Nervous System, and Digestive Disorders Using Advanced LLMs 19 Nov 2024 · 0 repositories · arXiv:2411.12712
-
Evaluating Tokenizer Performance of Large Language Models Across Official Indian Languages 19 Nov 2024 · 0 repositories · arXiv:2411.12240
-
Faster Multi-GPU Training with PPLL: A Pipeline Parallelism Framework Leveraging Local Learning 19 Nov 2024 · 0 repositories · arXiv:2411.12780
-
From Centralized RAN to Open RAN: A Survey on the Evolution of Distributed Antenna Systems 19 Nov 2024 · 0 repositories · arXiv:2411.12166
-
Graph Neural Network-Based Entity Extraction and Relationship Reasoning in Complex Knowledge Graphs 19 Nov 2024 · 0 repositories · arXiv:2411.15195
-
HEIGHT: Heterogeneous Interaction Graph Transformer for Robot Navigation in Crowded and Constrained Environments 19 Nov 2024 · 0 repositories · arXiv:2411.12150
-
Hypergraph p-Laplacian equations for data interpolation and semi-supervised learning 19 Nov 2024 · 0 repositories · arXiv:2411.12601
-
Leveraging Virtual Reality and AI Tutoring for Language Learning: A Case Study of a Virtual Campus Environment with OpenAI GPT Integration with Unity 3D 19 Nov 2024 · 0 repositories · arXiv:2411.12619
-
Med-2E3: A 2D-Enhanced 3D Medical Multimodal Large Language Model 19 Nov 2024 · 0 repositories · arXiv:2411.12783
-
Multi-Grained Preference Enhanced Transformer for Multi-Behavior Sequential Recommendation 19 Nov 2024 · 1 repository · arXiv:2411.12179
-
PoM: Efficient Image and Video Generation with the Polynomial Mixer 19 Nov 2024 · 1 repository · arXiv:2411.12663
-
Predicting User Intents and Musical Attributes from Music Discovery Conversations 19 Nov 2024 · 1 repository · arXiv:2411.12254
-
Residual Vision Transformer (ResViT) Based Self-Supervised Learning Model for Brain Tumor Classification 19 Nov 2024 · 0 repositories · arXiv:2411.12874
-
Robust 3D Semantic Occupancy Prediction with Calibration-free Spatial Transformation 19 Nov 2024 · 1 repository · arXiv:2411.12177
-
S3TU-Net: Structured Convolution and Superpixel Transformer for Lung Nodule Segmentation 19 Nov 2024 · 0 repositories · arXiv:2411.12547
-
Selective Attention: Enhancing Transformer through Principled Context Control 19 Nov 2024 · 1 repository · arXiv:2411.12892Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Self-Supervised Learning in Deep Networks: A Pathway to Robust Few-Shot Classification 19 Nov 2024 · 0 repositories · arXiv:2411.12151
-
Signformer is all you need: Towards Edge AI for Sign Language 19 Nov 2024 · 1 repository · arXiv:2411.12901
-
Strengthening Fake News Detection: Leveraging SVM and Sophisticated Text Vectorization Techniques. Defying BERT? 19 Nov 2024 · 0 repositories · arXiv:2411.12703
-
Transformer Neural Processes -- Kernel Regression 19 Nov 2024 · 0 repositories · arXiv:2411.12502
-
Ultra-Sparse Memory Network 19 Nov 2024 · 0 repositories · arXiv:2411.12364
-
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations 19 Nov 2024 · 0 repositories · arXiv:2411.12701
-
Advacheck at GenAI Detection Task 1: AI Detection Powered by Domain-Aware Multi-Tasking 18 Nov 2024 · 1 repository · arXiv:2411.11736
-
Attention-guided Spectrogram Sequence Modeling with CNNs for Music Genre Classification 18 Nov 2024 · 0 repositories · arXiv:2411.14474
-
BeautyBank: Encoding Facial Makeup in Latent Space 18 Nov 2024 · 0 repositories · arXiv:2411.11231
-
Can Open-source LLMs Enhance Data Synthesis for Toxic Detection?: An Experimental Study 18 Nov 2024 · 0 repositories · arXiv:2411.15175
-
Chapter 7 Review of Data-Driven Generative AI Models for Knowledge Extraction from Scientific Literature in Healthcare 18 Nov 2024 · 0 repositories · arXiv:2411.11635
-
CNMBERT: A Model for Converting Hanyu Pinyin Abbreviations to Chinese Characters 18 Nov 2024 · 1 repository · arXiv:2411.11770
-
DeforHMR: Vision Transformer with Deformable Cross-Attention for 3D Human Mesh Recovery 18 Nov 2024 · 0 repositories · arXiv:2411.11214
-
Edge-Enhanced Dilated Residual Attention Network for Multimodal Medical Image Fusion 18 Nov 2024 · 1 repository · arXiv:2411.11799
-
Enhancing Decision Transformer with Diffusion-Based Trajectory Branch Generation 18 Nov 2024 · 0 repositories · arXiv:2411.11327
-
Exploring Emerging Trends and Research Opportunities in Visual Place Recognition 18 Nov 2024 · 0 repositories · arXiv:2411.11481
-
Fast Convergence of Softmax Policy Mirror Ascent 18 Nov 2024 · 0 repositories · arXiv:2411.12042
-
FCC: Fully Connected Correlation for Few-Shot Segmentation 18 Nov 2024 · 0 repositories · arXiv:2411.11917
-
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training 18 Nov 2024 · 1 repository · arXiv:2411.11927
-
GLDesigner: Leveraging Multi-Modal LLMs as Designer for Enhanced Aesthetic Text Glyph Layouts 18 Nov 2024 · 0 repositories · arXiv:2411.11435
-
GPS-Gaussian+: Generalizable Pixel-wise 3D Gaussian Splatting for Real-Time Human-Scene Rendering from Sparse Views 18 Nov 2024 · 0 repositories · arXiv:2411.11363
-
Graph Neural Networks for Quantifying Compatibility Mechanisms in Traditional Chinese Medicine 18 Nov 2024 · 1 repository · arXiv:2411.11474
-
Harnessing Scale and Physics: A Multi-Graph Neural Operator Framework for PDEs on Arbitrary Geometries 18 Nov 2024 · 1 repository · arXiv:2411.15178
-
Higher Order Graph Attention Probabilistic Walk Networks 18 Nov 2024 · 0 repositories · arXiv:2411.12052
-
In-Situ Melt Pool Characterization via Thermal Imaging for Defect Detection in Directed Energy Deposition Using Vision Transformers 18 Nov 2024 · 0 repositories · arXiv:2411.12028
-
ITACLIP: Boosting Training-Free Semantic Segmentation with Image, Text, and Architectural Enhancements 18 Nov 2024 · 1 repository · arXiv:2411.12044
-
Item Association Factorization Mixed Markov Chains for Sequential Recommendation 18 Nov 2024 · 0 repositories · arXiv:2501.01429
-
LaVin-DiT: Large Vision Diffusion Transformer 18 Nov 2024 · 0 repositories · arXiv:2411.11505
-
LiTformer: Efficient Modeling and Analysis of High-Speed Link Transmitters Using Non-Autoregressive Transformer 18 Nov 2024 · 0 repositories · arXiv:2411.11699
-
Lung Disease Detection with Vision Transformers: A Comparative Study of Machine Learning Methods 18 Nov 2024 · 0 repositories · arXiv:2411.11376
-
Making Sigmoid-MSE Great Again: Output Reset Challenges Softmax Cross-Entropy in Neural Network Classification 18 Nov 2024 · 0 repositories · arXiv:2411.11213
-
Mechanism and Emergence of Stacked Attention Heads in Multi-Layer Transformers 18 Nov 2024 · 0 repositories · arXiv:2411.12118
-
Multi-Hyperbolic Space-based Heterogeneous Graph Attention Network 18 Nov 2024 · 0 repositories · arXiv:2411.11283
-
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback 18 Nov 2024 · 1 repository · arXiv:2412.03578Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Popular LLMs Amplify Race and Gender Disparities in Human Mobility 18 Nov 2024 · 0 repositories · arXiv:2411.14469
-
SeqProFT: Applying LoRA Finetuning for Sequence-only Protein Property Predictions 18 Nov 2024 · 0 repositories · arXiv:2411.11530
-
ST-Tree with Interpretability for Multivariate Time Series Classification 18 Nov 2024 · 0 repositories · arXiv:2411.11620
-
Suicide Risk Assessment on Social Media with Semi-Supervised Learning 18 Nov 2024 · 0 repositories · arXiv:2411.12767
-
Superpixel-informed Implicit Neural Representation for Multi-Dimensional Data 18 Nov 2024 · 0 repositories · arXiv:2411.11356
-
Enhancing LLM Reasoning with Reward-guided Tree Search 18 Nov 2024 · 2 repositories · arXiv:2411.11694
-
The ADUULM-360 Dataset -- A Multi-Modal Dataset for Depth Estimation in Adverse Weather 18 Nov 2024 · 0 repositories · arXiv:2411.11455
-
TimeFormer: Capturing Temporal Relationships of Deformable 3D Gaussians for Robust Reconstruction 18 Nov 2024 · 1 repository · arXiv:2411.11941
-
Towards a Practical Ethics of Generative AI in Creative Production Processes 18 Nov 2024 · 0 repositories · arXiv:2412.03579
-
TrojanRobot: Physical-World Backdoor Attacks Against VLM-based Robotic Manipulation 18 Nov 2024 · 0 repositories · arXiv:2411.11683
-
Uncovering the role of semantic and acoustic cues in normal and dichotic listening 18 Nov 2024 · 0 repositories · arXiv:2411.11308