Methods › General › Output Functions › Softmax › Papers, page 102
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 102 of 375: papers 10,101 to 10,200 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Micrometer: Micromechanics Transformer for Predicting Mechanical Responses of Heterogeneous Materials 23 Sep 2024 · 0 repositories · arXiv:2410.05281
-
Multi-Modal Generative AI: Multi-modal LLM, Diffusion and Beyond 23 Sep 2024 · 0 repositories · arXiv:2409.14993
-
Optimizing News Text Classification with Bi-LSTM and Attention Mechanism for Efficient Data Processing 23 Sep 2024 · 0 repositories · arXiv:2409.15576
-
PALLM: Evaluating and Enhancing PALLiative Care Conversations with Large Language Models 23 Sep 2024 · 1 repository · arXiv:2409.15188
-
Privacy Policy Analysis through Prompt Engineering for LLMs 23 Sep 2024 · 0 repositories · arXiv:2409.14879
-
Probabilistically Aligned View-unaligned Clustering with Adaptive Template Selection 23 Sep 2024 · 0 repositories · arXiv:2409.14882
-
RACER: Rich Language-Guided Failure Recovery Policies for Imitation Learning 23 Sep 2024 · 0 repositories · arXiv:2409.14674
-
Retrieval Augmented Generation (RAG) and Beyond: A Comprehensive Survey on How to Make your LLMs use External Data More Wisely 23 Sep 2024 · 0 repositories · arXiv:2409.14924
-
Robust and Flexible Omnidirectional Depth Estimation with Multiple 360° Cameras 23 Sep 2024 · 0 repositories · arXiv:2409.14766
-
RoWSFormer: A Robust Watermarking Framework with Swin Transformer for Enhanced Geometric Attack Resilience 23 Sep 2024 · 0 repositories · arXiv:2409.14829
-
Safe Guard: an LLM-agent for Real-time Voice-based Hate Speech Detection in Social Virtual Reality 23 Sep 2024 · 0 repositories · arXiv:2409.15623
-
Scaling Laws of Decoder-Only Models on the Multilingual Machine Translation Task 23 Sep 2024 · 0 repositories · arXiv:2409.15051
-
SDBA: A Stealthy and Long-Lasting Durable Backdoor Attack in Federated Learning 23 Sep 2024 · 1 repository · arXiv:2409.14805
-
SOFI: Multi-Scale Deformable Transformer for Camera Calibration with Enhanced Line Queries 23 Sep 2024 · 1 repository · arXiv:2409.15553
-
TransUKAN:Computing-Efficient Hybrid KAN-Transformer for Enhanced Medical Image Segmentation 23 Sep 2024 · 0 repositories · arXiv:2409.14676
-
Beyond Words: Evaluating Large Language Models in Transportation Planning 22 Sep 2024 · 0 repositories · arXiv:2409.14516
-
Can pre-trained language models generate titles for research papers? 22 Sep 2024 · 1 repository · arXiv:2409.14602
-
EchoAtt: Attend, Copy, then Adjust for More Efficient Large Language Models 22 Sep 2024 · 0 repositories · arXiv:2409.14595
-
EM-DARTS: Hierarchical Differentiable Architecture Search for Eye Movement Recognition 22 Sep 2024 · 0 repositories · arXiv:2409.14432
-
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks 22 Sep 2024 · 0 repositories · arXiv:2409.14488
-
EQ-CBM: A Probabilistic Concept Bottleneck with Energy-based Models and Quantized Vectors 22 Sep 2024 · 0 repositories · arXiv:2409.14630
-
Evaluating the Quality of Code Comments Generated by Large Language Models for Novice Programmers 22 Sep 2024 · 0 repositories · arXiv:2409.14368
-
GroupDiff: Diffusion-based Group Portrait Editing 22 Sep 2024 · 1 repository · arXiv:2409.14379
-
Investigating Layer Importance in Large Language Models 22 Sep 2024 · 0 repositories · arXiv:2409.14381
-
J2N -- Nominal Adjective Identification and its Application 22 Sep 2024 · 1 repository · arXiv:2409.14374
-
Large Model Based Agents: State-of-the-Art, Cooperation Paradigms, Security and Privacy, and Future Trends 22 Sep 2024 · 0 repositories · arXiv:2409.14457
-
LLMs are One-Shot URL Classifiers and Explainers 22 Sep 2024 · 0 repositories · arXiv:2409.14306
-
More Effective LLM Compressed Tokens with Uniformly Spread Position Identifiers and Compression Loss 22 Sep 2024 · 0 repositories · arXiv:2409.14364
-
OStr-DARTS: Differentiable Neural Architecture Search based on Operation Strength 22 Sep 2024 · 1 repository · arXiv:2409.14433
-
Patch Ranking: Efficient CLIP by Learning to Rank Local Patches 22 Sep 2024 · 1 repository · arXiv:2409.14607
-
Prior Knowledge Distillation Network for Face Super-Resolution 22 Sep 2024 · 0 repositories · arXiv:2409.14385
-
Proof Automation with Large Language Models 22 Sep 2024 · 0 repositories · arXiv:2409.14274
-
Sparse Low-Ranked Self-Attention Transformer for Remaining Useful Lifetime Prediction of Optical Fiber Amplifiers 22 Sep 2024 · 0 repositories · arXiv:2409.14378
-
Thinking in Granularity: Dynamic Quantization for Image Super-Resolution by Intriguing Multi-Granularity Clues 22 Sep 2024 · 1 repository · arXiv:2409.14330
-
TrackNetV4: Enhancing Fast Sports Object Tracking with Motion Attention Maps 22 Sep 2024 · 0 repositories · arXiv:2409.14543
-
UU-Mamba: Uncertainty-aware U-Mamba for Cardiovascular Segmentation 22 Sep 2024 · 1 repository · arXiv:2409.14305
-
ChemEval: A Comprehensive Multi-Level Chemical Evaluation for Large Language Models 21 Sep 2024 · 1 repository · arXiv:2409.13989Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 0 violated, 17 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
SMART-RAG: Selection using Determinantal Matrices for Augmented Retrieval 21 Sep 2024 · 0 repositories · arXiv:2409.13992
-
Graph Neural Network Framework for Sentiment Analysis Using Syntactic Feature 21 Sep 2024 · 0 repositories · arXiv:2409.14000
-
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators 21 Sep 2024 · 1 repository · arXiv:2409.14037
-
MultiMed: Multilingual Medical Speech Recognition via Attention Encoder Decoder 21 Sep 2024 · 1 repository · arXiv:2409.14074
-
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers 21 Sep 2024 · 0 repositories · arXiv:2409.14097
-
Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis 21 Sep 2024 · 2 repositories · arXiv:2409.14144Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Towards Building Efficient Sentence BERT Models using Layer Pruning 21 Sep 2024 · 0 repositories · arXiv:2409.14168
-
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling 21 Sep 2024 · 1 repository · arXiv:2409.14175
-
A Sinkhorn Regularized Adversarial Network for Image Guided DEM Super-resolution using Frequency Selective Hybrid Graph Transformer 21 Sep 2024 · 0 repositories · arXiv:2409.14198
-
AI Assistants for Spaceflight Procedures: Combining Generative Pre-Trained Transformer and Retrieval-Augmented Generation on Knowledge Graphs With Augmented Reality Cues 21 Sep 2024 · 0 repositories · arXiv:2409.14206
-
Are Music Foundation Models Better at Singing Voice Deepfake Detection? Far-Better Fuse them with Speech Foundation Models 21 Sep 2024 · 0 repositories · arXiv:2409.14131
-
Boolean Product Graph Neural Networks 21 Sep 2024 · 0 repositories · arXiv:2409.14001
-
Detecting Inpainted Video with Frequency Domain Insights 21 Sep 2024 · 0 repositories · arXiv:2409.13976
-
Developing a Thailand solar irradiance map using Himawari-8 satellite imageries and deep learning models 21 Sep 2024 · 1 repository · arXiv:2409.16320
-
Drift to Remember 21 Sep 2024 · 0 repositories · arXiv:2409.13997
-
Multilateral Cascading Network for Semantic Segmentation of Large-Scale Outdoor Point Clouds 21 Sep 2024 · 0 repositories · arXiv:2409.13983
-
FAMOUS: Flexible Accelerator for the Attention Mechanism of Transformer on UltraScale+ FPGAs 21 Sep 2024 · 0 repositories · arXiv:2409.14023
-
Generalizable Non-Line-of-Sight Imaging with Learnable Physical Priors 21 Sep 2024 · 0 repositories · arXiv:2409.14011
-
Knowledge in Triples for LLMs: Enhancing Table QA Accuracy with Semantic Extraction 21 Sep 2024 · 0 repositories · arXiv:2409.14192
-
Loop Neural Networks for Parameter Sharing 21 Sep 2024 · 0 repositories · arXiv:2409.14199
-
Monocular Event-Inertial Odometry with Adaptive decay-based Time Surface and Polarity-aware Tracking 21 Sep 2024 · 0 repositories · arXiv:2409.13971
-
MSDet: Receptive Field Enhanced Multiscale Detection for Tiny Pulmonary Nodule 21 Sep 2024 · 1 repository · arXiv:2409.14028
-
Multiple-Exit Tuning: Towards Inference-Efficient Adaptation for Vision Transformer 21 Sep 2024 · 0 repositories · arXiv:2409.13999
-
On Broad-Beam Reflection for Dual-Polarized RIS-Assisted MIMO Systems 21 Sep 2024 · 0 repositories · arXiv:2410.07134
-
ProTEA: Programmable Transformer Encoder Acceleration on FPGA 21 Sep 2024 · 0 repositories · arXiv:2409.13975
-
Semi-intrusive audio evaluation: Casting non-intrusive assessment as a multi-modal text prediction task 21 Sep 2024 · 0 repositories · arXiv:2409.14069
-
What is a Digital Twin Anyway? Deriving the Definition for the Built Environment from over 15,000 Scientific Publications 21 Sep 2024 · 0 repositories · arXiv:2409.19005
-
Window-based Channel Attention for Wavelet-enhanced Learned Image Compression 21 Sep 2024 · 0 repositories · arXiv:2409.14090
-
Contextual Compression in Retrieval-Augmented Generation for Large Language Models: A Survey 20 Sep 2024 · 1 repository · arXiv:2409.13385
-
3D-GSW: 3D Gaussian Splatting for Robust Watermarking 20 Sep 2024 · 0 repositories · arXiv:2409.13222
-
A Comparison between Financial and Gambling Markets 20 Sep 2024 · 0 repositories · arXiv:2409.13528
-
A Personalised 3D+t Mesh Generative Model for Unveiling Normal Heart Dynamics 20 Sep 2024 · 1 repository · arXiv:2409.13825
-
A Survey of 5G-Based Positioning for Industry 4.0: State of the Art and Enhanced Techniques 20 Sep 2024 · 0 repositories · arXiv:2409.13308
-
Aligning Language Models Using Follow-up Likelihood as Reward Signal 20 Sep 2024 · 1 repository · arXiv:2409.13948
-
Analysis of Gene Regulatory Networks from Gene Expression Using Graph Neural Networks 20 Sep 2024 · 1 repository · arXiv:2409.13664
-
Applying Pre-trained Multilingual BERT in Embeddings for Improved Malicious Prompt Injection Attacks Detection 20 Sep 2024 · 0 repositories · arXiv:2409.13331
-
AVG-LLaVA: A Large Multimodal Model with Adaptive Visual Granularity 20 Sep 2024 · 1 repository · arXiv:2410.02745Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Beyond the binary: Limitations and possibilities of gender-related speech technology research 20 Sep 2024 · 1 repository · arXiv:2409.13335
-
Brain-Cognition Fingerprinting via Graph-GCCA with Contrastive Learning 20 Sep 2024 · 0 repositories · arXiv:2409.13887
-
Cooperative Resilience in Artificial Intelligence Multiagent Systems 20 Sep 2024 · 1 repository · arXiv:2409.13187
-
Cross-Target Stance Detection: A Survey of Techniques, Datasets, and Challenges 20 Sep 2024 · 0 repositories · arXiv:2409.13594
-
Data Augmentation for Sequential Recommendation: A Survey 20 Sep 2024 · 1 repository · arXiv:2409.13545
-
DS2TA: Denoising Spiking Transformer with Attenuated Spatiotemporal Attention 20 Sep 2024 · 0 repositories · arXiv:2409.15375
-
EMMeTT: Efficient Multimodal Machine Translation Training 20 Sep 2024 · 0 repositories · arXiv:2409.13523
-
Enhancing Large Language Models with Domain-specific Retrieval Augment Generation: A Case Study on Long-form Consumer Health Question Answering in Ophthalmology 20 Sep 2024 · 0 repositories · arXiv:2409.13902
-
FAIR GPT: A virtual consultant for research data management in ChatGPT 20 Sep 2024 · 1 repository · arXiv:2410.07108
-
GAProtoNet: A Multi-head Graph Attention-based Prototypical Network for Interpretable Text Classification 20 Sep 2024 · 1 repository · arXiv:2409.13312
-
GASA-UNet: Global Axial Self-Attention U-Net for 3D Medical Image Segmentation 20 Sep 2024 · 0 repositories · arXiv:2409.13146
-
High-dimensional learning of narrow neural networks 20 Sep 2024 · 0 repositories · arXiv:2409.13904
-
HUT: A More Computation Efficient Fine-Tuning Method With Hadamard Updated Transformation 20 Sep 2024 · 0 repositories · arXiv:2409.13501
-
Imagine yourself: Tuning-Free Personalized Image Generation 20 Sep 2024 · 0 repositories · arXiv:2409.13346
-
Improved Unet brain tumor image segmentation based on GSConv module and ECA attention mechanism 20 Sep 2024 · 0 repositories · arXiv:2409.13626
-
Large Language Model Should Understand Pinyin for Chinese ASR Error Correction 20 Sep 2024 · 0 repositories · arXiv:2409.13262
-
Learning to Compare Hardware Designs for High-Level Synthesis 20 Sep 2024 · 1 repository · arXiv:2409.13138
-
Leveraging Knowledge Graphs and LLMs to Support and Monitor Legislative Systems 20 Sep 2024 · 0 repositories · arXiv:2409.13252
-
Localized Gaussians as Self-Attention Weights for Point Clouds Correspondence 20 Sep 2024 · 0 repositories · arXiv:2409.13291
-
Multiscale Encoder and Omni-Dimensional Dynamic Convolution Enrichment in nnU-Net for Brain Tumor Segmentation 20 Sep 2024 · 1 repository · arXiv:2409.13229
-
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks 20 Sep 2024 · 1 repository · arXiv:2409.13203Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Occupancy-Based Dual Contouring 20 Sep 2024 · 1 repository · arXiv:2409.13418
-
On-Device Collaborative Language Modeling via a Mixture of Generalists and Specialists 20 Sep 2024 · 1 repository · arXiv:2409.13931
-
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping 20 Sep 2024 · 1 repository · arXiv:2409.13912
-
Persistent Backdoor Attacks in Continual Learning 20 Sep 2024 · 0 repositories · arXiv:2409.13864
-
PlainUSR: Chasing Faster ConvNet for Efficient Super-Resolution 20 Sep 2024 · 1 repository · arXiv:2409.13435