Methods › General › Output Functions › Softmax › Papers, page 172
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 172 of 375: papers 17,101 to 17,200 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TAnet: A New Temporal Attention Network for EEG-based Auditory Spatial Attention Decoding with a Short Decision Window 11 Jan 2024 · 0 repositories · arXiv:2401.05819
-
The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models 11 Jan 2024 · 1 repository · arXiv:2401.05618Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Transforming Image Super-Resolution: A ConvFormer-based Efficient Approach 11 Jan 2024 · 1 repository · arXiv:2401.05633
-
YOLO-Former: YOLO Shakes Hand With ViT 11 Jan 2024 · 0 repositories · arXiv:2401.06244
-
Adaptive-avg-pooling based Attention Vision Transformer for Face Anti-spoofing 10 Jan 2024 · 0 repositories · arXiv:2401.04953
-
Advancing ECG Diagnosis Using Reinforcement Learning on Global Waveform Variations Related to P Wave and PR Interval 10 Jan 2024 · 0 repositories · arXiv:2401.04938
-
AdvMT: Adversarial Motion Transformer for Long-term Human Motion Prediction 10 Jan 2024 · 0 repositories · arXiv:2401.05018
-
Deep learning in motion deblurring: current status, benchmarks and future prospects 10 Jan 2024 · 1 repository · arXiv:2401.05055
-
AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning 10 Jan 2024 · 1 repository · arXiv:2401.05268Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
CADgpt: Harnessing Natural Language Processing for 3D Modelling to Enhance Computer-Aided Design Workflows 10 Jan 2024 · 0 repositories · arXiv:2401.05476
-
Can AI Write Classical Chinese Poetry like Humans? An Empirical Study Inspired by Turing Test 10 Jan 2024 · 0 repositories · arXiv:2401.04952
-
Derm-T2IM: Harnessing Synthetic Skin Lesion Data via Stable Diffusion Models for Enhanced Skin Disease Classification using ViT and CNN 10 Jan 2024 · 0 repositories · arXiv:2401.05159
-
Diffusion-based Pose Refinement and Muti-hypothesis Generation for 3D Human Pose Estimaiton 10 Jan 2024 · 1 repository · arXiv:2401.04921
-
Efficient Fine-Tuning with Domain Adaptation for Privacy-Preserving Vision Transformer 10 Jan 2024 · 0 repositories · arXiv:2401.05126
-
I am a Strange Dataset: Metalinguistic Tests for Language Models 10 Jan 2024 · 1 repository · arXiv:2401.05300
-
InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks 10 Jan 2024 · 1 repository · arXiv:2401.05507Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Knowledge-aware Graph Transformer for Pedestrian Trajectory Prediction 10 Jan 2024 · 0 repositories · arXiv:2401.04872
-
Knowledge Sharing in Manufacturing using Large Language Models: User Evaluation and Model Benchmarking 10 Jan 2024 · 0 repositories · arXiv:2401.05200
-
Leveraging Print Debugging to Improve Code Generation in Large Language Models 10 Jan 2024 · 0 repositories · arXiv:2401.05319
-
Can Active Label Correction Improve LLM-based Modular AI Systems? 10 Jan 2024 · 0 repositories · arXiv:2401.05467
-
MISS: Multiclass Interpretable Scoring Systems 10 Jan 2024 · 1 repository · arXiv:2401.05069
-
Monte Carlo Tree Search for Recipe Generation using GPT-2 10 Jan 2024 · 0 repositories · arXiv:2401.05199
-
Motion Guided Token Compression for Efficient Masked Video Modeling 10 Jan 2024 · 0 repositories · arXiv:2402.18577
-
Reinforcement Learning for Optimizing RAG for Domain Chatbots 10 Jan 2024 · 0 repositories · arXiv:2401.06800
-
SENet: Visual Detection of Online Social Engineering Attack Campaigns 10 Jan 2024 · 0 repositories · arXiv:2401.05569
-
SPT: Spectral Transformer for Red Giant Stars Age and Mass Estimation 10 Jan 2024 · 0 repositories · arXiv:2401.04900
-
An Assessment on Comprehending Mental Health through Large Language Models 9 Jan 2024 · 0 repositories · arXiv:2401.04592
-
Arabic Text Diacritization In The Age Of Transfer Learning: Token Classification Is All You Need 9 Jan 2024 · 0 repositories · arXiv:2401.04848
-
DebugBench: Evaluating Debugging Capability of Large Language Models 9 Jan 2024 · 1 repository · arXiv:2401.04621Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples)
-
DedustNet: A Frequency-dominated Swin Transformer-based Wavelet Network for Agricultural Dust Removal 9 Jan 2024 · 0 repositories · arXiv:2401.04750
-
DepressionEmo: A novel dataset for multilabel classification of depression emotions 9 Jan 2024 · 1 repository · arXiv:2401.04655
-
Fighting Fire with Fire: Adversarial Prompting to Generate a Misinformation Detection Dataset 9 Jan 2024 · 0 repositories · arXiv:2401.04481
-
Informed AI Regulation: Comparing the Ethical Frameworks of Leading LLM Chatbots Using an Ethics-Based Audit to Assess Moral Reasoning and Normative Values 9 Jan 2024 · 1 repository · arXiv:2402.01651
-
Iterative Feedback Network for Unsupervised Point Cloud Registration 9 Jan 2024 · 1 repository · arXiv:2401.04357
-
Language Detection for Transliterated Content 9 Jan 2024 · 0 repositories · arXiv:2401.04619
-
Lightning Attention-2: A Free Lunch for Handling Unlimited Sequence Lengths in Large Language Models 9 Jan 2024 · 1 repository · arXiv:2401.04658Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Phishing Website Detection through Multi-Model Analysis of HTML Content 9 Jan 2024 · 0 repositories · arXiv:2401.04820
-
Setting the Record Straight on Transformer Oversmoothing 9 Jan 2024 · 0 repositories · arXiv:2401.04301
-
Skin Cancer Segmentation and Classification Using Vision Transformer for Automatic Analysis in Dermatoscopy-based Non-invasive Digital System 9 Jan 2024 · 0 repositories · arXiv:2401.04746
-
T-PRIME: Transformer-based Protocol Identification for Machine-learning at the Edge 9 Jan 2024 · 1 repository · arXiv:2401.04837
-
WaveletFormerNet: A Transformer-based Wavelet Network for Real-world Non-homogeneous and Dense Fog Removal 9 Jan 2024 · 0 repositories · arXiv:2401.04550
-
A Philosophical Introduction to Language Models -- Part I: Continuity With Classic Debates 8 Jan 2024 · 0 repositories · arXiv:2401.03910
-
Advancing Spatial Reasoning in Large Language Models: An In-Depth Evaluation and Enhancement Using the StepGame Benchmark 8 Jan 2024 · 1 repository · arXiv:2401.03991
-
An Exploratory Study on Automatic Identification of Assumptions in the Development of Deep Learning Frameworks 8 Jan 2024 · 1 repository · arXiv:2401.03653
-
Anatomy of Neural Language Models 8 Jan 2024 · 1 repository · arXiv:2401.03797
-
Attention-Guided Erasing: A Novel Augmentation Method for Enhancing Downstream Breast Density Classification 8 Jan 2024 · 0 repositories · arXiv:2401.03912
-
Can Large Language Models Beat Wall Street? Unveiling the Potential of AI in Stock Selection 8 Jan 2024 · 0 repositories · arXiv:2401.03737
-
Comparative Analysis of Deep Convolutional Neural Networks for Detecting Medical Image Deepfakes 8 Jan 2024 · 0 repositories · arXiv:2406.08758
-
Distortions in Judged Spatial Relations in Large Language Models 8 Jan 2024 · 0 repositories · arXiv:2401.04218
-
Efficient Multiscale Multimodal Bottleneck Transformer for Audio-Video Classification 8 Jan 2024 · 0 repositories · arXiv:2401.04023
-
Efficient Selective Audio Masked Multimodal Bottleneck Transformer for Audio-Video Classification 8 Jan 2024 · 0 repositories · arXiv:2401.04154
-
Exploratory Evaluation of Speech Content Masking 8 Jan 2024 · 0 repositories · arXiv:2401.03936
-
GloTSFormer: Global Video Text Spotting Transformer 8 Jan 2024 · 1 repository · arXiv:2401.03694
-
Gramformer: Learning Crowd Counting via Graph-Modulated Transformer 8 Jan 2024 · 1 repository · arXiv:2401.03870Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Advancing bioinformatics with large language models: components, applications and perspectives 8 Jan 2024 · 0 repositories · arXiv:2401.04155
-
LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition 8 Jan 2024 · 1 repository · arXiv:2402.00033Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems 8 Jan 2024 · 1 repository · arXiv:2401.05443
-
MARG: Multi-Agent Review Generation for Scientific Papers 8 Jan 2024 · 1 repository · arXiv:2401.04259Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Mixtral of Experts 8 Jan 2024 · 6 repositories · arXiv:2401.04088Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
MoE-Mamba: Efficient Selective State Space Models with Mixture of Experts 8 Jan 2024 · 1 repository · arXiv:2401.04081
-
MS-DETR: Efficient DETR Training with Mixed Supervision 8 Jan 2024 · 1 repository · arXiv:2401.03989Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
Why Solving Multi-agent Path Finding with Large Language Model has not Succeeded Yet 8 Jan 2024 · 0 repositories · arXiv:2401.03630
-
Can generative AI and ChatGPT outperform humans on cognitive-demanding problem-solving tasks in science? 7 Jan 2024 · 0 repositories · arXiv:2401.15081
-
EAT: Self-Supervised Pre-Training with Efficient Audio Transformer 7 Jan 2024 · 1 repository · arXiv:2401.03497Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 3 pointer-only (licence)
-
Escalation Risks from Language Models in Military and Diplomatic Decision-Making 7 Jan 2024 · 1 repository · arXiv:2401.03408Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
InFoBench: Evaluating Instruction Following Ability in Large Language Models 7 Jan 2024 · 1 repository · arXiv:2401.03601Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
On Leveraging Large Language Models for Enhancing Entity Resolution: A Cost-efficient Approach 7 Jan 2024 · 0 repositories · arXiv:2401.03426
-
PEneo: Unifying Line Extraction, Line Grouping, and Entity Linking for End-to-end Document Pair Extraction 7 Jan 2024 · 1 repository · arXiv:2401.03472
-
RoBERTurk: Adjusting RoBERTa for Turkish 7 Jan 2024 · 0 repositories · arXiv:2401.03515
-
See360: Novel Panoramic View Interpolation 7 Jan 2024 · 1 repository · arXiv:2401.03431
-
SeTformer is What You Need for Vision and Language 7 Jan 2024 · 0 repositories · arXiv:2401.03540
-
The NPU-ASLP-LiAuto System Description for Visual Speech Recognition in CNVSRC 2023 7 Jan 2024 · 2 repositories · arXiv:2401.06788
-
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM 7 Jan 2024 · 0 repositories · arXiv:2401.03512
-
Towards Effective Multiple-in-One Image Restoration: A Sequential and Prompt Learning Strategy 7 Jan 2024 · 2 repositories · arXiv:2401.03379Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Exploring Defeasibility in Causal Reasoning 6 Jan 2024 · 0 repositories · arXiv:2401.03183
-
Multimodal Informative ViT: Information Aggregation and Distribution for Hyperspectral and LiDAR Classification 6 Jan 2024 · 1 repository · arXiv:2401.03179
-
PIXAR: Auto-Regressive Language Modeling in Pixel Space 6 Jan 2024 · 0 repositories · arXiv:2401.03321
-
PosDiffNet: Positional Neural Diffusion for Point Cloud Registration in a Large Field of View with Perturbations 6 Jan 2024 · 0 repositories · arXiv:2401.03167
-
Realism in Action: Anomaly-Aware Diagnosis of Brain Tumors from Medical Images Using YOLOv8 and DeiT 6 Jan 2024 · 0 repositories · arXiv:2401.03302
-
SecureReg: Combining NLP and MLP for Enhanced Detection of Malicious Domain Name Registrations 6 Jan 2024 · 0 repositories · arXiv:2401.03196
-
UGGNet: Bridging U-Net and VGG for Advanced Breast Cancer Diagnosis 6 Jan 2024 · 0 repositories · arXiv:2401.03173
-
Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors 6 Jan 2024 · 0 repositories · arXiv:2401.03238
-
Vision Transformers and Bi-LSTM for Alzheimer's Disease Diagnosis from 3D MRI 6 Jan 2024 · 0 repositories · arXiv:2401.03132
-
Web Diagnosis for COVID-19 and Pneumonia Based on Computed Tomography Scans and X-rays 6 Jan 2024 · 2 repositories
-
A Cost-Efficient FPGA Implementation of Tiny Transformer Model using Neural ODE 5 Jan 2024 · 0 repositories · arXiv:2401.02721
-
A Random Ensemble of Encrypted models for Enhancing Robustness against Adversarial Examples 5 Jan 2024 · 0 repositories · arXiv:2401.02633
-
AST-T5: Structure-Aware Pretraining for Code Generation and Understanding 5 Jan 2024 · 1 repository · arXiv:2401.03003Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Can Large Language Models Understand Molecules? 5 Jan 2024 · 2 repositories · arXiv:2402.00024
-
CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution 5 Jan 2024 · 1 repository · arXiv:2401.03065
-
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism 5 Jan 2024 · 1 repository · arXiv:2401.02954
-
Denoising Vision Transformers 5 Jan 2024 · 1 repository · arXiv:2401.02957Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
From LLM to Conversational Agent: A Memory Enhanced Architecture with Fine-Tuning of Large Language Models 5 Jan 2024 · 0 repositories · arXiv:2401.02777
-
Natural Language Programming in Medicine: Administering Evidence Based Clinical Workflows with Autonomous Agents Powered by Generative Large Language Models 5 Jan 2024 · 0 repositories · arXiv:2401.02851
-
Geometric-Facilitated Denoising Diffusion Model for 3D Molecule Generation 5 Jan 2024 · 1 repository · arXiv:2401.02683Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples)
-
German Text Embedding Clustering Benchmark 5 Jan 2024 · 1 repository · arXiv:2401.02709
-
Latte: Latent Diffusion Transformer for Video Generation 5 Jan 2024 · 4 repositories · arXiv:2401.03048Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
On the numerical reliability of nonsmooth autodiff: a MaxPool case study 5 Jan 2024 · 1 repository · arXiv:2401.02736
-
Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks 5 Jan 2024 · 2 repositories · arXiv:2401.02731
-
PeFoMed: Parameter Efficient Fine-tuning of Multimodal Large Language Models for Medical Imaging 5 Jan 2024 · 1 repository · arXiv:2401.02797
-
Prompt-driven Latent Domain Generalization for Medical Image Classification 5 Jan 2024 · 2 repositories · arXiv:2401.03002