Methods › General › Output Functions › Softmax › Papers, page 146
Softmax
Papers archive 2025-07-28
archive papers tagged: 37,443 · with a code link: 15,869 · where Syntology ran a sample: 4,578 (3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (4,578 of 37,443 tagged: 3,835 with a run with no instrument failure, 743 where every run was a failure of Syntology's instrument)
Page 146 of 375: papers 14,501 to 14,600 of 37,443, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Vision Transformers for End-to-End Vision-Based Quadrotor Obstacle Avoidance 16 May 2024 · 0 repositories · arXiv:2405.10391
-
ALPINE: Unveiling the Planning Capability of Autoregressive Learning in Language Models 15 May 2024 · 0 repositories · arXiv:2405.09220
-
An Embarrassingly Simple Approach to Enhance Transformer Performance in Genomic Selection for Crop Breeding 15 May 2024 · 2 repositories · arXiv:2405.09585
-
Bridging the gap in online hate speech detection: a comparative analysis of BERT and traditional models for homophobic content identification on X/Twitter 15 May 2024 · 0 repositories · arXiv:2405.09221
-
Comparing the Efficacy of GPT-4 and Chat-GPT in Mental Health Care: A Blind Assessment of Large Language Models for Psychological Support 15 May 2024 · 0 repositories · arXiv:2405.09300
-
Enhancing Maritime Trajectory Forecasting via H3 Index and Causal Language Modelling (CLM) 15 May 2024 · 0 repositories · arXiv:2405.09596
-
Exploring the Potential of Large Language Models for Automation in Technical Customer Service 15 May 2024 · 0 repositories · arXiv:2405.09161
-
IM-RAG: Multi-Round Retrieval-Augmented Generation Through Learning Inner Monologues 15 May 2024 · 0 repositories · arXiv:2405.13021
-
Improving Label Error Detection and Elimination with Uncertainty Quantification 15 May 2024 · 0 repositories · arXiv:2405.09602
-
Improving Transformers using Faithful Positional Encoding 15 May 2024 · 0 repositories · arXiv:2405.09061
-
Intelligent Tutor: Leveraging ChatGPT and Microsoft Copilot Studio to Deliver a Generative AI Student Support and Feedback System within Teams 15 May 2024 · 0 repositories · arXiv:2405.13024
-
Matching domain experts by training from scratch on domain knowledge 15 May 2024 · 0 repositories · arXiv:2405.09395
-
Modeling Bilingual Sentence Processing: Evaluating RNN and Transformer Architectures for Cross-Language Structural Priming 15 May 2024 · 0 repositories · arXiv:2405.09508
-
Perception- and Fidelity-aware Reduced-Reference Super-Resolution Image Quality Assessment 15 May 2024 · 0 repositories · arXiv:2405.09472
-
Positional Knowledge is All You Need: Position-induced Transformer (PiT) for Operator Learning 15 May 2024 · 1 repository · arXiv:2405.09285
-
Simulating Policy Impacts: Developing a Generative Scenario Writing Method to Evaluate the Perceived Effects of Regulation 15 May 2024 · 0 repositories · arXiv:2405.09679
-
SQL-to-Schema Enhances Schema Linking in Text-to-SQL 15 May 2024 · 0 repositories · arXiv:2405.09593
-
Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models 15 May 2024 · 1 repository · arXiv:2405.09454
-
Transfer Learning in Pre-Trained Large Language Models for Malware Detection Based on System Calls 15 May 2024 · 0 repositories · arXiv:2405.09318
-
Word Alignment as Preference for Machine Translation 15 May 2024 · 0 repositories · arXiv:2405.09223
-
A Click-Through Rate Prediction Method Based on Cross-Importance of Multi-Order Features 14 May 2024 · 0 repositories · arXiv:2405.08852
-
A Comprehensive Survey of Large Language Models and Multimodal Large Language Models in Medicine 14 May 2024 · 0 repositories · arXiv:2405.08603
-
A Timely Survey on Vision Transformer for Deepfake Detection 14 May 2024 · 0 repositories · arXiv:2405.08463
-
Abnormal Respiratory Sound Identification Using Audio-Spectrogram Vision Transformer 14 May 2024 · 0 repositories · arXiv:2405.08342
-
Airport Delay Prediction with Temporal Fusion Transformers 14 May 2024 · 0 repositories · arXiv:2405.08293
-
Amplifying Aspect-Sentence Awareness: A Novel Approach for Aspect-Based Sentiment Analysis 14 May 2024 · 0 repositories · arXiv:2405.13013
-
Beyond Scaling Laws: Understanding Transformer Performance with Associative Memory 14 May 2024 · 0 repositories · arXiv:2405.08707
-
Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis 14 May 2024 · 0 repositories · arXiv:2405.08944
-
Compositional Text-to-Image Generation with Dense Blob Representations 14 May 2024 · 0 repositories · arXiv:2405.08246
-
DGCformer: Deep Graph Clustering Transformer for Multivariate Time Series Forecasting 14 May 2024 · 0 repositories · arXiv:2405.08440
-
GPT-3.5 for Grammatical Error Correction 14 May 2024 · 0 repositories · arXiv:2405.08469
-
Harnessing the power of longitudinal medical imaging for eye disease prognosis using Transformer-based sequence modeling 14 May 2024 · 1 repository · arXiv:2405.08780
-
Improving Transformers with Dynamically Composable Multi-Head Attention 14 May 2024 · 2 repositories · arXiv:2405.08553Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
Refinement of an Epilepsy Dictionary through Human Annotation of Health-related posts on Instagram 14 May 2024 · 0 repositories · arXiv:2405.08784
-
Reinformer: Max-Return Sequence Modeling for Offline RL 14 May 2024 · 1 repository · arXiv:2405.08740Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 4 harvested samples) · 4 pointer-only (licence)
-
Rethinking Scanning Strategies with Vision Mamba in Semantic Segmentation of Remote Sensing Imagery: An Experimental Study 14 May 2024 · 0 repositories · arXiv:2405.08493
-
RMT-BVQA: Recurrent Memory Transformer-based Blind Video Quality Assessment for Enhanced Video Content 14 May 2024 · 0 repositories · arXiv:2405.08621
-
TFWT: Tabular Feature Weighting with Transformer 14 May 2024 · 0 repositories · arXiv:2405.08403
-
Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control 14 May 2024 · 0 repositories · arXiv:2405.08366
-
WaterMamba: Visual State Space Model for Underwater Image Enhancement 14 May 2024 · 0 repositories · arXiv:2405.08419
-
When Large Language Models Meet Optical Networks: Paving the Way for Automation 14 May 2024 · 0 repositories · arXiv:2405.17441
-
Can Language Models Explain Their Own Classification Behavior? 13 May 2024 · 1 repository · arXiv:2405.07436
-
CDFormer:When Degradation Prediction Embraces Diffusion Model for Blind Image Super-Resolution 13 May 2024 · 1 repository · arXiv:2405.07648Syntology official (archive's flag): 10 ran · 11 ran (of which 7 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 17 harvested samples) · 1 pointer-only (licence)
-
Coding historical causes of death data with Large Language Models 13 May 2024 · 1 repository · arXiv:2405.07560
-
Control Token with Dense Passage Retrieval 13 May 2024 · 0 repositories · arXiv:2405.13008
-
DEPTH: Discourse Education through Pre-Training Hierarchically 13 May 2024 · 1 repository · arXiv:2405.07788
-
Evaluation of Retrieval-Augmented Generation: A Survey 13 May 2024 · 1 repository · arXiv:2405.07437
-
Fighter flight trajectory prediction based on spatio-temporal graphcial attention network 13 May 2024 · 0 repositories · arXiv:2405.08034
-
FreeVA: Offline MLLM as Training-Free Video Assistant 13 May 2024 · 1 repository · arXiv:2405.07798Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
From Questions to Insightful Answers: Building an Informed Chatbot for University Resources 13 May 2024 · 0 repositories · arXiv:2405.08120
-
Ground-based image deconvolution with Swin Transformer UNet 13 May 2024 · 0 repositories · arXiv:2405.07842
-
Decision Mamba Architectures 13 May 2024 · 2 repositories · arXiv:2405.07943
-
HybridHash: Hybrid Convolutional and Self-Attention Deep Hashing for Image Retrieval 13 May 2024 · 1 repository · arXiv:2405.07524
-
MacBehaviour: An R package for behavioural experimentation on large language models 13 May 2024 · 1 repository · arXiv:2405.07495
-
Many-Shot Regurgitation (MSR) Prompting 13 May 2024 · 0 repositories · arXiv:2405.08134
-
MetaReflection: Learning Instructions for Language Agents using Past Reflections 13 May 2024 · 0 repositories · arXiv:2405.13009
-
NutritionVerse-Direct: Exploring Deep Neural Networks for Multitask Nutrition Prediction from Food Images 13 May 2024 · 0 repositories · arXiv:2405.07814
-
Open-vocabulary Auditory Neural Decoding Using fMRI-prompted LLM 13 May 2024 · 0 repositories · arXiv:2405.07840
-
Plot2Code: A Comprehensive Benchmark for Evaluating Multi-modal Large Language Models in Code Generation from Scientific Plots 13 May 2024 · 0 repositories · arXiv:2405.07990
-
PLUTO: Pathology-Universal Transformer 13 May 2024 · 0 repositories · arXiv:2405.07905
-
Quantifying and Optimizing Global Faithfulness in Persona-driven Role-playing 13 May 2024 · 1 repository · arXiv:2405.07726Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
RESTAD: REconstruction and Similarity based Transformer for time series Anomaly Detection 13 May 2024 · 1 repository · arXiv:2405.07509
-
SambaNova SN40L: Scaling the AI Memory Wall with Dataflow and Composition of Experts 13 May 2024 · 0 repositories · arXiv:2405.07518
-
BoQ: A Place is Worth a Bag of Learnable Queries 12 May 2024 · 1 repository · arXiv:2405.07364Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
CaFA: Global Weather Forecasting with Factorized Attention on Sphere 12 May 2024 · 1 repository · arXiv:2405.07395Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 2 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
DuetRAG: Collaborative Retrieval-Augmented Generation 12 May 2024 · 0 repositories · arXiv:2405.13002
-
Explainable Convolutional Neural Networks for Retinal Fundus Classification and Cutting-Edge Segmentation Models for Retinal Blood Vessels from Fundus Images 12 May 2024 · 1 repository · arXiv:2405.07338
-
ExplainableDetector: Exploring Transformer-based Language Modeling Approach for SMS Spam Detection with Explainability Analysis 12 May 2024 · 0 repositories · arXiv:2405.08026
-
HGTDR: Advancing Drug Repurposing with Heterogeneous Graph Transformers 12 May 2024 · 0 repositories · arXiv:2405.08031
-
Humor Mechanics: Advancing Humor Generation with Multistep Reasoning 12 May 2024 · 1 repository · arXiv:2405.07280
-
L(u)PIN: LLM-based Political Ideology Nowcasting 12 May 2024 · 0 repositories · arXiv:2405.07320
-
Learning Reward for Robot Skills Using Large Language Models via Self-Alignment 12 May 2024 · 0 repositories · arXiv:2405.07162
-
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis 12 May 2024 · 1 repository · arXiv:2405.07248Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
MedConceptsQA: Open Source Medical Concepts QA Benchmark 12 May 2024 · 1 repository · arXiv:2405.07348Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AraSpell: A Deep Learning Approach for Arabic Spelling Correction 11 May 2024 · 1 repository · arXiv:2405.06981
-
Automating Thematic Analysis: How LLMs Analyse Controversial Topics 11 May 2024 · 0 repositories · arXiv:2405.06919
-
CTRL: Continuous-Time Representation Learning on Temporal Heterogeneous Information Network 11 May 2024 · 0 repositories · arXiv:2405.08013
-
Identifying Key Terms in Prompts for Relevance Evaluation with GPT Models 11 May 2024 · 0 repositories · arXiv:2405.06931
-
Length-Aware Multi-Kernel Transformer for Long Document Classification 11 May 2024 · 1 repository · arXiv:2405.07052
-
Multi-agent Traffic Prediction via Denoised Endpoint Distribution 11 May 2024 · 0 repositories · arXiv:2405.07041
-
QMViT: A Mushroom is worth 16x16 Words 11 May 2024 · 0 repositories · arXiv:2407.04708
-
Quite Good, but Not Enough: Nationality Bias in Large Language Models -- A Case Study of ChatGPT 11 May 2024 · 1 repository · arXiv:2405.06996
-
Replication Study and Benchmarking of Real-Time Object Detection Models 11 May 2024 · 1 repository · arXiv:2405.06911
-
RETTA: Retrieval-Enhanced Test-Time Adaptation for Zero-Shot Video Captioning 11 May 2024 · 0 repositories · arXiv:2405.07046
-
RoTHP: Rotary Position Embedding-based Transformer Hawkes Process 11 May 2024 · 0 repositories · arXiv:2405.06985
-
Super-Resolving Blurry Images with Events 11 May 2024 · 0 repositories · arXiv:2405.06918
-
TacoERE: Cluster-aware Compression for Event Relation Extraction 11 May 2024 · 0 repositories · arXiv:2405.06890
-
A Lightweight Sparse Focus Transformer for Remote Sensing Image Change Captioning 10 May 2024 · 1 repository · arXiv:2405.06598
-
A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models 10 May 2024 · 0 repositories · arXiv:2405.06211
-
An Assessment of Model-On-Model Deception 10 May 2024 · 0 repositories · arXiv:2405.12999
-
CANAL -- Cyber Activity News Alerting Language Model: Empirical Approach vs. Expensive LLM 10 May 2024 · 0 repositories · arXiv:2405.06772
-
Characterizing the Accuracy -- Efficiency Trade-off of Low-rank Decomposition in Language Models 10 May 2024 · 0 repositories · arXiv:2405.06626
-
ChatGPTest: opportunities and cautionary tales of utilizing AI for questionnaire pretesting 10 May 2024 · 0 repositories · arXiv:2405.06329
-
Common Corruptions for Enhancing and Evaluating Robustness in Air-to-Air Visual Object Detection 10 May 2024 · 0 repositories · arXiv:2405.06765
-
Dual-Task Vision Transformer for Rapid and Accurate Intracerebral Hemorrhage CT Image Classification 10 May 2024 · 0 repositories · arXiv:2405.06814
-
Large Language Model in Financial Regulatory Interpretation 10 May 2024 · 0 repositories · arXiv:2405.06808
-
Linearizing Large Language Models 10 May 2024 · 1 repository · arXiv:2405.06640Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 3 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 6 pointer-only (licence)
-
Mesh Denoising Transformer 10 May 2024 · 0 repositories · arXiv:2405.06536
-
Multimodal LLMs Struggle with Basic Visual Network Analysis: a VNA Benchmark 10 May 2024 · 1 repository · arXiv:2405.06634
-
A Mixture of Experts Approach to 3D Human Motion Prediction 9 May 2024 · 1 repository · arXiv:2405.06088