Newest with code · page 5
Every paper with a repository link. Each page shows two streams, newest first within each, counted separately: papers newer than the archive snapshot come from Syntology's graph Syntology; the rest are archive rows archive 2025-07-28. The two are never added together.
Newer than the archive snapshot Syntology
Cards 61–75 of 9,581 graph papers newer than 2025-07-28; this feed shows the newest 150, newest arXiv id first. Dates and the abstract sentence are from arXiv's metadata (CC0) for 9,280 of 9,581; for the other 301 the month is read from the id.
From the archive archive 2025-07-28
Cards 61–75 of 218,469 archive papers with a code link; this feed shows the newest 150, archive date first (newest archive date 2025-09-24). Within a month, dated rows come first, then the 49,179 undated rows placed by the month in their arXiv id. 4 archive rows carry a date after the snapshot and are placed by that date. 382 archive papers with neither a date nor an arXiv id cannot be placed and are not listed. 1 code-linked slug has no paper row in the archive and is not listed (so 218,469 listed + 382 unplaceable + 1 = 218,852 papers with code).
UGC-VideoCaptioner: An Omni UGC Video Detail Caption Model and New Benchmarks
Real-world user-generated videos, especially on platforms like TikTok, often feature rich and intertwined audio visual content.
MonoMVSNet: Monocular Priors Guided Multi-View Stereo Network
Learning-based Multi-View Stereo (MVS) methods aim to predict depth maps for a sequence of calibrated images to recover dense point clouds.
SystolicAttention: Fusing FlashAttention within a Single Systolic Array
Transformer models rely heavily on scaled dot-product attention (SDPA), typically implemented using the FlashAttention algorithm.
KV-Latent: Dimensional-level KV Cache Reduction with Frequency-aware Rotary Positional Embedding
Large language models (LLMs) based on Transformer Decoders have become the preferred choice for conversational generative AI.
MFGDiffusion: Mask-Guided Smoke Synthesis for Enhanced Forest Fire Detection
Smoke is the first visible indicator of a wildfire.With the advancement of deep learning, image-based smoke detection has become a crucial method for detecting and preventing forest fires.
Fairness-Aware Grouping for Continuous Sensitive Variables: Application for Debiasing Face Analysis with respect to Skin Tone
Within a legal framework, fairness in datasets and models is typically assessed by dividing observations into predefined groups and then computing fairness measures (e.g., Disparate Impact or Equality of Odds with…
Optimal Sensor Scheduling and Selection for Continuous-Discrete Kalman Filtering with Auxiliary Dynamics
We study the Continuous-Discrete Kalman Filter (CD-KF) for State-Space Models (SSMs) where continuous-time dynamics are observed via multiple sensors with discrete, irregularly timed measurements.
Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities
Vision-language models (VLMs) are increasingly applied to identify unsafe or inappropriate images due to their internal ethical standards and powerful reasoning abilities.
Hashed Watermark as a Filter: Defeating Forging and Overwriting Attacks in Weight-based Neural Network Watermarking
As valuable digital assets, deep neural networks necessitate robust ownership protection, positioning neural network watermarking (NNW) as a promising solution.
Interpretable Bayesian Tensor Network Kernel Machines with Automatic Rank and Feature Selection
Tensor Network (TN) Kernel Machines speed up model learning by representing parameters as low-rank TNs, reducing computation and memory use.
MMOne: Representing Multiple Modalities in One Scene
Humans perceive the world through multimodal cues to understand and interact with the environment.
Try Harder: Hard Sample Generation and Learning for Clothes-Changing Person Re-ID
Hard samples pose a significant challenge in person re-identification (ReID) tasks, particularly in clothing-changing person Re-ID (CC-ReID).
The Devil behind the mask: An emergent safety vulnerability of Diffusion LLMs
Diffusion-based large language models (dLLMs) have recently emerged as a powerful alternative to autoregressive LLMs, offering faster inference and greater interactivity via parallel decoding and bidirectional modeling.
GKNet: Graph-based Keypoints Network for Monocular Pose Estimation of Non-cooperative Spacecraft
Monocular pose estimation of non-cooperative spacecraft is significant for on-orbit service (OOS) tasks, such as satellite maintenance, space debris removal, and station assembly.
Personalized Exercise Recommendation with Semantically-Grounded Knowledge Tracing
We introduce ExRec, a general framework for personalized exercise recommendation with semantically-grounded knowledge tracing.
The feed is static: 10 pages of up to 15 cards per stream, rebuilt with the site. Older papers are reachable from task, dataset and method pages and from search. No repository stars are tracked and nothing here is ranked by popularity. Machine-readable twin: JSON.