Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 8
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 8 of 56: papers 701 to 800 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Uncovering the Limitations of Query Performance Prediction: Failures, Insights, and Implications for Selective Query Processing 1 Apr 2025 · 0 repositories · arXiv:2504.01101
-
CaLiV: LiDAR-to-Vehicle Calibration of Arbitrary Sensor Setups via Object Reconstruction 31 Mar 2025 · 1 repository · arXiv:2504.01987
-
Implicit In-Context Learning: Evidence from Artificial Language Experiments 31 Mar 2025 · 0 repositories · arXiv:2503.24190
-
Learning a Canonical Basis of Human Preferences from Binary Ratings 31 Mar 2025 · 0 repositories · arXiv:2503.24150
-
MolGround: A Benchmark for Molecular Grounding 31 Mar 2025 · 0 repositories · arXiv:2503.23668
-
Style Quantization for Data-Efficient GAN Training 31 Mar 2025 · 0 repositories · arXiv:2503.24282
-
TransVFC: A Transformable Video Feature Compression Framework for Machines 31 Mar 2025 · 1 repository · arXiv:2503.23772
-
Uni-Render: A Unified Accelerator for Real-Time Rendering Across Diverse Neural Renderers 31 Mar 2025 · 0 repositories · arXiv:2503.23644
-
CoRanking: Collaborative Ranking with Small and Large Ranking Agents 30 Mar 2025 · 0 repositories · arXiv:2503.23427
-
GMapLatent: Geometric Mapping in Latent Space 30 Mar 2025 · 0 repositories · arXiv:2503.23407
-
Re-Aligning Language to Visual Objects with an Agentic Workflow 30 Mar 2025 · 0 repositories · arXiv:2503.23508
-
Spatiotemporal Learning of Brain Dynamics from fMRI Using Frequency-Specific Multi-Band Attention for Cognitive and Psychiatric Applications 30 Mar 2025 · 1 repository · arXiv:2503.23394
-
VideoFusion: A Spatio-Temporal Collaborative Network for Mutli-modal Video Fusion and Restoration 30 Mar 2025 · 0 repositories · arXiv:2503.23359
-
Evaluating how LLM annotations represent diverse views on contentious topics 29 Mar 2025 · 0 repositories · arXiv:2503.23243
-
Iterative VCG-based Mechanism Fosters Cooperation in Multi-Regional Network Design 29 Mar 2025 · 0 repositories · arXiv:2503.23255
-
Large Self-Supervised Models Bridge the Gap in Domain Adaptive Object Detection 29 Mar 2025 · 1 repository · arXiv:2503.23220
-
The geomagnetic storm and Kp prediction using Wasserstein transformer 29 Mar 2025 · 0 repositories · arXiv:2503.23102
-
A Cooperative Compliance Control Framework for Socially Optimal Mixed Traffic Routing 28 Mar 2025 · 0 repositories · arXiv:2503.22837
-
A Semantic-Enhanced Heterogeneous Graph Learning Method for Flexible Objects Recognition 28 Mar 2025 · 0 repositories · arXiv:2503.22079
-
Agent-Centric Personalized Multiple Clustering with Multi-Modal LLMs 28 Mar 2025 · 0 repositories · arXiv:2503.22241
-
DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness 28 Mar 2025 · 0 repositories · arXiv:2503.22677
-
Information Gain Is Not All You Need 28 Mar 2025 · 1 repository · arXiv:2504.01980
-
Mitigating Knowledge Discrepancies among Multiple Datasets for Task-agnostic Unified Face Alignment 28 Mar 2025 · 0 repositories · arXiv:2503.22359
-
Spatial Transport Optimization by Repositioning Attention Map for Training-Free Text-to-Image Synthesis 28 Mar 2025 · 0 repositories · arXiv:2503.22168
-
Cognitive Science-Inspired Evaluation of Core Capabilities for Object Understanding in AI 27 Mar 2025 · 0 repositories · arXiv:2503.21668
-
Collab: Controlled Decoding using Mixture of Agents for LLM Alignment 27 Mar 2025 · 0 repositories · arXiv:2503.21720
-
DynamiCtrl: Rethinking the Basic Structure and the Role of Text for High-quality Human Image Animation 27 Mar 2025 · 1 repository · arXiv:2503.21246
-
Evaluating book summaries from internal knowledge in Large Language Models: a cross-model and semantic consistency approach 27 Mar 2025 · 0 repositories · arXiv:2503.21613
-
Fusion of Graph Neural Networks via Optimal Transport 27 Mar 2025 · 0 repositories · arXiv:2503.21579
-
Multimodal Data Integration for Sustainable Indoor Gardening: Tracking Anyplant with Time Series Foundation Model 27 Mar 2025 · 0 repositories · arXiv:2503.21932
-
Rerouting Connection: Hybrid Computer Vision Analysis Reveals Visual Similarity Between Indus and Tibetan-Yi Corridor Writing Systems 27 Mar 2025 · 1 repository · arXiv:2503.21074
-
Retrieving Time-Series Differences Using Natural Language Queries 27 Mar 2025 · 0 repositories · arXiv:2503.21378
-
Reward Design for Reinforcement Learning Agents 27 Mar 2025 · 1 repository · arXiv:2503.21949
-
StyleMotif: Multi-Modal Motion Stylization using Style-Content Cross Fusion 27 Mar 2025 · 0 repositories · arXiv:2503.21775
-
Consistency Trajectory Matching for One-Step Generative Super-Resolution 26 Mar 2025 · 0 repositories · arXiv:2503.20349
-
Eyes Tell the Truth: GazeVal Highlights Shortcomings of Generative AI in Medical Imaging 26 Mar 2025 · 0 repositories · arXiv:2503.20967
-
GatedxLSTM: A Multimodal Affective Computing Approach for Emotion Recognition in Conversations 26 Mar 2025 · 0 repositories · arXiv:2503.20919
-
InfoBid: A Simulation Framework for Studying Information Disclosure in Auctions with Large Language Model-based Agents 26 Mar 2025 · 0 repositories · arXiv:2503.22726
-
Iterative Prompting with Persuasion Skills in Jailbreaking Large Language Models 26 Mar 2025 · 0 repositories · arXiv:2503.20320
-
MoRE-LLM: Mixture of Rule Experts Guided by a Large Language Model 26 Mar 2025 · 1 repository · arXiv:2503.22731
-
Omnidirectional Depth-Aided Occupancy Prediction based on Cylindrical Voxel for Autonomous Driving 26 Mar 2025 · 0 repositories · arXiv:2504.01023
-
Optimizing Case-Based Reasoning System for Functional Test Script Generation with Large Language Models 26 Mar 2025 · 0 repositories · arXiv:2503.20576
-
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics 26 Mar 2025 · 0 repositories · arXiv:2503.20308
-
Physics-Informed Neural Networks with Unknown Partial Differential Equations: an Application in Multivariate Time Series 26 Mar 2025 · 0 repositories · arXiv:2503.20144
-
Sacred or Secular? Religious Bias in AI-Generated Financial Advice 26 Mar 2025 · 0 repositories · arXiv:2504.07118
-
Towards Practical Emotion Recognition: An Unsupervised Source-Free Approach for EEG Domain Adaptation 26 Mar 2025 · 0 repositories · arXiv:2504.03707
-
Uncertainty Weighted Gradients for Model Calibration 26 Mar 2025 · 1 repository · arXiv:2503.22725Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
UWarp: A Whole Slide Image Registration Pipeline to Characterize Scanner-Induced Local Domain Shift 26 Mar 2025 · 0 repositories · arXiv:2503.20653
-
World Model Agents with Change-Based Intrinsic Motivation 26 Mar 2025 · 1 repository · arXiv:2503.21047
-
Zero-Shot Audio-Visual Editing via Cross-Modal Delta Denoising 26 Mar 2025 · 0 repositories · arXiv:2503.20782
-
A minimal electrical model of the human heart 25 Mar 2025 · 0 repositories · arXiv:2503.19578
-
AI Identity, Empowerment, and Mindfulness in Mitigating Unethical AI Use 25 Mar 2025 · 0 repositories · arXiv:2503.20099
-
Analyzable Chain-of-Musical-Thought Prompting for High-Fidelity Music Generation 25 Mar 2025 · 0 repositories · arXiv:2503.19611
-
Context-Aware Semantic Segmentation: Enhancing Pixel-Level Understanding with Large Language Models for Advanced Vision Applications 25 Mar 2025 · 0 repositories · arXiv:2503.19276
-
From Sparse to Dense: Camera Relocalization with Scene-Specific Detector from Feature Gaussian Splatting 25 Mar 2025 · 0 repositories · arXiv:2503.19358
-
Improved Alignment of Modalities in Large Vision Language Models 25 Mar 2025 · 0 repositories · arXiv:2503.19508
-
Kernel Learning Assisted Synthesis Condition Exploration for Ternary Spinel 25 Mar 2025 · 1 repository · arXiv:2503.19637
-
Large Language Models Meet Contrastive Learning: Zero-Shot Emotion Recognition Across Languages 25 Mar 2025 · 1 repository · arXiv:2503.21806
-
Towards Long-Range ENSO Prediction with an Explainable Deep Learning Model 25 Mar 2025 · 0 repositories · arXiv:2503.19502
-
Unpaired Translation of Chest X-ray Images for Lung Opacity Diagnosis via Adaptive Activation Masks and Cross-Domain Alignment 25 Mar 2025 · 0 repositories · arXiv:2503.19860
-
Vanishing Depth: A Depth Adapter with Positional Depth Encoding for Generalized Image Encoders 25 Mar 2025 · 1 repository · arXiv:2503.19947
-
Visuo-Tactile Object Pose Estimation for a Multi-Finger Robot Hand with Low-Resolution In-Hand Tactile Sensing 25 Mar 2025 · 0 repositories · arXiv:2503.19893
-
Bridging Writing Manner Gap in Visual Instruction Tuning by Creating LLM-aligned Instructions 24 Mar 2025 · 0 repositories · arXiv:2503.18320
-
InPO: Inversion Preference Optimization with Reparametrized DDIM for Efficient Diffusion Model Alignment 24 Mar 2025 · 1 repository · arXiv:2503.18454
-
Instruct-CLIP: Improving Instruction-Guided Image Editing with Automated Data Refinement Using Contrastive Learning 24 Mar 2025 · 1 repository · arXiv:2503.18406Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Mitigating Cache Noise in Test-Time Adaptation for Large Vision-Language Models 24 Mar 2025 · 0 repositories · arXiv:2503.18334
-
MonoInstance: Enhancing Monocular Priors via Multi-view Instance Alignment for Neural Rendering and Reconstruction 24 Mar 2025 · 0 repositories · arXiv:2503.18363
-
Neuro-symbolic Weak Supervision: Theory and Semantics 24 Mar 2025 · 0 repositories · arXiv:2503.18509
-
Regional House Price Dynamics in Australia: Insights into Lifestyle and Mining Dynamics through PCA 24 Mar 2025 · 0 repositories · arXiv:2503.18332
-
Coverage-Guaranteed Speech Emotion Recognition via Calibrated Uncertainty-Adaptive Prediction Sets 24 Mar 2025 · 0 repositories · arXiv:2503.22712
-
Safeguarding Mobile GUI Agent via Logic-based Action Verification 24 Mar 2025 · 0 repositories · arXiv:2503.18492
-
TARDIS: Mitigate Temporal Misalignment via Representation Steering 24 Mar 2025 · 0 repositories · arXiv:2503.18693
-
Towards Training-free Anomaly Detection with Vision and Language Foundation Models 24 Mar 2025 · 1 repository · arXiv:2503.18325Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
An AI-enabled dual-hormone model predictive control algorithm that delivers insulin and pramlintide 23 Mar 2025 · 0 repositories · arXiv:2503.17887
-
DualCP: Rehearsal-Free Domain-Incremental Learning via Dual-Level Concept Prototype 23 Mar 2025 · 0 repositories · arXiv:2503.18042
-
Human-AI Interaction and User Satisfaction: Empirical Evidence from Online Reviews of AI Products 23 Mar 2025 · 0 repositories · arXiv:2503.17955
-
LocDiffusion: Identifying Locations on Earth by Diffusing in the Hilbert Space 23 Mar 2025 · 0 repositories · arXiv:2503.18142
-
MedPlan:A Two-Stage RAG-Based System for Personalized Medical Plan Generation 23 Mar 2025 · 0 repositories · arXiv:2503.17900
-
MLLM-For3D: Adapting Multimodal Large Language Model for 3D Reasoning Segmentation 23 Mar 2025 · 0 repositories · arXiv:2503.18135
-
PHT-CAD: Efficient CAD Parametric Primitive Analysis with Progressive Hierarchical Tuning 23 Mar 2025 · 1 repository · arXiv:2503.18147
-
Aligning Foundation Model Priors and Diffusion-Based Hand Interactions for Occlusion-Resistant Two-Hand Reconstruction 22 Mar 2025 · 0 repositories · arXiv:2503.17788
-
Collaborative Temporal Consistency Learning for Point-supervised Natural Language Video Localization 22 Mar 2025 · 0 repositories · arXiv:2503.17651
-
Enhancing Persona Consistency for LLMs' Role-Playing using Persona-Aware Contrastive Learning 22 Mar 2025 · 0 repositories · arXiv:2503.17662
-
HiLoTs: High-Low Temporal Sensitive Representation Learning for Semi-Supervised LiDAR Segmentation in Autonomous Driving 22 Mar 2025 · 1 repository · arXiv:2503.17752
-
OMR-Diffusion:Optimizing Multi-Round Enhanced Training in Diffusion Models for Improved Intent Understanding 22 Mar 2025 · 0 repositories · arXiv:2503.17660
-
Towards Invisible Backdoor Attack on Text-to-Image Diffusion Model 22 Mar 2025 · 1 repository · arXiv:2503.17724
-
Echo-E³Net: Efficient Endo-Epi Spatio-Temporal Network for Ejection Fraction Estimation 21 Mar 2025 · 0 repositories · arXiv:2503.17543
-
From Faces to Voices: Learning Hierarchical Representations for High-quality Video-to-Speech 21 Mar 2025 · 0 repositories · arXiv:2503.16956
-
Generative Modeling of Class Probability for Multi-Modal Representation Learning 21 Mar 2025 · 0 repositories · arXiv:2503.17417
-
ProDehaze: Prompting Diffusion Models Toward Faithful Image Dehazing 21 Mar 2025 · 1 repository · arXiv:2503.17488
-
Understanding Social Support Needs in Questions: A Hybrid Approach Integrating Semi-Supervised Learning and LLM-based Data Augmentation 21 Mar 2025 · 0 repositories · arXiv:2503.17421
-
Vision-Language Gradient Descent-driven All-in-One Deep Unfolding Networks 21 Mar 2025 · 0 repositories · arXiv:2503.16930
-
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding 20 Mar 2025 · 1 repository · arXiv:2503.16707Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Cultural Alignment in Large Language Models Using Soft Prompt Tuning 20 Mar 2025 · 0 repositories · arXiv:2503.16094
-
Efficient ANN-Guided Distillation: Aligning Rate-based Features of Spiking Neural Networks through Hybrid Block-wise Replacement 20 Mar 2025 · 0 repositories · arXiv:2503.16572
-
FedAWA: Adaptive Optimization of Aggregation Weights in Federated Learning Using Client Vectors 20 Mar 2025 · 1 repository · arXiv:2503.15842
-
From Structured Prompts to Open Narratives: Measuring Gender Bias in LLMs Through Open-Ended Storytelling 20 Mar 2025 · 0 repositories · arXiv:2503.15904
-
GAIR: Improving Multimodal Geo-Foundation Model with Geo-Aligned Implicit Representations 20 Mar 2025 · 0 repositories · arXiv:2503.16683
-
InCo-DPO: Balancing Distribution Shift and Data Quality for Enhanced Preference Optimization 20 Mar 2025 · 0 repositories · arXiv:2503.15880
-
Learning 3D Scene Analogies with Neural Contextual Scene Maps 20 Mar 2025 · 0 repositories · arXiv:2503.15897