Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 9
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 9 of 56: papers 801 to 900 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MobiFuse: Learning Universal Human Mobility Patterns through Cross-domain Data Fusion 20 Mar 2025 · 0 repositories · arXiv:2503.15779
-
RL4Med-DDPO: Reinforcement Learning for Controlled Guidance Towards Diverse Medical Image Generation using Vision-Language Foundation Models 20 Mar 2025 · 0 repositories · arXiv:2503.15784
-
Towards Agentic Recommender Systems in the Era of Multimodal Large Language Models 20 Mar 2025 · 0 repositories · arXiv:2503.16734
-
UniCrossAdapter: Multimodal Adaptation of CLIP for Radiology Report Generation 20 Mar 2025 · 1 repository · arXiv:2503.15940
-
Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion 20 Mar 2025 · 1 repository · arXiv:2503.15851
-
Aligning Information Capacity Between Vision and Language via Dense-to-Sparse Feature Distillation for Image-Text Matching 19 Mar 2025 · 1 repository · arXiv:2503.14953
-
Enhancing Fault Detection and Isolation in an All-Electric Auxiliary Power Unit (APU) Gas Generator by Utilizing Starter/Generator Signal 19 Mar 2025 · 0 repositories · arXiv:2503.14986
-
FP4DiT: Towards Effective Floating Point Quantization for Diffusion Transformers 19 Mar 2025 · 1 repository · arXiv:2503.15465Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
ML-Triton, A Multi-Level Compilation and Language Extension to Triton GPU Programming 19 Mar 2025 · 0 repositories · arXiv:2503.14985
-
Optimizing Decomposition for Optimal Claim Verification 19 Mar 2025 · 1 repository · arXiv:2503.15354
-
POSTA: A Go-to Framework for Customized Artistic Poster Generation 19 Mar 2025 · 0 repositories · arXiv:2503.14908
-
Uncertainty-Aware Diffusion Guided Refinement of 3D Scenes 19 Mar 2025 · 0 repositories · arXiv:2503.15742
-
Bridging Past and Future: End-to-End Autonomous Driving with Historical Prediction and Planning 18 Mar 2025 · 1 repository · arXiv:2503.14182
-
EEG-CLIP : Learning EEG representations from natural language descriptions 18 Mar 2025 · 1 repository · arXiv:2503.16531
-
EvolvingGrasp: Evolutionary Grasp Generation via Efficient Preference Alignment 18 Mar 2025 · 0 repositories · arXiv:2503.14329
-
Iffy-Or-Not: Extending the Web to Support the Critical Evaluation of Fallacious Texts 18 Mar 2025 · 0 repositories · arXiv:2503.14412
-
KANITE: Kolmogorov-Arnold Networks for ITE estimation 18 Mar 2025 · 0 repositories · arXiv:2503.13912
-
Large Language Models for Virtual Human Gesture Selection 18 Mar 2025 · 0 repositories · arXiv:2503.14408
-
Limb-Aware Virtual Try-On Network with Progressive Clothing Warping 18 Mar 2025 · 1 repository · arXiv:2503.14074
-
Make the Most of Everything: Further Considerations on Disrupting Diffusion-based Customization 18 Mar 2025 · 0 repositories · arXiv:2503.13945
-
Marten: Visual Question Answering with Mask Generation for Multi-modal Document Understanding 18 Mar 2025 · 0 repositories · arXiv:2503.14140Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
MTLoc: A Confidence-Based Source-Free Domain Adaptation Approach For Indoor Localization 18 Mar 2025 · 0 repositories · arXiv:2503.14767
-
MusicInfuser: Making Video Diffusion Listen and Dance 18 Mar 2025 · 0 repositories · arXiv:2503.14505
-
Second language Korean Universal Dependency treebank v1.2: Focus on data augmentation and annotation scheme refinement 18 Mar 2025 · 1 repository · arXiv:2503.14718
-
See-Saw Modality Balance: See Gradient, and Sew Impaired Vision-Language Balance to Mitigate Dominant Modality Bias 18 Mar 2025 · 0 repositories · arXiv:2503.13834
-
VEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation 18 Mar 2025 · 0 repositories · arXiv:2503.14350
-
Advancing Chronic Tuberculosis Diagnostics Using Vision-Language Models: A Multi modal Framework for Precision Analysis 17 Mar 2025 · 0 repositories · arXiv:2503.14536
-
HiMTok: Learning Hierarchical Mask Tokens for Image Segmentation with Large Multimodal Model 17 Mar 2025 · 1 repository · arXiv:2503.13026Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)
-
Integrating AI for Human-Centric Breast Cancer Diagnostics: A Multi-Scale and Multi-View Swin Transformer Framework 17 Mar 2025 · 0 repositories · arXiv:2503.13309
-
Island-Based Evolutionary Computation with Diverse Surrogates and Adaptive Knowledge Transfer for High-Dimensional Data-Driven Optimization 17 Mar 2025 · 1 repository · arXiv:2503.12856
-
Optimal compound downselection to promote diversity and parallel chemistry 17 Mar 2025 · 0 repositories · arXiv:2503.13627
-
Progressive Human Motion Generation Based on Text and Few Motion Frames 17 Mar 2025 · 1 repository · arXiv:2503.13300
-
TablePilot: Recommending Human-Preferred Tabular Data Analysis with Large Language Models 17 Mar 2025 · 0 repositories · arXiv:2503.13262
-
Enhancing Visual Representation with Textual Semantics: Textual Semantics-Powered Prototypes for Heterogeneous Federated Learning 16 Mar 2025 · 0 repositories · arXiv:2503.13543
-
MagicID: Hybrid Preference Optimization for ID-Consistent and Dynamic-Preserved Video Customization 16 Mar 2025 · 0 repositories · arXiv:2503.12689
-
A State Alignment-Centric Approach to Federated System Identification: The FedAlign Framework 15 Mar 2025 · 0 repositories · arXiv:2503.12137
-
Aerial Vision-and-Language Navigation with Grid-based View Selection and Map Construction 14 Mar 2025 · 0 repositories · arXiv:2503.11091
-
Cyclic Contrastive Knowledge Transfer for Open-Vocabulary Object Detection 14 Mar 2025 · 1 repository · arXiv:2503.11005
-
Examples as the Prompt: A Scalable Approach for Efficient LLM Adaptation in E-Commerce 14 Mar 2025 · 0 repositories · arXiv:2503.13518
-
Optimal Transport and Adaptive Thresholding for Universal Domain Adaptation on Time Series 14 Mar 2025 · 1 repository · arXiv:2503.11217
-
PARIC: Probabilistic Attention Regularization for Language Guided Image Classification from Pre-trained Vison Language Models 14 Mar 2025 · 0 repositories · arXiv:2503.11360
-
Quantifying Interpretability in CLIP Models with Concept Consistency 14 Mar 2025 · 0 repositories · arXiv:2503.11103
-
T2I-FineEval: Fine-Grained Compositional Metric for Text-to-Image Evaluation 14 Mar 2025 · 1 repository · arXiv:2503.11481
-
TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools 14 Mar 2025 · 4 repositories · arXiv:2503.10970Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Zero-TIG: Temporal Consistency-Aware Zero-Shot Illumination-Guided Low-light Video Enhancement 14 Mar 2025 · 1 repository · arXiv:2503.11175
-
A Hierarchical Semantic Distillation Framework for Open-Vocabulary Object Detection 13 Mar 2025 · 1 repository · arXiv:2503.10152
-
Bayesian Prompt Flow Learning for Zero-Shot Anomaly Detection 13 Mar 2025 · 1 repository · arXiv:2503.10080Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Cosh-DiT: Co-Speech Gesture Video Synthesis via Hybrid Audio-Visual Diffusion Transformers 13 Mar 2025 · 0 repositories · arXiv:2503.09942
-
Finetuning Generative Trajectory Model with Reinforcement Learning from Human Feedback 13 Mar 2025 · 0 repositories · arXiv:2503.10434
-
GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing 13 Mar 2025 · 1 repository · arXiv:2503.10639Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
NeighborRetr: Balancing Hub Centrality in Cross-Modal Retrieval 13 Mar 2025 · 1 repository · arXiv:2503.10526
-
PluralLLM: Pluralistic Alignment in LLMs via Federated Learning 13 Mar 2025 · 0 repositories · arXiv:2503.09925
-
RankPO: Preference Optimization for Job-Talent Matching 13 Mar 2025 · 1 repository · arXiv:2503.10723
-
VMBench: A Benchmark for Perception-Aligned Video Motion Generation 13 Mar 2025 · 1 repository · arXiv:2503.10076Syntology official: no sample here; runs from other or unrecorded repositories · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Aligning to What? Limits to RLHF Based Alignment 12 Mar 2025 · 1 repository · arXiv:2503.09025
-
BAMBI: Developing Baby Language Models for Italian 12 Mar 2025 · 0 repositories · arXiv:2503.09481
-
Can A Society of Generative Agents Simulate Human Behavior and Inform Public Health Policy? A Case Study on Vaccine Hesitancy 12 Mar 2025 · 0 repositories · arXiv:2503.09639
-
Exploring the best way for UAV visual localization under Low-altitude Multi-view Observation Condition: a Benchmark 12 Mar 2025 · 1 repository · arXiv:2503.10692
-
Got Compute, but No Data: Lessons From Post-training a Finnish LLM 12 Mar 2025 · 0 repositories · arXiv:2503.09407
-
EEG relative phase-based analysis unveils the complexity and universality of human brain dynamics: integrative insights from general anesthesia and ADHD 12 Mar 2025 · 0 repositories · arXiv:2503.19924
-
Incomplete Multi-view Clustering via Diffusion Contrastive Generation 12 Mar 2025 · 0 repositories · arXiv:2503.09185
-
JBFuzz: Jailbreaking LLMs Efficiently and Effectively Using Fuzzing 12 Mar 2025 · 0 repositories · arXiv:2503.08990
-
Learning Cascade Ranking as One Network 12 Mar 2025 · 0 repositories · arXiv:2503.09492Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Online Language Splatting 12 Mar 2025 · 0 repositories · arXiv:2503.09447
-
OpenVidVRD: Open-Vocabulary Video Visual Relation Detection via Prompt-Driven Semantic Space Alignment 12 Mar 2025 · 0 repositories · arXiv:2503.09416
-
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latant Space 12 Mar 2025 · 0 repositories · arXiv:2503.09215
-
SciFi-Benchmark: How Would AI-Powered Robots Behave in Science Fiction Literature? 12 Mar 2025 · 0 repositories · arXiv:2503.10706
-
Towards Robust Model Evolution with Algorithmic Recourse 12 Mar 2025 · 0 repositories · arXiv:2503.09658
-
Decentralized Integration of Grid Edge Resources into Wholesale Electricity Markets via Mean-field Games 11 Mar 2025 · 0 repositories · arXiv:2503.07984
-
From Slices to Sequences: Autoregressive Tracking Transformer for Cohesive and Consistent 3D Lymph Node Detection in CT Scans 11 Mar 2025 · 0 repositories · arXiv:2503.07933
-
Hedonic Adaptation in the Age of AI: A Perspective on Diminishing Satisfaction Returns in Technology Adoption 11 Mar 2025 · 0 repositories · arXiv:2503.08074
-
LangTime: A Language-Guided Unified Model for Time Series Forecasting with Proximal Policy Optimization 11 Mar 2025 · 0 repositories · arXiv:2503.08271
-
MMRL: Multi-Modal Representation Learning for Vision-Language Models 11 Mar 2025 · 1 repository · arXiv:2503.08497Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
"Principal Components" Enable A New Language of Images 11 Mar 2025 · 1 repository · arXiv:2503.08685Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 3 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
ReviewAgents: Bridging the Gap Between Human and AI-Generated Paper Reviews 11 Mar 2025 · 0 repositories · arXiv:2503.08506
-
SQLCritic: Correcting Text-to-SQL Generation via Clause-wise Critic 11 Mar 2025 · 0 repositories · arXiv:2503.07996
-
Understanding and Mitigating Distribution Shifts For Machine Learning Force Fields 11 Mar 2025 · 0 repositories · arXiv:2503.08674
-
ADROIT: A Self-Supervised Framework for Learning Robust Representations for Active Learning 10 Mar 2025 · 0 repositories · arXiv:2503.07506
-
Aligning Instance-Semantic Sparse Representation towards Unsupervised Object Segmentation and Shape Abstraction with Repeatable Primitives 10 Mar 2025 · 0 repositories · arXiv:2503.06947
-
Beyond One-Size-Fits-All Summarization: Customizing Summaries for Diverse Users 10 Mar 2025 · 0 repositories · arXiv:2503.10675
-
Customized SAM 2 for Referring Remote Sensing Image Segmentation 10 Mar 2025 · 0 repositories · arXiv:2503.07266
-
Enhancing Time Series Forecasting via Logic-Inspired Regularization 10 Mar 2025 · 0 repositories · arXiv:2503.06867
-
Erase Diffusion: Empowering Object Removal Through Calibrating Diffusion Pathways 10 Mar 2025 · 0 repositories · arXiv:2503.07026
-
Generative AI in Transportation Planning: A Survey 10 Mar 2025 · 0 repositories · arXiv:2503.07158
-
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection 10 Mar 2025 · 0 repositories · arXiv:2503.07593
-
LLMs syntactically adapt their language use to their conversational partner 10 Mar 2025 · 0 repositories · arXiv:2503.07457
-
MADS: Multi-Attribute Document Supervision for Zero-Shot Image Classification 10 Mar 2025 · 0 repositories · arXiv:2503.06847
-
Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model 10 Mar 2025 · 1 repository · arXiv:2503.07703Syntology 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
The Impact of Generative AI Coding Assistants on Developers Who Are Visually Impaired 10 Mar 2025 · 0 repositories · arXiv:2503.16491
-
The Sustainable Future is now: a dynamic model to advance investments in PV and Energy Storage 10 Mar 2025 · 0 repositories · arXiv:2503.07131
-
Training Domain Draft Models for Speculative Decoding: Best Practices and Insights 10 Mar 2025 · 0 repositories · arXiv:2503.07807
-
Trustworthy Machine Learning via Memorization and the Granular Long-Tail: A Survey on Interactions, Tradeoffs, and Beyond 10 Mar 2025 · 0 repositories · arXiv:2503.07501
-
Advancing AI Negotiations: New Theory and Evidence from a Large-Scale Autonomous Negotiations Competition 9 Mar 2025 · 0 repositories · arXiv:2503.06416
-
Color Alignment in Diffusion 9 Mar 2025 · 0 repositories · arXiv:2503.06746
-
Dr Genre: Reinforcement Learning from Decoupled LLM Feedback for Generic Text Rewriting 9 Mar 2025 · 0 repositories · arXiv:2503.06781
-
ExGes: Expressive Human Motion Retrieval and Modulation for Audio-Driven Gesture Synthesis 9 Mar 2025 · 0 repositories · arXiv:2503.06499
-
FEDS: Feature and Entropy-Based Distillation Strategy for Efficient Learned Image Compression 9 Mar 2025 · 0 repositories · arXiv:2503.06399
-
GenDR: Lightning Generative Detail Restorator 9 Mar 2025 · 0 repositories · arXiv:2503.06790
-
Machine Learning meets Algebraic Combinatorics: A Suite of Datasets Capturing Research-level Conjecturing Ability in Pure Mathematics 9 Mar 2025 · 0 repositories · arXiv:2503.06366
-
Optimal Transport for Brain-Image Alignment: Unveiling Redundancy and Synergy in Neural Information Processing 9 Mar 2025 · 0 repositories · arXiv:2503.10663