Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 18
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 18 of 56: papers 1,701 to 1,800 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
VLSBench: Unveiling Visual Leakage in Multimodal Safety 29 Nov 2024 · 1 repository · arXiv:2411.19939Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
EzSQL: An SQL intermediate representation for improving SQL-to-text Generation 28 Nov 2024 · 0 repositories · arXiv:2411.18923
-
GRAPE: Generalizing Robot Policy via Preference Alignment 28 Nov 2024 · 0 repositories · arXiv:2411.19309
-
InstanceGaussian: Appearance-Semantic Joint Gaussian Representation for 3D Instance-Level Perception 28 Nov 2024 · 0 repositories · arXiv:2411.19235
-
LoRA of Change: Learning to Generate LoRA for the Editing Instruction from A Single Before-After Image Pair 28 Nov 2024 · 0 repositories · arXiv:2411.19156
-
Mapping Public Perception of Artificial Intelligence: Expectations, Risk-Benefit Tradeoffs, and Value As Determinants for Societal Acceptance 28 Nov 2024 · 0 repositories · arXiv:2411.19356
-
Personalized Federated Fine-Tuning for LLMs via Data-Driven Heterogeneous Model Architectures 28 Nov 2024 · 1 repository · arXiv:2411.19128
-
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation 28 Nov 2024 · 1 repository · arXiv:2411.19331Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
3D Scene Graph Guided Vision-Language Pre-training 27 Nov 2024 · 0 repositories · arXiv:2411.18666
-
ELEMENTAL: Interactive Learning from Demonstrations and Vision-Language Models for Reward Design in Robotics 27 Nov 2024 · 0 repositories · arXiv:2411.18825
-
FaithDiff: Unleashing Diffusion Priors for Faithful Image Super-resolution 27 Nov 2024 · 0 repositories · arXiv:2411.18824
-
Large Scale Evaluation of Deep Learning-based Explainable Solar Flare Forecasting Models with Attribution-based Proximity Analysis 27 Nov 2024 · 0 repositories · arXiv:2411.18070
-
Manual-PA: Learning 3D Part Assembly from Instruction Diagrams 27 Nov 2024 · 0 repositories · arXiv:2411.18011
-
Optimal payoff under Bregman-Wasserstein divergence constraints 27 Nov 2024 · 0 repositories · arXiv:2411.18397
-
Different Bias Under Different Criteria: Assessing Bias in LLMs with a Fact-Based Approach 26 Nov 2024 · 1 repository · arXiv:2411.17338Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding 26 Nov 2024 · 1 repository · arXiv:2411.17481
-
FTMoMamba: Motion Generation with Frequency and Text State Space Models 26 Nov 2024 · 0 repositories · arXiv:2411.17532
-
g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks 26 Nov 2024 · 1 repository · arXiv:2411.17030Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Low-rank Adaptation-based All-Weather Removal for Autonomous Navigation 26 Nov 2024 · 0 repositories · arXiv:2411.17814
-
MUSE-VL: Modeling Unified VLM through Semantic Discrete Encoding 26 Nov 2024 · 0 repositories · arXiv:2411.17762
-
sbi reloaded: a toolkit for simulation-based inference workflows 26 Nov 2024 · 1 repository · arXiv:2411.17337Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
Signs as Tokens: An Autoregressive Multilingual Sign Language Generator 26 Nov 2024 · 0 repositories · arXiv:2411.17799
-
ThreatModeling-LLM: Automating Threat Modeling using Large Language Models for Banking System 26 Nov 2024 · 0 repositories · arXiv:2411.17058
-
What Differentiates Educational Literature? A Multimodal Fusion Approach of Transformers and Computational Linguistics 26 Nov 2024 · 0 repositories · arXiv:2411.17593
-
Automated Registration of 3D Neurovascular Territory Atlas to 2D DSA for Targeted Quantitative Angiography Analysis 25 Nov 2024 · 0 repositories · arXiv:2411.16637
-
Beyond Sight: Towards Cognitive Alignment in LVLM via Enriched Visual Knowledge 25 Nov 2024 · 0 repositories · arXiv:2411.16824
-
Do Activists Align with Larger Mutual Funds? 25 Nov 2024 · 0 repositories · arXiv:2411.16553
-
DoubleCCA: Improving Foundation Model Group Robustness with Random Sentence Embeddings 25 Nov 2024 · 0 repositories · arXiv:2411.16236
-
Hyperspectral Image Cross-Domain Object Detection Method based on Spectral-Spatial Feature Alignment 25 Nov 2024 · 0 repositories · arXiv:2411.16772
-
Inference-Time Policy Steering through Human Interactions 25 Nov 2024 · 0 repositories · arXiv:2411.16627
-
Leveraging the Power of MLLMs for Gloss-Free Sign Language Translation 25 Nov 2024 · 0 repositories · arXiv:2411.16789
-
SEMU-Net: A Segmentation-based Corrector for Fabrication Process Variations of Nanophotonics with Microscopic Images 25 Nov 2024 · 0 repositories · arXiv:2411.16973
-
DiffBreak: Is Diffusion-Based Purification Robust? 25 Nov 2024 · 1 repository · arXiv:2411.16598
-
Detecting Turkish Synonyms Used in Different Time Periods 24 Nov 2024 · 0 repositories · arXiv:2411.15768
-
Editable-DeepSC: Reliable Cross-Modal Semantic Communications for Facial Editing 24 Nov 2024 · 0 repositories · arXiv:2411.15702
-
LeMoLE: LLM-Enhanced Mixture of Linear Experts for Time Series Forecasting 24 Nov 2024 · 0 repositories · arXiv:2412.00053
-
PriorDiffusion: Leverage Language Prior in Diffusion Models for Monocular Depth Estimation 24 Nov 2024 · 0 repositories · arXiv:2411.16750
-
TableTime: Reformulating Time Series Classification as Zero-Shot Table Understanding via Large Language Models 24 Nov 2024 · 1 repository · arXiv:2411.15737
-
A Preliminary Study of Multilingual Code Language Models for Code Generation Task Using Translated Benchmarks 23 Nov 2024 · 0 repositories · arXiv:2411.15470
-
ConsistentAvatar: Learning to Diffuse Fully Consistent Talking Head Avatar with Temporal Guidance 23 Nov 2024 · 0 repositories · arXiv:2411.15436
-
FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation 23 Nov 2024 · 1 repository · arXiv:2411.15413Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples)
-
Quantitative Analysis of IITs' Research Growth and SDG Contributions 23 Nov 2024 · 0 repositories · arXiv:2411.15451
-
Twin Trigger Generative Networks for Backdoor Attacks against Object Detection 23 Nov 2024 · 0 repositories · arXiv:2411.15439
-
Continual SFT Matches Multimodal RLHF with Negative Supervision 22 Nov 2024 · 0 repositories · arXiv:2411.14797
-
Derivative-Free Diffusion Manifold-Constrained Gradient for Unified XAI 22 Nov 2024 · 1 repository · arXiv:2411.15265Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
Fine-Grained Alignment in Vision-and-Language Navigation through Bayesian Optimization 22 Nov 2024 · 0 repositories · arXiv:2411.14811
-
FOCUS: Knowledge-enhanced Adaptive Visual Compression for Few-shot Whole Slide Image Classification 22 Nov 2024 · 1 repository · arXiv:2411.14743Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Grid and Road Expressions Are Complementary for Trajectory Representation Learning 22 Nov 2024 · 1 repository · arXiv:2411.14768
-
Unsupervised Multi-view UAV Image Geo-localization via Iterative Rendering 22 Nov 2024 · 0 repositories · arXiv:2411.14816
-
Beyond Monte Carlo: Harnessing Diffusion Models to Simulate Financial Market Dynamics 21 Nov 2024 · 0 repositories · arXiv:2412.00036
-
LEADRE: Multi-Faceted Knowledge Enhanced LLM Empowered Display Advertisement Recommender System 21 Nov 2024 · 0 repositories · arXiv:2411.13789
-
Lost in Inference: Rediscovering the Role of Natural Language Inference for Large Language Models 21 Nov 2024 · 0 repositories · arXiv:2411.14103
-
Optimizing Student Ability Assessment: A Hierarchy Constraint-Aware Cognitive Diagnosis Framework for Educational Contexts 21 Nov 2024 · 0 repositories · arXiv:2412.04488
-
Process and Policy Insights from an Intercomparison of Open Electricity System Capacity Expansion Models 21 Nov 2024 · 0 repositories · arXiv:2411.13783
-
Test-Time Adaptation of 3D Point Clouds via Denoising Diffusion Models 21 Nov 2024 · 1 repository · arXiv:2411.14495
-
Efficient and Physically-Consistent Modeling of Reconfigurable Electromagnetic Structures 20 Nov 2024 · 0 repositories · arXiv:2411.13475
-
ESARM: 3D Emotional Speech-to-Animation via Reward Model from Automatically-Ranked Demonstrations 20 Nov 2024 · 0 repositories · arXiv:2411.13089
-
Hints of Prompt: Enhancing Visual Representation for Multimodal LLMs in Autonomous Driving 20 Nov 2024 · 0 repositories · arXiv:2411.13076
-
Improving OOD Generalization of Pre-trained Encoders via Aligned Embedding-Space Ensembles 20 Nov 2024 · 0 repositories · arXiv:2411.13073
-
MEGL: Multimodal Explanation-Guided Learning 20 Nov 2024 · 0 repositories · arXiv:2411.13053
-
On the Consistency of Video Large Language Models in Temporal Comprehension 20 Nov 2024 · 1 repository · arXiv:2411.12951Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Conditional Distribution Learning on Graphs 20 Nov 2024 · 1 repository · arXiv:2411.15206Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Unlocking Historical Clinical Trial Data with ALIGN: A Compositional Large Language Model System for Medical Coding 20 Nov 2024 · 0 repositories · arXiv:2411.13163
-
VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models 20 Nov 2024 · 1 repository · arXiv:2411.13503
-
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation 20 Nov 2024 · 1 repository · arXiv:2411.13243Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples)
-
3D Reconstruction by Looking: Instantaneous Blind Spot Detector for Indoor SLAM through Mixed Reality 19 Nov 2024 · 0 repositories · arXiv:2411.12514
-
C²INet: Realizing Incremental Trajectory Prediction with Prior-Aware Continual Causal Intervention 19 Nov 2024 · 0 repositories · arXiv:2411.12313
-
CCIS-Diff: A Generative Model with Stable Diffusion Prior for Controlled Colonoscopy Image Synthesis 19 Nov 2024 · 0 repositories · arXiv:2411.12198
-
JuniperLiu at CoMeDi Shared Task: Models as Annotators in Lexical Semantics Disagreements 19 Nov 2024 · 1 repository · arXiv:2411.12147
-
Guide-to-Explain for Controllable Summarization 19 Nov 2024 · 0 repositories · arXiv:2411.12460
-
Intelligent Tutors for Adult Learners: An Analysis of Needs and Challenges 19 Nov 2024 · 0 repositories · arXiv:2412.04477
-
Joint Vision-Language Social Bias Removal for CLIP 19 Nov 2024 · 1 repository · arXiv:2411.12785
-
ProSec: Fortifying Code LLMs with Proactive Security Alignment 19 Nov 2024 · 1 repository · arXiv:2411.12882Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Analyzing and Improving the Skin Tone Consistency and Bias in Implicit 3D Relightable Face Generators 18 Nov 2024 · 0 repositories · arXiv:2411.12002
-
DeSiRe-GS: 4D Street Gaussians for Static-Dynamic Decomposition and Surface Reconstruction for Urban Driving Scenes 18 Nov 2024 · 1 repository · arXiv:2411.11921Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
LaVin-DiT: Large Vision Diffusion Transformer 18 Nov 2024 · 0 repositories · arXiv:2411.11505
-
Moral Persuasion in Large Language Models: Evaluating Susceptibility and Ethical Alignment 18 Nov 2024 · 1 repository · arXiv:2411.11731
-
SADDE: Semi-supervised Anomaly Detection with Dependable Explanations 18 Nov 2024 · 1 repository · arXiv:2411.11293
-
Stacking Brick by Brick: Aligned Feature Isolation for Incremental Face Forgery Detection 18 Nov 2024 · 1 repository · arXiv:2411.11396Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Value Imprint: A Technique for Auditing the Human Values Embedded in RLHF Datasets 18 Nov 2024 · 0 repositories · arXiv:2411.11937
-
Visual-Semantic Graph Matching Net for Zero-Shot Learning 18 Nov 2024 · 1 repository · arXiv:2411.11351
-
GeomCLIP: Contrastive Geometry-Text Pre-training for Molecules 16 Nov 2024 · 1 repository · arXiv:2411.10821
-
In silico discovery of representational relationships across visual cortex 16 Nov 2024 · 0 repositories · arXiv:2411.10872
-
SPICA: Retrieving Scenarios for Pluralistic In-Context Alignment 16 Nov 2024 · 1 repository · arXiv:2411.10912
-
AC-Informed DC Optimal Transmission Switching Problems via Parameter Optimization 15 Nov 2024 · 0 repositories · arXiv:2411.10528
-
Any2Any: Incomplete Multimodal Retrieval with Conformal Prediction 15 Nov 2024 · 0 repositories · arXiv:2411.10513
-
Boundary Attention Constrained Zero-Shot Layout-To-Image Generation 15 Nov 2024 · 0 repositories · arXiv:2411.10495
-
FedAli: Personalized Federated Learning with Aligned Prototypes through Optimal Transport 15 Nov 2024 · 1 repository · arXiv:2411.10595
-
Fill in the blanks: Rethinking Interpretability in vision 15 Nov 2024 · 0 repositories · arXiv:2411.10273
-
InvestESG: A multi-agent reinforcement learning benchmark for studying climate investment as a social dilemma 15 Nov 2024 · 1 repository · arXiv:2411.09856
-
Morpho-Aware Global Attention for Image Matting 15 Nov 2024 · 0 repositories · arXiv:2411.10251
-
Safe Text-to-Image Generation: Simply Sanitize the Prompt Embedding 15 Nov 2024 · 0 repositories · arXiv:2411.10329
-
Step-wise Distribution Alignment Guided Style Prompt Tuning for Source-free Cross-domain Few-shot Learning 15 Nov 2024 · 1 repository · arXiv:2411.10070
-
Towards Automatic Evaluation of Task-Oriented Dialogue Flows 15 Nov 2024 · 0 repositories · arXiv:2411.10416
-
Two-step registration method boosts sensitivity in longitudinal fixel-based analyses 15 Nov 2024 · 0 repositories · arXiv:2411.10116
-
A Self-Supervised Model for Multi-modal Stroke Risk Prediction 14 Nov 2024 · 1 repository · arXiv:2411.09822Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment 14 Nov 2024 · 0 repositories · arXiv:2411.09341
-
DyGASR: Dynamic Generalized Exponential Splatting with Surface Alignment for Accelerated 3D Mesh Reconstruction 14 Nov 2024 · 0 repositories · arXiv:2411.09156
-
Embedding Space Allocation with Angle-Norm Joint Classifiers for Few-Shot Class-Incremental Learning 14 Nov 2024 · 0 repositories · arXiv:2411.09250
-
How do Machine Learning Models Change? 14 Nov 2024 · 0 repositories · arXiv:2411.09645