Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 17
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 17 of 56: papers 1,601 to 1,700 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LatentSync: Audio Conditioned Latent Diffusion Models for Lip Sync 12 Dec 2024 · 1 repository · arXiv:2412.09262Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
LoRACLR: Contrastive Adaptation for Customization of Diffusion Models 12 Dec 2024 · 0 repositories · arXiv:2412.09622
-
Mojito: Motion Trajectory and Intensity Control for Video Generation 12 Dec 2024 · 0 repositories · arXiv:2412.08948
-
On Round-Off Errors and Gaussian Blur in Superresolution and in Image Registration 12 Dec 2024 · 0 repositories · arXiv:2412.09741
-
Radiology Report Generation via Multi-objective Preference Optimization 12 Dec 2024 · 0 repositories · arXiv:2412.08901
-
SPRec: Leveraging Self-Play to Debias Preference Alignment for Large Language Model-based Recommendations 12 Dec 2024 · 1 repository · arXiv:2412.09243Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Temporal Action Localization with Cross Layer Task Decoupling and Refinement 12 Dec 2024 · 1 repository · arXiv:2412.09202
-
Toward Foundation Model for Multivariate Wearable Sensing of Physiological Signals 12 Dec 2024 · 1 repository · arXiv:2412.09758
-
Adversarial Purification by Consistency-aware Latent Space Optimization on Data Manifolds 11 Dec 2024 · 0 repositories · arXiv:2412.08394
-
Bridging Relevance and Reasoning: Rationale Distillation in Retrieval-Augmented Generation 11 Dec 2024 · 0 repositories · arXiv:2412.08519
-
Collaborative Hybrid Propagator for Temporal Misalignment in Audio-Visual Segmentation 11 Dec 2024 · 0 repositories · arXiv:2412.08161
-
Coverage-based Fairness in Multi-document Summarization 11 Dec 2024 · 1 repository · arXiv:2412.08795
-
DocSum: Domain-Adaptive Pre-training for Document Abstractive Summarization 11 Dec 2024 · 0 repositories · arXiv:2412.08196
-
Generate Any Scene: Evaluating and Improving Text-to-Vision Generation with Scene Graph Programming 11 Dec 2024 · 1 repository · arXiv:2412.08221Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Learning to Reason via Self-Iterative Process Feedback for Small Language Models 11 Dec 2024 · 0 repositories · arXiv:2412.08393
-
Leveraging Graph-RAG and Prompt Engineering to Enhance LLM-Based Automated Requirement Traceability and Compliance Checks 11 Dec 2024 · 0 repositories · arXiv:2412.08593
-
TECO: Improving Multimodal Intent Recognition with Text Enhancement through Commonsense Knowledge Extraction 11 Dec 2024 · 0 repositories · arXiv:2412.08529
-
CapGen:An Environment-Adaptive Generator of Adversarial Patches 10 Dec 2024 · 0 repositories · arXiv:2412.07253
-
DiffSensei: Bridging Multi-Modal LLMs and Diffusion Models for Customized Manga Generation 10 Dec 2024 · 0 repositories · arXiv:2412.07589
-
Efficient Diversity-Preserving Diffusion Alignment via Gradient-Informed GFlowNets 10 Dec 2024 · 0 repositories · arXiv:2412.07775
-
Exploring What Why and How: A Multifaceted Benchmark for Causation Understanding of Video Anomaly 10 Dec 2024 · 1 repository · arXiv:2412.07183
-
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model 10 Dec 2024 · 0 repositories · arXiv:2412.07333
-
IntellectSeeker: A Personalized Literature Management System with the Probabilistic Model and Large Language Model 10 Dec 2024 · 1 repository · arXiv:2412.07213
-
LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models 10 Dec 2024 · 0 repositories · arXiv:2412.07746
-
Moving to the suburbs? Exploring the potential impact of work-from-home on suburbanization in Poland 10 Dec 2024 · 0 repositories · arXiv:2412.07459
-
Optimizing Alignment with Less: Leveraging Data Augmentation for Personalized Evaluation 10 Dec 2024 · 0 repositories · arXiv:2412.07429
-
RAZOR: Sharpening Knowledge by Cutting Bias with Unsupervised Text Rewriting 10 Dec 2024 · 1 repository · arXiv:2412.07675
-
StyleMaster: Stylize Your Video with Artistic Generation and Translation 10 Dec 2024 · 0 repositories · arXiv:2412.07744
-
AnyBimanual: Transferring Unimanual Policy for General Bimanual Manipulation 9 Dec 2024 · 1 repository · arXiv:2412.06779
-
Bridging Conversational and Collaborative Signals for Conversational Recommendation 9 Dec 2024 · 0 repositories · arXiv:2412.06949
-
Towards Brain Passage Retrieval -- An Investigation of EEG Query Representations 9 Dec 2024 · 0 repositories · arXiv:2412.06695
-
Driv3R: Learning Dense 4D Reconstruction for Autonomous Driving 9 Dec 2024 · 1 repository · arXiv:2412.06777
-
MSCrackMamba: Leveraging Vision Mamba for Crack Detection in Fused Multispectral Imagery 9 Dec 2024 · 0 repositories · arXiv:2412.06211
-
MVReward: Better Aligning and Evaluating Multi-View Diffusion Models with Human Preferences 9 Dec 2024 · 0 repositories · arXiv:2412.06614
-
Political-LLM: Large Language Models in Political Science 9 Dec 2024 · 0 repositories · arXiv:2412.06864
-
Proactive Agents for Multi-Turn Text-to-Image Generation Under Uncertainty 9 Dec 2024 · 1 repository · arXiv:2412.06771
-
VP-MEL: Visual Prompts Guided Multimodal Entity Linking 9 Dec 2024 · 0 repositories · arXiv:2412.06720
-
Latent-Reframe: Enabling Camera Control for Video Diffusion Model without Training 8 Dec 2024 · 0 repositories · arXiv:2412.06029
-
Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent 7 Dec 2024 · 0 repositories · arXiv:2412.05722
-
LABIIUM: AI-Enhanced Zero-configuration Measurement Automation System 7 Dec 2024 · 0 repositories · arXiv:2412.16172
-
On the effective transfer of knowledge from English to Hindi Wikipedia 7 Dec 2024 · 1 repository · arXiv:2412.05708
-
PromptRefine: Enhancing Few-Shot Performance on Low-Resource Indic Languages with Example Selection from Related Example Banks 7 Dec 2024 · 0 repositories · arXiv:2412.05710
-
Addressing Attribute Leakages in Diffusion-based Image Editing without Training 6 Dec 2024 · 0 repositories · arXiv:2412.04715
-
DreamColour: Controllable Video Colour Editing without Training 6 Dec 2024 · 1 repository · arXiv:2412.05180
-
Explingo: Explaining AI Predictions using Large Language Models 6 Dec 2024 · 1 repository · arXiv:2412.05145
-
KaLM: Knowledge-aligned Autoregressive Language Modeling via Dual-view Knowledge Graph Contrastive Learning 6 Dec 2024 · 0 repositories · arXiv:2412.04948
-
LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment 6 Dec 2024 · 0 repositories · arXiv:2412.04814
-
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment 6 Dec 2024 · 0 repositories · arXiv:2412.04835
-
Multi-Party Supervised Fine-tuning of Language Models for Multi-Party Dialogue Generation 6 Dec 2024 · 0 repositories · arXiv:2412.05342
-
Reconstruction of 3D lumbar spine models from incomplete segmentations using landmark detection 6 Dec 2024 · 0 repositories · arXiv:2412.05065
-
Beyond the Binary: Capturing Diverse Preferences With Reward Regularization 5 Dec 2024 · 0 repositories · arXiv:2412.03822
-
Frequency-Adaptive Low-Latency Object Detection Using Events and Frames 5 Dec 2024 · 0 repositories · arXiv:2412.04149
-
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning 5 Dec 2024 · 1 repository · arXiv:2412.04661
-
Inferring Leader-Follower Behavior from Presence Data in the Marine Environment: A Case Study on Reef Manta Rays 5 Dec 2024 · 1 repository · arXiv:2412.03990
-
Marvel: Accelerating Safe Online Reinforcement Learning with Finetuned Offline Policy 5 Dec 2024 · 1 repository · arXiv:2412.04426
-
Pinco: Position-induced Consistent Adapter for Diffusion Transformer in Foreground-conditioned Inpainting 5 Dec 2024 · 0 repositories · arXiv:2412.03812
-
Relationships between Keywords and Strong Beats in Lyrical Music 5 Dec 2024 · 0 repositories · arXiv:2412.04202
-
Safeguarding Text-to-Image Generation via Inference-Time Prompt-Noise Optimization 5 Dec 2024 · 1 repository · arXiv:2412.03876
-
Fully Distributed, Flexible Compositional Visual Representations via Soft Tensor Products 5 Dec 2024 · 1 repository · arXiv:2412.04671Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Training MLPs on Graphs without Supervision 5 Dec 2024 · 1 repository · arXiv:2412.03864Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
TransAdapter: Vision Transformer for Feature-Centric Unsupervised Domain Adaptation 5 Dec 2024 · 1 repository · arXiv:2412.04073
-
Advancing Conversational Psychotherapy: Integrating Privacy, Dual-Memory, and Domain Expertise with Large Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.02987
-
AI-Driven Day-to-Day Route Choice 4 Dec 2024 · 1 repository · arXiv:2412.03338
-
Align3R: Aligned Monocular Depth Estimation for Dynamic Videos 4 Dec 2024 · 0 repositories · arXiv:2412.03079
-
BIMCaP: BIM-based AI-supported LiDAR-Camera Pose Refinement 4 Dec 2024 · 1 repository · arXiv:2412.03434
-
ChatTS: Aligning Time Series with LLMs via Synthetic Data for Enhanced Understanding and Reasoning 4 Dec 2024 · 1 repository · arXiv:2412.03104Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
DIVE: Taming DINO for Subject-Driven Video Editing 4 Dec 2024 · 0 repositories · arXiv:2412.03347
-
Enhancing Recommendation Systems with GNNs and Addressing Over-Smoothing 4 Dec 2024 · 0 repositories · arXiv:2412.03097
-
Expanding Event Modality Applications through a Robust CLIP-Based Encoder 4 Dec 2024 · 2 repositories · arXiv:2412.03093
-
Multi-view Image Diffusion via Coordinate Noise and Fourier Attention 4 Dec 2024 · 0 repositories · arXiv:2412.03756
-
BANER: Boundary-Aware LLMs for Few-Shot Named Entity Recognition 3 Dec 2024 · 1 repository · arXiv:2412.02228
-
Crash Severity Risk Modeling Strategies under Data Imbalance 3 Dec 2024 · 0 repositories · arXiv:2412.02094
-
Cross-Attention Head Position Patterns Can Align with Human Visual Concepts in Text-to-Image Generative Models 3 Dec 2024 · 1 repository · arXiv:2412.02237Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Deep Matrix Factorization with Adaptive Weights for Multi-View Clustering 3 Dec 2024 · 0 repositories · arXiv:2412.02292
-
Diffusion-based Visual Anagram as Multi-task Learning 3 Dec 2024 · 1 repository · arXiv:2412.02693Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Flattering to Deceive: The Impact of Sycophantic Behavior on User Trust in Large Language Model 3 Dec 2024 · 0 repositories · arXiv:2412.02802
-
Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback 3 Dec 2024 · 0 repositories · arXiv:2412.02617
-
Optimizing Latent Goal by Learning from Trajectory Preference 3 Dec 2024 · 0 repositories · arXiv:2412.02125
-
Single-Shot Metric Depth from Focused Plenoptic Cameras 3 Dec 2024 · 0 repositories · arXiv:2412.02386
-
3DSceneEditor: Controllable 3D Scene Editing with Gaussian Splatting 2 Dec 2024 · 0 repositories · arXiv:2412.01583
-
An Efficient Unsupervised Framework for Convex Quadratic Programs via Deep Unrolling 2 Dec 2024 · 0 repositories · arXiv:2412.01051
-
Cerberus: Attribute-based person re-identification using semantic IDs 2 Dec 2024 · 0 repositories · arXiv:2412.01048
-
Enhancing Perception Capabilities of Multimodal LLMs with Training-Free Fusion 2 Dec 2024 · 0 repositories · arXiv:2412.01289
-
HDGS: Textured 2D Gaussian Splatting for Enhanced Scene Rendering 2 Dec 2024 · 0 repositories · arXiv:2412.01823
-
Misalignments in AI Perception: Quantitative Findings and Visual Mapping of How Experts and the Public Differ in Expectations and Risks, Benefits, and Value Judgments 2 Dec 2024 · 0 repositories · arXiv:2412.01459
-
VLsI: Verbalized Layers-to-Interactions from Large to Small Vision Language Models 2 Dec 2024 · 0 repositories · arXiv:2412.01822
-
Large Language Models as Mirrors of Societal Moral Standards 1 Dec 2024 · 0 repositories · arXiv:2412.00956
-
Linear Probe Penalties Reduce LLM Sycophancy 1 Dec 2024 · 0 repositories · arXiv:2412.00967
-
Techno-Economic Assessment of Net-Zero Energy Buildings: Financial Projections and Incentives for Achieving Energy Decarbonization Goals 1 Dec 2024 · 0 repositories · arXiv:2412.00874
-
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters 30 Nov 2024 · 0 repositories · arXiv:2412.00530
-
Jailbreak Large Vision-Language Models Through Multi-Modal Linkage 30 Nov 2024 · 1 repository · arXiv:2412.00473
-
Leveraging LLM for Automated Ontology Extraction and Knowledge Graph Generation 30 Nov 2024 · 0 repositories · arXiv:2412.00608
-
Linear Simple Cycle Reservoirs at the edge of stability perform Fourier decomposition of the input driving signals 30 Nov 2024 · 1 repository · arXiv:2412.00295
-
LMSeg: Unleashing the Power of Large-Scale Models for Open-Vocabulary Semantic Segmentation 30 Nov 2024 · 0 repositories · arXiv:2412.00364
-
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability 29 Nov 2024 · 1 repository · arXiv:2411.19943Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Dynamic EEG-fMRI mapping: Revealing the relationship between brain connectivity and cognitive state 29 Nov 2024 · 0 repositories · arXiv:2411.19922
-
Effective Fine-Tuning of Vision-Language Models for Accurate Galaxy Morphology Analysis 29 Nov 2024 · 0 repositories · arXiv:2411.19475
-
Knowledge-Data Fusion Based Source-Free Semi-Supervised Domain Adaptation for Seizure Subtype Classification 29 Nov 2024 · 0 repositories · arXiv:2411.19502
-
The AI Interface: Designing for the Ideal Machine-Human Experience (Editorial) 29 Nov 2024 · 0 repositories · arXiv:2412.09000
-
The Syncytial Mesh Model: A Biophysical Framework for Scale-Dependent Coherence in the Brain 29 Nov 2024 · 0 repositories · arXiv:2412.12106