Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 3
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 3 of 56: papers 201 to 300 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Harnessing Large Language Models for Scientific Novelty Detection 30 May 2025 · 0 repositories · arXiv:2505.24615
-
HELM: Hyperbolic Large Language Models via Mixture-of-Curvature Experts 30 May 2025 · 1 repository · arXiv:2505.24722Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Interactive Video Generation via Domain Adaptation 30 May 2025 · 0 repositories · arXiv:2505.24253
-
KEVER^2: Knowledge-Enhanced Visual Emotion Reasoning and Retrieval 30 May 2025 · 0 repositories · arXiv:2505.24342
-
Light as Deception: GPT-driven Natural Relighting Against Vision-Language Pre-training Models 30 May 2025 · 0 repositories · arXiv:2505.24227
-
MGS3: A Multi-Granularity Self-Supervised Code Search Framework 30 May 2025 · 0 repositories · arXiv:2505.24274
-
MiCRo: Mixture Modeling and Context-aware Routing for Personalized Preference Learning 30 May 2025 · 0 repositories · arXiv:2505.24846
-
PhySense: Principle-Based Physics Reasoning Benchmarking for Large Language Models 30 May 2025 · 0 repositories · arXiv:2505.24823
-
SCOUT: Teaching Pre-trained Language Models to Enhance Reasoning via Flow Chain-of-Thought 30 May 2025 · 0 repositories · arXiv:2505.24181
-
When Harry Meets Superman: The Role of The Interlocutor in Persona-Based Dialogue Generation 30 May 2025 · 0 repositories · arXiv:2505.24613
-
A Hetero-functional Graph Theory Perspective of Engineering Management of Mega-Projects 29 May 2025 · 0 repositories · arXiv:2505.24045
-
Beyond Optimal Transport: Model-Aligned Coupling for Flow Matching 29 May 2025 · 0 repositories · arXiv:2505.23346
-
Characterizing the Expressivity of Transformer Language Models 29 May 2025 · 0 repositories · arXiv:2505.23623
-
DeepTheorem: Advancing LLM Reasoning for Theorem Proving Through Natural Language and Reinforcement Learning 29 May 2025 · 2 repositories · arXiv:2505.23754Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis 29 May 2025 · 0 repositories · arXiv:2505.23325
-
FOLIAGE: Towards Physical Intelligence World Models Via Unbounded Surface Evolution 29 May 2025 · 0 repositories · arXiv:2506.03173
-
From Connectivity to Autonomy: The Dawn of Self-Evolving Communication Systems 29 May 2025 · 0 repositories · arXiv:2505.23710
-
Graph Random Walk with Feature-Label Space Alignment: A Multi-Label Feature Selection Method 29 May 2025 · 0 repositories · arXiv:2505.23228
-
Hallo4: High-Fidelity Dynamic Portrait Animation via Direct Preference Optimization and Temporal Motion Modulation 29 May 2025 · 1 repository · arXiv:2505.23525
-
Implicit Inversion turns CLIP into a Decoder 29 May 2025 · 1 repository · arXiv:2505.23161
-
Mobi-π: Mobilizing Your Robot Learning Policy 29 May 2025 · 0 repositories · arXiv:2505.23692
-
On-Policy RL with Optimal Reward Baseline 29 May 2025 · 1 repository · arXiv:2505.23585
-
OTPTO: Joint Product Selection and Inventory Optimization in Fresh E-commerce Front-End Warehouses 29 May 2025 · 0 repositories · arXiv:2505.23421
-
Proximalized Preference Optimization for Diverse Feedback Types: A Decomposed Perspective on DPO 29 May 2025 · 0 repositories · arXiv:2505.23316
-
Quantum computing and artificial intelligence: status and perspectives 29 May 2025 · 0 repositories · arXiv:2505.23860
-
Semantics-Aware Human Motion Generation from Audio Instructions 29 May 2025 · 0 repositories · arXiv:2505.23465
-
Spatio-Temporal Joint Density Driven Learning for Skeleton-Based Action Recognition 29 May 2025 · 1 repository · arXiv:2505.23012
-
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation 29 May 2025 · 1 repository · arXiv:2505.23368Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Zero-P-to-3: Zero-Shot Partial-View Images to 3D Object 29 May 2025 · 0 repositories · arXiv:2505.23054
-
Align-DA: Align Score-based Atmospheric Data Assimilation with Multiple Preferences 28 May 2025 · 0 repositories · arXiv:2505.22008
-
CAST: Contrastive Adaptation and Distillation for Semi-Supervised Instance Segmentation 28 May 2025 · 0 repositories · arXiv:2505.21904
-
Do Large Language Models Think Like the Brain? Sentence-Level Evidence from fMRI and Hierarchical Embeddings 28 May 2025 · 0 repositories · arXiv:2505.22563
-
Enhancing Paraphrase Type Generation: The Impact of DPO and RLHF Evaluated with Human-Ranked Data 28 May 2025 · 1 repository · arXiv:2506.02018
-
HiLDe: Intentional Code Generation via Human-in-the-Loop Decoding 28 May 2025 · 0 repositories · arXiv:2505.22906
-
Modeling and Optimizing User Preferences in AI Copilots: A Comprehensive Survey and Taxonomy 28 May 2025 · 0 repositories · arXiv:2505.21907
-
MoRE: A Mixture of Low-Rank Experts for Adaptive Multi-Task Learning 28 May 2025 · 1 repository · arXiv:2505.22694
-
Operationalizing CaMeL: Strengthening LLM Defenses for Enterprise Deployment 28 May 2025 · 0 repositories · arXiv:2505.22852
-
ValueSim: Generating Backstories to Model Individual Value Systems 28 May 2025 · 0 repositories · arXiv:2505.23827
-
Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment 27 May 2025 · 1 repository · arXiv:2505.21494
-
Constructing a bridge between functioning of oscillatory neuronal networks and quantum-like cognition along with quantum-inspired computation and AI 27 May 2025 · 0 repositories · arXiv:2506.00040
-
Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation 27 May 2025 · 1 repository · arXiv:2505.20897
-
Dub-S2ST: Textless Speech-to-Speech Translation for Seamless Dubbing 27 May 2025 · 0 repositories · arXiv:2505.20899
-
Emotion-aware Dual Cross-Attentive Neural Network with Label Fusion for Stance Detection in Misinformative Social Media Content 27 May 2025 · 1 repository · arXiv:2505.23812
-
FinTagging: An LLM-ready Benchmark for Extracting and Structuring Financial Information 27 May 2025 · 1 repository · arXiv:2505.20650
-
From prosthetic memory to prosthetic denial: Auditing whether large language models are prone to mass atrocity denialism 27 May 2025 · 0 repositories · arXiv:2505.21753
-
HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling 27 May 2025 · 0 repositories · arXiv:2505.20836
-
Herd Behavior: Investigating Peer Influence in LLM-based Multi-Agent Systems 27 May 2025 · 0 repositories · arXiv:2505.21588
-
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding 27 May 2025 · 1 repository · arXiv:2505.20715
-
Roboflow100-VL: A Multi-Domain Object Detection Benchmark for Vision-Language Models 27 May 2025 · 1 repository · arXiv:2505.20612
-
Simple yet Effective Graph Distillation via Clustering 27 May 2025 · 0 repositories · arXiv:2505.20807
-
Text-Queried Audio Source Separation via Hierarchical Modeling 27 May 2025 · 0 repositories · arXiv:2505.21025
-
Unified Alignment Protocol: Making Sense of the Unlabeled Data in New Domains 27 May 2025 · 0 repositories · arXiv:2505.21010
-
Align and Surpass Human Camouflaged Perception: Visual Refocus Reinforcement Fine-Tuning 26 May 2025 · 1 repository · arXiv:2505.19611
-
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection 26 May 2025 · 1 repository · arXiv:2505.19528
-
CCL-LGS: Contrastive Codebook Learning for 3D Language Gaussian Splatting 26 May 2025 · 0 repositories · arXiv:2505.20469
-
Customising Electricity Contracts at Scale with Large Language Models 26 May 2025 · 0 repositories · arXiv:2505.19551
-
Diversity-Driven Generative Dataset Distillation Based on Diffusion Model with Self-Adaptive Memory 26 May 2025 · 0 repositories · arXiv:2505.19469
-
From What to How: Attributing CLIP's Latent Components Reveals Unexpected Semantic Reliance 26 May 2025 · 1 repository · arXiv:2505.20229
-
Gradient Inversion Transcript: Leveraging Robust Generative Priors to Reconstruct Training Data from Gradient Leakage 26 May 2025 · 0 repositories · arXiv:2505.20026
-
Holes in Latent Space: Topological Signatures Under Adversarial Influence 26 May 2025 · 0 repositories · arXiv:2505.20435
-
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation 26 May 2025 · 0 repositories · arXiv:2505.20271
-
InfoCons: Identifying Interpretable Critical Concepts in Point Clouds via Information Theory 26 May 2025 · 1 repository · arXiv:2505.19820Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Knowledge-Aligned Counterfactual-Enhancement Diffusion Perception for Unsupervised Cross-Domain Visual Emotion Recognition 26 May 2025 · 0 repositories · arXiv:2505.19694
-
Kuramoto-FedAvg: Using Synchronization Dynamics to Improve Federated Learning Optimization under Statistical Heterogeneity 26 May 2025 · 0 repositories · arXiv:2505.19605
-
Large Language Models Meet Knowledge Graphs for Question Answering: Synthesis and Opportunities 26 May 2025 · 1 repository · arXiv:2505.20099
-
LLMs as Better Recommenders with Natural Language Collaborative Signals: A Self-Assessing Retrieval Approach 26 May 2025 · 0 repositories · arXiv:2505.19464
-
MAS-Zero: Designing Multi-Agent Systems with Zero Supervision 26 May 2025 · 1 repository
-
MLLM-Guided VLM Fine-Tuning with Joint Inference for Zero-Shot Composed Image Retrieval 26 May 2025 · 0 repositories · arXiv:2505.19707
-
Modeling Beyond MOS: Quality Assessment Models Must Integrate Context, Reasoning, and Multimodality 26 May 2025 · 0 repositories · arXiv:2505.19696
-
Monocle: Hybrid Local-Global In-Context Evaluation for Long-Text Generation with Uncertainty-Based Active Learning 26 May 2025 · 0 repositories · arXiv:2505.20195
-
Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval 26 May 2025 · 0 repositories · arXiv:2505.19952
-
OmniCharacter: Towards Immersive Role-Playing Agents with Seamless Speech-Language Personality Interaction 26 May 2025 · 1 repository · arXiv:2505.20277
-
Poison in the Well: Feature Embedding Disruption in Backdoor Attacks 26 May 2025 · 0 repositories · arXiv:2505.19821
-
ReasonPlan: Unified Scene Prediction and Decision Reasoning for Closed-loop Autonomous Driving 26 May 2025 · 1 repository · arXiv:2505.20024
-
Reconceptualizing Smart Microscopy: From Data Collection to Knowledge Creation by Multi-Agent Integration 26 May 2025 · 0 repositories · arXiv:2505.20466
-
Risk-aware Direct Preference Optimization under Nested Risk Measure 26 May 2025 · 1 repository · arXiv:2505.20359
-
SCIRGC: Multi-Granularity Citation Recommendation and Citation Sentence Preference Alignment 26 May 2025 · 0 repositories · arXiv:2505.20103
-
T^2Agent A Tool-augmented Multimodal Misinformation Detection Agent with Monte Carlo Tree Search 26 May 2025 · 0 repositories · arXiv:2505.19768
-
Task Memory Engine: Spatial Memory for Robust Multi-Step LLM Agents 26 May 2025 · 1 repository · arXiv:2505.19436
-
Ten Principles of AI Agent Economics 26 May 2025 · 0 repositories · arXiv:2505.20273
-
THiNK: Can Large Language Models Think-aloud? 26 May 2025 · 1 repository · arXiv:2505.20184
-
TTPA: Token-level Tool-use Preference Alignment Training Framework with Fine-grained Evaluation 26 May 2025 · 0 repositories · arXiv:2505.20016
-
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation 26 May 2025 · 1 repository · arXiv:2505.19462
-
A Joint Learning Framework with Feature Reconstruction and Prediction for Incomplete Satellite Image Time Series in Agricultural Semantic Segmentation 25 May 2025 · 1 repository · arXiv:2505.19159
-
ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment 25 May 2025 · 0 repositories · arXiv:2505.19241
-
Evaluating AI for Finance: Is AI Credible at Assessing Investment Risk? 25 May 2025 · 0 repositories · arXiv:2505.18953
-
Evaluating Steering Techniques using Human Similarity Judgments 25 May 2025 · 0 repositories · arXiv:2505.19333
-
Fluent but Culturally Distant: Can Regional Training Teach Cultural Understanding? 25 May 2025 · 0 repositories · arXiv:2505.21548
-
Language Models Surface the Unwritten Code of Science and Society 25 May 2025 · 0 repositories · arXiv:2505.18942
-
Nine Ways to Break Copyright Law and Why Our LLM Won't: A Fair Use Aligned Generation Framework 25 May 2025 · 0 repositories · arXiv:2505.23788
-
PolyPose: Localizing Deformable Anatomy in 3D from Sparse 2D X-ray Images using Polyrigid Transforms 25 May 2025 · 1 repository · arXiv:2505.19256
-
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning 25 May 2025 · 1 repository · arXiv:2505.19196Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
Training-free Stylized Text-to-Image Generation with Fast Inference 25 May 2025 · 0 repositories · arXiv:2505.19063
-
AI-Driven Climate Policy Scenario Generation for Sub-Saharan Africa 24 May 2025 · 1 repository · arXiv:2505.18694
-
Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation 24 May 2025 · 1 repository · arXiv:2505.18730
-
Building a Functional Machine Translation Corpus for Kpelle 24 May 2025 · 0 repositories · arXiv:2505.18905
-
Diffusion Blend: Inference-Time Multi-Preference Alignment for Diffusion Models 24 May 2025 · 1 repository · arXiv:2505.18547Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving 24 May 2025 · 0 repositories · arXiv:2505.18644
-
Generative RLHF-V: Learning Principles from Multi-modal Human Preference 24 May 2025 · 0 repositories · arXiv:2505.18531
-
Guiding the Experts: Semantic Priors for Efficient and Focused MoE Routing 24 May 2025 · 1 repository · arXiv:2505.18586