Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 6
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 6 of 56: papers 501 to 600 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
HiLLIE: Human-in-the-Loop Training for Low-Light Image Enhancement 4 May 2025 · 0 repositories · arXiv:2505.02134
-
Student Perspectives on the Benefits and Risks of AI in Education 4 May 2025 · 0 repositories · arXiv:2505.02198
-
Unaligned RGB Guided Hyperspectral Image Super-Resolution with Spatial-Spectral Concordance 4 May 2025 · 0 repositories · arXiv:2505.02109
-
3DWG: 3D Weakly Supervised Visual Grounding via Category and Instance-Level Alignment 3 May 2025 · 0 repositories · arXiv:2505.01809
-
A Dual-Task Synergy-Driven Generalization Framework for Pancreatic Cancer Segmentation in CT Scans 3 May 2025 · 1 repository · arXiv:2505.01644
-
Multimodal Graph Representation Learning for Robust Surgical Workflow Recognition with Adversarial Feature Disentanglement 3 May 2025 · 0 repositories · arXiv:2505.01766
-
Same evaluation, more tokens: On the effect of input length for machine translation evaluation using Large Language Models 3 May 2025 · 0 repositories · arXiv:2505.01761Syntology 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples)
-
Federated Adapter on Foundation Models: An Out-Of-Distribution Approach 2 May 2025 · 0 repositories · arXiv:2505.01075
-
FlowDubber: Movie Dubbing with LLM-based Semantic-aware Learning and Flow Matching based Voice Enhancing 2 May 2025 · 0 repositories · arXiv:2505.01263
-
Model See Model Do: Speech-Driven Facial Animation with Style Control 2 May 2025 · 0 repositories · arXiv:2505.01319
-
Multi-agents based User Values Mining for Recommendation 2 May 2025 · 0 repositories · arXiv:2505.00981
-
Diverse Semantics-Guided Feature Alignment and Decoupling for Visible-Infrared Person Re-Identification 1 May 2025 · 0 repositories · arXiv:2505.00619
-
Towards Scalable Human-aligned Benchmark for Text-guided Image Editing 1 May 2025 · 1 repository · arXiv:2505.00502
-
Characterizing AI Agents for Alignment and Governance 30 Apr 2025 · 0 repositories · arXiv:2504.21848
-
Clustering Internet Memes Through Template Matching and Multi-Dimensional Similarity 30 Apr 2025 · 1 repository · arXiv:2505.00056
-
Nexus-Gen: A Unified Model for Image Understanding, Generation, and Editing 30 Apr 2025 · 1 repository · arXiv:2504.21356Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Democratizing Differential Privacy: A Participatory AI Framework for Public Decision-Making 30 Apr 2025 · 0 repositories · arXiv:2504.21297
-
Vision-Language Model-Based Semantic-Guided Imaging Biomarker for Early Lung Cancer Detection 30 Apr 2025 · 0 repositories · arXiv:2504.21344
-
A Picture is Worth a Thousand Prompts? Efficacy of Iterative Human-Driven Prompt Refinement in Image Regeneration Tasks 29 Apr 2025 · 0 repositories · arXiv:2504.20340
-
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation 29 Apr 2025 · 0 repositories · arXiv:2504.20629
-
ChestX-Reasoner: Advancing Radiology Foundation Models with Reasoning through Step-by-Step Verification 29 Apr 2025 · 1 repository · arXiv:2504.20930
-
Cognitive maps are generative programs 29 Apr 2025 · 0 repositories · arXiv:2504.20628
-
FFCBA: Feature-based Full-target Clean-label Backdoor Attacks 29 Apr 2025 · 0 repositories · arXiv:2504.21054
-
Frequency Feature Fusion Graph Network For Depression Diagnosis Via fNIRS 29 Apr 2025 · 0 repositories · arXiv:2504.21064
-
Head-Tail-Aware KL Divergence in Knowledge Distillation for Spiking Neural Networks 29 Apr 2025 · 0 repositories · arXiv:2504.20445
-
Hierarchical Context Learning of object components for unsupervised semantic segmentation 29 Apr 2025 · 1 repository
-
HyPerAlign: Interpretable Personalized LLM Alignment via Hypothesis Generation 29 Apr 2025 · 0 repositories · arXiv:2505.00038
-
MemeBLIP2: A novel lightweight multimodal system to detect harmful memes 29 Apr 2025 · 0 repositories · arXiv:2504.21226
-
The Hidden Risks of LLM-Generated Web Application Code: A Security-Centric Evaluation of Code Generation Capabilities in Large Language Models 29 Apr 2025 · 0 repositories · arXiv:2504.20612
-
Accelerated 3D-3D rigid registration of echocardiographic images obtained from apical window using particle filter 28 Apr 2025 · 0 repositories · arXiv:2504.19930
-
Mapping the Italian Telegram Ecosystem: Communities, Toxicity, and Hate Speech 28 Apr 2025 · 0 repositories · arXiv:2504.19594
-
Contextual Online Uncertainty-Aware Preference Learning for Human Feedback 27 Apr 2025 · 0 repositories · arXiv:2504.19342
-
Explanatory Summarization with Discourse-Driven Planning 27 Apr 2025 · 0 repositories · arXiv:2504.19339
-
Harmonizing Generalization and Personalization in Ring-topology Decentralized Federated Learning 27 Apr 2025 · 0 repositories · arXiv:2504.19103
-
LLM-Evaluation Tropes: Perspectives on the Validity of LLM-Evaluations 27 Apr 2025 · 0 repositories · arXiv:2504.19076
-
Low-Rank Adaptive Structural Priors for Generalizable Diabetic Retinopathy Grading 27 Apr 2025 · 0 repositories · arXiv:2504.19362
-
Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs 26 Apr 2025 · 1 repository · arXiv:2504.19019
-
A Review of 3D Object Detection with Vision-Language Models 25 Apr 2025 · 0 repositories · arXiv:2504.18738
-
Exploring Personality-Aware Interactions in Salesperson Dialogue Agents 25 Apr 2025 · 0 repositories · arXiv:2504.18058
-
MAGI: Multi-Agent Guided Interview for Psychiatric Assessment 25 Apr 2025 · 0 repositories · arXiv:2504.18260
-
TRACE Back from the Future: A Probabilistic Reasoning Approach to Controllable Language Generation 25 Apr 2025 · 0 repositories · arXiv:2504.18535
-
BIM-Constrained Optimization for Accurate Localization and Deviation Correction in Construction Monitoring 24 Apr 2025 · 0 repositories · arXiv:2504.17693
-
Crisp: Cognitive Restructuring of Negative Thoughts through Multi-turn Supportive Dialogues 24 Apr 2025 · 0 repositories · arXiv:2504.17238
-
RefVNLI: Towards Scalable Evaluation of Subject-driven Text-to-image Generation 24 Apr 2025 · 0 repositories · arXiv:2504.17502
-
Tailored minimal reservoir computing: on the bidirectional connection between nonlinearities in the reservoir and in data 24 Apr 2025 · 0 repositories · arXiv:2504.17503
-
The Role of Open-Source LLMs in Shaping the Future of GeoAI 24 Apr 2025 · 0 repositories · arXiv:2504.17833
-
Approaches to Responsible Governance of GenAI in Organizations 23 Apr 2025 · 0 repositories · arXiv:2504.17044
-
Steering the CensorShip: Uncovering Representation Vectors for LLM "Thought" Control 23 Apr 2025 · 1 repository · arXiv:2504.17130Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A closer look at how large language models trust humans: patterns and biases 22 Apr 2025 · 0 repositories · arXiv:2504.15801
-
Achieving Distributive Justice in Federated Learning via Uncertainty Quantification 22 Apr 2025 · 1 repository · arXiv:2504.15924
-
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions 22 Apr 2025 · 1 repository · arXiv:2504.15918
-
Comparative Analysis of Evolutionary Algorithms for Energy-Aware Production Scheduling 22 Apr 2025 · 0 repositories · arXiv:2504.15672
-
Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback 22 Apr 2025 · 0 repositories · arXiv:2504.15804
-
Intent-aware Diffusion with Contrastive Learning for Sequential Recommendation 22 Apr 2025 · 1 repository · arXiv:2504.16077
-
SocialMOIF: Multi-Order Intention Fusion for Pedestrian Trajectory Prediction 22 Apr 2025 · 0 repositories · arXiv:2504.15616
-
Vision-Language Models Are Not Pragmatically Competent in Referring Expression Generation 22 Apr 2025 · 1 repository · arXiv:2504.16060
-
AlignRAG: Leveraging Critique Learning for Evidence-Sensitive Retrieval-Augmented Reasoning 21 Apr 2025 · 1 repository · arXiv:2504.14858Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Distribution-aware Dataset Distillation for Efficient Image Restoration 21 Apr 2025 · 0 repositories · arXiv:2504.14826
-
DSPO: Direct Semantic Preference Optimization for Real-World Image Super-Resolution 21 Apr 2025 · 0 repositories · arXiv:2504.15176
-
Establishing Reliability Metrics for Reward Models in Large Language Models 21 Apr 2025 · 0 repositories · arXiv:2504.14838
-
Improving Sound Source Localization with Joint Slot Attention on Image and Audio 21 Apr 2025 · 0 repositories · arXiv:2504.15118
-
Integrating Response Time and Attention Duration in Bayesian Preference Learning for Multiple Criteria Decision Aiding 21 Apr 2025 · 0 repositories · arXiv:2504.14938
-
Landmark-Free Preoperative-to-Intraoperative Registration in Laparoscopic Liver Resection 21 Apr 2025 · 0 repositories · arXiv:2504.15152
-
Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs 21 Apr 2025 · 1 repository · arXiv:2504.15280
-
SOLIDO: A Robust Watermarking Method for Speech Synthesis via Low-Rank Adaptation 21 Apr 2025 · 0 repositories · arXiv:2504.15035
-
Twin Co-Adaptive Dialogue for Progressive Image Generation 21 Apr 2025 · 0 repositories · arXiv:2504.14868
-
Connecting Parameter Magnitudes and Hessian Eigenspaces at Scale using Sketched Methods 20 Apr 2025 · 0 repositories · arXiv:2504.14701
-
Functional Abstraction of Knowledge Recall in Large Language Models 20 Apr 2025 · 0 repositories · arXiv:2504.14496
-
sEEG-based Encoding for Sentence Retrieval: A Contrastive Learning Approach to Brain-Language Alignment 20 Apr 2025 · 0 repositories · arXiv:2504.14468
-
Balancing Privacy and Action Performance: A Penalty-Driven Approach to Image Anonymization 19 Apr 2025 · 0 repositories · arXiv:2504.14301
-
Dual-channel Heterophilic Message Passing for Graph Fraud Detection 19 Apr 2025 · 0 repositories · arXiv:2504.14205
-
FedCIA: Federated Collaborative Information Aggregation for Privacy-Preserving Recommendation 19 Apr 2025 · 1 repository · arXiv:2504.14208
-
Know Me, Respond to Me: Benchmarking LLMs for Dynamic User Profiling and Personalized Responses at Scale 19 Apr 2025 · 1 repository · arXiv:2504.14225Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Manipulating Multimodal Agents via Cross-Modal Prompt Injection 19 Apr 2025 · 0 repositories · arXiv:2504.14348
-
Chain-of-Thought Textual Reasoning for Few-shot Temporal Action Localization 18 Apr 2025 · 0 repositories · arXiv:2504.13460
-
POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image Generation 18 Apr 2025 · 0 repositories · arXiv:2504.13392
-
Point-Driven Interactive Text and Image Layer Editing Using Diffusion Models 18 Apr 2025 · 0 repositories · arXiv:2504.14108
-
RAG Without the Lag: Interactive Debugging for Retrieval-Augmented Generation Pipelines 18 Apr 2025 · 0 repositories · arXiv:2504.13587
-
A Two-Phase Perspective on Deep Learning Dynamics 17 Apr 2025 · 0 repositories · arXiv:2504.12700
-
Aligning Constraint Generation with Design Intent in Parametric CAD 17 Apr 2025 · 0 repositories · arXiv:2504.13178
-
Benchmarking LLM-based Relevance Judgment Methods 17 Apr 2025 · 1 repository · arXiv:2504.12558
-
CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training Framework 17 Apr 2025 · 1 repository · arXiv:2504.12576Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Contour Field based Elliptical Shape Prior for the Segment Anything Model 17 Apr 2025 · 0 repositories · arXiv:2504.12556
-
FashionDPO:Fine-tune Fashion Outfit Generation Model using Direct Preference Optimization 17 Apr 2025 · 1 repository · arXiv:2504.12900
-
Fast Computation of the Discrete Fourier Transform Rectangular Index Coefficients 17 Apr 2025 · 1 repository · arXiv:2504.12551
-
Generate, but Verify: Reducing Hallucination in Vision-Language Models with Retrospective Resampling 17 Apr 2025 · 1 repository · arXiv:2504.13169Syntology official (archive's flag): 8 ran · 10 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 9 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Pandora: A Code-Driven Large Language Model Agent for Unified Reasoning Across Diverse Structured Knowledge 17 Apr 2025 · 0 repositories · arXiv:2504.12734
-
Post-pre-training for Modality Alignment in Vision-Language Foundation Models 17 Apr 2025 · 1 repository · arXiv:2504.12717
-
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval 17 Apr 2025 · 0 repositories · arXiv:2504.13035
-
SMARTe: Slot-based Method for Accountable Relational Triple extraction 17 Apr 2025 · 1 repository · arXiv:2504.12816
-
TraCeS: Trajectory Based Credit Assignment From Sparse Safety Feedback 17 Apr 2025 · 0 repositories · arXiv:2504.12557
-
Unsupervised Cross-Domain 3D Human Pose Estimation via Pseudo-Label-Guided Global Transforms 17 Apr 2025 · 0 repositories · arXiv:2504.12699
-
Climate-economy projections under shared socioeconomic pathways and net-zero scenarios 16 Apr 2025 · 1 repository · arXiv:2504.11721
-
DART: Disease-aware Image-Text Alignment and Self-correcting Re-alignment for Trustworthy Radiology Report Generation 16 Apr 2025 · 0 repositories · arXiv:2504.11786
-
DC-SAM: In-Context Segment Anything in Images and Videos via Dual Consistency 16 Apr 2025 · 1 repository · arXiv:2504.12080
-
Efficient and Adaptive Simultaneous Speech Translation with Fully Unidirectional Architecture 16 Apr 2025 · 0 repositories · arXiv:2504.11809
-
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models - 16 Apr 2025 · 0 repositories · arXiv:2504.12137
-
FedEPA: Enhancing Personalization and Modality Alignment in Multimodal Federated Learning 16 Apr 2025 · 0 repositories · arXiv:2504.12025
-
Generative Recommendation with Continuous-Token Diffusion 16 Apr 2025 · 0 repositories · arXiv:2504.12007
-
Learning Physics-Informed Color-Aware Transforms for Low-Light Image Enhancement 16 Apr 2025 · 0 repositories · arXiv:2504.11896