Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 27
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 27 of 56: papers 2,601 to 2,700 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Simplifying the Theory on Over-Smoothing 16 Jul 2024 · 1 repository · arXiv:2407.11876Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena 15 Jul 2024 · 0 repositories · arXiv:2407.10627
-
CLAVE: An Adaptive Framework for Evaluating Values of LLM Generated Responses 15 Jul 2024 · 0 repositories · arXiv:2407.10725
-
IDOL: Unified Dual-Modal Latent Diffusion for Human-Centric Joint Video-Depth Generation 15 Jul 2024 · 1 repository · arXiv:2407.10937
-
Learning to Unlearn for Robust Machine Unlearning 15 Jul 2024 · 0 repositories · arXiv:2407.10494
-
NoviCode: Generating Programs from Natural Language Utterances by Novices 15 Jul 2024 · 1 repository · arXiv:2407.10626
-
Bridging Sequence-Structure Alignment in RNA Foundation Models 15 Jul 2024 · 1 repository · arXiv:2407.11242
-
OVLW-DETR: Open-Vocabulary Light-Weighted Detection Transformer 15 Jul 2024 · 1 repository · arXiv:2407.10655
-
Improved Uncertainty Estimation of Graph Neural Network Potentials Using Engineered Latent Space Distances 15 Jul 2024 · 0 repositories · arXiv:2407.10844
-
Multi-Granularity Semantic Revision for Large Language Model Distillation 14 Jul 2024 · 0 repositories · arXiv:2407.10068
-
What Makes and Breaks Safety Fine-tuning? A Mechanistic Study 14 Jul 2024 · 0 repositories · arXiv:2407.10264
-
3D Weakly Supervised Semantic Segmentation with 2D Vision-Language Guidance 13 Jul 2024 · 1 repository · arXiv:2407.09826
-
DiffRect: Latent Diffusion Label Rectification for Semi-supervised Medical Image Segmentation 13 Jul 2024 · 1 repository · arXiv:2407.09918
-
Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control 12 Jul 2024 · 1 repository · arXiv:2407.09024Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Enhancing Emotion Recognition in Incomplete Data: A Novel Cross-Modal Alignment, Reconstruction, and Refinement Framework 12 Jul 2024 · 0 repositories · arXiv:2407.09029
-
New Desiderata for Direct Preference Optimization 12 Jul 2024 · 0 repositories · arXiv:2407.09072
-
Pronunciation Assessment with Multi-modal Large Language Models 12 Jul 2024 · 0 repositories · arXiv:2407.09209
-
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models 12 Jul 2024 · 0 repositories · arXiv:2407.09012
-
15M Multimodal Facial Image-Text Dataset 11 Jul 2024 · 0 repositories · arXiv:2407.08515
-
AddressCLIP: Empowering Vision-Language Models for City-wide Image Address Localization 11 Jul 2024 · 1 repository · arXiv:2407.08156
-
Bootstrapping Vision-language Models for Self-supervised Remote Physiological Measurement 11 Jul 2024 · 0 repositories · arXiv:2407.08507
-
Chromosomal Structural Abnormality Diagnosis by Homologous Similarity 11 Jul 2024 · 0 repositories · arXiv:2407.08204
-
Explainability of Sub-Field Level Crop Yield Prediction using Remote Sensing 11 Jul 2024 · 0 repositories · arXiv:2407.08274
-
Feature Diversification and Adaptation for Federated Domain Generalization 11 Jul 2024 · 0 repositories · arXiv:2407.08245
-
Fine-Tuning Stable Diffusion XL for Stylistic Icon Generation: A Comparison of Caption Size 11 Jul 2024 · 0 repositories · arXiv:2407.08513
-
GTA: A Benchmark for General Tool Agents 11 Jul 2024 · 1 repository · arXiv:2407.08713
-
MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine 11 Jul 2024 · 3 repositories · arXiv:2407.08739
-
The Career Interests of Large Language Models 11 Jul 2024 · 0 repositories · arXiv:2407.08564
-
A Machine Learning and Explainable AI Framework Tailored for Unbalanced Experimental Catalyst Discovery 10 Jul 2024 · 1 repository · arXiv:2407.18935
-
Cross Domain Object Detection via Multi-Granularity Confidence Alignment based Mean Teacher 10 Jul 2024 · 0 repositories · arXiv:2407.07780
-
Deformable Feature Alignment and Refinement for Moving Infrared Dim-small Target Detection 10 Jul 2024 · 0 repositories · arXiv:2407.07289
-
Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture 10 Jul 2024 · 0 repositories · arXiv:2407.07342
-
Raising the Ceiling: Conflict-Free Local Feature Matching with Dynamic View Switching 10 Jul 2024 · 0 repositories · arXiv:2407.07789
-
Video In-context Learning 10 Jul 2024 · 0 repositories · arXiv:2407.07356
-
CEIA: CLIP-Based Event-Image Alignment for Open-World Event-Based Understanding 9 Jul 2024 · 0 repositories · arXiv:2407.06611
-
Historical Review of Variants of Informal Semantics for Logic Programs under Answer Set Semantics: GL'88, GL'91, GK'14, D-V'12 9 Jul 2024 · 0 repositories · arXiv:2407.06814
-
Ada-adapter:Fast Few-shot Style Personlization of Diffusion Model with Pre-trained Image Encoder 8 Jul 2024 · 0 repositories · arXiv:2407.05552
-
ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation 8 Jul 2024 · 1 repository · arXiv:2407.06135
-
Enhancing Vision-Language Models with Scene Graphs for Traffic Accident Understanding 8 Jul 2024 · 0 repositories · arXiv:2407.05910
-
Exposing Privacy Gaps: Membership Inference Attack on Preference Data for LLM Alignment 8 Jul 2024 · 0 repositories · arXiv:2407.06443
-
Generation and De-Identification of Indian Clinical Discharge Summaries using LLMs 8 Jul 2024 · 1 repository · arXiv:2407.05887
-
Leveraging Transformers for Weakly Supervised Object Localization in Unconstrained Videos 8 Jul 2024 · 1 repository · arXiv:2407.06018
-
Link Representation Learning for Probabilistic Travel Time Estimation 8 Jul 2024 · 2 repositories · arXiv:2407.05895
-
TransMA: an explainable multi-modal deep learning model for predicting properties of ionizable lipid nanoparticles in mRNA delivery 8 Jul 2024 · 1 repository · arXiv:2407.05736
-
Beyond Binary Gender Labels: Revealing Gender Biases in LLMs through Gender-Neutral Name Predictions 7 Jul 2024 · 0 repositories · arXiv:2407.05271
-
Edge-guided and Cross-scale Feature Fusion Network for Efficient Multi-contrast MRI Super-Resolution 7 Jul 2024 · 1 repository · arXiv:2407.05307
-
Unlocking Textual and Visual Wisdom: Open-Vocabulary 3D Object Detection Enhanced by Comprehensive Guidance from Text and Image 7 Jul 2024 · 0 repositories · arXiv:2407.05256
-
Test-time Contrastive Concepts for Open-world Semantic Segmentation 6 Jul 2024 · 0 repositories · arXiv:2407.05061
-
Enhance the Robustness of Text-Centric Multimodal Alignments 6 Jul 2024 · 0 repositories · arXiv:2407.05036
-
Helios: An extremely low power event-based gesture recognition for always-on smart eyewear 6 Jul 2024 · 0 repositories · arXiv:2407.05206
-
Incremental Multiview Point Cloud Registration 6 Jul 2024 · 0 repositories · arXiv:2407.05021
-
Rethinking the Effectiveness of Graph Classification Datasets in Benchmarks for Assessing GNNs 6 Jul 2024 · 1 repository · arXiv:2407.04999Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language Models 6 Jul 2024 · 1 repository · arXiv:2407.05131Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments 5 Jul 2024 · 0 repositories · arXiv:2407.12847
-
Dude: Dual Distribution-Aware Context Prompt Learning For Large Vision-Language Model 5 Jul 2024 · 0 repositories · arXiv:2407.04489
-
MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation? 5 Jul 2024 · 1 repository · arXiv:2407.04842Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples)
-
VRSD: Rethinking Similarity and Diversity for Retrieval in Large Language Models 5 Jul 2024 · 0 repositories · arXiv:2407.04573
-
DGR-MIL: Exploring Diverse Global Representation in Multiple Instance Learning for Whole Slide Image Classification 4 Jul 2024 · 1 repository · arXiv:2407.03575Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples)
-
MAMA: Meta-optimized Angular Margin Contrastive Framework for Video-Language Representation Learning 4 Jul 2024 · 1 repository · arXiv:2407.03788
-
The Mysterious Case of Neuron 1512: Injectable Realignment Architectures Reveal Internal Characteristics of Meta's Llama 2 Model 4 Jul 2024 · 1 repository · arXiv:2407.03621
-
A Case Study on Context-Aware Neural Machine Translation with Multi-Task Learning 3 Jul 2024 · 0 repositories · arXiv:2407.03076
-
CogErgLLM: Exploring Large Language Model Systems Design Perspective Using Cognitive Ergonomics 3 Jul 2024 · 0 repositories · arXiv:2407.02885
-
Large language models, physics-based modeling, experimental measurements: the trinity of data-scarce learning of polymer properties 3 Jul 2024 · 0 repositories · arXiv:2407.02770
-
MuDiT & MuSiT: Alignment with Colloquial Expression in Description-to-Song Generation 3 Jul 2024 · 0 repositories · arXiv:2407.03188
-
Towards Federated RLHF with Aggregated Client Preference for LLMs 3 Jul 2024 · 0 repositories · arXiv:2407.03038Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Wild inference for wild SVARs with application to heteroscedasticity-based IV 3 Jul 2024 · 0 repositories · arXiv:2407.03265
-
Visual Grounding with Attention-Driven Constraint Balancing 3 Jul 2024 · 0 repositories · arXiv:2407.03243
-
Aligning Human Motion Generation with Human Perceptions 2 Jul 2024 · 0 repositories · arXiv:2407.02272
-
An End-to-End Speech Summarization Using Large Language Model 2 Jul 2024 · 0 repositories · arXiv:2407.02005
-
AXIAL: Attention-based eXplainability for Interpretable Alzheimer's Localized Diagnosis using 2D CNNs on 3D MRI brain scans 2 Jul 2024 · 1 repository · arXiv:2407.02418
-
Camera-LiDAR Cross-modality Gait Recognition 2 Jul 2024 · 0 repositories · arXiv:2407.02038
-
CFinBench: A Comprehensive Chinese Financial Benchmark for Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02301
-
GlyphDraw2: Automatic Generation of Complex Glyph Posters with Diffusion Models and Large Language Models 2 Jul 2024 · 1 repository · arXiv:2407.02252
-
Investigating Event-Based Cameras for Video Frame Interpolation in Sports 2 Jul 2024 · 0 repositories · arXiv:2407.02370
-
SADL: An Effective In-Context Learning Method for Compositional Visual QA 2 Jul 2024 · 0 repositories · arXiv:2407.01983
-
ScaleDreamer: Scalable Text-to-3D Synthesis with Asynchronous Score Distillation 2 Jul 2024 · 1 repository · arXiv:2407.02040Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 4 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Spectral Graph Reasoning Network for Hyperspectral Image Classification 2 Jul 2024 · 0 repositories · arXiv:2407.02647
-
Understanding Alignment in Multimodal LLMs: A Comprehensive Study 2 Jul 2024 · 0 repositories · arXiv:2407.02477
-
Unleash the Power of Local Representations for Few-Shot Classification 2 Jul 2024 · 0 repositories · arXiv:2407.01967
-
WTU-EVAL: A Whether-or-Not Tool Usage Evaluation Benchmark for Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.12823
-
Aligning Target-Aware Molecule Diffusion Models with Exact Energy Optimization 1 Jul 2024 · 1 repository · arXiv:2407.01648Syntology official (archive's flag): 6 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Cross-Modal Attention Alignment Network with Auxiliary Text Description for zero-shot sketch-based image retrieval 1 Jul 2024 · 0 repositories · arXiv:2407.00979
-
DaBiT: Depth and Blur informed Transformer for Joint Refocusing and Super-Resolution 1 Jul 2024 · 0 repositories · arXiv:2407.01230
-
Entropic Optimal Transport Eigenmaps for Nonlinear Alignment and Joint Embedding of High-Dimensional Datasets 1 Jul 2024 · 1 repository · arXiv:2407.01718
-
View From Above: A Framework for Evaluating Distribution Shifts in Model Behavior 1 Jul 2024 · 1 repository · arXiv:2407.00948Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Unaligning Everything: Or Aligning Any Text to Any Image in Multimodal Models 1 Jul 2024 · 0 repositories · arXiv:2407.01157
-
ZeroDDI: A Zero-Shot Drug-Drug Interaction Event Prediction Method with Semantic Enhanced Learning and Dual-Modal Uniform Alignment 1 Jul 2024 · 1 repository · arXiv:2407.00891Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A Deep Generative Framework for Joint Households and Individuals Population Synthesis 30 Jun 2024 · 0 repositories · arXiv:2407.01643
-
BAPO: Base-Anchored Preference Optimization for Overcoming Forgetting in Large Language Models Personalization 30 Jun 2024 · 0 repositories · arXiv:2407.00693
-
Causality-driven Sequence Segmentation for Enhancing Multiphase Industrial Process Data Analysis and Soft Sensing 30 Jun 2024 · 0 repositories · arXiv:2407.05954
-
Improving Real-Time Music Accompaniment Separation with MMDenseNet 30 Jun 2024 · 0 repositories · arXiv:2407.00657
-
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning 30 Jun 2024 · 1 repository · arXiv:2407.00782Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning 29 Jun 2024 · 1 repository · arXiv:2407.00324
-
Evaluating Human Alignment and Model Faithfulness of LLM Rationale 28 Jun 2024 · 0 repositories · arXiv:2407.00219
-
GM-DF: Generalized Multi-Scenario Deepfake Detection 28 Jun 2024 · 1 repository · arXiv:2406.20078
-
MetaDesigner: Advancing Artistic Typography Through AI-Driven, User-Centric, and Multilingual WordArt Synthesis 28 Jun 2024 · 0 repositories · arXiv:2406.19859
-
Parallax-tolerant Image Stitching via Segmentation-guided Multi-homography Warping 28 Jun 2024 · 1 repository · arXiv:2406.19922
-
Simulating Financial Market via Large Language Model based Agents 28 Jun 2024 · 0 repositories · arXiv:2406.19966
-
STLLaVA-Med: Self-Training Large Language and Vision Assistant for Medical Question-Answering 28 Jun 2024 · 1 repository · arXiv:2406.19973Syntology official (archive's flag): 9 ran · 9 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 10 unverified (of 19 harvested samples)
-
Aligning Teacher with Student Preferences for Tailored Training Data Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19227