Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 2
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 2 of 56: papers 101 to 200 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Anatomy-Grounded Weakly Supervised Prompt Tuning for Chest X-ray Latent Diffusion Models 12 Jun 2025 · 0 repositories · arXiv:2506.10633
-
Fine-Grained Perturbation Guidance via Attention Head Selection 12 Jun 2025 · 0 repositories · arXiv:2506.10978
-
Harmonizing Geometry and Uncertainty: Diffusion with Hyperspheres 12 Jun 2025 · 0 repositories · arXiv:2506.10576
-
LLM-as-a-Fuzzy-Judge: Fine-Tuning Large Language Models as a Clinical Evaluation Judge with Fuzzy Logic 12 Jun 2025 · 1 repository · arXiv:2506.11221
-
Adding simple structure at inference improves Vision-Language Compositionality 11 Jun 2025 · 1 repository · arXiv:2506.09691
-
CEM-FBGTinyDet: Context-Enhanced Foreground Balance with Gradient Tuning for tiny Objects 11 Jun 2025 · 0 repositories · arXiv:2506.09897
-
DreamCS: Geometry-Aware Text-to-3D Generation with Unpaired 3D Reward Supervision 11 Jun 2025 · 0 repositories · arXiv:2506.09814
-
Features-based embedding or Feature-grounding 11 Jun 2025 · 0 repositories · arXiv:2506.22442
-
Geometric Regularity in Deterministic Sampling of Diffusion-based Generative Models 11 Jun 2025 · 0 repositories · arXiv:2506.10177
-
How Do People Revise Inconsistent Beliefs? Examining Belief Revision in Humans with User Studies 11 Jun 2025 · 0 repositories · arXiv:2506.09977
-
Leveraging Depth and Language for Open-Vocabulary Domain-Generalized Semantic Segmentation 11 Jun 2025 · 1 repository · arXiv:2506.09881
-
Natural Language Guided Ligand-Binding Protein Design 11 Jun 2025 · 0 repositories · arXiv:2506.09332
-
Vectorized Region Based Brush Strokes for Artistic Rendering 11 Jun 2025 · 0 repositories · arXiv:2506.09969
-
Neighbors and relatives: How do speech embeddings reflect linguistic connections across the world? 10 Jun 2025 · 0 repositories · arXiv:2506.08564
-
Societal AI Research Has Become Less Interdisciplinary 10 Jun 2025 · 0 repositories · arXiv:2506.08738
-
POLARON: Precision-aware On-device Learning and Adaptive Runtime-cONfigurable AI acceleration 10 Jun 2025 · 0 repositories · arXiv:2506.08785
-
BioLangFusion: Multimodal Fusion of DNA, mRNA, and Protein Language Models 10 Jun 2025 · 0 repositories · arXiv:2506.08936
-
Convergence of Spectral Principal Paths: How Deep Networks Distill Linear Representations from Noisy Inputs 10 Jun 2025 · 0 repositories · arXiv:2506.08543
-
Genetic Transformer-Assisted Quantum Neural Networks for Optimal Circuit Design 10 Jun 2025 · 0 repositories · arXiv:2506.09205
-
In Crowd Veritas: Leveraging Human Intelligence To Fight Misinformation 10 Jun 2025 · 0 repositories · arXiv:2506.09221
-
Low-resource domain adaptation while minimizing energy and hardware resource consumption 10 Jun 2025 · 0 repositories · arXiv:2506.08433
-
Multimodal Representation Alignment for Cross-modal Information Retrieval 10 Jun 2025 · 0 repositories · arXiv:2506.08774
-
RS-MTDF: Multi-Teacher Distillation and Fusion for Remote Sensing Semi-Supervised Semantic Segmentation 10 Jun 2025 · 1 repository · arXiv:2506.08772
-
Tailored Architectures for Time Series Forecasting: Evaluating Deep Learning Models on Gaussian Process-Generated Data 10 Jun 2025 · 1 repository · arXiv:2506.08977
-
Towards Secure and Private Language Models for Nuclear Power Plants 10 Jun 2025 · 0 repositories · arXiv:2506.08746
-
Dynamic Diffusion Schrödinger Bridge in Astrophysical Observational Inversions 9 Jun 2025 · 1 repository · arXiv:2506.08065Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints 9 Jun 2025 · 1 repository · arXiv:2506.08266
-
Instruction-Tuned Video-Audio Models Elucidate Functional Specialization in the Brain 9 Jun 2025 · 1 repository · arXiv:2506.08277
-
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation 9 Jun 2025 · 1 repository · arXiv:2506.07530
-
CommSense: A Rapid and Accurate ISAC Paradigm 9 Jun 2025 · 0 repositories · arXiv:2506.07685
-
Diffusion Sequence Models for Enhanced Protein Representation and Generation 9 Jun 2025 · 1 repository · arXiv:2506.08293
-
Gradients: When Markets Meet Fine-tuning -- A Distributed Approach to Model Optimisation 9 Jun 2025 · 0 repositories · arXiv:2506.07940
-
LLM-driven Indoor Scene Layout Generation via Scaled Human-aligned Data Synthesis and Multi-Stage Preference Optimization 9 Jun 2025 · 0 repositories · arXiv:2506.07570
-
Reinforcement Learning via Implicit Imitation Guidance 9 Jun 2025 · 0 repositories · arXiv:2506.07505
-
RSafe: Incentivizing proactive reasoning to build robust and adaptive LLM safeguards 9 Jun 2025 · 1 repository · arXiv:2506.07736
-
Shapley-Coop: Credit Assignment for Emergent Cooperation in Self-Interested LLM Agents 9 Jun 2025 · 0 repositories · arXiv:2506.07388
-
Super Encoding Network: Recursive Association of Multi-Modal Encoders for Video Understanding 9 Jun 2025 · 0 repositories · arXiv:2506.07576
-
Filling the Missings: Spatiotemporal Data Imputation by Conditional Diffusion 8 Jun 2025 · 1 repository · arXiv:2506.07099Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
How Far Are We from Optimal Reasoning Efficiency? 8 Jun 2025 · 1 repository · arXiv:2506.07104
-
MAGNet: A Multi-Scale Attention-Guided Graph Fusion Network for DRC Violation Detection 8 Jun 2025 · 0 repositories · arXiv:2506.07126
-
Future of Work with AI Agents: Auditing Automation and Augmentation Potential across the U.S. Workforce 6 Jun 2025 · 0 repositories · arXiv:2506.06576
-
Connectome brain fingerprinting: terminology, measures, and target properties 6 Jun 2025 · 0 repositories · arXiv:2506.05769
-
CrimeMind: Simulating Urban Crime with Multi-Modal LLM Agents 6 Jun 2025 · 0 repositories · arXiv:2506.05981
-
Information Bargaining: Bilateral Commitment in Bayesian Persuasion 6 Jun 2025 · 1 repository · arXiv:2506.05876
-
Latent Diffusion Model Based Denoising Receiver for 6G Semantic Communication: From Stochastic Differential Theory to Application 6 Jun 2025 · 0 repositories · arXiv:2506.05710
-
On Inverse Problems, Parameter Estimation, and Domain Generalization 6 Jun 2025 · 0 repositories · arXiv:2506.06024
-
Towards an Explainable Comparison and Alignment of Feature Embeddings 6 Jun 2025 · 1 repository · arXiv:2506.06231
-
Unintended Harms of Value-Aligned LLMs: Psychological and Empirical Insights 6 Jun 2025 · 1 repository · arXiv:2506.06404Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Vid2Sim: Generalizable, Video-based Reconstruction of Appearance, Geometry and Physics for Mesh-free Simulation 6 Jun 2025 · 0 repositories · arXiv:2506.06440
-
You Only Estimate Once: Unified, One-stage, Real-Time Category-level Articulated Object 6D Pose Estimation for Robotic Grasping 6 Jun 2025 · 0 repositories · arXiv:2506.05719
-
Agentic AI for Intent-Based Industrial Automation 5 Jun 2025 · 1 repository · arXiv:2506.04980
-
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning 5 Jun 2025 · 0 repositories · arXiv:2506.05128
-
From Play to Replay: Composed Video Retrieval for Temporally Fine-Grained Videos 5 Jun 2025 · 1 repository · arXiv:2506.05274
-
Handle-based Mesh Deformation Guided By Vision Language Model 5 Jun 2025 · 0 repositories · arXiv:2506.04562
-
Identifying Reliable Evaluation Metrics for Scientific Text Revision 5 Jun 2025 · 1 repository · arXiv:2506.04772
-
Information Locality as an Inductive Bias for Neural Language Models 5 Jun 2025 · 1 repository · arXiv:2506.05136
-
Multi-scale Image Super Resolution with a Single Auto-Regressive Model 5 Jun 2025 · 0 repositories · arXiv:2506.04990
-
MVP-Shapley: Feature-based Modeling for Evaluating the Most Valuable Player in Basketball 5 Jun 2025 · 0 repositories · arXiv:2506.04602
-
NEAT and HyperNEAT based Design for Soft Actuator Controllers 5 Jun 2025 · 0 repositories · arXiv:2506.04698
-
Neural Network Reprogrammability: A Unified Theme on Model Reprogramming, Prompt Tuning, and Prompt Instruction 5 Jun 2025 · 0 repositories · arXiv:2506.04650
-
PUB: An LLM-Enhanced Personality-Driven User Behaviour Simulator for Recommender System Evaluation 5 Jun 2025 · 0 repositories · arXiv:2506.04551
-
Reasoning or Overthinking: Evaluating Large Language Models on Financial Sentiment Analysis 5 Jun 2025 · 0 repositories · arXiv:2506.04574
-
SAM-aware Test-time Adaptation for Universal Medical Image Segmentation 5 Jun 2025 · 0 repositories · arXiv:2506.05221
-
Seeing the Invisible: Machine learning-Based QPI Kernel Extraction via Latent Alignment 5 Jun 2025 · 0 repositories · arXiv:2506.05325
-
Single GPU Task Adaptation of Pathology Foundation Models for Whole Slide Image Analysis 5 Jun 2025 · 0 repositories · arXiv:2506.05184
-
SPARTA ALIGNMENT: Collectively Aligning Multiple Language Models through Combat 5 Jun 2025 · 0 repositories · arXiv:2506.04721
-
Vision-Based Autonomous MM-Wave Reflector Using ArUco-Driven Angle-of-Arrival Estimation 5 Jun 2025 · 0 repositories · arXiv:2506.05195
-
Crowd-SFT: Crowdsourcing for LLM Alignment 4 Jun 2025 · 0 repositories · arXiv:2506.04063
-
Generating Pedagogically Meaningful Visuals for Math Word Problems: A New Benchmark and Analysis of Text-to-Image Models 4 Jun 2025 · 1 repository · arXiv:2506.03735
-
MMR-V: What's Left Unsaid? A Benchmark for Multimodal Deep Reasoning in Videos 4 Jun 2025 · 0 repositories · arXiv:2506.04141
-
RewardAnything: Generalizable Principle-Following Reward Models 4 Jun 2025 · 1 repository · arXiv:2506.03637
-
Seeing in the Dark: Benchmarking Egocentric 3D Vision with the Oxford Day-and-Night Dataset 4 Jun 2025 · 0 repositories · arXiv:2506.04224
-
HGOT: Self-supervised Heterogeneous Graph Neural Network with Optimal Transport 3 Jun 2025 · 0 repositories · arXiv:2506.02619
-
Do Language Models Think Consistently? A Study of Value Preferences Across Varying Response Lengths 3 Jun 2025 · 0 repositories · arXiv:2506.02481
-
EgoVLM: Policy Optimization for Egocentric Video Understanding 3 Jun 2025 · 1 repository · arXiv:2506.03097
-
Enhancing Lyrics Transcription on Music Mixtures with Consistency Loss 3 Jun 2025 · 0 repositories · arXiv:2506.02339
-
GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents 3 Jun 2025 · 0 repositories · arXiv:2506.03143
-
Large Language Models Can Achieve Explainable and Training-Free One-shot HRRP ATR 3 Jun 2025 · 0 repositories · arXiv:2506.02465
-
The Reader is the Metric: How Textual Features and Reader Profiles Explain Conflicting Evaluations of AI Creative Writing 3 Jun 2025 · 1 repository · arXiv:2506.03310
-
An Empirical Study of Group Conformity in Multi-Agent Systems 2 Jun 2025 · 0 repositories · arXiv:2506.01332
-
Confidence intervals for forced alignment boundaries using model ensembles 2 Jun 2025 · 1 repository · arXiv:2506.01256
-
CVC: A Large-Scale Chinese Value Rule Corpus for Value Alignment of Large Language Models 2 Jun 2025 · 1 repository · arXiv:2506.01495
-
IF-GUIDE: Influence Function-Guided Detoxification of LLMs 2 Jun 2025 · 1 repository · arXiv:2506.01790Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Is PMBOK Guide the Right Fit for AI? Re-evaluating Project Management in the Face of Artificial Intelligence Projects 2 Jun 2025 · 1 repository · arXiv:2506.02214
-
MODS: Multi-source Observations Conditional Diffusion Model for Meteorological State Downscaling 2 Jun 2025 · 0 repositories · arXiv:2506.14798
-
Multi-Platform Methane Plume Detection via Model and Domain Adaptation 2 Jun 2025 · 0 repositories · arXiv:2506.06348
-
SAM2-LOVE: Segment Anything Model 2 in Language-aided Audio-Visual Scenes 2 Jun 2025 · 0 repositories · arXiv:2506.01558
-
The Promise of Spiking Neural Networks for Ubiquitous Computing: A Survey and New Perspectives 2 Jun 2025 · 0 repositories · arXiv:2506.01737
-
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition 2 Jun 2025 · 1 repository · arXiv:2506.01411
-
AuralSAM2: Enabling SAM2 Hear Through Pyramid Audio-Visual Feature Prompting 1 Jun 2025 · 1 repository · arXiv:2506.01015
-
Beyond Attention: Learning Spatio-Temporal Dynamics with Emergent Interpretable Topologies 1 Jun 2025 · 0 repositories · arXiv:2506.00770
-
SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers 1 Jun 2025 · 3 repositories · arXiv:2506.00830
-
Con Instruction: Universal Jailbreaking of Multimodal Large Language Models via Non-Textual Modalities 31 May 2025 · 1 repository · arXiv:2506.00548
-
Enhancing Multimodal Continual Instruction Tuning with BranchLoRA 31 May 2025 · 0 repositories · arXiv:2506.02041
-
M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction 31 May 2025 · 1 repository · arXiv:2506.00466Syntology official (archive's flag): 30 ran · 30 ran (of which 23 constructed an object rather than computing a result; 28 with no instrument failure: 0 honoured, 0 violated, 28 with no contract checked; 2 where Syntology's instrument failed) · 11 unverified (of 41 harvested samples) · 41 pointer-only (licence)
-
A Mathematical Perspective On Contrastive Learning 30 May 2025 · 0 repositories · arXiv:2505.24134
-
A Simple Linear Patch Revives Layer-Pruned Large Language Models 30 May 2025 · 0 repositories · arXiv:2505.24680
-
Contrast-Invariant Self-supervised Segmentation for Quantitative Placental MRI 30 May 2025 · 0 repositories · arXiv:2505.24739
-
Designing AI Tools for Clinical Care Teams to Support Serious Illness Conversations with Older Adults in the Emergency Department 30 May 2025 · 0 repositories · arXiv:2506.00241
-
Don't Reinvent the Wheel: Efficient Instruction-Following Text Embedding based on Guided Space Transformation 30 May 2025 · 1 repository · arXiv:2505.24754Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)