Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 7
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 7 of 56: papers 601 to 700 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Mirror: Multimodal Cognitive Reframing Therapy for Rolling with Resistance 16 Apr 2025 · 0 repositories · arXiv:2504.13211
-
On the calibration of Just-in-time Defect Prediction 16 Apr 2025 · 0 repositories · arXiv:2504.12051
-
Optimizing Utility-Scale Solar Siting for Local Economic Benefits and Regional Decarbonization 16 Apr 2025 · 0 repositories · arXiv:2504.12508
-
Real-World Depth Recovery via Structure Uncertainty Modeling and Inaccurate GT Depth Fitting 16 Apr 2025 · 0 repositories · arXiv:2504.11820
-
Standardization of Multi-Objective QUBOs 16 Apr 2025 · 0 repositories · arXiv:2504.12419
-
The Devil is in the Prompts: Retrieval-Augmented Prompt Optimization for Text-to-Video Generation 16 Apr 2025 · 0 repositories · arXiv:2504.11739
-
VGDFR: Diffusion-based Video Generation with Dynamic Latent Frame Rate 16 Apr 2025 · 1 repository · arXiv:2504.12259
-
3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians 15 Apr 2025 · 1 repository · arXiv:2504.11218
-
A Dual-Space Framework for General Knowledge Distillation of Large Language Models 15 Apr 2025 · 1 repository · arXiv:2504.11426
-
AFiRe: Anatomy-Driven Self-Supervised Learning for Fine-Grained Representation in Radiographic Images 15 Apr 2025 · 1 repository · arXiv:2504.10972
-
Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment 15 Apr 2025 · 0 repositories · arXiv:2504.10886
-
LOKA Protocol: A Decentralized Framework for Trustworthy and Ethical AI Agent Ecosystems 15 Apr 2025 · 0 repositories · arXiv:2504.10915
-
RAID: An In-Training Defense against Attribute Inference Attacks in Recommender Systems 15 Apr 2025 · 0 repositories · arXiv:2504.11510
-
Reinforcing Compositional Retrieval: Retrieving Step-by-Step for Composing Informative Contexts 15 Apr 2025 · 1 repository · arXiv:2504.11420
-
Revealing Covert Attention by Analyzing Human and Reinforcement Learning Agent Gameplay 15 Apr 2025 · 0 repositories · arXiv:2504.11118
-
REWARD CONSISTENCY: Improving Multi-Objective Alignment from a Data-Centric Perspective 15 Apr 2025 · 0 repositories · arXiv:2504.11337
-
Seedream 3.0 Technical Report 15 Apr 2025 · 0 repositories · arXiv:2504.11346
-
The Art of Audience Engagement: LLM-Based Thin-Slicing of Scientific Talks 15 Apr 2025 · 0 repositories · arXiv:2504.10768
-
A Survey of Personalization: From RAG to Agent 14 Apr 2025 · 1 repository · arXiv:2504.10147
-
BoTTA: Benchmarking on-device Test Time Adaptation 14 Apr 2025 · 0 repositories · arXiv:2504.10149
-
DataMosaic: Explainable and Verifiable Multi-Modal Data Analytics through Extract-Reason-Verify 14 Apr 2025 · 0 repositories · arXiv:2504.10036
-
DiffMOD: Progressive Diffusion Point Denoising for Moving Object Detection in Remote Sensing 14 Apr 2025 · 0 repositories · arXiv:2504.10278
-
Digital Staining with Knowledge Distillation: A Unified Framework for Unpaired and Paired-But-Misaligned Data 14 Apr 2025 · 1 repository · arXiv:2504.09899
-
Enhancing Multi-task Learning Capability of Medical Generalist Foundation Model via Image-centric Multi-annotation Data 14 Apr 2025 · 0 repositories · arXiv:2504.09967
-
From Prompting to Alignment: A Generative Framework for Query Recommendation 14 Apr 2025 · 0 repositories · arXiv:2504.10208
-
MorphTok: Morphologically Grounded Tokenization for Indian Languages 14 Apr 2025 · 0 repositories · arXiv:2504.10335
-
Plasticity-Aware Mixture of Experts for Learning Under QoE Shifts in Adaptive Video Streaming 14 Apr 2025 · 0 repositories · arXiv:2504.09906Syntology 4 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Self-Controlled Dynamic Expansion Model for Continual Learning 14 Apr 2025 · 0 repositories · arXiv:2504.10561
-
StePO-Rec: Towards Personalized Outfit Styling Assistant via Knowledge-Guided Multi-Step Reasoning 14 Apr 2025 · 0 repositories · arXiv:2504.09915
-
The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer 14 Apr 2025 · 1 repository · arXiv:2504.10462Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 9 harvested samples)
-
A highly maneuverable flying squirrel drone with agility-improving foldable wings 13 Apr 2025 · 0 repositories · arXiv:2504.09609
-
Causal integration of chemical structures improves representations of microscopy images for morphological profiling 13 Apr 2025 · 1 repository · arXiv:2504.09544
-
Distilling Transitional Pattern to Large Language Models for Multimodal Session-based Recommendation 13 Apr 2025 · 0 repositories · arXiv:2504.10538
-
Question Tokens Deserve More Attention: Enhancing Large Language Models without Training through Step-by-Step Reading and Question Attention Recalibration 13 Apr 2025 · 0 repositories · arXiv:2504.09402
-
BlockGaussian: Efficient Large-Scale Scene Novel View Synthesis via Adaptive Block-Based Gaussian Splatting 12 Apr 2025 · 1 repository · arXiv:2504.09048
-
MGS: Markov Greedy Sums for Accurate Low-Bitwidth Floating-Point Accumulation 12 Apr 2025 · 0 repositories · arXiv:2504.09072
-
An Evaluation of Cultural Value Alignment in LLM 11 Apr 2025 · 0 repositories · arXiv:2504.08863
-
Beyond Self-Reports: Multi-Observer Agents for Personality Assessment in Large Language Models 11 Apr 2025 · 0 repositories · arXiv:2504.08399
-
Designing Child-Friendly AI Interfaces: Six Developmentally-Appropriate Design Insights from Analysing Disney Animation 11 Apr 2025 · 0 repositories · arXiv:2504.08670
-
Exploring Cognitive Attributes in Financial Decision-Making 11 Apr 2025 · 0 repositories · arXiv:2504.08849
-
In-2-4D: Inbetweening from Two Single-View Images to 4D Generation 11 Apr 2025 · 0 repositories · arXiv:2504.08366
-
Large Language Model Empowered Recommendation Meets All-domain Continual Pre-Training 11 Apr 2025 · 0 repositories · arXiv:2504.08949
-
LGRPool: Hierarchical Graph Pooling Via Local-Global Regularisation 11 Apr 2025 · 0 repositories · arXiv:2504.08530
-
Scaling Up On-Device LLMs via Active-Weight Swapping Between DRAM and Flash 11 Apr 2025 · 0 repositories · arXiv:2504.08378
-
The Other Side of the Coin: Exploring Fairness in Retrieval-Augmented Generation 11 Apr 2025 · 1 repository · arXiv:2504.12323
-
VLMT: Vision-Language Multimodal Transformer for Multimodal Multi-hop Question Answering 11 Apr 2025 · 0 repositories · arXiv:2504.08269
-
Data Requirement Goal Modeling for Machine Learning Systems 10 Apr 2025 · 0 repositories · arXiv:2504.07664
-
From empirical brain networks towards modeling music perception -- a perspective 10 Apr 2025 · 0 repositories · arXiv:2504.07721
-
Gen3DEval: Using vLLMs for Automatic Evaluation of Generated 3D Objects 10 Apr 2025 · 0 repositories · arXiv:2504.08125
-
Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction 10 Apr 2025 · 1 repository · arXiv:2504.07961
-
Learning Higher-Order Interactions in Brain Networks via Topological Signal Processing 10 Apr 2025 · 0 repositories · arXiv:2504.07695
-
Learning Long Short-Term Intention within Human Daily Behaviors 10 Apr 2025 · 0 repositories · arXiv:2504.07597
-
Synthesizing High-Quality Programming Tasks with LLM-based Expert and Student Agents 10 Apr 2025 · 0 repositories · arXiv:2504.07655
-
Better Decisions through the Right Causal World Model 9 Apr 2025 · 0 repositories · arXiv:2504.07257
-
Image registration of 2D optical thin sections in a 3D porous medium: Application to a Berea sandstone digital rock image 9 Apr 2025 · 1 repository · arXiv:2504.06604
-
MultiADS: Defect-aware Supervision for Multi-type Anomaly Detection and Segmentation in Zero-Shot Learning 9 Apr 2025 · 0 repositories · arXiv:2504.06740
-
PinRec: Outcome-Conditioned, Multi-Token Generative Retrieval for Industry-Scale Recommendation Systems 9 Apr 2025 · 0 repositories · arXiv:2504.10507
-
Two by Two: Learning Multi-Task Pairwise Objects Assembly for Generalizable Robot Manipulation 9 Apr 2025 · 0 repositories · arXiv:2504.06961
-
A Training-Free Style-aligned Image Generation with Scale-wise Autoregressive Model 8 Apr 2025 · 0 repositories · arXiv:2504.06144
-
Text-to-Image Models and Their Representation of People from Different Nationalities Engaging in Activities 8 Apr 2025 · 0 repositories · arXiv:2504.06313
-
Autoencoder-Based Detection of Anomalous Stokes V Spectra in the Flare-Producing Active Region 13663 Using Hinode/SP Observations 8 Apr 2025 · 0 repositories · arXiv:2504.05962
-
Leanabell-Prover: Posttraining Scaling in Formal Reasoning 8 Apr 2025 · 1 repository · arXiv:2504.06122
-
On the merit principle in strategic exchange 8 Apr 2025 · 0 repositories · arXiv:2504.05678
-
PathGPT: Leveraging Large Language Models for Personalized Route Generation 8 Apr 2025 · 0 repositories · arXiv:2504.05846
-
The Hall of AI Fears and Hopes: Comparing the Views of AI Influencers and those of Members of the U.S. Public Through an Interactive Platform 8 Apr 2025 · 0 repositories · arXiv:2504.06016
-
Unified Generative Search and Recommendation 8 Apr 2025 · 0 repositories · arXiv:2504.05730
-
Why is Normalization Necessary for Linear Recommenders? 8 Apr 2025 · 1 repository · arXiv:2504.05805
-
LLM-Alignment Live-Streaming Recommendation 7 Apr 2025 · 0 repositories · arXiv:2504.05217
-
Low-Rate Semantic Communication with Codebook-based Conditional Generative Models 7 Apr 2025 · 0 repositories · arXiv:2504.04977
-
MedGNN: Capturing the Links Between Urban Characteristics and Medical Prescriptions 7 Apr 2025 · 0 repositories · arXiv:2504.04739
-
Riemannian Geometry for the classification of brain states with intracortical brain-computer interfaces 7 Apr 2025 · 0 repositories · arXiv:2504.05534
-
Exact Unlearning of Finetuning Data via Model Merging at Scale 6 Apr 2025 · 0 repositories · arXiv:2504.04626
-
FluentLip: A Phonemes-Based Two-stage Approach for Audio-Driven Lip Synthesis with Optical Flow Consistency 6 Apr 2025 · 0 repositories · arXiv:2504.04427
-
iADCPS: Time Series Anomaly Detection for Evolving Cyber-physical Systems via Incremental Meta-learning 6 Apr 2025 · 0 repositories · arXiv:2504.04374
-
CATS: Mitigating Correlation Shift for Multivariate Time Series Classification 5 Apr 2025 · 0 repositories · arXiv:2504.04283
-
Collaboration and Controversy Among Experts: Rumor Early Detection by Tuning a Comment Generator 5 Apr 2025 · 1 repository · arXiv:2504.04076
-
From Keypoints to Realism: A Realistic and Accurate Virtual Try-on Network from 2D Images 4 Apr 2025 · 0 repositories · arXiv:2504.03807
-
PIONM: A Generalized Approach to Solving Density-Constrained Mean-Field Games Equilibrium under Modified Boundary Conditions 4 Apr 2025 · 0 repositories · arXiv:2504.03209
-
A Framework for Situating Innovations, Opportunities, and Challenges in Advancing Vertical Systems with Large AI Models 3 Apr 2025 · 0 repositories · arXiv:2504.02793
-
A Sensorimotor Vision Transformer 3 Apr 2025 · 0 repositories · arXiv:2504.02536
-
How Deep Do Large Language Models Internalize Scientific Literature and Citation Practices? 3 Apr 2025 · 1 repository · arXiv:2504.02767
-
MAD: Makeup All-in-One with Cross-Domain Diffusion Model 3 Apr 2025 · 0 repositories · arXiv:2504.02545
-
Quantitative assessment of biological dynamics with aggregate data 3 Apr 2025 · 0 repositories · arXiv:2504.02581
-
Research Paper Recommender System by Considering Users' Information Seeking Behaviors 3 Apr 2025 · 0 repositories · arXiv:2504.02377
-
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models 3 Apr 2025 · 1 repository · arXiv:2504.02821Syntology official (archive's flag): 8 ran · 8 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
VoiceCraft-Dub: Automated Video Dubbing with Neural Codec Language Models 3 Apr 2025 · 0 repositories · arXiv:2504.02386
-
3DBonsai: Structure-Aware Bonsai Modeling Using Conditioned 3D Gaussian Splatting 2 Apr 2025 · 0 repositories · arXiv:2504.01619
-
Code Red! On the Harmfulness of Applying Off-the-shelf Large Language Models to Programming Tasks 2 Apr 2025 · 0 repositories · arXiv:2504.01850
-
Generative Retrieval and Alignment Model: A New Paradigm for E-commerce Retrieval 2 Apr 2025 · 0 repositories · arXiv:2504.01403
-
Leveraging Modality Tags for Enhanced Cross-Modal Video Retrieval 2 Apr 2025 · 0 repositories · arXiv:2504.01591
-
LVMed-R2: Perception and Reflection-driven Complex Reasoning for Medical Report Generation 2 Apr 2025 · 0 repositories · arXiv:2504.02885
-
OnRL-RAG: Real-Time Personalized Mental Health Dialogue System 2 Apr 2025 · 0 repositories · arXiv:2504.02894
-
Refining Interactions: Enhancing Anisotropy in Graph Neural Networks with Language Semantics 2 Apr 2025 · 0 repositories · arXiv:2504.01429
-
VideoScene: Distilling Video Diffusion Model to Generate 3D Scenes in One Step 2 Apr 2025 · 0 repositories · arXiv:2504.01956
-
Data-Driven Safety Verification using Barrier Certificates and Matrix Zonotopes 1 Apr 2025 · 1 repository · arXiv:2504.01007
-
Distilling Multi-view Diffusion Models into 3D Generators 1 Apr 2025 · 0 repositories · arXiv:2504.00457
-
FSSUWNet: Mitigating the Fragility of Pre-trained Models with Feature Enhancement for Few-Shot Semantic Segmentation in Underwater Images 1 Apr 2025 · 1 repository · arXiv:2504.00478
-
Multi-Agent LLM Judge: automatic personalized LLM judge design for evaluating natural language generation applications 1 Apr 2025 · 0 repositories · arXiv:2504.02867
-
OccludeNeRF: Geometric-aware 3D Scene Inpainting with Collaborative Score Distillation in NeRF 1 Apr 2025 · 0 repositories · arXiv:2504.02007
-
SPF-Portrait: Towards Pure Portrait Customization with Semantic Pollution-Free Fine-tuning 1 Apr 2025 · 0 repositories · arXiv:2504.00396