Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 35
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 35 of 56: papers 3,401 to 3,500 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
JSTR: Joint Spatio-Temporal Reasoning for Event-based Moving Object Detection 12 Mar 2024 · 0 repositories · arXiv:2403.07436
-
Learning Generalizable Feature Fields for Mobile Manipulation 12 Mar 2024 · 0 repositories · arXiv:2403.07563
-
Physics-constrained Active Learning for Soil Moisture Estimation and Optimal Sensor Placement 12 Mar 2024 · 0 repositories · arXiv:2403.07228
-
CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning 12 Mar 2024 · 2 repositories · arXiv:2403.07300Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
A Comparative Study of Perceptual Quality Metrics for Audio-driven Talking Head Videos 11 Mar 2024 · 1 repository · arXiv:2403.06421
-
A Logical Pattern Memory Pre-trained Model for Entailment Tree Generation 11 Mar 2024 · 1 repository · arXiv:2403.06410
-
Answering Diverse Questions via Text Attached with Key Audio-Visual Clues 11 Mar 2024 · 1 repository · arXiv:2403.06679
-
AuG-KD: Anchor-Based Mixup Generation for Out-of-Domain Knowledge Distillation 11 Mar 2024 · 1 repository · arXiv:2403.07030Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
CoRAL: Collaborative Retrieval-Augmented Large Language Models Improve Long-tail Recommendation 11 Mar 2024 · 0 repositories · arXiv:2403.06447
-
Enhancing Image Caption Generation Using Reinforcement Learning with Human Feedback 11 Mar 2024 · 0 repositories · arXiv:2403.06735
-
Enhancing Semantic Fidelity in Text-to-Image Synthesis: Attention Regulation in Diffusion Models 11 Mar 2024 · 1 repository · arXiv:2403.06381
-
Guiding Clinical Reasoning with Large Language Models via Knowledge Seeds 11 Mar 2024 · 0 repositories · arXiv:2403.06609
-
OMH: Structured Sparsity via Optimally Matched Hierarchy for Unsupervised Semantic Segmentation 11 Mar 2024 · 0 repositories · arXiv:2403.06546
-
PointSeg: A Training-Free Paradigm for 3D Scene Segmentation via Foundation Models 11 Mar 2024 · 0 repositories · arXiv:2403.06403
-
SPAWNing Structural Priming Predictions from a Cognitively Motivated Parser 11 Mar 2024 · 0 repositories · arXiv:2403.07202
-
Split to Merge: Unifying Separated Modalities for Unsupervised Domain Adaptation 11 Mar 2024 · 1 repository · arXiv:2403.06946Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Cooperative Classification and Rationalization for Graph Generalization 10 Mar 2024 · 1 repository · arXiv:2403.06239Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
What Matters When Repurposing Diffusion Models for General Dense Perception Tasks? 10 Mar 2024 · 1 repository · arXiv:2403.06090Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
FARPLS: A Feature-Augmented Robot Trajectory Preference Labeling System to Assist Human Labelers' Preference Elicitation 10 Mar 2024 · 0 repositories · arXiv:2403.06267
-
Low-dose CT Denoising with Language-engaged Dual-space Alignment 10 Mar 2024 · 1 repository · arXiv:2403.06128
-
Text-Guided Variational Image Generation for Industrial Anomaly Detection and Segmentation 10 Mar 2024 · 0 repositories · arXiv:2403.06247
-
Adaptive Multi-modal Fusion of Spatially Variant Kernel Refinement with Diffusion Model for Blind Image Super-Resolution 9 Mar 2024 · 0 repositories · arXiv:2403.05808
-
Aligning Speech to Languages to Enhance Code-switching Speech Recognition 9 Mar 2024 · 0 repositories · arXiv:2403.05887
-
Privacy-Preserving Diffusion Model Using Homomorphic Encryption 9 Mar 2024 · 1 repository · arXiv:2403.05794
-
Recurrent Aligned Network for Generalized Pedestrian Trajectory Prediction 9 Mar 2024 · 0 repositories · arXiv:2403.05810
-
S²IP-LLM: Semantic Space Informed Prompt Learning with LLM for Time Series Forecasting 9 Mar 2024 · 1 repository · arXiv:2403.05798Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Algorithmic Identification of Essential Exogenous Nodes for Causal Sufficiency in Brain Networks 8 Mar 2024 · 0 repositories · arXiv:2403.05407
-
Are Large Language Models Aligned with People's Social Intuitions for Human-Robot Interactions? 8 Mar 2024 · 1 repository · arXiv:2403.05701
-
ChatUIE: Exploring Chat-based Unified Information Extraction using Large Language Models 8 Mar 2024 · 0 repositories · arXiv:2403.05132
-
DiffChat: Learning to Chat with Text-to-Image Synthesis Models for Interactive Image Creation 8 Mar 2024 · 0 repositories · arXiv:2403.04997
-
PipeRAG: Fast Retrieval-Augmented Generation via Algorithm-System Co-design 8 Mar 2024 · 0 repositories · arXiv:2403.05676
-
Provable Multi-Party Reinforcement Learning with Diverse Human Feedback 8 Mar 2024 · 0 repositories · arXiv:2403.05006
-
StereoDiffusion: Training-Free Stereo Image Generation Using Latent Diffusion Models 8 Mar 2024 · 1 repository · arXiv:2403.04965
-
Aligners: Decoupling LLMs and Alignment 7 Mar 2024 · 2 repositories · arXiv:2403.04224Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Aligning GPTRec with Beyond-Accuracy Goals with Reinforcement Learning 7 Mar 2024 · 1 repository · arXiv:2403.04875
-
An Image-based Typology for Visualization 7 Mar 2024 · 0 repositories · arXiv:2403.05594
-
Anatomy-Guided Surface Diffusion Model for Alzheimer's Disease Normative Modeling 7 Mar 2024 · 0 repositories · arXiv:2403.04531
-
DecompOpt: Controllable and Decomposed Diffusion Models for Structure-based Molecular Optimization 7 Mar 2024 · 0 repositories · arXiv:2403.13829
-
Discriminative Probing and Tuning for Text-to-Image Generation 7 Mar 2024 · 0 repositories · arXiv:2403.04321
-
Explaining Bayesian Optimization by Shapley Values Facilitates Human-AI Collaboration 7 Mar 2024 · 0 repositories · arXiv:2403.04629
-
MedM2G: Unifying Medical Multi-Modal Generation via Cross-Guided Diffusion with Visual Invariant 7 Mar 2024 · 0 repositories · arXiv:2403.04290
-
On the Essence and Prospect: An Investigation of Alignment Approaches for Big Models 7 Mar 2024 · 0 repositories · arXiv:2403.04204
-
Proxy-RLHF: Decoupling Generation and Alignment in Large Language Model with Proxy 7 Mar 2024 · 0 repositories · arXiv:2403.04283
-
Yi: Open Foundation Models by 01.AI 7 Mar 2024 · 1 repository · arXiv:2403.04652Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 8 harvested samples)
-
Causal Disentanglement for Regulating Social Influence Bias in Social Recommendation 6 Mar 2024 · 0 repositories · arXiv:2403.03578
-
Causal Prototype-inspired Contrast Adaptation for Unsupervised Domain Adaptive Semantic Segmentation of High-resolution Remote Sensing Imagery 6 Mar 2024 · 0 repositories · arXiv:2403.03704
-
Designing Informative Metrics for Few-Shot Example Selection 6 Mar 2024 · 0 repositories · arXiv:2403.03861
-
HDRFlow: Real-Time HDR Video Reconstruction with Large Motions 6 Mar 2024 · 0 repositories · arXiv:2403.03447
-
Interactive Melody Generation System for Enhancing the Creativity of Musicians 6 Mar 2024 · 0 repositories · arXiv:2403.03395
-
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators 6 Mar 2024 · 1 repository · arXiv:2403.03894
-
Portraying the Need for Temporal Data in Flood Detection via Sentinel-1 6 Mar 2024 · 0 repositories · arXiv:2403.03671
-
PromptCharm: Text-to-Image Generation through Multi-modal Prompting and Refinement 6 Mar 2024 · 1 repository · arXiv:2403.04014
-
Drug Resistance Predictions Based on a Directed Flag Transformer 5 Mar 2024 · 0 repositories · arXiv:2403.02603
-
In-Memory Learning: A Declarative Learning Framework for Large Language Models 5 Mar 2024 · 0 repositories · arXiv:2403.02757
-
MADTP: Multimodal Alignment-Guided Dynamic Token Pruning for Accelerating Vision-Language Transformer 5 Mar 2024 · 1 repository · arXiv:2403.02991Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Modeling Collaborator: Enabling Subjective Vision Classification With Minimal Human Effort via LLM Tool-Use 5 Mar 2024 · 0 repositories · arXiv:2403.02626
-
Motion-Corrected Moving Average: Including Post-Hoc Temporal Information for Improved Video Segmentation 5 Mar 2024 · 0 repositories · arXiv:2403.03120
-
RulePrompt: Weakly Supervised Text Classification with Prompting PLMs and Self-Iterative Logical Rules 5 Mar 2024 · 1 repository · arXiv:2403.02932
-
Towards Geometric-Photometric Joint Alignment for Facial Mesh Registration 5 Mar 2024 · 0 repositories · arXiv:2403.02629
-
Zero-Shot Cross-Lingual Document-Level Event Causality Identification with Heterogeneous Graph Contrastive Transfer Learning 5 Mar 2024 · 0 repositories · arXiv:2403.02893
-
Balancing Enhancement, Harmlessness, and General Capabilities: Enhancing Conversational LLMs with Direct RLHF 4 Mar 2024 · 0 repositories · arXiv:2403.02513
-
DACO: Towards Application-Driven and Comprehensive Data Analysis via Code Generation 4 Mar 2024 · 1 repository · arXiv:2403.02528Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
DragTex: Generative Point-Based Texture Editing on 3D Mesh 4 Mar 2024 · 0 repositories · arXiv:2403.02217
-
Enhancing LLM Safety via Constrained Direct Preference Optimization 4 Mar 2024 · 0 repositories · arXiv:2403.02475
-
HyperPredict: Estimating Hyperparameter Effects for Instance-Specific Regularization in Deformable Image Registration 4 Mar 2024 · 1 repository · arXiv:2403.02069
-
Ice-Tide: Implicit Cryo-ET Imaging and Deformation Estimation 4 Mar 2024 · 1 repository · arXiv:2403.02182
-
KorMedMCQA: Multi-Choice Question Answering Benchmark for Korean Healthcare Professional Licensing Examinations 3 Mar 2024 · 0 repositories · arXiv:2403.01469
-
SCHEMA: State CHangEs MAtter for Procedure Planning in Instructional Videos 3 Mar 2024 · 0 repositories · arXiv:2403.01599
-
Extending Complex Logical Queries on Uncertain Knowledge Graphs 3 Mar 2024 · 0 repositories · arXiv:2403.01508
-
Face Swap via Diffusion Model 2 Mar 2024 · 1 repository · arXiv:2403.01108
-
GraphRCG: Self-Conditioned Graph Generation 2 Mar 2024 · 0 repositories · arXiv:2403.01071
-
Mitigating the Bias in the Model for Continual Test-Time Adaptation 2 Mar 2024 · 0 repositories · arXiv:2403.01344
-
Deformable One-shot Face Stylization via DINO Semantic Guidance 1 Mar 2024 · 1 repository · arXiv:2403.00459
-
Gradient Cuff: Detecting Jailbreak Attacks on Large Language Models by Exploring Refusal Loss Landscapes 1 Mar 2024 · 0 repositories · arXiv:2403.00867Syntology 9 ran (of which 2 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
Policy Optimization for PDE Control with a Warm Start 1 Mar 2024 · 0 repositories · arXiv:2403.01005
-
Provably Robust DPO: Aligning Language Models with Noisy Feedback 1 Mar 2024 · 0 repositories · arXiv:2403.00409
-
Spatial Cascaded Clustering and Weighted Memory for Unsupervised Person Re-identification 1 Mar 2024 · 0 repositories · arXiv:2403.00261
-
Aligning Knowledge Graph with Visual Perception for Object-goal Navigation 29 Feb 2024 · 1 repository · arXiv:2402.18892
-
Effective Two-Stage Knowledge Transfer for Multi-Entity Cross-Domain Recommendation 29 Feb 2024 · 0 repositories · arXiv:2402.19101
-
Enhancing Visual Document Understanding with Contrastive Learning in Large Visual-Language Models 29 Feb 2024 · 0 repositories · arXiv:2402.19014
-
EAMA : Entity-Aware Multimodal Alignment Based Approach for News Image Captioning 29 Feb 2024 · 0 repositories · arXiv:2402.19404
-
Global and Local Prompts Cooperation via Optimal Transport for Federated Learning 29 Feb 2024 · 1 repository · arXiv:2403.00041Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Gradient Alignment for Cross-Domain Face Anti-Spoofing 29 Feb 2024 · 1 repository · arXiv:2402.18817
-
How do Large Language Models Handle Multilingualism? 29 Feb 2024 · 1 repository · arXiv:2402.18815Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Impacts of electric carsharing on a power sector with variable renewables 29 Feb 2024 · 1 repository · arXiv:2402.19380
-
ProtoP-OD: Explainable Object Detection with Prototypical Parts 29 Feb 2024 · 0 repositories · arXiv:2402.19142
-
Boosting Neural Representations for Videos with a Conditional Decoder 28 Feb 2024 · 1 repository · arXiv:2402.18152Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Coarse-to-Fine Latent Diffusion for Pose-Guided Person Image Synthesis 28 Feb 2024 · 1 repository · arXiv:2402.18078Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
DecisionNCE: Embodied Multimodal Representations via Implicit Preference Learning 28 Feb 2024 · 1 repository · arXiv:2402.18137Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Do Large Language Models Mirror Cognitive Language Processing? 28 Feb 2024 · 0 repositories · arXiv:2402.18023
-
Dark energy reconstruction analysis with artificial neural networks: Application on simulated Supernova Ia data from Rubin Observatory 28 Feb 2024 · 0 repositories · arXiv:2402.18124
-
Random Silicon Sampling: Simulating Human Sub-Population Opinion Using a Large Language Model Based on Group-Level Demographic Information 28 Feb 2024 · 0 repositories · arXiv:2402.18144
-
WIKIGENBENCH: Exploring Full-length Wikipedia Generation under Real-World Scenario 28 Feb 2024 · 1 repository · arXiv:2402.18264
-
AlignMiF: Geometry-Aligned Multimodal Implicit Field for LiDAR-Camera Joint Synthesis 27 Feb 2024 · 1 repository · arXiv:2402.17483Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Comparison of the Effects of Interaction with Intentional Agent and Artificial Intelligence using fNIRS 27 Feb 2024 · 0 repositories · arXiv:2402.17650
-
Efficiently Leveraging Linguistic Priors for Scene Text Spotting 27 Feb 2024 · 0 repositories · arXiv:2402.17134
-
Neural Networks for Portfolio-Level Risk Management: Portfolio Compression, Static Hedging, Counterparty Credit Risk Exposures and Impact on Capital Requirement 27 Feb 2024 · 0 repositories · arXiv:2402.17941
-
Against Filter Bubbles: Diversified Music Recommendation via Weighted Hypergraph Embedding Learning 26 Feb 2024 · 0 repositories · arXiv:2402.16299
-
Incremental Concept Formation over Visual Images Without Catastrophic Forgetting 26 Feb 2024 · 1 repository · arXiv:2402.16933
-
Read and Think: An Efficient Step-wise Multimodal Language Model for Document Understanding and Reasoning 26 Feb 2024 · 0 repositories · arXiv:2403.00816