Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 40
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 40 of 56: papers 3,901 to 4,000 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Generative AI in Higher Education: Seeing ChatGPT Through Universities' Policies, Resources, and Guidelines 8 Dec 2023 · 0 repositories · arXiv:2312.05235
-
User-Aware Prefix-Tuning is a Good Learner for Personalized Image Captioning 8 Dec 2023 · 0 repositories · arXiv:2312.04793
-
Combining inherent knowledge of vision-language models with unsupervised domain adaptation through strong-weak guidance 7 Dec 2023 · 1 repository · arXiv:2312.04066
-
Enhancing the Rationale-Input Alignment for Self-explaining Rationalization 7 Dec 2023 · 1 repository · arXiv:2312.04103Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Improved Face Representation via Joint Label Classification and Supervised Contrastive Clustering 7 Dec 2023 · 0 repositories · arXiv:2312.04029
-
Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations 7 Dec 2023 · 1 repository · arXiv:2312.06674
-
OT-Attack: Enhancing Adversarial Transferability of Vision-Language Models via Optimal Transport Optimization 7 Dec 2023 · 0 repositories · arXiv:2312.04403
-
Residual Graph Convolutional Network for Bird's-Eye-View Semantic Segmentation 7 Dec 2023 · 0 repositories · arXiv:2312.04044
-
TeMO: Towards Text-Driven 3D Stylization for Multi-Object Meshes 7 Dec 2023 · 0 repositories · arXiv:2312.04248
-
Validation and Comparison of Non-Stationary Cognitive Models: A Diffusion Model Application 7 Dec 2023 · 1 repository · arXiv:2401.08626Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
AnimateZero: Video Diffusion Models are Zero-Shot Image Animators 6 Dec 2023 · 1 repository · arXiv:2312.03793
-
Blueprinting the Future: Automatic Item Categorization using Hierarchical Zero-Shot and Few-Shot Classifiers 6 Dec 2023 · 0 repositories · arXiv:2312.03561
-
Diffusion Illusions: Hiding Images in Plain Sight 6 Dec 2023 · 0 repositories · arXiv:2312.03817
-
Indirect Gradient Matching for Adversarial Robust Distillation 6 Dec 2023 · 0 repositories · arXiv:2312.03286
-
Lite-Mind: Towards Efficient and Robust Brain Representation Network 6 Dec 2023 · 1 repository · arXiv:2312.03781Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
OneLLM: One Framework to Align All Modalities with Language 6 Dec 2023 · 1 repository · arXiv:2312.03700Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
GPT vs Human for Scientific Reviews: A Dual Source Review on Applications of ChatGPT in Science 5 Dec 2023 · 0 repositories · arXiv:2312.03769
-
Enhancing Content Moderation with Culturally-Aware Models 5 Dec 2023 · 0 repositories · arXiv:2312.02401
-
LLaRA: Large Language-Recommendation Assistant 5 Dec 2023 · 1 repository · arXiv:2312.02445Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Machine Vision Therapy: Multimodal Large Language Models Can Enhance Visual Robustness via Denoising In-Context Learning 5 Dec 2023 · 2 repositories · arXiv:2312.02546Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
ULMA: Unified Language Model Alignment with Human Demonstration and Point-wise Preference 5 Dec 2023 · 1 repository · arXiv:2312.02554Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Bootstrapping SparseFormers from Vision Foundation Models 4 Dec 2023 · 1 repository · arXiv:2312.01987
-
Distilled Self-Critique of LLMs with Synthetic Data: a Bayesian Perspective 4 Dec 2023 · 1 repository · arXiv:2312.01957Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Few-Shot Anomaly Detection with Adversarial Loss for Robust Feature Representations 4 Dec 2023 · 0 repositories · arXiv:2312.03005
-
Generating Action-conditioned Prompts for Open-vocabulary Video Action Recognition 4 Dec 2023 · 0 repositories · arXiv:2312.02226
-
Hot PATE: Private Aggregation of Distributions for Diverse Task 4 Dec 2023 · 0 repositories · arXiv:2312.02132
-
Simultaneous Alignment and Surface Regression Using Hybrid 2D-3D Networks for 3D Coherent Layer Segmentation of Retinal OCT Images with Full and Sparse Annotations 4 Dec 2023 · 1 repository · arXiv:2312.01726
-
The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning 4 Dec 2023 · 1 repository · arXiv:2312.01552
-
VideoSwap: Customized Video Subject Swapping with Interactive Semantic Point Correspondence 4 Dec 2023 · 0 repositories · arXiv:2312.02087
-
X-Adapter: Adding Universal Compatibility of Plugins for Upgraded Diffusion Model 4 Dec 2023 · 0 repositories · arXiv:2312.02238
-
Effectively Fine-tune to Improve Large Multimodal Models for Radiology Report Generation 3 Dec 2023 · 0 repositories · arXiv:2312.01504
-
Few-shot Shape Recognition by Learning Deep Shape-aware Features 3 Dec 2023 · 0 repositories · arXiv:2312.01315
-
Personality of AI 3 Dec 2023 · 0 repositories · arXiv:2312.02998
-
A New Learning Paradigm for Foundation Model-based Remote Sensing Change Detection 2 Dec 2023 · 2 repositories · arXiv:2312.01163
-
Axiomatic Preference Modeling for Longform Question Answering 2 Dec 2023 · 0 repositories · arXiv:2312.02206
-
Bootstrapping Interactive Image-Text Alignment for Remote Sensing Image Captioning 2 Dec 2023 · 1 repository · arXiv:2312.01191
-
Harnessing Discrete Representations For Continual Reinforcement Learning 2 Dec 2023 · 1 repository · arXiv:2312.01203Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
LinguaLinked: A Distributed Large Language Model Inference System for Mobile Devices 1 Dec 2023 · 0 repositories · arXiv:2312.00388
-
Segment and Caption Anything 1 Dec 2023 · 1 repository · arXiv:2312.00869
-
StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter 1 Dec 2023 · 3 repositories · arXiv:2312.00330Syntology official (archive's flag): 14 ran · 17 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 3 honoured, 3 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (of 23 harvested samples) · 7 pointer-only (licence)
-
Dataset Distillation via Curriculum Data Synthesis in Large Data Era 30 Nov 2023 · 1 repository · arXiv:2311.18838
-
Each Test Image Deserves A Specific Prompt: Continual Test-Time Adaptation for 2D Medical Image Segmentation 30 Nov 2023 · 1 repository · arXiv:2311.18363
-
Geometry-Aware Normalizing Wasserstein Flows for Optimal Causal Inference 30 Nov 2023 · 0 repositories · arXiv:2311.18826
-
Layered Rendering Diffusion Model for Controllable Zero-Shot Image Synthesis 30 Nov 2023 · 1 repository · arXiv:2311.18435
-
MicroCinema: A Divide-and-Conquer Approach for Text-to-Video Generation 30 Nov 2023 · 0 repositories · arXiv:2311.18829
-
mPLUG-PaperOwl: Scientific Diagram Analysis with the Multimodal Large Language Model 30 Nov 2023 · 1 repository · arXiv:2311.18248Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Leveraging Local Patch Alignment to Seam-cutting for Large Parallax Image Stitching 30 Nov 2023 · 1 repository · arXiv:2311.18564
-
VTimeLLM: Empower LLM to Grasp Video Moments 30 Nov 2023 · 1 repository · arXiv:2311.18445Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
C3Net: Compound Conditioned ControlNet for Multimodal Content Generation 29 Nov 2023 · 0 repositories · arXiv:2311.17951
-
Contrastive Vision-Language Alignment Makes Efficient Instruction Learner 29 Nov 2023 · 1 repository · arXiv:2311.17945
-
Grounding Foundation Models through Federated Transfer Learning: A General Framework 29 Nov 2023 · 0 repositories · arXiv:2311.17431
-
Contextual Knowledge Pursuit for Faithful Visual Synthesis 29 Nov 2023 · 1 repository · arXiv:2311.17898
-
Object-based (yet Class-agnostic) Video Domain Adaptation 29 Nov 2023 · 0 repositories · arXiv:2311.17942
-
SAMPro3D: Locating SAM Prompts in 3D for Zero-Shot Scene Segmentation 29 Nov 2023 · 1 repository · arXiv:2311.17707Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples) · 2 pointer-only (licence)
-
ShapeGPT: 3D Shape Generation with A Unified Multi-modal Language Model 29 Nov 2023 · 0 repositories · arXiv:2311.17618
-
SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis 29 Nov 2023 · 1 repository · arXiv:2311.17590Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
VBench: Comprehensive Benchmark Suite for Video Generative Models 29 Nov 2023 · 1 repository · arXiv:2311.17982
-
A Case for Competent AI Systems - A Concept Note 28 Nov 2023 · 0 repositories · arXiv:2312.00052
-
Embodied Multi-Modal Agent trained by an LLM from a Parallel TextWorld 28 Nov 2023 · 1 repository · arXiv:2311.16714Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
HandyPriors: Physically Consistent Perception of Hand-Object Interactions with Differentiable Priors 28 Nov 2023 · 0 repositories · arXiv:2311.16552
-
Large Model Based Referring Camouflaged Object Detection 28 Nov 2023 · 0 repositories · arXiv:2311.17122
-
Parameter Efficient Fine-tuning via Cross Block Orchestration for Segment Anything Model 28 Nov 2023 · 0 repositories · arXiv:2311.17112
-
A Two-Stage Adaptation of Large Language Models for Text Ranking 28 Nov 2023 · 1 repository · arXiv:2311.16720
-
Text-Driven Image Editing via Learnable Regions 28 Nov 2023 · 1 repository · arXiv:2311.16432Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples) · 2 pointer-only (licence)
-
Semi-supervised Segmentation of Histopathology Images with Noise-Aware Topological Consistency 28 Nov 2023 · 1 repository · arXiv:2311.16447
-
ArGue: Attribute-Guided Prompt Tuning for Vision-Language Models 27 Nov 2023 · 0 repositories · arXiv:2311.16494
-
Check, Locate, Rectify: A Training-Free Layout Calibration System for Text-to-Image Generation 27 Nov 2023 · 0 repositories · arXiv:2311.15773
-
DGR: Tackling Drifted and Correlated Noise in Quantum Error Correction via Decoding Graph Re-weighting 27 Nov 2023 · 0 repositories · arXiv:2311.16214
-
Fully Authentic Visual Question Answering Dataset from Online Communities 27 Nov 2023 · 1 repository · arXiv:2311.15562
-
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint 27 Nov 2023 · 1 repository · arXiv:2311.15864Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 8 pointer-only (licence)
-
Can Out-of-Domain data help to Learn Domain-Specific Prompts for Multimodal Misinformation Detection? 27 Nov 2023 · 1 repository · arXiv:2311.16496
-
LLMGA: Multimodal Large Language Model based Generation Assistant 27 Nov 2023 · 1 repository · arXiv:2311.16500
-
Quantum-classical simulation of quantum field theory by quantum circuit learning 27 Nov 2023 · 0 repositories · arXiv:2311.16297
-
Reinforcement Learning from Diffusion Feedback: Q* for Image Search 27 Nov 2023 · 0 repositories · arXiv:2311.15648
-
TFMQ-DM: Temporal Feature Maintenance Quantization for Diffusion Models 27 Nov 2023 · 1 repository · arXiv:2311.16503
-
Forking paths in financial economics 25 Nov 2023 · 0 repositories · arXiv:2401.08606
-
Data-Efficient Alignment of Large Language Models with Human Feedback Through Natural Language 24 Nov 2023 · 0 repositories · arXiv:2311.14543
-
Deciphering and integrating invariants for neural operator learning with various physical mechanisms 24 Nov 2023 · 1 repository · arXiv:2311.14361
-
Deformable multi-modal image registration for the correlation between optical measurements and histology images 24 Nov 2023 · 0 repositories · arXiv:2311.14414
-
Electric Vehicles coordination for grid balancing using multi-objective Harris Hawks Optimization 24 Nov 2023 · 0 repositories · arXiv:2311.14563
-
GeoChat: Grounded Large Vision-Language Model for Remote Sensing 24 Nov 2023 · 1 repository · arXiv:2311.15826Syntology official (archive's flag): 7 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
UniHPE: Towards Unified Human Pose Estimation via Contrastive Learning 24 Nov 2023 · 0 repositories · arXiv:2311.16477
-
Universal Jailbreak Backdoors from Poisoned Human Feedback 24 Nov 2023 · 2 repositories · arXiv:2311.14455Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Posterior Distillation Sampling 23 Nov 2023 · 0 repositories · arXiv:2311.13831
-
SySMOL: Co-designing Algorithms and Hardware for Neural Networks with Heterogeneous Precisions 23 Nov 2023 · 0 repositories · arXiv:2311.14114
-
Machine Translation to Control Formality Features in the Target Language 22 Nov 2023 · 0 repositories · arXiv:2311.13475
-
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning 22 Nov 2023 · 0 repositories · arXiv:2311.13721
-
A Baseline Analysis of Reward Models' Ability To Accurately Analyze Foundation Models Under Distribution Shift 21 Nov 2023 · 0 repositories · arXiv:2311.14743
-
Diffusion Model Alignment Using Direct Preference Optimization 21 Nov 2023 · 2 repositories · arXiv:2311.12908Syntology 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Fast calculation of Counterparty Credit exposures and associated sensitivities using fourier series expansion 21 Nov 2023 · 0 repositories · arXiv:2311.12575
-
Novel OCT mosaicking pipeline with Feature- and Pixel-based registration 21 Nov 2023 · 1 repository · arXiv:2311.13052
-
Stable Diffusion For Aerial Object Detection 21 Nov 2023 · 0 repositories · arXiv:2311.12345
-
SuGaR: Surface-Aligned Gaussian Splatting for Efficient 3D Mesh Reconstruction and High-Quality Mesh Rendering 21 Nov 2023 · 2 repositories · arXiv:2311.12775Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Enhancing Visual Grounding and Generalization: A Multi-Task Cycle Training Approach for Vision-Language Models 21 Nov 2023 · 0 repositories · arXiv:2311.12327
-
BadCLIP: Dual-Embedding Guided Backdoor Attack on Multimodal Contrastive Learning 20 Nov 2023 · 1 repository · arXiv:2311.12075Syntology 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
InfiMM-Eval: Complex Open-Ended Reasoning Evaluation For Multi-Modal Large Language Models 20 Nov 2023 · 0 repositories · arXiv:2311.11567
-
Human Learning by Model Feedback: The Dynamics of Iterative Prompting with Midjourney 20 Nov 2023 · 1 repository · arXiv:2311.12131
-
Measuring and Mitigating Biases in Motor Insurance Pricing 20 Nov 2023 · 0 repositories · arXiv:2311.11900
-
MILA: Memory-Based Instance-Level Adaptation for Cross-Domain Object Detection 20 Nov 2023 · 1 repository
-
MoVideo: Motion-Aware Video Generation with Diffusion Models 19 Nov 2023 · 0 repositories · arXiv:2311.11325