Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 16
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 16 of 56: papers 1,501 to 1,600 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
First-frame Supervised Video Polyp Segmentation via Propagative and Semantic Dual-teacher Network 21 Dec 2024 · 1 repository · arXiv:2412.16503
-
Towards Graph Foundation Models: Learning Generalities Across Graphs via Task-Trees 21 Dec 2024 · 0 repositories · arXiv:2412.16441Syntology 0 ran · 3 unverified (of 3 harvested samples)
-
LLaVA-SLT: Visual Language Tuning for Sign Language Translation 21 Dec 2024 · 0 repositories · arXiv:2412.16524
-
RFUDS -- A Brain Metastases Imaging Dataset of Radiotherapy Follow-Up 21 Dec 2024 · 0 repositories · arXiv:2412.16568
-
SubData: Bridging Heterogeneous Datasets to Enable Theory-Driven Evaluation of Political and Demographic Perspectives in LLMs 21 Dec 2024 · 0 repositories · arXiv:2412.16783
-
Transducer-Llama: Integrating LLMs into Streamable Transducer-based Speech Recognition 21 Dec 2024 · 0 repositories · arXiv:2412.16464
-
Trusted Mamba Contrastive Network for Multi-View Clustering 21 Dec 2024 · 1 repository · arXiv:2412.16487
-
3D Shape Tokenization via Latent Flow Matching 20 Dec 2024 · 0 repositories · arXiv:2412.15618
-
Contrastive Learning for Task-Independent SpeechLLM-Pretraining 20 Dec 2024 · 1 repository · arXiv:2412.15712
-
Deliberative Alignment: Reasoning Enables Safer Language Models 20 Dec 2024 · 0 repositories · arXiv:2412.16339
-
DINOv2 Meets Text: A Unified Framework for Image- and Pixel-Level Vision-Language Alignment 20 Dec 2024 · 1 repository · arXiv:2412.16334
-
Gaze Label Alignment: Alleviating Domain Shift for Gaze Estimation 20 Dec 2024 · 0 repositories · arXiv:2412.15601
-
GCA-3D: Towards Generalized and Consistent Domain Adaptation of 3D Generators 20 Dec 2024 · 0 repositories · arXiv:2412.15491
-
HoVLE: Unleashing the Power of Monolithic Vision-Language Models with Holistic Vision-Language Embedding 20 Dec 2024 · 0 repositories · arXiv:2412.16158
-
Humanlike Cognitive Patterns as Emergent Phenomena in Large Language Models 20 Dec 2024 · 0 repositories · arXiv:2412.15501
-
Modeling Autonomous Shifts Between Focus State and Mind-Wandering Using a Predictive-Coding-Inspired Variational RNN Model 20 Dec 2024 · 0 repositories · arXiv:2412.15620
-
MotiF: Making Text Count in Image Animation with Motion Focal Loss 20 Dec 2024 · 0 repositories · arXiv:2412.16153
-
Multi-Pair Temporal Sentence Grounding via Multi-Thread Knowledge Transfer Network 20 Dec 2024 · 0 repositories · arXiv:2412.15678
-
NeSyCoCo: A Neuro-Symbolic Concept Composer for Compositional Generalization 20 Dec 2024 · 1 repository · arXiv:2412.15588
-
RESQUE: Quantifying Estimator to Task and Distribution Shift for Sustainable Model Reusability 20 Dec 2024 · 1 repository · arXiv:2412.15511
-
Score-based Generative Diffusion Models for Social Recommendations 20 Dec 2024 · 1 repository · arXiv:2412.15579Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SeagrassFinder: Deep Learning for Eelgrass Detection and Coverage Estimation in the Wild 20 Dec 2024 · 0 repositories · arXiv:2412.16147
-
Towards Safe and Honest AI Agents with Neural Self-Other Overlap 20 Dec 2024 · 0 repositories · arXiv:2412.16325
-
A Comparative Study of DSPy Teleprompter Algorithms for Aligning Large Language Models Evaluation Metrics to Human Evaluation 19 Dec 2024 · 0 repositories · arXiv:2412.15298
-
A Super-pixel-based Approach to the Stable Interpretation of Neural Networks 19 Dec 2024 · 0 repositories · arXiv:2412.14509
-
CitaLaw: Enhancing LLM with Citations in Legal Domain 19 Dec 2024 · 0 repositories · arXiv:2412.14556
-
Dynamic User Interface Generation for Enhanced Human-Computer Interaction Using Variational Autoencoders 19 Dec 2024 · 0 repositories · arXiv:2412.14521
-
GenHMR: Generative Human Mesh Recovery 19 Dec 2024 · 0 repositories · arXiv:2412.14444
-
Knowing Where to Focus: Attention-Guided Alignment for Text-based Person Search 19 Dec 2024 · 0 repositories · arXiv:2412.15106
-
Multimodal Hypothetical Summary for Retrieval-based Multi-image Question Answering 19 Dec 2024 · 1 repository · arXiv:2412.14880
-
Northeastern Uni at Multilingual Counterspeech Generation: Enhancing Counter Speech Generation with LLM Alignment through Direct Preference Optimization 19 Dec 2024 · 0 repositories · arXiv:2412.15453
-
PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization 19 Dec 2024 · 1 repository · arXiv:2412.14510
-
A Generative Framework for Probabilistic, Spatiotemporally Coherent Downscaling of Climate Simulation 19 Dec 2024 · 1 repository · arXiv:2412.15361
-
Tree-of-Code: A Tree-Structured Exploring Framework for End-to-End Code Generation and Execution in Complex Task Handling 19 Dec 2024 · 0 repositories · arXiv:2412.15305
-
Circuits-Informed Machine Learning Technique for Blind Open-Loop Digital Calibration of SAR ADC 18 Dec 2024 · 0 repositories · arXiv:2412.14051
-
Context-DPO: Aligning Language Models for Context-Faithfulness 18 Dec 2024 · 1 repository · arXiv:2412.15280Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Dynamic Adapter with Semantics Disentangling for Cross-lingual Cross-modal Retrieval 18 Dec 2024 · 1 repository · arXiv:2412.13510
-
Exploring Query Efficient Data Generation towards Data-free Model Stealing in Hard Label Setting 18 Dec 2024 · 0 repositories · arXiv:2412.15276
-
Few-shot Steerable Alignment: Adapting Rewards and LLM Policies with Neural Processes 18 Dec 2024 · 1 repository · arXiv:2412.13998
-
LLMs can realize combinatorial creativity: generating creative ideas via LLMs for scientific research 18 Dec 2024 · 0 repositories · arXiv:2412.14141
-
Look Inside for More: Internal Spatial Modality Perception for 3D Anomaly Detection 18 Dec 2024 · 1 repository · arXiv:2412.13461
-
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation 18 Dec 2024 · 0 repositories · arXiv:2412.14148
-
Personalized Clustering via Targeted Representation Learning 18 Dec 2024 · 1 repository · arXiv:2412.13690Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Semantic Convergence: Harmonizing Recommender Systems via Two-Stage Alignment and Behavioral Semantic Tokenization 18 Dec 2024 · 0 repositories · arXiv:2412.13771
-
Text2Relight: Creative Portrait Relighting with Text Guidance 18 Dec 2024 · 0 repositories · arXiv:2412.13734
-
3DGUT: Enabling Distorted Cameras and Secondary Rays in Gaussian Splatting 17 Dec 2024 · 1 repository · arXiv:2412.12507Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Knowledge-enhanced Pathology Vision-language Foundation Model for Cancer Diagnosis 17 Dec 2024 · 2 repositories · arXiv:2412.13126
-
AnalogXpert: Automating Analog Topology Synthesis by Incorporating Circuit Design Expertise into Large Language Models 17 Dec 2024 · 0 repositories · arXiv:2412.19824
-
Boosting Fine-Grained Visual Anomaly Detection with Coarse-Knowledge-Aware Adversarial Learning 17 Dec 2024 · 1 repository · arXiv:2412.12850
-
Can Large Language Models Understand You Better? An MBTI Personality Detection Dataset Aligned with Population Traits 17 Dec 2024 · 1 repository · arXiv:2412.12510
-
Differential Alignment for Domain Adaptive Object Detection 17 Dec 2024 · 1 repository · arXiv:2412.12830
-
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment 17 Dec 2024 · 0 repositories · arXiv:2412.12902
-
MotionBridge: Dynamic Video Inbetweening with Flexible Controls 17 Dec 2024 · 0 repositories · arXiv:2412.13190
-
Multi-Dimensional Insights: Benchmarking Real-World Personalization in Large Multimodal Models 17 Dec 2024 · 0 repositories · arXiv:2412.12606
-
SAUGE: Taming SAM for Uncertainty-Aligned Multi-Granularity Edge Detection 17 Dec 2024 · 1 repository · arXiv:2412.12892
-
A Deep Learning Approach for Trading Factor Residuals 16 Dec 2024 · 0 repositories · arXiv:2412.11432
-
A Survey on Large Language Models for Communication, Network, and Service Management: Application Insights, Challenges, and Future Directions 16 Dec 2024 · 0 repositories · arXiv:2412.19823
-
ACE-M³: Automatic Capability Evaluator for Multimodal Medical Models 16 Dec 2024 · 0 repositories · arXiv:2412.11453
-
Achieving Collective Welfare in Multi-Agent Reinforcement Learning via Suggestion Sharing 16 Dec 2024 · 0 repositories · arXiv:2412.12326
-
Information-Geometric Barycenters for Bayesian Federated Learning 16 Dec 2024 · 0 repositories · arXiv:2412.11646
-
CLDA-YOLO: Visual Contrastive Learning Based Domain Adaptive YOLO Detector 16 Dec 2024 · 0 repositories · arXiv:2412.11812
-
Emergence of Power-Law and Other Wealth Distributions in Crowd of Heterogeneous Agents 16 Dec 2024 · 0 repositories · arXiv:2412.12393
-
Evaluating the Efficacy of Vectocardiographic and ECG Parameters for Efficient Tertiary Cardiology Care Allocation Using Decision Tree Analysis 16 Dec 2024 · 0 repositories · arXiv:2412.11839
-
Event-based Motion Deblurring via Multi-Temporal Granularity Fusion 16 Dec 2024 · 0 repositories · arXiv:2412.11866
-
EvoLlama: Enhancing LLMs' Understanding of Proteins via Multimodal Structure and Sequence Representations 16 Dec 2024 · 0 repositories · arXiv:2412.11618
-
Gramian Multimodal Representation Learning and Alignment 16 Dec 2024 · 2 repositories · arXiv:2412.11959Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 1 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
IGR: Improving Diffusion Model for Garment Restoration from Person Image 16 Dec 2024 · 0 repositories · arXiv:2412.11513
-
Large Language Models as Realistic Microservice Trace Generators 16 Dec 2024 · 1 repository · arXiv:2502.17439
-
Multimodal LLM for Intelligent Transportation Systems 16 Dec 2024 · 0 repositories · arXiv:2412.11683
-
Oriented Tiny Object Detection: A Dataset, Benchmark, and Dynamic Unbiased Learning 16 Dec 2024 · 0 repositories · arXiv:2412.11582
-
Probabilistic Behavioral Aggregation: A Case Study on the Nordic Power Grid 16 Dec 2024 · 0 repositories · arXiv:2412.11899
-
Re-Attentional Controllable Video Diffusion Editing 16 Dec 2024 · 2 repositories · arXiv:2412.11710
-
Self-Adaptive Paraphrasing and Preference Learning for Improved Claim Verifiability 16 Dec 2024 · 0 repositories · arXiv:2412.11653
-
Sequence Matters: Harnessing Video Models in 3D Super-Resolution 16 Dec 2024 · 0 repositories · arXiv:2412.11525
-
Temporal Contrastive Learning for Video Temporal Reasoning in Large Vision-Language Models 16 Dec 2024 · 0 repositories · arXiv:2412.11391
-
UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models 16 Dec 2024 · 1 repository · arXiv:2412.11803Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Understanding Knowledge Hijack Mechanism in In-context Learning through Associative Memory 16 Dec 2024 · 0 repositories · arXiv:2412.11459
-
Vertical Federated Unlearning via Backdoor Certification 16 Dec 2024 · 1 repository · arXiv:2412.11476
-
Combating Multimodal LLM Hallucination via Bottom-Up Holistic Reasoning 15 Dec 2024 · 0 repositories · arXiv:2412.11124
-
OccScene: Semantic Occupancy-based Cross-task Mutual Learning for 3D Scene Generation 15 Dec 2024 · 0 repositories · arXiv:2412.11183
-
Representation learning of dynamic networks 15 Dec 2024 · 0 repositories · arXiv:2412.11065
-
Smaller Language Models Are Better Instruction Evolvers 15 Dec 2024 · 1 repository · arXiv:2412.11231
-
Cocoa: Co-Planning and Co-Execution with AI Agents 14 Dec 2024 · 0 repositories · arXiv:2412.10999
-
Enhance Vision-Language Alignment with Noise 14 Dec 2024 · 1 repository · arXiv:2412.10817Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
PEARL: Input-Agnostic Prompt Enhancement with Negative Feedback Regulation for Class-Incremental Learning 14 Dec 2024 · 1 repository · arXiv:2412.10900
-
Pop-out vs. Glue: A Study on the pre-attentive and focused attention stages in Visual Search tasks 14 Dec 2024 · 0 repositories · arXiv:2412.12198
-
RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors 14 Dec 2024 · 0 repositories · arXiv:2412.10713
-
RWKV-Lite: Deeply Compressed RWKV for Resource-Constrained Devices 14 Dec 2024 · 0 repositories · arXiv:2412.10856
-
Video Diffusion Transformers are In-Context Learners 14 Dec 2024 · 1 repository · arXiv:2412.10783
-
Enhancing Fine-Grained Vision-Language Pretraining with Negative Augmented Samples 13 Dec 2024 · 0 repositories · arXiv:2412.10029
-
Generating 3D Pseudo-Healthy Knee MR Images to Support Trochleoplasty Planning 13 Dec 2024 · 1 repository · arXiv:2412.09962
-
Script-Based Dialog Policy Planning for LLM-Powered Conversational Agents: A Basic Architecture for an "AI Therapist" 13 Dec 2024 · 0 repositories · arXiv:2412.15242
-
Timealign: A multi-modal object detection method for time misalignment fusing in autonomous driving 13 Dec 2024 · 0 repositories · arXiv:2412.10033
-
Towards Unified Benchmark and Models for Multi-Modal Perceptual Metrics 13 Dec 2024 · 1 repository · arXiv:2412.10594
-
AI Predicts AGI: Leveraging AGI Forecasting and Peer Review to Explore LLMs' Complex Reasoning Capabilities 12 Dec 2024 · 1 repository · arXiv:2412.09385
-
ATPrompt: Textual Prompt Learning with Embedded Attributes 12 Dec 2024 · 1 repository · arXiv:2412.09442Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Dial-In LLM: Human-Aligned LLM-in-the-loop Intent Clustering for Customer Service Dialogues 12 Dec 2024 · 0 repositories · arXiv:2412.09049
-
Dynamic Contrastive Knowledge Distillation for Efficient Image Restoration 12 Dec 2024 · 1 repository · arXiv:2412.08939
-
Enhancing Facial Consistency in Conditional Video Generation via Facial Landmark Transformation 12 Dec 2024 · 0 repositories · arXiv:2412.08976
-
Federated Foundation Models on Heterogeneous Time Series 12 Dec 2024 · 1 repository · arXiv:2412.08906Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)