Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 34
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 34 of 56: papers 3,301 to 3,400 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation 25 Mar 2024 · 1 repository · arXiv:2403.16990Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
Camera-aware Label Refinement for Unsupervised Person Re-identification 25 Mar 2024 · 1 repository · arXiv:2403.16450
-
CLHA: A Simple yet Effective Contrastive Learning Framework for Human Alignment 25 Mar 2024 · 1 repository · arXiv:2403.16649Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
CMViM: Contrastive Masked Vim Autoencoder for 3D Multi-modal Representation Learning for AD classification 25 Mar 2024 · 0 repositories · arXiv:2403.16520
-
Cross-lingual Contextualized Phrase Retrieval 25 Mar 2024 · 1 repository · arXiv:2403.16820Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
EDUE: Expert Disagreement-Guided One-Pass Uncertainty Estimation for Medical Image Segmentation 25 Mar 2024 · 0 repositories · arXiv:2403.16594
-
If CLIP Could Talk: Understanding Vision-Language Model Representations Through Their Preferred Concept Descriptions 25 Mar 2024 · 1 repository · arXiv:2403.16442Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Joint enhancement of automatic chest X-ray diagnosis and radiological gaze prediction with multi-stage cooperative learning 25 Mar 2024 · 0 repositories · arXiv:2403.16970
-
ProCQA: A Large-scale Community-based Programming Question Answering Dataset for Code Search 25 Mar 2024 · 1 repository · arXiv:2403.16702
-
RCBEVDet: Radar-camera Fusion in Bird's Eye View for 3D Object Detection 25 Mar 2024 · 1 repository · arXiv:2403.16440
-
Re2LLM: Reflective Reinforcement Large Language Model for Session-based Recommendation 25 Mar 2024 · 0 repositories · arXiv:2403.16427
-
Resolution Limit of Single-Photon LiDAR 25 Mar 2024 · 0 repositories · arXiv:2403.17719
-
Revisiting Boehmer et al. (2021): Recent Period, Alternative Method, Different Conclusions 25 Mar 2024 · 0 repositories · arXiv:2403.17095
-
SYNAPSE: SYmbolic Neural-Aided Preference Synthesis Engine 25 Mar 2024 · 0 repositories · arXiv:2403.16689
-
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation 25 Mar 2024 · 0 repositories · arXiv:2403.17001
-
Knowledge-aware Dual-side Attribute-enhanced Recommendation 24 Mar 2024 · 1 repository · arXiv:2403.16037
-
PVALane: Prior-Guided 3D Lane Detection with View-Agnostic Feature Alignment 24 Mar 2024 · 0 repositories
-
Anticipatory Gains and Event-Driven Losses in Blockchain-Based Fan Tokens: Evidence from the FIFA World Cup 23 Mar 2024 · 0 repositories · arXiv:2403.15810
-
Identifiable Latent Neural Causal Models 23 Mar 2024 · 0 repositories · arXiv:2403.15711
-
An Optimization Framework to Enforce Multi-View Consistency for Texturing 3D Meshes 22 Mar 2024 · 0 repositories · arXiv:2403.15559
-
Bilateral Unsymmetrical Graph Contrastive Learning for Recommendation 22 Mar 2024 · 0 repositories · arXiv:2403.15075
-
Brain-aligning of semantic vectors improves neural decoding of visual stimuli 22 Mar 2024 · 0 repositories · arXiv:2403.15176
-
ESG Classification by Implicit Rule Learning via GPT-4 22 Mar 2024 · 0 repositories · arXiv:2403.15040
-
Evidence-Driven Retrieval Augmented Response Generation for Online Misinformation 22 Mar 2024 · 0 repositories · arXiv:2403.14952
-
Multimodal Fusion with Pre-Trained Model Features in Affective Behaviour Analysis In-the-wild 22 Mar 2024 · 0 repositories · arXiv:2403.15044
-
Risk and Response in Large Language Models: Evaluating Key Threat Categories 22 Mar 2024 · 0 repositories · arXiv:2403.14988
-
Unifying Large Language Model and Deep Reinforcement Learning for Human-in-Loop Interactive Socially-aware Navigation 22 Mar 2024 · 0 repositories · arXiv:2403.15648
-
ThemeStation: Generating Theme-Aware 3D Assets from Few Exemplars 22 Mar 2024 · 1 repository · arXiv:2403.15383
-
Transactive Local Energy Markets Enable Community-Level Resource Coordination Using Individual Rewards 22 Mar 2024 · 0 repositories · arXiv:2403.15617
-
CFPL-FAS: Class Free Prompt Learning for Generalizable Face Anti-spoofing 21 Mar 2024 · 0 repositories · arXiv:2403.14333
-
DreamReward: Text-to-3D Generation with Human Preference 21 Mar 2024 · 0 repositories · arXiv:2403.14613
-
Learning Quadruped Locomotion Using Differentiable Simulation 21 Mar 2024 · 0 repositories · arXiv:2403.14864
-
OTSeg: Multi-prompt Sinkhorn Attention for Zero-Shot Semantic Segmentation 21 Mar 2024 · 1 repository · arXiv:2403.14183Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Reinforcement Learning from Reflective Feedback (RLRF): Aligning and Improving LLMs via Fine-Grained Self-Reflection 21 Mar 2024 · 0 repositories · arXiv:2403.14238
-
A Large Language Model Enhanced Sequential Recommender for Joint Video and Comment Recommendation 20 Mar 2024 · 1 repository · arXiv:2403.13574Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
AgentGroupChat: An Interactive Group Chat Simulacra For Better Eliciting Emergent Behavior 20 Mar 2024 · 1 repository · arXiv:2403.13433Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
BlendScape: Enabling End-User Customization of Video-Conferencing Environments through Generative AI 20 Mar 2024 · 0 repositories · arXiv:2403.13947
-
Building Optimal Neural Architectures using Interpretable Knowledge 20 Mar 2024 · 1 repository · arXiv:2403.13293
-
CLIPSwarm: Generating Drone Shows from Text Prompts with Vision-Language Models 20 Mar 2024 · 0 repositories · arXiv:2403.13467
-
Isometric Neural Machine Translation using Phoneme Count Ratio Reward-based Reinforcement Learning 20 Mar 2024 · 0 repositories · arXiv:2403.15469
-
Polaris: A Safety-focused LLM Constellation Architecture for Healthcare 20 Mar 2024 · 0 repositories · arXiv:2403.13313
-
RewardBench: Evaluating Reward Models for Language Modeling 20 Mar 2024 · 2 repositories · arXiv:2403.13787
-
S2DM: Sector-Shaped Diffusion Models for Video Generation 20 Mar 2024 · 0 repositories · arXiv:2403.13408
-
When Cars meet Drones: Hyperbolic Federated Learning for Source-Free Domain Adaptation in Adverse Weather 20 Mar 2024 · 1 repository · arXiv:2403.13762
-
Dated Data: Tracing Knowledge Cutoffs in Large Language Models 19 Mar 2024 · 1 repository · arXiv:2403.12958Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
DreamDA: Generative Data Augmentation with Diffusion Models 19 Mar 2024 · 1 repository · arXiv:2403.12803Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
EAS-SNN: End-to-End Adaptive Sampling and Representation for Event-based Detection with Recurrent Spiking Neural Networks 19 Mar 2024 · 1 repository · arXiv:2403.12574
-
ERASE: Benchmarking Feature Selection Methods for Deep Recommender Systems 19 Mar 2024 · 2 repositories · arXiv:2403.12660
-
Non-negative Contrastive Learning 19 Mar 2024 · 1 repository · arXiv:2403.12459Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Semantics, Distortion, and Style Matter: Towards Source-free UDA for Panoramic Segmentation 19 Mar 2024 · 0 repositories · arXiv:2403.12505
-
UniBind: LLM-Augmented Unified and Balanced Representation Space to Bind Them All 19 Mar 2024 · 0 repositories · arXiv:2403.12532
-
Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers 19 Mar 2024 · 0 repositories · arXiv:2403.12943
-
AICL: Action In-Context Learning for Video Diffusion Model 18 Mar 2024 · 1 repository · arXiv:2403.11535
-
Embracing the Generative AI Revolution: Advancing Tertiary Education in Cybersecurity with GPT 18 Mar 2024 · 0 repositories · arXiv:2403.11402
-
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection 18 Mar 2024 · 0 repositories · arXiv:2403.11848
-
Impart: An Imperceptible and Effective Label-Specific Backdoor Attack 18 Mar 2024 · 0 repositories · arXiv:2403.13017
-
Latent CLAP Loss for Better Foley Sound Synthesis 18 Mar 2024 · 1 repository · arXiv:2403.12182
-
Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning 18 Mar 2024 · 0 repositories · arXiv:2403.11401
-
Specific Emitter Identification Handling Modulation Variation with Margin Disparity Discrepancy 18 Mar 2024 · 1 repository · arXiv:2403.11531
-
WIA-LD2ND: Wavelet-based Image Alignment for Self-supervised Low-Dose CT Denoising 18 Mar 2024 · 1 repository · arXiv:2403.11672
-
Driving Style Alignment for LLM-powered Driver Agent 17 Mar 2024 · 2 repositories · arXiv:2403.11368
-
Training A Small Emotional Vision Language Model for Visual Art Comprehension 17 Mar 2024 · 2 repositories · arXiv:2403.11150
-
GazeFusion: Saliency-Guided Image Generation 16 Mar 2024 · 0 repositories · arXiv:2407.04191
-
Improving the Robustness of Dense Retrievers Against Typos via Multi-Positive Contrastive Learning 16 Mar 2024 · 1 repository · arXiv:2403.10939
-
LuoJiaHOG: A Hierarchy Oriented Geo-aware Image Caption Dataset for Remote Sensing Image-Text Retrival 16 Mar 2024 · 0 repositories · arXiv:2403.10887
-
Optimizing Language Augmentation for Multilingual Large Language Models: A Case Study on Korean 16 Mar 2024 · 0 repositories · arXiv:2403.10882
-
Benchmarking Adversarial Robustness of Image Shadow Removal with Shadow-adaptive Attacks 15 Mar 2024 · 0 repositories · arXiv:2403.10076
-
Leveraging Synthetic Data for Generalizable and Fair Facial Action Unit Detection 15 Mar 2024 · 0 repositories · arXiv:2403.10737
-
MeDSLIP: Medical Dual-Stream Language-Image Pre-training for Fine-grained Alignment 15 Mar 2024 · 1 repository · arXiv:2403.10635
-
Parameter Efficient Reinforcement Learning from Human Feedback 15 Mar 2024 · 0 repositories · arXiv:2403.10704
-
RadCLIP: Enhancing Radiologic Image Analysis through Contrastive Language-Image Pre-training 15 Mar 2024 · 1 repository · arXiv:2403.09948
-
Whose Side Are You On? Investigating the Political Stance of Large Language Models 15 Mar 2024 · 1 repository · arXiv:2403.13840
-
3D-VLA: A 3D Vision-Language-Action Generative World Model 14 Mar 2024 · 0 repositories · arXiv:2403.09631
-
Annotation Free Semantic Segmentation with Vision Foundation Models 14 Mar 2024 · 0 repositories · arXiv:2403.09307
-
Attention-based Class-Conditioned Alignment for Multi-Source Domain Adaptation of Object Detectors 14 Mar 2024 · 1 repository · arXiv:2403.09918
-
CodeUltraFeedback: An LLM-as-a-Judge Dataset for Aligning Large Language Models to Coding Preferences 14 Mar 2024 · 2 repositories · arXiv:2403.09032Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Task-Specific Adaptation of Segmentation Foundation Model via Prompt Learning 14 Mar 2024 · 0 repositories · arXiv:2403.09199
-
Video Editing via Factorized Diffusion Distillation 14 Mar 2024 · 0 repositories · arXiv:2403.09334
-
A Moral Imperative: The Need for Continual Superalignment of Large Language Models 13 Mar 2024 · 0 repositories · arXiv:2403.14683
-
AI coach for badminton 13 Mar 2024 · 0 repositories · arXiv:2403.08956
-
Automatic Interactive Evaluation for Large Language Models with State Aware Patient Simulator 13 Mar 2024 · 4 repositories · arXiv:2403.08495
-
CoIN: A Benchmark of Continual Instruction tuNing for Multimodel Large Language Model 13 Mar 2024 · 1 repository · arXiv:2403.08350Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Developing and Deploying Industry Standards for Artificial Intelligence in Education (AIED): Challenges, Strategies, and Future Directions 13 Mar 2024 · 0 repositories · arXiv:2403.14689
-
DialogGen: Multi-modal Interactive Dialogue System for Multi-turn Text-to-Image Generation 13 Mar 2024 · 1 repository · arXiv:2403.08857
-
Diffusion Models with Implicit Guidance for Medical Anomaly Detection 13 Mar 2024 · 1 repository · arXiv:2403.08464
-
Improving Implicit Regularization of SGD with Preconditioning for Least Square Problems 13 Mar 2024 · 0 repositories · arXiv:2403.08585
-
PathM3: A Multimodal Multi-Task Multiple Instance Learning Framework for Whole Slide Image Classification and Captioning 13 Mar 2024 · 0 repositories · arXiv:2403.08967
-
Skipformer: A Skip-and-Recover Strategy for Efficient Speech Recognition 13 Mar 2024 · 0 repositories · arXiv:2403.08258
-
Distract Large Language Models for Automatic Jailbreak Attack 13 Mar 2024 · 1 repository · arXiv:2403.08424Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Fourier Transform Framework for Domain Adaptation 12 Mar 2024 · 0 repositories · arXiv:2403.07798
-
Accurate Spatial Gene Expression Prediction by integrating Multi-resolution features 12 Mar 2024 · 1 repository · arXiv:2403.07592Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences 12 Mar 2024 · 0 repositories · arXiv:2403.07230
-
Debatrix: Multi-dimensional Debate Judge with Iterative Chronological Analysis Based on LLM 12 Mar 2024 · 2 repositories · arXiv:2403.08010
-
Decomposing Disease Descriptions for Enhanced Pathology Detection: A Multi-Aspect Vision-Language Pre-training Framework 12 Mar 2024 · 2 repositories · arXiv:2403.07636Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Deep Generative Domain Adaptation with Temporal Relation Knowledge for Cross-User Activity Recognition 12 Mar 2024 · 0 repositories · arXiv:2403.14682
-
Dynamic U-Net: Adaptively Calibrate Features for Abdominal Multi-organ Segmentation 12 Mar 2024 · 1 repository · arXiv:2403.07303
-
Empowering Sequential Recommendation from Collaborative Signals and Semantic Relatedness 12 Mar 2024 · 1 repository · arXiv:2403.07623
-
Federated Learning of Socially Appropriate Agent Behaviours in Simulated Home Environments 12 Mar 2024 · 1 repository · arXiv:2403.07586
-
Gujarati-English Code-Switching Speech Recognition using ensemble prediction of spoken language 12 Mar 2024 · 0 repositories · arXiv:2403.08011
-
Improving Reinforcement Learning from Human Feedback Using Contrastive Rewards 12 Mar 2024 · 0 repositories · arXiv:2403.07708