Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 33
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 33 of 56: papers 3,201 to 3,300 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
M-scan: A Multi-Scenario Causal-driven Adaptive Network for Recommendation 11 Apr 2024 · 0 repositories · arXiv:2404.07581
-
Language Models Meet Anomaly Detection for Better Interpretability and Generalizability 11 Apr 2024 · 1 repository · arXiv:2404.07622
-
Parameter Hierarchical Optimization for Visible-Infrared Person Re-Identification 11 Apr 2024 · 0 repositories · arXiv:2404.07930
-
Adapting LLaMA Decoder to Vision Transformer 10 Apr 2024 · 1 repository · arXiv:2404.06773Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 2 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
PoliTune: Analyzing the Impact of Data Selection and Fine-Tuning on Economic and Political Biases in Large Language Models 10 Apr 2024 · 3 repositories · arXiv:2404.08699
-
A predictive machine learning force field framework for liquid electrolyte development 10 Apr 2024 · 0 repositories · arXiv:2404.07181
-
Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion Models 10 Apr 2024 · 0 repositories · arXiv:2404.07389
-
Oxygen, Angiogenesis, Cancer and Immune Interplay in Breast Tumor Micro-Environment: A Computational Investigation 10 Apr 2024 · 0 repositories · arXiv:2404.06699
-
Unified Language-driven Zero-shot Domain Adaptation 10 Apr 2024 · 1 repository · arXiv:2404.07155Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Latent Distance Guided Alignment Training for Large Language Models 9 Apr 2024 · 0 repositories · arXiv:2404.06390
-
Prompt-driven Universal Model for View-Agnostic Echocardiography Analysis 9 Apr 2024 · 0 repositories · arXiv:2404.05916
-
Reconstructing Hand-Held Objects in 3D from Images and Videos 9 Apr 2024 · 0 repositories · arXiv:2404.06507
-
Rethinking How to Evaluate Language Model Jailbreak 9 Apr 2024 · 1 repository · arXiv:2404.06407
-
DLoRA: Distributed Parameter-Efficient Fine-Tuning Solution for Large Language Model 8 Apr 2024 · 0 repositories · arXiv:2404.05182
-
Learning Topology Uniformed Face Mesh by Volume Rendering for Multi-view Reconstruction 8 Apr 2024 · 0 repositories · arXiv:2404.05606
-
Mapping Network-Coordinated Stacked Gated Recurrent Units for Turbulence Prediction 8 Apr 2024 · 1 repository
-
PORTULAN ExtraGLUE Datasets and Models: Kick-starting a Benchmark for the Neural Processing of Portuguese 8 Apr 2024 · 0 repositories · arXiv:2404.05333
-
SpeechAlign: Aligning Speech Generation to Human Preferences 8 Apr 2024 · 1 repository · arXiv:2404.05600
-
The Hallucinations Leaderboard -- An Open Effort to Measure Hallucinations in Large Language Models 8 Apr 2024 · 0 repositories · arXiv:2404.05904
-
Towards Explainable Automated Neuroanatomy 8 Apr 2024 · 0 repositories · arXiv:2404.05814
-
AI for DevSecOps: A Landscape and Future Opportunities 7 Apr 2024 · 0 repositories · arXiv:2404.04839
-
AnimateZoo: Zero-shot Video Generation of Cross-Species Animation via Subject Alignment 7 Apr 2024 · 0 repositories · arXiv:2404.04946
-
Bootstrapping Chest CT Image Understanding by Distilling Knowledge from X-ray Expert Models 7 Apr 2024 · 0 repositories · arXiv:2404.04936
-
DWE+: Dual-Way Matching Enhanced Framework for Multimodal Entity Linking 7 Apr 2024 · 1 repository · arXiv:2404.04818
-
FGAIF: Aligning Large Vision-Language Models with Fine-grained AI Feedback 7 Apr 2024 · 0 repositories · arXiv:2404.05046
-
GauU-Scene V2: Assessing the Reliability of Image-Based Metrics with Expansive Lidar Image Dataset Using 3DGS and NeRF 7 Apr 2024 · 0 repositories · arXiv:2404.04880
-
Light the Night: A Multi-Condition Diffusion Framework for Unpaired Low-Light Enhancement in Autonomous Driving 7 Apr 2024 · 0 repositories · arXiv:2404.04804
-
Regularized Conditional Diffusion Model for Multi-Task Preference Alignment 7 Apr 2024 · 0 repositories · arXiv:2404.04920
-
A Comparison of Methods for Evaluating Generative IR 5 Apr 2024 · 1 repository · arXiv:2404.04044
-
Does Biomedical Training Lead to Better Medical Performance? 5 Apr 2024 · 1 repository · arXiv:2404.04067Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Idea23D: Collaborative LMM Agents Enable 3D Model Generation from Interleaved Multimodal Inputs 5 Apr 2024 · 1 repository · arXiv:2404.04363
-
Context-Aware Aerial Object Detection: Leveraging Inter-Object and Background Relationships 5 Apr 2024 · 0 repositories · arXiv:2404.04140
-
Distributionally Robust Alignment for Medical Federated Vision-Language Pre-training Under Data Heterogeneity 5 Apr 2024 · 0 repositories · arXiv:2404.03854
-
Pixel-wise RL on Diffusion Models: Reinforcement Learning from Rich Feedback 5 Apr 2024 · 0 repositories · arXiv:2404.04356
-
Verifiable by Design: Aligning Language Models to Quote from Pre-Training Data 5 Apr 2024 · 0 repositories · arXiv:2404.03862
-
BanglaAutoKG: Automatic Bangla Knowledge Graph Construction with Semantic Neural Graph Filtering 4 Apr 2024 · 1 repository · arXiv:2404.03528Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Designing for Human-Agent Alignment: Understanding what humans want from their agents 4 Apr 2024 · 0 repositories · arXiv:2404.04289
-
OpenNeRF: Open Set 3D Neural Scene Segmentation with Pixel-Wise Features and Rendered Novel Views 4 Apr 2024 · 0 repositories · arXiv:2404.03650
-
Personalized Federated Learning for Spatio-Temporal Forecasting: A Dual Semantic Alignment-Based Contrastive Approach 4 Apr 2024 · 0 repositories · arXiv:2404.03702
-
PRobELM: Plausibility Ranking Evaluation for Language Models 4 Apr 2024 · 0 repositories · arXiv:2404.03818
-
Social Media Emotions and Market Behavior 4 Apr 2024 · 0 repositories · arXiv:2404.03792
-
3DStyleGLIP: Part-Tailored Text-Guided 3D Neural Stylization 3 Apr 2024 · 1 repository · arXiv:2404.02634
-
Cross-Modal Conditioned Reconstruction for Language-guided Medical Image Segmentation 3 Apr 2024 · 2 repositories · arXiv:2404.02845Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Towards Explainable Traffic Flow Prediction with Large Language Models 3 Apr 2024 · 1 repository · arXiv:2404.02937
-
Exploring the Trade-off Between Model Performance and Explanation Plausibility of Text Classifiers Using Human Rationales 3 Apr 2024 · 1 repository · arXiv:2404.03098
-
GenN2N: Generative NeRF2NeRF Translation 3 Apr 2024 · 1 repository · arXiv:2404.02788
-
Independently Keypoint Learning for Small Object Semantic Correspondence 3 Apr 2024 · 0 repositories · arXiv:2404.02678
-
Weakly-Supervised 3D Scene Graph Generation via Visual-Linguistic Assisted Pseudo-labeling 3 Apr 2024 · 1 repository · arXiv:2404.02527
-
Stereotype Detection in LLMs: A Multiclass, Explainable, and Benchmark-Driven Approach 2 Apr 2024 · 0 repositories · arXiv:2404.01768
-
DELAN: Dual-Level Alignment for Vision-and-Language Navigation by Cross-Modal Contrastive Learning 2 Apr 2024 · 1 repository · arXiv:2404.01994
-
Disentangled Pre-training for Human-Object Interaction Detection 2 Apr 2024 · 1 repository · arXiv:2404.01725Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models 2 Apr 2024 · 0 repositories · arXiv:2404.02928
-
Polarity Calibration for Opinion Summarization 2 Apr 2024 · 1 repository · arXiv:2404.01706
-
Remote sensing framework for geological mapping via stacked autoencoders and clustering 2 Apr 2024 · 1 repository · arXiv:2404.02180
-
Segment Any 3D Object with Language 2 Apr 2024 · 0 repositories · arXiv:2404.02157
-
Test-Time Model Adaptation with Only Forward Passes 2 Apr 2024 · 1 repository · arXiv:2404.01650Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
CosmicMan: A Text-to-Image Foundation Model for Humans 1 Apr 2024 · 0 repositories · arXiv:2404.01294
-
Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models? 1 Apr 2024 · 1 repository · arXiv:2404.01399Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Evaluating Text-to-Visual Generation with Image-to-Text Generation 1 Apr 2024 · 3 repositories · arXiv:2404.01291
-
iMD4GC: Incomplete Multimodal Data Integration to Advance Precise Treatment Response Prediction and Survival Analysis for Gastric Cancer 1 Apr 2024 · 1 repository · arXiv:2404.01192
-
OVFoodSeg: Elevating Open-Vocabulary Food Image Segmentation via Image-Informed Textual Representation 1 Apr 2024 · 0 repositories · arXiv:2404.01409
-
Scalable 3D Registration via Truncated Entry-wise Absolute Residuals 1 Apr 2024 · 1 repository · arXiv:2404.00915
-
Slightly Shift New Classes to Remember Old Classes for Video Class-Incremental Learning 1 Apr 2024 · 0 repositories · arXiv:2404.00901
-
SyncMask: Synchronized Attentional Masking for Fashion-centric Vision-Language Pretraining 1 Apr 2024 · 0 repositories · arXiv:2404.01156
-
The Double-Edged Sword of Input Perturbations to Robust Accurate Fairness 1 Apr 2024 · 0 repositories · arXiv:2404.01356
-
Absolute-Unified Multi-Class Anomaly Detection via Class-Agnostic Distribution Alignment 31 Mar 2024 · 0 repositories · arXiv:2404.00724
-
Modeling State Shifting via Local-Global Distillation for Event-Frame Gaze Tracking 31 Mar 2024 · 1 repository · arXiv:2404.00548
-
DiffAgent: Fast and Accurate Text-to-Image API Selection with Large Language Model 31 Mar 2024 · 1 repository · arXiv:2404.01342
-
Dual DETRs for Multi-Label Temporal Action Detection 31 Mar 2024 · 0 repositories · arXiv:2404.00653
-
3DGSR: Implicit Surface Reconstruction with 3D Gaussian Splatting 30 Mar 2024 · 0 repositories · arXiv:2404.00409
-
Aurora-M: Open Source Continual Pre-training for Multilingual Language and Code 30 Mar 2024 · 0 repositories · arXiv:2404.00399
-
Enhancing Content-based Recommendation via Large Language Model 30 Mar 2024 · 1 repository · arXiv:2404.00236
-
Monocular Identity-Conditioned Facial Reflectance Reconstruction 30 Mar 2024 · 0 repositories · arXiv:2404.00301
-
TTD: Text-Tag Self-Distillation Enhancing Image-Text Alignment in CLIP to Alleviate Single Tag Bias 30 Mar 2024 · 1 repository · arXiv:2404.00384Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
An Interpretable Cross-Attentive Multi-modal MRI Fusion Framework for Schizophrenia Diagnosis 29 Mar 2024 · 0 repositories · arXiv:2404.00144
-
ELITR-Bench: A Meeting Assistant Benchmark for Long-Context Language Models 29 Mar 2024 · 1 repository · arXiv:2403.20262
-
GDA: Generalized Diffusion for Robust Test-time Adaptation 29 Mar 2024 · 0 repositories · arXiv:2404.00095
-
A Real-Time Framework for Domain-Adaptive Underwater Object Detection with Image Enhancement 28 Mar 2024 · 0 repositories · arXiv:2403.19079
-
Burst Super-Resolution with Diffusion Models for Improving Perceptual Quality 28 Mar 2024 · 1 repository · arXiv:2403.19428
-
Channel Deduction: A New Learning Framework to Acquire Channel from Outdated Samples and Coarse Estimate 28 Mar 2024 · 0 repositories · arXiv:2403.19409
-
Developing Healthcare Language Model Embedding Spaces 28 Mar 2024 · 0 repositories · arXiv:2403.19802
-
InterDreamer: Zero-Shot Text to 3D Dynamic Human-Object Interaction 28 Mar 2024 · 0 repositories · arXiv:2403.19652
-
JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models 28 Mar 2024 · 3 repositories · arXiv:2404.01318Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Jointly Training and Pruning CNNs via Learnable Agent Guidance and Alignment 28 Mar 2024 · 0 repositories · arXiv:2403.19490
-
X-MIC: Cross-Modal Instance Conditioning for Egocentric Action Generalization 28 Mar 2024 · 1 repository · arXiv:2403.19811
-
Teaching AI the Anatomy Behind the Scan: Addressing Anatomical Flaws in Medical Image Segmentation with Learnable Prior 27 Mar 2024 · 0 repositories · arXiv:2403.18878
-
Aiming for Relevance 27 Mar 2024 · 0 repositories · arXiv:2403.18668
-
FlexEdit: Flexible and Controllable Diffusion-based Object-centric Image Editing 27 Mar 2024 · 0 repositories · arXiv:2403.18605
-
Growth rate of liquidity provider's wealth in G3Ms 27 Mar 2024 · 0 repositories · arXiv:2403.18177
-
Learning CNN on ViT: A Hybrid Model to Explicitly Class-specific Boundaries for Domain Adaptation 27 Mar 2024 · 2 repositories · arXiv:2403.18360
-
PLOT-TAL -- Prompt Learning with Optimal Transport for Few-Shot Temporal Action Localization 27 Mar 2024 · 0 repositories · arXiv:2403.18915
-
SemRoDe: Macro Adversarial Training to Learn Representations That are Robust to Word-Level Attacks 27 Mar 2024 · 1 repository · arXiv:2403.18423Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
"Sorry, Come Again?" Prompting -- Enhancing Comprehension and Diminishing Hallucination with [PAUSE]-injected Optimal Paraphrasing 27 Mar 2024 · 0 repositories · arXiv:2403.18976
-
IDGenRec: LLM-RecSys Alignment with Textual ID Learning 27 Mar 2024 · 1 repository · arXiv:2403.19021Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
What are human values, and how do we align AI to them? 27 Mar 2024 · 0 repositories · arXiv:2404.10636
-
COIG-CQIA: Quality is All You Need for Chinese Instruction Fine-tuning 26 Mar 2024 · 0 repositories · arXiv:2403.18058
-
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis 26 Mar 2024 · 1 repository · arXiv:2403.17936
-
WordRobe: Text-Guided Generation of Textured 3D Garments 26 Mar 2024 · 0 repositories · arXiv:2403.17541
-
Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators 25 Mar 2024 · 1 repository · arXiv:2403.16950Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
An Intermediate Fusion ViT Enables Efficient Text-Image Alignment in Diffusion Models 25 Mar 2024 · 0 repositories · arXiv:2403.16530