Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 44
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 44 of 56: papers 4,301 to 4,400 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
XATU: A Fine-grained Instruction-based Benchmark for Explainable Text Updates 20 Sep 2023 · 1 repository · arXiv:2309.11063
-
You Only Look at Screens: Multimodal Chain-of-Action Agents 20 Sep 2023 · 7 repositories · arXiv:2309.11436Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
MBR and QE Finetuning: Training-time Distillation of the Best and Most Expensive Decoding Methods 19 Sep 2023 · 0 repositories · arXiv:2309.10966
-
PICK: Polished & Informed Candidate Scoring for Knowledge-Grounded Dialogue Systems 19 Sep 2023 · 1 repository · arXiv:2309.10413
-
Semi-supervised Domain Adaptation in Graph Transfer Learning 19 Sep 2023 · 0 repositories · arXiv:2309.10773
-
What is the Best Automated Metric for Text to Motion Generation? 19 Sep 2023 · 0 repositories · arXiv:2309.10248
-
Collaborative Three-Stream Transformers for Video Captioning 18 Sep 2023 · 0 repositories · arXiv:2309.09611
-
Face-Driven Zero-Shot Voice Conversion with Memory-based Face-Voice Alignment 18 Sep 2023 · 0 repositories · arXiv:2309.09470
-
RaLF: Flow-based Global and Metric Radar Localization in LiDAR Maps 18 Sep 2023 · 0 repositories · arXiv:2309.09875
-
Scaling the time and Fourier domains to align periodically and their convolution 18 Sep 2023 · 0 repositories · arXiv:2309.09645
-
Wait, That Feels Familiar: Learning to Extrapolate Human Preferences for Preference Aligned Path Planning 18 Sep 2023 · 0 repositories · arXiv:2309.09912
-
DiffusionWorldViewer: Exposing and Broadening the Worldview Reflected by Generative Text-to-Image Models 18 Sep 2023 · 1 repository · arXiv:2309.09944
-
Exploring the impact of low-rank adaptation on the performance, efficiency, and regularization of RLHF 16 Sep 2023 · 1 repository · arXiv:2309.09055Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
Has Sentiment Returned to the Pre-pandemic Level? A Sentiment Analysis Using U.S. College Subreddit Data from 2019 to 2022 16 Sep 2023 · 1 repository · arXiv:2309.08845
-
Adaptive Communications in Collaborative Perception with Domain Alignment for Autonomous Driving 15 Sep 2023 · 0 repositories · arXiv:2310.00013
-
Beyond Domain Gap: Exploiting Subjectivity in Sketch-Based Person Retrieval 15 Sep 2023 · 1 repository · arXiv:2309.08372
-
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer 15 Sep 2023 · 1 repository · arXiv:2309.08583
-
IHT-Inspired Neural Network for Single-Snapshot DOA Estimation with Sparse Linear Arrays 15 Sep 2023 · 0 repositories · arXiv:2309.08429
-
Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context 15 Sep 2023 · 2 repositories · arXiv:2309.08105Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
"Merge Conflicts!" Exploring the Impacts of External Distractors to Parametric Knowledge Graphs 15 Sep 2023 · 1 repository · arXiv:2309.08594Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MusiLingo: Bridging Music and Text with Pre-trained Language Models for Music Captioning and Query Response 15 Sep 2023 · 1 repository · arXiv:2309.08730
-
Semi-supervised Sound Event Detection with Local and Global Consistency Regularization 15 Sep 2023 · 0 repositories · arXiv:2309.08355
-
Structural Self-Supervised Objectives for Transformers 15 Sep 2023 · 1 repository · arXiv:2309.08272
-
Topological Node2vec: Enhanced Graph Embedding via Persistent Homology 15 Sep 2023 · 1 repository · arXiv:2309.08241
-
A general Framework for Utilizing Metaheuristic Optimization for Sustainable Unrelated Parallel Machine Scheduling: A concise overview 14 Sep 2023 · 0 repositories · arXiv:2311.12802
-
Aligning Speakers: Evaluating and Visualizing Text-based Diarization Using Efficient Multiple Sequence Alignment (Extended Version) 14 Sep 2023 · 0 repositories · arXiv:2309.07677
-
Adapted Large Language Models Can Outperform Medical Experts in Clinical Text Summarization 14 Sep 2023 · 1 repository · arXiv:2309.07430Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
L1-aware Multilingual Mispronunciation Detection Framework 14 Sep 2023 · 0 repositories · arXiv:2309.07719
-
What Matters to Enhance Traffic Rule Compliance of Imitation Learning for End-to-End Autonomous Driving 14 Sep 2023 · 0 repositories · arXiv:2309.07808
-
RAIN: Your Language Models Can Align Themselves without Finetuning 13 Sep 2023 · 1 repository · arXiv:2309.07124Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
UnifiedGesture: A Unified Gesture Synthesis Model for Multiple Skeletons 13 Sep 2023 · 1 repository · arXiv:2309.07051
-
Mitigating the Alignment Tax of RLHF 12 Sep 2023 · 1 repository · arXiv:2309.06256Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploiting Machine Unlearning for Backdoor Attacks in Deep Learning System 12 Sep 2023 · 0 repositories · arXiv:2310.10659
-
Federated Learning for Large-Scale Scene Modeling with Neural Radiance Fields 12 Sep 2023 · 0 repositories · arXiv:2309.06030
-
ChemSpaceAL: An Efficient Active Learning Methodology Applied to Protein-Specific Molecular Generation 11 Sep 2023 · 2 repositories · arXiv:2309.05853
-
Instabilities in Convnets for Raw Audio 11 Sep 2023 · 1 repository · arXiv:2309.05855
-
Cross-tokamak Disruption Prediction based on Physics-Guided Feature Extraction and domain adaptation 11 Sep 2023 · 0 repositories · arXiv:2309.05361
-
Large Language Models for Difficulty Estimation of Foreign Language Content with Application to Language Learning 10 Sep 2023 · 0 repositories · arXiv:2309.05142
-
Real-time Learning of Driving Gap Preference for Personalized Adaptive Cruise Control 10 Sep 2023 · 0 repositories · arXiv:2309.05115
-
The Effect of Alignment Objectives on Code-Switching Translation 10 Sep 2023 · 0 repositories · arXiv:2309.05044
-
BiLMa: Bidirectional Local-Matching for Text-based Person Re-identification 9 Sep 2023 · 0 repositories · arXiv:2309.04675
-
Learning Spiking Neural Network from Easy to Hard task 9 Sep 2023 · 0 repositories · arXiv:2309.04737
-
DeformToon3D: Deformable 3D Toonification from Neural Radiance Fields 8 Sep 2023 · 1 repository · arXiv:2309.04410
-
MaskDiffusion: Boosting Text-to-Image Consistency with Conditional Mask 8 Sep 2023 · 0 repositories · arXiv:2309.04399
-
MoEController: Instruction-based Arbitrary Image Manipulation with Mixture-of-Expert Controllers 8 Sep 2023 · 0 repositories · arXiv:2309.04372
-
SegmentAnything helps microscopy images based automatic and quantitative organoid detection and analysis 8 Sep 2023 · 1 repository · arXiv:2309.04190
-
Automatic Concept Embedding Model (ACEM): No train-time concepts, No issue! 7 Sep 2023 · 0 repositories · arXiv:2309.03970
-
Autoregressive Omni-Aware Outpainting for Open-Vocabulary 360-Degree Image Generation 7 Sep 2023 · 1 repository · arXiv:2309.03467Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
DiffusionEngine: Diffusion Model is Scalable Data Engine for Object Detection 7 Sep 2023 · 0 repositories · arXiv:2309.03893
-
ImageBind-LLM: Multi-modality Instruction Tuning 7 Sep 2023 · 2 repositories · arXiv:2309.03905Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Multimodal Learning Framework for Comprehensive 3D Mineral Prospectivity Modeling with Jointly Learned Structure-Fluid Relationships 6 Sep 2023 · 0 repositories · arXiv:2309.02911
-
ETP: Learning Transferable ECG Representations via ECG-Text Pre-training 6 Sep 2023 · 0 repositories · arXiv:2309.07145
-
Parameter Efficient Audio Captioning With Faithful Guidance Using Audio-text Shared Latent Representation 6 Sep 2023 · 0 repositories · arXiv:2309.03340
-
3D View Prediction Models of the Dorsal Visual Stream 4 Sep 2023 · 0 repositories · arXiv:2309.01782
-
DCAlign v1.0: Aligning biological sequences using co-evolution models and informed priors 4 Sep 2023 · 0 repositories · arXiv:2309.01540
-
Learning Residual Elastic Warps for Image Stitching under Dirichlet Boundary Condition 4 Sep 2023 · 1 repository · arXiv:2309.01406
-
MDSC: Towards Evaluating the Style Consistency Between Music and Dance 4 Sep 2023 · 1 repository · arXiv:2309.01340
-
Multispectral Indices for Wildfire Management 4 Sep 2023 · 0 repositories · arXiv:2309.01751
-
Open Sesame! Universal Black Box Jailbreaking of Large Language Models 4 Sep 2023 · 0 repositories · arXiv:2309.01446
-
Text-Only Domain Adaptation for End-to-End Speech Recognition through Down-Sampling Acoustic Representation 4 Sep 2023 · 0 repositories · arXiv:2309.02459
-
Benchmarking Autoregressive Conditional Diffusion Models for Turbulent Flow Simulation 4 Sep 2023 · 1 repository · arXiv:2309.01745Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Unified Pre-training with Pseudo Texts for Text-To-Image Person Re-identification 4 Sep 2023 · 1 repository · arXiv:2309.01420Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 3 pointer-only (licence)
-
BDC-Adapter: Brownian Distance Covariance for Better Vision-Language Reasoning 3 Sep 2023 · 0 repositories · arXiv:2309.01256
-
EdaDet: Open-Vocabulary Object Detection Using Early Dense Alignment 3 Sep 2023 · 0 repositories · arXiv:2309.01151
-
MILA: Memory-Based Instance-Level Adaptation for Cross-Domain Object Detection 3 Sep 2023 · 1 repository · arXiv:2309.01086
-
Contrastive Grouping with Transformer for Referring Image Segmentation 2 Sep 2023 · 1 repository · arXiv:2309.01017Syntology official (archive's flag): 22 ran · 22 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 9 where Syntology's instrument failed) · 1 unverified (of 23 harvested samples) · 10 pointer-only (licence)
-
Toward Value-oriented Renewable Energy Forecasting: An Iterative Learning Approach 2 Sep 2023 · 1 repository · arXiv:2309.00803
-
Deep learning in medical image registration: introduction and survey 1 Sep 2023 · 0 repositories · arXiv:2309.00727
-
Fine-Grained Spatiotemporal Motion Alignment for Contrastive Video Representation Learning 1 Sep 2023 · 1 repository · arXiv:2309.00297Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Human-Inspired Facial Sketch Synthesis with Dynamic Adaptation 1 Sep 2023 · 1 repository · arXiv:2309.00216Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Let the Models Respond: Interpreting Language Model Detoxification Through the Lens of Prompt Dependence 1 Sep 2023 · 1 repository · arXiv:2309.00751
-
Reinforcement Learning with Human Feedback for Realistic Traffic Simulation 1 Sep 2023 · 0 repositories · arXiv:2309.00709
-
Trust your Good Friends: Source-free Domain Adaptation by Reciprocal Neighborhood Clustering 1 Sep 2023 · 0 repositories · arXiv:2309.00528
-
What Makes Good Open-Vocabulary Detector: A Disassembling Perspective 1 Sep 2023 · 0 repositories · arXiv:2309.00227
-
Using machine learning to understand causal relationships between urban form and travel CO2 emissions across continents 31 Aug 2023 · 1 repository · arXiv:2308.16599
-
Fine-Grained Cross-View Geo-Localization Using a Correlation-Aware Homography Estimator 31 Aug 2023 · 1 repository · arXiv:2308.16906Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
ViLTA: Enhancing Vision-Language Pre-training through Textual Augmentation 31 Aug 2023 · 0 repositories · arXiv:2308.16689
-
Physics-Informed DeepMRI: Bridging the Gap from Heat Diffusion to k-Space Interpolation 30 Aug 2023 · 0 repositories · arXiv:2308.15918
-
CLIPTrans: Transferring Visual Knowledge with Pre-trained Models for Multimodal Machine Translation 29 Aug 2023 · 1 repository · arXiv:2308.15226Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Learning Cross-modality Information Bottleneck Representation for Heterogeneous Person Re-Identification 29 Aug 2023 · 0 repositories · arXiv:2308.15063
-
Lifelike Agility and Play in Quadrupedal Robots using Reinforcement Learning and Generative Pre-trained Models 29 Aug 2023 · 0 repositories · arXiv:2308.15143
-
Adversarial Attacks on Foundational Vision Models 28 Aug 2023 · 0 repositories · arXiv:2308.14597
-
Gender bias and stereotypes in Large Language Models 28 Aug 2023 · 0 repositories · arXiv:2308.14921
-
Priority-Centric Human Motion Generation in Discrete Latent Space 28 Aug 2023 · 0 repositories · arXiv:2308.14480
-
A Unified Transformer-based Network for multimodal Emotion Recognition 27 Aug 2023 · 0 repositories · arXiv:2308.14160
-
Nonrigid Object Contact Estimation With Regional Unwrapping Transformer 27 Aug 2023 · 0 repositories · arXiv:2308.14074
-
Bias in Unsupervised Anomaly Detection in Brain MRI 26 Aug 2023 · 0 repositories · arXiv:2308.13861
-
Chunk, Align, Select: A Simple Long-sequence Processing Method for Transformers 25 Aug 2023 · 1 repository · arXiv:2308.13191
-
Decoding Natural Images from EEG for Object Recognition 25 Aug 2023 · 4 repositories · arXiv:2308.13234Syntology official (archive's flag): 6 ran · 9 ran (of which 5 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
Position-Enhanced Visual Instruction Tuning for Multimodal Large Language Models 25 Aug 2023 · 1 repository · arXiv:2308.13437Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Match-And-Deform: Time Series Domain Adaptation through Optimal Transport and Temporal Alignment 24 Aug 2023 · 1 repository · arXiv:2308.12686
-
On the Consistency of Average Embeddings for Item Recommendation 24 Aug 2023 · 1 repository · arXiv:2308.12767
-
POLCA: Power Oversubscription in LLM Cloud Providers 24 Aug 2023 · 1 repository · arXiv:2308.12908
-
Sentence Embedding Models for Ancient Greek Using Multilingual Knowledge Distillation 24 Aug 2023 · 2 repositories · arXiv:2308.13116
-
ToonTalker: Cross-Domain Face Reenactment 24 Aug 2023 · 0 repositories · arXiv:2308.12866
-
Aligning Language Models with Offline Learning from Human Feedback 23 Aug 2023 · 2 repositories · arXiv:2308.12050
-
DR-Tune: Improving Fine-tuning of Pretrained Visual Models by Distribution Regularization with Semantic Calibration 23 Aug 2023 · 1 repository · arXiv:2308.12058
-
From Instructions to Intrinsic Human Values -- A Survey of Alignment Goals for Big Models 23 Aug 2023 · 0 repositories · arXiv:2308.12014
-
Knowledge-injected Prompt Learning for Chinese Biomedical Entity Normalization 23 Aug 2023 · 0 repositories · arXiv:2308.12025
-
Understanding Dark Scenes by Contrasting Multi-Modal Observations 23 Aug 2023 · 1 repository · arXiv:2308.12320