Methods › Computer Vision › Vision and Language Pre-Trained Models › ALIGN › Papers, page 25
ALIGN
Papers archive 2025-07-28
archive papers tagged: 5,527 · with a code link: 2,162 · where Syntology ran a sample: 726 (628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (726 of 5,527 tagged: 628 with a run with no instrument failure, 98 where every run was a failure of Syntology's instrument)
Page 25 of 56: papers 2,401 to 2,500 of 5,524, newest first by the archive's date (ties by slug), in archive order.
3 tagged papers are not listed: the archive title is spam (see /not-shown).
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Federated Clustering: An Unsupervised Cluster-Wise Training for Decentralized Data Distributions 20 Aug 2024 · 0 repositories · arXiv:2408.10664
-
Investigating Context Effects in Similarity Judgements in Large Language Models 20 Aug 2024 · 0 repositories · arXiv:2408.10711
-
Large Language Models for Multimodal Deformable Image Registration 20 Aug 2024 · 1 repository · arXiv:2408.10703
-
Minor SFT loss for LLM fine-tune to increase performance and reduce model deviation 20 Aug 2024 · 0 repositories · arXiv:2408.10642
-
MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval 20 Aug 2024 · 1 repository · arXiv:2408.10575
-
QUITO-X: A New Perspective on Context Compression from the Information Bottleneck Theory 20 Aug 2024 · 0 repositories · arXiv:2408.10497
-
Towards Efficient Large Language Models for Scientific Text: A Review 20 Aug 2024 · 0 repositories · arXiv:2408.10729
-
Can an unsupervised clustering algorithm reproduce a categorization system? 19 Aug 2024 · 0 repositories · arXiv:2408.10340
-
Caption-Driven Explorations: Aligning Image and Text Embeddings through Human-Inspired Foveated Vision 19 Aug 2024 · 0 repositories · arXiv:2408.09948
-
DELIA: Diversity-Enhanced Learning for Instruction Adaptation in Large Language Models 19 Aug 2024 · 0 repositories · arXiv:2408.10841
-
Demystifying Reinforcement Learning in Production Scheduling via Explainable AI 19 Aug 2024 · 0 repositories · arXiv:2408.09841
-
Hear Your Face: Face-based voice conversion with F0 estimation 19 Aug 2024 · 1 repository · arXiv:2408.09802
-
Minor DPO reject penalty to increase training robustness 19 Aug 2024 · 0 repositories · arXiv:2408.09834
-
MSDiagnosis: A Benchmark for Evaluating Large Language Models in Multi-Step Clinical Diagnosis 19 Aug 2024 · 0 repositories · arXiv:2408.10039
-
Pose-GuideNet: Automatic Scanning Guidance for Fetal Head Ultrasound from Pose Estimation 19 Aug 2024 · 0 repositories · arXiv:2408.09931
-
RealCustom++: Representing Images as Real-Word for Real-Time Customization 19 Aug 2024 · 0 repositories · arXiv:2408.09744
-
SpaRP: Fast 3D Object Reconstruction and Pose Estimation from Sparse Views 19 Aug 2024 · 0 repositories · arXiv:2408.10195
-
Value Alignment from Unstructured Text 19 Aug 2024 · 0 repositories · arXiv:2408.10392
-
Efficient Budget Allocation for Large-Scale LLM-Enabled Virtual Screening 18 Aug 2024 · 0 repositories · arXiv:2408.09537
-
Towards Boosting LLMs-driven Relevance Modeling with Progressive Retrieved Behavior-augmented Prompting 18 Aug 2024 · 0 repositories · arXiv:2408.09439
-
Quality Assessment in the Era of Large Models: A Survey 17 Aug 2024 · 0 repositories · arXiv:2409.00031
-
SA-GDA: Spectral Augmentation for Graph Domain Adaptation 17 Aug 2024 · 0 repositories · arXiv:2408.09189
-
Trust-Oriented Adaptive Guardrails for Large Language Models 16 Aug 2024 · 0 repositories · arXiv:2408.08959
-
An End-to-End Model for Photo-Sharing Multi-modal Dialogue Generation 16 Aug 2024 · 1 repository · arXiv:2408.08650
-
Constructing Domain-Specific Evaluation Sets for LLM-as-a-judge 16 Aug 2024 · 0 repositories · arXiv:2408.08808
-
CoSEC: A Coaxial Stereo Event Camera Dataset for Autonomous Driving 16 Aug 2024 · 0 repositories · arXiv:2408.08500
-
Evaluating the Evaluator: Measuring LLMs' Adherence to Task Evaluation Instructions 16 Aug 2024 · 0 repositories · arXiv:2408.08781
-
From Lazy to Prolific: Tackling Missing Labels in Open Vocabulary Extreme Classification by Positive-Unlabeled Sequence Learning 16 Aug 2024 · 0 repositories · arXiv:2408.08981
-
Math-PUMA: Progressive Upward Multimodal Alignment to Enhance Mathematical Reasoning 16 Aug 2024 · 1 repository · arXiv:2408.08640Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Multi Teacher Privileged Knowledge Distillation for Multimodal Expression Recognition 16 Aug 2024 · 1 repository · arXiv:2408.09035
-
Optimal Symmetries in Binary Classification 16 Aug 2024 · 0 repositories · arXiv:2408.08823
-
SEAL: Systematic Error Analysis for Value ALignment 16 Aug 2024 · 1 repository · arXiv:2408.10270
-
A review of the calculation methods of optimal power flow in integrated energy systems 15 Aug 2024 · 0 repositories · arXiv:2408.08919
-
Coarse-to-fine Alignment Makes Better Speech-image Retrieval 15 Aug 2024 · 0 repositories · arXiv:2408.13119
-
Csi-LLM: A Novel Downlink Channel Prediction Method Aligned with LLM Pre-Training 15 Aug 2024 · 0 repositories · arXiv:2409.00005
-
CT4D: Consistent Text-to-4D Generation with Animatable Meshes 15 Aug 2024 · 0 repositories · arXiv:2408.08342
-
Modeling Domain and Feedback Transitions for Cross-Domain Sequential Recommendation 15 Aug 2024 · 0 repositories · arXiv:2408.08209
-
SLCA++: Unleash the Power of Sequential Fine-tuning for Continual Learning with Pre-training 15 Aug 2024 · 1 repository · arXiv:2408.08295Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Training Spatial-Frequency Visual Prompts and Probabilistic Clusters for Accurate Black-Box Transfer Learning 15 Aug 2024 · 0 repositories · arXiv:2408.07944
-
3D Gaussian Editing with A Single Image 14 Aug 2024 · 0 repositories · arXiv:2408.07540
-
Assessing the Role of Lexical Semantics in Cross-lingual Transfer through Controlled Manipulations 14 Aug 2024 · 1 repository · arXiv:2408.07599
-
Bridging and Modeling Correlations in Pairwise Data for Direct Preference Optimization 14 Aug 2024 · 1 repository · arXiv:2408.07471Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
CMU's IWSLT 2024 Simultaneous Speech Translation System 14 Aug 2024 · 0 repositories · arXiv:2408.07452
-
Interpretable Graph Neural Networks for Heterogeneous Tabular Data 14 Aug 2024 · 2 repositories · arXiv:2408.07661Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
MedTsLLM: Leveraging LLMs for Multimodal Medical Time Series Analysis 14 Aug 2024 · 1 repository · arXiv:2408.07773
-
One Step Diffusion-based Super-Resolution with Time-Aware Distillation 14 Aug 2024 · 1 repository · arXiv:2408.07476
-
Supervised and Unsupervised Alignments for Spoofing Behavioral Biometrics 14 Aug 2024 · 0 repositories · arXiv:2408.08918
-
Amuro and Char: Analyzing the Relationship between Pre-Training and Fine-Tuning of Large Language Models 13 Aug 2024 · 0 repositories · arXiv:2408.06663
-
Causal Agent based on Large Language Model 13 Aug 2024 · 1 repository · arXiv:2408.06849Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples)
-
GeoFormer: Learning Point Cloud Completion with Tri-Plane Integrated Transformer 13 Aug 2024 · 1 repository · arXiv:2408.06596
-
What should I wear to a party in a Greek taverna? Evaluation for Conversational Agents in the Fashion Domain 13 Aug 2024 · 0 repositories · arXiv:2408.08907
-
Anchored Preference Optimization and Contrastive Revisions: Addressing Underspecification in Alignment 12 Aug 2024 · 1 repository · arXiv:2408.06266
-
Audit-LLM: Multi-Agent Collaboration for Log-based Insider Threat Detection 12 Aug 2024 · 0 repositories · arXiv:2408.08902
-
MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection 12 Aug 2024 · 0 repositories · arXiv:2408.05945
-
Pattern-Matching Dynamic Memory Network for Dual-Mode Traffic Prediction 12 Aug 2024 · 1 repository · arXiv:2408.07100
-
Towards Adversarial Robustness via Debiased High-Confidence Logit Alignment 12 Aug 2024 · 0 repositories · arXiv:2408.06079
-
Decoder Pre-Training with only Text for Scene Text Recognition 11 Aug 2024 · 1 repository · arXiv:2408.05706
-
GPT-4 Emulates Average-Human Emotional Cognition from a Third-Person Perspective 11 Aug 2024 · 0 repositories · arXiv:2408.13718
-
High-fidelity and Lip-synced Talking Face Synthesis via Landmark-based Diffusion Model 10 Aug 2024 · 0 repositories · arXiv:2408.05416
-
Multimodal generative semantic communication based on latent diffusion model 10 Aug 2024 · 0 repositories · arXiv:2408.05455
-
Towards a Quantitative Analysis of Coarticulation with a Phoneme-to-Articulatory Model 10 Aug 2024 · 0 repositories · arXiv:2408.05641
-
Audio-visual cross-modality knowledge transfer for machine learning-based in-situ monitoring in laser additive manufacturing 9 Aug 2024 · 0 repositories · arXiv:2408.05307
-
Evaluating the capability of large language models to personalize science texts for diverse middle-school-age learners 9 Aug 2024 · 0 repositories · arXiv:2408.05204
-
Masked adversarial neural network for cell type deconvolution in spatial transcriptomics 9 Aug 2024 · 1 repository · arXiv:2408.05065
-
Surgical-VQLA++: Adversarial Contrastive Learning for Calibrated Robust Visual Question-Localized Answering in Robotic Surgery 9 Aug 2024 · 1 repository · arXiv:2408.04958Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Are Social Sentiments Inherent in LLMs? An Empirical Study on Extraction of Inter-demographic Sentiments 8 Aug 2024 · 0 repositories · arXiv:2408.04293
-
Deep Learning-based Unsupervised Domain Adaptation via a Unified Model for Prostate Lesion Detection Using Multisite Bi-parametric MRI Datasets 8 Aug 2024 · 0 repositories · arXiv:2408.04777
-
Mathfish: Evaluating Language Model Math Reasoning via Grounding in Educational Curricula 8 Aug 2024 · 1 repository · arXiv:2408.04226
-
Synthetic SQL Column Descriptions and Their Impact on Text-to-SQL Performance 8 Aug 2024 · 0 repositories · arXiv:2408.04691
-
Physical prior guided cooperative learning framework for joint turbulence degradation estimation and infrared video restoration 8 Aug 2024 · 0 repositories · arXiv:2408.04227
-
A broken duet: multistable dynamics of dyadic interactions 7 Aug 2024 · 1 repository · arXiv:2408.03809
-
A Comparison of LLM Finetuning Methods & Evaluation Metrics with Travel Chatbot Use Case 7 Aug 2024 · 0 repositories · arXiv:2408.03562
-
AutoFAIR : Automatic Data FAIRification via Machine Reading 7 Aug 2024 · 0 repositories · arXiv:2408.04673
-
No-Reference Image Quality Assessment with Global-Local Progressive Integration and Semantic-Aligned Quality Transfer 7 Aug 2024 · 1 repository · arXiv:2408.03885
-
Non-Causal to Causal SSL-Supported Transfer Learning: Towards a High-Performance Low-Latency Speech Vocoder 7 Aug 2024 · 0 repositories · arXiv:2408.11842
-
Patchview: LLM-Powered Worldbuilding with Generative Dust and Magnet Visualization 7 Aug 2024 · 0 repositories · arXiv:2408.04112
-
RepoMasterEval: Evaluating Code Completion via Real-World Repositories 7 Aug 2024 · 0 repositories · arXiv:2408.03519
-
Unlocking Exocentric Video-Language Data for Egocentric Video Representation Learning 7 Aug 2024 · 0 repositories · arXiv:2408.03567
-
Empathy Level Alignment via Reinforcement Learning for Empathetic Response Generation 6 Aug 2024 · 1 repository · arXiv:2408.02976
-
TextIM: Part-aware Interactive Motion Synthesis from Text 6 Aug 2024 · 0 repositories · arXiv:2408.03302
-
FastEdit: Fast Text-Guided Single-Image Editing via Semantic-Aware Diffusion Fine-Tuning 6 Aug 2024 · 0 repositories · arXiv:2408.03355
-
On the Generalization of Preference Learning with DPO 6 Aug 2024 · 0 repositories · arXiv:2408.03459
-
Probabilistic Scores of Classifiers, Calibration is not Enough 6 Aug 2024 · 1 repository · arXiv:2408.03421
-
SNFinLLM: Systematic and Nuanced Financial Domain Adaptation of Chinese Large Language Models 5 Aug 2024 · 0 repositories · arXiv:2408.02302
-
MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization 5 Aug 2024 · 1 repository · arXiv:2408.02555Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Latent-INR: A Flexible Framework for Implicit Representations of Videos with Discriminative Semantics 5 Aug 2024 · 0 repositories · arXiv:2408.02672
-
Text Conditioned Symbolic Drumbeat Generation using Latent Diffusion Models 5 Aug 2024 · 1 repository · arXiv:2408.02711
-
SiCo: An Interactive Size-Controllable Virtual Try-On Approach for Informed Decision-Making 5 Aug 2024 · 1 repository · arXiv:2408.02803
-
Progressively Label Enhancement for Large Language Model Alignment 5 Aug 2024 · 0 repositories · arXiv:2408.02599
-
Strong and weak alignment of large language models with human values 5 Aug 2024 · 1 repository · arXiv:2408.04655Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Cross-layer Attention Sharing for Large Language Models 4 Aug 2024 · 0 repositories · arXiv:2408.01890
-
Model Hijacking Attack in Federated Learning 4 Aug 2024 · 0 repositories · arXiv:2408.02131
-
A Comparative Analysis of Wealth Index Predictions in Africa between three Multi-Source Inference Models 3 Aug 2024 · 1 repository · arXiv:2408.01631
-
Transforming Slot Schema Induction with Generative Dialogue State Inference 3 Aug 2024 · 1 repository · arXiv:2408.01638
-
Piculet: Specialized Models-Guided Hallucination Decrease for MultiModal Large Language Models 2 Aug 2024 · 0 repositories · arXiv:2408.01003
-
Prototypical Partial Optimal Transport for Universal Domain Adaptation 2 Aug 2024 · 0 repositories · arXiv:2408.01089
-
A Robotics-Inspired Scanpath Model Reveals the Importance of Uncertainty and Semantic Object Cues for Gaze Guidance in Dynamic Scenes 2 Aug 2024 · 1 repository · arXiv:2408.01322
-
DebateQA: Evaluating Question Answering on Debatable Knowledge 2 Aug 2024 · 1 repository · arXiv:2408.01419Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Aligning Multiple Knowledge Graphs in a Single Pass 1 Aug 2024 · 0 repositories · arXiv:2408.00662
-
Modeling stochastic eye tracking data: A comparison of quantum generative adversarial networks and Markov models 1 Aug 2024 · 0 repositories · arXiv:2408.00673