Methods › Computer Vision › Vision and Language Pre-Trained Models › CLIP › Papers, page 17
Contrastive Language-Image Pre-training
CLIP
Papers archive 2025-07-28
archive papers tagged: 3,094 · with a code link: 1,617 · where Syntology ran a sample: 649 (554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (649 of 3,094 tagged: 554 with a run with no instrument failure, 95 where every run was a failure of Syntology's instrument)
Page 17 of 31: papers 1,601 to 1,700 of 3,094, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models 4 Mar 2024 · 1 repository · arXiv:2403.01849Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Rethinking CLIP-based Video Learners in Cross-Domain Open-Vocabulary Action Recognition 3 Mar 2024 · 1 repository · arXiv:2403.01560
-
Data-free Multi-label Image Recognition via LLM-powered Prompt Tuning 2 Mar 2024 · 0 repositories · arXiv:2403.01209
-
Abductive Ego-View Accident Video Understanding for Safe Driving Perception 1 Mar 2024 · 0 repositories · arXiv:2403.00436
-
G3DR: Generative 3D Reconstruction in ImageNet 1 Mar 2024 · 1 repository · arXiv:2403.00939Syntology official (archive's flag): 8 ran · 8 ran (of which 8 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; every one of the 8 samples that ran constructed an object rather than computing a result (of 13 harvested samples) · 13 pointer-only (licence)
-
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model 1 Mar 2024 · 1 repository · arXiv:2403.00376
-
Multi-modal Attribute Prompting for Vision-Language Models 1 Mar 2024 · 0 repositories · arXiv:2403.00219
-
TAMM: TriAdapter Multi-Modal Learning for 3D Shape Understanding 28 Feb 2024 · 0 repositories · arXiv:2402.18490
-
Measuring Vision-Language STEM Skills of Neural Models 27 Feb 2024 · 1 repository · arXiv:2402.17205Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Impression-CLIP: Contrastive Shape-Impression Embedding for Fonts 26 Feb 2024 · 1 repository · arXiv:2402.16350
-
Infrared and visible Image Fusion with Language-driven Loss in CLIP Embedding Space 26 Feb 2024 · 1 repository · arXiv:2402.16267
-
MIP: CLIP-based Image Reconstruction from PEFT Gradients 26 Feb 2024 · 0 repositories · arXiv:2403.07901
-
Fine-tuning CLIP Text Encoders with Two-step Paraphrasing 23 Feb 2024 · 0 repositories · arXiv:2402.15120Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Grasp, See, and Place: Efficient Unknown Object Rearrangement with Policy Structure Prior 23 Feb 2024 · 1 repository · arXiv:2402.15402
-
Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding 23 Feb 2024 · 2 repositories · arXiv:2402.15300Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Balanced Data Sampling for Language Model Training with Clustering 22 Feb 2024 · 1 repository · arXiv:2402.14526Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
CLoVe: Encoding Compositional Language in Contrastive Vision-Language Models 22 Feb 2024 · 1 repository · arXiv:2402.15021
-
Exploring and Applying Audio-Based Sentiment Analysis in Music 22 Feb 2024 · 1 repository · arXiv:2403.17379
-
Generalizable Semantic Vision Query Generation for Zero-shot Panoptic and Semantic Segmentation 21 Feb 2024 · 0 repositories · arXiv:2402.13697
-
On Large Visual Language Models for Medical Imaging Analysis: An Empirical Study 21 Feb 2024 · 0 repositories · arXiv:2402.14162
-
Analysis of Using Sigmoid Loss for Contrastive Learning 20 Feb 2024 · 0 repositories · arXiv:2402.12613
-
CLIPping the Deception: Adapting Vision-Language Models for Universal Deepfake Detection 20 Feb 2024 · 1 repository · arXiv:2402.12927Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
CounterCurate: Enhancing Physical and Semantic Visio-Linguistic Compositional Reasoning via Counterfactual Examples 20 Feb 2024 · 1 repository · arXiv:2402.13254Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Slot-VLM: SlowFast Slots for Video-Language Modeling 20 Feb 2024 · 0 repositories · arXiv:2402.13088
-
Learning the Unlearned: Mitigating Feature Suppression in Contrastive Learning 19 Feb 2024 · 1 repository · arXiv:2402.11816
-
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models 19 Feb 2024 · 1 repository · arXiv:2402.12336Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Assessing News Thumbnail Representativeness: Counterfactual text can enhance the cross-modal matching ability 17 Feb 2024 · 1 repository · arXiv:2402.11159
-
ZeroG: Investigating Cross-dataset Zero-shot Transferability in Graphs 17 Feb 2024 · 1 repository · arXiv:2402.11235Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Interpreting CLIP with Sparse Linear Concept Embeddings (SpLiCE) 16 Feb 2024 · 1 repository · arXiv:2402.10376Syntology official (archive's flag): 7 ran · 7 ran (of which 1 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Any-Shift Prompting for Generalization over Distributions 15 Feb 2024 · 0 repositories · arXiv:2402.10099
-
Mind the Modality Gap: Towards a Remote Sensing Vision-Language Model via Cross-modal Alignment 15 Feb 2024 · 0 repositories · arXiv:2402.09816
-
CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding 14 Feb 2024 · 1 repository · arXiv:2402.08994Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Open-Vocabulary Segmentation with Unpaired Mask-Text Supervision 14 Feb 2024 · 2 repositories · arXiv:2402.08960Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Quantified Task Misalignment to Inform PEFT: An Exploration of Domain Generalization and Catastrophic Forgetting in CLIP 14 Feb 2024 · 0 repositories · arXiv:2402.09613
-
Captions Are Worth a Thousand Words: Enhancing Product Retrieval with Pretrained Image-to-Text Models 13 Feb 2024 · 0 repositories · arXiv:2402.08532
-
A Closer Look at the Robustness of Contrastive Language-Image Pre-Training (CLIP) 12 Feb 2024 · 0 repositories · arXiv:2402.07410
-
SpeechCLIP+: Self-supervised multi-task representation learning for speech via CLIP and speech-image data 10 Feb 2024 · 1 repository · arXiv:2402.06959
-
Beyond DAGs: A Latent Partial Causal Model for Multimodal Learning 9 Feb 2024 · 0 repositories · arXiv:2402.06223
-
ColorSwap: A Color and Word Order Dataset for Multimodal Evaluation 7 Feb 2024 · 1 repository · arXiv:2402.04492Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
λ-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent Space 7 Feb 2024 · 1 repository · arXiv:2402.05195
-
OV-NeRF: Open-vocabulary Neural Radiance Fields with Vision and Language Foundation Models for 3D Semantic Understanding 7 Feb 2024 · 1 repository · arXiv:2402.04648Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
A Hard-to-Beat Baseline for Training-free CLIP-based Adaptation 6 Feb 2024 · 1 repository · arXiv:2402.04087Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters 6 Feb 2024 · 2 repositories · arXiv:2402.04252Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples) · 4 pointer-only (licence)
-
CLIP Can Understand Depth 5 Feb 2024 · 0 repositories · arXiv:2402.03251
-
Enhancing Compositional Generalization via Compositional Feature Alignment 5 Feb 2024 · 1 repository · arXiv:2402.02851Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
FROSTER: Frozen CLIP Is A Strong Teacher for Open-Vocabulary Action Recognition 5 Feb 2024 · 1 repository · arXiv:2402.03241
-
AI Art Neural Constellation: Revealing the Collective and Contrastive State of AI-Generated and Human Art 4 Feb 2024 · 1 repository · arXiv:2402.02453
-
Generalizable Entity Grounding via Assistance of Large Language Model 4 Feb 2024 · 0 repositories · arXiv:2402.02555
-
Video Editing for Video Retrieval 4 Feb 2024 · 0 repositories · arXiv:2402.02335
-
Variance Alignment Score: A Simple But Tough-to-Beat Data Selection Method for Multimodal Contrastive Learning 3 Feb 2024 · 2 repositories · arXiv:2402.02055Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 3 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Probabilistic Model Behind Self-Supervised Learning 2 Feb 2024 · 1 repository · arXiv:2402.01399Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Can Shape-Infused Joint Embeddings Improve Image-Conditioned 3D Diffusion? 2 Feb 2024 · 0 repositories · arXiv:2402.01241
-
ConRF: Zero-shot Stylization of 3D Scenes with Conditioned Radiation Fields 2 Feb 2024 · 1 repository · arXiv:2402.01950
-
Cross-modality debiasing: using language to mitigate sub-population shifts in imaging 2 Feb 2024 · 0 repositories · arXiv:2403.07888
-
SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training? 2 Feb 2024 · 1 repository · arXiv:2402.01832Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
GaussianStyle: Gaussian Head Avatar via StyleGAN 1 Feb 2024 · 1 repository · arXiv:2402.00827
-
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks 1 Feb 2024 · 1 repository · arXiv:2402.00626Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
M2-RAAP: A Multi-Modal Recipe for Advancing Adaptation-based Pre-training towards Effective and Efficient Zero-shot Video-text Retrieval 31 Jan 2024 · 1 repository · arXiv:2401.17797
-
Embracing Language Inclusivity and Diversity in CLIP through Continual Language Learning 30 Jan 2024 · 1 repository · arXiv:2401.17186Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
M2-Encoder: Advancing Bilingual Image-Text Understanding by Large-scale Efficient Pretraining 29 Jan 2024 · 1 repository · arXiv:2401.15896
-
Cross-Modal Coordination Across a Diverse Set of Input Modalities 29 Jan 2024 · 0 repositories · arXiv:2401.16347
-
NFT1000: A Cross-Modal Dataset for Non-Fungible Token Retrieval 29 Jan 2024 · 1 repository · arXiv:2402.16872
-
Data-Free Generalized Zero-Shot Learning 28 Jan 2024 · 1 repository · arXiv:2401.15657Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
FreeStyle: Free Lunch for Text-guided Style Transfer using Diffusion Models 28 Jan 2024 · 1 repository · arXiv:2401.15636Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Taiyi-Diffusion-XL: Advancing Bilingual Text-to-Image Generation with Large Vision-Language Model Support 26 Jan 2024 · 1 repository · arXiv:2401.14688
-
Image Synthesis with Graph Conditioning: CLIP-Guided Diffusion Models for Scene Graphs 25 Jan 2024 · 0 repositories · arXiv:2401.14111
-
UrbanGenAI: Reconstructing Urban Landscapes using Panoptic Segmentation and Diffusion Models 25 Jan 2024 · 0 repositories · arXiv:2401.14379
-
Enhancing Image Retrieval : A Comprehensive Study on Photo Search using the CLIP Mode 24 Jan 2024 · 0 repositories · arXiv:2401.13613
-
SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval 24 Jan 2024 · 1 repository · arXiv:2401.13478Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
ClipSAM: CLIP and SAM Collaboration for Zero-Shot Anomaly Segmentation 23 Jan 2024 · 1 repository · arXiv:2401.12665
-
On the Efficacy of Text-Based Input Modalities for Action Anticipation 23 Jan 2024 · 0 repositories · arXiv:2401.12972
-
RAW: A Robust and Agile Plug-and-Play Watermark Framework for AI-Generated Images with Provable Guarantees 23 Jan 2024 · 0 repositories · arXiv:2403.18774
-
The Neglected Tails in Vision-Language Models 23 Jan 2024 · 0 repositories · arXiv:2401.12425
-
UniHDA: A Unified and Versatile Framework for Multi-Modal Hybrid Domain Adaptation 23 Jan 2024 · 0 repositories · arXiv:2401.12596
-
M2-CLIP: A Multimodal, Multi-task Adapting Framework for Video Action Recognition 22 Jan 2024 · 0 repositories · arXiv:2401.11649
-
Semantic Prompt Learning for Weakly-Supervised Semantic Segmentation 22 Jan 2024 · 1 repository · arXiv:2401.11791
-
Zoom-shot: Fast and Efficient Unsupervised Zero-Shot Transfer of CLIP to Vision Encoders with Multimodal Loss 22 Jan 2024 · 0 repositories · arXiv:2401.11633
-
A Novel Benchmark for Few-Shot Semantic Segmentation in the Era of Foundation Models 20 Jan 2024 · 1 repository · arXiv:2401.11311
-
Prompting Large Vision-Language Models for Compositional Reasoning 20 Jan 2024 · 1 repository · arXiv:2401.11337Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
DGL: Dynamic Global-Local Prompt Tuning for Text-Video Retrieval 19 Jan 2024 · 2 repositories · arXiv:2401.10588Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
On mitigating stability-plasticity dilemma in CLIP-guided image morphing via geodesic distillation loss 19 Jan 2024 · 1 repository · arXiv:2401.10526
-
Weakly Supervised Gaussian Contrastive Grounding with Large Multimodal Models for Video Question Answering 19 Jan 2024 · 1 repository · arXiv:2401.10711Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
CLIP Model for Images to Textual Prompts Based on Top-k Neighbors 18 Jan 2024 · 0 repositories · arXiv:2401.09763
-
CPCL: Cross-Modal Prototypical Contrastive Learning for Weakly Supervised Text-based Person Re-Identification 18 Jan 2024 · 1 repository · arXiv:2401.10011
-
Supervised Fine-tuning in turn Improves Visual Foundation Models 18 Jan 2024 · 1 repository · arXiv:2401.10222Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
Temporal Insight Enhancement: Mitigating Temporal Hallucination in Multimodal Large Language Models 18 Jan 2024 · 0 repositories · arXiv:2401.09861
-
Concept-Guided Prompt Learning for Generalization in Vision-Language Models 15 Jan 2024 · 0 repositories · arXiv:2401.07457
-
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding 15 Jan 2024 · 0 repositories · arXiv:2401.07572
-
FiGCLIP: Fine-Grained CLIP Adaptation via Densely Annotated Videos 15 Jan 2024 · 0 repositories · arXiv:2401.07669
-
Image Similarity using An Ensemble of Context-Sensitive Models 15 Jan 2024 · 1 repository · arXiv:2401.07951
-
Towards A Better Metric for Text-to-Video Generation 15 Jan 2024 · 0 repositories · arXiv:2401.07781
-
Domain Adaptation for Large-Vocabulary Object Detectors 13 Jan 2024 · 0 repositories · arXiv:2401.06969
-
APLe: Token-Wise Adaptive for Multi-Modal Prompt Learning 12 Jan 2024 · 0 repositories · arXiv:2401.06827
-
Application Of Vision-Language Models For Assessing Osteoarthritis Disease Severity 12 Jan 2024 · 0 repositories · arXiv:2401.06331
-
Synthetic Data Generation Framework, Dataset, and Efficient Deep Model for Pedestrian Intention Prediction 12 Jan 2024 · 1 repository · arXiv:2401.06757
-
UMG-CLIP: A Unified Multi-Granularity Vision Generalist for Open-World Understanding 12 Jan 2024 · 1 repository · arXiv:2401.06397
-
CLIP-Driven Semantic Discovery Network for Visible-Infrared Person Re-Identification 11 Jan 2024 · 1 repository · arXiv:2401.05806
-
Cross-modal Retrieval for Knowledge-based Visual Question Answering 11 Jan 2024 · 1 repository · arXiv:2401.05736Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Exploring Self- and Cross-Triplet Correlations for Human-Object Interaction Detection 11 Jan 2024 · 0 repositories · arXiv:2401.05676
-
Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs 11 Jan 2024 · 1 repository · arXiv:2401.06209Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)