Skip to content
Papers.

Jina CLIP: Your CLIP Model Is Also Your Text Retriever research paper by Jina AI, 2024

Jina AI · May 30, 2024 · Alignment and safety · 61 citations · 37 upvotes · unverified

Read on arXiv

What it shows

A novel multi-task contrastive training method improves CLIP model performance on both text-image and text-text retrieval tasks.

UnverifiedHugging Face's summary; not yet checked by hand.

More from Jina AI

All 3
Topic
About this paper
Authors
Andreas Koukounas, Georgios Mastrapas, Michael Günther and 11 more
arXiv
2405.20204 · PDF
Venue
arXiv.org
Citations
61, 10 influential · Semantic Scholar
Upvotes
37 · Hugging Face
Lab
Jina AI · on Companies · on Acquisitions

Changes

What changed
Influential citationsfirst count: 10Sep 25, 2026
Citationsfirst count: 61Sep 25, 2026
New paperFound by the weekly scan, unverifiedSep 25, 2026

Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.

New papers by email

Monday afternoons, only in weeks with new papers from the labs.

Double opt-in. Unsubscribe any time.