SentenceTransformers provides models that allow to embed images and text into the same vector space.
202
SentenceTransformers provides models that allow to embed images and text into the same vector space.
This allows to find similar images as well as to implement image search.
This is the Image & Text model CLIP, which maps text and images to a shared vector space
Content type
Image
Digest
sha256:0e91bffa3…
Size
883 MB
Last updated
almost 3 years ago
docker pull gibbo96/text2image