rom1504/clip-retrieval
Easily compute clip embeddings and build a clip retrieval system with them
[NAACL 2022]Mobile Text-to-Image search powered by multimodal semantic representation models(e.g., OpenAI's CLIP)
$ git clone https://github.com/DRSY/MoTIS.gitEasily compute clip embeddings and build a clip retrieval system with them
Pocket-Sized Multimodal AI for content understanding and generation across multilingual texts, images, and ๐ video, up to 5x faster than OpenAI CLIP and LLaVA ๐ผ๏ธ & ๐๏ธ
Offline semantic Text-to-Image and Image-to-Image search on Android powered by quantized state-of-the-art vision-language pretrained CLIP model and ONNX Runtime inference engine
World's fastest and most compact embedded vector database: exact by default, multimodal, local-first, and GPU-accelerated
AI-powered media search โ find images and videos using natural language or visual queries
21 Lessons, Get Started Building with Generative AI
Data from GitHub ยท snapshot Sep 24, 2026