Skip to content
#

image-description

Here are 25 public repositories matching this topic...

A new package that processes user-submitted text descriptions of images or videos containing watermarks and returns structured, watermark-free descriptions. It uses an LLM to reinterpret the content w

  • Updated Dec 21, 2025
  • Python

给 DeepSeek Harness 纯文本模型装上原生视觉(Windows):粘贴即看图——预注入描述,模型首轮就看见,不用选模型、不用调工具;see_image 精查;自定义视觉后端(任意 OpenAI 兼容模型)+ 四后端容灾;换主模型视觉自动跟随。| Give text-only DeepSeek Harness models native-feeling vision on Windows: paste and the model just sees it — pre-injected descriptions, see_image tool, custom backends, 4-backend failover.

  • Updated Aug 15, 2026
  • JavaScript

📊 Multi-Modal RAG 2.0 — An enterprise-grade RAG system that processes PDFs, tables, charts/images, and audio files together. Uses Unstructured.io, Camelot, Gemini Vision, Whisper, ChromaDB, FastAPI, and Streamlit. Increases information retrieval accuracy by 35% over text-only RAG.

  • Updated Aug 26, 2026
  • Python

It is an innovative repository housing a sophisticated Large Language Model (LLM) project, showcasing the intersection of advanced natural language processing and cutting-edge artificial intelligence. This repository serves as a comprehensive platform for the development, experimentation, and application of state-of-the-art language models.

  • Updated Jun 13, 2024
  • Python

Add this topic to your repo

To associate your repository with the image-description topic, visit your repo's landing page and select "manage topics."

Learn more