All repositories

RAG & vector DBsMintplex-Labs

anything-llm

A ready-made «chat with your documents» app: plug in any model and vector database, upload files and talk to them. Has a desktop version, Docker and agents.

Who it is for

For people who want a local ChatGPT over their own files without coding.

How to start

  1. Open the Self-Hosting section of the README and pick a method: Docker, cloud or desktop.
  2. Launch the app and connect a model and a vector database.
  3. Upload documents to a workspace and ask questions about them.

Steps are taken from the README. Check the current version in the repository before running them.

Stars over the last 30 days

+1,385Sep 6 — Oct 5
65,33466,719

Author's description

Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience

microsoft

markitdown

Growing fastest

A small Microsoft utility that converts PDF, Word, PowerPoint, Excel and other files into Markdown that is easy to hand to a language model.

189Kstars+11K in 30 dPython

firecrawl

firecrawl

Growing fastest

A service that turns any website into clean Markdown or structured data for models. It can search, scrape, crawl a whole site and even click through a page.

189Kstars+12K in 30 dTypeScript

infiniflow

ragflow

A ready-made RAG engine with a UI: it parses documents with templates, chunks them, searches with citations and supports agentic retrieval. Deploys with Docker.

92Kstars+1.8K in 30 dGo

unclecode

crawl4ai

An open-source Python crawler: it visits pages with a real browser and returns clean Markdown for LLMs. Runs locally, in Docker or as a cloud service.

85Kstars+3.2K in 30 dPython

opendatalab

MinerU

Parses complex PDFs, scans and Office files into Markdown or JSON while keeping tables and formulas. Comes with a CLI, a Python SDK and an agent skill.

81Kstars+2K in 30 dPython

docling-project

docling

An IBM library for document parsing: PDF, Office, HTML and images become one unified structure from which Markdown is easy to get. Supports vision models for hard pages.

68Kstars+2.5K in 30 dPython