AnythingLLM lets you point a language model at your own documents and query them, running either as a desktop application or a self-hosted server. It handles document ingestion, chunking and vector storage, and works with local models through Ollama as well as hosted APIs. Workspaces keep different document sets separate. It is a practical option for private retrieval where sending documents to a third-party service is not acceptable.
Pricing: Free