Khoj
Open-source AI second brain with local SLMs, documents, and web search
About Khoj
Khoj is an open-source AI desktop assistant and personal second brain that allows users to search, chat, and synthesize insights across their personal notes, documents, and code repositories. Operating seamlessly with local SLMs (via Ollama) or hosted frontier models, Khoj prioritizes user privacy and offline autonomy. Whether indexing Obsidian markdown vaults, PDF research papers, Emacs org-mode files, or browser bookmarks, Khoj acts as a unified knowledge retrieval agent that answers complex multi-hop questions directly from your private workspace.
Khoj supports scheduled automated research agents that browse the web, compile daily intelligence briefings, and synthesize relevant industry updates directly to your inbox or WhatsApp. The architecture combines dense embeddings with hybrid local vector indexes, ensuring sub-second response times on standard consumer laptops. Developers can run Khoj as a native desktop client, Emacs package, Obsidian plugin, or private self-hosted Docker server with complete offline privacy.
Install Khoj via Docker, the native desktop application, or the cloud hosted dashboard.
Connect your local knowledge sources (Obsidian vault, PDF library, GitHub repos, or Notion workspace).
Select your preferred AI model: local Ollama models (Llama 3, Mistral) for privacy or cloud models for complex reasoning.
Ask questions in natural language to search across all your documents simultaneously.
Configure autonomous research agents to monitor web topics and generate recurring synthesized summaries.
Capabilities & Features
Common Use Cases
personal-knowledge-management
offline-ai-chat
document-search
automated-research
obsidian-ai
Frequently Asked Questions
Can Khoj run completely offline?
Yes, Khoj can run 100% offline on your local computer when paired with Ollama or local small language models without transmitting any data over the internet.
Does Khoj integrate with Obsidian?
Yes, Khoj provides an official Obsidian community plugin that allows you to chat with your markdown vault directly inside Obsidian.
Is Khoj free?
Yes, the self-hosted open-source version of Khoj is completely free with no usage limits. A hosted cloud subscription is available for $8/month.
Free Plan
100% free and open-source self-hosted version with unlimited local document indexing
Paid Plan
Cloud hosted plan at $8/mo with hosted GPT-4o, Claude 3.5, and automated web research agents
Direct link · Verified & reader-supported
Pros & Cons
100% open source with complete local offline privacy and zero telemetry
Native plugins for Obsidian, Emacs, and desktop operating systems
Supports local models via Ollama as well as hosted frontier LLMs
Autonomous recurring web research agents deliver briefings to email and chat
Blazing fast hybrid semantic retrieval across private document collections
Self-hosting local LLMs requires modern computer hardware with sufficient RAM and VRAM
Cloud plan required for users who do not want to manage local Docker containers
Alternatives
View allNotebookLM
AI-powered research assistant
NotebookLM is Google's AI-powered research and note-taking tool that lets you upload your own documents and interact with them using natural language. It helps summarize sources, generate insights, and answer questions based on your uploaded content.
Mem0
The universal persistent memory layer for AI agents & LLM apps
Mem0 (formerly Embedchain) is a universal, persistent memory architecture designed to solve the critical context amnesia problem in modern AI applications. While foundational LLMs forget user preferences and past interactions the moment a session ends, Mem0 maintains a continuous, self-improving memory graph across user sessions, agents, and applications. With Mem0, developers can build personalized AI assistants, customer support agents, and autonomous workflow bots that remember user preferences, past project decisions, and communication styles over months and years. Mem0 operates as both an open-source self-hostable Python/TypeScript library and a managed cloud platform, providing sub-100ms vector search, episodic memory extraction, and automated memory consolidation without manual prompt engineering.
RunPod
Globally distributed GPU cloud and serverless platform for AI inference and training
RunPod is a leading globally distributed GPU cloud and serverless computing platform engineered specifically for artificial intelligence workloads. It provides developers, AI researchers, and enterprises with on-demand access to top-tier NVIDIA GPUs (including H100, A100, L40S, and RTX 4090) at up to 80% lower cost than traditional legacy hyperscalers.
Qdrant
High-performance vector database and similarity search engine for AI
Qdrant is an open-source, high-performance vector database and similarity search engine engineered in Rust for production AI systems, semantic search engines, and Retrieval-Augmented Generation (RAG) pipelines. It provides lightning-fast nearest-neighbor search with rich payload filtering and custom distance metrics. Unlike traditional databases adapted for vectors, Qdrant was designed from day one to handle high-dimensional neural embeddings at scale. Its Rust engine provides memory-efficient vector quantization (scalar, product, and binary), allowing engineering teams to search billions of vectors on cost-effective cloud hardware.
LlamaIndex
Leading data framework for connecting custom data sources to LLMs and Agentic RAG workflows.
LlamaIndex is the premier open-source data framework designed to bridge private, enterprise, and unstructured data with large language models. By providing sophisticated data connectors, automated parser modules, semantic chunking algorithms, and multi-document index structures, LlamaIndex enables developers to build context-augmented LLM applications and autonomous knowledge retrieval engines with minimal boilerplate. From parsing complex multi-page PDF documents and financial spreadsheets to orchestrating complex Agentic RAG workflows that query multiple disparate databases, LlamaIndex handles the complete data ingestion, indexing, and query evaluation lifecycle.
Stanford STORM
Autonomous deep research system synthesizing multi-perspective, citation-backed Wikipedia-style reports
Stanford STORM (Synthesis of Topic Outlines through Retrieval and Multi-perspective Question Asking) is an open-source AI research system developed by Stanford University. It autonomously conducts deep web investigations, generates diverse stakeholder interview perspectives, and synthesizes exhaustive, citation-backed long-form reports. Unlike standard search engines that produce single-paragraph summaries, STORM simulates a collaborative expert panel to research complex topics thoroughly and assemble referenced research dossiers.
Compare Khoj with Alternatives
Side-by-side feature, pricing, and pros & cons breakdowns
Featured in In-Depth Guides & Comparisons
Read our hands-on technical evaluations and workflow guides mentioning Khoj

