A local AI search engine for your Obsidian vault: a private, self-hosted NotebookLM alternative with chat, one-click local models, and web crawling

lilbee for Obsidian v0.6.66b
In the community plugin store Runs offline 100% coverage MIT

Obsidian's interface to lilbee, a local AI search engine: find, run, and manage local AI models, and search your vault and the sites you crawl with them.

Browse a model catalog and pull a model that runs on your own machine, then ask questions about anything in your vault (notes, PDFs, ebooks, code, scans, 150+ file types in all) and get answers with the source one click away.

one plugin, inside obsidian chat sidebar cited answers source preview model catalog memory auto wiki

Everything runs on your computer. Your notes stay in your vault; nothing leaves unless you ask it to. Cloud models are optional, and only when you choose one.

Heads up: it downloads to your computer. The lilbee server (a few hundred MB) on first launch, plus whatever models you pick from the catalog (a few hundred MB up to several GB each). It all stays local.

see it

The plugin's own source code added to the library on an Apple M1 Pro: the Task Center embeds all 58 files beside the live GPU placement view, then what is lilbee for Obsidian? gets a cited answer and the citation opens the README at the source.

install
1lilbee is in the official Obsidian community plugin store. In Obsidian, go to Settings → Community plugins → Browse, search for lilbee, then Install and Enable. The link above jumps straight to the plugin page in Obsidian.
2A setup wizard opens: pick a chat model and run the first sync. That's it; lilbee brings its own engine, nothing else to install.

lilbee is in active development, with frequent releases; updates arrive through Obsidian's plugin updater. view the store listing →

Coming from Google's NotebookLM? lilbee does the same job on your own hardware. see how it compares →

It's early days. If it's useful to you, a ★ on GitHub helps other Obsidian users find it, and bug reports are very welcome.

the difference
without it
  • a chatbot in a browser tab doesn't know what's in your vault
  • Obsidian's search finds the word, not the answer
  • when something matters, you open the document and read it yourself
  • your notes and your chatbot live in separate apps
a chatbot here, a search box there, you in the middle.
with lilbee
  • ask in plain English; get a real answer
  • every answer points to the exact line it came from; one click to see it
  • it works offline, on your computer
  • it's right there in Obsidian, next to your notes
one plugin. pick a model, ask.
what it does

ask your vault anything

Type a question in plain English; lilbee reads your notes and files and answers it. It's like having your own private Encarta, built from everything you've collected.

every answer shows its work

Each reply comes with footnotes. Click one and the exact spot opens, right where it came from, so you can trust the answer or check it for yourself.

it reads more than notes

PDFs, ebooks, spreadsheets, code, even scanned pages and photos. Over 150 file types in all. If it's in your vault, lilbee can search it.

pick a model, no account needed

Browse a built-in model catalog, straight from Hugging Face Hub: featured picks up top, or search the full list. Download one with a click; it runs on your computer. 50+ settings to tinker with, good defaults if you'd rather not.

private by default

Your files, the models, the search: it all stays on your machine and works offline. Want a cloud model for one job? Plug one in, and the plugin tells you whenever it's in use.

remembers what you tell it

Turn on memory and lilbee holds onto durable facts about you and how you like your answers, then recalls the relevant ones in later chats, no matter which conversation they came from. Off by default, managed from a Memories view, and never mixed into your citations.

writes a wiki of your knowledge experimental

lilbee can draft linked wiki pages from your library; they land in your vault as ordinary notes and show up in the graph view, alongside your own.

across your GPUs

one box, many GPUs, no flags

Got more than one graphics card? lilbee finds them all. The placement view draws every GPU and what's running on it, chat, embedding, vision, reranking, and lets you spread a model across cards or pin a worker to one with a click.

Spread one model's layers across several cards, or pin embedding, vision and reranking workers to the GPUs you choose, then apply and the fleet rebuilds to match. The usage bars update live, so you can see each card's load in real time. On a single machine it stays simple: one card, everything together, nothing to configure.

go deeper
questions
Do I need Ollama or LM Studio to use lilbee in Obsidian?

No. lilbee downloads and runs the AI models for you; it is a complete model manager, so there is no separate runner to set up. If you already use Ollama or LM Studio, you can point lilbee at them instead.

Does my vault leave my computer?

No. Indexing and search run on your own machine, and your notes stay on disk. lilbee uses a cloud model only if you pick one.

What can it search in my vault?

Your notes and markdown, plus PDFs, code, ebooks, and scanned images through OCR, and whole websites you crawl into the vault. Over 150 file types, with answers that cite the exact source.

Does it work offline?

Yes. With local models in place, lilbee searches, asks, and chats with your vault with no internet connection.

How do I install it?

From the Obsidian community plugin store: open Settings, then Community plugins, search for lilbee, and install. lilbee sets up the models for you on first run.

obsidian-lilbee  .  MIT License