1.1.3 — Local model options and reliability fixes

This patch release expands the local model catalog and improves how VaultRAG displays file metadata, model limits, currency ranges, and Windows file actions. Highlights:

  • Vinci Piccolo local model. Vinci Piccolo 1.0 Q4_K_M is now available in the managed local model catalog for chat and document summaries.
  • Clearer local model limits. Context-window sizes and output limits are now reported separately, avoiding misleading maximum output values for long-context models.
  • Reliable currency rendering. Prices and ranges such as $499-$699, $8K-$15K, and $1.5M-$3M remain readable text in chat paragraphs, lists, and tables instead of being mistaken for math.
  • Native Windows Open With. Windows now uses the system Open With chooser, including for files that already have a default app.

1.1.2 — Chat tabs, pop-out windows, and local AI polish

This release adds multi-conversation chat with tabs and pop-out windows, better conversation history organization, and a round of local AI and chat rendering fixes. Highlights:

  • Chat tabs. Run multiple conversations at once, each with its own model and live streaming state. Open a new tab with Cmd+T / Ctrl+T — the tab strip stays hidden until you have more than one conversation.
  • Pop-out chat windows. Pop any conversation out into its own window for side-by-side use, styled to match the main window. Closing a pop-out re-docks it as a tab.
  • Better conversation history. Pin important conversations, search your history, and browse it grouped by date. Deleting a conversation that is open now asks for confirmation and closes its tab safely.
  • Smarter local AI errors and stats. Request timeouts now come with guidance to raise the timeout instead of a generic connection error, and chat stats report real token usage from local models instead of estimates.
  • Vision-aware attachments. Image attachments are now blocked up front for local models that can't see images, instead of being silently dropped.
  • Smoother streaming. Answers stream in live again when web search and other tools are enabled, without stray text flashing before a tool call starts.

1.1.1 — PDF image parsing fix

This patch release fixes PDF parsing for documents that include images, improving indexing and chat context for image-heavy PDFs.

1.1.0 — Local AI and more

The biggest update since launch. VaultRAG can now run entirely on your own computer, chat responses stream live, and the file manager, themes, and onboarding all got a lot more polished. Highlights:

  • Local AI. Run chat, summaries, and embeddings entirely on your own machine — no API key and nothing leaves your computer. Install models from a built-in catalog that flags which ones fit your hardware, with GPU acceleration (including CUDA on Windows) and a one-click setup that picks sensible defaults for you.
  • Live streaming responses. Answers now stream in as they are generated, for both cloud and local models.
  • Web search. Use OpenRouter's built-in web search or your own Tavily key to bring live results into chat.
  • Choose what gets embedded. Index documents using an AI summary for the best search quality, or embed document text directly for faster indexing — especially with local models.
  • Conversation history and search. Browse past conversations and search within a chat, with matches highlighted inline.
  • Richer chat. Math and LaTeX rendering, syntax-highlighted code blocks, inline images, and improved vision support for images, PDFs, and email attachments.
  • More themes and appearance options. Dozens of new light and dark themes, a font picker, a compact minimal mode, and an option to center the window when it opens.
  • File manager improvements. Progress indicators for large copy, move, and delete operations, Finder-style icons, a customizable vault toolbar, file previews, and a native macOS share sheet.
  • Smoother onboarding. A guided setup that lets you choose cloud, local, or a mix, with one-click local AI setup built in.

1.0.0 — Initial release

The first public build of VaultRAG. Highlights of this release:

  • Local vault. Point VaultRAG at any folder on your machine and it builds a searchable index of your documents in place.
  • Embedding-based search. Find documents by what they actually contain, with optional filtering by file or folder name.
  • Automatic summaries. Every document is summarized on import. Summaries can be regenerated, edited by hand, or cleared at any time.
  • Chat with your documents. Ask questions against your vault and get answers grounded in the documents you have indexed.
  • Bring your own model. Connect to any provider through OpenRouter using your own API key.
  • Customizable system prompt. Tune the model's tone, focus, and behavior to match how you work.
  • One-time license. Pay once for the app, activate on up to two devices, and pay your AI provider directly for tokens.

Have feedback or want to report an issue? Reach out at [email protected].