Runs local LLM, vision-language, ASR, OCR, and image-generation models across NPU, GPU, and CPU from one command. Differs from Ollama and llama.cpp with first-class Qualcomm Hexagon NPU support and day-0 coverage of new models like Qwen3-VL.
Runs a native, extensible AI agent on desktop, CLI, or API to automate code, workflows, research, and writing. Built in Rust, supports 15+ LLM providers and 70+ extensions via the Model Context Protocol — designed for local-first automation and developer workflows.
Desktop finance analytics terminal that combines CFA-level models, real-time trading and 100+ data connectors with embedded Python for analytics; includes 37 AI agents and local/multi-provider LLM support for automated research and decision workflows.
Implements the Model Context Protocol in TypeScript, providing server and client libraries to expose tools, resources, and prompts to LLM hosts. Ships Streamable HTTP and stdio transports, optional middleware for Express/Fastify/Hono, and runnable examples for Node/Bun/Deno.
Runs AI models on user devices with native SDKs, optimized model management, hardware acceleration, and OpenAI-compatible APIs for apps that need offline, private inference.
Self-hostable alternative to Google NotebookLM: organize PDFs, videos, audio, web pages, and Office docs, then chat over them, take AI-assisted notes, and search via full-text and vector. Routes to 18+ model providers and generates 1-4 speaker podcasts.
Extends the Wand (WeMod) desktop client’s local configuration and UI with a remote web panel, injected renderer scripts, automated compatibility patches and client-side AI features; runs entirely locally and does not publish official executables (build your own).
Brings an agentic chat experience to the terminal: describe a task in natural language and it plans, edits files, and runs commands to build the app. Written in Rust, ships on macOS and Linux. Now succeeded by the closed-source Kiro CLI.
Expose Python functions as MCP‑compliant servers and clients so LLMs can call tools and resources directly; includes automatic schema generation, input validation, transport negotiation, authentication, and in‑conversation interactive UIs.
A self-hostable virtual companion: a VRM or Live2D character you own that voice-chats in real time, plays Minecraft and Factorio, and runs models in-browser via WebGPU or across 25+ LLM providers like Ollama, OpenAI, and Claude.
Generates structured, streaming UIs from LLM output and renders them in React using a compact OpenUI Lang, built component libraries, and chat surfaces; claims up to ~67% token savings vs JSON and includes a playground and CLI.
Runs iterative, fully-local web research loops using locally hosted LLMs (via Ollama or LMStudio): it auto-generates search queries, gathers and summarizes results, reflects to find gaps, re-queries, and emits a final markdown report with sources.