Translates scientific PDFs while keeping the original layout intact: parses text, tables, and figures, then re-renders bilingual or monolingual output via any OpenAI-compatible LLM. Tuned for English-to-Chinese papers, with CSV glossary support.
Converts PDF, Office docs, EPUB, images, audio, HTML and ZIP archives into structured Markdown for LLM pipelines, preserving headings, tables and links instead of visual layout. Adds optional OCR, audio transcription and LLM image captions.
Extends the Wand (WeMod) desktop client’s local configuration and UI with a remote web panel, injected renderer scripts, automated compatibility patches and client-side AI features; runs entirely locally and does not publish official executables (build your own).
Custom ComfyUI nodes that run Lightricks' LTX-Video diffusion-transformer models for text-to-video and image-to-video, adding IC-LoRA control over depth, pose, edges, and motion plus distilled and low-VRAM variants for node-based workflows.
Official remote MCP servers that let AI agents read and change Cloudflare config in natural language — managing Workers and bindings, querying observability and DNS analytics, searching docs. Each capability is a separate scoped server.
Runs text-to-speech, speech-to-text, and speech-to-speech models natively on Apple Silicon via MLX — no CUDA or cloud. Supports 20+ TTS and 15+ STT models (Kokoro, Whisper, Qwen3), low-bit quantization, an OpenAI-compatible API, and a Swift package.
Expose Python functions as MCP‑compliant servers and clients so LLMs can call tools and resources directly; includes automatic schema generation, input validation, transport negotiation, authentication, and in‑conversation interactive UIs.
A self-hostable virtual companion: a VRM or Live2D character you own that voice-chats in real time, plays Minecraft and Factorio, and runs models in-browser via WebGPU or across 25+ LLM providers like Ollama, OpenAI, and Claude.
Bridges AI assistants to Jira and Confluence via the Model Context Protocol, exposing ~72 tools for JQL search, issue/page CRUD, status transitions, and comments. Supports Cloud and Server/Data Center with API-token, PAT, or OAuth 2.0 auth.
A library of specialized AI agents that automate data science steps: loading, cleaning, wrangling, feature engineering, SQL queries, EDA, and ML modeling via H2O and MLflow. Higher-level analyst workflows chain these under a supervisor agent.
Native desktop client unifying many model providers (OpenAI, Gemini, Anthropic, Ollama, local LLMs) in one app on Windows, macOS, and Linux. Adds 300+ preset assistants, document/PDF chat, MCP server integration, and WebDAV backup, with no subscription.
Captures, transcribes, and summarizes meetings entirely on the user's machine with real-time local transcription and speaker diarization. Privacy-first design keeps audio, transcripts, and models local; supports Ollama, Claude, Groq, OpenRouter or custom OpenAI-compatible endpoints.