Drop-in Jinja chat templates for Qwen 3.5/3.6 that fix rendering errors, token waste, and tool-calling failures across runtimes (LM Studio, llama.cpp, vLLM, MLX). Adds a think-on/think-off toggle, auto-closes broken thinking tags, robust tool-argument handling, and a graceful fallback for missing user queries.
Self-hosted sales CRM that runs native AI agents (RAG per tenant) to handle WhatsApp conversations, qualify leads, trigger automations and move deals through configurable pipelines. Multi-tenant with LGPD-minded controls and a one-command VPS installer for full data ownership.
Provides a drop-in Jinja chat template for Qwen 3.5/3.6/3.8 that reduces reasoning-token waste, enforces a concise terseness system prompt, and preserves in-chat reasoning and tool-call rendering across turns. Terseness is on by default but switchable per request; no model weights are changed.
Provides a web chat and app front end for Alibaba's Qwen model family, with open-weight language, coding, vision, audio, image, and reasoning models. Its appeal is breadth; its tradeoffs are policy constraints and shifting model availability.