AIAny
AI Model2026
Icon for item

Nex-N2.5-mini

A multimodal, agentic LLM optimized for long‑horizon, visually grounded workflows — capable of operating browsers and terminals and autonomously executing and testing code. Open‑source weights are available and the family ships in mini, Pro and Max variants for different compute/quality tradeoffs.

Introduction

Most recent model releases treat vision as an additional input for classification or captioning. Nex‑N2.5 instead treats vision as an active control channel in closed perception→action→verification loops, aiming to sustain multi‑step agentic tasks such as web automation, terminal-based development, and program execution & debugging over long horizons.

Key Capabilities
  • Visually grounded agenting — integrates visual observations into continuous action planning and self-correction, so the model can verify outcomes and iterate rather than rely on single-step replies.
  • Computer and browser control — supports autonomous interactions with GUIs, webpages and terminals, making it suitable for tasks like automated testing, data extraction, and scripted workflows.
  • Agentic coding and execution — can generate, run, and debug code in a loop, enabling end‑to‑end development workflows where the model tests and refines its outputs.
  • Multi-scale open weights — released in mini/Pro/Max sizes so teams can choose between lower-latency deployment (mini) and higher-capability, large‑scale MoE configurations (Max).
Who it's for and tradeoffs

Great fit if you want an open‑source agent model for automating multi‑step real‑world workflows (browser/terminal automation, programmatic testing, agentic research) and can provide accelerated compute for deployment. Look elsewhere if you only need a lightweight chat model, require a fully managed enterprise SLA, or cannot meet the GPU/memory requirements for the desired model tier. The family emphasizes agentic, multimodal capabilities over minimal-resource conversational latency.

More Items

Hugging Face

Provides ~483K agent instruction‑tuning trajectories for supervised fine‑tuning, including tool calls, environment feedback, errors/retries and verification across search, code, office and general agent workflows; static snapshots for SFT and mix‑ratio studies.

Hugging Face
AI Model2026

Compact causal LLM for on-device assistants, coding agents and long-context tool use — ~2.52B parameters with a 131,072-token context, trained with SFT + RL + OPD and released with its UltraData training corpora and multi-format deployment checkpoints.

Hugging Face
AI Model2026

Provides a cybersecurity-focused CRACK variant of GLM-5.3 FP8 that reduces refusals for offensive-security, red-team, exploit-development and malware-analysis queries while retaining native FP8 speed on Hopper GPUs; MIT-licensed for authorized security work.