Nex-N2.5-mini logo

Nex-N2.5-mini

Visit

Nex-N2.5-mini is Nex-AGI's Apache-2.0 multimodal agent model for computer use, browsing, and coding.

Share:
View alternatives

Nex-N2.5-mini is the small SKU in Nex-AGI's Nex-N2.5 agent family. The Hugging Face card lists license Apache-2.0, pipeline image-text-to-text, and architecture qwen3_5_moe. Hugging Face showed 62 likes and 2 downloads in the last month on 2026-09-09. createdAt is 2026-09-08. Safetensors on the card total about 35.1B parameters. r/LocalLLaMA had the card on the hot list the same day.

The family also ships Pro and Max. Max is described as a 1.6T text-only MoE. This page is the mini checkpoint, not Max.

Compare Qwen3.8-Flash-Next if you wanted Alibaba's sparse open MoE, DeepSeek-V4-Flash-Vision-Exp if you wanted a screenshot-and-text Flash VLM, or GLM-5.3-Flash if you wanted a cheap hosted fast SKU.

Key Features

  • Agentic multimodal: card says mini and Pro keep Nex-N2's computer-use, browsing, and visually grounded agent path. Vision is framed as a way to verify outcomes, not only an input.
  • Thinking modes: reasoning_effort of "none", "medium" (default adaptive), or "high". Chat template uses that field, not a generic enable_thinking flag.
  • Serving: customized SGLang image nexagi/sglang:v0.5.18-nex-patch. Mini recipe is a single node with 2 x H100, --tp 2, --reasoning-parser qwen3, --tool-call-parser qwen3_coder.
  • Sampling on the card: temperature 0.7, top_p 0.95, top_k 40.
  • Hosted path: OpenRouter listed nex-agi/nex-n2.5-mini:free at $0 / $0 on 2026-09-09. Recheck that slug before you promise a free seat.

Limitation: the GitHub mirror nex-agi/Nex-N2.5 had only 18 stars on 2026-09-09. Benchmark tables on the card are first-party, often run with NexAU / NexCUA harnesses. We did not rerun them. 62 likes on day one is a heat signal, not an audit.

Specs

Item Value Source
Parameters (safetensors) ~35.1B HF API, 2026-09-09
License Apache-2.0 HF card
Likes / last-month downloads 62 / 2 HF API, 2026-09-09
Serve recipe 2 x H100, tp 2 same card
OpenRouter mini $0 in / $0 out (:free) OpenRouter API, 2026-09-09
Software price $0 weights Apache-2.0

Card-cited mini scores (treat as vendor numbers): Terminal-Bench 2.1 73.4, SWE-Bench Pro 43.8, OSWorld-Verified 71.2, BrowseComp 83.4.

Use Cases

  • Local agent teams who can spare two H100s and want an Apache-2.0 computer-use checkpoint.
  • People comparing Nex-N2.5 sizes who need mini, not the 1.6T Max.
  • r/LocalLLaMA readers who saw the HF link on the hot list.

If you needed a dense local 27B chat model, start with Qwen3.8-27B.

Getting Started

  1. Open nex-agi/Nex-N2.5-mini.
  2. Pull nexagi/sglang:v0.5.18-nex-patch and mount the weights.
  3. Launch with --tp 2 --reasoning-parser qwen3 --tool-call-parser qwen3_coder and the mini chat template.
  4. Set reasoning_effort to "medium" unless you want always-on or no thinking.

First-party resource: Nex-N2.5-mini model card. Website: nex-agi.com.

Frequently Asked Questions

Is mini the same as Max?

No. Mini and Pro are multimodal. Max is the trillion-scale text-only MoE.

Can I run it on one 4090?

The published recipe is 2 x H100. Confirm VRAM before you promise a desktop box.

Are the SWE numbers comparable to Claude Opus 5?

The card table puts mini below Opus 5 on SWE-Bench Pro (43.8 vs 79.2). Those are vendor rows.

Alternatives

Tips

  1. Use the Nex SGLang image. A stock wheel may miss the chat template.
  2. Recheck OpenRouter :free before you budget $0 inference.
  3. Do not copy Max or Pro scores onto this mini page.

Nex-N2.5-mini is the Apache-2.0 35B agent checkpoint r/LocalLLaMA linked on 2026-09-08. Start at the Hugging Face card, then decide whether Pro or a hosted :free seat already covers the work.

Comments

No comments yet. Be the first to comment!