model release

Nex AGI Releases Nex-N2-Pro: 397B Parameter MoE Model With 262K Context, Available Free

TL;DR

Nex AGI has released Nex-N2-Pro, an agentic mixture-of-experts model with 397B total parameters and 17B active parameters. The model features a 262K token context window and is available free via OpenRouter's API.

1 min read
0

Nex AGI Releases Nex-N2-Pro: 397B Parameter MoE Model With 262K Context, Available Free

Nex AGI has released Nex-N2-Pro, a mixture-of-experts (MoE) model with 397B total parameters and 17B active parameters. The model is available free through OpenRouter's API.

Model Specifications

Nex-N2-Pro is built on the Qwen3.5 architecture and accepts both text and image inputs, producing text output. The model features a 262K token context window, significantly larger than many competing models.

The MoE architecture activates only 17B parameters per inference while maintaining access to the full 397B parameter base. Nex AGI describes the model as "agentic," though specific details about its agentic capabilities have not been disclosed.

Availability and Pricing

The model is accessible through OpenRouter's API under the endpoint nex-agi/Nex-N2-Pro:free at no cost. Pricing for input and output tokens: $0 per 1M tokens.

No benchmark scores or performance comparisons have been released at the time of publication. Training data cutoff date and release date have not been specified by the company.

What This Means

The release of a 397B parameter MoE model at zero cost represents an unusual pricing strategy in the AI model market. The 262K context window positions it among the longer-context models available, though competitors like Claude 3.5 (200K) and GPT-4 Turbo (128K) have established context length benchmarks. Without published benchmark scores, actual performance relative to other models remains unverified. The Qwen3.5 architecture base suggests technical lineage from Alibaba's open-source work, though the extent of modifications by Nex AGI is unclear.

Related Articles

model release

Meituan launches LongCat 2.0: 1.6T parameter MoE model with 1M+ context window at $0.30 per 1M input tokens

Meituan has released LongCat 2.0, a sparse mixture-of-experts language model with 48 billion active parameters out of 1.6 trillion total. The model features a 1,049,000 token context window and costs $0.30 per 1M input tokens and $1.20 per 1M output tokens.

model release

Alibaba releases Qwen 3.8, a 2.4 trillion parameter open-weight model claiming second place behind Fable 5

Alibaba has released Qwen 3.8, a 2.4 trillion parameter open-weight model that the company claims trails only Fable 5. The multimodal model processes images, videos, and documents, with a preview available through Alibaba's platforms at 10 percent of standard pricing.

model release

Thinking Machines releases Inkling: 975B-parameter MoE model with Apache 2.0 license, first major US open-weight multimo

Thinking Machines Lab released Inkling, a mixture-of-experts model with 975B total parameters and 41B active parameters, trained on 45 trillion tokens across text, images, audio, and video. The Apache 2.0-licensed model supports up to 1M context and debuts alongside Inkling-Small (276B-A12B), marking what observers call the strongest US-based open-weight release to date.

model release

Poolside Releases Laguna S 2.1, an 8B-Active-Parameter Open Coding Model That Rivals Systems 20x Its Size

Poolside has released Laguna S 2.1, a mixture-of-experts coding model with 8 billion active parameters out of 118 billion total, its third coding model release in three months. The company claims it outperforms open-weight models 10 to 20 times its size on agentic coding benchmarks like Terminal-Bench 2.1 and DeepSWE.

Comments

Loading...

Nex-N2-Pro: Free 397B Parameter MoE Model With 262K Context | TPS