Sakana AI Releases Namazu, a Japanese-Specialized Reasoning Model Built on Kimi K2.6
Sakana AI has released Namazu, a reasoning model built on Kimi K2.6 and fine-tuned for Japanese language and business contexts. The model offers a 262K token context window at $0.95 per 1M input tokens and $4 per 1M output tokens.
Sakana Namazu — Quick Specs
Sakana AI has released Namazu, a reasoning model specialized for Japanese language and business use cases, built on top of Moonshot AI's Kimi K2.6 architecture with additional training data.
What Namazu Does
According to Sakana AI, Namazu is designed for Japanese instruction following, business writing, mathematics, coding, research, and multi-step agent workflows. The model inherits its base capabilities from Kimi K2.6 and adds targeted fine-tuning for Japanese linguistic and business-context tasks — a niche that general-purpose frontier models often handle less precisely than dedicated regional models.
Pricing and Specs
Namazu is priced at $0.95 per 1 million input tokens and $4 per 1 million output tokens. The model supports a context window of 262,000 tokens, putting it in line with other long-context reasoning models currently available through OpenRouter.
No published benchmark scores accompanied the release. Sakana AI has not disclosed a parameter count or training cutoff date for Namazu.
Data Handling Notice
Sakana AI states that it may use user inputs for model training and improvement by default, with an opt-out available through the Sakana Console. The company also notes that processing of requests is not guaranteed to occur entirely within Japan, a detail relevant to enterprises with data residency requirements — a common consideration for Japanese business customers evaluating AI vendors.
What This Means
Namazu is a fine-tune rather than a from-scratch model, built on Moonshot AI's Kimi K2.6 base. This approach lets Sakana AI compete in the Japanese-language and business-context niche without the cost of training a foundation model from zero. The pricing — $0.95/$4 per 1M tokens — positions Namazu as a mid-tier option, more expensive than budget open models but well below flagship frontier pricing.
The lack of published benchmarks makes it difficult to independently verify Sakana AI's claims about Namazu's performance on Japanese instruction following or coding tasks. Buyers evaluating the model for production use, particularly in regulated industries, should also weigh the data residency caveat: Sakana AI explicitly does not guarantee in-country processing, which could matter for Japanese enterprises with strict compliance requirements. As with any derivative model, its ceiling is ultimately bound by the capabilities of its Kimi K2.6 base plus whatever additional fine-tuning Sakana AI applied.
Related Articles
Meta Releases Muse Spark 1.3 Contributor, a Low-Cost Multimodal Reasoning Model With 1M Context Window
Meta has released Muse Spark 1.3 Contributor, described as the cost-efficient contributor tier of its multimodal reasoning model line. The model offers a 1 million token context window at $0.10 per 1M input tokens and $0.20 per 1M output tokens, targeting experimentation and early-stage agentic workflows.
Meta Releases Muse Spark 1.3, a Free Multimodal Reasoning Model with 1M-Token Context
Meta has released Muse Spark 1.3, a multimodal reasoning model with a 1M-token context window, listed as free on OpenRouter. The model targets long-running agentic, multi-agent, and coding workflows, though audio input support remains incomplete.
InclusionAI Releases Ling 3.0 Flash Fin, a Finance-Focused MoE Model with 5.1B Active Parameters
InclusionAI has released Ling 3.0 Flash Fin, a finance-specialized mixture-of-experts model built on Ling 3.0 Flash. The model activates 5.1B of its 124B total parameters and targets long-horizon investment planning tasks while retaining general reasoning, coding, and math capabilities.
Google Lists Gemini 3.8 Flash on OpenRouter With 1M-Token Context, September 2026 Release Date
Google's Gemini 3.8 Flash has surfaced on OpenRouter with a 1-million-token context window and discounted pricing of $0.75 per 1M input tokens and $3.75 per 1M output tokens. Google has not issued a separate public announcement, and the listed release date of September 2, 2026 is unusually far out, leaving key details unconfirmed.
Comments
Loading...