Nexith Core は自社主権 GPU クラスタで稼働開始 —— キャッシュヒットは最大90%オフ、OpenAI 互換。詳細を見る
Nexith
登録
All models
Available now

Nexith Core

Language & reasoning

Our flagship model — built for reasoning, writing, and code.

Try Nexith CoreRead the docs
Abstract luminous neural lattice representing Nexith Core
Core system
● Sovereign inference · online
Why Core

A model that keeps the whole problem in view.

Nexith Core is our flagship model, trained from the ground up for reasoning, writing, and code. It holds a 1M-token context, so it works over whole documents and codebases at once instead of fragments — and it reads images alongside text.

Context window
1M tokens
Max output
128K tokens
Input
Text · Image
Knowledge cutoff
Feb 2026
API
OpenAI-compatible
Cache hits
Up to 90% off
Pricing
Input
$4.00
per 1M tokens
Cached input
$0.40
per 1M tokens
Output
$20.00
per 1M tokens

Semantic cache hits save up to 90% on input tokens. No credit card required to start.

Modalities
Text in

Prompts, documents, and code — up to 1M tokens in a single request.

Image in

Read images alongside text in the same message.

Text out

Streamed token-by-token over SSE, up to 128K per response.

Endpoints
POST /v1/chat/completions

Primary chat endpoint — streaming and non-streaming.

POST /v1/responses

Multi-step responses with tools and structured output.

POST /v1/embeddings

Vector embeddings for search and retrieval.

Features
Streaming (SSE)
Function calling
Structured outputs
Vision (image in)
Persistent memory
Semantic cache
Tools
Function calling

Define tools in JSON schema; the model calls them with typed arguments.

Structured outputs

Constrain responses to a schema so you get valid JSON every time.

Memory

Persist context across sessions, scoped per user and per app.

Rate limits
Tier
RPM
TPM
Daily limit
Free
10
20K
1M tokens
Pro
300
500K
10M tokens
Enterprise
Custom
Custom
Unlimited
Overview

Nexith Core is our flagship model, trained from the ground up for reasoning, writing, and code. It holds a 1M-token context, so it works over whole documents and codebases at once instead of fragments — and it reads images alongside text.

It runs on our own sovereign GPU cluster behind a clean, OpenAI-compatible API — with streaming everywhere, a semantic cache that makes repeated work instant, and memory that carries across sessions.

What makes it good

Serious coding

Writes, reads, and refactors code across languages — and knows when a task is actually done.

Long-horizon tasks

Breaks a goal into steps, tracks its own budget, and drives to a clean finish instead of spinning.

Persistent memory

Remembers across sessions — scoped per user and per app, so it picks up where you left off.

Semantic cache

Recognizes repeated work and answers instantly at up to 90% off, with no repeated compute.

Start building with Nexith Core

The same OpenAI-compatible API, up to 90% off on cache hits, streaming everywhere. No seats, no minimums.

Try Nexith CoreRead the docs
More from the family
Preview

Nexith

NX-ONE

One endpoint that routes each request to the right capability — automatically.

Learn more
Coming soon

Nexith Image

NX-I1

Generate and edit images from a prompt — coming to Apps and the API.

Learn more
Nexith — フロンティア AI モデル | Nexith