Nebius Token Factory logo

Nebius Token Factory

Netherlands EU EU hosted

Nebius Token Factory

Nebius Token Factory provides managed inference for open-source language models through an OpenAI-compatible API. Developers can use it for chat completions, reasoning, code generation, embeddings, and related application workloads without managing model-serving infrastructure. Its public serverless endpoints use token-based billing, while optional dedicated endpoints provide reserved capacity and regional deployment controls. The service supports API-key authentication, a browser-based playground, and model-specific performance tiers.

Pricing

Public serverless inference is billed per input and output token, with prices varying by model and performance tier. Free starter credits are promotional and do not constitute an ongoing free plan. Dedicated endpoints use separate capacity-based pricing.

Read more on the pricing page of the service.

Hosting

The core API domain resolves to Nebius B.V. Nebius documents fixed, user-selected deployment regions for dedicated endpoints, including eu-north1 in Finland. Public serverless endpoints are billed per token but can change region; only a fixed EU deployment meets the category's regional hosting condition.

Domain name Usage type Lookup type Hosting provider
api.tokenfactory.nebius.com Core Service Web
  • Nebius B.V. (AS213291)
Report
Other products in category Large language model (LLM) APIs
Scaleway Generative APIs logo

Scaleway Generative APIs

France EU EU hosted Free plan

Scaleway Generative APIs is a managed model-inference service from France. Its serverless APIs provide OpenAI-compatible access to language, reasoning, vision, audio, embedding, and reranking models, while dedicated deployments can serve curated or customer-selected models. Text-generation endpoints support chat completions, structured outputs, tool calling, streaming, and batch processing. Serverless usage is billed by token or task and includes an ongoing free tier.

OVHcloud AI Endpoints is a managed generative AI API service from France. It gives developers serverless access to pre-trained language, reasoning, vision, embedding, image, and speech models without requiring them to deploy inference infrastructure. Its OpenAI-compatible chat-completions and responses routes support common application patterns including text generation, structured output, and function calling. Access can be authenticated with API keys, and usage is billed according to the selected model.


Any suggestions?

Sign up for an account to suggest changes or new products.

Sign Up