Scaleway Generative APIs logo

Scaleway Generative APIs

France EU EU hosted Free plan

Scaleway Generative APIs

Scaleway Generative APIs is a managed model-inference service from France. Its serverless APIs provide OpenAI-compatible access to language, reasoning, vision, audio, embedding, and reranking models, while dedicated deployments can serve curated or customer-selected models. Text-generation endpoints support chat completions, structured outputs, tool calling, streaming, and batch processing. Serverless usage is billed by token or task and includes an ongoing free tier.

Pricing

Serverless usage is billed by model and task, usually per million input and output tokens. An ongoing free tier includes a monthly allowance for token-based models and audio transcription, while batch requests receive discounted pricing. Dedicated deployments are billed by provisioned instance time.

Read more on the pricing page of the service.

Hosting

Scaleway states that all serverless models are currently hosted in a Paris data center operated by OPCORE and guarantees that models reached through api.scaleway.ai remain hosted in Europe.

Domain name Usage type Lookup type Hosting provider
api.scaleway.ai Core Service Web Report
Other products in category Large language model (LLM) APIs
Nebius Token Factory logo

Nebius Token Factory

Netherlands EU EU hosted

Nebius Token Factory provides managed inference for open-source language models through an OpenAI-compatible API. Developers can use it for chat completions, reasoning, code generation, embeddings, and related application workloads without managing model-serving infrastructure. Its public serverless endpoints use token-based billing, while optional dedicated endpoints provide reserved capacity and regional deployment controls. The service supports API-key authentication, a browser-based playground, and model-specific performance tiers.

OVHcloud AI Endpoints is a managed generative AI API service from France. It gives developers serverless access to pre-trained language, reasoning, vision, embedding, image, and speech models without requiring them to deploy inference infrastructure. Its OpenAI-compatible chat-completions and responses routes support common application patterns including text generation, structured output, and function calling. Access can be authenticated with API keys, and usage is billed according to the selected model.


Any suggestions?

Sign up for an account to suggest changes or new products.

Sign Up