IONOS AI Model Hub logo

IONOS AI Model Hub

Germany EU EU hosted

IONOS AI Model Hub

IONOS AI Model Hub is a managed generative AI platform from Germany. It provides OpenAI-compatible and native APIs for text and image generation, embeddings, document collections, semantic search, and retrieval-augmented generation. Developers can use curated open-source language models for chat, text, and code generation without managing GPU infrastructure. Usage is billed according to the selected model and the tokens or images processed.

Pricing

Usage is billed according to the selected model. Language models are priced per million input and output tokens, embedding models per processed tokens, and image models per generated image. Current rates are published on the IONOS Cloud pricing page.

Read more on the pricing page of the service.

Hosting

IONOS states that all AI Model Hub services, including inference endpoints and managed vector databases, are hosted in IONOS Cloud data centers in Germany. The documented OpenAI-compatible endpoint is located in the IONOS-operated Berlin data center.

Domain name Usage type Lookup type Hosting provider
openai.inference.de-txl.ionos.com Core Service Web Report
Other products in category Large language model (LLM) APIs
Scaleway Generative APIs logo

Scaleway Generative APIs

France EU EU hosted Free plan

Scaleway Generative APIs is a managed model-inference service from France. Its serverless APIs provide OpenAI-compatible access to language, reasoning, vision, audio, embedding, and reranking models, while dedicated deployments can serve curated or customer-selected models. Text-generation endpoints support chat completions, structured outputs, tool calling, streaming, and batch processing. Serverless usage is billed by token or task and includes an ongoing free tier.

Nebius Token Factory logo

Nebius Token Factory

Netherlands EU EU hosted

Nebius Token Factory provides managed inference for open-source language models through an OpenAI-compatible API. Developers can use it for chat completions, reasoning, code generation, embeddings, and related application workloads without managing model-serving infrastructure. Its public serverless endpoints use token-based billing, while optional dedicated endpoints provide reserved capacity and regional deployment controls. The service supports API-key authentication, a browser-based playground, and model-specific performance tiers.


Any suggestions?

Sign up for an account to suggest changes or new products.

Sign Up