API Compatibility

OpenAI-Compatible API. Zero Code Changes.

Switch from OpenAI or Anthropic in minutes. Same SDKs, same code, up to 10x faster inference on EU sovereign infrastructure.

Switch in 5 Minutes

Change one line of code. Keep everything else.

01

Sign up & get API key

Create your account at cloud.infercom.ai.

02

Change base URL

Point your existing OpenAI or Anthropic SDK to api.infercom.ai

03

Choose your model

Use MiniMax M2.7 Ultraspeed (flagship), gpt-oss-120b (fastest), or any model from our catalog.

OpenAI SDK
from openai import OpenAI

client = OpenAI(
    base_url="https://api.infercom.ai/v1",
    api_key="your-api-key"
)

response = client.chat.completions.create(
    model="MiniMax-M2.7",
    messages=[{"role": "user", "content": "Hello"}],
    stream=True
)

for chunk in response:
    print(chunk.choices[0].delta.content, end="")
Responses API
from openai import OpenAI

client = OpenAI(
    base_url="https://api.infercom.ai/v1",
    api_key="your-api-key"
)

# Responses API - built for agentic workflows
response = client.responses.create(
    model="MiniMax-M2.7",
    input="Hello",
    reasoning={"effort": "high"}
)

print(response.output_text)
Anthropic SDK
from anthropic import Anthropic

client = Anthropic(
    base_url="https://api.infercom.ai",
    api_key="your-api-key"
)

response = client.messages.create(
    model="MiniMax-M2.7",
    max_tokens=1024,
    messages=[{"role": "user", "content": "Hello"}]
)

print(response.content[0].text)
curl
curl https://api.infercom.ai/v1/chat/completions \
  -H "Authorization: Bearer your-api-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "MiniMax-M2.7",
    "messages": [{"role": "user", "content": "Hello"}],
    "stream": true
  }'

Choose Your SDK

Full compatibility with OpenAI and Anthropic client libraries

OpenAI SDK

OpenAI SDK Compatible

Drop-in replacement for OpenAI. Works with your existing code, tools, and frameworks.

  • Chat Completions and Responses API
  • Works with LangChain, LlamaIndex, CrewAI
  • Python, JavaScript, TypeScript, REST
  • Streaming, function calling, JSON mode

View OpenAI compatibility docs

Anthropic SDK

Anthropic Messages API

Use the Anthropic SDK directly. Same Messages API format, EU hosted.

  • Standard /v1/messages endpoint
  • Works with existing Anthropic code
  • Tool use on MiniMax and gpt-oss-120b
  • Same authentication patterns

Vision, extended thinking, and prompt caching are not currently supported.

View Anthropic compatibility docs

Model-Specific Guides

Detailed setup instructions for specific models

MiniMax APIM2.7 models via OpenAI SDK

More Examples

Works with your favorite languages and frameworks

JavaScript / TypeScript
import OpenAI from 'openai';

const client = new OpenAI({
  baseURL: 'https://api.infercom.ai/v1',
  apiKey: 'your-api-key',
});

const response = await client.chat.completions.create({
  model: 'MiniMax-M2.7',
  messages: [{ role: 'user', content: 'Hello' }],
});

console.log(response.choices[0].message.content);
LangChain
from langchain_openai import ChatOpenAI

llm = ChatOpenAI(
    base_url="https://api.infercom.ai/v1",
    api_key="your-api-key",
    model="MiniMax-M2.7"
)

response = llm.invoke("Explain quantum computing")
print(response.content)

View full quickstart guide

Faster Than the Alternatives

SambaNova's dataflow architecture delivers up to 10x faster inference than GPU-based providers.

428tok/s · MiniMax M2.7
713tok/s · gpt-oss-120b
Up to 10xfaster than GPU inference

Output Throughput, p50, 10K input / 1K output, single request. Last measured: July 2026.

See performance benchmarks

Supported Endpoints

  • /v1/chat/completions

    Chat completions with streaming support

  • /v1/responses

    Responses API for agentic workflows - tools, reasoning, streaming

  • /v1/messages

    Anthropic Messages API format

  • /v1/models

    List available models and metadata

  • /v1/embeddings

    Text embeddings (selected models)

Full API reference

EU Sovereignty Included

OpenAI-compatible API, European jurisdiction.

  • EU sovereign models run in Germany (Equinix MU4)
  • No US CLOUD Act exposure
  • GDPR compliant, ISO 27001 certified
  • No persistent storage of prompts or outputs

Learn about EU sovereignty

ISO 27001Certified
GDPRCompliant
AI ActReady
Tier III+Datacenter

Ready to Build the Future of AI in Europe?

Join forward-thinking organizations deploying sovereign AI with world-class performance