RootInference

Integrations

The inference layer plugs into your existing stack

500+ models from every major provider on one side; the OpenAI SDK, Vercel AI SDK, LangChain, LlamaIndex, your observability and automation tools on the other.

Model providers

OpenAI

Available

GPT model family — chat, reasoning, vision and embeddings.

Anthropic

Available

Claude models with tool use and long-context reasoning.

Google Gemini

Available

Gemini models — multimodal, long context and embeddings.

Meta Llama

Available

Llama open-weight models via multiple hosted providers.

Mistral

Available

Mistral and Mixtral models for chat, code and embeddings.

DeepSeek

Available

DeepSeek chat and reasoning models at economy pricing.

xAI Grok

Beta

Grok models for chat and reasoning workloads.

Qwen

Beta

Qwen open-weight models for chat, code and vision.

Cohere

Beta

Command models plus rerank and embedding endpoints.

Amazon Bedrock

Beta

Bedrock-hosted models through your unified API key.

Perplexity

Coming soon

Search-grounded Sonar models with citations.

Moonshot Kimi

Coming soon

Kimi long-context chat and reasoning models.

SDKs & frameworks

OpenAI SDK compatibility

Available

Point any official OpenAI SDK at our base URL — no code changes.

Vercel AI SDK

Available

First-class provider for streaming UI and tool calling.

LangChain

Available

Chat model integration for chains, agents and tools.

LlamaIndex

Available

LLM and embedding backends for RAG pipelines.

Haystack

Beta

Generator components for Haystack pipelines.

Semantic Kernel

Coming soon

Connector for Microsoft Semantic Kernel apps.

Observability

Langfuse

Available

Trace requests, prompts and costs into Langfuse.

Helicone

Beta

Request logging and analytics passthrough.

Datadog

Coming soon

Latency, error and spend metrics in your dashboards.

Automation

Zapier

Beta

Trigger inference from thousands of connected apps.

n8n

Beta

Model nodes for self-hosted workflow automation.

Make

Coming soon

Inference modules for Make scenarios.

Missing a model or provider you depend on?

We add providers continuously, and Enterprise plans support private and fine-tuned model endpoints behind the same API.