XAI1 — FULL-SERVICE AI COMPANY

Your own AI.
Built end to end.

XAI1 designs and builds ChatGPT-class AI for your organization — the assistant itself, the platform around it, and the GPU infrastructure underneath. Run it on our cloud, or on hardware we install and support in your building.

One partner from model to metal — with 24/7 technical support.
WHAT WE BUILD

Three layers. We deliver all of them —
or just the ones you need.

Most AI projects die between the demo and production. We take responsibility for the whole distance: the model, the product around it, and the compute it runs on.

01
Your AI
ChatGPT-class assistant on your data, your brand, your rules
02
Your platform
Chat & API surface, admin console, SSO, analytics, integrations
03a
XAI1 Cloud
Our GPUs, pay-as-you-go
03b
On-premises
Your building, we install it
ONE STACK · ONE PARTNER · 24/7 SUPPORT
01 — CUSTOM AI

Your own ChatGPT

An AI assistant that knows your business and answers like your best employee.

  • Model selection & fine-tuning on your data and tone
  • Retrieval over your knowledge — documents, wikis, tickets, databases
  • Guardrails & evaluation suites before anything reaches users
  • Multilingual chat, voice, and document understanding
  • You own it — weights, data, and code stay yours
Discuss your assistant →
02 — PLATFORM BUILD-OUT

The product around it

A complete, branded platform your teams and customers actually use.

  • Chat interface & APIs under your brand and domain
  • Admin console — users, roles, usage, content controls
  • SSO/SAML, audit logs, permissions, data retention policies
  • Integrations: Slack, Teams, CRM, ERP, internal tools
  • Analytics on questions, answers, and business impact
Scope your platform →
03 — HARDWARE & COMPUTE

The metal underneath

Serious AI needs serious compute. Choose where it lives.

  • XAI1 Cloud — run on our GPU fleet, pay only for what you use
  • On-premises — we design, procure, install, and tune a GPU cluster in your facility
  • Air-gapped and data-residency deployments
  • We maintain it — monitoring, updates, capacity planning
  • 24/7 support with named engineers, not ticket queues
Compare cloud vs on-prem →
HOW AN ENGAGEMENT WORKS

From first call to running system.

WEEK 1–2DiscoverWe map your use cases, data sources, security requirements, and success metrics. You get a concrete proposal: scope, architecture, timeline, and cost.
WEEK 2–4DesignModel strategy, platform architecture, and infrastructure sizing — cloud, on-prem, or hybrid. A working prototype on your real data, so decisions are made on evidence.
WEEK 4–10BuildWe fine-tune, integrate, and harden. Your team is involved throughout — with training, documentation, and evaluation reports at every milestone.
ONGOINGRunLaunch, monitor, and improve. Model refreshes, capacity planning, and 24/7 technical support — under SLA, with engineers who know your system by name.
INFRASTRUCTURE, YOUR WAY

Our GPUs or yours. Same stack, same support.

Everything we build runs identically on XAI1 Cloud and on hardware we deploy at your site — so you can start in our cloud and move in-house later, or run both at once.

XAI1 CloudOn-Premises by XAI1
Where it runsOur managed GPU fleet — serverless or dedicated capacityYour datacenter or office — a cluster we design and install
Data residencyEncrypted in transit and at rest; region pinning availableData never leaves your network; air-gapped options
Cost modelPay per token or GPU-hour — $0 idle, no capexOne-time build + predictable support contract — no per-token fees
Time to launchDaysWeeks — including procurement, racking, networking, and burn-in
ScalingAutomatic, zero to thousands of GPUsSized to your workload; expandable by design
MaintenanceFully managed, includedWe monitor and maintain it — updates, health, capacity planning
Support24/7, SLA-backed24/7, SLA-backed, with named engineers and on-site options

Already own GPUs? We also integrate and optimize existing hardware into the same stack.

3.4×
throughput on identical hardware after XAI1 kernel optimization*
$0
billed for idle time on XAI1 Cloud — ever
99.9%
uptime SLA with service credits
24/7
technical support with named engineers

*Illustrative pre-launch figure from internal benchmarks; methodology published at GA.

FOR DEVELOPERS — THE XAI1 PLATFORM

The same engine, self-serve.

The inference platform that powers our client builds is open to every developer: deploy any open-source Hugging Face model as a production API in minutes.

  • Any open-source model, one API
    LLMs, speech-to-text, text-to-speech, vision, and embeddings — all OpenAI-compatible, so switching is one line of code.
  • It optimizes itself
    XAI1 profiles your model, picks the most cost-effective hardware, quantizes where quality allows, and compiles tuned CUDA kernels — automatically.
  • Honest economics
    Published per-token rates, batch −50%, cached input −80%, and never a cent for idle. A pre-deploy estimate before anything runs.
  • Free to start
    $25/month in recurring free credits for every developer. No credit card required.
quickstart.pyOpenAI-compatible
from openai import OpenAI client = OpenAI( base_url="https://api.xai1.cloud/v1", api_key="XAI1_KEY", ) out = client.chat.completions.create( model="meta-llama/Llama-3.3-70B-Instruct", messages=[{"role": "user", "content": "Hello, XAI1."}], )
$0.03
/1M tokens · LLM 8B
$0.05
/audio-hour · ASR
$12
/1M chars · TTS
$0.008
/1M tokens · embeddings
ENGAGEMENT MODELS

Start self-serve, or start with a conversation.

PLATFORM

$0 to start
Self-serve inference APIs for developers and teams.
  • $25/mo recurring free credits
  • Published usage rates, all modalities
  • Pro plan with priority capacity & SOC 2 access
  • Scale to dedicated endpoints anytime
Explore the platform

MANAGED AI BUILD

Scoped per project
We build your assistant and platform, end to end.
  • Fixed-scope proposal after discovery
  • Prototype on your data before you commit
  • You own the result — weights, code, data
  • Training and hand-over included
  • Typical build: 4–12 weeks
Start your project

ENTERPRISE PARTNERSHIP

Annual
Build + run: infrastructure, support, and evolution under one contract.
  • Cloud, on-prem, or hybrid deployment
  • On-site cluster design & installation
  • 99.9%+ SLA · 24/7 named engineers
  • SOC 2 · HIPAA BAA · data residency
  • Quarterly model & capacity reviews
Talk to us
QUESTIONS, ANSWERED

What clients ask before they start.

Who owns the AI you build for us?

You do. Model weights, fine-tuning data, prompts, and code are delivered as your property. If we part ways, everything keeps running — and your team is trained to operate it.

Can it really work like ChatGPT?

Yes — a conversational assistant with the same fluency, but grounded in your knowledge, speaking in your brand's voice, and bound by your rules. It cites your documents, respects user permissions, and refuses what you tell it to refuse.

Our data is sensitive. How do you handle that?

Your data is never used to train anything outside your own system. For strict environments we deploy fully on-premises — including air-gapped setups where nothing touches the internet at all.

What does the on-premises option include?

Everything: workload sizing, hardware procurement, racking, networking, the full software stack, burn-in testing, and staff training. Afterwards we monitor and maintain the cluster under a support contract — remotely or with on-site visits.

We already have GPUs. Can you use them?

Yes. We audit your existing hardware, integrate it into the stack, and typically unlock significant extra throughput through kernel-level optimization before recommending any new purchases.

How long until we have something working?

A prototype on your real data typically lands within the first two to four weeks. Full production builds run four to twelve weeks depending on scope; cloud deployments launch in days.

What does support actually mean?

24/7 coverage with named engineers who know your system — not a ticket queue. Monitoring, incident response under SLA, security updates, model refreshes, and capacity planning are part of the contract.

START THE CONVERSATION

Tell us what you want to build.

A 30-minute call with an engineer — not a sales deck. You'll leave with an honest read on scope, timeline, and cost.

hello@xai1.io · response within one business day