bitHuman for enterprise

Real-time AI avatars that run where your data lives

bitHuman makes real-time AI avatars — visual AI agents, or digital humans — that turn audio into a lip-synced, talking face. The same avatar runs in the bitHuman cloud, on your own servers, on the device (iPhone, iPad, Mac, Android, Linux, and in the browser with WebGPU), or, for Business and Enterprise customers, completely offline. When the avatar renders on your hardware and you use your own voice and language services, bitHuman receives usage metering only — never audio, video or conversation text.

Put a lifelike, lip-synced avatar in front of customers — on the device in their hand, on the terminal in your branch, or on servers you control. You decide what, if anything, reaches us.

Building a consumer or companion app instead? Build an app →

Deployment

Four ways to deploy

The same avatar runs in each mode. What changes is where it renders, where the conversation runs, and what reaches bitHuman.

bitHuman cloud

The avatar renders
On bitHuman's servers, in the US
What reaches bitHuman
The session's audio and conversation, to run it. Transcripts are kept with your agent; deleting the agent deletes them.
Best for
The fastest start: web pages, apps and APIs

Your servers

The avatar renders
On your own Mac or Linux machines, on-premises or in your cloud account; a standard Linux PC needs no GPU
What reaches bitHuman
A credential check when a session starts, the avatar download, and usage reports with no audio, video or conversation text
Best for
On-premises, data-center and private-cloud deployments

On the device

The avatar renders
On the iPhone, iPad, Mac or Android device in front of the user, or in a WebGPU browser tab
What reaches bitHuman
With your own voice and language services, usage metering only — never audio, video or conversation text
Best for
Apps, kiosks and embedded screens

Fully offline

The avatar renders
On your Linux PCs and terminals
What reaches bitHuman
Nothing while it runs: usage is metered on the machine, and no reconnection is required
Best for
Kiosks, ATM machines, trade shows and embedded screens — Business and Enterprise

Compare deployment options

Use cases

Built for places where data has to stay put

Performance

Measured on everyday hardware

Every configuration we publish renders faster than real time — including a standard Linux PC with no GPU and 10-minute runs on phones.
Seconds of avatar video rendered per second, per device and model
DeviceEssence 2Expression 2
macOS · CLIApple M44.24× real time2026-09-278.4× real time2026-09-27
macOS · PythonApple M46.96× real time2026-09-258.45× real time2026-09-25
macOS · Swift packageApple M44.8× real time2026-09-248.85× real time2026-09-24
Linux · CLIIntel Core i7-13700F (x86_64)2× real time2026-09-272.2× real time2026-09-27
Linux · PythonIntel Core i7-13700F (x86_64)1.96× real time2026-09-262.35× real time2026-09-26
iPhone · Swift packageiPhone 152.16× real time2026-09-275.55× real time2026-09-27
AndroidSamsung Galaxy S25+2.08× real time2026-09-252.4× real time2026-09-23
Web browser (WebGPU)Chrome on Apple M41.72× real time2026-09-271.95× real time2026-09-27
iPhone · Swift package · held 10 miniPhone 151.32× real time2026-09-255.15× real time2026-09-25
Android · held 10 minSamsung Galaxy S25+1.48× real time2026-09-272.2× real time2026-09-27
Web browser (WebGPU) · held 10 minChrome on Apple M42.16× real time2026-09-272.05× real time2026-09-25
“× real time” is seconds of avatar video rendered per second of wall-clock time: end-to-end audio in to frame out, one session, unpaced; the slowest of three quiet runs at least ten minutes apart on the published release. Read from docs performance.json (generated 2026-09-27). How we measure →

Security and privacy

Private by architecture

  • When the avatar renders on your hardware, its audio and video stay on your hardware.
  • Usage reports carry no audio, video, images or conversation text.
  • We do not use your content to train our AI models unless you explicitly opt in, and we do not sell your personal information.

Read the details

In production

Already running on site

  • MINT Museum of Toys

    The museum's AI ambassador gives visitors multilingual guidance and ticketing, running on-premises on one Mac mini.

  • NRF: Retail's Big Show

    Ten AI-powered conference guides worked as interactive kiosks, helping attendees navigate the venue and find sessions.

  • Ivoclar Vivadent

    A “Virtual Einstein” avatar delivers 24/7 customer care across trade shows and online.

FAQ

Questions security and procurement teams ask

Can bitHuman run on the device?

Yes. Essence 2 and Expression 2 render on iPhone and iPad, Android (arm64), Mac, Linux and in browsers with WebGPU. Measured results for each device are published at docs.bithuman.ai/performance.

Can bitHuman run without an internet connection?

Offline license is only available to Business and Enterprise clients who want to run realtime avatars completely locally, off the internet — e.g. kiosks, trade shows, ATM machines, embedded screens. Linux PCs and terminals; arranged through sales. Essence 1 runs fully offline on Linux (x86_64 and ARM64) today. Essence 2 and Expression 2 offline come later. Expression 1 runs in the bitHuman cloud only. Packs start at 100,000 credits. Other self-hosted and on-device sessions check in when they start and keep rendering through a network drop of up to 5 minutes.

Does customer audio or video leave our network?

When the avatar renders on your hardware, its audio and video stay there. The conversation runs where you choose: the CLI's local conversation brain, your own speech and language services, or bitHuman's cloud voice service. With the first two, bitHuman receives usage metering only — never audio, video or conversation text. In the bitHuman cloud, and with the web embed, audio and conversation are processed by bitHuman and its service providers to run the session.

Is bitHuman HIPAA compliant?

HIPAA compliance is a property of your program and deployment, not of a single software component. bitHuman can be deployed so that no protected health information reaches bitHuman — on the device, on your own servers with speech and language services you control, or fully offline — and sessions that render on your hardware store no transcript with us. Healthcare deployments are set up under an enterprise agreement and review; contact us to walk through your data flows.

How does bitHuman help with GDPR and CCPA?

With on-device, self-hosted or offline deployment, your customers' conversations stay in the systems and regions you choose, and bitHuman receives usage metering only. Creating an avatar from a portrait happens in the bitHuman cloud. We do not sell your personal information, and we do not use your content to train our AI models unless you explicitly opt in.

All questions

Talk to us

Tell us where the avatar needs to run

Share your devices, your network and your requirements. We'll walk your team through each data flow and the deployment that fits.

Or email hello@bithuman.ai.