CPU only (no GPU)

Both models run live on a standard Linux PC with no GPU.

Creator plan or higher No GPU Your servers Linux x86_64 / arm64

What it is

Essence 2 and Expression 2 render live on the processor of an ordinary Linux PC, with no graphics card: with the CLI or the Python SDK, on Linux x86_64 or arm64. It suits screens that run all day where a GPU is not practical: kiosks, lobby screens, and servers without GPUs.

Where it renders

QuestionCPU only (no GPU)
Where the avatar renderson a standard Linux PC's CPU, with no GPU
Where the conversation runsyour choice: the CLI's local conversation brain, your own services, or bitHuman's
What reaches bitHumana credential check, the avatar download, and usage reports with no audio, video or text
Networkto start; rendering continues through a drop of up to 5 minutes
A LINUX PC · CPU ONLY, NO GPUYour appthe CLI, the Python SDK or the LiveKit pluginThe avatar renders on the CPUits audio and video stay on this PCThe conversationthe local brain, your own services, orbitHuman'sBITHUMANCredential check and usageusage: noaudio, video ortextthe avatar,once
CPU only (no GPU). On a standard Linux PC with no GPU, both models render live on the CPU, and the avatar's audio and video stay on the PC. The conversation runs where you choose: the CLI's local conversation brain, your own services, or bitHuman's. For the rendering, bitHuman receives a credential check, the avatar download and usage reports with no audio, video, images or conversation text.

When the avatar renders on your hardware, its audio and video stay there. With the CLI’s local conversation brain, speech recognition, the language model and the voice run on the machine too; the session still reports usage online.

Models available here

ModelLinux, no GPUHow
Essence 2YesCLI, Python
Expression 2YesCLI, Python
Essence 1YesCLI (run), Python
Expression 1—

Speed

Measured on a desktop CPU with no GPU:

ConfigurationEssence 2Expression 2
Intel Core i7-13700F (x86_64) Linux · CLI CPU only (no GPU)
2.0× real timeIntel Core i7-13700F (x86_64), CPU only (no GPU) · CLI 2.8.1 · measured 2026-09-27
2.2× real timeIntel Core i7-13700F (x86_64), CPU only (no GPU) · CLI 2.8.1 · measured 2026-09-27
Intel Core i7-13700F (x86_64) Linux · Python CPU only (no GPU)
1.9× real timeIntel Core i7-13700F (x86_64), CPU only (no GPU) · bithuman 2.11.13 · measured 2026-09-26
2.3× real timeIntel Core i7-13700F (x86_64), CPU only (no GPU) · bithuman 2.11.13 · measured 2026-09-26

Times real time: seconds of avatar video rendered per second. At 1.0× or more, an avatar holds a live conversation. Select a figure for its release and date. All configurations and how we measure.

Price

2 credits per minute of active session time for Essence 2 and Expression 2, about $0.02 a minute at the top-up rate of $1 = 100 credits. Realtime usage bills active session time, talking or idle, to the second. Every rate: Pricing and credits.

Limits

  • Network: a session checks your credential when it starts and keeps rendering through a network drop of up to 5 minutes. Usage reports carry no audio, video, images or conversation text.
  • Operating system: Linux on x86_64 or arm64. On Windows, use WSL2 or the web embed; Intel Macs are not supported.
  • Sessions: self-hosted sessions are limited by credits.

First command

CLI

curl -fsSL https://install.bithuman.ai | sh
bithuman login
curl -fsSLo speech.wav https://docs.bithuman.ai/samples/speech.wav
bithuman render wise-pup speech.wav -o out.mp4

Python

python3 -m venv .venv && source .venv/bin/activate
pip install "bithuman[expression-2]"
export BITHUMAN_API_SECRET="<your API secret>"
curl -fL -o wise-pup.imx "https://api.bithuman.ai/v1/agent/A23WJF0199/model/download?model=expression-2"
curl -fsSLo speech.wav https://docs.bithuman.ai/samples/speech.wav
python -c 'import bithuman
with bithuman.open("wise-pup.imx") as a: print(sum(1 for _ in a.render("speech.wav")), "frames")'
# → 300 frames

bithuman run wise-pup opens a live conversation instead of a file (CLI).

Choosing between modes