CPU only (no GPU)
More ▾
Both models run live on a standard Linux PC with no GPU.
What it is
Essence 2 and Expression 2 render live on the processor of an ordinary Linux PC, with no graphics card: with the CLI or the Python SDK, on Linux x86_64 or arm64. It suits screens that run all day where a GPU is not practical: kiosks, lobby screens, and servers without GPUs.
Where it renders
| Question | CPU only (no GPU) |
|---|---|
| Where the avatar renders | on a standard Linux PC's CPU, with no GPU |
| Where the conversation runs | your choice: the CLI's local conversation brain, your own services, or bitHuman's |
| What reaches bitHuman | a credential check, the avatar download, and usage reports with no audio, video or text |
| Network | to start; rendering continues through a drop of up to 5 minutes |
When the avatar renders on your hardware, its audio and video stay there. With the CLI’s local conversation brain, speech recognition, the language model and the voice run on the machine too; the session still reports usage online.
Models available here
| Model | Linux, no GPU | How |
|---|---|---|
| Essence 2 | Yes | CLI, Python |
| Expression 2 | Yes | CLI, Python |
| Essence 1 | Yes | CLI (run), Python |
| Expression 1 | — |
Speed
Measured on a desktop CPU with no GPU:
| Configuration | Essence 2 | Expression 2 |
|---|---|---|
| Intel Core i7-13700F (x86_64) Linux · CLI CPU only (no GPU) | 2.0× real timeIntel Core i7-13700F (x86_64), CPU only (no GPU) · CLI 2.8.1 · measured 2026-09-27 | 2.2× real timeIntel Core i7-13700F (x86_64), CPU only (no GPU) · CLI 2.8.1 · measured 2026-09-27 |
| Intel Core i7-13700F (x86_64) Linux · Python CPU only (no GPU) | 1.9× real timeIntel Core i7-13700F (x86_64), CPU only (no GPU) · bithuman 2.11.13 · measured 2026-09-26 | 2.3× real timeIntel Core i7-13700F (x86_64), CPU only (no GPU) · bithuman 2.11.13 · measured 2026-09-26 |
Times real time: seconds of avatar video rendered per second. At 1.0× or more, an avatar holds a live conversation. Select a figure for its release and date. All configurations and how we measure.
Price
2 credits per minute of active session time for Essence 2 and Expression 2, about $0.02 a minute at the top-up rate of $1 = 100 credits. Realtime usage bills active session time, talking or idle, to the second. Every rate: Pricing and credits.
Limits
- Network: a session checks your credential when it starts and keeps rendering through a network drop of up to 5 minutes. Usage reports carry no audio, video, images or conversation text.
- Operating system: Linux on x86_64 or arm64. On Windows, use WSL2 or the web embed; Intel Macs are not supported.
- Sessions: self-hosted sessions are limited by credits.
First command
CLI
curl -fsSL https://install.bithuman.ai | sh
bithuman login
curl -fsSLo speech.wav https://docs.bithuman.ai/samples/speech.wav
bithuman render wise-pup speech.wav -o out.mp4
Python
python3 -m venv .venv && source .venv/bin/activate
pip install "bithuman[expression-2]"
export BITHUMAN_API_SECRET="<your API secret>"
curl -fL -o wise-pup.imx "https://api.bithuman.ai/v1/agent/A23WJF0199/model/download?model=expression-2"
curl -fsSLo speech.wav https://docs.bithuman.ai/samples/speech.wav
python -c 'import bithuman
with bithuman.open("wise-pup.imx") as a: print(sum(1 for _ in a.render("speech.wav")), "frames")'
# → 300 frames
bithuman run wise-pup opens a live conversation instead of a file (CLI).
Choosing between modes
- Off the internet, on Linux PCs and terminals: Fully offline, for Business and Enterprise.
- On your own Macs, or Linux machines you already run: Your servers.
- Inside an app on the phone or in the browser: On the device.
- Nothing to run yourself: bitHuman cloud.
- All five side by side: Deployment options.