Skip to content

Licenses in bitHuman Live

bitHuman Live listens and speaks with models that run on your device. Here are their licenses, and the use restrictions you agree to.

Last updated: October 2026

Use restrictions for the Supertonic voice

The characters' on-device voice is Supertonic 3, a model by Supertone Inc., licensed under the BigScience Open RAIL-M License (dated August 18, 2022). Its use restrictions are part of our Terms of Service: you agree not to use bitHuman Live, the model or anything derived from it for any of the uses below. This is Attachment A of that license, word for word:

Attachment A

Use Restrictions

You agree not to use the Model or Derivatives of the Model:

  • (a) In any way that violates any applicable national, federal, state, local or international law or regulation;
  • (b) For the purpose of exploiting, harming or attempting to exploit or harm minors in any way;
  • (c) To generate or disseminate verifiably false information and/or content with the purpose of harming others;
  • (d) To generate or disseminate personal identifiable information that can be used to harm an individual;
  • (e) To generate or disseminate information and/or content (e.g. images, code, posts, articles), and place the information and/or content in any context (e.g. bot generating tweets) without expressly and intelligibly disclaiming that the information and/or content is machine generated;
  • (f) To defame, disparage or otherwise harass others;
  • (g) To impersonate or attempt to impersonate (e.g. deepfakes) others without their consent;
  • (h) For fully automated decision making that adversely impacts an individual’s legal rights or otherwise creates or modifies a binding, enforceable obligation;
  • (i) For any use intended to or which has the effect of discriminating against or harming individuals or groups based on online or offline social behavior or known or predicted personal or personality characteristics;
  • (j) To exploit any of the vulnerabilities of a specific group of persons based on their age, social, physical or mental characteristics, in order to materially distort the behavior of a person pertaining to that group in a manner that causes or is likely to cause that person or another person physical or psychological harm;
  • (k) For any use intended to or which has the effect of discriminating against individuals or groups based on legally protected characteristics or categories;
  • (l) To provide medical advice and medical results interpretation;
  • (m) To generate or disseminate information for the purpose to be used for administration of justice, law enforcement, immigration or asylum processes, such as predicting an individual will commit fraud/crime commitment (e.g. by text profiling, drawing causal relationships between assertions made in documents, indiscriminate and arbitrarily-targeted use).

Read the full license on Hugging Face. The app has the full text too, under Account, then Licenses.

Characters are AI, and so is what they say

Everything a character says, and the voice it says it in, is generated by AI. The app tells you so before your first call and on screen during every call. If you share anything a character said or how it sounded, such as a recording, a clip or a quote, say clearly that it is machine-generated.

Changes we made to the voice model

We don't use the Supertonic files exactly as Supertone published them. On iPhone, iPad and Mac, we store the model's weights at half precision (16-bit) and convert them back to 32-bit when they load; each changed file records the change. On Android, we use the 8-bit (INT8) build published by the sherpa-onnx project.

Speech recognition and the other parts

Speech recognition uses NVIDIA Parakeet TDT-CTC 110M by NVIDIA Corporation, licensed under CC BY 4.0. We use the version converted to ONNX and quantized to 8 bits (INT8) by the sherpa-onnx project, without further changes.

  • NVIDIA Parakeet TDT-CTC 110M (speech recognition): CC BY 4.0
  • Silero VAD (tells when you have finished speaking): MIT
  • sherpa-onnx (runs speech recognition, and the voice on Android): Apache 2.0
  • ONNX Runtime (runs the on-device models): MIT

Speech recognition: NVIDIA Parakeet (CC BY 4.0). Voices: Supertone Supertonic (OpenRAIL-M).

Contact

Questions about these licenses? Write to us at hello@bithuman.ai.