OpenAI Launch Review: Dots Just Dropped — Here’s What It Actually Does
Last month I needed a quick way to add a talking avatar to a client demo, and when OpenAI launched Dots this week I got exactly that. Dots is a bubbly agentic avatar that syncs facial expressions and lip‑movement to your voice in real time, letting you drop a talking character into videos or live streams with no extra software. It matters because it cuts the production time for personalized avatars from hours to minutes, which is a big deal for anyone who makes frequent short‑form content.
OpenAI launch review: what it actually does
Dots works as a web‑based tool you open in Chrome or Edge. You record or upload a short audio clip, pick a base avatar from the library, and the system generates a matching video where the avatar’s head, eyes, and mouth move in sync with the sound. The output is a transparent‑background WebM file you can overlay on any background. The service runs entirely in the cloud; nothing installs on your machine. You can also drive the avatar live via a microphone, which makes it usable for virtual events or Twitch streams.
Under the hood Dots uses a diffusion model trained on a mix of synthetic and real‑world facial data, fine‑tuned for real‑time latency under 200 ms on a typical broadband connection. The avatar library includes a range of styles—from cartoonish to semi‑realistic—each with adjustable parameters for eye blink frequency, head tilt intensity, and mouth openness. You can export the result at up to 4K resolution, though the free tier caps output at 1080p.
Who this is for
Solo creators who need a fast way to add a talking head to tutorials, product demos, or social clips will find Dots useful. Small agencies that produce a lot of client‑specific video intros can use it to generate a personalized avatar per project without hiring a motion‑graphics freelancer. Educators who want to put a friendly face on lecture snippets without appearing on camera also benefit. It’s less suited for large studios that need deep rigging control or complex body animation; those teams will still rely on traditional pipelines.
What to try in the first 15 minutes
- Record a 10‑second voice memo directly in the browser and watch the avatar lip‑sync instantly.
- Swap between two avatars from the library and compare how each style handles the same audio.
- Adjust the “eye blink” slider from 0 to 100 % and notice the difference in perceived liveliness.
- Export the result as a WebM and drop it into a simple HTML page to see the transparent background work.
- Try the live microphone mode: speak for 30 seconds while the avatar follows your tone and pitch in real time.
