For the complete documentation index, see llms.txt. This page is also available as Markdown.

Welcome

Build AI Agents That Feel Human. In Minutes.

Live in real time. Speech in, synchronized video out.

Get started in 2 minutes

The fastest way to deploy a conversational AI agent with a lifelike avatar:

1. Get your API keyCreate one here

2. Create an agent in the dashboard. Pick a face, voice, and personality.

3. Embed on your site:

Vite or Create React App. Load the script once, then drop in the tag.

Vite uses the project-root index.html. Create React App uses public/index.html.

App Router. Load the script on the client only. The widget uses window at mount time:

That's it. Your users get a live conversational agent with video, audio, and a real-time avatar. Full Human Agent docs →


What Ojin offers

Apps: deploy in minutes

Human Agent

A complete conversational AI agent with a realistic visual avatar. Speech in, speech out, with synchronized lip movements and expressions. Two modes:

  • Ojin Agent: Ojin handles everything (STT, LLM, TTS, avatar). You configure the personality.

  • Third-Party Agent: bring your own speech-to-speech provider (Hume, ElevenLabs, Ultravox). Ojin adds the face.

Get started →

Models: build custom pipelines

Drive Ojin's real-time face models from your own stack. Both are powered by the same Python SDK and Pipecat integration. Pick the model with its config_id.

ojin/human-presence

Our flagship face model: a fully expressive, generative presence with rich expressions and natural movement (including the hands) that goes beyond lip-sync.

Learn more →

ojin/human-portrait

A cost-effective, real-time lipsync model. Transforms a single reference image into a natural animated persona with audio-synchronized lip movements and expressions. Streams at 25 fps, up to 720p.

Learn more →


Core features

  • Human Agent: end-to-end conversational AI with a lifelike visual avatar. One widget embed, no pipeline assembly.

  • Real-time streaming: WebSocket and WebRTC transport built for live, conversational latency.

  • One-shot personas: create a lifelike persona from a single reference image. No training, ready immediately.

  • Python SDK and Pipecat: drive the face models from ojin-client, or drop pipecat-ojin into a Pipecat pipeline. Raw WebSocket and REST when you need them.

  • Auto-scale: our infrastructure takes care of the scale for you, your users will never see an unavailable service.

  • Cost-effective: competitive per-minute pricing with no commitments. $10 free credits to start.


Use cases

  • Customer Support: deploy lifelike agents for personalized 24/7 support

  • Sales: greet, qualify, and convert leads with conversational avatars

  • Education: build interactive tutors with natural speech and expressions

  • Onboarding and Training: conversational AI for employee learning

  • Brand Ambassador: always-on, always-on-brand digital representative

  • Healthcare: empathetic virtual health assistants


LLM-Ready Docs

This documentation is optimized for LLM access:

Note: llms.txt and llms-full.txt are auto-generated from the published sitemap. Human Agent pages will appear once they are published in the navigation.

Last updated

Was this helpful?