Production

Puppet

Jalaran Puppet is a local lip-sync renderer that animates a character image so it speaks your audio, rendered on your own GPU with blinks, gaze and breathing driven by the sound itself.

Who Puppet is for

For people making narrated character video — explainers, series, presenters — who do not want a person’s likeness or voice sitting on someone else’s servers. Illustrated-character creators in particular, because that path is built specifically for artwork rather than adapted from a photo-realistic model.

What Puppet does

Give Puppet a character image and an audio clip and it renders a video of that character speaking. Motion is not decoration on a timer: one beat track is derived from the audio and every channel — brows, head, gaze, breathing, blinks — is a function of it, so the head dips and the brows lift on the same stressed syllable rather than drifting independently. Illustrated characters take a completely different path from photographic ones, because a model trained on real faces paints a photoreal mouth onto cel artwork; for drawn characters the mouth is constructed from colours sampled out of the character’s own pixels instead.

  • Separate rendering paths for photographic and illustrated characters
  • Blinks, gaze, head motion and breathing all driven by one audio-derived beat track
  • Runs entirely on your own GPU — likeness and voice never leave the machine
  • A job completes only when the output file itself verifies

How Puppet works

  1. Upload a character image

    A face scan runs first and fails fast if there is no usable face, rather than letting you wait minutes for a job that cannot work.

  2. Choose photographic or illustrated

    You set this; it is never guessed. A wrong guess is invisible until playback, which is the worst time to find out.

  3. Add the audio

    Audio from Voice carries a phoneme timeline and lip-syncs exactly. Any other clip falls back to estimating mouth shapes from the waveform.

  4. Render on your GPU

    The job runs on a local worker. Its estimate is based on measured throughput for the path you chose, not a generic guess.

  5. Verify and download

    A job is only marked complete after the finished file is re-inspected — duration, frame count and decodability all have to match.

What Puppet does not do

Puppet needs a real GPU; there is no cloud fallback, so on a machine without one it will not run. Audio length is capped, and the cap differs by path because the cost does. Illustrated characters combined with a driving performance clip do not yet get the drawn mouth — for artwork, the automatic mode is the one to use. Generated media holds a person’s likeness and voice, so everything is purged on a retention schedule rather than kept. It animates one still character; it is not a video editor.

Common questions

Does my character image or voice get uploaded anywhere?

No. Rendering happens on your own GPU through a local worker that has no credentials and no database access. Files stay in your private storage and are purged automatically on a retention schedule.

Does it work with anime or illustrated characters?

Yes, and that path is built for them specifically. A model trained on real faces generates a photoreal mouth on cel artwork, so for illustrated characters the mouth is drawn from colours sampled out of the character’s own pixels instead.

How do I know the video is not silently truncated?

Because a job is not marked complete until the output file is re-inspected — it has to exist, decode, and match the audio duration and the composited frame count. A wrongly-failed job is honest and retryable; a wrongly-completed one is indistinguishable from a good one.