AI Research

Research behind the talking head.

Clone13's product is built on a generative-video pipeline that fuses neural photo animation, instant voice cloning, and frame-accurate lip-sync. Here's what we're researching and why it matters for automotive retail.

Photo animation

Turning a single static headshot into a natural, on-camera talking head — preserving identity, expression, and head movement without artifacts or distortion.

Voice cloning

Instant, high-fidelity voice cloning from a short sample. The cloned voice retains the speaker's accent, tone, and cadence across languages.

Lip-sync

Frame-accurate alignment of generated speech to mouth movement, so the video feels genuine — not a dubbed foreign film.

Multilingual synthesis

A single cloned voice can speak in many languages while keeping the speaker's timbre, letting dealerships serve diverse communities.

Consent & provenance

Research into watermarking and provenance so cloned material is traceable, and consent is verifiable — the foundation of responsible deployment.

Evaluation

We benchmark our pipeline on realism, identity fidelity, and lip-sync accuracy — iterating continuously to close the gap to a real camera.

Our principles

  • • Consent first. Every voice and photo is uploaded by its owner, scoped to their dealership, and fully deletable on request.
  • • Narrow deployment. We build for dealership sales communication — not open-ended impersonation.
  • • Transparency. Videos generated by Clone13 are understood by recipients as a salesperson's intro, sent through branded channels.
  • • Continuous improvement. Real-world dealership usage feeds back into our research to make the pipeline more natural and reliable.

Get in touch

We're a Chicago-based research and product team. If you're a researcher or partner interested in collaborating on responsible generative video for automotive retail, we'd love to talk.