02 — The intelligence layer

The engine that changes everything but you.

Manifest Engine turns a live camera stream into a live brand world. It is built on leading real-time video foundation models, orchestrated so that the guest's identity is the one thing that never drifts.

Real-timeVideo stream
2Modes
1Identity, locked
60+Moments

What makes it hard

Generative video is easy to demo and hard to trust.

Most systems will happily turn a guest into a different person. That is the failure the engine is designed around. The prompt stack puts identity preservation first and last, and every world published through Manifest Studio is tested against it before it reaches a mirror.

Identity lock

Identity lock

Face, proportions, expression: kept.

The engine is instructed, before anything else and after everything else, to keep the person. Wardrobe, setting and lighting are what change.

  • Face and skin tone preserved. No idealised stranger in the mirror.
  • Expression tracked live. A smile in front of the mirror is a smile in the world.
  • Body proportions kept. The outfit fits the person, not a template.
Outfit-only mode

Two modes

Keep the store. Change the outfit.

In outfit-only mode the real background from the camera stays exactly as it is and only the garments change: a virtual try-on that respects the room. In full-world mode the setting is transformed as well. The venue chooses per mirror.

  • Outfit only. Retail, fashion, sport: the guest stays in the store.
  • Full world. Museums, cinema, travel: the guest steps somewhere else.
  • One switch. Toggled in the settings menu on the unit.
Prompt stack

Orchestration

World, moment, look, identity.

A session prompt is layered: the brand world sets tone and palette, the moment sets wardrobe and scene, the look sets the interface, and the identity lock wraps it all. Studio authors the layers; the engine assembles them per session.

  • Deterministic assembly. Same moment, same result, every session.
  • Hallucination guardrails. Worlds are reviewed by people before publishing.
  • Vendor-independent. The orchestration layer sits above the model.

At a glance

Manifest Engine in numbers and words.

RenderingContinuous real-time video, not still frames
LatencySub-second on a standard venue connection
IdentityLocked: face, skin tone, age, expression, proportions
ModesOutfit only (real background kept) or full world
Session capConfigurable centrally; 90 seconds by default
ConcurrencyMultiple mirrors per venue, leases per session
TrainingGuest footage is never used to train models
TransparencyAI notice shown to every guest

Questions

What venues and brands ask first.

Manifest Engine orchestrates leading real-time video foundation models and can move between them. The identity lock, the prompt stack and the operations layer are ours.

It is designed to. Glasses, headscarves, beards, wheelchairs and children (with a guardian's consent) are all handled by the identity lock rather than fought by it.

No. Stylisation applies to wardrobe, setting and lighting. The face is not a creative variable.

The stream is processed live. The recording the guest requests is kept for seven days and then deleted.

Next step

See yourself in it. Then decide.

Ninety seconds in front of the mirror explains more than any deck. Book a live session in Cologne or run the browser demo now.