Script
Each line names who speaks, and can carry a note such as a smile or a pause.
We give AVSP a script, a picture and a voice for each person, and it renders the video itself. Every frame is decided by us.
A script, a picture and a voice for each person, and a picture of the studio go in at one end. A finished video comes out the other. Inside, the episode moves through stations in a set order, each handing its work to the next, so a better voice tool next year changes only that one station.
Each line names who speaks, and can carry a note such as a smile or a pause.
A picture of every person in the show. The face each viewer watches is built from it.
A voice reference for each person, so every line is spoken in that voice.
A picture of the room. It becomes the set the show plays on.
The script, the portraits, the voices, and the studio picture.
Speaks every line in that person's voice, and records when each word is said.
Turns each line's audio into face movement: lips, jaw, brows, blinks, and emotion.
Posture, gesture, and the listener: nodding, looking at the speaker, reacting.
Who sits where, how the light falls, which camera shows whom, and when to cut.
The frame is drawn by our own pipeline, not generated from a prompt.
The shots and the sound are joined into the final video.
Automatic checks first, such as lip sync. Then a person watches the whole video before anything ships.
Picture and sound, joined and checked, ready to watch.
Ask a model to make your show and it hands you something you cannot direct or keep consistent, so we render instead.
Every camera move and cut is chosen and written down. Change one line of dialogue and only that shot re-renders; the rest stands.
The host's face and voice stay the same in every episode. That consistency is part of why we render.
The renders and the archive all belong to the studio.
We use AI where it works well: it speaks every script in each person's voice, and it reads a face's movement from sound. Nothing in the frame is left to chance.
Each voice is made once and then reused in every episode.
Movement comes from the words and the voice, and every frame is rendered under our direction.
The first thing the engine is making is a thirty-second announcement. See the show it opens.
About the studio: the studio page · Want to know when the show is out? [email protected]