The Prompt Does Not End the Conversation. It Starts It.
Every AI video tool on the market right now works the same way. You type a prompt, hit generate, wait, and receive a clip. If you want something different, you start over. The output is fixed. The process is closed.
HappyOyster 1.0, launched today by Alibaba Token Hub (ATH), does not work that way. It is not a video generator. It is a world model. You type a prompt, it builds a digital environment in real time, and then you walk around inside it. You can change direction, switch camera angles, add new characters, and alter the storyline while the world keeps generating around you. The scene does not freeze. It keeps going.
This is the first product that lets anyone create, explore, and direct an interactive 3D world from a text prompt or an image. And it is live right now at happyoyster.cn with free daily credits until July 17.
What HappyOyster Actually Does
HappyOyster has two modes, and both work differently from anything else on the market.
Wander Mode: Explore a World You Just Described
Type a prompt like “a rainy Paris street in the 1980s” and HappyOyster builds the scene in roughly ten seconds. Wet pavement reflecting amber streetlights. Cars moving through the frame. Shop fronts lining both sides. Then you can move through it. First-person, third-person, whatever perspective you want. Walk forward, turn corners, look around. The world keeps generating as you explore, staying visually coherent for over a minute of continuous movement.
You can also start from an image. Upload a Ghibli-style cycling scene and HappyOyster generates the full 3D environment around it. Move forward and the road extends, the flower fields spread out, the distant hills layer with depth. The style stays consistent throughout.
It auto-generates background music that matches the mood of the scene. Sound and visuals are produced together, not layered on afterward.
Directing Mode: Direct a Video in Real Time
This is the feature that changes what is possible. Directing mode gives you up to three minutes of continuous 720p video. You start with a scene, then you intervene with text, voice, or image prompts at any point to change what happens.
Start with a Ghibli-style girl walking along a country road. Partway through, type “a cute kitten runs up to the girl.” The kitten appears in the scene, running alongside her. Then type “the girl crouches down to pet the kitten.” The character bends down and reaches out. The motion is smooth. The transition is seamless. No re-rendering. No starting over. The world just incorporates the change and keeps rolling.
You can switch camera angles, redirect characters, change the environment, and alter the narrative arc. All in real time. The output is exportable as shareable video.
Why This Is Different From Sora, Kling, or Runway
Text-to-video models like Sora, Kling, and Runway are one-shot systems. You provide input, they produce output, and the process ends. There is no room to intervene. You cannot pause midway and change the weather, add a character, or redirect a car that was going left to go right instead.
HappyOyster runs on a completely different architecture. It is a streaming generative world model. Instead of rendering a fixed clip, it continuously evolves a world state. When you inject a new instruction, it recomputes the future from the current state. No reset. No re-prompt. The model stays where it is and adapts.
Under the hood, HappyOyster compresses high-dimensional video and multimodal information into a compact dynamic latent state. This keeps per-step computation low enough for real-time generation. A persistent state reuse mechanism transfers historical attention states across generation steps, maintaining scene consistency over longer durations. Audio and visuals are produced under the same world state, so the soundtrack naturally aligns with what is happening on screen.
This is not a text-to-video model with an interactive coat of paint. It is a fundamentally different thing. The world keeps going whether you intervene or not.
What This Unlocks for Creators
For anyone making content, this opens up workflows that did not exist before.
- Game prototyping. Build a playable 3D scene from a prompt, move through it, see if the spatial logic works, then iterate. No code. No engine. Ten seconds from idea to exploration.
- Interactive storytelling. Create a scene, direct the characters, change the story mid-stream, and export the result. You are the director and the editor at the same time.
- Film concept validation. Test whether a visual idea holds up in motion before committing resources to production. Walk through the set. See the lighting. Feel the pace.
- Brand and marketing content. Build immersive brand environments that viewers can explore, not just watch. Create short-form video that reacts to narrative direction in real time.
- Virtual tourism and cultural experiences. Generate walk-throughs of real or imagined locations. Deep sea. The moon. A 1980s Tokyo backstreet. If you can describe it, HappyOyster builds it.
What to Look Out For
HappyOyster is in early access. The 720p output in Directing mode is functional but not cinematic resolution. Long generation sessions can introduce scene drift, where the environment gradually loses structural consistency. Complex physics interactions (water, fire, large-scale destruction) are still rough. The model sometimes takes creative license with causal chains in ways that look impressive but do not hold up to scrutiny.
But the core capability is real. The ability to generate a world, walk through it, and reshape it in real time is something no other product offers right now. And the free daily credits make it easy to try without commitment.
The gap between “type a prompt, get a clip” and “type a prompt, enter a world” just closed. HappyOyster is the first product to actually deliver on the world model promise in a way you can use today. Whether the output is production-ready or not, the interaction model is the future. This is what video generation looks like when it stops being a one-shot pipeline and starts being a living environment.
Want more like this? Join MOKU Club for free. Weekly resources, early access to new guides, and occasional templates you can actually use. Join below.



