PixVerse R2 - A real-time world model you can explore and change

PixVerse R2 is a real-time world model that generates continuously evolving audiovisual worlds instead of fixed video clips. It accepts text, images, audio and actions while generating, remembers what happened earlier in the session, and carries those changes forward in real time. R2 scales to longer, more coherent and controllable experiences — powering everything from interactive stories and characters to playable generative worlds.

Add a comment

Replies

Best

What kind of response time are you targeting when someone interacts with the world?

Good question, I don’t want to quote a number without checking with the team lol. We’re aiming for the interaction to feel as close to real time as possible. Let me confirm the current target with our engineers then get back to u

This is just crazy where we're heading 😮 This looks really great! Congrats on the launch! The landing hero kinda made me feel a bit dizzy because some are really moving fast visuals so it was hard to focus on the text (but this just might be me). Otherwise you do a really good job at explaining the product upfront. The one other minor detail I noticed is the "muted" icon when I go to the gallery, that stripped musical note is not very common and I honestly though it looked a broken icon until I clicked it and saw the complete icon. Maybe use the same icons as you find in YouTube etc.? This is my 2c, wish you all the best with the launch!

Thanks Owen, for your support and suggestions😊

That's pretty damn cool! I was thinking on doing an emergence experiment with AI, perhaps this could be exactly the tool I need! Will certainly check this out

Thank u for loving it lol🫶🏻

 🫰🫰🫰

Hey everyone, Loqi from PixVerse here 👋

R2 lets you explore an AI-generated world with WASD and shape what happens next. We’re working on making those interactions feel natural and keeping the world consistent as you go.

Give it a try. If something feels off, tell us what you tried and what happened. That would help us a lot.

If you’d like to try more of PixVerse, feel free to DM me. Happy to share some extra trial credits.

Thanks for spending some time with it.

I wonder how much control creators have over keeping characters consistent as the world keeps evolving?

Great question, character consistency is definitely an important part of this. Let me check with our tech team on how much control R2 currently gives creators

​what are the system requirements or cloud specs needed to stream these sessions with zero lag?

Good question. Let me check with our tech team on the exact setup and I’ll get back to you here.

Nice launch, huge congrats 👏

Curious to see what creators build with this once they start pushing it beyond the demo worlds.

 Thank you Ristan, your support mean a lot to us!

Congrats on the launch, PixVerse team!

What interests me most about R2 is the chance to explore a generated world and change it as I go. That could be useful for interactive stories and early game prototypes. I’d be curious to see how well the world holds together as people move around and make changes. How long should a change, like a new character or a shift in the weather, stay consistent? A typical response time would also help set expectations.

Thank u Gene, I’m glad to hear you like it. Yes, it works for game prototypes. You’re spot on that consistency over time is one of the key challenges we’re working on with R2. The goal is for changes, whether it’s a new character, weather, or something in the environment, to persist naturally as you keep exploring the same world. Response time can vary depending on the scene and the change you’re asking for, and we’re actively working on making the interaction faster and more consistent. Would love to hear what you think if you get a chance to try it, especially how far you’re able to push the world before it starts to feel less coherent.

​shifting from static video generation to playable generative worlds opens up insane possibilities for indie game creators.

Yep, exactly. I keep thinking about indie game creators too, being able to jump into a world, mess with it, and see where it goes feels way more fun than starting from a blank scene

the "remembers what happened earlier and carries it forward" part is the actual hard problem here, most generative video demos reset context every clip. curious how far back the memory actually reaches before a long session starts drifting or forgetting earlier changes, is there a practical session length where consistency breaks down