Google Launches New Model That ‘Generates Worlds’ from Text

Google DeepMind, the tech giant’s AI research lab, announced the launch of Genie 3, which can create 3D environments at 720p resolution from text prompts.

This level of quality in the new Genie model is an upgrade from Genie 2, which operated at 10 to 20 frames per second instead of Genie 3’s 24 frames per second.

Genie 3’s environment generations can stay physically consistent over time because the model remembers what was previously generated.

Using the model, users will be able to infinitely adapt virtual worlds for industries such as gaming, robotics, training, disaster preparedness, education, and for creating immersive experiences.

A First Step Towards AGI

One of the standout features in the product is real-time interactivity – which enables users to change virtual worlds in real-time, and dynamic environment modification – which allows users to reshape virtual worlds continuously.

Google says the Genie 3 model could be used to train robots and autonomous vehicles to interact with realistic environments.

The company also highlights the advent of the model being a first step to achieving artificial general intelligence (AGI) – an advanced form of AI that would be able to reason and carry out tasks at the same level of humans.

This is because the model could change the way AI interacts with and understands the world.

During a brief, Jack Parker-Holder, a research scientist at DeepMind, said: “We think world models are key on the path to AGI, specifically for embodied agents, where simulating real world scenarios is particularly challenging.”

Subscribe to our newsletter for updates

Join thousands of media and marketing professionals by signing up for our newsletter.

"*" indicates required fields

This field is for validation purposes and should be left unchanged.

Share

Related Posts

Popular Articles

Featured Posts

Menu