❌

Reading view

Reka AI's omni-model Rho-1 handles text, images, video, and robot control in a single model

Reka AI's Rho-1 is a 19-billion-parameter omni-model that processes and generates text, images, video, and robot control actions in a single neural network. Trained on 320 H100 GPUs in about three months, it uses a fraction of the compute today's top models need. Instead of routing tasks to specialized systems, Rho-1 runs all modalities as tokens in one shared context window.

The article Reka AI's omni-model Rho-1 handles text, images, video, and robot control in a single model appeared first on The Decoder.

πŸ’Ύ

  •  

Runway moves into robotics with open-weight Praxis-1 AI model

Runway, one of the best-known developers of generative AI video technology, is moving into robotics with Praxis-1, an open-weight AI model designed to use knowledge learned from video to control physical robots. β€œOpen-weight” means developers can download the model and run or adapt it themselves, rather than only accessing it through Runway’s servers. The company […]
  •  

How to Make Your First World Model from Scratch

A beginner-friendly guide to building a world model in Python, letting it daydream its way through CartPole, and accurately measuring when the illusion collapses.

The post How to Make Your First World Model from Scratch appeared first on Towards Data Science.

  •  

LTX launches new free-to-use open world model for video and physical AI

LTX has launched LTX-2.5, the latest version of its open-weights world model, introducing new capabilities for video generation, real-time applications and physical AI. The company says the new model delivers improvements in visual quality, prompt understanding, generation speed and efficiency, while allowing developers and enterprises to run and customize the model on their own hardware. […]
  •  
❌