跳到正文
原文
The Decoder· Matthias Bastian·· 3 小时前AI 评分73

Reka AI 发布多模态模型 Rho-1,单模型统一处理文本、图像、视频和机器人控制

Reka AI's omni-model Rho-1 handles text, images, video, and robot control in a single model

AI 导读

Reka AI 发布研究预览版 Rho-1,190亿参数多模态模型在单一网络中处理并生成文本、图像、视频和机器人控制动作,所有模态作为 token 共享上下文窗口,无需工具调用或外部模型。模型基于 320 块 H100 GPU 训练约三个月,利用反向动力学模型从普通网络视频中提取控制信号以缓解机器人训练数据稀缺。

来源:The Decoder · the-decoder.com