레카AI가 텍스트, 이미지, 영상, 로봇 제어 행동을 하나의 신경망에서 처리하고 생성하는 190억 파라미터 옴니모델 '로-1(Rho-1)'을 공개했다. 320개의 H100 GPU로 약 3개월간 학습했으며, 현재 최상위 모델들이 필요로 하는 연산량의 일부만 사용한다. 특정 작업을 전문 시스템으로 분기하는 대신, 로-1은 모든 모달리티를 하나의 공유 컨텍스트 윈도우 안에서 토큰으로 처리한다.
- •레카AI, 텍스트·이미지·영상·로봇제어를 한 모델에서 처리하는 190억 파라미터 옥니모델 '로-1' 공개
- •320개 H100 GPU로 약 3개월 학습, 최상위 모델 대비 적은 연산량 사용
- •모든 모달리티를 하나의 공유 컨텍스트 윈도우에서 토큰으로 통합 처리
0단 자동
AI가 규칙대로 쓰고 그대로 게시했습니다. 사람이 따로 보지 않았습니다.
- 규칙 판
- 규칙 판 도입 이전 기사입니다.
- 남기는 것
- 규칙 판 · 모델 · 시각
- 판 기록
- 아직 없습니다.
Reka AI's omni-model Rho-1 handles text, images, video, and robot control in a single model
본문 미리보기
Reka AI's Rho-1 is a 19-billion-parameter omni-model that processes and generates text, images, video, and robot control actions in a single neural network. Trained on 320 H100 GPUs in about three months, it uses a fraction of the compute today's top models need. Instead of routing tasks to specialized systems, Rho-1 runs all modalities as tokens in one shared context window. The article Reka AI's omni-model Rho-1 handles text, images, video, and robot control in a single model appeared first on
전체 내용이 궁금하다면?
원문을 직접 읽어보세요
이 글이 만들어진 과정
- 10:33AI 초안

