샤오미가 MiMo-V2.6을 공개하며, 멀티모달 모델과 강화학습 과정을 선보였다.
샤오미는 MiMo-V2.6을 발표하며 텍스트, 이미지, 영상, 오디오를 함께 이해하는 멀티모달 모델을 공개했다. 이 모델은 코딩 및 도구 사용, 컴퓨터 조작을 수행할 수 있다. Pro 모델은 46.32점의 인공지능 지능 지수로, 공개된 모델 중 1위를 기록하며 API 가격도 공개되었다.
Xiaomi unveiled MiMo-V2.6, showcasing a multimodal model that utilizes reinforcement learning.
Xiaomi has announced MiMo-V2.6, introducing a multimodal model capable of understanding text, images, videos, and audio. This model performs tasks such as coding and tool usage, alongside computer manipulation. The Pro variant achieved an artificial intelligence score of 46.32, ranking first among public models, and its API pricing has also been revealed.