Overview
Baidu’s unified multimodal foundation model jointly modeling text, images, audio, and video, first announced in November 2025.
Baidu’s unified multimodal foundation model jointly modeling text, images, audio, and video, first announced in November 2025.
Only published specifications are shown.
Baidu’s unified multimodal foundation model jointly modeling text, images, audio, and video, first announced in November 2025.
ERNIE 5.0 was announced.
View sourceERNIE 5.0-0110 dropped its Preview label on LMArena; this does not establish same-day availability across every channel.
View sourceSelect a card to view the related model.