11/17/2025 AI Express | What's New in AI: MiroThinker Open Source, Gemini and Grok Features Upgraded

MiroMind team released the open source bAgent model MiroThinker v1.0, proposing the concept of "Deep Interaction Scaling". Google for ...

young dragon
November 18, 2025

Baidu MuseSteamer in-depth analysis: a new milestone in domestic AI video generation

Baidu's commercial R&D team launched MuseSteamer, a multimodal generative large model, which achieved the world's first place in the VBench graph-generated video review, and synchronized audio and video in Chinese...

young dragon
July 5, 2025

Qwen-VLo: A major release in the field of multimodal AI from AliCloud

AliCloud recently released its latest multimodal AI model, Qwen-VLo, whose image generation and editing capabilities were highly rated by users and even surpassed GPT-4o. The model has a detailed...

young dragon
July 3, 2025

OmniGen2: A breakthrough in next-generation multimodal AI

OmniGen2 is a multimodal generative model based on the Qwen-VL-2.5 architecture with 7 billion parameters, of which 3 billion are used for text processing and 4 billion for...

young dragon
July 3, 2025

Google Gemini 2.5 Pro: a multimodal evolution from video to interactive apps

Google releases Gemini version 2.5 Pro, a major realization in the field of multimodal understanding and code generation. The model surpasses competitor Cl ...

young dragon
May 7, 2025