AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Video AI

82 article(s) found · Clear tag

2 months ago

ABot-World Studio – A General-Purpose World Model Workshop Launched by AutoNavi

ABot-World Studio is a general-purpose world model workshop launched by AutoNavi Maps. It innovatively unifies interactive video generation with 3DGS (3D Gaussian Splatting) scene generation within th...

Video AIModel InferenceEdge Deployment
Jul 20, 2026Read more →
2 months ago

MOSS-VL-Realtime – OpenMOSS's Open-Source Vision-Language Model

MOSS-VL-Realtime is an open-source 11B parameter streaming vision-language model developed by OpenMOSS, specifically designed for real-time video understanding. The model supports answering while watc...

Video AIAI CodingBenchmark
Jul 20, 2026Read more →
2 months ago

Xiaomi-Robotics-U0 – Xiaomi's Unified Embodied Synthesis Model

Xiaomi-Robotics-U0 is Xiaomi's unified embodied synthesis model with 38 billion parameters, trained continuously on the world foundation model, and jointly optimized for five major tasks: text-to-imag...

Image GenerationVideo AIEmbodied AI
Jul 20, 2026Read more →
2 months ago

X2.0 – Xmax AI Launches the World's First Real-Time Interactive Video Generation Model

X2.0 is the world's first real-time interactive video generation model introduced by Xmax AI, supporting millisecond-level streaming generation. Users can replace character clothing in real time via a...

Video AIModel Inference
Jul 20, 2026Read more →
2 months ago

Qwen-Audio-3.0-Realtime – Alibaba's Real-Time Speech Interaction Model

Qwen-Audio-3.0-Realtime is a new generation of real-time speech interaction dialogue model introduced by Alibaba Cloud's Tongyi team, offering Plus and Flash versions. It maintains high inference dept...

AI AgentSpeech AIVideo AI
Jul 20, 2026Read more →
2 months ago

Wan-Streamer v0.2 – A Full-Modal Understanding and Generation Model from Alibaba Tongyi

Wan-Streamer v0.2 is an end-to-end full-modal understanding and generation model introduced by Alibaba Tongyi Lab, designed for real-time full-duplex interaction. This model processes real-time unders...

Speech AIVideo AILLM
Jul 20, 2026Read more →
2 months ago

AudioX-Turbo – A Unified and Efficient Audio Generation Framework Jointly Released by Noiz AI and Tsinghua University

AudioX-Turbo is a unified and efficient audio generation framework jointly developed by Noiz AI, the Hong Kong University of Science and Technology, and Tsinghua University. Based on a multimodal diff...

MultimodalSpeech AIVideo AI
Jul 7, 2026Read more →
2 months ago

LiveWorld – Generative Video World Model from University of Adelaide and Others

LiveWorld is a generative video world model jointly developed by the University of Adelaide, the Australian National University, and other institutions. Its core focus is solving the problem of out-of...

MultimodalVideo AIEmbodied AI
Jul 1, 2026Read more →
2 months ago

Wan-Streamer – Alibaba's Open-Source Real-Time Full-Duplex Multimodal Foundation Model

Wan-Streamer is an end-to-end real-time full-duplex multimodal foundation model open-sourced by Alibaba DAMO Academy. It unifies text, audio, and video input/output tokens into a single causal sequenc...

MultimodalSpeech AIVideo AI
Jun 30, 2026Read more →
2 months ago

Agent-Reach – Open Source AI Agent Tool for One-Click Internet Content Retrieval

Agent-Reach is an open-source, free AI Agent internet capability scaffolding tool designed to install web access for mainstream AI Agents such as Claude Code, Cursor, and OpenClaw with a single natura...

AI AgentVideo AIAI Coding
Jun 29, 2026Read more →
Page 6 of 9 (82 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.