AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Video AI

82 article(s) found · Clear tag

4 weeks ago

Zing-0.5 – Loopit's Open-Source Real-Time Interactive Video World Model

Zing-0.5 is an open-source real-time interactive video world model developed by the Loopit team, with 5B parameters. It is trained on the Wan2.2-TI2V-5B base model and released under the Apache 2.0 li...

Video AIModel InferenceBenchmark
Aug 29, 2026Read more →
4 weeks ago

shuohao-skills – Open-Source AI Short Film Production Skill Collection, Featuring a Complete Workflow

shuohao-skills is an open-source collection of AI skills specifically designed for agents like Claude Code and Codex. The tool provides a complete workflow for short film production, ranging from nove...

AI AgentSpeech AIVideo AI
Aug 28, 2026Read more →
4 weeks ago

Gemini Omni 1.1 Flash — Google's Multi-Tier AI Video Generation Model for Professional Creators

Gemini Omni 1.1 Flash is an AI video generation model launched by Google, targeting developers and professional creators. It supports multi-tier output ranging from 360p quick previews to 4K ultra-hig...

Video AIAI Tools
Aug 28, 2026Read more →
1 months ago

EgoSuite-Open100K – Guanglun Intelligence's Open-Source Multimodal Human Behavior Dataset

EgoSuite-Open100K is Guanglun Intelligence's globally first open-source multimodal human behavior dataset at the 100,000-hour scale, aimed at the fields of physical AI and embodied intelligence. The d...

MultimodalVideo AIEmbodied AI
Aug 26, 2026Read more →
1 months ago

In-Depth Review of Breeze TTS 2 – BreezeBlue's Leading Text-to-Speech Model

Breeze TTS 2 is a next-generation text-to-speech (TTS) model launched by BreezeBlue. It supports zero-shot character voice design through natural language and precisely controls performance details su...

AI AgentSpeech AIVideo AI
Aug 26, 2026Read more →
1 months ago

PixVerse R2 – A Real-Time Multimodal World Model from Aise Tech

PixVerse R2 is a real-time multimodal world model launched by Aise Tech, an upgraded version of PixVerse R1. The model supports multimodal inputs such as text, images, audio, and action signals, and c...

MultimodalSpeech AIVideo AI
Aug 26, 2026Read more →
1 months ago

OpenStory: Open Source AI Video Production Platform, Generate Style-Consistent Short Drama Videos with One Click

OpenStory is an open-source AI video production platform that specializes in converting scripts into style-consistent short drama videos with a single click. The platform can automatically split scene...

Speech AIVideo AIDocument AI
Aug 25, 2026Read more →
1 months ago

CoinVE-200K – Tencent's Open-Source Large-Scale Composite Instruction Video Editing Dataset

CoinVE-200K is a large-scale composite instruction video editing dataset open-sourced by Tencent Video's Intelligent Creation Team. It contains 200,000 pairs of 1080p high-definition video editing sam...

Video AIBenchmark
Aug 22, 2026Read more →
1 months ago

Mage – Microsoft's Open-Source 4B Parameter Multimodal Model Family

Mage is a family of open-source 4B parameter multimodal models developed by Microsoft, consisting of the streaming video/image understanding model Mage-VL and the image generation and editing model Ma...

MultimodalVideo AIAI Coding
Aug 19, 2026Read more →
1 months ago

Gemini 3.7 Flash – Google DeepMind Launches Its Leading AI Model

Gemini 3.7 Flash is a new generation of leading AI model introduced by Google DeepMind, specifically designed for coding and Agent workflows. This model achieves significant performance improvements i...

AI AgentVideo AIAI Coding
Aug 14, 2026Read more →
Page 3 of 9 (82 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.