AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Video AI

82 article(s) found · Clear tag

3 months ago

Xiaomi OneVL – Open-Source Autonomous Driving Model from Xiaomi Embodied Intelligence

Xiaomi OneVL is an open-source autonomous driving model from Xiaomi's embodied intelligence team—the first framework to unify VLA (vision-language-action), world modeling, and latent-space reasoning i...

MultimodalVideo AIAI Coding
Jun 21, 2026Read more →
3 months ago

OpenMontage – Open-Source Agentic Video Production System

OpenMontage is the first open-source agentic video production system, orchestrated by AI coding assistants from concept to final cut. It integrates 12 production pipelines, 52 professional tools, and ...

AI AgentVideo AIAI Coding
Jun 21, 2026Read more →
3 months ago

MiniCPM-V 4.6 – OpenBMB's Open-Source Edge Multimodal LLM

MiniCPM-V 4.6 is an edge multimodal large language model from ModelBest (OpenBMB) with a 1.3B-parameter LLM backbone, deeply optimized for on-device deployment on mobile hardware. Built on the llama.c...

MultimodalVideo AIDocument AI
Jun 21, 2026Read more →
3 months ago

Violin – Oxford's Kevin Lin Open-Source End-to-End AI Video Translation Tool

Violin is an open-source end-to-end AI video translation tool led by Oxford postdoc Kevin Lin, built to break language barriers for high-quality video. It combines OpenAI Whisper for speech recognitio...

MultimodalSpeech AIVideo AI
Jun 21, 2026Read more →
3 months ago

Lance – ByteDance's Lightweight Native Unified Multimodal Model

Lance is a lightweight native unified multimodal model open-sourced by ByteDance's Intelligent Creation team. With only 3B active parameters, it supports the full pipeline of image and video understan...

AI AgentMultimodalVideo AI
Jun 21, 2026Read more →
3 months ago

Xiaomi Auto World Model – Xiaomi's Assisted-Driving World Model

Xiaomi Auto's assisted-driving world model (Xiaomi Auto World Model) is the first to deeply couple 3D reconstruction (WorldRec) with video generation (WorldGen) into one driving-scene understanding an...

Video AIModel Inference
Jun 21, 2026Read more →
3 months ago

WBench – Meituan’s Interactive Video World Model Multi-Turn Benchmark

WBench is Meituan LongCat’s first systematic multi-turn benchmark for interactive video world models—289 test cases, 1,058 interaction rounds, six scene types (nature, city, indoor, workspace, fantasy...

Video AIModel InferenceBenchmark
Jun 21, 2026Read more →
3 months ago

Toonflow – Open-Source All-in-One AI Short-Drama Creation Tool

Toonflow is an open-source, multi-agent platform that turns novels or creative text into structured scripts, storyboards, character visuals, and animated video—full pipeline from words to finished sho...

AI AgentVideo AIAI Coding
Jun 21, 2026Read more →
3 months ago

Seedance 2.0 Mini – ByteDance's Lightweight Video Generation Model

Seedance 2.0 Mini is a cost-efficient lightweight video generation model from ByteDance Volcano Engine for high-frequency short-video production, marketing asset iteration, and early-stage drafts. It ...

MultimodalVideo AIModel Inference
Jun 21, 2026Read more →
3 months ago

SCAIL-2 – Zhipu AI and Tsinghua's Open-Source Character Animation Model

SCAIL-2 is the second-generation film-grade character animation framework open-sourced jointly by Zhipu AI and Professor Liu Yongjin's research group at Tsinghua University. Built on a Diffusion Trans...

Video AIEmbodied AIOpen Source
Jun 21, 2026Read more →
Page 7 of 9 (82 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.