AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Video AI

82 article(s) found · Clear tag

2 weeks ago

ffmpeg-skill: Open-Source Toolkit for Programming Agents to Achieve Professional Local Video Editing Capabilities

ffmpeg-skill is an open-source Agent Skill designed for programming agents (such as Claude Code, Cursor, Codex). Its core value lies in enabling AI to invoke local FFmpeg just like a professional vide...

AI AgentVideo AIAI Coding
Sep 10, 2026Read more →
2 weeks ago

Ling-3.0-flash-VL – Ant Group's Open-Source Native Multimodal Large Model

Ling-3.0-flash-VL is the first open-source native multimodal large model in Ant Group's InclusionAI Bailing series. It is an extension of the MoE architecture from Ling-3.0-flash, with a total paramet...

AI AgentMultimodalVideo AI
Sep 9, 2026Read more →
2 weeks ago

H3-World – A World Model with Action Controllability Developed by Tencent in Collaboration with Universities

H3-World is an action-controllable world model developed collaboratively by Tencent, National University of Singapore, and The Hong Kong Polytechnic University. It is based on the MiniMax-H3 video gen...

Video AIEmbedding & RAGAI Safety
Sep 9, 2026Read more →
3 weeks ago

Orbis – Visko's Real-Time World Model That Makes AI Video Streaming Like Live Broadcasting

Orbis is Visko's first Live Model (real-time world model), fundamentally changing the way AI videos are generated—no longer requiring minutes or even tens of minutes of offline rendering, but instead ...

Video AI
Sep 5, 2026Read more →
3 weeks ago

MiniMax H3 Max: Live-level Speed and Ecosystem Evolution in Real-time Video Generation

MiniMax H3 Max is a real-time video generation model introduced by MiniMax, based on the open-source H3 model, with post-training and inference optimization. This model supports two input methods: tex...

Speech AIVideo AIModel Inference
Sep 4, 2026Read more →
3 weeks ago

E-Commerce Bench – A Long-Term E-Commerce Business Evaluation Benchmark Jointly Open-Sourced by Qwen and Taotian

E-Commerce Bench is a long-term e-commerce business evaluation benchmark jointly open-sourced by Alibaba's Qwen and Taotian Group. This benchmark enables LLM Agents to operate an online store in a sim...

AI AgentVideo AILLM
Sep 4, 2026Read more →
3 weeks ago

Atlas – The World's First Multimodal World Model from World Labs

Atlas is the world's first multimodal world model introduced by World Labs, founded by Fei-Fei Li. This model natively understands text, images, videos, and 3D spatial information. By anchoring visual...

MultimodalVideo AIEmbodied AI
Sep 4, 2026Read more →
3 weeks ago

MiniMax H3 Max: Real-Time Video Generation Model Jointly Released by MiniMax and fal.ai

MiniMax H3 Max is a real-time video generation model jointly released by MiniMax and fal.ai. Its main selling point is speed: it can generate a 768p video from a 5-second input in under 3 seconds, whi...

Video AIAI CodingBenchmark
Sep 1, 2026Read more →
3 weeks ago

WeMM-Embedding: Tencent Open-Sources a General Multimodal Embedding Model

Tencent's WeChat Vision Team has open-sourced WeMM-Embedding, which encodes text, images, and videos into a unified vector space. Teams working on content retrieval, recommendation recall, or agent me...

AI AgentMultimodalVideo AI
Sep 1, 2026Read more →
3 weeks ago

MiniMax H3 Max: fal's Post-trained Real-time Video Model — Faster than Real-time, No Benchmarks Published

H3 Max is a hosted service built on the open-source H3 base model. It suits teams that want to integrate video generation into real-time content production workflows, as well as individual users looki...

MultimodalImage GenerationVideo AI
Aug 31, 2026Read more →
Page 2 of 9 (82 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.