AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Video AI

82 article(s) found · Clear tag

1 months ago

JoyAI-Video-Edit – JD.com's Open-Source Real-Time Streaming Video Editing Model

JoyAI-Video-Edit is a real-time streaming video editing model developed and open-sourced by JD.com. It is based on a self-regressive diffusion architecture with 16B parameters, achieving end-to-end in...

Image GenerationVideo AIModel Inference
Aug 5, 2026Read more →
1 months ago

SeedRealtime – ByteDance's Native Audio-Video Full-Duplex Large Model

SeedRealtime is a native audio-video full-duplex large model introduced by ByteDance's Seed team. It integrates audio, video, and text within a unified architecture, enabling real-time, full-modal int...

Speech AIVideo AIBenchmark
Aug 5, 2026Read more →
1 months ago

MAGI-2-preview – Sand.ai's Open-Source Multimodal Video Generation Model

MAGI-2-preview is a multimodal video generation model developed and open-sourced by Sand.ai, utilizing a Mixture of Experts (MoE) architecture with a total parameter count of 114B, activating only 6B ...

MultimodalSpeech AIVideo AI
Aug 5, 2026Read more →
1 months ago

UniWorld-View – RabbitZoom Intelligence Collaborates with Peking University and Others to Open-Source a World Model

UniWorld-View is an open-source world model jointly developed by RabbitZoom Intelligence, Peking University, and the鹏城实验室 (Pengcheng Laboratory). It has topped the WorldScore world model evaluation le...

Video AIAI CodingBenchmark
Aug 1, 2026Read more →
1 months ago

MiniMax H3 – A General-Purpose Multimodal Generation Model from MiniMax

MiniMax H3 is a general-purpose multimodal generation model officially released by MiniMax on July 31, 2026. This model breaks the boundaries between traditional tasks and modalities, achieving unifie...

MultimodalSpeech AIVideo AI
Aug 1, 2026Read more →
2 months ago

MAI-Voice-2-Flash – Microsoft's High-Speed Text-to-Speech Model

MAI-Voice-2-Flash is a high-speed text-to-speech (TTS) model introduced by Microsoft's AI team for high-concurrency, low-latency scenarios. While maintaining natural tone and high audio quality, its i...

Speech AIVideo AIModel Inference
Jul 28, 2026Read more →
2 months ago

DramaClaw – Industrial-Grade AI Video Production Tool, Offering an End-to-End Pipeline

DramaClaw is an industrial-grade AIGC video production tool designed for scenarios such as AI short dramas, animated series, and novel adaptations. It is built upon the self-developed Ling Shan AI Dir...

Video AIAI Tools
Jul 20, 2026Read more →
2 months ago

HyOCR-1.5 – Tencent Hunyuan's Lightweight End-to-End OCR Expert Model Open Sourced

HyOCR-1.5 is a lightweight end-to-end Optical Character Recognition (OCR) expert large model introduced by the Tencent Hunyuan team. With only 1B parameters, it integrates full-stack capabilities incl...

Video AIAI CodingDocument AI
Jul 20, 2026Read more →
2 months ago

PixVerse Game – The First Real-Time Video Game Engine from PixVerse Tech

PixVerse Game is the first real-time video game engine introduced by PixVerse Tech, built upon its self-developed PixVerse R1 real-time world model. This engine constructs a complete gaming experience...

Video AI
Jul 20, 2026Read more →
2 months ago

Wan-Dancer – The Open-Source Human Figure Dance Video Generation Model from Alibaba Tongyi Wanxiang

Wan-Dancer is an open-source music-driven human figure dance video generation model developed by Alibaba Tongyi Wanxiang. Users only need to provide a single正面 (front-facing) portrait photo and a piec...

Video AI
Jul 20, 2026Read more →
Page 5 of 9 (82 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.