AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Latest Articles

Total 467 articles

1 months ago

Hunyuan3D-Buffalo 1.0 – Tencent Hunyuan's Unified 3D Multimodal Framework

Hunyuan3D-Buffalo 1.0 is a unified 3D multimodal framework introduced by the Tencent Hunyuan team. It integrates multiple tasks—such as 3D question answering, spatial localization, text-to-3D generati...

MultimodalTool CallingEdge Deployment
Aug 6, 2026Read more →
1 months ago

InstructAV2AV – An Open-Source Audio-Visual Joint Editing Model Developed by BAAI and Peking University

InstructAV2AV is an open-source audio-visual joint editing model jointly developed by the Beijing Academy of Artificial Intelligence (BAAI) and Peking University. With just a single natural language i...

Speech AIVideo AIBenchmark
Aug 6, 2026Read more →
1 months ago

JoyAI-Video-Edit – JD.com's Open-Source Real-Time Streaming Video Editing Model

JoyAI-Video-Edit is a real-time streaming video editing model developed and open-sourced by JD.com. It is based on a self-regressive diffusion architecture with 16B parameters, achieving end-to-end in...

Image GenerationVideo AIModel Inference
Aug 5, 2026Read more →
1 months ago

SeedRealtime – ByteDance's Native Audio-Video Full-Duplex Large Model

SeedRealtime is a native audio-video full-duplex large model introduced by ByteDance's Seed team. It integrates audio, video, and text within a unified architecture, enabling real-time, full-modal int...

Speech AIVideo AIBenchmark
Aug 5, 2026Read more →
1 months ago

MAGI-2-preview – Sand.ai's Open-Source Multimodal Video Generation Model

MAGI-2-preview is a multimodal video generation model developed and open-sourced by Sand.ai, utilizing a Mixture of Experts (MoE) architecture with a total parameter count of 114B, activating only 6B ...

MultimodalSpeech AIVideo AI
Aug 5, 2026Read more →
1 months ago

Shieldstral – Mistral AI's Open-Source Multimodal Content Safety Classification Model

Shieldstral is an open-source 3B parameter multimodal content safety classification model launched by Mistral AI, built upon the Ministral-3B foundation. This model redefines traditional fixed-categor...

MultimodalEmbedding & RAGBenchmark
Aug 5, 2026Read more →
1 months ago

Hy ASR 3.0 Preview – The New Generation Speech Recognition Model from Tencent Hunyuan

Hy ASR 3.0 Preview is a new generation speech recognition model launched by Tencent Hunyuan, built upon the Hy3 large language model and employing a Mixture of Experts (MoE) architecture. It integrate...

Speech AILLMBenchmark
Aug 5, 2026Read more →
1 months ago

MemHarness – A Memory Reconstruction Framework for LLM Agents Introduced by the Shanghai AI Lab and Others

MemHarness is a memory reconstruction framework for LLM Agents introduced jointly by the Shanghai Artificial Intelligence Lab and universities such as Zhejiang University, Fudan University, and Shangh...

AI AgentEmbedding & RAGLLM
Aug 5, 2026Read more →
1 months ago

Qwen-CUA – The Native Computer Use Agent Introduced by Alibaba Qwen and Others

Qwen-CUA is a native Computer Use Agent introduced jointly by the Qwen team and XLang Lab. Based on a 397B-A17B Mixture-of-Experts (MoE) architecture, it perceives the interface state solely through s...

AI AgentLLMBenchmark
Aug 5, 2026Read more →
1 months ago

Orchard – Microsoft Research's Open-Source Agentic AI Modeling Framework

Orchard is an open-source Agentic AI modeling framework introduced by Microsoft Research. At its core is the Kubernetes-based Orchard Env environment service, which enables cross-domain reuse of sandb...

AI AgentModel InferenceBenchmark
Aug 4, 2026Read more →
Page 18 of 47 (467 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.