AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Model Inference

66 article(s) found · Clear tag

2 days ago

In-Depth Review of Spark-ASR-2.0: A New Paradigm in Speech Recognition with Non-Autoregressive Architecture

Spark-ASR-2.0 is the latest generation speech recognition large model launched by iFLYTEK based on its proprietary Spark-Audio speech foundation model. This model continues the non-autoregressive para...

Speech AIAI CodingModel Inference
Sep 25, 2026Read more →
1 weeks ago

In-Depth Review of Step 5 Preview: A 600B Sparse MoE Flagship with 1M Token Context and 1/8 Cost Advantage

Step 5 Preview is a new-generation flagship foundation model launched by StepFun, designed for real-world Agentic tasks. Based on a sparse MoE architecture, the model has a total of 600B parameters bu...

AI AgentMultimodalModel Inference
Sep 20, 2026Read more →
1 weeks ago

Review: GLM-5.3-FlashX — Zhipu AI's High-Speed Inference Model, Setting a New Benchmark for Real-Time Interaction at 200 tokens/s

GLM-5.3-FlashX is a high-speed inference model launched by Zhipu AI in 2026, serving as an accelerated upgrade of GLM-5.3-Flash. Its core selling point lies in its maximum output speed of up to 200 to...

AI AgentModel InferenceBenchmark
Sep 18, 2026Read more →
1 weeks ago

Open-RAIL Evaluation: China Mobile's Open-Source General-Purpose Engineering Foundation for Embodied Intelligence, Bridging the "Last Mile" for VLA/WAM Model Deployment

Open-RAIL is a general-purpose engineering foundation for embodied intelligence that China Mobile has open-sourced globally. It is positioned as the industry's first universal "nervous system" connect...

AI CodingEmbodied AIModel Inference
Sep 17, 2026Read more →
1 weeks ago

Step Audio 3 – StepFun's Voice Large Model Series

Step Audio 3 is a new generation of voice large model series introduced by StepFun. It derives five specialized models—Realtime, ASR, TTS, Gen, and Music—from a single technical foundation, covering t...

Speech AIModel InferenceBenchmark
Sep 15, 2026Read more →
2 weeks ago

VDN-MiniMax-H3 – OpenVDN's Open-Source Video Generation Acceleration Solution

VDN-MiniMax-H3 is an open-source video generation acceleration solution developed by the OpenVDN team, based on architectural modifications and deep optimization of the MiniMax-H3 model. This solution...

Video AIModel InferenceEdge Deployment
Sep 10, 2026Read more →
2 weeks ago

LLaDA-Image – A Unified Image Generation and Editing Model Open-Sourced by Ant Group

LLaDA-Image is a 6B parameter unified image generation and editing model open-sourced by the inclusionAI Lab at Ant Group. This model adopts an innovative training approach, first pre-training purely ...

Image GenerationModel InferenceBenchmark
Sep 8, 2026Read more →
3 weeks ago

MiniMax H3 Max: Live-level Speed and Ecosystem Evolution in Real-time Video Generation

MiniMax H3 Max is a real-time video generation model introduced by MiniMax, based on the open-source H3 model, with post-training and inference optimization. This model supports two input methods: tex...

Speech AIVideo AIModel Inference
Sep 4, 2026Read more →
3 weeks ago

In-Depth Review of Claude Fable 5.1 – Anthropic's Flagship Large Model

Claude Fable 5.1 is a flagship large model launched by Anthropic, positioned as one of the high-performance AI systems available to the public. Designed for complex reasoning, long-running Agent tasks...

AI AgentModel InferenceBenchmark
Sep 4, 2026Read more →
3 weeks ago

Claude Mythos 5.1: Anthropic's Flagship Model with Controlled Access for High-Risk Research Fields

Claude Mythos 5.1 is the flagship model of the Claude 5.1 series launched by Anthropic. It shares the exact same underlying model weights and inference capabilities with the publicly available Claude ...

AI for ScienceAI SafetyModel Inference
Sep 4, 2026Read more →
Prev12345...7
Page 1 of 7 (66 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.