AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Reasoning Model

23 article(s) found · Clear tag

1 weeks ago

In-Depth Review of Gemini 3.8 Live – Google's Native Real-Time Speech Dialogue Model

Gemini 3.8 Live is a series of native real-time speech dialogue models launched by Google, which includes two variants: Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. This series employs an en...

Speech AIReasoning ModelAI Tools
Sep 16, 2026Read more →
1 weeks ago

Xiaomi-CocktailASR-1: In-Depth Evaluation of a Target Speaker ASR Model Based on an End-to-End LLM Architecture

Xiaomi-CocktailASR-1 is Xiaomi's open-source Target Speaker ASR (TS-ASR) large model, designed using an end-to-end LLM architecture. It uses a reference speech as a speaker embedding prompt to accurat...

Speech AIReasoning ModelEmbedding & RAG
Sep 14, 2026Read more →
1 months ago

Mureka V9.5 – A New Generation AI Music Generation Model from Kunlunwanwei

Mureka V9.5 is a new generation AI music generation model launched by Kunlunwanwei, built upon its self-developed MusiCoT music reasoning framework. It first constructs a global musical structure befo...

Speech AIReasoning ModelBenchmark
Aug 18, 2026Read more →
2 months ago

VibeThinker-3B – Weibo's Open-Source 3-Billion-Parameter Dense Reasoning Model

VibeThinker-3B is a 3-billion-parameter dense reasoning model open-sourced by the AI team at Weibo. Built upon the Qwen2.5-Coder-3B base model, it undergoes an enhanced Spectrum-to-Signal post-trainin...

AI CodingReasoning ModelLLM
Jul 7, 2026Read more →
2 months ago

Leanstral 1.5 – Mistral AI's Open-Source Formal Verification Large Model

Leanstral 1.5 is an open-source formal verification large model from Mistral AI, deeply optimized for Lean 4 automated theorem proving. The model adopts a sparse mixture of experts (MoE) architecture ...

MultimodalAI CodingReasoning Model
Jul 6, 2026Read more →
2 months ago

GeneBench-Pro – OpenAI's Research-Grade Benchmark for Computational Biology

GeneBench-Pro is a research-grade benchmark developed by OpenAI, specifically designed to evaluate AI models' ability to handle judgment-intensive analysis in computational biology. The benchmark comp...

AI AgentReasoning ModelAI for Science
Jul 2, 2026Read more →
2 months ago

yuxinlu1 Gemma4-12B – Open-Source Coding & Agentic Model Series

yuxinlu1 Gemma4-12B is an open-source coding and Agentic model series fine-tuned by individual developer Lu Yuxin based on Google's Gemma 4 12B instruction model, comprising the V1 Code version and V2...

AI AgentAI CodingReasoning Model
Jul 1, 2026Read more →
2 months ago

Claude Sonnet 5 – Anthropic's Most Powerful Agent Model

Claude Sonnet 5 is the most capable agent model in Anthropic's Sonnet series. Its performance in benchmarks for agentic coding, terminal operations, browser search, and computer use approaches that of...

AI AgentAI CodingReasoning Model
Jul 1, 2026Read more →
3 months ago

Intern-S2-Preview – Shanghai AI Lab Open Scientific Multimodal LLM

Intern-S2-Preview is Shanghai AI Laboratory's open scientific multimodal LLM preview, delivering trillion-class scientific capability at 35B parameters. Its "general–specialist fusion" training pipeli...

AI AgentMultimodalReasoning Model
Jun 21, 2026Read more →
3 months ago

General365 – Meituan LongCat Team's Open General Reasoning Benchmark

General365 is an open general reasoning benchmark from Meituan's LongCat team, designed to evaluate large language models (LLMs) purely on logical reasoning in everyday scenarios. The benchmark includ...

Reasoning ModelLLMBenchmark
Jun 21, 2026Read more →
Prev123
Page 1 of 3 (23 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.