AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: AI Agent

224 article(s) found · Clear tag

2 months ago

Ling 3.0 Flash – A Lightweight MoE Inference Model Launched by Ant Bailing

Ling 3.0 Flash is a lightweight Mixture-of-Experts (MoE) inference model introduced by the Bailing large model team at Ant Group. With a total parameter count of 124B, it activates only 5.1B parameter...

AI AgentAI CodingDocument AI
Jul 28, 2026Read more →
2 months ago

WorkBuddy Bench – Tencent's Open-Source Code Agent Evaluation Suite

WorkBuddy Bench is an open-source evaluation suite for code agents developed by Tencent, designed to provide standardized capability assessments for AI agents in four real-world scenarios: code, web d...

AI AgentAI CodingAI Safety
Jul 28, 2026Read more →
2 months ago

MineExplorer – Meituan's Open-World Minute-Level Long-Horizon Task Evaluation Benchmark

MineExplorer is the first open-world minute-level long-horizon task evaluation benchmark introduced by Meituan's LongCat team, based on Minecraft. This benchmark includes 813 manually verified instanc...

AI AgentAI CodingBenchmark
Jul 28, 2026Read more →
2 months ago

OpenWorker – Andrew Ng's Open-Source, Free, Local-First AI Desktop Agent

OpenWorker is an open-source AI desktop agent released by Andrew Ng, fully free under the MIT license. It is not a traditional chatbot, but rather a local-first AI colleague oriented toward "deliverin...

AI AgentDocument AIOpen Source
Jul 28, 2026Read more →
2 months ago

last30days-skill – Open-Source Cross-Platform AI Agent for Real-Time Comment Research

last30days-skill is an open-source AI Agent research skill that enables agents to automatically scrape real discussions from major overseas social platforms over the past 30 days and synthesize them i...

AI AgentAI Tools
Jul 28, 2026Read more →
2 months ago

Nanbeige4.2-3B – A General-Purpose Agent Small Model from Nanbeige Lab

Nanbeige4.2-3B is a general-purpose agent small model from BOSS Zhipin’s Nanbeige Lab. With only 3B parameters, it surpasses larger models such as Qwen3.5-9B and Gemma4-12B on code agents, office work

AI AgentAI CodingModel Inference
Jul 27, 2026Read more →
2 months ago

Macaron-V1 – MindLab's Post-training Multi-capability Model System

Macaron-V1 is a post-training multi-capability model system introduced by MindLab, built upon the GLM 5.2 base model. It employs LoRA fine-tuning on approximately 0.5% of core parameters through reinf...

AI AgentAI CodingAI Tools
Jul 24, 2026Read more →
2 months ago

OpenAI Presence – Enterprise AI Agent Platform for Trusted Voice and Chat Workflows

OpenAI Presence is OpenAI’s enterprise platform for deploying and operating governed AI agents in production. Announced on July 22, 2026, it targets high-volume, high-stakes workflows such as customer

AI AgentSpeech AIAI Safety
Jul 24, 2026Read more →
2 months ago

CodeBuddy NPC – Tencent Cloud's AI Agent for Enterprise R&D Workflows

CodeBuddy NPC is a cloud-based AI agent introduced by Tencent Cloud, designed specifically for enterprise R&D workflows. It deeply integrates with the CNB platform, actively participating in code repo...

AI AgentAI CodingTool Calling
Jul 24, 2026Read more →
2 months ago

Step Edge – The Edge Model Suite from StepFusion

Step Edge is StepFusion's edge model suite, comprising four core components: Base, Audio, GUI, and Gen, designed for mobile phones and automotive terminals. Through an edge-cloud collaborative archite...

AI AgentSpeech AIModel Inference
Jul 20, 2026Read more →
Page 10 of 23 (224 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.