AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: AI Agent

224 article(s) found · Clear tag

3 months ago

Qwen3.7-Plus – Alibaba Tongyi's Agentic Multimodal Large Model

Qwen3.7-Plus is Alibaba Cloud Tongyi Qwen team's next-generation multimodal large model positioned as an "agent foundation," integrating visual perception, language understanding, code generation, and...

AI AgentMultimodalAI Coding
Jun 21, 2026Read more →
3 months ago

Ponytail – Open-Source AI Agent Plugin That Slashes Generated Code Volume

Ponytail is an open-source AI Agent code-minimization plugin. It injects a “senior lazy developer” minimalist mindset into 10+ mainstream AI coding tools—including Claude Code, Codex, and Cursor—forci...

AI AgentAI CodingAI Safety
Jun 21, 2026Read more →
3 months ago

Polar – NVIDIA's Open-Source Agentic Reinforcement Learning Training Framework

Polar is NVIDIA's open-source agentic reinforcement learning (RL) training framework whose core innovation lets existing agent frameworks plug into GRPO and other RL algorithms without modifying inter...

AI AgentAI CodingLLM
Jun 21, 2026Read more →
3 months ago

PlanningBench – Open LLM Planning Evaluation Framework by Tencent Hunyuan and Partners

PlanningBench is an open framework from Tencent Hunyuan with Renmin University of China Gaoling School of Artificial Intelligence and partners, focused on evaluating and training large language model ...

AI AgentLLMBenchmark
Jun 21, 2026Read more →
3 months ago

PilotDeck – Tsinghua and ModelBest's Open-Source Agent Operating System

PilotDeck is a next-generation agent operating system jointly open-sourced by Tsinghua University's NLP Lab (THUNLP), ModelBest, OpenBMB, and AI9stars. Centered on the WorkSpace paradigm, it establish...

AI AgentOpen SourceEdge Deployment
Jun 21, 2026Read more →
3 months ago

PawBench – Tongyi Lab's General Agent Evaluation Benchmark

PawBench is a general agent evaluation benchmark from Tongyi Lab for personal assistant and agent scenarios, evaluating base models and runtime frameworks (Harness) together. PawBench v1.0 includes 15...

AI AgentModel InferenceBenchmark
Jun 21, 2026Read more →
3 months ago

opera-browser-cli – Opera Neon's Open-Source Command-Line Tool

opera-browser-cli is an open-source CLI from the Opera Neon team, built on opera-devtools, letting local AI agents (e.g., Claude Code) control the browser from the terminal. No cloud relay or heavy OA...

AI AgentAI CodingAI Safety
Jun 21, 2026Read more →
3 months ago

OpenSquilla – Open-Source Microkernel AI Agent Framework That Cuts Token Costs

OpenSquilla is an open-source, self-hostable, token-efficient microkernel AI Agent runtime framework developed by the community, built around the idea of “more intelligence density for the same budget...

AI AgentAI CodingAI Safety
Jun 21, 2026Read more →
3 months ago

openPangu 2.0 – Huawei's Open-Source Upgraded Pangu Large Model

openPangu 2.0 is a major upgrade of Huawei's Pangu large model, offering a 505B-parameter Pro variant and a 92B Flash variant, both with a unified 512K context window. From algorithms through training...

AI AgentAI CodingModel Inference
Jun 21, 2026Read more →
3 months ago

OpenClacky – Li Yafei Team's Open-Source Low-Cost AI Agent

OpenClacky is an open-source AI Agent from the Li Yafei team aimed at professional users who need low ongoing API cost. Through a lean architecture and smart scheduling, it sharply cuts continuous-run...

AI AgentAI CodingModel Inference
Jun 21, 2026Read more →
Page 18 of 23 (224 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.