AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: AI Agent

224 article(s) found · Clear tag

1 months ago

In-Depth Review of Grok 4.6: The Triangular Balance of Performance, Speed, and Cost-Effectiveness in the Flagship Model

Grok 4.6 is the latest generation flagship large language model launched by SpaceXAI (xAI), based on a 1.5T parameter Mixture-of-Experts (MoE) architecture, with an ultra-long context window of 500K t...

AI AgentAI CodingLLM
Aug 13, 2026Read more →
1 months ago

Palmier Pro – Open-Source AI Video Editor That Generates Video and Images During Editing

Palmier Pro is a macOS-native video editor built from scratch using Swift, designed for the AI era. It is open-source and offers its core features for free. Its key innovation lies in embedding genera...

AI AgentSpeech AIVideo AI
Aug 12, 2026Read more →
1 months ago

DeepSeek Harness – DeepSeek's AI Agent Execution Framework

DeepSeek Harness is an AI agent execution framework launched by DeepSeek, based on the core concept of "Model + Harness = Agent." Through engineering modules such as context management, tool calling o...

AI AgentAI CodingTool Calling
Aug 12, 2026Read more →
1 months ago

Nemotron 3.5 Lightning – NVIDIA's Open-Source MoE Model

Nemotron 3.5 Lightning is an open-source Mixture-of-Experts (MoE) model with 30B parameters launched by NVIDIA, specifically optimized for multi-agent systems. The model employs a sparse activation ar...

AI AgentModel InferenceLLM
Aug 12, 2026Read more →
1 months ago

PAST-Bench – Princeton's Benchmark for Performance Attribution in Personal Agents

PAST-Bench is a benchmark introduced by Princeton University's team led by Mengdi Wang, specifically designed to evaluate the recursive self-improvement capabilities of personal AI agents. By comparin...

AI AgentModel InferenceBenchmark
Aug 12, 2026Read more →
1 months ago

Muse Glimmer: In-Depth Evaluation of Meta's Open-Source 300 Billion Parameter Local Agent Large Language Model

Muse Glimmer is an open-source large language model with 30 billion parameters developed by Meta, specifically optimized for agentic workflows that operate continuously on local devices. Using 4bit qu...

AI AgentAI CodingModel Inference
Aug 11, 2026Read more →
1 months ago

WorldClaw – A 3D World Generation Framework Introduced by Tencent Hunyuan Team

WorldClaw is an Agent-driven open-world 3D generation framework introduced by Tencent's Hunyuan team. With a natural language description as input, the system can automatically plan terrain, regions, ...

AI AgentEdge Deployment
Aug 10, 2026Read more →
1 months ago

MemHarness – A Memory Reconstruction Framework for LLM Agents Introduced by the Shanghai AI Lab and Others

MemHarness is a memory reconstruction framework for LLM Agents introduced jointly by the Shanghai Artificial Intelligence Lab and universities such as Zhejiang University, Fudan University, and Shangh...

AI AgentEmbedding & RAGLLM
Aug 5, 2026Read more →
1 months ago

Qwen-CUA – The Native Computer Use Agent Introduced by Alibaba Qwen and Others

Qwen-CUA is a native Computer Use Agent introduced jointly by the Qwen team and XLang Lab. Based on a 397B-A17B Mixture-of-Experts (MoE) architecture, it perceives the interface state solely through s...

AI AgentLLMBenchmark
Aug 5, 2026Read more →
1 months ago

Orchard – Microsoft Research's Open-Source Agentic AI Modeling Framework

Orchard is an open-source Agentic AI modeling framework introduced by Microsoft Research. At its core is the Kubernetes-based Orchard Env environment service, which enables cross-domain reuse of sandb...

AI AgentModel InferenceBenchmark
Aug 4, 2026Read more →
Page 8 of 23 (224 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.