AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: AI Agent

224 article(s) found · Clear tag

2 months ago

MemGUI-Agent – A Long-Horizon Mobile GUI Agent Jointly Developed by Zhejiang University and Kuaishou

MemGUI-Agent is a long-horizon mobile GUI agent jointly developed by Zhejiang University and Kuaishou, specifically designed for cross-app, multi-step, and long-chain mobile automation tasks. Traditio...

AI AgentMultimodalAI for Science
Jul 7, 2026Read more →
2 months ago

Octo – Mininglamp's Open-Source AI-Native Team Collaboration Platform

Octo is an open-source AI-native team collaboration platform developed by Mininglamp. It aims to aggregate dispersed AI Agents into a unified space, enabling efficient orchestration and collaboration ...

AI AgentOpen Source
Jul 7, 2026Read more →
2 months ago

GenEvolve – Self-Evolving Image Generation Agent by Meituan and Others

GenEvolve is a self-evolving image generation agent jointly developed by the Hong Kong University of Science and Technology (Guangzhou), Meituan, and the National University of Singapore. It formalize...

AI AgentMultimodalImage Generation
Jul 7, 2026Read more →
2 months ago

Hy3 – Tencent Hunyuan's Open-Source Mixture of Experts Model

Hy3 is a 295B-parameter Mixture of Experts (MoE) model open-sourced by the Tencent Hunyuan team. It demonstrates significant improvements in agent capabilities, reasoning, and long-context tasks, with...

AI AgentAI CodingLLM
Jul 6, 2026Read more →
2 months ago

Elements Claw – Alibaba DAMO Academy's AI Agent for Superconducting Material Discovery

Elements Claw is the industry's first AI agent for superconducting material discovery, jointly launched by Alibaba DAMO Academy, Renmin University of China, and the University of Chinese Academy of Sc...

AI AgentAI for ScienceLLM
Jul 6, 2026Read more →
2 months ago

EdgeBench – ByteDance's AI Learning Capability Benchmark Framework

EdgeBench is a benchmark framework developed by ByteDance's Seed team, specifically designed to evaluate the long-term learning capabilities of autonomous AI Agents in real-world environments. The fra...

AI AgentAI CodingAI for Science
Jul 5, 2026Read more →
2 months ago

GeneBench-Pro – OpenAI's Research-Grade Benchmark for Computational Biology

GeneBench-Pro is a research-grade benchmark developed by OpenAI, specifically designed to evaluate AI models' ability to handle judgment-intensive analysis in computational biology. The benchmark comp...

AI AgentReasoning ModelAI for Science
Jul 2, 2026Read more →
2 months ago

yuxinlu1 Gemma4-12B – Open-Source Coding & Agentic Model Series

yuxinlu1 Gemma4-12B is an open-source coding and Agentic model series fine-tuned by individual developer Lu Yuxin based on Google's Gemma 4 12B instruction model, comprising the V1 Code version and V2...

AI AgentAI CodingReasoning Model
Jul 1, 2026Read more →
2 months ago

Nano Banana 2 Lite – Google's Lightweight AI Image Generation Model

Nano Banana 2 Lite is Google's self-developed lightweight AI image generation model, positioned as a speed-first ultra-fast version capable of generating a single image in 4 seconds, with a cost of on...

AI AgentMultimodalImage Generation
Jul 1, 2026Read more →
2 months ago

Claude Sonnet 5 – Anthropic's Most Powerful Agent Model

Claude Sonnet 5 is the most capable agent model in Anthropic's Sonnet series. Its performance in benchmarks for agentic coding, terminal operations, browser search, and computer use approaches that of...

AI AgentAI CodingReasoning Model
Jul 1, 2026Read more →
Page 12 of 23 (224 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.