AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: LLM

72 article(s) found · Clear tag

3 months ago

SkillOpt – Microsoft's Open-Source Agent Skill Documentation Optimization Tool

SkillOpt is an open-source Agent skill documentation optimization tool developed by Microsoft. Its core concept is introducing the deep learning training paradigm into the text space, treating the Age...

AI AgentModel InferenceLLM
Jun 26, 2026Read more →
3 months ago

HiCAD – Open AI Parametric 3D CAD Modeling Platform

HiCAD is an open AI parametric 3D CAD platform built for makers and 3D printing enthusiasts. Natural language—in Chinese or English—generates editable JSCAD parametric code in seconds, with live 3D pr...

AI CodingLLMBenchmark
Jun 21, 2026Read more →
3 months ago

General365 – Meituan LongCat Team's Open General Reasoning Benchmark

General365 is an open general reasoning benchmark from Meituan's LongCat team, designed to evaluate large language models (LLMs) purely on logical reasoning in everyday scenarios. The benchmark includ...

Reasoning ModelLLMBenchmark
Jun 21, 2026Read more →
3 months ago

Mirage – strukto-ai's Unified Virtual Filesystem for AI Agents

Mirage from strukto-ai is a unified virtual filesystem for AI Agents. It mounts heterogeneous backends—S3, Slack, Gmail, GitHub, MongoDB, and more—as one virtual tree so Agents read, write, query, and...

AI AgentEmbedding & RAGLLM
Jun 21, 2026Read more →
3 months ago

HiDream-O1-Image-Pro – HiDream.ai's Flagship Image Model

HiDream-O1-Image-Pro is HiDream.ai's flagship image generation model built on the native full-modal UiT (Unified Transformer) architecture with over 200B parameters. It sets new SOTA across text-to-im...

MultimodalImage GenerationLLM
Jun 21, 2026Read more →
3 months ago

VitaBench 2.0 – Meituan LongCat’s Long-Horizon Dynamic Agent Benchmark

VitaBench 2.0 from Meituan’s LongCat team is the first benchmark for long-horizon dynamic user modeling in real-life agent scenarios. It ships 56 statistically grounded synthetic users, 819 lifecycle-...

AI AgentModel InferenceLLM
Jun 21, 2026Read more →
3 months ago

SenseNova-U1-8B-MoT-Infographic – SenseTime's Open-Source Infographic-Enhanced Model

SenseNova-U1-8B-MoT-Infographic is SenseTime's open-source infographic-enhanced model built on the unified SenseNova-U1-8B-MoT architecture at 8B parameters. Through targeted data training and reinfor...

MultimodalAI CodingLLM
Jun 21, 2026Read more →
3 months ago

Qwen-VLA – Alibaba Tongyi's General Vision-Language-Action Model

Qwen-VLA is Tongyi Lab's general vision-language-action model: Qwen3.5-4B VLM backbone plus 1.15B DiT action decoder. A unified action trajectory prediction framework merges manipulation, navigation, ...

MultimodalEmbodied AILLM
Jun 21, 2026Read more →
3 months ago

Qwen-Robot Suite – Alibaba Tongyi's Physical-World Foundation Model Suite

Qwen-Robot Suite is Alibaba Tongyi Lab's foundation model suite for physical-world intelligence, comprising Qwen-RobotNav (navigation), Qwen-RobotManip (manipulation), and Qwen-RobotWorld (world model...

MultimodalEmbodied AILLM
Jun 21, 2026Read more →
3 months ago

Qwen-Image-Bench – Qwen Team's Text-to-Image Model Evaluation Benchmark

Qwen-Image-Bench is a standardized benchmark dataset from Alibaba's Qwen team (Tongyi Qianwen) for evaluating text-to-image models—1,000 carefully designed test samples covering Chinese and English pr...

MultimodalImage GenerationLLM
Jun 21, 2026Read more →
Page 5 of 8 (72 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.