AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Embodied AI

25 article(s) found · Clear tag

2 months ago

LiveWorld – Generative Video World Model from University of Adelaide and Others

LiveWorld is a generative video world model jointly developed by the University of Adelaide, the Australian National University, and other institutions. Its core focus is solving the problem of out-of...

MultimodalVideo AIEmbodied AI
Jul 1, 2026Read more →
2 months ago

LocateAnything – NVIDIA's Visual Language Grounding Model

LocateAnything is a visual language grounding model developed by NVIDIA, based on Parallel Box Decoding (PBD) technology. Users can input natural language to precisely select targets in images. With 3...

MultimodalDocument AIEmbodied AI
Jul 1, 2026Read more →
2 months ago

BrowserBC – An Open-Source Browser Operation Trajectory Generation Skill by Einsia AI

BrowserBC is an open-source project released by Navers Lab under Einsia AI. Its core goal is to transform human browser operation trajectories into reusable natural language skills, enabling Web Agent...

AI AgentEmbodied AIModel Inference
Jun 29, 2026Read more →
3 months ago

OpenHuman – Open Desktop AI Assistant with Proactive Work Context

OpenHuman is an open desktop "personal AI super intelligence" from tinyhumansai—not a passive chatbot but a desktop agent that proactively senses work context. Every 20 minutes it syncs 118+ third-par...

AI AgentEmbodied AIOpen Source
Jun 21, 2026Read more →
3 months ago

Agora-1 – Odyssey's First Multi-Agent World Model

Agora-1 is Odyssey's first multi-agent world model, breaking past the single-user limits of traditional world models by enabling humans and AI to interact in the same real-time generated world simulat...

AI AgentEmbodied AIBenchmark
Jun 21, 2026Read more →
3 months ago

Wall-OSS-0.5 – X Square Robot’s Open Embodied Intelligence Model

Wall-OSS-0.5 is X Square Robot’s open vision-language-action (VLA) model—a 4B-parameter stack on a 3B Qwen2.5-VL backbone achieving zero-shot real-robot deployment without per-task fine-tuning. Gradie...

MultimodalEmbodied AIOpen Source
Jun 21, 2026Read more →
3 months ago

SCAIL-2 – Zhipu AI and Tsinghua's Open-Source Character Animation Model

SCAIL-2 is the second-generation film-grade character animation framework open-sourced jointly by Zhipu AI and Professor Liu Yongjin's research group at Tsinghua University. Built on a Diffusion Trans...

Video AIEmbodied AIOpen Source
Jun 21, 2026Read more →
3 months ago

Qwen-VLA – Alibaba Tongyi's General Vision-Language-Action Model

Qwen-VLA is Tongyi Lab's general vision-language-action model: Qwen3.5-4B VLM backbone plus 1.15B DiT action decoder. A unified action trajectory prediction framework merges manipulation, navigation, ...

MultimodalEmbodied AILLM
Jun 21, 2026Read more →
3 months ago

Qwen-Robot Suite – Alibaba Tongyi's Physical-World Foundation Model Suite

Qwen-Robot Suite is Alibaba Tongyi Lab's foundation model suite for physical-world intelligence, comprising Qwen-RobotNav (navigation), Qwen-RobotManip (manipulation), and Qwen-RobotWorld (world model...

MultimodalEmbodied AILLM
Jun 21, 2026Read more →
3 months ago

Microsoft Scout – Microsoft's AI Personal Assistant

Microsoft Scout is an enterprise-grade AI personal assistant from Microsoft. It is not a traditional chatbot but an Autopilot agent built on the OpenClaw open-source stack with its own Entra identity....

AI AgentEmbodied AIOpen Source
Jun 21, 2026Read more →
Page 2 of 3 (25 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.