AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Document AI

32 article(s) found · Clear tag

1 months ago

OpenStory: Open Source AI Video Production Platform, Generate Style-Consistent Short Drama Videos with One Click

OpenStory is an open-source AI video production platform that specializes in converting scripts into style-consistent short drama videos with a single click. The platform can automatically split scene...

Speech AIVideo AIDocument AI
Aug 25, 2026Read more →
1 months ago

TabTin – In-Depth Review of Open-Source Full-Stack AI Agent Team Collaboration Platform

TabTin is a full-stack open-source AI Agent team collaboration platform designed to integrate intelligent agents into team workflows. Tasks can be initiated on desktop, viewed on mobile, executed by a...

AI AgentAI CodingDocument AI
Aug 24, 2026Read more →
1 months ago

SenseNova U1.5 Lite – A Lightweight Native Unified Multimodal Large Model Open-Sourced by SenseTime

SenseNova U1.5 Lite is a lightweight native unified multimodal large model open-sourced by SenseTime, specifically designed for real-world visual creation workflows. This model achieves a native multi...

MultimodalDocument AITool Calling
Aug 21, 2026Read more →
1 months ago

MyContext – In-Depth Review of Alibaba Qwen Office's Open-Source Agent Context Infrastructure

MyContext is an open-source Agent context infrastructure introduced by the Alibaba Cloud Qwen Office team. It employs a local-first architecture to automatically process multi-source work data from in...

AI AgentDocument AIEdge Deployment
Aug 18, 2026Read more →
1 months ago

PixelRAG – Berkeley's Open-Source Vision-Native RAG Framework

PixelRAG is an open-source vision-native RAG framework developed by Berkeley's SkyLab/BAIR. It moves beyond the traditional RAG paradigm of extracting text from web pages or PDFs for retrieval, instea...

Embedding & RAGDocument AIBenchmark
Aug 13, 2026Read more →
1 months ago

HOMIE – An Open-Source Digital Human Video Generation Framework from The Hong Kong University of Science and Technology

HOMIE is an open-source digital human video generation framework developed by the Department of Computer Science and Engineering at The Hong Kong University of Science and Technology. It is built upon...

MultimodalVideo AIDocument AI
Aug 12, 2026Read more →
1 months ago

Grok Imagine Image 2.0 – The Image Generation Model from SpaceXAI

Grok Imagine Image 2.0 is a new-generation image generation and editing model launched by xAI (published under the name SpaceXAI). It is integrated into the Grok web interface and iOS/Android apps in ...

Document AIEdge DeploymentAI Tools
Aug 9, 2026Read more →
1 months ago

Wan3.0 – Alibaba's WanXiang Latest Video Generation Large Model

Wan3.0 is the latest video generation large model launched by Alibaba Cloud's WanXiang team, achieving comprehensive upgrades in generation duration, multimodal input, consistency maintenance, and rea...

MultimodalVideo AIDocument AI
Aug 7, 2026Read more →
1 months ago

SenseNova U1.5-Lite-Preview – A Lightweight Multimodal Model Open-Sourced by SenseTime

SenseNova U1.5-Lite-Preview is the preview version of a lightweight, native unified multimodal model open-sourced by SenseTime, developed iteratively based on the NEO-Unify architecture. With only 8B-...

MultimodalDocument AIEdge Deployment
Aug 4, 2026Read more →
2 months ago

Ling 3.0 Flash – A Lightweight MoE Inference Model Launched by Ant Bailing

Ling 3.0 Flash is a lightweight Mixture-of-Experts (MoE) inference model introduced by the Bailing large model team at Ant Group. With a total parameter count of 124B, it activates only 5.1B parameter...

AI AgentAI CodingDocument AI
Jul 28, 2026Read more →
Page 2 of 4 (32 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.