AI Model Library: LLMs, Agents & Dev Tools

Curated reviews of mainstream AI models, agents, and dev tools — capabilities, use cases, and how they compare.

Expert analysisUpdated regularlyRSS Feed

Tag: Speech AI

88 article(s) found · Clear tag

1 months ago

dots3-note preview – In-Depth Review of Xiaohongshu's Open-Source Multimodal MoE Model

dots3-note preview is an open-source multimodal MoE model developed by Xiaohongshu's dots model lab. As the first version of the dots3 series, it shares the same technical lineage as the IMO 2026 full...

AI AgentMultimodalSpeech AI
Aug 14, 2026Read more →
1 months ago

MiniMax Music 3.0 – MiniMax's Open-Weight Music Generation Model

MiniMax Music 3.0 is the next-generation open-weight music generation model introduced by MiniMax. The model employs a hierarchical architecture combining an 8B Global LLM and a 0.6B Local LLM, and in...

Speech AILLMEdge Deployment
Aug 14, 2026Read more →
1 months ago

dots.tts – Xiaohongshu and Shanghai Jiao Tong University Open-Source Base Model for Text-to-Speech

dots.tts is a 20-billion-parameter fully continuous autoregressive text-to-speech base model jointly open-sourced by Xiaohongshu's dots team and the X-LANCE Lab at Shanghai Jiao Tong University. The m...

Speech AIAI CodingBenchmark
Aug 13, 2026Read more →
1 months ago

LTX-2.5 – LTX's Open-Source AI Video Generation Foundation Model

LTX-2.5 is a professional-grade AI video generation foundation model open-sourced by Lightricks, an Israeli AI company. With 220 billion parameters, it provides full model weights upon release. The mo...

Speech AIVideo AILLM
Aug 13, 2026Read more →
1 months ago

LTX-2.5 – Lightricks' Open-Source AI Video Generation Base Model

LTX-2.5 is a professional-level AI video generation base model open-sourced by the Israeli company Lightricks. With 22B parameters, the model weights are made available immediately upon release. It su...

Speech AIVideo AITool Calling
Aug 13, 2026Read more →
1 months ago

Palmier Pro – Open-Source AI Video Editor That Generates Video and Images During Editing

Palmier Pro is a macOS-native video editor built from scratch using Swift, designed for the AI era. It is open-source and offers its core features for free. Its key innovation lies in embedding genera...

AI AgentSpeech AIVideo AI
Aug 12, 2026Read more →
1 months ago

IndexTTS-2.5: In-Depth Review of Bilibili's Open-Source Industrial-Grade Zero-Shot Voice Cloning Model

IndexTTS-2.5 is an industrial-grade zero-shot voice cloning model open-sourced by the Bilibili Index Speech team, with only 0.8B parameters. This model supports cross-lingual transfer across five lang...

Speech AIVideo AIModel Inference
Aug 12, 2026Read more →
1 months ago

Shotcut – Open-Source Video Editing Software with Multi-Track Support

Shotcut is an open-source video editing software built on FFmpeg, offering completely free usage with no watermarks or time restrictions. The software supports direct editing of hundreds of audio and ...

Speech AIVideo AIAI Tools
Aug 12, 2026Read more →
1 months ago

SmartSub – Open-Source All-in-One Desktop Tool for Audio and Video Subtitling

SmartSub (MiaoMu) is an open-source all-in-one desktop tool for audio and video subtitling, developed by independent developer Buxuku. This tool integrates end-to-end functionalities such as speech tr...

Speech AIVideo AIEdge Deployment
Aug 10, 2026Read more →
1 months ago

SALMONN-2 – A General Audio Large Language Model Open-Sourced by Tsinghua University and Others

SALMONN-2 is a general audio large language model open-sourced by Tsinghua University, the Shanghai AI Laboratory, and the University of Cambridge. The model employs the SPEAR unified self-supervised ...

Speech AIAI CodingBenchmark
Aug 7, 2026Read more →
Page 4 of 9 (88 articles total)

Subscribe to AI Model Reviews

Get in-depth reviews the moment a new AI model drops. Know its capabilities and use cases.

We never share your email. Unsubscribe anytime.