Daily curated AI news: breakthroughs, product launches, industry trends and more.
Total 157 articles
OpenAI launched an upgraded voice mode solution for ChatGPT on September 10. This version allows users to actively select GPT-5.6 Sol or GPT-6 Astra models for complex reasoning tasks during conversations, introducing dynamic model switching mechanisms and a unified billing system.
HiDream ZhiXiang Future announces the global launch of its content creation agent vivago R1. The product breaks through the time limitations of traditional AI video generation tools, achieving 5-minute high-quality video output in a single generation and supporting multiple extensions.
On September 8, Tencent Hunyuan team announced the completion of iterative upgrades for the Hy4 Preview model. The new version reduces average reasoning rounds by 23% and token consumption for input/output by 18% through algorithm architecture adjustments and reasoning process reengineering, while maintaining task completion quality. This optimization has been fully deployed across two core scenarios—financial analysis and code generation—verified by Bench benchmark tests and human evaluation.
xAI has expanded Grok Bot to iPad and Android, completing coverage across Mac, iPhone, iPad, and Android. The platform supports round-the-clock AI agent collaboration, cross-application access, retained task context, and information sharing between agents.
NVIDIA announced a cash acquisition of the AI open-source community Hugging Face for approximately $13 billion. The transaction structure allocates $11.9 billion to institutional investors and $1 billion to employee equity incentive plans. This acquisition will significantly enhance NVIDIA's influence within the open AI ecosystem.
The Tencent Hunyuan team launched the Hy4 Preview large model on August 28th, achieving multiple technical breakthroughs at the foundational architecture level. This model enables efficient computation with a total parameter scale of 770B through dynamic activation mechanisms and has optimized long-text processing capabilities for multimodal reasoning scenarios. Its research outcomes have been applied across interdisciplinary fields including game engine optimization, biopharmaceutical simulations, and materials science experimental design.
Zhipu AI launches GLM-5.3-Flash multimodal model, achieving performance and cost balance through hybrid attention architecture
xAI has released Grok Voice Think Fast 2.0, a next-generation voice model built on an end-to-end speech-to-speech architecture. It directly processes spoken input and generates spoken output, eliminating the need for separate speech recognition, language model inference, and spee...
The Tencent HunYuan team has recently open-sourced AngelSpec, an end-to-end speculative decoding training framework built on TorchSpec. The framework decouples inference from training: the inference engine streams hidden states to training workers via Mooncake/RDMA. It supports u...
Get curated AI news delivered to your inbox daily. Never miss an update.
We never share your email. Unsubscribe anytime.