Alibaba’s Qwen Audio 3.0 TTS Tops the Speech Leaderboard
Alibaba released Qwen Audio 3.0 TTS, a speech model that ranks first on an independent text-to-speech leaderboard, clones voices across 16 languages and costs a third of ElevenLabs.
Daniel Mercer covers foundation models, generative AI systems, and applied machine learning deployments across enterprise and consumer platforms. He reports on model launches, performance benchmarks, inference pricing, and the competitive dynamics shaping leading AI labs. His work evaluates compute efficiency, data sourcing strategies, and integration pathways that determine scalability. Daniel takes a systems-level approach, connecting technical documentation and release notes to business outcomes and adoption signals. He frequently analyzes safety mechanisms, model limitations, and reliability claims through testing data and operational evidence. Based in San Francisco, he spends his free time studying urban design and cycling.
Alibaba released Qwen Audio 3.0 TTS, a speech model that ranks first on an independent text-to-speech leaderboard, clones voices across 16 languages and costs a third of ElevenLabs.
Alibaba previewed Qwen 3.8, a 2.4-trillion-parameter model it claims is second only to Anthropic's Fable 5, days after Moonshot's Kimi K3 rattled markets, though no benchmarks...
Chinese startup Moonshot released Kimi K3, a 2.8-trillion-parameter open-weight model it says rivals Anthropic's Claude, as the US-China AI gap narrows.
OpenAI launched Codex Micro, a $230 mechanical keypad built with Work Louder that gives developers physical controls for its Codex coding agent, its first hardware to...
Visa, Mastercard, Stripe, Google and 36 others formed a Linux Foundation body to steward x402, an open standard letting AI agents pay for services directly over...
OpenAI's new GPT-5.6 prompting guide urges developers to write shorter, outcome-first prompts, saying leaner instructions raised scores 10-15% while cutting tokens and cost sharply.
SpaceXAI's Grok 4.5 tied GPT-5.6 Sol for first on the SWE-Atlas-QnA coding benchmark, a narrow win that highlights the model's real edge: efficiency and low cost.
Anthropic extended free Claude Fable 5 access to July 19, its second extension in a week, as OpenAI removed usage caps and made GPT-5.6 Sol cheaper...
OpenAI launched ChatGPT Work, a GPT-5.6 agent that pulls context from a user's apps to produce finished docs, slides and sites, as it merges Codex into...
SpaceXAI released Grok 4.5, a coding and agentic model trained jointly with Cursor that undercuts rivals on price and token use, though it trails Fable 5...