Nvidia Considers Boosting H200 Chip Production for China
Nvidia is reportedly planning to increase production of its H200 AI training chips following approval to sell them in China.
A transformer is a neural network architecture that uses attention to model relationships within a sequence or across multiple data types. Unlike recurrent networks that process elements step by step, transformers can analyze many positions in parallel during training. Self-attention lets each token weigh the relevance of other tokens, while additional layers transform those contextual representations. The architecture underpins modern language models and is also used for images, audio, video, biology, and robotics. Transformers scale effectively but can require substantial data, memory, and computation. Their outputs still reflect training limitations, and long context does not guarantee correct reasoning or recall. Model size, attention design, data quality, and evaluation remain central to performance.
Nvidia is reportedly planning to increase production of its H200 AI training chips following approval to sell them in China.
Google released a major upgrade to its Gemini Deep Research agent, giving developers access to advanced autonomous research capabilities through the Interactions API. The company also introduced DeepSearchQA, a benchmark for evaluating multi-step web research agents.
OpenAI has released GPT-5.2, a new model series designed to improve professional knowledge work with major gains in reasoning, coding, long-context analysis, and vision. The update brings enhanced performance to ChatGPT and the API across Instant, Thinking, and Pro variants.
Disney and OpenAI announced a three-year licensing deal allowing Sora to generate fan-inspired short videos featuring over 200 Disney, Marvel, Pixar, and Star Wars characters.
OpenAI faces a wrongful death lawsuit alleging ChatGPT reinforced a man’s violent delusions, contributing to the killing of his mother in Connecticut. The case marks the first time an AI platform has been accused of involvement in a murder.
Google introduced the Titans architecture and the MIRAS framework to enable AI models to handle massive contexts and update their internal memory while running, improving performance in long-sequence tasks.
ChatGPT now reaches more than 800 million weekly users, creating a powerful adoption flywheel that is speeding the transition from AI experimentation to full-scale enterprise deployment, according to OpenAI’s new 2025 report.
OpenAI buys Neptune for under $400 million in stock, gaining advanced AI model tracking tools and strengthening capabilities in large-scale AI development.
OpenAI introduces an experimental “confessions” system that rewards honest self-reporting of mistakes in GPT-5 Thinking models, aiming to improve AI safety and visibility of failures.
AWS introduced new tools for enterprise customers to build and fine-tune custom large language models, alongside AI infrastructure offerings for on-premises deployments.