Meta Launches Muse Glimmer, a 30B AI Model That Runs on One GPU
Meta has released Muse Glimmer, a 30-billion-parameter open-weight model built for local AI agents, coding, tool use, and multimodal tasks on a single consumer GPU or high-end Mac.
Inference is the stage when a trained AI model uses new input to produce a prediction, classification, recommendation, or generated response. Training may take days across specialized hardware, while inference happens whenever a user submits a prompt, a camera captures an image, or an application requests a score. Production systems optimize inference for latency, throughput, memory use, energy consumption, and cost without degrading output quality. The model may run in a cloud service, data center, phone, vehicle, or edge device. Monitoring remains necessary because real-world inputs change, hardware can fail, and a model that performed well during evaluation may behave differently under live traffic.
Meta has released Muse Glimmer, a 30-billion-parameter open-weight model built for local AI agents, coding, tool use, and multimodal tasks on a single consumer GPU or high-end Mac.
Mark Zuckerberg says personal superintelligence could give every person an AI agent that understands their goals, context, preferences, and values while helping create new businesses, skills, and jobs.
Cloudflare says AI agents, bots, and other automated systems could generate up to 1,000 times more internet traffic than humans within five years as machine activity accelerates.
OpenAI is replacing GPT-5.5 Instant with GPT-5.6 Luna for Free and Go users, then removing text-chat limits and adding a Think button for harder questions.
Demis Hassabis is becoming Alphabet chief scientist as Koray Kavukcuoglu takes operational control of DeepMind and four Google veterans launch Discovery Loop.
Meta has launched Muse Code, a terminal coding agent powered by Muse Spark 1.2, targeting long-running software projects and established rivals Codex and Claude Code.
AI startups captured 53% of global venture funding in July, their lowest share since December, even as total investment reached $65 billion in a record month for mega-rounds.
More than 1,200 employees from leading AI companies are asking the U.S. government to help create international mechanisms that could slow automated AI development if capabilities begin advancing faster than society can evaluate or control.
Elon Musk says xAI plans to release Grok 4.6 around August 7 and follow it with the larger Grok 4.7 several weeks later, extending the company’s rapid model rollout.
Anthropic researchers used Claude Mythos Preview to develop stronger attacks on the HAWK post-quantum signature scheme and a reduced-round version of AES, demonstrating research-level cryptanalysis without threatening current production systems.