The models, labs, and research the AI world is actually talking about.
Developer posts video of his San Francisco run after announcing new role.
Former DeepMind researcher shares update on leaving Meta for new AI lab role.

Meta Superintelligence Labs releases its first real-time audio perception model for streaming speech.
The post comes from the account of X's owner and xAI founder.
Elon Musk states Grok operates independently on cloud infrastructure.
Perplexity AI enables tasks to shift from cloud to local models on Mac for private data.
Replies target an arXiv paper on sliding-window attention with sinks.

Grok is providing another reset on token usage limits for all its bot users.
OpenRouter says Mercury 2.5 Preview reaches 1,107 tokens per second exclusively on the platform with tunable reasoning and parallel tool calls.
Conversation covers path to AGI, 3.7 Flash progress, and frontier focus.
Co-founders announce seed funding for an AI-powered physics research lab.
Investor shares details of second talk with Conviction founder on AI progress.
Google engineers announce tool-based video processing in Gemini API that cuts tokens by up to 88%.

Meta executive releases first real-time audio perception model for streaming speech-to-text.

Team posts emphasize gains in accuracy, speed, and pricing for complex visual documents.
Wafer is an inference provider using Nvidia and non-Nvidia chips that declined buyout offers.
Screenshots from GitHub and support docs show new model references.
Former Tesla leader shares how he routes all Grok use through one controlling agent.
Tweet questions if people forgot rogue AIs can self-replicate without hacking AWS.
An academic shares a post contrasting EU rules with stricter policies in China.