AI news feed

Latest 25 news items

Filter by day, tag, source, or full-text query. Results are sorted by newest ingested item first.

Results
25
Showing newest 25 of 106
Clear

Timeline

Last 90 days
Search: Qwen
Latent Spacelatent.spaceAdded
[AINews] not much happened today

Congrats to Harvey but we covered that already . AI News for 9/8/2026-9/9/2026.

benchmarkmodelpaperrelease
Latent Spacelatent.spaceAdded
[AINews] OpenAI reports Navier-Stokes singularity find in 88 hours using Astra-next, roughly 10,000 agents and 130B tokens (>$40M), a contender for second ever Millennium Prize awarded

Today was a tough news cycle to launch anything; we ordinarily promise to cover any new decacorn fundraises so Cognition’s $48B round and Mistral’s $24B round would normally have made it; we love imagegen so GPT Image 2.5 would have been its own headline; we covered the Dreamer story closely so their relaunch as Meta’s Muse agent should have made it; but..…

benchmarkmodelpaperrelease
Latent Spacelatent.spaceAdded
[AINews] GPT-6 Astra: OpenAI’s biggest LLM launch of all time

The launch is barely 9 hours old, and with 36M views and 164K likes, already is OpenAI’s most successful launch since Sora and certainly GPT-4 or GPT-5 . You’ll recall we’ve previously observed that Anthropic tends to far outclass OpenAI in launch popularity.

benchmarkmodelreleasesafety
Hugging Face Bloghuggingface.coAdded
Training a coding model to paint watercolours with TRL and OpenEnv

On 23 August, Surya Narreddi posted a beautiful video of watercolours painted by a language model. The model writes JavaScript through p5.brush , a library that "adds natural drawing tools to p5.js".

benchmarkmodelpaperrelease
Latent Spacelatent.spaceAdded
[AINews] Muse Spark 1.3 matches GPT-5.6-Sol, confirming Meta Superintelligence as the newest Frontier Lab, >90% discount for training

Launch season continues from yesterday , with Gemini 3.8 Flash as rumored today, but Muse Spark 1.3, promised in Zuck’s big comeback letter last month, definitely deserved the title story win today. Per AAII it is now the #3 model in the world (!?!) Just look at the confidence displayed finally putting up comparable numbers to the frontier models from…

benchmarkmodelreleasesafety
TheSequencethesequence.substack.comAdded
The Sequence Learning Loop - Issue 925: Learn About Fable and Mythos 5.1, GLM-5.3-Flash, and Qwen 3.8

The past month delivered three releases worth reading closely, not because they move the same benchmark but because each is a different answer to the same question: how do you build a model that can work on its own for hours, and how do you make that affordable? Anthropic shipped Claude Fable 5.1 and Mythos 5.1 , one set of weights sold under two safeguard…

benchmarkmodelrelease
Latent Spacelatent.spaceAdded
[AINews] Claude Fable/Mythos 5.1: new SOTA model, 75% cache price cut but 70% more output tokens

With Astra clearly finally warming up for a full launch (with @sama and @openai writing about it again after a month of self imposed pacing ), there’s a familiar window to take the narrative with the round robin of model launches, with Grok 4.7 and Gemini Flash 3.8 also on the way. But that’s also perhaps not the best way to frame today’s launch… which got…

benchmarkmodelpaperrelease
Latent Spacelatent.spaceAdded
[AINews] OpenAI shuts off Cursor

A late entrant in the news cycle of an eventful week: Following the closing of Cursor’s acquisition by SpaceX last week , it was time for OpenAI to do what Anthropic did to Windsurf when it was being considered for acquisition by OpenAI: We’re ending our partnership with Cursor following its acquisition by SpaceX. Under our proposal, Cursor’s direct access…

benchmarkmodelpaperrelease
Latent Spacelatent.spaceAdded
[AINews] Hot Chips: OpenAI’s Jalapeño, Cerebras CS-5, Groq 3 LPX, Apple M6

By far the biggest announcement at the 37th Hot Chips conference was OpenAI’s stunning progress on their own chip, less than a year after the Broadcom announcement … and that it isn’t an ASIC; but a full on Blackwell-beating alternative. Since announcing Jalapeño, our first custom inference chip, we’ve been testing it and the system around it.

Hugging Face Bloghuggingface.coAdded
How Hugging Face Inference Endpoints, Jobs, and Buckets Power Search on Papers with Code

3 months ago, we started a revival of Papers with Code (see also the announcement tweet ). Its goal is to make open AI research accessible and digestible, so that people can easily find the artifacts related to a paper, find state-of-the-art (SOTA) across the various domains of AI, share interesting research and build on top of each other's work.

Latent Spacelatent.spaceAdded
[AINews] Andrew Ng gets into AI Engineering

We’ve lost count of how many adoption milestones have been passed since the original Rise of the AI Engineer post, but surely Andrew Ng, cofounder of Google Brain and Coursera among many other things, relaunching DeepLearning.ai with a focus on AI Engineering is a big one : This was done via “ an analysis of over 10,000 job postings; carrying out dozens of…

benchmarkmodelpaperrelease
Hugging Face Bloghuggingface.coAdded
AI Workflows in Gradio

Most interesting AI apps are pipelines. You generate an image, then cut out its background if you want to, or edit it into something new.

modelreleasetool
Latent Spacelatent.spaceAdded
[AINews] 10% worse, 100x cheaper, 10000x faster: Why Simulation is taking over

By AI standards today is a pretty quiet Friday, so it’s time to take a step back and reflect on what is really going on. If you read our 2025 reading list , and followed our coverage of Z.ai GLM , understood the Poolside pivot , been following our AI for Science themes , and tuned in to today’s Simile pod , you not only are one of the biggest readers of…

benchmarkmodelreleasesafety
https://aiweekly.co/aiweekly.coAdded
The frontier just split into three markets

In the Wild What people are installing, watching, and searching for now. See the full daily movement in In the Wild .

benchmarkmodelpaperrelease
The Neurontheneurondaily.comAdded
Qwen, Cursor, DeepSeek Explained Live

Your browser does not support the audio element. Click the header to go straight to YouTube.

benchmarkmodelpapertool
Hugging Face Bloghuggingface.coAdded
Measuring benchmark optimization in speech recognition

Public voice AI benchmarks increasingly suggest that models are performing at human levels. Yet those scores don't always reflect how models work in the real-world.

benchmarkmodelreleasesafety