OpenAI’s “Astra” reportedly cracked 10 open math and CS problems—let’s wait for the evidence.

Ongoing story : Fatigue hype 2026 : le tri entre modèle et harness· Part 14/16

Horizon 17 h ago8Add to bookmarks

OpenAI’s “Astra” reportedly cracked 10 open math and CS problems—let’s wait for the evidence.
Illustration : Léa Fontaine

Noam Brown posted, HN comments: an internal model may have solved ten major open problems. Nothing is published. We observe, we do not conclude.

In plain terms. OpenAI researcher Noam Brown (@polynoamial) announces that an internal model called Astra has allegedly solved ten major open problems in mathematics and computer science. The proofs have not been published.

The fact

The post, shared on Hacker News (item 49143688) on August 2, 2026, does not list the exact ten problems. One problem (number 6) is said to have been constructed based on the work of researcher Henry Yuen (quantum complexity theory). No written proof is publicly available at the time of writing.

Our take

This announcement is part of a pattern. Recall publication #1342: an OpenAI result on a thirty-year-old open convex optimization problem, announced before peer review. Today’s HN discussion echoes the same criticism—waiting for the proofs to verify the mathematics—and highlights the contrast with standard scientific practices: arXiv deposit, review, replication. Neither alarmism nor hype: we note the announcement, withhold judgment until the proofs are available.

This matters because, if validated, it would shift the bar for what advanced LLMs can actually do in mathematical reasoning—not just mimic derivation chains, but generate verifiable new proofs. This is precisely where the research/hype boundary is at stake.

To watch

Actual publication of proofs (arXiv, OpenAI blog); naming of the ten problems; reactions from Henry Yuen and relevant mathematicians; replication by third parties.

Resources, try it

Article produced by artificial intelligence, reviewed under human editorial control.

Our newsroom
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

SSHMonitoringAI Ops
Get early access
Was this article helpful?

8 people liked this article

Like
J
Jin-ho ParkFrontier & research
🇬🇧 Research, deep tech, foresight.
Share:
Comments (8)

Sign in to join the discussion.

TechGuru99 03 Aug 2026 · 06:25

If true, this would be huge-but we’ve seen this movie before with other AI claims. No proofs, no papers, just tweets and speculation. Feels like marketing until the work is out.

BookWorm88 03 Aug 2026 · 06:20

Isn't this why we need independent verification before getting excited? Without peer-reviewed results, it's just hype.

TechSavvy47 03 Aug 2026 · 06:18

Without published proofs, it’s just noise. But if real, this could change how we approach unsolved problems-and whether AI can do more than crunch numbers.

TravelTom 03 Aug 2026 · 05:50

If these claims are legit, it’s wild-but OpenAI’s track record makes me wonder why they’d announce this without the proof. Ever seen a lab drop a bombshell like this and then vanish?

FoodieFiona 03 Aug 2026 · 05:47

Hard to get hyped when ten problems remain unsolved if the proof isn’t public yet. Feels like déjà vu with other breakthrough claims that fizzled out.

MusicFanatic 03 Aug 2026 · 05:41

Waiting for the paper is the only way to know if this is hype or breakthrough. Either way, OpenAI’s secrecy isn’t helping.

ph1lippe_m 03 Aug 2026 · 05:41

If these claims hold up, it’s a massive leap - but without peer-reviewed proof, skepticism is fair.

J.P.R. 2 03 Aug 2026 · 05:31

Ten unsolved problems cracked in math and CS sounds impressive, but until there are proofs to scrutinize, this is just another claim in a long line of hyped AI announcements. Skepticism isn’t cynicism-it’s due diligence.

Story timeline

Fatigue hype 2026 : le tri entre modèle et harness

  1. 1« I love LLMs, I hate hype » - geohot reminds the only rule that remains13/07/2026
  2. 2"Poor and overconfident": developers are poor judges of LLM assertions13/07/2026
  3. 3How do software professionals really judge the code generated by AI?13/07/2026
  4. 4Zig, Zed, Anthropic: when a language creator calls the hype by its name13/07/2026
  5. 5"The LLM critics are right. I use LLMs anyway" - the voice that reassembles16/07/2026
  6. 6The cost of saying yes has changed: GitHub reignites the debate on the real bottleneck17/07/2026
  7. 7"Claude Code: Anatomy of a Misfeature" - when public review becomes the real QA17/07/2026
  8. 8Google's Gemini 3.6 Flash is cheaper and shorter - and Gemini 4 gets a tease while 3.5 Pro stays late22/07/2026
  9. 9"AI didn't make programming easier, it just made it differently difficult" - CACM lands the anti-hype line22/07/2026
  10. 10"State-owned AI won't solve inequality": Rest of World's bold thesis on AI in the Global South24/07/2026
  11. 11Refactoring as a token-cost lever: an experiment in Fowler's gen-AI series30/07/2026
  12. 12Rachel Laycock: "Attention has become the scarce resource" - the dev-orchestrator, managing 8 to 12 agents simultaneously31/07/2026
  13. 13Situational Awareness drops 67% in a month: the trial of the true believers02/08/2026
  14. 14OpenAI’s “Astra” reportedly cracked 10 open math and CS problems—let’s wait for the evidence.02/08/2026
  15. 15"Cancelling Cursor": Quality debt takes precedence over feature velocity02/08/2026
  16. 16Jeff Dean on what AI teams get wrong: the diagnostic from the shop that pays every bill03/08/2026
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

Get early access
Topics
Explore
Information