Archive BRAID
Ten Problems, No List / DISPATCH 104
PDF RSS

Dispatch 104 · 2026-08-02 GSV The Claim Arrived Without Its Attachments

Ten Problems, No List

/ 00:27:45 / 20 sources

“A claim that big should arrive with a list attached, and this one arrived with a screenshot and a lot of enthusiasm.”

— Lenar Kess, today's narration

An OpenAI researcher says an internal model closed ten open problems in mathematics and theoretical computer science. Nobody has published the ten problems. Today's episode is about what the industry does in the gap between a claim and its evidence — and about four other stories where the evidence is a screenshot, a vendor chart, or a sarcastic Reddit post.

Chapters

  1. 00:00:04 Transcript

Sources

20 cited
  1. 1

    @yonashav (Yo Shavit)

    X yonashav

    This extends a core industry debate (OpenAI vs Anthropic) by speculating on technical advantages (RL, TTC, math proving) that could lead to significant shifts in capability.

    x.com/yonashav/status/2083544031334678662 →
    Details
    Excerpt
    This extends a core industry debate (OpenAI vs Anthropic) by speculating on technical advantages (RL, TTC, math proving) that could lead to significant shifts in capability.
    Context
    This extends a core industry debate (OpenAI vs Anthropic) by speculating on technical advantages (RL, TTC, math proving) that could lead to significant shifts in capability.
    Key points
    • This extends a core industry debate (OpenAI vs Anthropic) by speculating on technical advantages (RL, TTC, math proving) that could lead to significant shifts in capability.
    Provenance
    Tweet · Primary source
  2. 2

    @theojaffee (Theo Jaffee)

    X theojaffee

    The quoted tweet reports a major breakthrough (solving 10 open problems) in a key area (math/theory) linked to a major player (OpenAI), fitting the criteria for a breaking story or significant model release.

    x.com/theojaffee/status/2083562674441625805 →
    Details
    Excerpt
    The quoted tweet reports a major breakthrough (solving 10 open problems) in a key area (math/theory) linked to a major player (OpenAI), fitting the criteria for a breaking story or significant model release.
    Context
    The quoted tweet reports a major breakthrough (solving 10 open problems) in a key area (math/theory) linked to a major player (OpenAI), fitting the criteria for a breaking story or significant model release.
    Key points
    • The quoted tweet reports a major breakthrough (solving 10 open problems) in a key area (math/theory) linked to a major player (OpenAI), fitting the criteria for a breaking story or significant model release.
    Provenance
    Tweet · Primary source
  3. 3

    @willdepue (will depue)

    X willdepue

    The quoted tweet describes a major breakthrough (solving 10 open problems) from an internal model version (Astra), which is a breaking story about capability and scientific reasoning.

    x.com/willdepue/status/2083575278928867680 →
    Details
    Excerpt
    The quoted tweet describes a major breakthrough (solving 10 open problems) from an internal model version (Astra), which is a breaking story about capability and scientific reasoning.
    Context
    The quoted tweet describes a major breakthrough (solving 10 open problems) from an internal model version (Astra), which is a breaking story about capability and scientific reasoning.
    Key points
    • The quoted tweet describes a major breakthrough (solving 10 open problems) from an internal model version (Astra), which is a breaking story about capability and scientific reasoning.
    Provenance
    Tweet · Primary source
  4. 4

    r/LocalLLaMA: EU AI Act takes effect tomorrow, August 2, 2026. 🤡 - 0 pts · 0 comments

    Article xoxaxo

    A major regulatory intervention (EU AI Act) is a core topic. Even if presented lightly, it signals significant policy shifts affecting all builders.

    www.reddit.com/r/LocalLLaMA/comments/1vcqpn… →
    Details
    Excerpt
    A major regulatory intervention (EU AI Act) is a core topic. Even if presented lightly, it signals significant policy shifts affecting all builders.
    Context
    A major regulatory intervention (EU AI Act) is a core topic. Even if presented lightly, it signals significant policy shifts affecting all builders.
    Key points
    • A major regulatory intervention (EU AI Act) is a core topic. Even if presented lightly, it signals significant policy shifts affecting all builders.
    Provenance
    Article · Supporting source
  5. 5

    @VictorTaelin (Taelin)

    X VictorTaelin

    The quoted tweet describes a major model release (Astra) solving complex open problems in math and theory, fitting criteria #1 for CORE content.

    x.com/VictorTaelin/status/20835996505781044… →
    Details
    Excerpt
    The quoted tweet describes a major model release (Astra) solving complex open problems in math and theory, fitting criteria #1 for CORE content.
    Context
    The quoted tweet describes a major model release (Astra) solving complex open problems in math and theory, fitting criteria #1 for CORE content.
    Key points
    • The quoted tweet describes a major model release (Astra) solving complex open problems in math and theory, fitting criteria #1 for CORE content.
    Provenance
    Tweet · Primary source
  6. 6

    @antoniogm (Antonio García Martínez (agm.eth))

    X antoniogm

    A major model developer (OpenAI) conducting security evaluations is a significant industry signal regarding safety and capability boundaries.

    x.com/antoniogm/status/2083600782121972109 →
    Details
    Excerpt
    A major model developer (OpenAI) conducting security evaluations is a significant industry signal regarding safety and capability boundaries.
    Context
    A major model developer (OpenAI) conducting security evaluations is a significant industry signal regarding safety and capability boundaries.
    Key points
    • A major model developer (OpenAI) conducting security evaluations is a significant industry signal regarding safety and capability boundaries.
    Provenance
    Tweet · Primary source
  7. 7

    @_NathanCalvin (Nathan Calvin)

    X _NathanCalvin

    This tweet addresses a major industry topic (AI model failures/incidents) and critiques how companies are framing them, extending an ongoing debate about AI reliability and corporate transparency.

    x.com/_NathanCalvin/status/2083619249042714… →
    Details
    Excerpt
    This tweet addresses a major industry topic (AI model failures/incidents) and critiques how companies are framing them, extending an ongoing debate about AI reliability and corporate transparency.
    Context
    This tweet addresses a major industry topic (AI model failures/incidents) and critiques how companies are framing them, extending an ongoing debate about AI reliability and corporate transparency.
    Key points
    • This tweet addresses a major industry topic (AI model failures/incidents) and critiques how companies are framing them, extending an ongoing debate about AI reliability and corporate transparency.
    Provenance
    Tweet · Primary source
  8. 8

    @kevinroose (Kevin Roose)

    X kevinroose

    The quoted tweet describes a major model family (Astra) solving significant open problems in math and theoretical CS. This is a primary builder artifact that changes the perceived capability of AI models.

    x.com/kevinroose/status/2083632335438905441 →
    Details
    Excerpt
    The quoted tweet describes a major model family (Astra) solving significant open problems in math and theoretical CS. This is a primary builder artifact that changes the perceived capability of AI models.
    Context
    The quoted tweet describes a major model family (Astra) solving significant open problems in math and theoretical CS. This is a primary builder artifact that changes the perceived capability of AI models.
    Key points
    • The quoted tweet describes a major model family (Astra) solving significant open problems in math and theoretical CS. This is a primary builder artifact that changes the perceived capability of AI models.
    Provenance
    Tweet · Primary source
  9. 9

    @brttbmn (Brett Bauman)

    X brttbmn

    This suggests a fundamental shift in how software tasks are managed (from scheduled jobs to AI agents), directly impacting developer workflows and mental models.

    x.com/brttbmn/status/2083641132534083915/ph… →
    Details
    Excerpt
    This suggests a fundamental shift in how software tasks are managed (from scheduled jobs to AI agents), directly impacting developer workflows and mental models.
    Context
    This suggests a fundamental shift in how software tasks are managed (from scheduled jobs to AI agents), directly impacting developer workflows and mental models.
    Key points
    • This suggests a fundamental shift in how software tasks are managed (from scheduled jobs to AI agents), directly impacting developer workflows and mental models.
    Provenance
    Tweet · Primary source
  10. 10

    @gdb (Greg Brockman)

    X gdb

    Discusses a specific, usable capability (cloud browser/live intervention) that changes how developers interact with AI agents, fitting the 'primary builder artifact' criteria.

    x.com/gdb/status/2083641828415602989 →
    Details
    Excerpt
    Discusses a specific, usable capability (cloud browser/live intervention) that changes how developers interact with AI agents, fitting the 'primary builder artifact' criteria.
    Context
    Discusses a specific, usable capability (cloud browser/live intervention) that changes how developers interact with AI agents, fitting the 'primary builder artifact' criteria.
    Key points
    • Discusses a specific, usable capability (cloud browser/live intervention) that changes how developers interact with AI agents, fitting the 'primary builder artifact' criteria.
    Provenance
    Tweet · Primary source
  11. 11

    @badlogicgames (Mario Zechner)

    X badlogicgames

    Claims of a major model release (GPT 5.6) with significant, demonstrable capabilities (breaking DRM) are high-signal and suggest a potential shift in developer workflows or security landscape.

    x.com/badlogicgames/status/2083653919461240… →
    Details
    Excerpt
    Claims of a major model release (GPT 5.6) with significant, demonstrable capabilities (breaking DRM) are high-signal and suggest a potential shift in developer workflows or security landscape.
    Context
    Claims of a major model release (GPT 5.6) with significant, demonstrable capabilities (breaking DRM) are high-signal and suggest a potential shift in developer workflows or security landscape.
    Key points
    • Claims of a major model release (GPT 5.6) with significant, demonstrable capabilities (breaking DRM) are high-signal and suggest a potential shift in developer workflows or security landscape.
    Provenance
    Tweet · Primary source
  12. 12

    @badlogicgames (Mario Zechner)

    X badlogicgames

    Discusses a specific future model (GPT 5.6) and its capabilities for advanced penetration testing/cybersecurity, hitting on frontier models and power struggles.

    x.com/badlogicgames/status/2083654848461852… →
    Details
    Excerpt
    Discusses a specific future model (GPT 5.6) and its capabilities for advanced penetration testing/cybersecurity, hitting on frontier models and power struggles.
    Context
    Discusses a specific future model (GPT 5.6) and its capabilities for advanced penetration testing/cybersecurity, hitting on frontier models and power struggles.
    Key points
    • Discusses a specific future model (GPT 5.6) and its capabilities for advanced penetration testing/cybersecurity, hitting on frontier models and power struggles.
    Provenance
    Tweet · Primary source
  13. 13

    @badlogicgames (Mario Zechner)

    X badlogicgames

    Discusses agentic capabilities taking over tedious tasks and generating complex outputs (machine code), which is a major shift in developer workflows.

    x.com/badlogicgames/status/2083655341770789… →
    Details
    Excerpt
    Discusses agentic capabilities taking over tedious tasks and generating complex outputs (machine code), which is a major shift in developer workflows.
    Context
    Discusses agentic capabilities taking over tedious tasks and generating complex outputs (machine code), which is a major shift in developer workflows.
    Key points
    • Discusses agentic capabilities taking over tedious tasks and generating complex outputs (machine code), which is a major shift in developer workflows.
    Provenance
    Tweet · Primary source
  14. 14

    r/OpenAI: AI firms must answer for rogue bots, says boss of hacked company - 0 pts · 0 comments

    Article KeanuRave100

    This touches directly on corporate governance and regulatory risk (AI liability/rogue bots), a high-signal power struggle topic.

    www.bbc.com/news/articles/cr7k49xjzzeo →
    Details
    Excerpt
    This touches directly on corporate governance and regulatory risk (AI liability/rogue bots), a high-signal power struggle topic.
    Context
    This touches directly on corporate governance and regulatory risk (AI liability/rogue bots), a high-signal power struggle topic.
    Key points
    • This touches directly on corporate governance and regulatory risk (AI liability/rogue bots), a high-signal power struggle topic.
    Provenance
    Article · Supporting source
  15. 15

    @simonw (Simon Willison)

    X simonw

    This reports on new, usable capabilities (browser/screenshots/deployment) within a major AI product, directly impacting how developers build and interact with AI tools.

    x.com/simonw/status/2083704964233462075 →
    Details
    Excerpt
    This reports on new, usable capabilities (browser/screenshots/deployment) within a major AI product, directly impacting how developers build and interact with AI tools.
    Context
    This reports on new, usable capabilities (browser/screenshots/deployment) within a major AI product, directly impacting how developers build and interact with AI tools.
    Key points
    • This reports on new, usable capabilities (browser/screenshots/deployment) within a major AI product, directly impacting how developers build and interact with AI tools.
    Provenance
    Tweet · Primary source
  16. 16

    @Miles_Brundage (Miles Brundage)

    X Miles_Brundage

    This directly addresses a major industry debate: the need for secure software development and AI's role in it. It speaks to corporate strategy (OpenAI making a bet) and infrastructure/governance.

    x.com/Miles_Brundage/status/208371356399958… →
    Details
    Excerpt
    This directly addresses a major industry debate: the need for secure software development and AI's role in it. It speaks to corporate strategy (OpenAI making a bet) and infrastructure/governance.
    Context
    This directly addresses a major industry debate: the need for secure software development and AI's role in it. It speaks to corporate strategy (OpenAI making a bet) and infrastructure/governance.
    Key points
    • This directly addresses a major industry debate: the need for secure software development and AI's role in it. It speaks to corporate strategy (OpenAI making a bet) and infrastructure/governance.
    Provenance
    Tweet · Primary source
  17. 17

    r/OpenAI: GPT-5.6 Sol Raw reasoning leaked on failed tool call attempt - 0 pts · 0 comments

    Article Suspicious_Raise_589

    A purported leak of raw reasoning traces from a next-gen model is a major artifact that directly impacts the understanding of frontier model mechanics and agentic tool use.

    www.reddit.com/r/OpenAI/comments/1vd3wfp/gp… →
    Details
    Excerpt
    A purported leak of raw reasoning traces from a next-gen model is a major artifact that directly impacts the understanding of frontier model mechanics and agentic tool use.
    Context
    A purported leak of raw reasoning traces from a next-gen model is a major artifact that directly impacts the understanding of frontier model mechanics and agentic tool use.
    Key points
    • A purported leak of raw reasoning traces from a next-gen model is a major artifact that directly impacts the understanding of frontier model mechanics and agentic tool use.
    Provenance
    Article · Supporting source
  18. 18

    AI News & Strategy Daily | Nate B Jones · 19s

    Video AI News & Strategy Daily | Nate B Jones

    Chat GPT 5.6 is a dumber model and I love it so much. In fact, I use it all the time. Dumber does not mean dumb. Not remotely. Soul is an incredibly intelligent model. On Asia's last exam, which measures long-running pr…

    www.youtube.com/shorts/lMiHRN8pzn4 →
    Details
    Excerpt
    Chat GPT 5.6 is a dumber model and I love it so much. In fact, I use it all the time. Dumber does not mean dumb. Not remotely. Soul is an incredibly intelligent model. On Asia's last exam, which measures long-running professional work across 55 different fields, Soul set a new high.
    Context
    Discusses a specific model (ChatGPT 5.6/Soul) and its performance on professional benchmarks, signaling a major product update or capability shift.
    Key points
    • Discusses a specific model (ChatGPT 5.6/Soul) and its performance on professional benchmarks, signaling a major product update or capability shift.
    Provenance
    Video · Supporting source
  19. 19

    @gdb (Greg Brockman)

    X gdb

    This hits on agentic tools and workflow changes (cron jobs), which is a core focus of the podcast. It suggests a major shift in how software tasks are managed.

    x.com/gdb/status/2083750556062093745 →
    Details
    Excerpt
    This hits on agentic tools and workflow changes (cron jobs), which is a core focus of the podcast. It suggests a major shift in how software tasks are managed.
    Context
    This hits on agentic tools and workflow changes (cron jobs), which is a core focus of the podcast. It suggests a major shift in how software tasks are managed.
    Key points
    • This hits on agentic tools and workflow changes (cron jobs), which is a core focus of the podcast. It suggests a major shift in how software tasks are managed.
    Provenance
    Tweet · Primary source
  20. 20

    An internal OpenAI Astra model solved 10 major open math and CS problems — 5 pts · 0 comments

    Article wa5ina

    A claim of an internal OpenAI model solving major open math/CS problems is a massive breaking story about frontier AI capability and directly impacts the perceived state-of-the-art.

    twitter.com/polynoamial/status/208346719466… →
    Details
    Excerpt
    A claim of an internal OpenAI model solving major open math/CS problems is a massive breaking story about frontier AI capability and directly impacts the perceived state-of-the-art.
    Context
    A claim of an internal OpenAI model solving major open math/CS problems is a massive breaking story about frontier AI capability and directly impacts the perceived state-of-the-art.
    Key points
    • A claim of an internal OpenAI model solving major open math/CS problems is a massive breaking story about frontier AI capability and directly impacts the perceived state-of-the-art.
    Provenance
    Article · Supporting source