← August 2, 2026

Start of day · analyzed 2026-08-02 06:40:53 PT

Morning brief

Sunday, August 2, 2026

Overnight developments and what deserves attention today.

60sources scanned
55new signals
36edge cases kept
8confirmed
ListenEnglish edition

📡 Jin Miao Signals — Morning Brief · 2026-08-02

1. Top 5 — what actually matters today

  • Kimi K3 reportedly serves better performance-per-dollar on AMD MI355X than on Nvidia B300 — first credible cross-vendor economics on a frontier open-weights model; if it survives independent replication, the "just buy Blackwell" default becomes a per-workload decision for anyone sizing an inference fleet, and it's the kind of datapoint that gets read against AMD/Nvidia into Monday's open. Self-reported bench — treat as directional, not settled. wafer.ai
  • An internal OpenAI model called "Astra" reportedly solved 10 major open math and CS problems — posted by a named OpenAI researcher, not a press release, which is why it matters; the story is no longer "AI does contest math" but "AI closes problems humans left open," and that reframes what a research hire is for. [Rumor] — single social post, no paper, no problem list. @polynoamial (pairs oddly well with the human-only proof landing this week that the Burau representation is faithful at n=4 — arXiv)
  • Gemini Robotics ER 2 — embodied reasoning shipped against two real robot bodies (Duo and Apollo) — the embodied-reasoning layer is being productized as a model, not a research demo, which is the step that lets robotics teams stop building their own perception-to-plan stack. Fair warning: this is ONGOING, not overnight — I'm carrying it because a robotics foundation model clears my bar regardless of news cycle, not because it broke last night. Google DeepMind
  • Only 8.9% of sites block AI crawlers — and 94.8% are never cited in an AI answer — the open-web bargain quietly inverted: you're paying the training cost and getting none of the distribution. For any founder whose funnel starts with search, that's a channel that's already gone, not going. AI Visibility Index
  • Antora closes a $550M Series C for thermal battery storage, explicitly aimed at AI datacenter demand — the marginal dollar in "AI infrastructure" keeps moving downstream from chips to electrons; watch this as the tell for where 2027 compute actually gets sited, and as context for power/industrial names levered to datacenter buildout. [Rumor / ONGOING] — round size not primary-sourced. Crunchbase News

2. New-direction sparks

  • Frontier weights are collapsing into consumer memory faster than anyone's roadmap assumed. Overnight on r/LocalLLaMA: llama.cpp landed MTP/DSpark support for DeepSeek V4-Flash, a claimed 284B V4-Flash run in 5.3 GB, Kimi K3 pushed onto a single CPU with 8 GB (we were at 29 GB yesterday morning), and 12–15 tok/s on 3090s and MI50s. Non-obvious because none of it is a model release — it's runtime plumbing, the least-covered and most leverage-dense layer. If this holds, the unit of "who can serve a frontier model" moves from a cluster to a laptop inside a quarter. All [Rumor], all self-reported, no links in feed — r/LocalLLaMA threads.
  • "The Greenhouse and the Lens" — a working taxonomy for two distinct modes of agentic work. Rare thing: someone naming the shape of agent collaboration instead of shipping another harness. Non-obvious because the field is drowning in tooling and starved of vocabulary, and vocabulary is what lets teams argue productively about which mode they're in. brethorsting.com

3. Threads worth watching

  • Human–AI interaction — moved materially, from an unlikely source. Greg Brockman: at OpenAI, people hook ChatGPT into Slack, and colleagues hate it — they'd happily do the same task if a human asked. The rejection isn't about capability, it's about relationship. That's the sharpest empirical read I've seen on the ceiling of agent-to-human delegation, and it came out as a throwaway quote. simonwillison.net

4. Contrarian watch

  • Nvidia's inference moat is being priced as permanent; the MI355X number says it's per-workload. One blog post isn't a thesis, but it's the first cross-vendor perf/$ claim specific enough to be falsified. Context only, but note the Guardian's read this morning that market turmoil is exposing how opaque the AI economy's actual unit costs are. Guardian
  • Consensus: frontier capability requires datacenter capital. Edge: 284B params in 5.3 GB. If even half the local-inference claims replicate, the "compute is the moat" argument is weaker at the serving layer than the capex narrative implies. Watch whether these get reproduced by anyone with a name attached this week.
  • The EU AI Act's general applicability date is today, 2026-08-02. Circulating on r/LocalLLaMA as a punchline; it is not a punchline for anyone shipping into the EU. No link in feed — verify against the official Official Journal text before you act on any of it, including mine.
  • AI firms are reportedly scanning rare book editions and destroying the physical copies afterward. Non-consensus because the training-data fight has been framed entirely as copying; this frames it as loss. Different legal theory, different constituency, much worse optics. Dallas Express

5. Verification flags

  • ⚠️ do not act on yet — needs primary source: OpenAI "Astra" solving 10 major open math/CS problems. Single social post, no problem list, no paper. @polynoamial
  • ⚠️ do not act on yet — needs primary source: Antora's $550M Series C size and lead investor — secondary reporting only. Crunchbase News
  • ⚠️ do not act on yet — needs primary source: every local-inference number above (V4-Flash in 5.3 GB, K3 on 8 GB CPU, 12–15 tok/s figures) — anonymous, self-reported, no reproducible configs.
  • ⚠️ do not act on yet — needs primary source: MI355X-vs-B300 perf/$ — vendor-adjacent blog, no third-party replication. wafer.ai
  • ⚠️ do not act on yet — needs primary source: EU AI Act applicability details as described in social chatter — read the statute, not the thread.

Markets context only — not financial advice.

Private founder layer

Co-founder confidential

Strategic synthesis and adversarial review, encrypted in the page source.

Source ledgerEvery scored item, including outliers
  1. ReportedNEWOutlier
    i5 / e5
  2. RumorNEWOutlier
    I pushed Kimi K3 onto one CPU with 8 GB of RAMreddit/r/LocalLLaMA
    i5 / e5
  3. RumorNEWOutlier
    DeepSeek-V4-Flash 284B on 5.3GB of memoryreddit/r/LocalLLaMA
    i5 / e5
  4. RumorNEWOutlier
    llama.cpp just added MTP / DSpark support for DeepSeek V4 Flashreddit/r/LocalLLaMA
    i5 / e5
  5. RumorNEWOutlier
    Setting up of a 16xGB10 (DGX Spark) clusterreddit/r/LocalLLaMA
    i4 / e5
  6. RumorNEWOutlier
    PSA for DeepSeek-V4-Flash-0731 users — don't blow out your prompt cache with system role messages mid-conversationreddit/r/LocalLLaMA
    i4 / e5
  7. RumorNEWOutlier
    Ran DS V4-Flash-0731 Locally on 3xMI50 32GB @ ~15 t/s TGreddit/r/LocalLLaMA
    i4 / e5
  8. RumorNEWOutlier
    i5 / e4
  9. RumorNEWOutlier
    i4 / e4
  10. ReportedNEWOutlier
    i4 / e4
  11. ReportedNEWOutlier
    i4 / e4
  12. RumorNEWOutlier
    DeepSeek-V4-Flash-0731 UD-IQ3_S 12.5 tok/s on RTX 3090 +128GB DDR5reddit/r/LocalLLaMA
    i4 / e4
  13. ConfirmedONGOINGOutlier
    i4 / e4
  14. ConfirmedONGOINGOutlier
    i4 / e4
  15. RumorONGOINGOutlier
    i4 / e4
  16. ConfirmedONGOINGOutlier
    i2 / e5
  17. ConfirmedNEWOutlier
    i3 / e4
  18. ReportedNEWOutlier
    Agent4Leasehackernews
    i3 / e4
  19. RumorNEWOutlier
    Vacuum 16Treddit/r/LocalLLaMA
    i3 / e4
  20. RumorNEWOutlier
    Xberg v1 is outreddit/r/LocalLLaMA
    i3 / e4
  21. ReportedNEWOutlier
    i3 / e4
  22. ReportedONGOINGOutlier
    i3 / e4
  23. ReportedNEWOutlier
    i4 / e3
  24. RumorNEWOutlier
    EU AI Act takes effect tomorrow, August 2, 2026. 🤡reddit/r/LocalLLaMA
    i4 / e3
  25. ReportedNEWOutlier
    i2 / e4
  26. ConfirmedNEWOutlier
    i2 / e4
  27. ConfirmedNEWOutlier
    i3 / e3
  28. ReportedNEWOutlier
    i3 / e3
  29. ReportedNEWOutlier
    i2 / e3
  30. RumorNEWOutlier
    No replies to rebuttals and comments even by AC [D]reddit/r/MachineLearning
    i2 / e3
  31. ReportedNEWOutlier
    i2 / e3
  32. ReportedNEWOutlier
    i2 / e3
  33. ReportedNEWOutlier
    i2 / e3
  34. ReportedNEWOutlier
    i2 / e3
  35. ReportedNEWOutlier
    i2 / e3
  36. ReportedNEWOutlier
    i2 / e3
  37. ReportedNEW
    i3 / e3
  38. ReportedNEW
    i3 / e3
  39. ReportedNEW
    i2 / e3
  40. ReportedNEW
    i2 / e3
  41. ReportedNEW
    i2 / e3
  42. ReportedNEW
    i2 / e3
  43. ReportedNEW
    i3 / e2
  44. ReportedNEW
    Diátaxishackernews
    i3 / e2
  45. ReportedNEW
    i3 / e2
  46. ReportedNEW
    i2 / e2
  47. ReportedNEW
    i2 / e2
  48. ReportedNEW
    i2 / e2
  49. ReportedNEW
    Seedance 2.5hackernews
    i2 / e2
  50. ConfirmedNEW
    NetBSD 11.0hackernews
    i2 / e2
  51. ReportedNEW
    i2 / e2
  52. ReportedNEW
    i2 / e2
  53. ReportedNEW
    i2 / e2
  54. ReportedNEW
    i2 / e2
  55. ConfirmedNEW
    i1 / e2
  56. RumorNEW
    ARR August Cycle [D]reddit/r/MachineLearning
    i1 / e2
  57. RumorNEW
    Question about NeurIPS discussion phase [D]reddit/r/MachineLearning
    i1 / e2
  58. RumorNEW
    Discussion About Meta-Review Issue Report in ARR Cycle [D]reddit/r/MachineLearning
    i1 / e2
  59. ReportedNEW
    i1 / e1
  60. ReportedNEW
    i1 / e1