← July 12, 2026

Start of day · analyzed 2026-07-12 06:39:52 PT

Morning brief

Sunday, July 12, 2026

Overnight developments and what deserves attention today.

40sources scanned
21new signals
22edge cases kept
13confirmed
ListenEnglish edition

📡 Jin Miao Signals — Morning Brief · 2026-07-12

1. Top 5 — what actually matters today

  • Terry Tao publishes on building real apps with coding agents — a Fields Medalist writing field notes on what modern coding agents actually do to a working developer's loop is the rare seminal-voice post that outranks a hundred product launches; if you build software, read this before your next sprint. terrytao.wordpress.com (NEW, Confirmed)
  • AgentLens: trajectory-level evals for coding agents — [PRIORITY] the shift from "did the task pass?" to grading how the agent used tools, recovered from mistakes, and talked to you is the eval paradigm operators have been missing — this is the layer teams will buy once they stop trusting single-bit pass rates. huggingface.co/papers (carrying over; the production-assessed framing is what advanced it past yesterday's chatter)
  • Confessor: replay what private info Claude Code touched on your machine — the average power-user has zero visibility into what a local coding agent read or exfiltrated; a one-click "what did it access" replay is a direct cognitive-sovereignty play and a preview of where agent governance goes. github.com/ninjahawk (NEW, Confirmed)
  • Google rolls AlphaEvolve out widely to Cloud customers — evolutionary algorithm-discovery (chip design, routing, research) moving from lab demo to a self-serve enterprise capability is a real founder/markets signal — watch which optimization-heavy verticals it undercuts. blog.google (ONGOING — what changed: general availability to Cloud customers)
  • Mercor in talks for a $20B valuation — [PRIORITY] a 2× step-up from October's $10B in months tells you the market is pricing AI-labor/data-labeling as core infrastructure, not a side business — a tell for where the next wave of talent-and-data capital flows. techcrunch.com ⚠️ Rumor — valuation unconfirmed.

2. New-direction sparks

  • "AI agent forensics" is quietly becoming a category. Three independent builders shipped tools to replay what an agent did: Confessor (what Claude Code accessed) github, a reverse-engineered dump of what Grok Build CLI sends to xAI gist, and Mindwalk (replay agent sessions on a 3D map of your codebase) github. Non-obvious because everyone is racing to build agents; almost no one is building the audit/trust layer underneath them.

3. Threads worth watching

  • Embodied / tactile robotics (radar: embodied AI) — two fresh papers push touch as a first-class modality: OmniTacTune (policy-agnostic real-world RL for tactile residual adaptation) huggingface.co and Splash (mask-isolated tactile alignment in MLLMs) huggingface.co. Contact-rich manipulation is where vision-only priors keep failing — worth tracking as the sensor-fusion bottleneck for humanoids.

4. Contrarian watch

  • Non-LLM foundation models go zero-shot-local. [OUTLIER] Someone wrapped Google's TabFM & TimesFM as a 100% local MCP server for zero-shot forecasting/classification/regression [r/MachineLearning]. Consensus is "LLM for everything"; the edge is that small, specialized structured-data FMs may quietly own tabular/time-series tasks LLMs are bad at. ⚠️ Rumor/self-post — unverified.
  • The human backlash is organizing. [OUTLIER] WSJ on hard-line anti-AI activists "ramping up for the war with AI" wsj.com. Underpriced sentiment/regulatory risk for anyone shipping consumer AI.
  • Interpretability is getting eerie. [OUTLIER] Anthropic's "Jacobian lens" claims the clearest look yet inside Claude — findings ranging "from the mundane to the unnerving" technologyreview.com. Watch whether this becomes a safety-marketing edge or a liability.

5. Verification flags

  • ⚠️ Mercor $20B valuation — do not act on yet — needs primary source. [Rumor] techcrunch.com
  • ⚠️ Gradium $100M seed, Nvidia-backed — do not act on yet — needs primary source. [Rumor] techcrunch.com
  • ⚠️ TabFM/TimesFM zero-shot MCP (Zer0Fit) claims — do not act on yet — self-reported, needs primary source. [Rumor] [r/MachineLearning]
  • ⚠️ Qwen3.5-122B "daily driver on Mac Studio" bugfix claims — do not act on yet — needs reproduction. [Rumor] mrzk.io

Markets context only — not financial advice.

Private founder layer

Co-founder confidential

Strategic synthesis and adversarial review, encrypted in the page source.

Source ledgerEvery scored item, including outliers
  1. RumorNEWOutlier
    Zer0Fit: I took Google's new TabFM & TimesFM ML foundation models and made them available as an MCP server for zero-shot ML tasks (forecasts / classifications / regressions). 100% local. [P]reddit/r/MachineLearning
    i5 / e5
  2. ConfirmedONGOINGOutlier
    i5 / e5
  3. ConfirmedNEWOutlier
    i4 / e5
  4. ConfirmedNEWOutlier
    i4 / e5
  5. RumorNEWOutlier
    i4 / e5
  6. ConfirmedNEWOutlier
    i4 / e5
  7. ReportedONGOINGOutlier
    i5 / e4
  8. ConfirmedONGOINGOutlier
    i3 / e5
  9. ConfirmedONGOINGOutlier
    i3 / e5
  10. RumorONGOINGOutlier
    i4 / e4
  11. RumorONGOINGOutlier
    i4 / e4
  12. ConfirmedNEWOutlier
    i3 / e4
  13. ConfirmedNEWOutlier
    i3 / e4
  14. ReportedONGOINGOutlier
    i3 / e4
  15. ConfirmedONGOINGOutlier
    i3 / e4
  16. ConfirmedONGOINGOutlier
    i3 / e4
  17. ConfirmedONGOINGOutlier
    i4 / e3
  18. ReportedONGOINGOutlier
    i4 / e3
  19. ReportedNEWOutlier
    i2 / e3
  20. RumorNEWOutlier
    Context and average best linear mappings [D]reddit/r/MachineLearning
    i2 / e3
  21. RumorNEWOutlier
    Obtaining Irregular Learning Curves with HyberBand Tuned ANN model for Price Prediction [P]reddit/r/MachineLearning
    i2 / e3
  22. ReportedONGOINGOutlier
    i2 / e3
  23. ReportedNEW
    i4 / e4
  24. ConfirmedNEW
    i4 / e4
  25. ReportedNEW
    i4 / e4
  26. ReportedONGOING
    i5 / e3
  27. ReportedNEW
    i3 / e3
  28. ReportedNEW
    i3 / e3
  29. ReportedNEW
    i3 / e3
  30. ReportedNEW
    i3 / e3
  31. ReportedONGOING
    i4 / e2
  32. ConfirmedONGOING
    GPT-5.6hackernews
    i5 / e1
  33. ReportedNEW
    i2 / e3
  34. ReportedNEW
    i3 / e2
  35. ReportedONGOING
    i3 / e2
  36. ReportedONGOING
    i3 / e2
  37. ReportedONGOING
    i3 / e1
  38. ReportedONGOING
    i2 / e1
  39. ReportedNEW
    i1 / e1
  40. ReportedNEW
    i1 / e1