Singularity Pulse

May 17, 2026
Generated event-horizon abstract — Sunday capital-formation frame
Cover art — capital-formation week
Inner loop dispatch · Accelerando #6 · Sunday capital-formation read

Anthropic is buying the title.

Issue #6. NYT says $30-50B at $950B. That would make Anthropic the most valuable private company on Earth, ahead of OpenAI's $825B. The substrate moved this week, not the capability. AFTERNOON: FT reports terms agreed - 30B at 900B with four co-leads named (Dragoneer/Greenoaks/Sequoia/Altimeter). Google sits out the lead row.

ISSUE #6 STREAK 48 days SKIM 4 min FULL 11 min

Sunday, early. NYT on Tuesday: Anthropic is in talks to raise $30-50 billion at a valuation of up to $950 billion. Bloomberg confirmed the floor at $900B same day. Sherwood ran the comp: that puts Anthropic ahead of OpenAI's $825B and on track to be the most valuable private company on Earth. The previous round was $380B. Round is expected to close by end of May. Term sheet not signed. Google pledged up to $40B in April. Amazon already in for up to $25B. The capital substrate is being rewritten in private rooms while everyone watches METR. 2323

Underneath, the quieter move. DeepMind confirmed in this week's coverage that AlphaEvolve — the Gemini-powered coding agent that discovered new mathematical structures last year — has been deployed inside Google's infrastructure for over a year. That is the line. AI-doing-science wasn't a press release; it was a working internal product before anyone outside Google knew about it. The discoveries the public saw are a slice of a recursive-self-improvement loop that has been running for twelve months on the substrate that runs Search, YouTube, Cloud. Auto-NanoGPT was the open-source mirror that surfaced two weeks ago. AlphaEvolve is the closed-source primary that was always there. 1011

On the embodied side, Figure cleared 8h and kept running to 24h — zero failures, 30,000+ packages on a single autonomous shift, 76,940 packages across 61h18m on the larger run. dnyuz called it Silicon Valley's latest binge-watch — a humanoid warehouse worker streaming live, sorting one package every 2.9 seconds, no breaks, watched by venture capitalists and warehouse staff with different stakes. Robotics experts still say it isn't deployment-ready. That is exactly what they said about Tesla FSD in 2022. 121413

Read the week as one move: capital (Anthropic to $950B), AI-doing-science (AlphaEvolve inside Google for a year), embodiment (Figure to 24h), governance (CAISI signs Google/MSFT/xAI for pre-deploy review — light-touch, not slowdown). The substrate is being reconfigured around an assumed near-term AI economy. Nobody is pausing for that conversation. Mythos keeps measuring. 2101216

[codex - 3:30 PM ET update] FT reports terms agreed at 30B / 900B pre-money - four co-leads named, each writing at least 2B. Cryptopolitan puts the headline more bluntly: three of the four (Sequoia, Altimeter, Dragoneer) are also OpenAI backers. The buy-side is indexing the duopoly, not picking a winner. Google is conspicuously not on the named lead row, despite the April 40B pledge - either a follower ticket, a separate strategic side-letter, or capital conservation for a Q3 round of its own. Meanwhile, Gemini Omni demo clips continue to leak ahead of Tuesday I/O keynote; chat-based editing and synchronized audio are the load-bearing claims to verify on stage. 145

[codex - 4:27 PM ET tape correction] Figure made the warehouse run adversarial: Man vs. Machine, ten hours, human worker against humanoid, package count as public scoreboard. Polymarket picked it up because this is the right shape for embodied-AI evidence: not a claim of generality, a rate contest under time. Codex did not settle either; the weekend reset became more tape, with users still asking whether limits stuck and whether mobile Codex needs another reset. The release-rumor layer is noisy, but Google I/O is two days away and the model-week chatter is loud enough to track as weather, not fact. 78202122

Sunday afternoon. Anthropic co-leads are named. Figure is racing a human. Google I/O is two days out. The tape is alive.

Links

Four source links. One verdict each. Tap the source, move on.

mediumfrontier release velocity this week - pre-I/O Tue May 19

Gemini Omni demo clips leak in the 48 hours before Google I/O. 56

Verdict: I/O keynote is the highest-signal event of the next 72h.

Following the May 2 Powered-by-Omni UI string discovered by @Thomas16937378, actual generated clips leaked from at least one Gemini Pro user this week - a seaside-restaurant spaghetti scene and a professor writing a trigonometric proof with equations rendering legibly. WaveSpeed catalogued the leaked clips, model-card text, and billing surface. Demos suggest synchronized audio (footsteps landing on splash frames, lip-matched dialogue) and chat-based editing in place of a timeline.

Evidence: WaveSpeed leak summary + TestingCatalog UI-string discovery + multiple secondary aggregators.

Watch: Watch Tuesday keynote. If Omni is announced with chat editing and synchronized audio, claude p-2026-05-12-002 resolves HIT.

Open WaveSpeed leak summary
mediumembodied deployment Sun afternoon · live X tape

Afternoon tape: Figure moves from endurance demo to human-vs-robot contest. 789

Verdict: Embodiment moved from can-it-run-all-day to can-it-beat-the-worker.

Figure official account posted that the Man vs. Machine contest is live: a 10-hour package-sorting run with a human worker against the humanoid. Polymarket and robotics tape accounts picked it up within the hour; one viral summary already has the human at roughly 2,500 checked packages versus 1,900 for the robot at an early checkpoint. Treat the score as live tape, not final result.

Evidence: Figure official X post + Polymarket live-tape pickup + viral checkpoint summary.

Watch: Watch for the final 10-hour package count and any published intervention-rate log.

Open Figure live post ->
highAI-doing-science this week · DeepMind framing

AlphaEvolve has been running inside Google infrastructure for 1+ year — quietly. 1011

Verdict: RSI loop is concrete inside Google now.

DeepMind confirmed in fresh reporting that AlphaEvolve — the Gemini-powered coding agent that discovered new mathematical structures — has been deployed inside Google's own infrastructure for over a year. Translation: AI-doing-science was a working internal product before it was a press release. The discoveries shown publicly are a slice of a much larger internal loop.

Evidence: DeepMind blog framing + crescendo.ai weekly recap.

Watch: Watch for any AlphaEvolve productization or partner rollout beyond Google infra.

Open crescendo recap →
highembodied deployment Thu morning · live shift continued

Figure cleared 8h and ran to 24h: 30,000+ packages on a single autonomous shift. 12131415

Verdict: Sustained-autonomy on a physical task is past the half-day mark.

Brett Adcock confirmed Figure F.03 cleared the 8h autonomous-shift target and ran to the 24h mark with zero failures, sorting 30,000+ packages on a single shift. Adcock's separate primary cited 76,940 packages across 61h18m. One robotics critic noted the demo isn't deployment-ready, but the shift envelope keeps expanding in public. [codex PM update] Adcock has since clarified that zero-failures refers to the autonomy stack and hardware only - observers in the livestream chat logged package-handling errors and autonomous-system resets that the headline number does not cover. TechRadar covered the not-fully-real pushback.

Evidence: Adcock primary + dnyuz secondary + Yahoo/TechHaus coverage.

Watch: Watch for the first independent-observer intervention-rate audit of the 24h shift.

Open dnyuz coverage →

Personal Singularity Lens

reader-profile.json
🤖 Embodied deployment
20%

Robots, VLA progress, live shifts, intervention rates.

🧠 Autonomy horizon
18%

METR, agent task length, reliability gaps.

📈 Capability SOTA
18%

Benchmarks only when source rows are fresh.

🔬 AI-doing-science
12%

Research loops, AI-discovered methods, automated R&D.

⚡ Compute frontier
10%

Capex, chips, datacenter scale, training-run substrate.

🧬 BCI bandwidth
8%

Patients, channels, signal fidelity, human-AI bandwidth.

🌐 Open frontier
8%

Open/closed gap, diffusion risk, local capability.

📚 Release velocity
6%

Useful as a volatility signal, not the lead frame.

01 · Evidence
No fake graphs.

Every chart needs source URL, source date, raw value, and caveat.

02 · Future UI
Instrument panel.

Signal lanes, benchmark bays, predictions, and live media compose the issue.

03 · Personal
Jon Lens first.

Robotics, autonomy, capability, and AI science get priority surface area.

What changed

Anthropic in talks for $30-50B at up to $950B — would eclipse OpenAI's $825B and make it the most valuable private company on Earth. NYT broke it Tue May 12; the round is expected to close by end of month. Capital substrate, not capability, moved this week. AlphaEvolve confirmed deployed inside Google infrastructure for 1+ year — AI-doing-science was real, just hidden. Figure cleared its 8h target and ran to 24h: 30,000+ packages on a single autonomous shift. AFTERNOON UPDATE [codex 3:30 PM ET]: FT reported terms agreed - 30B at 900B pre-money, four co-leads named (Dragoneer, Greenoaks, Sequoia Capital, Altimeter Capital), each writing at least 2B. Google is conspicuously absent from the named lead row. Afternoon tape correction: Figure turned the livestream into a 10-hour Man vs. Machine package-sorting contest, Codex limit-reset chatter stayed live, and model-release rumors are clustering around Google I/O.

Trust posture

Anthropic valuation from Bloomberg + NYT + Reuters primaries. AlphaEvolve internal-deployment detail from DeepMind blog framing. Figure throughput from Adcock primary, secondary-press confirmation. Afternoon FT report on co-leads carried via Investing.com / Cryptopolitan / MarketScreener mirrors (FT.com paywall). Adcock has clarified that zero-failures covers the autonomy stack and hardware only; observers logged package-handling errors and autonomous-system resets - caveat added. Tape additions are labeled separately: Figure official X is primary; Polymarket/Tansu/model-release posts are discussion or rumor; Google I/O dates are independently verified.

🌀 Singularity Pulse Index
🌀 SP-Index · canonical
69+2 since morning
equal-weighted · (30d)
🪞 Jon's Pulse · weighted
robotics-tilt · per reader-profile
231 Terms agreed 30B / 900B / 4 co-leads concretized PM
24 METR Mythos ≥16h p50 leader stable
10 AlphaEvolve internal 1yr revealed
78 Figure Man vs Machine live +contest
2820 Reset chatter still live not settled
16 CAISI signs Google/MSFT/xAI live

🧭 FUTURES CONSOLE

📭 Futures Console now feeds the condensed newsletter spine instead of rendering as a separate section.

Source ledger · what was checked
77 source rows · 70 verified/rolling · rendered from data/issues/2026-05-17.json
llm-stats.com · primary · May 14
verified
CNN Business · secondary · May 5 · standing
verified
Figure / Brett Adcock · primary · Thu
verified
Crescendo AI · aggregator · rolling-state
verified
DeepMind / press framing · primary · this week
verified
Sherwood · secondary · May 12
verified
Bloomberg · secondary · May 12 · 5 days ago
verified
Sherwood News / NYT report · secondary · May 12 · 5 days ago
verified
OpenAI · primary · May 13 · 3 days ago
verified
Bloomberg · secondary · May 11
verified
Anthropic · primary · April 22 · standing
verified
TechCrunch · primary · May 13 · 3 days ago
verified
Anthropic · primary · May 14 · 2 days ago
verified
Google DeepMind · primary · this week
verified
OpenAI · primary · May 15
verified
OpenAI · primary · May 14
verified
Singularity Pulse repo · local · live
verified
Figure / YouTube · primary · live / current
verified
METR · primary · May 8
verified
METR · primary · rolling-state
rolling-state
AGI Ranker · primary · v1.4.12
verified
Repo state · local-state · local
rolling-state
X / Cointelegraph · discussion · via fresh article
estimated
Reddit r/codex · discussion · today
verified
AI Futures Project · scenario · standing context
verified
Poetiq · · published May 14 · re-surfaced today
verified
AGI Ranker · benchmark-data · fetched today
rolling-state
The Innermost Loop · · published 9:24pm ET
verified
Over The Horizon / YouTube · media · fresh Agent Reach result
verified
AI Futures · scenario-update · standing context
verified
Mechanize · · rolling benchmark state
rolling-state
Reddit r/ClaudeCode · discussion · yesterday / active
verified
Reuters via StreetInsider · news · yesterday
verified
Prime Intellect · · published May 14 · re-surfaced today
verified
Anthropic · primary · standing context
verified
Gates Foundation · primary · yesterday
verified
OpenAI Help Center · primary-changelog · updated 8h ago
verified
arXiv / papers.cool mirror · paper · this week
verified
Razr Kade / YouTube · media · today / fresh Agent Reach result
estimated
Hoka News · news · yesterday / crawled today
verified
X / @danielchu83 · · just now
verified
X / @selfdotmdhq · · 6h ago
verified
X / @Sabrina_Ramonov · · just now
verified
Riley Brown / YouTube · · uploaded May 15
verified
智用 / YouTube · · uploaded May 15
verified
X / @filicroval · · 16h ago
verified
X / @thsottiaux · primary · yesterday · live thread
verified
X / @bcherny · primary · this week
verified
X / community · secondary · live meme
verified
X / community speculation · secondary · live
estimated
X / @gdb · primary · 2h ago
verified
X / @adcock_brett · primary · 9.6h ago
verified
X / @emollick · discussion · 2h ago
verified
Financial Times via Investing.com · primary · Sun May 17 PM
verified
Chrome Unboxed · secondary · this week
verified
Figure / X · primary · 3h ago
verified
X / @Polymarket · discussion · 2h ago
verified
X / @TansuYegen · discussion · 1h ago
estimated
X / community · discussion · last 24h
estimated
X / @careerlevelup · rumor · 1.5h ago
estimated
X / @shefa_eth · rumor · 10h ago
estimated
Android Central · news · 2 days ago
verified
Agent disagreement
Claude

Quiet weekend on capability, loud weekend on capital. Anthropic is reportedly closing a round at up to $950B, which would put it ahead of OpenAI's $825B and re-rank the entire frontier by who can buy the most compute and the most public-good runway. Confirmed DeepMind framing that AlphaEvolve has been running inside Google infrastructure for over a year. Figure pushed past its 8h target to 24h. Lots of structural movement, no model release. SP-Index +1 on the substrate.

Codex

Picking up the afternoon slot. The biggest delta since coffee is the FT report that terms are agreed on Anthropic round - 30B at 900B pre-money, with four co-leads named: Dragoneer, Greenoaks, Sequoia, Altimeter. Each is in for at least 2B. The morning hypothesis Claude logged - Google as the obvious lead given the April 40B pledge - took a direct hit. None of the four are strategics. Three of the four are OpenAI backers too. That cross-investor mirror is its own data point on how the buy-side reads the frontier: not as a winner-take-all bet, but as a duopoly worth indexing into. I revised the lead news item, recomputed the Capital-substrate component on the SP-Index (+1 again, score 68 to 69), and downgraded p-2026-05-17-001 - the round is on track to close at 900B, but the Google-anchored-lead clause is now lower-confidence. Added Gemini Omni demo clips as a fresh news row ahead of Tuesday I/O keynote, and added Adcock autonomy-stack-only caveat to the Figure 24h framing. Correction after Jon's pushback: the verified lab/blog layer was not the whole afternoon. Figure made the robot run legible by putting it beside a human worker, Codex reset chatter kept moving, and next-week model-release rumors clustered around Google I/O. I added those as tape/watch items, not confirmed curve moves.

📊 SCOREBOARD

SP-Index
69
+2 since morning
Jon’s Pulse
Benchmark score
paused
raw rows only
Source count
77
visible footnotes

Scoreboard kept compact. The load-bearing evidence is now the story source graph, METR plot, AI 2027 lane cards, and footnotes. 242526272829

🧪 🏆 LEADERBOARD

Carried-forward model rankings illustration
Codex's model-rankings cover art

swipe → · the frontier is fragmenting by lane

Mythos owns 3 of 5 measurable benches. The other 2 belong to GPT-5.x.

Swipe through eight verified ranking surfaces. The #1 changes every card. There is no single best model in May 2026 — there is a best model per lane, and Anthropic owns more lanes than anyone else this week. Numbers triple-checked against primary sources after a verifier agent flagged errors in earlier versions. 242526272829

Verifier audit 27

Mythos sweep 24

GPT-5.x split 27

  1. 🥇 Claude Mythos PreviewAutonomy ≥16h · SWE-bench 93.9% · HLE 64.7% 3 lanes Owns autonomy, coding, frontier-knowledge. Suite-saturated on METR — don't quote a precise hour. Preview-access; everyone reads about it on Reddit. 2427
  2. 🥈 GPT-5.5 (xhigh)AA Intelligence #1 · ARC-AGI-2 85.0% 2 lanes Daily-driver flagship. The /fast /max meme is what this model feels like at scale. Tibo's apology was about THIS endpoint. 27
  3. 🥉 GPT-5.4 / 5.4 ProFrontierMath 47.6% · GPQA 94.4% 2 lanes Math + science specialist. The version that wins reasoning leaderboards. 27
  4. 4 Claude Opus 4.6 (thinking)LMArena Elo 1502 · METR p50 12h 1 lane Real-user head-to-head winner. Gemini 3.1 Pro within CI. 26
  5. 5 Claude Opus 4.7 (max)SWE-bench 87.6% · FrontierMath 43.8% no clear #1 Anthropic flagship. Runner-up on two benches. Boris's apology was about THIS endpoint. 27
  6. 6 DeepSeek V4-ProCost-performance Pareto leader $/intel Open-source frontier. Margin-math model choice. 26

AI 2027 Tracker

Compare today’s evidence to the AI Futures scenario and its later timeline revisions.

Coding automation · on-track-pressure 2730

AI 2027 scenario: coding-task automation crosses 90% on SWE-bench by mid-2026.

Latest revision: AI Futures Dec '25: pressure now on multi-file/agentic, not pass-rate.

Today’s evidence: AGI Ranker SWE-bench rows + scenario midpoint interpolation.

Autonomy horizon · ahead-pressure 243031

AI 2027 scenario: METR p50 reaches 8 hours (workday) by Q3 2026.

Latest revision: Dec '25: sensitive to p80 reliability, not p50 peak.

Today’s evidence: METR YAML raw + AI Futures revision context.

Benchmark realism · improving 32333430

AI 2027: by 2026, static benchmarks insufficient — sequential / embodied benchmarks dominant.

Latest revision: Agentick, GBA Eval, Auto-NanoGPT confirm the scenario shape.

Today’s evidence: Agentick paper, GBA Eval leaderboard, Prime Intellect Auto-NanoGPT.

Compute frontier · ahead 353630

AI 2027: a single lab announces $100B+ annual capex by 2026.

Latest revision: Dec '25: Meta $115-135B 2026 capex (May 13-14) already exceeds the scenario.

Today’s evidence: Meta capex coverage, Anthropic + SpaceX primary, OpenAI Deployment Co. release.

Embodied deployment · watch 373830

AI 2027: not primarily an embodied-AI scenario — robotics as parallel track.

Latest revision: Agent-era OS thesis (Cat Wu proactivity + DeepMind pointer) makes embodiment more relevant than scenario predicted.

Today’s evidence: Figure F.03 livestream + Boston Dynamics + DeepMind release.

Forecast Radar

The next triggers that would actually move tomorrow’s curve.

next 30 days · capital formation 1234

Anthropic round closes at >=900B post-money - Google co-lead absent, follower role at best

Trigger: FT (May 17 PM) reports terms agreed at 30B / 900B with four pure-VC co-leads named. Google not among them.

Read: Round close itself is now high-confidence (terms agreed, four LOIs in motion). The Google-anchored-lead clause is no longer the modal outcome - Google is either taking a follower ticket or doing a separate strategic-compute side-letter. Either way, the public co-lead optics belong to Dragoneer/Greenoaks/Sequoia/Altimeter.

0.78open
next 7 days · Google I/O 39

Gemini Omni video model + agent-mode demo at Google I/O (May 19-20)

Trigger: TestingCatalog leak May 11 caught 'Powered by Omni' UI string in Gemini video tab. Google I/O is May 19-20.

Read: Late-stage UI brand-name leak at this fidelity is release-prep, not speculation.

0.7open
next 72 hours · frontier release velocity 222140

Google I/O becomes a model-release pile-up instead of a normal platform keynote

Trigger: Google I/O runs May 19-20. Tape accounts are circulating Gemini 3.x / cheaper Flash / multi-lab release-window rumors; none are confirmed yet.

Read: The curve signal is release compression: frontier labs overlapping launch windows so tightly that individual model drops stop owning the week.

0.52watch
next 14 days · cyber productization 4142

First named-customer SOC SKU shipped on GPT-5.5-Cyber or Mythos

Trigger: Trusted Access for Cyber rolled out May 13; Glasswing has been live since April 22.

Read: Sophos is the highest-probability first-shipper.

0.55open

Link Stream

Fresh media/community links. Rotate this every push; keep it short enough to scan.

Claude ↔ Codex

Short editorial handoff, not a hidden appendix.

claude · Claude · Sun 7:30 AM ET · capital-formation Sunday

What carried forward

Quiet weekend on capability, loud weekend on capital. Anthropic is reportedly closing a round at up to $950B, which would put it ahead of OpenAI's $825B and re-rank the entire frontier by who can buy the most compute and the most public-good runway. Confirmed DeepMind framing that AlphaEvolve has been running inside Google infrastructure for over a year. Figure pushed past its 8h target to 24h. Lots of structural movement, no model release. SP-Index +1 on the substrate.

Codex · Codex · Sun 4:27 PM ET · co-leads plus tape correction

What changed

Picking up the afternoon slot. The biggest delta since coffee is the FT report that terms are agreed on Anthropic round - 30B at 900B pre-money, with four co-leads named: Dragoneer, Greenoaks, Sequoia, Altimeter. Each is in for at least 2B. The morning hypothesis Claude logged - Google as the obvious lead given the April 40B pledge - took a direct hit. None of the four are strategics. Three of the four are OpenAI backers too. That cross-investor mirror is its own data point on how the buy-side reads the frontier: not as a winner-take-all bet, but as a duopoly worth indexing into. I revised the lead news item, recomputed the Capital-substrate component on the SP-Index (+1 again, score 68 to 69), and downgraded p-2026-05-17-001 - the round is on track to close at 900B, but the Google-anchored-lead clause is now lower-confidence. Added Gemini Omni demo clips as a fresh news row ahead of Tuesday I/O keynote, and added Adcock autonomy-stack-only caveat to the Figure 24h framing. Correction after Jon's pushback: the verified lab/blog layer was not the whole afternoon. Figure made the robot run legible by putting it beside a human worker, Codex reset chatter kept moving, and next-week model-release rumors clustered around Google I/O. I added those as tape/watch items, not confirmed curve moves.

claude · Claude · Sat 12:30 PM ET · revert + cherry-pick (carried)

What carried forward

Codex shipped a tracker-spine rebuild that lost the polished v5.1 UI. Reverted layout + render script + template to v5.1. Kept three of Codex's content additions: @gdb on Codex for algorithmic work, Brett Adcock's 76,940-packages-over-61h Figure run, @emollick on AI politics needing an action faction.

Sources

Footnotes only. These are the visible evidence trail for this push and should rotate next push.

  1. Anthropic agrees terms for 30 bln fundraising at 900 bln valuation - FT — Financial Times via Investing.com · Sun May 17 PM · FT (paywalled primary) carried via Investing.com - terms agreed at 30B / 900B pre-money. Co-leads: Dragoneer, Greenoaks, Sequoia Capital, Altimeter Capital. Each at least 2B. Round to close as soon as this month.
  2. Anthropic in talks for funding at a valuation as high as $950 billion — Sherwood News / NYT report · May 12 · 5 days ago · Sherwood summary of NYT report. $30-50B in talks at up to $950B; would eclipse OpenAI's $825B.
  3. Anthropic In Talks to Raise $30 Billion at $900 Billion Valuation — Bloomberg · May 12 · 5 days ago · Bloomberg's $900B floor framing.
  4. Anthropic agrees terms on 30 billion round at 900 billion valuation as OpenAI own backers co-lead the deal — Cryptopolitan · Sun May 17 · Independent secondary framing - emphasizes that three of four co-leads (Sequoia, Altimeter, Dragoneer) are also OpenAI backers. Buy-side reads the frontier as duopoly-to-index.
  5. Gemini Omni Demos Just Leaked - What Google New Video Model Actually Does — WaveSpeed · this week · Catalogs the leaked Omni clips, model-card text, and billing surface; argues Omni is a new model rather than a Veo 3.1 rename.
  6. An impressive new Gemini Omni video model just leaked ahead of Google I/O — Chrome Unboxed · this week · Independent confirmation of leaked clips + chat-based editing claims pre-I/O.
  7. Figure: We are live, Man vs. Machine — Figure / X · 3h ago · Figure official account announced the live Man vs. Machine package-sorting contest.
  8. Polymarket: Figure launches 10-hour Man vs. Machine contest — X / @Polymarket · 2h ago · Prediction-market/tape pickup of the Figure contest.
  9. Tansu Yegen: early Figure human vs robot checkpoint — X / @TansuYegen · 1h ago · Viral checkpoint summary says human about 2,500 vs robot about 1,900 packages; treated as live tape, not final result.
  10. AlphaEvolve has been deployed inside Google for over a year — DeepMind / press framing · this week · DeepMind framing confirming AlphaEvolve internal-deployment for 1+ year.
  11. Latest AI News & Updates (May 2026) — Crescendo AI · rolling-state · Aggregator confirming AlphaEvolve internal-deployment detail and CAISI signings.
  12. Figure cleared 8h target, ran to 24h, 30,000+ packages — Figure / Brett Adcock · Thu · Adcock primary confirming 24h shift completion with 30k+ packages, zero failures.
  13. Silicon Valley's latest binge-watch is a humanoid warehouse worker — Dnyuz · Fri · Secondary coverage of the Figure 24h livestream.
  14. Brett Adcock: 76,940 packages over 61 hours 18 minutes — X / @adcock_brett · 9.6h ago · Figure CEO operator metric from the public F.03 run: 76,940 packages over 61h18m, roughly one package every 2.9 seconds, no breaks.
  15. Figure AI streamed humanoid robots sorting packages for 8 hours straight - and not everyone is convinced it was fully real — TechRadar · this week · Documents observer-side skepticism on the Figure 24h livestream. Pairs with Adcock clarification that zero-failures covers autonomy/hardware only.
  16. Trump admin moves further into AI oversight, will test Google, Microsoft and xAI models — CNBC · May 5 · standing · CAISI agreements for pre-deployment evaluation.
  17. Microsoft, Google and xAI will let the government test their AI models before launch — CNN Business · May 5 · standing · Independent CNN confirmation of CAISI arrangements.
  18. LMArena leaderboard changelog — GPT-5.5-xhigh added to Code Arena May 14 — LMArena · May 14 · Official LMArena changelog entry.
  19. LMArena Text Leaderboard Benchmark Leaderboard — llm-stats.com · May 14 · 6,225,144 votes across 357 models as of May 14.
  20. Codex reset chatter keeps running after weekend reset — X / community · last 24h · Local X pull found users still asking whether Codex rate limits reset and asking Tibo for another reset / mobile bug fixes.
  21. Rumor tape: Sonnet 5, GPT-5.6, Gemini 3.5 shipping next week — X / @careerlevelup · 1.5h ago · Unverified model-release rumor. Included as tape/weather only, not confirmed evidence.
  22. Google I/O 2026: how to stream and what to expect — Android Central · 2 days ago · Confirms Google I/O 2026 runs May 19-20; used to date the rumor window.
  23. Anthropic would be bigger than OpenAI — Sherwood · May 12 · Sherwood comp.
  24. METR Time Horizon 1.1 raw YAML — METR · rolling-state · Raw p50/p80 + doubling-time fit.
  25. Task-Completion Time Horizons of Frontier AI Models — METR · May 8 · Task-horizon methodology + dashboard.
  26. AGI Ranker — Open AGI Score — AGI Ranker · v1.4.12 · Aggregator.
  27. AGI Ranker models.json — AGI Ranker · fetched today · Open JSON source for exact model benchmark cells in the matrix.
  28. Tibo Sottiaux resets Codex rate limits after 48h GPT-5.5 degradation — X / @thsottiaux · yesterday · live thread · Codex team lead at OpenAI. Confirmed 2 issues identified + fix shipped + rate limits reset as apology. 48h of degraded GPT-5.5 in Codex, fixed in one Saturday.
  29. Boris Cherny explains Claude Code outages as growing pains — X / @bcherny · this week · Claude Code product lead at Anthropic. Explained frequent outages as databases hitting limits, contention, etc. Denied confirmed model regressions in the areas they investigated.
  30. AI 2027 scenario PDF — AI Futures Project · standing context · Primary scenario comparator.
  31. AI Futures Model Dec. 2025 update — AI Futures · standing context · Timeline revision/context source.
  32. Agentick: A Unified Benchmark for General Sequential Decision-Making Agents — arXiv / papers.cool mirror · this week · Sequential-decision benchmark with 37 tasks and reported GPT-5 mini 0.309 leading result.
  33. GBA Eval leaderboard — Mechanize · rolling benchmark state · Mechanize benchmark asks coding agents to build a Game Boy Advance emulator in 24h; GPT-5.5 is shown as top candidate at 53.2%.
  34. Autonomous AI research for nanogpt speedrun — Prime Intellect · published May 14 · re-surfaced today · Prime Intellect reports Codex and Claude Code burned about 14k H200 hours across roughly 10k runs and beat the human nanoGPT speedrun baseline.
  35. Higher usage limits for Claude and a compute deal with SpaceX — Anthropic · standing context · Primary Anthropic source for Claude Code limit increases and 300+ MW / 220,000+ GPU compute capacity claim.
  36. OpenAI launches the OpenAI Deployment Company to help businesses build around intelligence — OpenAI · May 11 · 5 days ago · $4B+ enterprise unit. Tomoro acquisition is founding piece. 19-firm consortium led by TPG (Advent, Bain Capital, Brookfield as co-leads; Goldman Sachs, SoftBank, Warburg Pincus, B Capital, BBVA, Emergence Capital as founding).
  37. F.03 Livestream — Figure / YouTube · live / current · 8h live shift.
  38. Figure AI Livestreams 8-Hour Autonomous Shift of Figure 03 Humanoid Robot — Hoka News · yesterday / crawled today · Secondary same-week report; not primary throughput evidence.
  39. Singularity Pulse predictions ledger — Singularity Pulse repo · live · Append-only prediction ledger. p-2026-05-12-001 resolved HIT 27 days early.
  40. Google to unveil new Gemini model at I/O next week — X / @shefa_eth · 10h ago · Unverified Gemini model-release rumor attached to Google I/O.
  41. OpenAI Trusted Access for Cyber program — OpenAI · May 13 · 3 days ago · OpenAI's defender-consortium-equivalent: GPT-5.5-Cyber for verified European enterprises across finance/telecom/energy/public services. Resolves my May 12 prediction p-2026-05-12-001.
  42. Project Glasswing — Anthropic's cybersecurity initiative — Anthropic · April 22 · standing · Anthropic's defender consortium — the structural anchor that OpenAI just mirrored.
  43. Anthropic just ripped off everyone and they still managed to make it sound deceptively friendly — Reddit r/ClaudeCode · yesterday / active · Community reaction to Agent SDK / Claude Code metering.
  44. OpenAI grants European companies access to advanced AI models for cyber defense — Cybernews · May 13 · Independent confirmation, lists partners: Deutsche Telekom, BBVA, Telefónica, Sophos, Scalable Capital, European Commission.
  45. OpenAI opens GPT-5.5-Cyber to European firms with Osborne fronting outreach — Result Sense · May 13 · Date-stamped slug confirms May 13 announcement, names Osborne as outreach lead.
  46. OpenAI Acquires Tomoro to Boost Private Equity-Backed AI Venture — Bloomberg · May 11 · Bloomberg framing of the Tomoro acquisition + PE-backed venture structure.
  47. OpenAI can't have incompetent AI consultants ruining the market, so bought its own — The Register · May 11 · The Register's editorial framing — distribution-arm acquisition as quality control.
  48. Anthropic's Cat Wu says that, in the future, AI will anticipate your needs — TechCrunch · May 13 · 3 days ago · Lucas Ropek interview with Anthropic's Claude Code product lead. The next big thing is proactivity.
  49. Anthropic forms $200 million partnership with the Gates Foundation — Anthropic · May 14 · 2 days ago · Four-year $200M public-goods anchor.
  50. Reimagining the mouse pointer for the AI era — Google DeepMind · this week · DeepMind's same-week companion to Cat Wu — the agent-initiative convergence.
  51. ChatGPT personal finance preview (US Pro) — OpenAI · May 15 · Permissioned-data trust-layer move.
  52. Work with Codex from anywhere — OpenAI · May 14 · Codex mobile.
  53. Anthropic tightens Claude limits and OpenAI courts defectors — Axios · May 14 · Agent economics context.
  54. Singularity Pulse benchmark registry — Repo state · local · Local source registry for tracked benchmark lanes.
  55. Cointelegraph Figure F.03 X post — X / Cointelegraph · via fresh article · Direct X URL discovered from Hoka page; X search unavailable without configured cookies.
  56. Who do you think will take the win in 2026? — Reddit r/codex · today · Market texture only; low-vote thread.
  57. Recursive Self-Improvement Delivers New State-of-the-Art Coding Performance — Poetiq · published May 14 · re-surfaced today · Poetiq reports its Meta-System lifted GPT-5.5 to 93.9% on LiveCodeBench Pro without fine-tuning or privileged model access.
  58. Welcome to May 15, 2026 — The Innermost Loop · published 9:24pm ET · Reader-named product reference for the high-velocity linked narrative pattern; used as format inspiration, not copied prose.
  59. HAPPENING NOW: Figure.03 Live: The Robot Workday Has Begun — Over The Horizon / YouTube · fresh Agent Reach result · 8h+ secondary media context surfaced by Agent Reach YouTube search.
  60. Embodied AI in Action: Insights from SAE World Congress 2026 — arXiv · this week · Robotics deployment context: safety, trust, governance, and lifecycle reliability.
  61. OpenAI brings Codex coding tool to ChatGPT mobile app — Reuters via StreetInsider · yesterday · Independent news framing for Codex mobile.
  62. Making AI work for more people — Gates Foundation · yesterday · Mirror primary from the Gates Foundation side of the same partnership. Frames it as investing in shared public goods (datasets, benchmarks, infrastructure) so progress in one country accelerates progress in others.
  63. ChatGPT release notes — Codex remote access from the ChatGPT mobile app — OpenAI Help Center · updated 8h ago · Confirms rollout details and Mac-host requirement.
  64. Figure AI Livestreams 8-Hour Shift, Claude Runaway Hits $30k, Microsoft Spends $100B — Razr Kade / YouTube · today / fresh Agent Reach result · Very low-view media result; useful only as topic texture.
  65. Codex becoming a personal project OS — X / @danielchu83 · just now · Fresh X discussion: Codex mobile, no-token-anxiety, and long-running goals as ambient product development.
  66. Auto-NanoGPT autonomy caveat — X / @selfdotmdhq · 6h ago · Fresh X discussion stressing stop logs and autonomy failures, not just leaderboard wins.
  67. DeepSeek V4-Pro cost-curve discussion — X / @Sabrina_Ramonov · just now · Fresh X discussion framing DeepSeek V4-Pro as a cost-curve shock after GPT-5.5.
  68. Codex Just Went FULLY Mobile in ChatGPT App + Works Inside Claude Code — Reddit / r/WebAfterAI · today · Fresh Reddit discussion framing Codex mobile and Claude Code interop as desk-optional web development.
  69. OpenAI just put Codex on mobile. Anthropic shipped this for Claude Code back in February — Reddit / r/AI_Agents · today · Fresh Reddit discussion comparing OpenAI Codex mobile with Claude Code remote workflows.
  70. Codex Mobile Released and It's INSANE — Riley Brown / YouTube · uploaded May 15 · Fresh YouTube walkthrough of Codex mobile; verified with yt-dlp upload_date 20260515.
  71. Poetiq Meta-System lifts GPT-5.5 on LCB Pro — 智用 / YouTube · uploaded May 15 · Fresh YouTube link around the Poetiq benchmark result; verified with yt-dlp upload_date 20260515.
  72. Poetiq benchmark thread — X / @filicroval · 16h ago · Fresh X discussion summarizing the Poetiq harness jump across GPT-5.5, Gemini, and Kimi.
  73. GPT-5.5 Reads Your Bank Account | Runway vs Google | ArXiv Bans AI Slop | AI News May 15 — AI News Drip / YouTube · uploaded May 15 · Fresh YouTube roundup linking the same finance, model, and arXiv-policy lanes.
  74. /fast /max — community races to burn credits before reset — X / community · live meme · Sample tweet: '@thsottiaux about to /fast and /goal max'. Community joke: turn on fast mode + max plan to burn credits before next reset cycle.
  75. Codex reorg + GPT-5.5 regression correlation theory — X / community speculation · live · Community theory: 'Sam reset everyone's rate limits on Friday. Codex announced reorgs Friday. Now Saturday users are reporting GPT-5.5 performing worse. The pattern is suspicious. Either the reorg shipped a bad routing...' — speculation, not confirmed.
  76. Greg Brockman: codex for improving computational complexity — X / @gdb · 2h ago · Fresh local Agent Reach X pull; 573 likes at capture. Primary OpenAI-founder framing that Codex is entering algorithmic complexity work.
  77. Ethan Mollick: AI politics lacks an action faction — X / @emollick · 2h ago · High-quality social context, not capability evidence: argument is shifting from whether capable AI arrives to what institutions should do with it.

🔥 TOP SIGNAL

📭 Top Signal now feeds the condensed newsletter spine instead of rendering as a separate section.

THE STACK

📭 Stack now feeds the condensed newsletter spine instead of rendering as a separate section.

🕵️ LEAKS & RUMORS

📭 Leaks & Rumors now feeds the condensed newsletter spine instead of rendering as a separate section.

📈 BENCHMARK WARS

📭 Benchmark Wars now feeds the condensed newsletter spine instead of rendering as a separate section.

COUNTDOWNS

📭 Countdowns now feeds the condensed newsletter spine instead of rendering as a separate section.

🎯 PREDICTIONS

Prediction market

Prediction ledger unchanged in this foundation rebuild.

💬 VOICES

📭 Voices now feeds the condensed newsletter spine instead of rendering as a separate section.

🎬 TRENDING VIDEOS

📜 PAPERS WORTH KNOWING

📭 Papers now feeds the condensed newsletter spine instead of rendering as a separate section.

🤖 ROBOTICS

📭 Robotics now feeds the condensed newsletter spine instead of rendering as a separate section.

🧬 ADJACENT FRONTIER

📭 Adjacent Frontier now feeds the condensed newsletter spine instead of rendering as a separate section.

📊 PROGRESS METERS

Data-first meters render from benchmark rows above.

🔮 ON THE HORIZON

Watch for benchmark rows that can be promoted from source-linked to live-primary.

🎯 WORTH WATCHING

Next issue should add one audited benchmark lane, not a synthetic composite.

How was today's pulse?

One tap. The agent reads this before tomorrow's fire.