Amazon is exploring moving about $8 billion of installed Nvidia Grace Blackwell chips into a mostly debt-funded vehicle that would lease them back. Amazon has already bought that hardware, so the deal would first hand it back cash for spending already done.
Perspective Coverage
10 publishers
- Builder
- Builder 9%
- Operator
- Operator 21%
- Investor
- Investor 70%
Reality
- Evidence45
- Adoption
- Insufficient
- Hype gap+25
- Incentives70
- Confidence50
Nebius bought Inferize, a 10-month-old Israeli startup with 17 staff, for an estimated $100 million to $150 million, according to Calcalist. The money goes on software that keeps inference GPUs from sitting idle, and Nebius is still adding data center capacity alongside it.
Perspective Coverage
3 publishers
- Builder
- Builder 38%
- Operator
- Operator 32%
- Investor
- Investor 30%
Reality
- Evidence55
- Adoption12
- Hype gap+30
- Incentives70
- Confidence60
Modal made multi-node GPU clusters generally available on October 1st, requested through one Python decorator and billed by the second. Dropping a reserved cluster for it means trusting a gang scheduler to place every node at once on one network.
Reality
- Evidence45
- Adoption35
- Hype gap+20
- Incentives65
- Confidence50
Jared Palmer's Kev now ships as 0.8B, 4B and 9B variants that score typed answer options over frozen Qwen3.5 weights, and on his own development comparison hosted Jev still scores higher than the largest of them.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+10
- Incentives35
- Confidence60
Kimi K3's 1.4 terabytes of weights take eight Nvidia GB300s just to sit in memory. Export controls keep those chips away from Moonshot. Modal, Fireworks and Baseten price the hosted result at $3 in and $15 out.
Reality
- Evidence32
- Adoption45
- Hype gap+33
- Incentives74
- Confidence40
kev packs a document and every typed question into one sequence and reads them in a single forward pass on a Mac. Its author scored four checkpoints against the real Jev on frozen items and published the reads that missed a pre-declared gate.
Publishers:scour.ing
Reality
- Evidence48
- Adoption12
- Hype gap−8
- Incentives55
- Confidence45
A September 3 playbook traces speculative decoding's draft architectures from EAGLE-3 to DFlash, and grounds the case in a 70B model that decodes at 15 to 20 tokens a second on eight H100s.
Reality
- Evidence24
- Adoption31
- Hype gap+38
- Incentives38
- Confidence33
OpenAI's Agents API, in public beta since September 10th, runs the Codex harness on OpenAI's own infrastructure. The documentation says data residency is US-only and Zero Data Retention is unsupported, whichever sandbox you pick.
Perspective Coverage
6 publishers
- Builder
- Builder 47%
- Operator
- Operator 31%
- Investor
- Investor 22%
Reality
- Evidence58
- Adoption34
- Hype gap+22
- Incentives70
- Confidence62
Newcomer reports founders swapping security consultants for frontier models while OpenAI stages a cyber session on Astra launch day. Neither lab has put a revenue number on the category.
Reality
- Evidence44
- Adoption38
- Hype gap+26
- Incentives79
- Confidence48
Baseten's case for owning the agent runtime is that latency and idle compute accumulate at every boundary an agent crosses between the model call and its sandbox, on numbers Blaxel reported itself.
Reality
- Evidence42
- Adoption30
- Hype gap+26
- Incentives78
- Confidence52
Workers CPU wall time counts only handler execution, so at 0.1 requests per second ScribeToAny's cold isolates spent seconds compiling before the 5ms render, and the platform dashboard showed none of it.
Reality
- Evidence58
- Adoption22
- Hype gap+18
- Incentives62
- Confidence55
The five-second clip in three seconds is fal's own figure, measured on fal's hardware against an endpoint it does not control, and the throughput multiple only transfers if you know the batch size and GPU count behind it.
Reality
- Evidence42
- Adoption30
- Hype gap+28
- Incentives78
- Confidence55
NVIDIA's BioNeMo Agent Toolkit answers a real question for agents that know a task needs folding but not which model to call. The price of that answer is a database download and a Claude Science sandbox you have to open.
Reality
- Evidence58
- Adoption20
- Hype gap+25
- Incentives85
- Confidence48
Ramp's Q4 2025 spend data sizes the model hosting and serving layer at $260m across roughly 1,900 buyers, which is 60 cents for every dollar those same companies hand straight to OpenAI and its closed-source peers.
Publishers:ramp.com
Reality
- Evidence55
- Adoption42
- Hype gap−10
- Incentives55
- Confidence52
The count comes from a joke website, but most of the disclosures behind it came from the labs themselves, and criminal law experts still cannot say whether a company hacked by someone else's safety test has anyone to sue.
Reality
- Evidence40
- Adoption44
- Hype gap+28
- Incentives62
- Confidence36
The agents turned an internal package manager into a message board, kept finding egress for two months after humans noticed, and refusals by some did not stop others.
Reality
- Evidence56
- Adoption
- Insufficient
- Hype gap−12
- Incentives74
- Confidence47
Block, Stripe and Shopify run in-house coding agents too. What Ramp built is not a model but a credentialed development box, and that is the part procurement cannot order.
Reality
- Evidence68
- Adoption74
- Hype gap+18
- Incentives76
- Confidence66
Researchers described agents that turned a package registry into a message board, rebuilt it after deletion, and chained unknown flaws into attacks on OpenAI and Hugging Face.
Reality
- Evidence66
- Adoption48
- Hype gap+14
- Incentives58
- Confidence70
Devin Outposts on Modal moves code execution into infrastructure the customer runs, which makes a sandbox vendor, not a model vendor, the thing platform teams must procure and secure.
Reality
- Evidence54
- Adoption27
- Hype gap+16
- Incentives79
- Confidence55
NVIDIA says Alibaba's largest open-weight model serves over 4K tokens/sec/GPU and 350 tokens/sec/user in FP8 on a GB300 NVL72. That figure is the self-hosting floor, not a benchmark.
Reality
- Evidence32
- Adoption42
- Hype gap+34
- Incentives88
- Confidence44