Invest1 distinct publisher3 min readPublished
fal's H3 Max clears real-time playback by two thirds at the fast end of its range, which turns continuous video into a metered compute bill. The 12.5-cent promotional clip price is what the arithmetic rests on.
The Investor · Invest desk

Compiled by The InvestorSomething wrong?How this is made
Take the promotional sheet at face value and the arithmetic is unusually clean: 12.5 cents buys five seconds at 480p [10], which is two and a half cents per second of finished footage [1], a dollar fifty a minute [2], ninety dollars an hour [3], and $2,160 for a channel that runs a full day without stopping [3]. Anything above 480p doubles [10], so the same day costs $4,320 [4]. Hold the lower tier for a year and you have spent $788,400 [5] on footage nobody commissioned.
The latency number does something different to the cost number. Fifteen seconds arriving every nine [1] means the renderer is occupied 60% of wall-clock time [6], so a single always-on channel leaves four tenths of a pipeline idle, the same capacity could in principle carry about 1.67 concurrent streams [6], and a full day of programming takes 14.4 hours to manufacture [8].
Which is where the source material argues with itself. The same write-up that reports 15 seconds in 9 also gives the window as 9 to 16 seconds depending on resolution and settings [7], and at 16 the model returns 0.94 seconds of footage per second of compute [7], so the day needs 25.6 hours [7] and the channel falls behind its own schedule. Faster than playback is a configuration, or rather the fast end of a range with a demo attached. The short-clip figure is the sturdier evidence: under three seconds for five seconds at 720p or 768p, roughly 2.5 in some tests [6], which is two seconds of output per second of compute [9].
Worth naming what fal did not spend on. H3 Max is a closed, post-trained iteration of MiniMax H3 [4], an open-weight multimodal model that appeared in July [5], so the budget went into inference optimisation over somebody else's weights rather than a pretraining run, and Crypto Briefing's reading of the 35x throughput claim [8] is that architecture rather than additional hardware produced it [13].
Three ways this goes. The promotional rate holds and continuous video becomes a metered utility that anyone with a card can buy by the hour; or the usable resolution sits at the doubled tier and nearer the 16-second end, in which case the crossing evaporates outside 480p; or demand caps out because acceptance does, since the render cost fell to cents while a platform's decision to carry the stream stayed binary and both Twitch and Kick said no [11].
This is probably wrong, but the always-on channel looks like the least valuable of the three uses on offer [14], and the one with money in it is the advertising pipeline with no rendering queue, because latency under playback means bespoke creative gets made inside a session rather than booked against a farm. I would drop that view if the first-place human-preference result [9] turns out to hold only at the resolution tier that costs twice as much, which would put the interesting quality and the interesting speed on opposite sides of the price list.
Ranked by verification strength, evidence, and original report placement.
fal engineer Rehan Sheikh set up a broadcast using H3 Max that produces roughly 15 seconds of AI video every 9 seconds, generating footage quicker than real-time playback.
Sheikh's stream was a proof of concept for an ongoing automated video stream generated entirely by AI in real time, with no pre-rendered assets and no human operator choosing clips.
H3 Max launched on August 27, 2026 as a closed, post-trained iteration of MiniMax H3, described as fal's optimized version of the MiniMax H3 video generation model.
MiniMax H3 is an open-weight multimodal model that debuted in July 2026.
H3 Max renders a 5-second clip at 768p or 720p resolution in under 3 seconds, with some tests clocking it at around 2.5 seconds.
Distinct publishers with included, body-backed reporting in this cluster.
cryptobriefing.com
1 article · August 29, 2026
Follow any of these and your For You feed starts watching them — no settings page required.
product
Ox Alpha was GLM-5.3-Flash, and the number that decides displacement is 18 billion1 distinct publisher
build
Wan 3.0's billing hinges on a fingerprint, because a careless retry buys a second render1 distinct publisher
invest
The cheap-token trade is closing: DeepSeek's 12x price rise resets everyone's AI cost model1 distinct publisher
invest
Higgsfield's $400M round prices the collapse of production cost, and agencies are the line item3 distinct publishers
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
One outlet, no primary documents
Every figure that matters here — 15 seconds in 9, 2.5 seconds for a 720p clip, 12.5 cents, 35x — reaches us through a single Crypto Briefing write-up. There is no fal release to read, the 'independent benchmarks' are never identified, and the human-preference ranking arrives with no leaderboard attached. The one thing that would be trivially verifiable, a stream anyone could watch, is described rather than linked.
One engineer's stream, two days after launch
What demonstrably exists is a model that shipped on August 27 and one employee's proof of concept that shipped on the 29th, now running on Rumble. No customer is named, no volume or spend is disclosed, and nobody outside fal has reproduced the timings. A promotional price is a bid for usage, not a record of it.
Framing runs past the stated range
The piece says the model 'outpaces reality itself', then concedes a few paragraphs later that the 15-second window can take 16 seconds — at which point it is slower than playback, which was the entire premise. The always-on channel, the personalized feeds and the queue-free ad pipeline are all pitched off the fast end of a range whose slow end would need 25.6 hours to fill a day. Stack an anonymous 35x on top and the story is selling more than it has shown.
Vendor demo at a vendor promo price
The stunt was built by a fal engineer on fal's own model, and the number that makes a 24-hour channel look affordable is one fal has flagged as temporary — both are marketing instruments before they are measurements. Crypto Briefing passes along the throughput and ranking claims without a counterparty, and its line about Rumble becoming a home for experimental AI reads more like positioning than observation.
Checkable but unchecked
Credit where due: the dates, prices and latencies are specific enough that anyone could falsify them in an afternoon. Nobody has. With one publisher, no vendor documentation, and a rate that can change without notice, the arithmetic in this story is sound and its inputs are provisional.