/ Blogs & Newsletters

Powering an Unlimited AI Economy That Works

Private, unrestricted inference.
Without corporate cloud costs.

Scroll
Sogni Unlimited · $20/moFair-use unlimited image, video, music & chat on open models.
Upwards of 51% of network revenueEarned by the GPU operators who render the work.
Sogni Supernet158 million renders in its first public year.
Essay · Unlimited economics

Powering an Unlimited AI Economy That Works

Every few months, an AI company launches an unlimited generation plan. Creators celebrate. A few months later, the plan quietly disappears — closed to new subscribers, converted to credits, "evolved" into something metered. This has now happened enough times that it's a pattern:

December 2024

Pika discontinues its Unlimited plan

moving users to a credit-based tier. The stated reason: "unsustainable due to misuse." source

June 2025

Hailuo (MiniMax) closes its $94.99 Unlimited plan to new buyers

"unlimited" now covers only the older Hailuo-01 model, with flagship models gated behind credits on $124.99–$199.99 tiers. source

March 2026

OpenAI announces the shutdown of its Sora app

after Forbes reported the free video feed was costing up to an estimated $15 million per day in compute. source

April 2026

Freepik

which launched unlimited AI generation to real acclaim in July 2025, reintroduces credits for premium models — roughly ten models stay free as it rebrands to Magnific. Unlimited lasted nine months. source

June 1, 2026

Runway closes its $95/month Unlimited plan to new subscribers

existing customers migrate to metered credits ("Max") on September 1. Two details worth knowing: the Unlimited plan reportedly carried the same 2,250 monthly credits as the $35 Pro plan for fast generations, and the migration removes Explore Mode — the throttled free-generation queue that made "unlimited" true at all. source

And it's not just creative tools. On June 1, 2026, GitHub Copilot moved every plan to usage-based billing (source). OpenAI's Nick Turley has publicly compared unlimited AI plans to selling "unlimited electricity." The industry consensus is congealing: the AI subsidy era is ending, flat-rate is dead, everything becomes a meter.

Against those odds, Sogni is doubling down by announcing unlimited usage plans, starting at $20/mo.

Making the math work

When a centralized AI company sells you unlimited generation, it is arbitraging a fixed, brutal cost: enterprise GPUs rented from clouds, or bought outright. An NVIDIA H100 runs roughly $30,000; the cloud rental meter never stops. Every marginal render is real marginal cost against that bill, so a flat-rate plan is a bet that users won't use it — and creative tools are iteration machines. The heaviest users are the most passionate customers. The plan works right up until the product succeeds, and then the GPU bill eats it. That's the story of every retreat above.

Sogni's cost structure is inverted. We render on the Sogni Supernet — a network of independent GPU operators, most running consumer cards like the RTX 4090 and 5090. We don't rent their hardware at a fixed rate. They stake their idle compute on the decentralized network and earn upwards of 51% of network revenue. There is no fixed cloud bill to underrun. When network activity grows, so does the earnings pool, which recruits more supply — the same demand-attracts-supply dynamic that lets a rideshare network absorb demand spikes without owning cars.

Consumer GPUs are the right tool for this specific job

The obvious objection: consumer cards can't compete with data-center hardware. For some workloads that's true. For this one, the economics run the other way:

The open-model wave makes the timing

The second thing the retreats have in common: they were built on closed, per-render-licensed or self-trained frontier models — the most expensive possible substrate for an unlimited promise. Meanwhile, open-weight models crossed a quality threshold that most of the market hasn't priced in. Lightricks' LTX-2.3 — open weights, synchronized audio+video, up to 4K — was downloaded more than 2.1 million times on Hugging Face last month (source). Alibaba's Z-Image Turbo: nearly one million (source). Krea 2 went open-weights on June 22 and was live on Sogni within days (source).

A cleaner example is image generation on the same Sogni pricing reference. One 1024×1024 GPT Image 2 render at medium quality with no references is estimated at $0.0685 (13.70 Spark), while a 1024px Krea 2 Turbo render at 8 steps is estimated at $0.0122 (2.43 Spark) — making GPT Image 2 about 5.6× the cost for these example jobs (source). These are not interchangeable models, and output quality is subjective; the point is not that every job should use Krea 2 Turbo. It is that open models can make high-volume iteration economically viable, while external frontier models remain available when their specific capabilities justify the premium.

Example job · one 1024×1024 image · Sogni posted estimates
GPT Image 2 · external frontier
$0.0685
vs
Krea 2 Turbo · open weights
$0.0122
GPT Image 2 is about 5.6× the price at the posted assumptions. Same output dimensions; different model capabilities.

That's the design of Sogni Unlimited: $20/month, fair-use unlimited across the open models we host — image, video, music, and LLM chat — with the three frontier vendor models (GPT Image 2, Seedance 2.0, HappyHorse 1.1) as metered add-ons at a member discount.

"Unlimited with fair use" — what we mean, precisely

Skepticism is warranted; the word has been abused. So here's the mechanics, published rather than buried: covered renders spend no credits — there is no monthly meter, nothing you buy and burn down. Fair use governs scheduling: each tier has published concurrency limits, exceptionally heavy sustained use queues behind other jobs at peak, and heavy-usage throttles reset each UTC day. Queued work gets done. The difference between this and "unlimited" plans that quietly suspend their heaviest users is that we're telling you the mechanics up front — they're the reason the promise is sustainable.

Can demand outrun the network? Three answers, in order: today there is substantial headroom; the worker earnings pool recruits supply as demand grows (and we can provision additional support nodes through rental partners like Salad profitably, because pool economics cover it); and the priority ladder protects everyone — paid pay-as-you-go jobs first, then Unlimited by tier, then free-tier jobs.

The priority ladder. Paid pay-as-you-go jobs first, then Unlimited by tier, then free-tier jobs. Queued work gets done.

The bet

Every company that retreated from unlimited was making the same wager against its own users: please don't use this too much. We're making the opposite wager: that if compute costs scale with revenue instead of against it, heavy use is the business model, not the threat to it. One year and 158 million renders into the Supernet's public life, that's a bet we're structurally built to make — and the people who own the GPUs earn upwards of 51% of network revenue to help us win it.

Iteration is where creative work actually happens. It shouldn't have a taxi meter running.

Sogni Unlimited

$20/month. No meter on the making.

Fair-use unlimited across the open models we host — image, video, music, and chat — or $199/year, with a 3-day trial for eligible new subscribers.

Sogni Unlimited is $20/month or $199/year, with a 3-day trial for eligible new subscribers at sogni.ai/unlimited. If you have a GPU, it can join the network.

Sources & archives
  1. Pika discontinues its Unlimited plan (Dec 2024) — original · archived
  2. Hailuo (MiniMax) payment policy — $94.99 Unlimited closed to new buyers (Jun 2025) — original · archived
  3. Freepik AI pricing — credits reintroduced, Magnific rebrand (Apr 2026) — original · archived
  4. OpenAI — what to know about the Sora discontinuation (Mar 2026) — original · archived
  5. Runway — Unlimited plan is switching to Max (Jun 2026) — original · archived corroboration
  6. GitHub Copilot is moving to usage-based billing (Jun 2026) — original · archived
  7. ACM — Idle Consumer GPUs as a Complement to Enterprise Hardware for LLM Inference (AIBC 2025) — original · archived corroboration
  8. Lightricks LTX-2.3 on Hugging Face — original · archived
  9. Alibaba Z-Image Turbo on Hugging Face — original · archived
  10. VentureBeat — Krea 2 Raw and Turbo available as open weights (Jun 2026) — original · archived
  11. Sogni pricing reference — GPT Image 2 and Krea 2 Turbo example jobs — current pricing table