Every few months, an AI company launches an unlimited generation plan. Creators celebrate. A few months later, the plan quietly disappears — closed to new subscribers, converted to credits, "evolved" into something metered. This has now happened enough times that it's a pattern:
moving users to a credit-based tier. The stated reason: "unsustainable due to misuse." source
"unlimited" now covers only the older Hailuo-01 model, with flagship models gated behind credits on $124.99–$199.99 tiers. source
after Forbes reported the free video feed was costing up to an estimated $15 million per day in compute. source
which launched unlimited AI generation to real acclaim in July 2025, reintroduces credits for premium models — roughly ten models stay free as it rebrands to Magnific. Unlimited lasted nine months. source
existing customers migrate to metered credits ("Max") on September 1. Two details worth knowing: the Unlimited plan reportedly carried the same 2,250 monthly credits as the $35 Pro plan for fast generations, and the migration removes Explore Mode — the throttled free-generation queue that made "unlimited" true at all. source
And it's not just creative tools. On June 1, 2026, GitHub Copilot moved every plan to usage-based billing (source). OpenAI's Nick Turley has publicly compared unlimited AI plans to selling "unlimited electricity." The industry consensus is congealing: the AI subsidy era is ending, flat-rate is dead, everything becomes a meter.
Against those odds, Sogni is doubling down by announcing unlimited usage plans, starting at $20/mo.
When a centralized AI company sells you unlimited generation, it is arbitraging a fixed, brutal cost: enterprise GPUs rented from clouds, or bought outright. An NVIDIA H100 runs roughly $30,000; the cloud rental meter never stops. Every marginal render is real marginal cost against that bill, so a flat-rate plan is a bet that users won't use it — and creative tools are iteration machines. The heaviest users are the most passionate customers. The plan works right up until the product succeeds, and then the GPU bill eats it. That's the story of every retreat above.
Sogni's cost structure is inverted. We render on the Sogni Supernet — a network of independent GPU operators, most running consumer cards like the RTX 4090 and 5090. We don't rent their hardware at a fixed rate. They stake their idle compute on the decentralized network and earn upwards of 51% of network revenue. There is no fixed cloud bill to underrun. When network activity grows, so does the earnings pool, which recruits more supply — the same demand-attracts-supply dynamic that lets a rideshare network absorb demand spikes without owning cars.
The obvious objection: consumer cards can't compete with data-center hardware. For some workloads that's true. For this one, the economics run the other way:
Turbo isn't the whole catalog: distillation and numerical precision are separate choices. Sogni offers distilled and turbo models alongside full-model variants in FP8, FP16, and BF16. Our higher-precision FP16/BF16 options include Z-Image, Z-Image Turbo, Dark Beast Z-Image Turbo v9, and One Obsession.
See our recent Krea 2 Turbo comparison rendered on Sogni, or try it yourself and be the judge.
The second thing the retreats have in common: they were built on closed, per-render-licensed or self-trained frontier models — the most expensive possible substrate for an unlimited promise. Meanwhile, open-weight models crossed a quality threshold that most of the market hasn't priced in. Lightricks' LTX-2.3 — open weights, synchronized audio+video, up to 4K — was downloaded more than 2.1 million times on Hugging Face last month (source). Alibaba's Z-Image Turbo: nearly one million (source). Krea 2 went open-weights on June 22 and was live on Sogni within days (source).
A cleaner example is image generation on the same Sogni pricing reference. One 1024×1024 GPT Image 2 render at medium quality with no references is estimated at $0.0685 (13.70 Spark), while a 1024px Krea 2 Turbo render at 8 steps is estimated at $0.0122 (2.43 Spark) — making GPT Image 2 about 5.6× the cost for these example jobs (source). These are not interchangeable models, and output quality is subjective; the point is not that every job should use Krea 2 Turbo. It is that open models can make high-volume iteration economically viable, while external frontier models remain available when their specific capabilities justify the premium.
That's the design of Sogni Unlimited: $20/month, fair-use unlimited across the open models we host — image, video, music, and LLM chat — with the three frontier vendor models (GPT Image 2, Seedance 2.0, HappyHorse 1.1) as metered add-ons at a member discount.
Skepticism is warranted; the word has been abused. So here's the mechanics, published rather than buried: covered renders spend no credits — there is no monthly meter, nothing you buy and burn down. Fair use governs scheduling: each tier has published concurrency limits, exceptionally heavy sustained use queues behind other jobs at peak, and heavy-usage throttles reset each UTC day. Queued work gets done. The difference between this and "unlimited" plans that quietly suspend their heaviest users is that we're telling you the mechanics up front — they're the reason the promise is sustainable.
Can demand outrun the network? Three answers, in order: today there is substantial headroom; the worker earnings pool recruits supply as demand grows (and we can provision additional support nodes through rental partners like Salad profitably, because pool economics cover it); and the priority ladder protects everyone — paid pay-as-you-go jobs first, then Unlimited by tier, then free-tier jobs.
Every company that retreated from unlimited was making the same wager against its own users: please don't use this too much. We're making the opposite wager: that if compute costs scale with revenue instead of against it, heavy use is the business model, not the threat to it. One year and 158 million renders into the Supernet's public life, that's a bet we're structurally built to make — and the people who own the GPUs earn upwards of 51% of network revenue to help us win it.
Iteration is where creative work actually happens. It shouldn't have a taxi meter running.
Fair-use unlimited across the open models we host — image, video, music, and chat — or $199/year, with a 3-day trial for eligible new subscribers.
Sogni Unlimited is $20/month or $199/year, with a 3-day trial for eligible new subscribers at sogni.ai/unlimited. If you have a GPU, it can join the network.