AI Models Are Going the Uber Route: Peak Pricing Is Here, and You're Paying Double During Work Hours
📖 Glossary
AI Box (also known as Agent Computer / Agent PC), is a dedicated local hardware device that runs AI Agents. Pre-installed with an AI agent management system, plug-and-play, running 24/7. Users can remotely command AI to work via Discord, Slack, Telegram, WhatsApp, and more.
Abstract: DeepSeek V4 just announced surge pricing: double the cost during weekday work hours, discounts only late at night. Alibaba Meoo followed with a Night Plan — 80% off flagship models between 10 PM and 8 AM, but only half off during the day without covering peak hours. The AI industry is copying Uber's playbook: the hours you need AI the most are exactly the hours they charge you the most. This article breaks down the economics behind the latest AI pricing wave and explains the only way to sidestep this system entirely: bring the models home.
You tolerated Uber's surge pricing because you had no choice. But AI models pulling the same move?
On June 29, DeepSeek announced that V4's official release in mid-July would come with time-of-day pricing. Weekdays from 9 AM to noon and 2 PM to 6 PM — the price doubles. The Pro model's output jumps from 6 yuan to 12 yuan per million tokens. Three days later on July 2, Alibaba Meoo launched a Night Plan — flagship models at 80% off from 10 PM to 8 AM Beijing time. During the day? Half off, and peak hours from 10 AM to 4 PM aren't even included.
The AI world erupted.
Not because of the price increase itself — everyone knew paid models were coming. It's the structure. The hours you actually need AI are the most expensive hours. The hours you don't need it, you get a discount.
This playbook is painfully familiar. Uber's dynamic pricing. Food delivery peak surcharges. Hotel seasonal rates. Now AI models have joined the club.
Why Are AI Models Adopting Surge Pricing?
One word: congestion.
Large language model inference isn't "type a prompt, get a response." Every conversation consumes GPU memory, runs continuous computation, and maintains context. When tens of millions of users hit the servers simultaneously during peak hours, there's only so much compute to go around. DeepSeek's published numbers: 120 trillion tokens served daily, over 200 million active users. With that many people piling on during weekday work hours, GPU clusters are running at near capacity.
From the platform's perspective, surge pricing is a "regulator." Higher prices at peak discourage some users. Lower prices off-peak monetize idle compute. Sounds logical.
From the user's perspective, it's a different story. Your work hours are 9-to-6. DeepSeek's peak pricing window? Exactly 9-to-6. The moments you need AI to draft proposals, research topics, run analysis — those are the moments they tell you it costs double.
This isn't a discount program. This is a "necessity tax."

Surge Pricing Is Just the Beginning: Three Data Points
If this were just DeepSeek, you could switch. But look at the broader landscape.
First, no major model provider is cutting prices. OpenAI's GPT-5 holds steady at $15-30 per million tokens. Claude Sonnet 5 at $2.50-15. Domestic players like Doubao and Qianwen may look cheaper, but free tiers are steadily shrinking.
Second, compute costs keep rising. NVIDIA H100 GPUs have gone from roughly $30,000 in 2023 to over $40,000. B200 cards are stratospherically expensive. Major AI companies are pouring tens of billions into infrastructure — ByteDance at an estimated 160 billion yuan this year, Alibaba Cloud in the hundreds of billions, OpenAI and Microsoft's joint data center project surpassing $100 billion.
Third, the free model is collapsing. Doubao and Qianwen killing their agent features was the canary in the coal mine. Kimi and Tencent Yuanbao may still be free today, but industry-wide, the shift to paid is inevitable.
The bottom line: the era of "use AI as much as you want, whenever you want, for free" is ending. And it's ending fast.
The Only Escape: Bring the Models Home
Faced with this pricing architecture, individual users have two choices: accept ever-growing API bills, or change the model entirely — stop running AI in the cloud.
Kaihe AIBOX takes the second path. It's a dedicated AI microcomputer — about the size of your palm — that runs OpenClaw and Hermes agent frameworks locally. The architecture is edge-cloud collaborative: local agents handle scheduling and execution, while cloud models handle inference. The difference from buying API access directly: you're not paying per token. There's no peak surcharge. You use it whenever you want — 3 AM or 10 AM on a Tuesday, same cost: just the electricity and bandwidth.
More importantly, a locally deployed AI doesn't get rug-pulled by platform policy changes. Your agent configs, memory, and knowledge base live on your own hard drive. Whatever pricing games the cloud providers play, they don't touch you.

Let's Run the Numbers
A typical content creator uses AI roughly 100 times a day, with about 3,000 tokens of output per request. Under DeepSeek V4 Pro's peak pricing (12 yuan per million output tokens), that's about 108 yuan per month just for output. Input tokens add more — the real monthly bill lands somewhere between 150 and 200 yuan.
If you're a peak-hour user year-round, you're looking at roughly 1,300 yuan annually just for DeepSeek's output costs. A Kaihe AIBOX is a one-time hardware purchase. Model access isn't billed per token. Use it for a year, it pays for itself. Year two is pure savings.
And you never have to worry about peak windows, token counting, or dragging yourself to work at midnight just to save a few bucks.
A Final Thought
The fact that AI models are introducing surge pricing isn't surprising — no commercial company burns cash forever. What's unsettling is the logic: they've precisely targeted the hours you need AI the most and priced them the highest.
Some will say, "Just use it at night — it's 80% off." But think about that. Should your workflow have to bend around an AI company's pricing schedule? You're at your most productive during the day, but you can't freely use AI then. You have to wait until midnight for the discount. At that point, who's serving whom?
That's what makes surge pricing so uncomfortable. It's not the cost. It's the attempt to control your rhythm.
If you don't want your rhythm dictated by someone else's pricing model — bring the AI home. You decide when to use it.
Further Reading
- Kaihe AIBOX-A1 Product Details — Local AI agent hardware
- Kaihe AIBOX Store — Full lineup of Agent Computers
- Doubao & Qianwen Kill Agents: Your AI Could Disappear
- Kaihe AIBOX Review: 5 Hours Saved Every Day
Kaihe AIBOX Website: https://agentaibox.com/ Email: [email protected] #KaiheAIBOX #AIAgent #OpenSource #ArtificialIntelligence #DeepSeek #LLM #AIPricing
Kaihe AIBOX | The Agent Computer That Works 7x24 for You · AI Frontier