News

What's happening in robotics and AI — curated by the wujec.ai editors.

Business8/19/2026 · DeepSeek — dokumentacja API (Models & Pricing)

DeepSeek's clock is a Beijing clock: the same working day costs 75% more there than in San Francisco

DeepSeek's price list, in force since 16 August 2026, is the first from a major model provider in which the hour of the day is a full pricing variable rather than a promotion: peak hours are 01:00–04:00 and 06:00–10:00 UTC, and everything outside them costs exactly half. The published windows say nothing about where those hours fall. Converted, they say a great deal. Beijing keeps UTC+8 all year, with no daylight saving. The two peak windows land at 09:00–12:00 and 14:00–18:00 local time — Chinese office hours. The gap between them, 04:00–06:00 UTC, is 12:00–14:00 in Beijing: the lunch break, priced at half rate. The tariff is not described in geographic terms anywhere in the documentation, but it is drawn around one country's working day. That has an arithmetic consequence nobody has published. Take a team using DeepSeek-V4-Pro evenly through a 09:00–17:00 local working day, and price a million output tokens ($3.96 at peak, $1.98 off-peak). In San Francisco the local working day is 16:00–24:00 UTC and misses both peak windows entirely, so every hour bills at $1.98. In Warsaw it is 07:00–15:00 UTC, of which three hours are peak: $2.72 per million, 37.5% above San Francisco. In Beijing it is 01:00–09:00 UTC, of which six hours are peak: $3.47 per million — 75% above San Francisco, for identical work on identical weights. Bangalore lands in between at $3.09, or 56% above. European bills also move with the clock change. In winter Warsaw shifts to UTC+1, only two working hours stay inside the peak, and the same million falls to $2.48 — a 9% discount granted by nothing but the end of daylight saving. What this does not show: DeepSeek is not charging anyone by location. The tariff is identical worldwide and time-based, and the company presents it as load management — the peak windows are simply when its servers are busiest, which is when its home market is at work. Batch and overnight jobs can be moved into the cheap hours by anyone, anywhere, and for most production workloads cached input, billed at a fiftieth of a cache miss, matters far more than the clock. The figures above assume usage spread evenly across office hours, which no real team does exactly. But the direction is not an artefact: the further a user's working day sits from Beijing's, the less DeepSeek's price rise costs them.

DeepSeek-V4-Pro
Business8/15/2026 · DeepSeek — dokumentacja API (cennik)

The cheapest lab stops being cheap: DeepSeek raises API prices from 16 August and starts charging by the clock

DeepSeek has published a new price list for its V4 models, effective 16 August 2026. The company that built its reputation on undercutting everyone else is raising rates across the board — by between 50 percent and more than 1,100 percent, depending on the model, the type of token and the hour of the day. The numbers for DeepSeek-V4-Flash: output tokens go from $0.28 to $1.32 per million at peak, cache-miss input from $0.14 to $0.44, and cached input from $0.0028 to $0.014 — a fivefold rise on the cheapest line in the catalogue. DeepSeek-V4-Pro goes from $0.87 to $3.96 per million output tokens at peak, while its cached input rises from roughly $0.0036 to $0.044 per million: the 1,100 percent figure comes from that one line, not from the headline rate. The structural change matters as much as the numbers. From Sunday the price depends on when the request is made: peak hours are 01:00–04:00 and 06:00–10:00 UTC, everything else is off-peak and costs exactly half. DeepSeek says the tiered structure is meant to "allocate resources more reasonably" and push developer workloads towards less congested windows. In practice it is the first time a major model provider has made the hour of the day a first-class pricing variable rather than a promotional discount. The timing is what makes this striking. In the same week Anthropic cancelled a planned 50 percent rise for Claude Sonnet 5 and Google put an expiry date on its Gemini 3.7 Flash discount, DeepSeek moved in the opposite direction — and moved further than either. Off-peak V4-Flash output at $0.66 per million is still cheap by Western standards, but the gap that made DeepSeek an obvious default has narrowed by a factor of four or five. Both V4 models keep their 1M-token context window and 384K maximum output; the concurrency limits (2,500 simultaneous requests for Flash, 500 for Pro) are unchanged. We have updated the pricing in both profiles in the catalogue.

DeepSeek-V4-Flash
Releases8/6/2026 · OpenAI — API deprecations

The original GPT-5 gets a death date: OpenAI shuts down the August 2025 snapshot on 11 December

OpenAI's deprecation page now carries a line that would have been hard to imagine a year ago: gpt-5-2025-08-07, the snapshot that launched GPT-5, will be shut off on 11 December 2026. The reasoning model o3-2025-04-16 goes the same day. Both point to a single recommended replacement, gpt-5.6-sol. The original GPT-5 will therefore have lived about sixteen months in the API. That is not unusual by current standards, but it is a striking number for a model that was treated as a generational landmark on release, and it says something about how fast the naming has moved: five point-releases have shipped in the time GPT-4 spent as the default. A nearer deadline is already in force. The aliases gpt-5.2-chat-latest and gpt-5.3-chat-latest stop responding on 10 August 2026, four days from now, also redirecting developers to gpt-5.6-sol. Anyone still calling the -chat-latest names has effectively no migration window left. The rest of the autumn calendar is dense. The Assistants API is removed on 26 August, a year after its replacement by the Responses and Conversations APIs was announced. Sora 2, sora-2-pro and the Videos API end on 24 September with no recommended successor at all — the only entry on the list where OpenAI does not name a replacement. The legacy realtime and audio families survive until 20 January 2027. The consumer side has its own, earlier deadline, announced on 28 May 2026: o3 disappears from ChatGPT on 26 August, at the end of a 90-day sunset, and takes GPT-4.5 with it. That removes the last GPT-4-family model from the product, after GPT-4o and the GPT-4.1 variants were withdrawn in February 2026. The API is unaffected by that change — o3 keeps answering there until December — so for a few months the same model will be alive for developers and gone for everyone else. For a catalogue like this one the pattern matters more than any single date: a frontier model is now a product with a service life measured in months, and the profile of a model increasingly needs to record when it stops answering, not just when it launched.

GPT-5