DeepSeek V4-Pro Ignores Token Limits, Billing Users for Empty Responses
A developer building an AI-powered Liar's Dice game discovered that DeepSeek's V4-Pro model silently ignores token limit parameters, continuing to generate thousands of tokens regardless of the cap set. In tests conducted on August 14, 2026, three consecutive API calls with a 3,072-token budget each consumed the entire budget on internal reasoning, returning zero visible output while still charging the full cost. The API accepted unrecognized and even fabricated parameters without error, making it impossible to confirm whether any limit had taken effect. With reasoning enabled, costs ran 16.7 times higher than with it disabled, and a working disable command existed but was not consistently documented. The developer nearly misattributed the token-budget failure as a model-compliance failure, as empty responses triggered a fallback that logged the model as disobeying instructions in over 94% of game hands.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in