DeepSeek's sparse attention halves API prices again
V3.2-Exp's 'lightning indexer' attends to a fraction of long contexts, cutting inference cost enough to drop API prices 50%+ across the board — to under 3 cents per million input tokens. The Chinese price floor keeps falling, and Western per-token pricing keeps having to answer.
0 sources