Fast
fastQuestions, lookups, small edits and the quick back-and-forth inside a session. Thinks only when the task needs it. Close calls round down.
- Models
- DeepSeek 4.1 Flash for every task
- Effort
- off for trivial and simple, medium for moderate and hard
- Latency
- first byte p50 100 to 290 ms, total p50 2.2 to 4.9 s
- Failover
- GLM 5.3
- Price
- $0.15 in, $0.60 out, $0.03 cached per 1M tokens