Fixed 13 Aug 2026

A provider having a bad minute no longer ends your answer

When the model you asked for is down or overloaded, the request now moves to a capable sibling instead of failing: chat replies, the tool steps behind them, Continue, deep research, prompt improvements, workflow text and music renders. A reply that has already started writing is left alone, because switching mid-sentence would contradict what you have already read. The charge follows the model that actually answered, at that model's rates.