On August 27 between 18:10 and 18:22 UTC, audio transcription processing was delayed for our customers leveraging our pooled inference infrastructure.
Jobs submitted during this window were queued and processed automatically once capacity was restored — no data was lost, no job was rejected, and no action is required from our customers and users.
The delay occurred during a routine scale-up operation when our cloud GPU provider had a temporary capacity shortage.
We have restored full capacity and are widening the range of hardware we can fall back to so this cannot recur.