Brief AI generation latency
April 22, 2026 · 23 minutes · minor
- ResolvedApril 22, 14:18 UTCAI generation latency back to baseline. Root cause: an upstream rate limit at our AI provider; we've added local caching for prompts to reduce upstream calls during similar conditions.
- IdentifiedApril 22, 14:02 UTCIdentified upstream rate limit as the cause. Engaging the AI provider's support team and routing prompts through a fallback model.
- InvestigatingApril 22, 13:55 UTCInvestigating elevated latency on /api/ai/* endpoints. Image-to-system requests are queueing.