aiAI Status Monitor

Model Performance Issue: llama-3.1-8b-instant

DegradedResolved · started Nov 4, 2025, 2:11 AM UTC (336d ago)

Was Groq down? Groq reported “Model Performance Issue: llama-3.1-8b-instant” on Nov 4, 2025, 2:11 AM UTC. Full timeline below. This incident affects Groq — see live Groq status. Source: official vendor status feed.

Timeline

  1. Nov 4, 2025, 2:47 AM UTC

    The issues affecting llama-3.1-8b-instant have been **resolved**. The model is operating normally. Root cause: Two recent changes to the inference-engine-instances repository that were contributing to the elevated Orion LoRA latencies were reverted . We apologize for the disruption and thank you for your patience.

  2. Nov 4, 2025, 2:33 AM UTC

    We have **implemented a fix** for llama-3.1-8b-instant service and performance is improving. We’re now **monitoring** the model to ensure stability. If all remains normal, we will resolve the incident in the next update.

  3. Nov 4, 2025, 2:11 AM UTC

    We have **identified an issue** affecting the lama-3.1-8b-instant service. A fix is in progress. Users may still experience latency until the fix completes.

More Groq outages