qwen-3-32b — Deprecated
For qwen-3-32b, provider is Cerebras; lifecycle status is Deprecated; deprecation announced is 2026-02-16; recommended replacement is GPT OSS 120B; platform / deployment scope is Cerebras Inference API, recorded from its source on 2026-09-09.
- Provider
- Cerebras our reading
- Model ID / API name
- qwen-3-32b verified
- Lifecycle status
- Deprecated verified
- Deprecation announced
- 2026-02-16 verified
- Recommended replacement
- GPT OSS 120B verified
- Platform / deployment scope
- Cerebras Inference API our reading
- Shutdown / retirement date
- 4/6/2026 per docs.sambanova.ai
Values marked our reading are our classification of what the source says — the source does not print them in those words. The quote below is the evidence for each one; judge it yourself.
What the source says
2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .
— inference-docs.cerebras.ai, retrieved 2026-09-09
Sources disagree
More than one authority states this, and they do not state the same thing. Both are reproduced with the source each came from — deciding between them is yours, not ours.
Provider
inference-docs.cerebras.ai says provider is Cerebras, as of 2026-09-09.
2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .
https://inference-docs.cerebras.ai/support/deprecation
docs.sambanova.ai says provider is SambaNova, as of 2026-09-15.
| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |
https://docs.sambanova.ai/docs/en/models/deprecations.md
console.groq.com says provider is Groq, as of 2026-09-09.
July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |
https://console.groq.com/docs/deprecations
docs.tokenfactory.nebius.com says provider is Nebius Token Factory, as of 2026-09-16.
| [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) |
https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md
Lifecycle status
inference-docs.cerebras.ai says lifecycle status is Deprecated, as of 2026-09-09.
2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .
https://inference-docs.cerebras.ai/support/deprecation
docs.sambanova.ai says lifecycle status is Deprecated, as of 2026-09-15.
| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |
https://docs.sambanova.ai/docs/en/models/deprecations.md
console.groq.com says lifecycle status is End-of-Life, as of 2026-09-09.
July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |
https://console.groq.com/docs/deprecations
docs.tokenfactory.nebius.com says lifecycle status is Removed, as of 2026-09-16.
| [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) |
https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md
Deprecation announced
inference-docs.cerebras.ai says deprecation announced is 2026-02-16, as of 2026-09-09.
2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .
https://inference-docs.cerebras.ai/support/deprecation
docs.sambanova.ai says deprecation announced is March 31, 2026, as of 2026-09-15.
| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |
https://docs.sambanova.ai/docs/en/models/deprecations.md
console.groq.com says deprecation announced is June 17, 2026, as of 2026-09-09.
July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |
https://console.groq.com/docs/deprecations
Recommended replacement
inference-docs.cerebras.ai says recommended replacement is GPT OSS 120B, as of 2026-09-09.
2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .
https://inference-docs.cerebras.ai/support/deprecation
docs.sambanova.ai says recommended replacement is MiniMax-M2.5, as of 2026-09-15.
| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |
https://docs.sambanova.ai/docs/en/models/deprecations.md
console.groq.com says recommended replacement is openai/gpt-oss-120b, as of 2026-09-09.
July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |
https://console.groq.com/docs/deprecations
docs.tokenfactory.nebius.com says recommended replacement is nvidia/Nemotron-3_5-Lightning, as of 2026-09-16.
| [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) |
https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md
Platform / deployment scope
inference-docs.cerebras.ai says platform / deployment scope is Cerebras Inference API, as of 2026-09-09.
2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .
https://inference-docs.cerebras.ai/support/deprecation
docs.sambanova.ai says platform / deployment scope is SambaNova SambaCloud, as of 2026-09-15.
| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |
https://docs.sambanova.ai/docs/en/models/deprecations.md
console.groq.com says platform / deployment scope is Free and developer-tier usage — enterprise customers with a committed-spend contract are not affected, as of 2026-09-09.
July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |
https://console.groq.com/docs/deprecations
docs.tokenfactory.nebius.com says platform / deployment scope is Nebius Token Factory serverless (dedicated endpoints unaffected), as of 2026-09-16.
| [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) |
https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md
Sources
- inference-docs.cerebras.aihttps://inference-docs.cerebras.ai/support/deprecation
- docs.sambanova.aihttps://docs.sambanova.ai/docs/en/models/deprecations.md
- console.groq.comhttps://console.groq.com/docs/deprecations
- docs.tokenfactory.nebius.comhttps://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md