# qwen-3-32b — Deprecated For qwen-3-32b, provider is Cerebras; lifecycle status is Deprecated; deprecation announced is 2026-02-16; recommended replacement is GPT OSS 120B; platform / deployment scope is Cerebras Inference API, recorded from its source on 2026-09-09. - **Provider:** Cerebras _(our reading, not quoted from the source)_ - **Model ID / API name:** qwen-3-32b _(verified: appears in the quote below)_ - **Lifecycle status:** Deprecated _(verified: appears in the quote below)_ - **Deprecation announced:** 2026-02-16 _(verified: appears in the quote below)_ - **Recommended replacement:** GPT OSS 120B _(verified: appears in the quote below)_ - **Platform / deployment scope:** Cerebras Inference API _(our reading, not quoted from the source)_ - **Shutdown / retirement date:** 4/6/2026 _(per docs.sambanova.ai, not stated by the source above)_ ## What the source says > 2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B . ## Sources disagree More than one authority states this, and they do not state the same thing. Both are reproduced with the source each came from. ### Provider inference-docs.cerebras.ai says provider is **Cerebras**, as of 2026-09-09. > 2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B . Source: https://inference-docs.cerebras.ai/support/deprecation docs.sambanova.ai says provider is **SambaNova**, as of 2026-09-15. > | `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` | Source: https://docs.sambanova.ai/docs/en/models/deprecations.md console.groq.com says provider is **Groq**, as of 2026-09-09. > July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b | Source: https://console.groq.com/docs/deprecations docs.tokenfactory.nebius.com says provider is **Nebius Token Factory**, as of 2026-09-16. > | [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) | Source: https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md ### Lifecycle status inference-docs.cerebras.ai says lifecycle status is **Deprecated**, as of 2026-09-09. > 2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B . Source: https://inference-docs.cerebras.ai/support/deprecation docs.sambanova.ai says lifecycle status is **Deprecated**, as of 2026-09-15. > | `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` | Source: https://docs.sambanova.ai/docs/en/models/deprecations.md console.groq.com says lifecycle status is **End-of-Life**, as of 2026-09-09. > July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b | Source: https://console.groq.com/docs/deprecations docs.tokenfactory.nebius.com says lifecycle status is **Removed**, as of 2026-09-16. > | [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) | Source: https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md ### Deprecation announced inference-docs.cerebras.ai says deprecation announced is **2026-02-16**, as of 2026-09-09. > 2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B . Source: https://inference-docs.cerebras.ai/support/deprecation docs.sambanova.ai says deprecation announced is **March 31, 2026**, as of 2026-09-15. > | `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` | Source: https://docs.sambanova.ai/docs/en/models/deprecations.md console.groq.com says deprecation announced is **June 17, 2026**, as of 2026-09-09. > July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b | Source: https://console.groq.com/docs/deprecations ### Recommended replacement inference-docs.cerebras.ai says recommended replacement is **GPT OSS 120B**, as of 2026-09-09. > 2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B . Source: https://inference-docs.cerebras.ai/support/deprecation docs.sambanova.ai says recommended replacement is **MiniMax-M2.5**, as of 2026-09-15. > | `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` | Source: https://docs.sambanova.ai/docs/en/models/deprecations.md console.groq.com says recommended replacement is **openai/gpt-oss-120b**, as of 2026-09-09. > July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b | Source: https://console.groq.com/docs/deprecations docs.tokenfactory.nebius.com says recommended replacement is **nvidia/Nemotron-3_5-Lightning**, as of 2026-09-16. > | [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) | Source: https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md ### Platform / deployment scope inference-docs.cerebras.ai says platform / deployment scope is **Cerebras Inference API**, as of 2026-09-09. > 2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B . Source: https://inference-docs.cerebras.ai/support/deprecation docs.sambanova.ai says platform / deployment scope is **SambaNova SambaCloud**, as of 2026-09-15. > | `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` | Source: https://docs.sambanova.ai/docs/en/models/deprecations.md console.groq.com says platform / deployment scope is **Free and developer-tier usage — enterprise customers with a committed-spend contract are not affected**, as of 2026-09-09. > July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b | Source: https://console.groq.com/docs/deprecations docs.tokenfactory.nebius.com says platform / deployment scope is **Nebius Token Factory serverless (dedicated endpoints unaffected)**, as of 2026-09-16. > | [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) | Source: https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md ## Source - https://inference-docs.cerebras.ai/support/deprecation - https://docs.sambanova.ai/docs/en/models/deprecations.md - https://console.groq.com/docs/deprecations - https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md Last verified: 2026-09-09. Review by: 2026-10-09. Part of [AI model deprecation and retirement dates by provider](https://referencesource.org/ai-model-deprecation-and-retirement/).