Reference Source

qwen-3-32b — Deprecated

For qwen-3-32b, provider is Cerebras; lifecycle status is Deprecated; deprecation announced is 2026-02-16; recommended replacement is GPT OSS 120B; platform / deployment scope is Cerebras Inference API, recorded from its source on 2026-09-09.

Provider
Cerebras our reading
Model ID / API name
qwen-3-32b verified
Lifecycle status
Deprecated verified
Deprecation announced
2026-02-16 verified
Recommended replacement
GPT OSS 120B verified
Platform / deployment scope
Cerebras Inference API our reading
Shutdown / retirement date
4/6/2026 per docs.sambanova.ai
Sourceinference-docs.cerebras.ai
Verified
Review by
DatasetAI model deprecation and retirement dates by provider

Values marked our reading are our classification of what the source says — the source does not print them in those words. The quote below is the evidence for each one; judge it yourself.

What the source says

2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .

— inference-docs.cerebras.ai, retrieved 2026-09-09

Sources disagree

More than one authority states this, and they do not state the same thing. Both are reproduced with the source each came from — deciding between them is yours, not ours.

Provider

inference-docs.cerebras.ai says provider is Cerebras, as of 2026-09-09.

2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .

https://inference-docs.cerebras.ai/support/deprecation

docs.sambanova.ai says provider is SambaNova, as of 2026-09-15.

| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |

https://docs.sambanova.ai/docs/en/models/deprecations.md

console.groq.com says provider is Groq, as of 2026-09-09.

July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |

https://console.groq.com/docs/deprecations

docs.tokenfactory.nebius.com says provider is Nebius Token Factory, as of 2026-09-16.

| [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) |

https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md

Lifecycle status

inference-docs.cerebras.ai says lifecycle status is Deprecated, as of 2026-09-09.

2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .

https://inference-docs.cerebras.ai/support/deprecation

docs.sambanova.ai says lifecycle status is Deprecated, as of 2026-09-15.

| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |

https://docs.sambanova.ai/docs/en/models/deprecations.md

console.groq.com says lifecycle status is End-of-Life, as of 2026-09-09.

July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |

https://console.groq.com/docs/deprecations

docs.tokenfactory.nebius.com says lifecycle status is Removed, as of 2026-09-16.

| [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) |

https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md

Deprecation announced

inference-docs.cerebras.ai says deprecation announced is 2026-02-16, as of 2026-09-09.

2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .

https://inference-docs.cerebras.ai/support/deprecation

docs.sambanova.ai says deprecation announced is March 31, 2026, as of 2026-09-15.

| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |

https://docs.sambanova.ai/docs/en/models/deprecations.md

console.groq.com says deprecation announced is June 17, 2026, as of 2026-09-09.

July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |

https://console.groq.com/docs/deprecations

Recommended replacement

inference-docs.cerebras.ai says recommended replacement is GPT OSS 120B, as of 2026-09-09.

2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .

https://inference-docs.cerebras.ai/support/deprecation

docs.sambanova.ai says recommended replacement is MiniMax-M2.5, as of 2026-09-15.

| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |

https://docs.sambanova.ai/docs/en/models/deprecations.md

console.groq.com says recommended replacement is openai/gpt-oss-120b, as of 2026-09-09.

July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |

https://console.groq.com/docs/deprecations

docs.tokenfactory.nebius.com says recommended replacement is nvidia/Nemotron-3_5-Lightning, as of 2026-09-16.

| [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) |

https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md

Platform / deployment scope

inference-docs.cerebras.ai says platform / deployment scope is Cerebras Inference API, as of 2026-09-09.

2026-02-16 Deprecated qwen-3-32b and llama-3.3-70b We recommend migrating to GPT OSS 120B .

https://inference-docs.cerebras.ai/support/deprecation

docs.sambanova.ai says platform / deployment scope is SambaNova SambaCloud, as of 2026-09-15.

| `Qwen3-32B` | 4/6/2026 | `MiniMax-M2.5` |

https://docs.sambanova.ai/docs/en/models/deprecations.md

console.groq.com says platform / deployment scope is Free and developer-tier usage — enterprise customers with a committed-spend contract are not affected, as of 2026-09-09.

July 17, 2026: qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct In line with our commitment to bringing you cutting-edge models, on June 17, 2026, we emailed users to announce the deprecation of qwen/qwen3-32b and meta-llama/llama-4-scout-17b-16e-instruct. We recommend migrating to openai/gpt-oss-120b (for Qwen 3 32B) and openai/gpt-oss-120b or qwen/qwen3.6-27b (for Llama 4 Scout 17B), which deliver exceptional performance with faster inference. This deprecation applies to free and developer-tier usage; enterprise customers with a committed-spend contract are not affected. Deprecated Model | Shutdown Date | Recommended Replacement Model ID | qwen/qwen3-32b | 07/17/26 | openai/gpt-oss-120b |

https://console.groq.com/docs/deprecations

docs.tokenfactory.nebius.com says platform / deployment scope is Nebius Token Factory serverless (dedicated endpoints unaffected), as of 2026-09-16.

| [`Qwen/Qwen3-32B`](https://tokenfactory.nebius.com/models/catalog/text2text/Qwen%2FQwen3-32B) | [`nvidia/Nemotron-3_5-Lightning`](https://tokenfactory.nebius.com/models/catalog/text2text/nvidia%2FNemotron-3_5-Lightning) |

https://docs.tokenfactory.nebius.com/august-2026-deprecation-notice.md

Sources

Last verified against source: . Due for re-check by . This page as Markdown · OKF bundle · full dataset as JSON.

This one changes, and we watch it.AI model deprecation and retirement dates by provider is re-read on a schedule and every change is dated. Subscribe: Atom feed · JSON · what has changed so far.