Import Qwen 3.8 into OCI Generative AI

You can now import the Qwen/Qwen3.8-2.4T-A95B-FP8 model into OCI Generative AI, create endpoints for it, and use it in the Generative AI service.

Qwen3.8-2.4T-A95B-FP8 is an FP8 Mixture-of-Experts (MoE) text-generation model with 2.4 trillion total parameters and 95 billion active parameters. The model is designed for coding, professional work, research, and long-horizon agentic tasks and supports a native context length of 262,144 tokens that can be extended to 1,010,000 tokens.

In OCI Generative AI, use the following details for this model:

  • Model capability: TEXT_TO_TEXT
  • Minimum dedicated AI cluster unit shape: B200_X16

For the complete list of models that you can import, see Models for Import. For available hardware and deployment steps, see Managing Imported Models.