Choose Your Platform
Smart Studio offers two deployment paths. Choose the one that fits your use case:
- BYO-GPU (Model Service Platform) — Deploy open-source models on your own GPUs for internal use or integration into an existing platform.
- BYO-Key (API Router Platform) — Resell Model APIs through a white-labelled User Console and Admin Console.
Both paths share the same activation workflow: choose a billing model, install the agent, verify your machines, and deploy the platform.
Start with the outcome you need
Choose one maintained path. Each path connects the prerequisite, Console action, successful result, recovery guidance, and next step.
Activate a GPU cluster
Result: Your cluster is Activated, its expected machines are visible, and its GPU capacity is ready for model deployment.
- 1Check requirementsConfirm supported infrastructure, network access, and operator permissions.Prerequisite
- 2Choose the billing modelSelect Model Service Platform and confirm the billing model (Token Sharing or Fixed Amount Fee).Decision
- 3Activate the clusterInstall the platform agent, submit machine configuration, and complete the guided checks.Core action
- 4Verify machines and capacityConfirm the activation state, expected nodes, available GPUs, and resource health.Verify
- 5Deploy a modelChoose a platform model in Model Gallery or upload your own, then create a deployment.Next step
BYO-GPU Path
| Stage | What you do | Result |
|---|---|---|
| 1. Prepare | Check the network, CPU server, GPU nodes, drivers, and account permissions. | Your environment is ready for validation. |
| 2. Choose billing model | Select Model Service Platform and confirm Token Sharing or Fixed Amount Fee. | Billing model is locked for this cluster. |
| 3. Activate | Install the Smart Studio agent and submit your machine configuration. | Smart Studio verifies every server and deploys the platform. |
| 4. Verify | Open GPU Dashboard from an activated cluster. | You can inspect capacity, utilization, nodes, and model-serving health. |
| 5. Deploy | Browse Model Gallery or upload a custom model, then create a deployment. | The model appears in Deployments and progresses to Ready. |
BYO-Key Path
| Stage | What you do | Result |
|---|---|---|
| 1. Prepare | Check the network, CPU server, and account permissions. | Your environment is ready for validation. |
| 2. Choose billing model | Select API Router Platform and confirm Token Sharing. | Billing model is locked for this cluster. |
| 3. Activate | Install the Smart Studio agent and submit your machine configuration. | Smart Studio verifies every server and deploys the platform. |
| 4. Access consoles | Log in to Admin Console to onboard models, or User Console for end-users. | Platform is accessible and ready for API reselling. |
Before You Start
You need:
- An Alibaba Cloud account.
- A CPU server that can reach the internet during activation.
- One or more supported GPU servers that meet the BYO-GPU requirements.
- SSH access and the network details required by the activation wizard.
Docs 2.0 follows the current public Console menu. If a capability is not available in the Console, it is not presented as a supported 2.0 workflow.
Choose Your Next Step
New to Smart Studio
Start with Activate Cluster. It covers billing model selection, agent installation, machine verification, platform deployment, and recovery from common failures.
Your cluster is already activated
Open GPU Dashboard to confirm resource health, then go to Model Gallery or My Models.
You already have a model
Follow Create Deployment, then use Manage Deployments to verify its state and manage its lifecycle.
Where Each Console Menu Leads
| Console menu | Use it for |
|---|---|
| Activate Cluster | Choose a billing model, connect GPU infrastructure, and deploy the platform. |
| Model Gallery | Discover platform models and continue to deployment. |
| My Models | Upload and manage custom or fine-tuned model assets. |
| Deployments | Create, monitor, stop, restart, edit, or delete deployments. |
| Datasets / Fine-Tuning / Evaluations | Prepare data, improve models, and measure quality. |
| Usage | Review consumption by self-deployed workloads. |
| Billing | Review charges and billing details. |