CNLab.ai Solutions

Three deployment models. One operating model.

Choose the cluster topology that matches your scale today. Expand without rewrites tomorrow. Every CNLab deployment shares the same scheduler, the same APIs, and the same admin console — only the resource boundary changes.

Supported Environments

Your code asks for a GPU. We route the rest.

Auto-detection of deep-learning GPU resource allocation timing — across every environment your researchers already use.

📓

Jupyter Notebook

Open a notebook, import torch, call .cuda() — CNLab routes the request to a sliced or full GPU on the closest cluster. No code changes, no kernel surprises.

💻

VS Code & IDE Forks

Cursor, Windsurf, JetBrains — the CNLab agent runs as a remote runtime. Workspace files, secrets, and SSH keys are forwarded transparently.

⌨️

CLI

cnlab run train.py --gpu h100-25 queues your script on the next available 25%-block H100. Logs stream back; the GPU releases on exit.

Decision Helper

Platform Selection Matrix

Choose by org size, GPU pool, and security posture. Every CNLab tier is forward-compatible — Single grows into Multi grows into Hybrid.

Dimension On-Premise Single Multi On-Premise Hybrid Cloud
Recommended forLabs, small AI teamsMulti-campus universities, mid-sized AIEnterprise, AI scale-ups
Typical pool size4–32 GPUs32–256 GPUs128+ GPUs + cloud burst
Initial costLowestMediumMedium + cloud opex
TCOLowestLow−50%+ vs. cloud-only
FailoverSingle clusterCross-cluster live migrationCross-cluster + cloud
Heterogeneous GPUNVIDIA + AMDNVIDIA + AMD + cloud
Data residency100% on-prem100% on-premOn-prem first, encrypted egress
Min. allocation unit1% Block1% Block1% Block
Setup timeDaysWeeksWeeks

Not sure which fits? Talk to an engineer →

Common Core

One scheduler. Three topologies.

Every CNLab deployment runs the same orchestration primitives. Switching topology is a configuration change, not a rewrite.

📐

GPU Sharing

1% Block + MIG. 100 simultaneous tenants per H100.

Scheduling

Priority + deadline queues. Proactive memory control.

🔄

Live Migration

Zero-downtime workload hand-off between any two pools.

🛡️

RBAC

User & project-level access control with full audit.

Ready to Optimize Your Infrastructure?

Book a 30-minute architecture review. We'll map your current GPU footprint, model the savings against your cloud spend, and recommend the right tier.