Compatibility

Built for the GPUs you already run.

GPU architectures, precisions, and runtimes supported by the Kernova AI optimization engine.

GPUArchitectureFP8BF16FP16INT8
H100 / H200 (SXM, PCIe)HopperBetaSupportedSupportedSupported
A100 40GB / 80GBAmperePlannedSupportedSupportedSupported
L40S, L4Ada LovelaceBetaSupportedSupportedBeta
RTX 4090 / 6000 AdaAda LovelacePlannedBetaSupportedBeta
B200BlackwellPlannedPlannedPlannedPlanned

Frameworks & runtimes

  • PyTorch 2.1+Supported
  • Triton 2.x / 3.xSupported
  • CUDA 12.xSupported
  • vLLMBeta
  • Hugging Face TGIBeta
  • TensorRT-LLMPlanned

Support matrix reflects the current roadmap and should be confirmed against your release notes before launch.