High-performance runtime for running generative AI models on-device with ONNX
Kubernetes infrastructure for managing isolated AI agent runtimes with stable identity and persistent storage