Aperta Scientia Crest
← All technologies
KServe

KServe

Category : AI & MLOps

Standardized, serverless model serving on Kubernetes with autoscaling.

Official website / GitHub ↗ Visit
Chouette Rôle & Utilité

What is this technology for?

Its role in production environments and why it is taught in our curriculums.

KServe standardizes model serving on Kubernetes: it deploys, scales and manages inference endpoints for ML and LLM models with serverless capabilities.

Chouette Pédagogie & Compétences

What you will learn

The hands-on skills you will gain on this technology in our curriculums.

  • Deploy models as KServe InferenceServices
  • Configure autoscaling and canary rollouts
  • Serve models with vLLM, OpenVINO and Triton runtimes
  • Secure and monitor inference endpoints
Chouette Veille Technologique

Latest News & Ecosystem Updates

Recent innovations, major releases, and key industry milestones in the ecosystem.

Model Serving 2026-05

KServe v0.15: v2 Data Plane GA & Intelligent Multi-Model Routing

Declarative serverless model serving with scale-to-zero, GPU fractional sharing, and real-time drift detection.

OpenShift AI Stack 2026-02

Automated Inference Pipeline Delivery with Knative & Istio

End-to-end mTLS encryption and fine-grained canary traffic shaping for production AI microservices.

Featured in the curriculum

Find this technology in the following modules of our Red Hat certified curriculums.

Mascotte AS300 AS300 — AI Platform Engineer

View curriculum →
M7

Red Hat OpenShift AI (RHOAI)

AI267 · Module 7 — AS300

In the same category