model serving
model serving

Projects by Tags — model-serving: Scalable, Low-Latency Model-Serving Solutions for Production Inference

Explore projects tagged with model-serving in our projects list under the tags pillar to discover open-source and commercial solutions for production inference, including Kubernetes-native frameworks (KServe/KFServing), TensorFlow Serving, TorchServe, and serverless or edge model serving for low-latency, GPU-accelerated workloads. This curated collection emphasizes long-tail capabilities like autoscaling, model versioning and registries, canary rollouts, A/B testing, observability, and latency/throughput optimization; use filters to compare deployment platform, runtime, hardware support (GPU/CPU), and integration options. Filter, sort, and review each project's maturity, license, and benchmarked performance to find the best model-serving projects for your ML pipelines, then click through to view code, deployment guides, and implementation details to accelerate production inference adoption.
Categories
Other Filters