Triton
Triton

Projects Tagged 'triton' - GPU-Accelerated ML Kernels, Triton Inference Server Integrations, and Model Acceleration

Explore projects tagged triton that implement high-performance GPU kernels, Triton Inference Server deployments, and custom kernel optimizations for PyTorch and TensorFlow models. This curated list of projects using triton highlights real-world implementations for low-level GPU kernel development, mixed-precision and tensor-core optimization, model serving at scale, and end-to-end inference pipelines; use long-tail searches like "projects using triton for GPU kernels" or "triton inference server projects" to find benchmarked examples. Filter the results by GPU architecture, framework integration, model type (transformers, CNNs), or deployment target to surface source code, performance benchmarks, and deployment recipes that you can reproduce. Browse and filter these projects to accelerate development, optimize inference performance with Triton, and deploy production-ready ML systems today.
Categories
Other Filters