- 43kRaydeployment
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
43kApache-2.0 - 24kncnndeployment
ncnn is a high-performance neural network inference framework optimized for the mobile platform
24kNot specified - 21kONNX Runtimedeployment
ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator
21kMIT - 11kTriton Inference Serverdeployment
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
11kBSD-3-Clause - 11kOpenVINOdeployment
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
11kApache-2.0 - 8.7kBentoMLdeployment
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
8.7kApache-2.0 - SponsorReach 50,000+ buyers
Enterprise buyers looking for private AI solutions see your brand here.
- 5.7kKServedeployment
Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes
5.7kApache-2.0 - 4.8kSeldon Core 2deployment
An MLOps framework to package, deploy, monitor and manage thousands of production machine learning models
4.8kNot specified
Stop paying for AI APIs. Everything here runs on your hardware.
Enterprise buyers looking for private AI solutions see your brand here.
No fluff. No spam. Join 12,000+ builders.