Bento is an inference platform designed for deploying, managing, and optimizing AI and machine learning models at scale. It supports self-hosted and cloud-managed deployments, offers a CLI and API, and provides tools for observability, access control, and efficient scaling. Bento is aimed at machine learning engineers and teams seeking robust model serving infrastructure.
Bento is an AI & ML product. It simplifies deploying, managing, and scaling AI/ML model inference for production workloads. It is built as a B2B product for machine learning engineers. There is a free tier. Bento is available on the web, the command line, and API, and it can be self-hosted.
BentoML builds and maintains Bento, and the product first shipped in 2019. Development happens publicly on GitHub with 8.7k stars and 6 commits in the last 90 days. Key capabilities include model deployment, inference optimization, and scalable serving. It exposes integrations via a public API.
Latest indexed changes and source events
Other apps tracked under the same category.