About This Resource

A self-hosted service for running language, vision, speech, and other AI models through compatible APIs. It supports multiple inference backends and hardware configurations, allowing deployments to choose the models and runtimes they need.