Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Goome-projects/ome

ome

Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

78.9/100
507Forks: 96
View on GitHubHomepage →
Loading report...

Similar Projects

ollama

95

Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.

Go180.4K

DeepSeek-Reasonix

89

DeepSeek-native AI coding agent for your terminal. Engineered around prefix-cache stability — leave it running.

Go35.4K

beta9

85

Ultrafast serverless GPU inference, sandboxes, and background jobs

Go1.8K

kubeai

82

AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.

Go1.3K
Back to List