← Back to List
⚠
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Pythonpredibase/lorax

lorax

Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs

67.4/100
★ 3.8KForks: 326
View on GitHub →Homepage →
Loading report...

Similar Projects

vllm

93

A high-throughput and memory-efficient inference and serving engine for LLMs

Python★ 92.8K

OpenLLM

89

Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.

Python★ 12.5K

rtp-llm

80

RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.

Python★ 1.4K

peft

91

🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.

Python★ 21.7K
← Back to List