Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Jupyter Notebookefeslab/Nanoflow

Nanoflow

A throughput-oriented high-performance serving framework for LLMs

41.7/100
975Forks: 53
View on GitHubHomepage →
Loading report...

Similar Projects

vllm

93

A high-throughput and memory-efficient inference and serving engine for LLMs

Python91.2K

llama-cookbook

71

Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services

Jupyter Notebook18.6K

LLMs-from-scratch

81

Implement a ChatGPT-like LLM in PyTorch from scratch, step by step

Jupyter Notebook104.5K

llm-app

70

Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more.

Jupyter Notebook59.0K
Back to List