:metal: TT-NN operator library, and TT-Metalium low level kernel programming model.
Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.
Fast Multimodal LLM on Mobile Devices
Local AI Assistant on your phone
High-speed Large Language Model Serving for Local Deployment