Python AxmlParser
Fast CUDA Kernels for ResNet Inference. Using Winograd algorithm to optimize the efficiency of co...
AI-powered drum removal tool using Meta's Demucs. Drop in any song, get a drumless backing track ...
A high-throughput and memory-efficient inference and serving engine for LLMs
SGLang is a fast serving framework for large language models and vision language models.
More token goodput from frontier models. A distributed, SLA-aware serving mesh — disaggregated pr...
Simple app to learning the lure finish tech
tabnotes
super repo for rocm systems projects
AI Tensor Engine for ROCm