Distributed LLM inference — pool GPUs across multiple devices to run models no single machine can handle
Homepage PyPI
pip install distributed-llm==0.4.1
Login to resync this project