Factual package intelligence from PyPI
PetaLLM allows single 4GB GPU card to run 70B large language models without quantization, distillation or pruning. 8GB vmem to run 405B Llama3.1.
pip install petallm
PyPI declares 8 unique dependency rules for this release. Environment markers are shown when supplied by the project.
petallm publishes 1 wheel and 0 source archives for version 2.11.0. Wheel platform tags: any.
No version-specific Python classifiers are declared.
PyPI lists 1 release with files. The first dated release is ; 1 release falls within the 365 days preceding the latest dated release. The current release files were uploaded on .