Hugging Face (Twitter)
RT @alvarobartt: 👾 `hf-mem` is all you need to estimate the required VRAM for inference of any model on @huggingface based on Safetensors metadata.
- Written in Python
- Lightweight, only depends on `httpx`
- Runs w/ @astral_sh `uvx` as `uvx hf-mem --model-id ...`
- Works with any Safetensors repository
- Output inspired by @usgraphics TR-100 Machine Report
Post #2138
20
