Huggingface Introduces Enhanced ZeroGPU Pro with 5x Usage Quota and Highest GPU Queue Priority, Alongside New L40S GPU Options
What changed
Centralized token control and approval added to Enterprise Hub
Nvidia H100 removed from Spaces Hardware on-demand section (was $10.00/hr Custom tier with 24 vCPU, 250GB RAM, 80GB VRAM). H100 remains available in Inference Endpoints.
ZeroGPU Pro feature description updated: Changed from 'Use distributed A100 hardware on your Spaces' to 'Get 5x usage quota and highest GPU queue priority' with separate Spaces Hosting feature for A100 hardware
L40S GPU options added to Spaces Hardware: 1x L40S at $1.80/hr (8 vCPU, 62GB, 48GB VRAM), 4x L40S at $8.30/hr (48 vCPU, 382GB, 192GB VRAM), 8x L40S at $23.50/hr (192 vCPU, 1534GB, 384GB VRAM)
Before and after
Jul 1, 2024 vs Nov 1, 2024Centralized token control and approval added to Enterprise Hub
Sign in to view the full comparison
See the before/after screenshots and every marked-up change.
More from Huggingface
Full history →- Q2 2026·improvedEnterprise SCIM provisioning & Team SSO OIDC added + 2 more changes
- Week 24, 2026·packagingHuggingface reframes PRO pitch from crowd count to value promise
- Week 23, 2026·improvedHuggingface upgrades free GPUs to RTX Pro 6000 with 96GB VRAM
- Week 22, 2026·improvedHuggingface adds NVIDIA RTX PRO 6000 GPU to AWS Inference Endpoints
- Week 17, 2026·improvedHuggingface adds SCIM provisioning and OIDC SSO to enterprise plans
