Qwen2.5-VL-32B-Instruct

Description
Pricing Details
Hugging Face is a platform specializing in the development and sharing of artificial intelligence and machine learning models. It provides an integrated environment for developers and researchers to host models, datasets, and AI applications, and to run inference operations. The platform offers plans for individuals, teams, and organizations, as well as storage, GPU computing, and cloud-based model deployment services. Hugging Face operates on a freemium model with paid plans; the platform can be used for free to access thousands of models, datasets, and public spaces, while paid plans offer additional benefits such as increased private storage, computing priority, private space deployment, collaboration tools, and enterprise capabilities. Free Plan: Allows you to use Hugging Face Hub to explore models, download public models and datasets, create projects and share work with the community, and use some free computing resources such as ZeroGPU and public spaces, with limits on private storage and advanced usage. Hugging Face PRO: This plan costs $9 per month and is designed for individual users who need greater capabilities for developing AI projects. The plan offers a 10-fold increase in private storage, a 2-fold increase in public storage, 20 times the inference credits, higher priority in ZeroGPU queues, the ability to host ZeroGPU, Gradio, and Docker spaces, as well as development mode for spaces, a personal blog, and private data viewing. Hugging Face Team: The cost is $20 per month per user, and it is designed for teams and startups. It offers single sign-on (SSO) support, control over data storage location, audit logs, permission management via Resource Groups, repository usage analytics, advanced security policies, centralized code control, and the ability to create Gradio and Docker spaces with advanced computing options. All team members also receive the ZeroGPU and Inference Providers benefits included in the PRO plan. Hugging Face Enterprise: Priced at $50 per user per month, this plan is designed for large enterprises requiring advanced infrastructure. It includes all the benefits of the Team plan, plus higher limits for storage, bandwidth, and API usage rates; automated user management via SCIM; advanced security and control tools, custom billing with annual contracts, support for legal and compliance operations, and dedicated enterprise support. Hugging Face also offers data-volume-based storage services, with Hub storage starting at a base price of approximately $12 per terabyte per month for public repositories and $18 per terabyte for private repositories, with discounts available for storage volumes exceeding 500 terabytes. Cloud computing services and Spaces start with free usage via CPU Basic and ZeroGPU, while paid GPU resources are available depending on the processor type. Prices for some GPU units, such as the NVIDIA T4, start at approximately $0.40 to $0.50 per hour, the NVIDIA L4 at around $0.80 per hour, and the NVIDIA A100 at around $2.50 per hour, while advanced resources such as the NVIDIA H200 and B200 command higher prices depending on the number of cards used. Hugging Face also offers an Inference Endpoints service for running AI models on dedicated servers starting at about $0.033 per hour, with support for CPUs, GPUs such as the T4, L4, A100, and H100, and advanced options for production applications that require consistent performance and scalability.