GitHub hosts HunyuanVideo, Tencent's open-source framework for large-scale video generation models, enabling AI-driven video creation.
CostPerPrompt
About CostPerPrompt
CostPerPrompt provides real-time per-token pricing for over 300 AI model variants from leading providers including OpenAI, Anthropic, DeepSeek, Google, Meta, Mistral, Alibaba and others. The platform displays both input and output token costs, as well as cached-context pricing, enabling users to compare models and assess their financial impact based on usage patterns. Pricing data is updated automatically, with the most recent refresh timestamp clearly indicated to ensure accuracy. A dedicated calculator allows users to model monthly expenses by inputting daily request volumes and token usage, while also supporting batching and caching features to refine estimates for production environments. Beyond model pricing, the service includes GPU rental pricing tables that compare on-demand rates for NVIDIA H100, A100, RTX 4090, and L40S across multiple cloud providers such as Vast.ai, RunPod, Lambda, Cudo Compute, CoreWeave, Google Cloud, and Azure. This dual focus on AI model and GPU costs makes it a practical resource for developers, researchers, and businesses evaluating infrastructure expenses for AI workloads.
Key features
- Live per-token pricing
- Cached-context pricing
- Batch request calculator
- Monthly cost estimation
- GPU rental pricing tables
- Vendor comparison across 300+ models
- Price refresh timestamp
- Input/output token cost breakdown
Use cases
- Comparing LLM API costs across vendors
- Estimating monthly cloud spend for AI workloads
- Evaluating GPU rental options for training and inference
Pros
- Aggregates live pricing from multiple vendors
- Supports more than 300 model variants
- Includes cached-context and batching options in calculator
- Provides GPU rental pricing comparisons
- Displays last update timestamp for transparency
Cons
- No free tier or public API access
- GPU pricing limited to four NVIDIA models
- Prices vary by region and availability
Frequently asked questions about CostPerPrompt
What is CostPerPrompt and what does it do?
CostPerPrompt is a service that provides live per-token pricing for over 300 AI model variants from major vendors, including OpenAI, Anthropic, DeepSeek, Google, Meta, Mistral, Alibaba, and others. It aggregates input and output token costs, as well as cached-context pricing, to help users compare models and estimate monthly expenses based on request volumes and token usage.
Who should use CostPerPrompt?
CostPerPrompt is designed for developers, data scientists, product managers, and businesses that need to estimate or optimize AI model costs for production workloads, research, or budgeting purposes. It is particularly useful for those evaluating multiple model options or planning scalable deployments.
How does CostPerPrompt calculate monthly costs?
The service includes a full calculator that allows users to input daily request volumes, token usage per request, and batching or caching settings to generate detailed monthly cost estimates. Prices are refreshed automatically, and the last update timestamp is displayed to ensure accuracy.
Does CostPerPrompt provide pricing for GPU rentals?
Yes, CostPerPrompt includes GPU rental pricing tables that compare on-demand rates for NVIDIA GPUs such as H100, A100, RTX 4090, and L40S across providers like Vast.ai, RunPod, Lambda, Cudo Compute, CoreWeave, Google Cloud, and Azure.
Can I compare models across different vendors on CostPerPrompt?
Yes, the platform allows users to compare per-token pricing, cached-context costs, and other metrics across models from multiple vendors in a single interface, making it easier to evaluate alternatives and make informed decisions.
How often are the prices updated on CostPerPrompt?
Prices are refreshed automatically, and the last update timestamp is displayed on the site to ensure users have access to the most current pricing information available.
CostPerPrompt Website Engagement
Last Update: 9 days ago