

Fireworks
High-speed generative AI for product innovation
- Website
- fireworks.ai
- Category
- Chatbot › AI Chatbot
- Pricing
- Open source
- Platforms
- website
Pricing to seamlessly scale from idea to enterprise | Start building in seconds, self-serve. Contact us for enterprise deployments with faster speeds, lower costs, and higher rate limits. Get started Contact Us | Start building in seconds, self-serve. Contact us for enterprise deployments with faster speeds, lower costs, and higher rate limits. | Get started Contact Us | Serverless Inference Get started in seconds with per token pricing, zero setup and no cold starts See Pricing | Serverless Inference Get started in seconds with per token pricing, zero setup and no cold starts | Serverless…
About Fireworks
Join us for our inaugural conference, Forge 2026 Pricing to seamlessly scale from idea to enterprise Start building in seconds, self-serve. Contact us for enterprise deployments with faster speeds, lower costs, and higher rate limits. Get started in seconds with per token pricing, zero setup and no cold starts Customize open models with your own data with minimal setup Pay per GPU second for faster speeds, higher rate limits, and lower costs at scale Pay per token with pre-paid, usage-based billing. Simply add a payment method to buy credits, and usage is then deducted from your balance. Usage is billed on output, input, and cached tokens. Turn on Auto Reload to top up automatically when your balance runs low, or set a monthly spend limit to cap usage. To view the current pricing for our most popular serverless models across Standard, Priority, and Fast serving paths, visit our full pricing documentation .
Screenshots






Reviews
“why did Cursor rollout Composer 2 with @FireworksAI_HQ?”




