Tools
Fireworks AI vs RunPod Pods
Fireworks AI or RunPod Pods? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.
In inglese
In short
- Both: Command line, API, Official SDKs, Runs models for you
- Only Fireworks AI states: Choice of models
Fireworks AI
Fireworks AI hosts, fine-tunes, and serves open models through serverless and dedicated APIs.
Plans
Serverless Inference — Embeddings (up to 150M parameters)
- $0.008 / 1M input tokens
- Per-token pricing
- Zero setup
- No cold starts
- High rate limits
- Postpaid billing
Serverless Inference — Embeddings (150M–350M parameters)
- $0.016 / 1M input tokens
- Per-token pricing
- Zero setup
- No cold starts
- High rate limits
- Postpaid billing
Serverless Inference — Qwen3 8B
- $0.1 / 1M input tokens
- Per-token pricing
- Zero setup
- No cold starts
- High rate limits
- Postpaid billing
Managed Training — Models up to 16B parameters
- $0.50 / 1M training tokens
- $1.00 / 1M training tokens
- $2.00 / 1M training tokens
- Supervised and preference fine-tuning
- Serve fine-tuned models for the same price as base models
Managed Training — Models 16.1B–80B
- $3.00 / 1M training tokens
- $6.00 / 1M training tokens
- $12.00 / 1M training tokens
- Supervised and preference fine-tuning
- Serve fine-tuned models for the same price as base models
Managed Training — Models 80B–300B
- $6.00 / 1M training tokens
- $12.00 / 1M training tokens
- $24.00 / 1M training tokens
- Supervised and preference fine-tuning
- Serve fine-tuned models for the same price as base models
Managed Training — Models over 300B
- $10.00 / 1M training tokens
- $20.00 / 1M training tokens
- $40.00 / 1M training tokens
- Supervised and preference fine-tuning
- Serve fine-tuned models for the same price as base models
Serverless Training API — GLM 5.3
- $4.86 / 1M Prefill
- $0.972 / 1M Cached Prefill
- $12.15 / 1M Sample
- $14.58 / 1M Train
- Shared, always-on trainer pool for LoRA training
- No provisioning or idle cost
- Pay only for tokens prefetched, sampled, and trained
Serverless Training API — Qwen 3.8 27B
- $1.86 / 1M Prefill
- $0.372 / 1M Cached Prefill
- $5.595 / 1M Sample
- $4.103 / 1M Train
- Shared, always-on trainer pool for LoRA training
- No provisioning or idle cost
- Pay only for tokens prefetched, sampled, and trained
Serverless Training API — Kimi K3
- $10.87 / 1M Prefill
- $2.17 / 1M Cached Prefill
- $27.11 / 1M Sample
- $32.55 / 1M Train
- Shared, always-on trainer pool for LoRA training
- No provisioning or idle cost
- Pay only for tokens prefetched, sampled, and trained
Prices checked 2026-09-25 on the maker’s page.
Capabilities
- Command line — “Developers Model Library Docs CLI API Changelog” source
- Choice of models — “Route to the best open or closed model for every task, and cut your AI coding spend 50 to 75%.” source
- API — “Serverless. Pay per token with Priority and Fast options to meet your requirements. OpenAI and Anthropic compatible.” source
- Official SDKs — “The Fireworks Training SDK lets us focus on our research instead of wrestling with infrastructure.” source
- Runs models for you — “Serve the latest open models, or your own trained versions.” source
Security
- SOC 2 Type II — “SOC 2 Type 2” source
- SOC 2 — “SOC 2 Type 2” source
- ISO 27001 — “ISO 27001 Certificate” source
- ISO 42001 — “ISO 42001 Certificate” source
- GDPR — “Compliance SOC 2 Type 2 HIPAA GDPR” source
- HIPAA — “SOC 2 Type 2 HIPAA” source
Latest updates
- Serverless pricing update: DeepSeek V4.1 Flash
Serverless pricing for DeepSeek V4.1 Flash changes; dedicated deployment and Reserved Throughput pricing is unaffected.
- New deployment creation flags: deploymentShape: "default" and acceptShapelessRisk
Create Deployment adds deploymentShape: "default" to pick a validated deployment shape and acceptShapelessRisk to create without a shape.
- Upcoming Serverless deprecation: older DeepSeek, GLM, Muse, and Kimi models
Several older Serverless models will be decommissioned on September 25, 2026; migrate to a recommended replacement before then.
RunPod Pods
RunPod Pods are dedicated cloud GPU instances for AI development, training, inference, batch jobs, and long-running workloads.
Plans
B300
- $6.94 /hr (Community Cloud)
- $ 7.89 /hr (Secure Cloud)
- 288 GB HBM3e
- 251 GB RAM
- 32 vCPUs
H200
- $3.59 /hr (Community Cloud)
- $ 4.59 /hr (Secure Cloud)
- 141 GB VRAM
- 276 GB RAM
- 24 vCPUs
B200
- $5.98 /hr (Community Cloud)
- $ 6.79 /hr (Secure Cloud)
- 180 GB VRAM
- 283 GB RAM
- 28 vCPUs
RTX Pro 6000
- $1.69 /hr (Community Cloud)
- $ 2.09 /hr (Secure Cloud)
- 96 GB VRAM
- 188 GB RAM
- 16 vCPUs
H100 NVL
- $2.59 /hr (Community Cloud)
- $ 3.19 /hr (Secure Cloud)
- 94 GB VRAM
- 94 GB RAM
- 16 vCPUs
H100 PCIe
- $1.99 /hr (Community Cloud)
- $ 2.89 /hr (Secure Cloud)
- 80 GB VRAM
- 188 GB RAM
- 16 vCPUs
H100 SXM
- $2.69 /hr (Community Cloud)
- $ 3.49 /hr (Secure Cloud)
- 80 GB VRAM
- 125 GB RAM
- 20 vCPUs
A100 PCIe
- $1.19 /hr (Community Cloud)
- $ 1.59 /hr (Secure Cloud)
- 80 GB VRAM
- 117 GB RAM
- 8 vCPUs
A100 SXM
- $1.39 /hr (Community Cloud)
- $ 1.59 /hr
- 80 GB VRAM
- 125 GB RAM
- 16 vCPUs
Pro 6000 MIG 48GB
- $ 1.09 /hr
- 48 GB VRAM
- 62 GB RAM
- 8 vCPUs
Prices checked 2026-09-25 on the maker’s page.
Capabilities
- Command line — “CLI & SDKs. Deploy and manage directly from your terminal.” source
- API — “Full API access. Automate everything with a simple, flexible API.” source
- Official SDKs — “CLI & SDKs. Deploy and manage directly from your terminal.” source
- Runs models for you — “Run API-based AI workloads with serverless GPU endpoints.” source
Latest updates
- Global Volumes - Beta
Mount it at startup across any region; no copying data between data centers required.
- Nano Banana Edit
The Nano Banana Edit endpoint will be retired September 28, 2026. Migrate to Nano Banana 2 Edit.
- Sales tax and tax ID support
Runpod collects sales tax on credit purchases. Add a business tax ID at checkout or in your account settings.