Vai al contenuto
AI.info

Tools

Fireworks AI vs RunPod Pods

Fireworks AI or RunPod Pods? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.

In inglese

In short

  • Both: Command line, API, Official SDKs, Runs models for you
  • Only Fireworks AI states: Choice of models

Fireworks AI

Fireworks AI hosts, fine-tunes, and serves open models through serverless and dedicated APIs.

Plans

Serverless Inference — Embeddings (up to 150M parameters)

  • $0.008 / 1M input tokens
  • Per-token pricing
  • Zero setup
  • No cold starts
  • High rate limits
  • Postpaid billing

Serverless Inference — Embeddings (150M–350M parameters)

  • $0.016 / 1M input tokens
  • Per-token pricing
  • Zero setup
  • No cold starts
  • High rate limits
  • Postpaid billing

Serverless Inference — Qwen3 8B

  • $0.1 / 1M input tokens
  • Per-token pricing
  • Zero setup
  • No cold starts
  • High rate limits
  • Postpaid billing

Managed Training — Models up to 16B parameters

  • $0.50 / 1M training tokens
  • $1.00 / 1M training tokens
  • $2.00 / 1M training tokens
  • Supervised and preference fine-tuning
  • Serve fine-tuned models for the same price as base models

Managed Training — Models 16.1B–80B

  • $3.00 / 1M training tokens
  • $6.00 / 1M training tokens
  • $12.00 / 1M training tokens
  • Supervised and preference fine-tuning
  • Serve fine-tuned models for the same price as base models

Managed Training — Models 80B–300B

  • $6.00 / 1M training tokens
  • $12.00 / 1M training tokens
  • $24.00 / 1M training tokens
  • Supervised and preference fine-tuning
  • Serve fine-tuned models for the same price as base models

Managed Training — Models over 300B

  • $10.00 / 1M training tokens
  • $20.00 / 1M training tokens
  • $40.00 / 1M training tokens
  • Supervised and preference fine-tuning
  • Serve fine-tuned models for the same price as base models

Serverless Training API — GLM 5.3

  • $4.86 / 1M Prefill
  • $0.972 / 1M Cached Prefill
  • $12.15 / 1M Sample
  • $14.58 / 1M Train
  • Shared, always-on trainer pool for LoRA training
  • No provisioning or idle cost
  • Pay only for tokens prefetched, sampled, and trained

Serverless Training API — Qwen 3.8 27B

  • $1.86 / 1M Prefill
  • $0.372 / 1M Cached Prefill
  • $5.595 / 1M Sample
  • $4.103 / 1M Train
  • Shared, always-on trainer pool for LoRA training
  • No provisioning or idle cost
  • Pay only for tokens prefetched, sampled, and trained

Serverless Training API — Kimi K3

  • $10.87 / 1M Prefill
  • $2.17 / 1M Cached Prefill
  • $27.11 / 1M Sample
  • $32.55 / 1M Train
  • Shared, always-on trainer pool for LoRA training
  • No provisioning or idle cost
  • Pay only for tokens prefetched, sampled, and trained

Prices checked 2026-09-25 on the maker’s page.

Capabilities

  • Command line — “Developers Model Library Docs CLI API Changelog” source
  • Choice of models — “Route to the best open or closed model for every task, and cut your AI coding spend 50 to 75%.” source
  • API — “Serverless. Pay per token with Priority and Fast options to meet your requirements. OpenAI and Anthropic compatible.” source
  • Official SDKs — “The Fireworks Training SDK lets us focus on our research instead of wrestling with infrastructure.” source
  • Runs models for you — “Serve the latest open models, or your own trained versions.” source

Security

  • SOC 2 Type II — “SOC 2 Type 2” source
  • SOC 2 — “SOC 2 Type 2” source
  • ISO 27001 — “ISO 27001 Certificate” source
  • ISO 42001 — “ISO 42001 Certificate” source
  • GDPR — “Compliance SOC 2 Type 2 HIPAA GDPR” source
  • HIPAA — “SOC 2 Type 2 HIPAA” source

Latest updates

About Fireworks AI

RunPod Pods

RunPod Pods are dedicated cloud GPU instances for AI development, training, inference, batch jobs, and long-running workloads.

Plans

B300

  • $6.94 /hr (Community Cloud)
  • $ 7.89 /hr (Secure Cloud)
  • 288 GB HBM3e
  • 251 GB RAM
  • 32 vCPUs

H200

  • $3.59 /hr (Community Cloud)
  • $ 4.59 /hr (Secure Cloud)
  • 141 GB VRAM
  • 276 GB RAM
  • 24 vCPUs

B200

  • $5.98 /hr (Community Cloud)
  • $ 6.79 /hr (Secure Cloud)
  • 180 GB VRAM
  • 283 GB RAM
  • 28 vCPUs

RTX Pro 6000

  • $1.69 /hr (Community Cloud)
  • $ 2.09 /hr (Secure Cloud)
  • 96 GB VRAM
  • 188 GB RAM
  • 16 vCPUs

H100 NVL

  • $2.59 /hr (Community Cloud)
  • $ 3.19 /hr (Secure Cloud)
  • 94 GB VRAM
  • 94 GB RAM
  • 16 vCPUs

H100 PCIe

  • $1.99 /hr (Community Cloud)
  • $ 2.89 /hr (Secure Cloud)
  • 80 GB VRAM
  • 188 GB RAM
  • 16 vCPUs

H100 SXM

  • $2.69 /hr (Community Cloud)
  • $ 3.49 /hr (Secure Cloud)
  • 80 GB VRAM
  • 125 GB RAM
  • 20 vCPUs

A100 PCIe

  • $1.19 /hr (Community Cloud)
  • $ 1.59 /hr (Secure Cloud)
  • 80 GB VRAM
  • 117 GB RAM
  • 8 vCPUs

A100 SXM

  • $1.39 /hr (Community Cloud)
  • $ 1.59 /hr
  • 80 GB VRAM
  • 125 GB RAM
  • 16 vCPUs

Pro 6000 MIG 48GB

  • $ 1.09 /hr
  • 48 GB VRAM
  • 62 GB RAM
  • 8 vCPUs

Prices checked 2026-09-25 on the maker’s page.

Capabilities

  • Command line — “CLI & SDKs. Deploy and manage directly from your terminal.” source
  • API — “Full API access. Automate everything with a simple, flexible API.” source
  • Official SDKs — “CLI & SDKs. Deploy and manage directly from your terminal.” source
  • Runs models for you — “Run API-based AI workloads with serverless GPU endpoints.” source

Latest updates

  • Global Volumes - Beta

    Mount it at startup across any region; no copying data between data centers required.

  • Nano Banana Edit

    The Nano Banana Edit endpoint will be retired September 28, 2026. Migrate to Nano Banana 2 Edit.

  • Sales tax and tax ID support

    Runpod collects sales tax on credit purchases. Add a business tax ID at checkout or in your account settings.

About RunPod Pods