Vai al contenuto
AI.info

Tools

RunPod Pods vs Together AI

RunPod Pods or Together AI? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.

In inglese

In short

  • Both: API, Runs models for you
  • Only RunPod Pods states: Command line, Official SDKs
  • Only Together AI states: Runs commands, Choice of models, Builds agents and workflows, Traces and evaluates

RunPod Pods

RunPod Pods are dedicated cloud GPU instances for AI development, training, inference, batch jobs, and long-running workloads.

Plans

B300

  • $6.94 /hr (Community Cloud)
  • $ 7.89 /hr (Secure Cloud)
  • 288 GB HBM3e
  • 251 GB RAM
  • 32 vCPUs

H200

  • $3.59 /hr (Community Cloud)
  • $ 4.59 /hr (Secure Cloud)
  • 141 GB VRAM
  • 276 GB RAM
  • 24 vCPUs

B200

  • $5.98 /hr (Community Cloud)
  • $ 6.79 /hr (Secure Cloud)
  • 180 GB VRAM
  • 283 GB RAM
  • 28 vCPUs

RTX Pro 6000

  • $1.69 /hr (Community Cloud)
  • $ 2.09 /hr (Secure Cloud)
  • 96 GB VRAM
  • 188 GB RAM
  • 16 vCPUs

H100 NVL

  • $2.59 /hr (Community Cloud)
  • $ 3.19 /hr (Secure Cloud)
  • 94 GB VRAM
  • 94 GB RAM
  • 16 vCPUs

H100 PCIe

  • $1.99 /hr (Community Cloud)
  • $ 2.89 /hr (Secure Cloud)
  • 80 GB VRAM
  • 188 GB RAM
  • 16 vCPUs

H100 SXM

  • $2.69 /hr (Community Cloud)
  • $ 3.49 /hr (Secure Cloud)
  • 80 GB VRAM
  • 125 GB RAM
  • 20 vCPUs

A100 PCIe

  • $1.19 /hr (Community Cloud)
  • $ 1.59 /hr (Secure Cloud)
  • 80 GB VRAM
  • 117 GB RAM
  • 8 vCPUs

A100 SXM

  • $1.39 /hr (Community Cloud)
  • $ 1.59 /hr
  • 80 GB VRAM
  • 125 GB RAM
  • 16 vCPUs

Pro 6000 MIG 48GB

  • $ 1.09 /hr
  • 48 GB VRAM
  • 62 GB RAM
  • 8 vCPUs

Prices checked 2026-09-25 on the maker’s page.

Capabilities

  • Command line — “CLI & SDKs. Deploy and manage directly from your terminal.” source
  • API — “Full API access. Automate everything with a simple, flexible API.” source
  • Official SDKs — “CLI & SDKs. Deploy and manage directly from your terminal.” source
  • Runs models for you — “Run API-based AI workloads with serverless GPU endpoints.” source

Latest updates

  • Global Volumes - Beta

    Mount it at startup across any region; no copying data between data centers required.

  • Nano Banana Edit

    The Nano Banana Edit endpoint will be retired September 28, 2026. Migrate to Nano Banana 2 Edit.

  • Sales tax and tax ID support

    Runpod collects sales tax on credit purchases. Add a business tax ID at checkout or in your account settings.

About RunPod Pods

Together AI

Together AI provides APIs and infrastructure for running, fine-tuning, and training open AI models.

Plans

Serverless Inference

  • MiniMax M3 — $0.30 per 1M tokens (Input)
  • MiniMax M3 — $1.20 per 1M tokens (output)
  • Kimi K3 — $3.00 per 1M tokens (Input)
  • Kimi K3 — $15.00 per 1M tokens (output)
  • GLM-5.3-Flash — $0.15 per 1M tokens (Input)
  • GLM-5.3-Flash — $0.50 per 1M tokens (output)
  • GPT Image 2 — $0.053 per image
  • Wan 2.6 Image — $0.03 per image
  • ByteDance Seedance 2.5 — $0.115 per video
  • ByteDance Seedance 2.0 — $0.16 per video
  • NVIDIA Nemotron 3 ASR Streaming 0.6B — $0.0015 per audio minute
  • Whisper Large v3 — $0.0015 per audio minute
  • High-performance inference as APIs
  • Prices vary by model and task; the page also lists batch API prices.

Dedicated Inference — NVIDIA HGX H100

  • $5.49 per gpu per hour
  • $3.99 per gpu per hour
  • Single-tenant GPU instances
  • Guaranteed performance (no sharing)
  • Support for custom models
  • Autoscaling & traffic spike handling

Dedicated Inference — NVIDIA HGX B200

  • $8.99 per gpu per hour
  • Single-tenant GPU instances
  • Guaranteed performance (no sharing)
  • Support for custom models
  • Autoscaling & traffic spike handling

Dedicated Inference — other hardware

Price on request

  • NVIDIA HGX H200, NVIDIA HGX B300, NVIDIA GB200 NVL72, and NVIDIA GB300 NVL72
  • Contact sales

GPU Clusters — On-demand

  • NVIDIA HGX B200 $8.19 per GPU per hour
  • NVIDIA HGX B300 $9.99 per GPU per hour
  • NVIDIA HGX H100 $3.99 per GPU per hour
  • NVIDIA HGX H200 $5.99 per GPU per hour
  • Pay-as-you-go GPU capacity on an hourly basis

GPU Clusters — Preemptible and reserved

  • NVIDIA HGX H100 Preemptible Compute $1.99 per GPU per hour
  • NVIDIA HGX H100 ON-Demand $3.99 per GPU per hour
  • NVIDIA HGX H100 7-30 days $3.69 per GPU per hour
  • NVIDIA HGX H100 31-90 days $3.45 per GPU per hour
  • NVIDIA HGX H100 91-180 days $3.19 per GPU per hour
  • On-demand hourly rates and reserved capacity
  • Reservation terms shown as 7-30, 31-90, and 91-180 days, and 181+ days

Code Sandbox

  • Per vCPU $0.0446 per hour
  • Per GiB RAM $0.0149 per hour
  • Customize a deployment of VM sandboxes for large development environments

Code Interpreter

  • Session (60 minutes) $0.03 per session
  • Execute LLM-generated code securely using the API

Managed Storage

  • Shared Filesystem $0.16 GiB/month
  • High-bandwidth, parallel filesystem colocated with your compute

Prices checked 2026-09-25 on the maker’s page.

Capabilities

  • Runs commands — “await client.commands.run("npm install && npm run build")” source
  • Choice of models — “Scale to 30 billion tokens per model with any serverless model or private deployment.” source
  • API — “High-performance inference as APIs” source
  • Runs models for you — “The fastest way to run open-source models on demand.” source
  • Builds agents and workflows — “Build voice agents for production” source
  • Traces and evaluates — “Measure model quality” source

Latest updates

About Together AI