Tools
RunPod Pods vs Together AI
RunPod Pods or Together AI? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.
In inglese
In short
- Both: API, Runs models for you
- Only RunPod Pods states: Command line, Official SDKs
- Only Together AI states: Runs commands, Choice of models, Builds agents and workflows, Traces and evaluates
RunPod Pods
RunPod Pods are dedicated cloud GPU instances for AI development, training, inference, batch jobs, and long-running workloads.
Plans
B300
- $6.94 /hr (Community Cloud)
- $ 7.89 /hr (Secure Cloud)
- 288 GB HBM3e
- 251 GB RAM
- 32 vCPUs
H200
- $3.59 /hr (Community Cloud)
- $ 4.59 /hr (Secure Cloud)
- 141 GB VRAM
- 276 GB RAM
- 24 vCPUs
B200
- $5.98 /hr (Community Cloud)
- $ 6.79 /hr (Secure Cloud)
- 180 GB VRAM
- 283 GB RAM
- 28 vCPUs
RTX Pro 6000
- $1.69 /hr (Community Cloud)
- $ 2.09 /hr (Secure Cloud)
- 96 GB VRAM
- 188 GB RAM
- 16 vCPUs
H100 NVL
- $2.59 /hr (Community Cloud)
- $ 3.19 /hr (Secure Cloud)
- 94 GB VRAM
- 94 GB RAM
- 16 vCPUs
H100 PCIe
- $1.99 /hr (Community Cloud)
- $ 2.89 /hr (Secure Cloud)
- 80 GB VRAM
- 188 GB RAM
- 16 vCPUs
H100 SXM
- $2.69 /hr (Community Cloud)
- $ 3.49 /hr (Secure Cloud)
- 80 GB VRAM
- 125 GB RAM
- 20 vCPUs
A100 PCIe
- $1.19 /hr (Community Cloud)
- $ 1.59 /hr (Secure Cloud)
- 80 GB VRAM
- 117 GB RAM
- 8 vCPUs
A100 SXM
- $1.39 /hr (Community Cloud)
- $ 1.59 /hr
- 80 GB VRAM
- 125 GB RAM
- 16 vCPUs
Pro 6000 MIG 48GB
- $ 1.09 /hr
- 48 GB VRAM
- 62 GB RAM
- 8 vCPUs
Prices checked 2026-09-25 on the maker’s page.
Capabilities
- Command line — “CLI & SDKs. Deploy and manage directly from your terminal.” source
- API — “Full API access. Automate everything with a simple, flexible API.” source
- Official SDKs — “CLI & SDKs. Deploy and manage directly from your terminal.” source
- Runs models for you — “Run API-based AI workloads with serverless GPU endpoints.” source
Latest updates
- Global Volumes - Beta
Mount it at startup across any region; no copying data between data centers required.
- Nano Banana Edit
The Nano Banana Edit endpoint will be retired September 28, 2026. Migrate to Nano Banana 2 Edit.
- Sales tax and tax ID support
Runpod collects sales tax on credit purchases. Add a business tax ID at checkout or in your account settings.
Together AI
Together AI provides APIs and infrastructure for running, fine-tuning, and training open AI models.
Plans
Serverless Inference
- MiniMax M3 — $0.30 per 1M tokens (Input)
- MiniMax M3 — $1.20 per 1M tokens (output)
- Kimi K3 — $3.00 per 1M tokens (Input)
- Kimi K3 — $15.00 per 1M tokens (output)
- GLM-5.3-Flash — $0.15 per 1M tokens (Input)
- GLM-5.3-Flash — $0.50 per 1M tokens (output)
- GPT Image 2 — $0.053 per image
- Wan 2.6 Image — $0.03 per image
- ByteDance Seedance 2.5 — $0.115 per video
- ByteDance Seedance 2.0 — $0.16 per video
- NVIDIA Nemotron 3 ASR Streaming 0.6B — $0.0015 per audio minute
- Whisper Large v3 — $0.0015 per audio minute
- High-performance inference as APIs
- Prices vary by model and task; the page also lists batch API prices.
Dedicated Inference — NVIDIA HGX H100
- $5.49 per gpu per hour
- $3.99 per gpu per hour
- Single-tenant GPU instances
- Guaranteed performance (no sharing)
- Support for custom models
- Autoscaling & traffic spike handling
Dedicated Inference — NVIDIA HGX B200
- $8.99 per gpu per hour
- Single-tenant GPU instances
- Guaranteed performance (no sharing)
- Support for custom models
- Autoscaling & traffic spike handling
Dedicated Inference — other hardware
Price on request
- NVIDIA HGX H200, NVIDIA HGX B300, NVIDIA GB200 NVL72, and NVIDIA GB300 NVL72
- Contact sales
GPU Clusters — On-demand
- NVIDIA HGX B200 $8.19 per GPU per hour
- NVIDIA HGX B300 $9.99 per GPU per hour
- NVIDIA HGX H100 $3.99 per GPU per hour
- NVIDIA HGX H200 $5.99 per GPU per hour
- Pay-as-you-go GPU capacity on an hourly basis
GPU Clusters — Preemptible and reserved
- NVIDIA HGX H100 Preemptible Compute $1.99 per GPU per hour
- NVIDIA HGX H100 ON-Demand $3.99 per GPU per hour
- NVIDIA HGX H100 7-30 days $3.69 per GPU per hour
- NVIDIA HGX H100 31-90 days $3.45 per GPU per hour
- NVIDIA HGX H100 91-180 days $3.19 per GPU per hour
- On-demand hourly rates and reserved capacity
- Reservation terms shown as 7-30, 31-90, and 91-180 days, and 181+ days
Code Sandbox
- Per vCPU $0.0446 per hour
- Per GiB RAM $0.0149 per hour
- Customize a deployment of VM sandboxes for large development environments
Code Interpreter
- Session (60 minutes) $0.03 per session
- Execute LLM-generated code securely using the API
Managed Storage
- Shared Filesystem $0.16 GiB/month
- High-bandwidth, parallel filesystem colocated with your compute
Prices checked 2026-09-25 on the maker’s page.
Capabilities
- Runs commands — “await client.commands.run("npm install && npm run build")” source
- Choice of models — “Scale to 30 billion tokens per model with any serverless model or private deployment.” source
- API — “High-performance inference as APIs” source
- Runs models for you — “The fastest way to run open-source models on demand.” source
- Builds agents and workflows — “Build voice agents for production” source
- Traces and evaluates — “Measure model quality” source
Latest updates
- How to train your own Jev for $17
Launched the together/Tev1-4B-experimental classifier on Together’s serverless platform.
- Canary rollouts: upgrade models in production without downtime
Dedicated inference supports staged traffic ramps, metric gates, and automatic rollback for model upgrades.
- Together AI expands fine-tuning service with more models, live metrics, and finer controls
Together Fine-Tuning added more models, live experiment tracking, Expert LoRA, early stopping, dataset previews, and pre-flight validation.