Tools
Krutrim Cloud vs Together AI
Krutrim Cloud or Together AI? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.
In inglese
In short
- Both: Runs commands, Choice of models, API, Runs models for you, Builds agents and workflows, Traces and evaluates
- Only Krutrim Cloud states: VS Code, Command line, Official SDKs
Krutrim Cloud
Krutrim Cloud provides GPU and CPU compute, storage, networking, Kubernetes, AI Pods, and AI model services.
Plans
Sandbox Compute — nano
- ₹1.56 /hr
- 0.25 vCPUs
- 0.5 GiB memory
- 20 GB disk
Sandbox Compute — small
- ₹2.91 /hr
- 0.5 vCPUs
- 1 GiB memory
- 20 GB disk
Sandbox Compute — medium
- ₹5.61 /hr
- 1 vCPU
- 2 GiB memory
- 20 GB disk
Sandbox Compute — large
- ₹11.01 /hr
- 2 vCPUs
- 4 GiB memory
- 20 GB disk
Sandbox Compute — x-Large
- ₹21.81 /hr
- 4 vCPUs
- 8 GiB memory
- 20 GB disk
Sandbox Volume storage
- ₹0.011 /GB-hour
- ₹8 /GB-month
GPU Instance — A100 80 GB × 1
- ₹189 /hr
- ₹148 /hr (Monthly)
- ₹132 /hr (6-month)
- ₹98 /hr (1-year)
- 96 GB RAM
- 80 GB GPU memory
- 24 vCPUs
GPU Instance — H100 × 1
- ₹213 /hr
- ₹198 /hr (Monthly)
- ₹186 /hr (6-month)
- ₹173 /hr (1-year)
- 200 GB RAM
- 80 GB GPU memory
- 24 vCPUs
GPU Instance — H100 × 2
- ₹426 /hr
- ₹396 /hr (Monthly)
- ₹372 /hr (6-month)
- ₹346 /hr (1-year)
- 400 GB RAM
- 160 GB GPU memory
- 48 vCPUs
GPU Instance — H100 × 4
- ₹852 /hr
- ₹792 /hr (Monthly)
- ₹744 /hr (6-month)
- ₹692 /hr (1-year)
- 800 GB RAM
- 320 GB GPU memory
- 96 vCPUs
Prices checked 2026-09-25 on the maker’s page.
Capabilities
- Runs commands — “Create a sandbox, run a command, stream the output, and you’re done:” source
- VS Code — “Connect VS Code” source
- Command line — “Python MCP CLI Terraform Live terminal” source
- Choice of models — “Krutrim Cloud's AI Studio offers access to a diverse catalogue of open-source and in-house AI models for text generation, embeddings, speech, and multimodal use cases.” source
- API — “AWS-compatible APIs for zero-friction migration.” source
- Official SDKs — “Python, Go, Java, Rust, and MCP integrations.” source
- Runs models for you — “Run distributed training on GPU clusters, deploy low-latency inference, and fine-tune models with fully managed pipelines.” source
- Builds agents and workflows — “MCP server support, AI agent frameworks, and real-time streaming inference APIs.” source
- Traces and evaluates — “The cost of running Model Evaluations and Performance Evaluations on the Krutrim platform is the same as inference — based on the number of tokens processed.” source
Together AI
Together AI provides APIs and infrastructure for running, fine-tuning, and training open AI models.
Plans
Serverless Inference
- MiniMax M3 — $0.30 per 1M tokens (Input)
- MiniMax M3 — $1.20 per 1M tokens (output)
- Kimi K3 — $3.00 per 1M tokens (Input)
- Kimi K3 — $15.00 per 1M tokens (output)
- GLM-5.3-Flash — $0.15 per 1M tokens (Input)
- GLM-5.3-Flash — $0.50 per 1M tokens (output)
- GPT Image 2 — $0.053 per image
- Wan 2.6 Image — $0.03 per image
- ByteDance Seedance 2.5 — $0.115 per video
- ByteDance Seedance 2.0 — $0.16 per video
- NVIDIA Nemotron 3 ASR Streaming 0.6B — $0.0015 per audio minute
- Whisper Large v3 — $0.0015 per audio minute
- High-performance inference as APIs
- Prices vary by model and task; the page also lists batch API prices.
Dedicated Inference — NVIDIA HGX H100
- $5.49 per gpu per hour
- $3.99 per gpu per hour
- Single-tenant GPU instances
- Guaranteed performance (no sharing)
- Support for custom models
- Autoscaling & traffic spike handling
Dedicated Inference — NVIDIA HGX B200
- $8.99 per gpu per hour
- Single-tenant GPU instances
- Guaranteed performance (no sharing)
- Support for custom models
- Autoscaling & traffic spike handling
Dedicated Inference — other hardware
Price on request
- NVIDIA HGX H200, NVIDIA HGX B300, NVIDIA GB200 NVL72, and NVIDIA GB300 NVL72
- Contact sales
GPU Clusters — On-demand
- NVIDIA HGX B200 $8.19 per GPU per hour
- NVIDIA HGX B300 $9.99 per GPU per hour
- NVIDIA HGX H100 $3.99 per GPU per hour
- NVIDIA HGX H200 $5.99 per GPU per hour
- Pay-as-you-go GPU capacity on an hourly basis
GPU Clusters — Preemptible and reserved
- NVIDIA HGX H100 Preemptible Compute $1.99 per GPU per hour
- NVIDIA HGX H100 ON-Demand $3.99 per GPU per hour
- NVIDIA HGX H100 7-30 days $3.69 per GPU per hour
- NVIDIA HGX H100 31-90 days $3.45 per GPU per hour
- NVIDIA HGX H100 91-180 days $3.19 per GPU per hour
- On-demand hourly rates and reserved capacity
- Reservation terms shown as 7-30, 31-90, and 91-180 days, and 181+ days
Code Sandbox
- Per vCPU $0.0446 per hour
- Per GiB RAM $0.0149 per hour
- Customize a deployment of VM sandboxes for large development environments
Code Interpreter
- Session (60 minutes) $0.03 per session
- Execute LLM-generated code securely using the API
Managed Storage
- Shared Filesystem $0.16 GiB/month
- High-bandwidth, parallel filesystem colocated with your compute
Prices checked 2026-09-25 on the maker’s page.
Capabilities
- Runs commands — “await client.commands.run("npm install && npm run build")” source
- Choice of models — “Scale to 30 billion tokens per model with any serverless model or private deployment.” source
- API — “High-performance inference as APIs” source
- Runs models for you — “The fastest way to run open-source models on demand.” source
- Builds agents and workflows — “Build voice agents for production” source
- Traces and evaluates — “Measure model quality” source
Latest updates
- How to train your own Jev for $17
Launched the together/Tev1-4B-experimental classifier on Together’s serverless platform.
- Canary rollouts: upgrade models in production without downtime
Dedicated inference supports staged traffic ramps, metric gates, and automatic rollback for model upgrades.
- Together AI expands fine-tuning service with more models, live metrics, and finer controls
Together Fine-Tuning added more models, live experiment tracking, Expert LoRA, early stopping, dataset previews, and pre-flight validation.