Skip to content
AI.info

Tools

Krutrim Cloud vs Together AI

Krutrim Cloud or Together AI? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.

In short

  • Both: Runs commands, Choice of models, API, Runs models for you, Builds agents and workflows, Traces and evaluates
  • Only Krutrim Cloud states: VS Code, Command line, Official SDKs

Krutrim Cloud

Krutrim Cloud provides GPU and CPU compute, storage, networking, Kubernetes, AI Pods, and AI model services.

Plans

Sandbox Compute — nano

  • ₹1.56 /hr
  • 0.25 vCPUs
  • 0.5 GiB memory
  • 20 GB disk

Sandbox Compute — small

  • ₹2.91 /hr
  • 0.5 vCPUs
  • 1 GiB memory
  • 20 GB disk

Sandbox Compute — medium

  • ₹5.61 /hr
  • 1 vCPU
  • 2 GiB memory
  • 20 GB disk

Sandbox Compute — large

  • ₹11.01 /hr
  • 2 vCPUs
  • 4 GiB memory
  • 20 GB disk

Sandbox Compute — x-Large

  • ₹21.81 /hr
  • 4 vCPUs
  • 8 GiB memory
  • 20 GB disk

Sandbox Volume storage

  • ₹0.011 /GB-hour
  • ₹8 /GB-month

    GPU Instance — A100 80 GB × 1

    • ₹189 /hr
    • ₹148 /hr (Monthly)
    • ₹132 /hr (6-month)
    • ₹98 /hr (1-year)
    • 96 GB RAM
    • 80 GB GPU memory
    • 24 vCPUs

    GPU Instance — H100 × 1

    • ₹213 /hr
    • ₹198 /hr (Monthly)
    • ₹186 /hr (6-month)
    • ₹173 /hr (1-year)
    • 200 GB RAM
    • 80 GB GPU memory
    • 24 vCPUs

    GPU Instance — H100 × 2

    • ₹426 /hr
    • ₹396 /hr (Monthly)
    • ₹372 /hr (6-month)
    • ₹346 /hr (1-year)
    • 400 GB RAM
    • 160 GB GPU memory
    • 48 vCPUs

    GPU Instance — H100 × 4

    • ₹852 /hr
    • ₹792 /hr (Monthly)
    • ₹744 /hr (6-month)
    • ₹692 /hr (1-year)
    • 800 GB RAM
    • 320 GB GPU memory
    • 96 vCPUs

    Prices checked 2026-09-25 on the maker’s page.

    Capabilities

    • Runs commands — “Create a sandbox, run a command, stream the output, and you’re done:” source
    • VS Code — “Connect VS Code” source
    • Command line — “Python MCP CLI Terraform Live terminal” source
    • Choice of models — “Krutrim Cloud's AI Studio offers access to a diverse catalogue of open-source and in-house AI models for text generation, embeddings, speech, and multimodal use cases.” source
    • API — “AWS-compatible APIs for zero-friction migration.” source
    • Official SDKs — “Python, Go, Java, Rust, and MCP integrations.” source
    • Runs models for you — “Run distributed training on GPU clusters, deploy low-latency inference, and fine-tune models with fully managed pipelines.” source
    • Builds agents and workflows — “MCP server support, AI agent frameworks, and real-time streaming inference APIs.” source
    • Traces and evaluates — “The cost of running Model Evaluations and Performance Evaluations on the Krutrim platform is the same as inference — based on the number of tokens processed.” source
    About Krutrim Cloud

    Together AI

    Together AI provides APIs and infrastructure for running, fine-tuning, and training open AI models.

    Plans

    Serverless Inference

    • MiniMax M3 — $0.30 per 1M tokens (Input)
    • MiniMax M3 — $1.20 per 1M tokens (output)
    • Kimi K3 — $3.00 per 1M tokens (Input)
    • Kimi K3 — $15.00 per 1M tokens (output)
    • GLM-5.3-Flash — $0.15 per 1M tokens (Input)
    • GLM-5.3-Flash — $0.50 per 1M tokens (output)
    • GPT Image 2 — $0.053 per image
    • Wan 2.6 Image — $0.03 per image
    • ByteDance Seedance 2.5 — $0.115 per video
    • ByteDance Seedance 2.0 — $0.16 per video
    • NVIDIA Nemotron 3 ASR Streaming 0.6B — $0.0015 per audio minute
    • Whisper Large v3 — $0.0015 per audio minute
    • High-performance inference as APIs
    • Prices vary by model and task; the page also lists batch API prices.

    Dedicated Inference — NVIDIA HGX H100

    • $5.49 per gpu per hour
    • $3.99 per gpu per hour
    • Single-tenant GPU instances
    • Guaranteed performance (no sharing)
    • Support for custom models
    • Autoscaling & traffic spike handling

    Dedicated Inference — NVIDIA HGX B200

    • $8.99 per gpu per hour
    • Single-tenant GPU instances
    • Guaranteed performance (no sharing)
    • Support for custom models
    • Autoscaling & traffic spike handling

    Dedicated Inference — other hardware

    Price on request

    • NVIDIA HGX H200, NVIDIA HGX B300, NVIDIA GB200 NVL72, and NVIDIA GB300 NVL72
    • Contact sales

    GPU Clusters — On-demand

    • NVIDIA HGX B200 $8.19 per GPU per hour
    • NVIDIA HGX B300 $9.99 per GPU per hour
    • NVIDIA HGX H100 $3.99 per GPU per hour
    • NVIDIA HGX H200 $5.99 per GPU per hour
    • Pay-as-you-go GPU capacity on an hourly basis

    GPU Clusters — Preemptible and reserved

    • NVIDIA HGX H100 Preemptible Compute $1.99 per GPU per hour
    • NVIDIA HGX H100 ON-Demand $3.99 per GPU per hour
    • NVIDIA HGX H100 7-30 days $3.69 per GPU per hour
    • NVIDIA HGX H100 31-90 days $3.45 per GPU per hour
    • NVIDIA HGX H100 91-180 days $3.19 per GPU per hour
    • On-demand hourly rates and reserved capacity
    • Reservation terms shown as 7-30, 31-90, and 91-180 days, and 181+ days

    Code Sandbox

    • Per vCPU $0.0446 per hour
    • Per GiB RAM $0.0149 per hour
    • Customize a deployment of VM sandboxes for large development environments

    Code Interpreter

    • Session (60 minutes) $0.03 per session
    • Execute LLM-generated code securely using the API

    Managed Storage

    • Shared Filesystem $0.16 GiB/month
    • High-bandwidth, parallel filesystem colocated with your compute

    Prices checked 2026-09-25 on the maker’s page.

    Capabilities

    • Runs commands — “await client.commands.run("npm install && npm run build")” source
    • Choice of models — “Scale to 30 billion tokens per model with any serverless model or private deployment.” source
    • API — “High-performance inference as APIs” source
    • Runs models for you — “The fastest way to run open-source models on demand.” source
    • Builds agents and workflows — “Build voice agents for production” source
    • Traces and evaluates — “Measure model quality” source

    Latest updates

    About Together AI