Tools
Baseten vs Replicate
Baseten or Replicate? Their plans monthly and yearly, the capabilities their makers state, security and latest updates, side by side — read from the makers' own pages.
In short
- Both: Choice of models, API, Runs models for you
- Only Baseten states: Self-hosted
- Only Replicate states: Official SDKs, Traces and evaluates
Baseten
Baseten provides managed infrastructure to deploy, serve, train, and distribute open-source, custom, and fine-tuned AI models.
Plans
Basic
Free
- Dedicated deployments
- Model APIs
- Training
- Fast cold starts
- SOC 2 Type II and HIPAA compliant
- Email and in-app chat support
Dedicated Deployments — T4
- $0.01052 per minute
- 16 GiB VM
Dedicated Deployments — L4
- $0.01414 per minute
- 24 GiB VRAM
Dedicated Deployments — A10G
- $0.02012 per minute
- 24 GiB VM
Dedicated Deployments — A100
- $0.06667 per minute
- 80 GiB VRAM
Dedicated Deployments — H100 MIG
- $0.0625 per minute
- 40 GiB VRAM
Dedicated Deployments — H100
- $0.10833 per minute
- 80 GiB VRAM
Dedicated Deployments — B200
- $0.16633 per minute
- 180 GiB VRAM
Prices checked 2026-09-24 on the maker’s page.
Capabilities
- Choice of models — “You can deploy open source and custom models on Baseten.” source
- Self-hosted — “Yes, you can self-host Baseten in order to manage security and use your own cloud commitments.” source
- API — “Model APIs” source
- Runs models for you — “Instant access to pre-optimized models running on the Baseten Inference Stack.” source
Security
- SOC 2 Type II — “Baseten Labs, Inc. SOC 2 Type 2 Report 5.31.26.pdf” source
- SOC 2 — “Baseten Labs, Inc. SOC 2 Type 2 Report 5.31.26.pdf” source
- ISO 27001 — “Baseten Labs, Inc. ISO 27001 Certificate.pdf” source
- GDPR — “SOC 2 ISO 27001:2022 HIPAA CCPA GDPR PCI DSS - SAQ D” source
- HIPAA — “Baseten Labs, Inc. HIPAA Report 5.31.26.pdf” source
Latest updates
- Web search with Baseten Hosted Tools
Baseten Hosted Tools now brings server-side web search to Model APIs through Baseten Grounded Inference.
- Baseten CLI 1.0.0 (1.0.0)
The Baseten CLI is now generally available. Version 1.0.0 marks the command surface as stable.
- Model API Deprecation (GLM 4.7, Kimi K2.7, Kimi K2.6, Inkling, Inkling Small, DeepSeek v4 Pro)
GLM 4.7, Kimi K2.7, Kimi K2.6, Inkling, Inkling Small, and DeepSeek v4 Pro will be deprecated at 5pm PT September 25th.
Replicate
Replicate lets developers run, deploy, and scale machine learning models through a web interface or cloud API.
Plans
black-forest-labs / flux-1.1-pro
- $0.04 / output image
- Text-to-image model
- Excellent image quality, prompt adherence, and output diversity
black-forest-labs / flux-dev
- $0.025 / output image
- 12 billion parameter rectified flow transformer
- Generates images from text descriptions
black-forest-labs / flux-schnell
- $3.00 / thousand output images
- Fast image generation
- Tailored for local development and personal use
deepseek-ai / deepseek-r1
- $0.01 / thousand output tokens
- $3.75 / million input tokens
- Reasoning model trained with reinforcement learning
- On par with OpenAI o1
ideogram-ai / ideogram-v3-quality
- $0.09 / output image
- Highest quality Ideogram v3 model
- Creates images with realism, creative designs, and consistent styles
recraft-ai / recraft-v3
- $0.04 / output image
- Text-to-image model
- Generates long texts and images in a wide list of styles
wavespeedai / wan-2.1-i2v-480p
- $0.09 / second of output video
- Accelerated inference for Wan 2.1 14B image to video
wavespeedai / wan-2.1-i2v-720p
- $0.25 / second of output video
- Accelerated inference for Wan 2.1 14B image to video with high resolution
CPU (Small)
- $ 0.000025 /sec
- $ 0.09 /hr
- cpu-small
- CPU 1x
- RAM 2GB
CPU
- $ 0.36 /hr
- cpu
- CPU 4x
- RAM 8GB
Prices checked 2026-09-25 on the maker’s page.
Capabilities
- Choice of models — “Compare models in the Playground” source
- API — “Replicate - Run AI with an API” source
- Official SDKs — “import Replicate from "replicate"” source
- Runs models for you — “Cog takes care of generating an API server and deploying it on a big cluster in the cloud.” source
- Traces and evaluates — “Metrics let you keep an eye on how your models are performing, and logs let you zoom in on particular predictions to debug how your model is behaving.” source
Latest updates
- Agent skills for Replicate
Replicate now publishes agent skills, markdown instruction files that give coding assistants knowledge about working with AI models on Replicate.
- Fallback model for Nano Banana Pro
Nano Banana Pro can fall back to Seedream 5.0 lite when Google’s API is at capacity.
- MCP server auto-discovery
Replicate’s MCP server can now be discovered automatically through the official MCP Registry.