tools
Helicone
Helicone routes, monitors, debugs, and analyzes AI applications through an open-source gateway and observability platform.
In inglese
Helicone helps AI engineers and teams log LLM requests, trace multi-step workflows, monitor costs and performance, manage prompts, run evaluations, and route traffic across model providers.
It offers an OpenAI-compatible AI Gateway with routing, caching, rate limits, automatic failover, and observability. A free Hobby plan is available; Pro and Team plans add usage-based charges, while Enterprise pricing is custom. Helicone was acquired by Mintlify on March 3, 2026, and its services remain live in maintenance mode.
Features
- Route requests across 100+ models and multiple AI providers
- Trace and debug multi-step LLM and agent workflows
- Monitor requests, costs, latency, users, and sessions
- Manage, version, test, and deploy prompts without code changes
- Run experiments, datasets, playground tests, and evaluations
- Use caching, rate limits, load balancing, and automatic failover
- Self-host the observability platform with Docker
- Open-source AI Gateway and observability platform
Use cases
- Monitor production LLM requests and identify errors
- Compare model costs, latency, and performance across providers
- Test prompt changes against production data before release
- Route AI traffic through a unified OpenAI-compatible API
- Debug multi-step agents with sessions and request traces
- Deploy observability inside an organization’s infrastructure
Pros
Cons
Latest updates
- Claude Sonnet 4 and Sonnet 4.5 now support 1M context window
Sonnet 4 and Sonnet 4.5 models on the AI Gateway now support a 1M token context window by default across Anthropic API, AWS Bedrock, and Google Vertex AI.
- Control Reasoning Effort in Playground and better feedback on thinking models
Added the reasoning effort parameter, a minimal option, existing reasoning levels, and visual reasoning display when available.
- OpenAI GPT-5 Models Pricing and Playground Support
Added pricing for GPT-5, GPT-5-Mini, GPT-5-Nano, and GPT-5-Chat-Latest models across multiple providers.
- OpenAI GPT OSS Models Pricing
Cost tracking is available for GPT-OSS models across Fireworks, Groq, and OpenRouter, with automatic cost calculation and dashboard integration.
- Claude Opus 4.1 Pricing
Cost tracking is available for Claude Opus 4.1 models, with automatic cost calculation and dashboard integration.
Capabilities
- Choice of models — “Request Routing : Route by model, cost, or custom rules” source
- Self-hosted — “Now you can deploy our powerful observability platform directly within your own infrastructure with a single Docker command.” source
- API — “With this setup, any calls to the OpenAI Responses API will be automatically logged and monitored by Helicone.” source
- Official SDKs — “We’re thrilled to announce that we now have a Go SDK for Helicone’s Helpers Package .” source
- Traces and evaluates — “Built-in Observability : Full integration with Helicone’s analytics” source
Get it
Pricing
- Starting price
- $79/mo
- Prices checked
- 2026-09-25
Hobby
Free
- 10,000 free requests
- 1 GB storage
- 1 seat, 1 organization
Pro
- $79 per month
- Everything in Hobby
- Unlimited seats
- Alerts & reports
- HQL (Query Language)
Team
- $799 per month
- Everything in Pro
- 5 organizations
- SOC-2 & HIPAA compliance
- Dedicated Slack channel
Enterprise
Price on request
- Everything in Team
- Custom MSA
- SAML SSO
- On-prem deployment
- Bulk cloud discounts