tools
Together AI - Open-Source Model API & Fine-Tuning Platform | AI.info
Together AI provides API access to 100+ open-source models including Llama 4 and DeepSeek. OpenAI-compatible with fine-tuning and dedicated endpoints.

Together AI is a developer platform for serverless and dedicated model inference, batch processing, fine-tuning, GPU clusters, sandboxes, and managed storage. It supports text, image, audio, video, transcription, embeddings, reranking, and moderation workloads.
Developers and AI teams use it to build applications, agents, media features, and production model deployments. Pricing is usage-based rather than subscription-based; dedicated infrastructure, compute, storage, and fine-tuning cost extra, and the platform requires at least $5 in purchased credits.
Features
- Run serverless inference for text, image, audio, video, transcription, and embedding models
- Process large workloads asynchronously with Batch Inference
- Deploy models on dedicated, single-tenant GPU infrastructure
- Fine-tune open models with supervised fine-tuning and preference optimization
- Run GPU clusters for training and inference workloads
- Create secure code sandboxes and code interpreter sessions through an API
- Use managed storage with parallel filesystems and no egress fees
Use cases
- Build production AI applications with model APIs
- Deploy coding, research, and customer-support agents
- Fine-tune open models for domain-specific behavior
- Process batch text, image, audio, or video workloads
- Run training and inference jobs on managed GPU clusters
- Add voice, media generation, or multimodal features to products
Pros
Cons
Pricing
- Starting price
- $0.0015
- Pricing checked
- 2026-09-19
Serverless Inference
$0.30
- MiniMax M3 input price per 1M tokens; output price is $1.20
Dedicated Inference
$3.99
- NVIDIA HGX H100 on-demand price per GPU per hour
GPU Clusters
$1.99
- NVIDIA HGX H100 preemptible compute price per GPU per hour
Sandbox
$0.03
- Code Interpreter session lasting 60 minutes
Storage
$0.16
- Shared filesystem price per GiB/month
Fine-Tuning
$0.48
- Supervised fine-tuning with LoRA for models up to 16B per 1M tokens