Vai al contenuto
AI.info

tools

SambaNova DataScale

SambaNova DataScale is an on-premises or hosted AI system for training and running large models with dataflow processors.

SambaNova DataScale

In inglese

SambaNova DataScale is an integrated hardware and software system for AI training and inference. The DataScale SN40L uses Reconfigurable Dataflow Units, a dataflow architecture, and three-tier memory to run multiple models and switch between them quickly.

It is aimed at enterprises, government agencies, and research organizations. Deployments can be on-premises or in hosted data centers. Hardware pricing is not published on the reviewed pricing page; that page lists SambaNova Cloud token usage instead.

Features

  • Runs AI model training and inference
  • Uses SN40L Reconfigurable Dataflow Units
  • Supports on-premises and hosted data-center deployment
  • Runs multiple models on one system
  • Switches between models in microseconds
  • Uses three-tier memory with SRAM, HBM, and DDR
  • Supports custom and chained models
  • Runs Red Hat Enterprise Linux

Use cases

  • Train and deploy large language models
  • Run generative and agentic AI workloads
  • Analyze high-resolution computer-vision images
  • Support scientific research and AI-for-science workloads
  • Deploy private AI infrastructure for enterprises
  • Run multiple models from a shared system

Pros

    Cons

      Latest updates

      Capabilities

      • Choice of models — “SambaStack switches between multiple frontier-scale models, enabling complex agentic AI workflows to execute end-to-end on one node.” source
      • Self-hosted — “SambaRack™ is a state-of-the-art system that can be set up easily in data centers to run Al inference workloads.” source
      • API — “SambaNova provides simple-to-integrate APIs for Al inference, making it easy to onboard applications.” source
      • Runs models for you — “SambaNova provides simple-to-integrate APIs for Al inference, making it easy to onboard applications.” source
      • Builds agents and workflows — “Build or use pre-built AI agents — all with business-aware intelligence.” source

      Get it

      Pricing

      Prices checked
      2026-09-25

      MiniMax-M2.7

      • 0.06 USD per 1M cached input tokens
      • 0.60 USD per 1M input tokens
      • 2.40 USD per 1M output tokens

        DeepSeek-V3.1

        • 3 USD per 1M input tokens
        • 4.50 USD per 1M output tokens

          DeepSeek-V3.2

          • 3 USD per 1M input tokens
          • 4.50 USD per 1M output tokens

            gemma-4-31B-it

            • 0.38 USD per 1M input tokens
            • 1.15 USD per 1M output tokens

              gpt-oss-120b

              • 0.22 USD per 1M input tokens
              • 0.59 USD per 1M output tokens

                Meta-Llama-3.3-70B-Instruct

                • 0.60 USD per 1M input tokens
                • 1.20 USD per 1M output tokens

                  MiniMax-M3

                  • 0.60 USD per 1M input tokens
                  • 2.40 USD per 1M output tokens
                    Official website