AI solutions provider SambaNova has launched a turnkey AI inference data center product, dubbed "SambaManaged."

According to the company, SambaManaged can be deployed in just 90 days and enables existing data centers to begin offering AI inference services with "minimal infrastructure modification."

sambanova_Hardware_Racks1_1600x900_72dpi
SambaRacks – SambaNova

SambaManaged is a modular solution, scalable up to a 1MW "token factory" with 100 racks and 1,600 chips, or even larger. The chips in question are SambaNova's SN40L AI chips, 16 of which are in each "SambaRack".

The SN40Ls are reconfigurable dataflow units (RDUs) and are manufactured by TSMC.

The company claims to already have a "major US public company" as a customer that will use SambaManaged to run DeepSeek and other similar models.

“Data centers are struggling with power, cooling, and expertise challenges as AI demand grows,” said Abhi Ingle, chief product and strategy officer at SambaNova. “SambaManaged delivers high-performance AI with just 10kW of air-cooled power and minimal infrastructure changes — making rapid deployment simple for any data center.”

“While others talk about the future of AI, we’re delivering it — today,” added Rodrigo Liang, CEO and co-founder of SambaNova. “SambaManaged is a game-changer for organizations that want to accelerate their AI initiatives without compromising on speed, scale, or efficiency. Anywhere you have power and networking, we can bring your AI infrastructure online in record time.”

Founded in 2017 and headquartered in Palo Alto, California, SambaNova previously focused on training workloads, but pivoted earlier this year to being an AI cloud services provider. Its Suite is touted as an on-prem or cloud-based offering that allows companies to train their own generative AI models based on popular foundation models.

The company first unveiled its SambaNova Cloud in September 2024. It provides cloud-based AI inference services using the company’s SN40L AI chip, launched in September 2023 and capable of running models with up to five trillion parameters.