Optimizing the compute cluster at scale
Combining Cerebras wafer-scale compute with Gimlet's inference cloud to deliver up to 3,000 tokens per second
A rack scale proving ground that extends Giga Computing from server manufacturer to full AI infrastructure partner
Brings validated partner integrations into one curated place - Spans the AI loop, anchored by years-deep collaborations with Reflection, Vast Data, ClickHouse, CrowdStrike and more
New high-efficiency power supply targeting high demanding environments
New service embeds CTERA engineers inside customer environments to prepare governed data, build AI workflows around real business processes, and enable operational success
v12 delivers multi-tenancy, API-first automation, context-aware tiering and 2x+ performance per watt, now qualified on Supermicro building blocks
Continuing to address both on-premise and cloud environments
Doubling production capacity to meet the new challenges of AI
Former OpenAI CEO of AGI Deployment appointed as independent director