SambaNova
8 articles in the TruthFoundry index. Each links out to the original.
- SambaCloud Adds Prompt Caching for MiniMax M3 to Cut Latency and Costs · SambaCloud launches automatic prompt caching for MiniMax M3, reducing latency by up to 88% and billing cached tokens at 90% discount.
- SambaNova Launches SambaRack SN50 Platform For Sovereign AI Deployments · SambaNova introduces SambaRack SN50, an energy-efficient platform for sovereign AI control.
- SambaNova SN50 Chip Enables Six-Month ROI for Neoclouds · SambaNova's SN50 chip allows neoclouds to achieve six-month ROI by adding premium AI inference capacity to existing data centers.
- AI Inference Explained: Operation, Phases and Enterprise Optimization · SambaNova explains AI inference, its phases, and enterprise scaling requirements.
- SambaNova SN50 RDU Optimizes AI Inference Bandwidth for Agents · SambaNova's SN50 RDU uses dataflow and high-speed networking to maximize bandwidth utilization for agentic AI workloads.
- MiniMax M3 Open-Weight Model Delivers 1M Context and Fastest Performance on SambaCloud · MiniMax M3 achieves top open-weight coding benchmarks and runs fastest on SambaCloud with 1M-token context.
- SambaNova and Adaption Labs Discuss Efficiency Constraints in Large AI Models · SambaNova and Adaption Labs executives argue that running large AI models efficiently is the industry's next critical challenge.
- Agentic AI Drives Demand for Premium Inference Speed Tiers · AI providers are launching premium fast inference tiers to meet latency demands of agentic workflows.