Blackbox's collaboration with SambaNova means developers get tools that work as fast as they think.
The Challenge:
Scaling smarter, faster agents
Blackbox.ai is redefining coding with autonomous agents that empower developers to build custom applications effortlessly. Their flagship agent, CyberCoder, helps developers edit multiple files, create applications from scratch, and auto-document code—all in real-time. But as Blackbox's user base skyrocketed to over 10 million monthly active users, including Fortune 500 companies, they faced a new challenge: How to keep up with the growing performance demands for faster, smarter coding experiences?
The Solution:
Low latency, high performance with SambaNova
Enter SambaCloud. Known for delivering lightning-fast inference across open-source models like Llama, DeepSeek, and Qwen, SambaCloud became the performance engine Blackbox needed.
"We were drawn to SambaNova after seeing Andrew Ng’s post about their cloud platform," Rizk explained. "We’re using Llama models from 8 B to 405 B for everything from simple edits to complex coding challenges. The low latency and high performance are game-changers."
3-4X
Faster model completions
100s
Requests per second
>10M
Monthly active users
“We’ve experimented extensively with Meta's Llama models since their release. They’re great for diverse use cases, and we fine-tune them to excel at specific tasks. But scaling that performance efficiently became a hurdle."
— Robert Rizk, Blackbox.ai Co-founder and CEO
Blackbox CyberCoder powered by SambaNova
FAQs
SambaCloud is SambaNova's fully managed cloud inference service, powered by the Reconfigurable Dataflow Unit (RDU). Ricoh uses it to host and serve custom AI models fine-tuned for Japanese business documents and industry-specific workflows.
SambaCloud delivers 10x faster inference than Ricoh's existing GPU infrastructure, achieving over 700 tokens per second on 70B-class models compared to tens of tokens per second previously.
Yes. SambaCloud is compatible with a wide range of open-weight models including Llama, Qwen, and Gemma, and supports fine-tuned variants, preserving accuracy and cultural relevance for Ricoh's Japanese business use cases.
SambaCloud's speed enables complex agentic workflows involving multiple model calls to complete in around ten seconds, compared to approximately one minute on previous infrastructure.


-1.png?width=380&height=220&name=2025-04_SambaNova+Llama4_1600x900_v1.0%20(1)-1.png)