Blogs

Qwen QwQ Now on SambaNova Cloud - Try 32B Preview
Qwen QwQ Now on SambaNova Cloud - Try 32B Preview

Qwen QwQ Now on SambaNova Cloud - Try 32B Preview

December 18, 2024
Meta Llama 3.3 70B Now Available Today for Developers and Enterprises
Meta Llama 3.3 70B Now Available Today for Developers and Enterprises

Meta Llama 3.3 70B Now Available Today for Developers and Enterprises

December 11, 2024
The SambaNova Startup Accelerator: Helping AI Innovators Realize Their Vision
The SambaNova Startup Accelerator: Helping AI Innovators Realize Their Vision

The SambaNova Startup Accelerator: Helping AI Innovators Realize Their Vision

December 10, 2024
Run Qwen 2.5 32B-Coder on SambaNova Cloud - 5X GPU Speed
Run Qwen 2.5 32B-Coder on SambaNova Cloud - 5X GPU Speed

Run Qwen 2.5 32B-Coder on SambaNova Cloud - 5X GPU Speed

December 06, 2024
How Gradio Makes Building Apps on SambaNova Cloud Super Easy
How Gradio Makes Building Apps on SambaNova Cloud Super Easy

How Gradio Makes Building Apps on SambaNova Cloud Super Easy

December 05, 2024
Hugging Face Makes it Faster to Review Papers with SambaNova
Hugging Face Makes it Faster to Review Papers with SambaNova

Hugging Face Makes it Faster to Review Papers with SambaNova

December 05, 2024
Zilliz: Powering AI RAG Applications with Vector Embeddings
Zilliz: Powering AI RAG Applications with Vector Embeddings

Zilliz: Powering AI RAG Applications with Vector Embeddings

December 04, 2024
Outperforming GPT-4o with Llama 3 8B: Domain Specific Fine Tuning for RAG
Outperforming GPT-4o with Llama 3 8B: Domain Specific Fine Tuning for RAG

Outperforming GPT-4o with Llama 3 8B: Domain Specific Fine Tuning for RAG

November 20, 2024
Correcting Common AI Benchmarking Errors with AI Starter Kits
Correcting Common AI Benchmarking Errors with AI Starter Kits

Correcting Common AI Benchmarking Errors with AI Starter Kits

November 11, 2024
Accelerating Coding with SambaNova Cloud
Accelerating Coding with SambaNova Cloud

Accelerating Coding with SambaNova Cloud

October 10, 2024
Developer Tips: Creating Valuable AI
Developer Tips: Creating Valuable AI

Developer Tips: Creating Valuable AI

October 03, 2024
Replacing the Judge: Can Llama 405B Outperform GPT4 in the Court of AI?
Replacing the Judge: Can Llama 405B Outperform GPT4 in the Court of AI?

Replacing the Judge: Can Llama 405B Outperform GPT4 in the Court of AI?

September 20, 2024
Advanced AI Apps Need Fast Inference. SambaNova Cloud Delivers It
Advanced AI Apps Need Fast Inference. SambaNova Cloud Delivers It

Advanced AI Apps Need Fast Inference. SambaNova Cloud Delivers It

September 10, 2024
Why SambaNova's SN40L Chip Is the Best for Inference
Why SambaNova's SN40L Chip Is the Best for Inference

Why SambaNova's SN40L Chip Is the Best for Inference

September 10, 2024
SubgoalXL: Pushing the Boundaries of LLM in Formal Theorem Proving
SubgoalXL: Pushing the Boundaries of LLM in Formal Theorem Proving

SubgoalXL: Pushing the Boundaries of LLM in Formal Theorem Proving

September 03, 2024
SambaNova Holds Speed Record on Llama 3.1 405B - 4X faster than the rest
SambaNova Holds Speed Record on Llama 3.1 405B - 4X faster than the rest

SambaNova Holds Speed Record on Llama 3.1 405B - 4X faster than the rest

July 29, 2024
Does reduced precision hurt? A bit about losing bits.
Does reduced precision hurt? A bit about losing bits.

Does reduced precision hurt? A bit about losing bits.

June 20, 2024
Introducing Fugaku-LLM in Composition of Experts
Introducing Fugaku-LLM in Composition of Experts

Introducing Fugaku-LLM in Composition of Experts

May 13, 2024
Sovereign AI: Full-Stack Infrastructure for AI Autonomy
Sovereign AI: Full-Stack Infrastructure for AI Autonomy

Sovereign AI: Full-Stack Infrastructure for AI Autonomy

May 06, 2024
Tokens Per Second is Not All You Need
Tokens Per Second is Not All You Need

Tokens Per Second is Not All You Need

May 01, 2024
Samba-CoE v0.3: The Power of Routing ML Models at Scale
Samba-CoE v0.3: The Power of Routing ML Models at Scale

Samba-CoE v0.3: The Power of Routing ML Models at Scale

April 11, 2024
Responsible AI
Responsible AI

Responsible AI

April 10, 2024
SambaLingo hits 15,000+ downloads, now integrated with Samba-CoE-v0.2
SambaLingo hits 15,000+ downloads, now integrated with Samba-CoE-v0.2

SambaLingo hits 15,000+ downloads, now integrated with Samba-CoE-v0.2

April 08, 2024
Using Mixed Precision on RDUs
Using Mixed Precision on RDUs

Using Mixed Precision on RDUs

March 21, 2024