# Google Cloud C4 Brings a 70% TCO improvement on GPT OSS with Intel and Hugging Face

DevFeed: [Google Cloud C4 Brings a 70% TCO improvement on GPT OSS with Intel and Hugging Face](<https://devfeed.tech/articles/google-cloud-c4-brings-a-70-tco-improvement-on-gpt-oss-with-intel-and-hugging-face-7220.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/gpt-oss-on-intel-xeon>)

Author: Jiqing.Feng; Matrix Yao; Ke Ding; Ilyas Moutawwakil

Published: 2025-10-16T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [moe](<https://devfeed.tech/topics/moe.md>), [cpu](<https://devfeed.tech/topics/cpu.md>), [AI, ML & Data Engineering](<https://devfeed.tech/topics/ai-ml-data-engineering.md>)

Tags: [cpu](<https://devfeed.tech/tags/cpu.md>), [google-cloud](<https://devfeed.tech/tags/google-cloud.md>), [gpt-oss](<https://devfeed.tech/tags/gpt-oss.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference](<https://devfeed.tech/tags/inference.md>), [intel](<https://devfeed.tech/tags/intel.md>), [llm](<https://devfeed.tech/tags/llm.md>), [mixture-of-experts-moe](<https://devfeed.tech/tags/mixture-of-experts-moe.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [performance](<https://devfeed.tech/tags/performance.md>)

## AI overview

The article benchmarks GPT OSS mixture-of-experts text generation on Google Cloud Intel Xeon virtual machines. It describes an expert-execution optimization and reports throughput, latency, and total-cost-of-ownership comparisons between Xeon generations.

## Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.