Blog
Engineering

Artemis on Intel AI Tiber Cloud: Whisper 20% faster on Xeon

Same model, same weights, same accuracy. Artemis optimised OpenAI's Whisper to run 20% faster on Intel Xeon 8380P with Gaudi accelerators — code changes only.

Artemis on Intel AI Tiber Cloud: Whisper 20% faster on Xeon
1 Apr 2025 · 2 min read

We ran Artemis against OpenAI's Whisper on Intel's AI Tiber Cloud, targeting Xeon 8380P processors with Gaudi accelerators.

The setup

The constraint was deliberately strict: the model architecture does not change. No swapping layers, no quantisation trade-offs presented as free wins, no retraining. Only the code around the model is allowed to move.

No architecture changes

Under that constraint Artemis found 20% faster inference. The value of the constraint is that the result is boring to adopt — there is no accuracy conversation to have with a risk committee, because the model is byte-for-byte the one you already approved.

More blogs

Discover the ROI hiding in your stack.

Point Artemis at a system you already run, and see the improvement it finds, validated, before you change a thing.