
Paiton serves Ornith 1.5 35B A3B at 44.63 output tok/s on one Radeon AI PRO R9700, 27% faster than tuned stock vLLM, with 21.3% lower modeled cost.
FIELD NOTES FROM COMPUTING, AI, AND OPTIMIZATION.
Stay updated on the latest in computing, IT optimization, and sustainable technology. Read ElioVP’s expert insights and industry trends.

Paiton serves Ornith 1.5 35B A3B at 44.63 output tok/s on one Radeon AI PRO R9700, 27% faster than tuned stock vLLM, with 21.3% lower modeled cost.

Paiton serves AMD’s Qwen3.8 27B at 39.77 output tokens/s on one Radeon AI PRO R9700, delivering 21% more throughput and 17.4% lower modeled cost.

Why do AI data-center proposals quote different GPU capacities? See how PUE, peak loads, storage, networking and cooling determine deployable compute.

When we first started building Paiton, one of our earliest focus areas was optimizing diffusion models. Stable Diffusion XL was one of the first large models where we showed that fused operators, efficient execution, and hardware-aware kernels could make a real difference. Now we are returning to those origins.With the growing interest in text-to-video generation, ...

It has been some incredible weeks for the team here at Eliovp. We are extremely proud to share that our company was recently featured on the front page of De Tijd, Belgium’s leading business newspaper. Seeing our story, from our founder’s early days tinkering with wires in an attic to generating €215 million in revenue, ...