GeeCoolGuest
GeeCoolIt does — this kind of high base clock exists for AI services that stay resident but run light. With 64 cores and 128 threads holding dozens of containers that mostly wait for requests, the CPU idles on the 2.2GHz base frequency, and snappiness at low load lives or dies on that number. 320MB of L3 keeps each container's small model and index on-die, so the waiting gaps do not burn memory bandwidth.
Tenant-style services with many containers and light individual requests fit it better than its two siblings; a single 24-hour batch job runs on boost clocks, making extra base frequency a poor spend. Prune idle containers to a minimum so those 2.2GHz land on real requests.