Their statement about sharing an Ultra for an office LLM server is a little misleading. The tokens/second would be too slow for any business use case. You'd want to buy multiple GPUs for that purpose, probably Nvidia, despite the extra cost.