Puget Systems tests dual AMD Radeon AI PRO R9700 for local LLM inference

Puget Systems benchmarked two AMD Radeon AI PRO R9700 GPUs (32GB each, $1,880 per card) for local LLM inference, finding stock vLLM multi-GPU doesn't work on these cards. They used llama.cpp's ROCm backend for a 27B-parameter model requiring both cards.
1 source
Developers by email
Get an email when there's news on Developers
No news that day, no email.
More stories today
- Security researcher changes mind on AI guardrails
- Forescout uses Claude AI to port PLC exploit in hours
- Weaviate shows how to extract meaning from charts and tables in PDFs
- Alok launches AI-powered personalized music video campaign for WAAW headphones
- X launches MCP server for advertiser tools