llama-server followups
Misc fixes for #16031
Add back dropped ROCm build flag for multi-GPU support on windows
Fix amdhip64_*.dll version detection for "latest" selection
Fix embeddings API for consistent normalize behavior with prior versions
ci: set up for automated llama.cpp update testing
reduce batch for fa-disabled, and constrained...
🛡️ VERIFIED CYBER INTELLIGENCE ID: #3535807