llama-server followups

Misc fixes for #16031

Add back dropped ROCm build flag for multi-GPU support on windows
Fix amdhip64_*.dll version detection for "latest" selection
Fix embeddings API for consistent normalize behavior with prior versions



ci: set up for automated llama.cpp update testing


reduce batch for fa-disabled, and constrained...