Verify the AI proxy stack on the real R9700 box (LiteLLM scheduler smoke test) #17
Notifications
Due Date
No due date set.
Depends on
#14 Author the docker-compose service for the chosen proxy
haylan/LLM-Server
Reference: haylan/LLM-Server#17
Reference in New Issue
Block a user
Part of #9
Question
Deploy and verify the docker-compose changes from #14 (litellm + litellm-db services, litellm-config.yaml) on the real Radeon R9700 hardware — mirrors map #1's real-hardware verification ticket (#5).
Specifically confirm:
docker compose up -dwith realLITELLM_MASTER_KEY/LITELLM_SALT_KEYset)./key/infoand the Admin UI dashboard after a real call.priorityfield doesn't leak into llama.cpp's request, and that high-priority requests actually get dispatched ahead of low-priority ones under concurrent load. If broken, flag it — #16's fallback (a queuing shim) becomes live work.proxy.ai.home/proxy.ai.haylan.chNPM Proxy Hosts route correctly, and the external host's/uideny rule actually blocks Admin UI access externally (per docs/network-access.md).