aa55634571cfe5a6d87e9d6f5a4493d32cf181fa
litellm-health: added GPU fleet topology table, model routing map, 3-tier GPU checks (reachability, VRAM, temp, test inference) litellm-self-heal: added GPU fleet reference, rules for GPU unreachable and model not responding Infrastructure topology now fully documented: - amdpve: Strix Halo CPU (35B, 16 threads, 262K ctx, llama-server) - llm-gpu: RTX 3090 24GB (Dense tier) - ocu-llm: RTX 5070 12GB (MoE + Light tiers)
Description
OpenProse contracts for Syslog agent mesh — health checks, deployments, workflows
335 KiB
Languages
Markdown
100%