621a897bec279732448dbe15c78eb31e80c994f2
More conversations now route to VLM as primary. 9B VLM has 262K context window and 88 tok/s average — well suited for moderate conversations. Dense absorbs overflow and heavy reasoning.
Description
SyslogAI Inference Harness — 3-GPU router, dashboard, LiteLLM proxy
916 KiB
Languages
Python
77%
HTML
22%
Shell
0.8%
Dockerfile
0.2%