This website requires JavaScript.
Explore
Help
Sign In
wmantly
/
turing-multi-gpu-llm-server
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
Packages
Projects
Releases
Wiki
Activity
Files
096fb421d19f4409f679c72085f1c67c7b7ae102
turing-multi-gpu-llm-server
/
systemd
T
History
wmantly
9ab0583cab
feat: implement session-managed persistent KV cache architecture with slot persistence and management API
2026-09-04 02:53:54 +00:00
..
llama-server.service
feat: implement session-managed persistent KV cache architecture with slot persistence and management API
2026-09-04 02:53:54 +00:00
ollama-proxy.service
Initial commit: Complete deployment scripts, power governor, systemd units, and architecture documentation for Turing multi-GPU LLM rig
2026-09-01 01:55:19 +00:00