This website requires JavaScript.
Explore
Help
Sign In
wmantly
/
turing-multi-gpu-llm-server
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
Packages
Projects
Releases
Wiki
Activity
Files
d1353e9b0d1fbf2030a9a1607e9ea7a97093974b
turing-multi-gpu-llm-server
/
scripts
T
History
wmantly
d1353e9b0d
Perf: Add cmp-tune insane profile, update Proxmox LXC privileged container docs, and update benchmarks to 28.8 tok/s with PCIe Gen2 and unlocked 610.43.03 driver
2026-09-02 00:56:37 +00:00
..
build-nccl-llama.sh
Initial commit: Complete deployment scripts, power governor, systemd units, and architecture documentation for Turing multi-GPU LLM rig
2026-09-01 01:55:19 +00:00
ollama-proxy.py
Fix power governor active retention during long prompt evaluations and remove redundant proxy governor thread
2026-09-01 14:40:19 +00:00
power-governor.py
Enhancement: Integrate NVAPI P8 deep idle downclocking into power governor script
2026-09-02 00:05:23 +00:00
start-server.sh
Perf: Add cmp-tune insane profile, update Proxmox LXC privileged container docs, and update benchmarks to 28.8 tok/s with PCIe Gen2 and unlocked 610.43.03 driver
2026-09-02 00:56:37 +00:00
tune-gpus.sh
Initial commit: Complete deployment scripts, power governor, systemd units, and architecture documentation for Turing multi-GPU LLM rig
2026-09-01 01:55:19 +00:00