Logo
Explore Help
Sign In
wmantly/turing-multi-gpu-llm-server
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions Packages Projects Releases Wiki Activity
30 Commits 1 Branch 0 Tags
aebc82e40380fdddc00fcf8afee995b074631649
Commit Graph
7 Commits
Author SHA1 Message Date
wmantly cd989b3e70 feat: update to Q5_K_P @ 128K ctx, non-blocking proxy sessions, and full multi-GPU benchmark suite 2026-09-13 02:14:25 +00:00
wmantly 5daafe24e8 refactor: remove container-side power governor in favor of host-level cmp-tune 2026-09-02 21:01:52 +00:00
wmantly 38b308c6fd Docs: Update README with 38.7-44.6 tok/s decode, 502 tok/s prefill, and 33W cluster idle power benchmarks 2026-09-02 01:04:49 +00:00
wmantly d726501d71 Docs: Add master CMP Reverse Engineering Ecosystem & Community Map (projects, techniques, hardware mods, and AI relevance) 2026-09-02 00:07:59 +00:00
wmantly 2a21eae435 Hardware: Add MSI CMP 50HX VBIOS ROM and document the 100% Video Engine Pinning bug and cross-flash idle power drop 2026-09-01 19:24:00 +00:00
wmantly 5f261c191f Docs: Update HARDWARE_LEARNINGS and README with low-latency NCCL ring buffer benchmarks and speculative decoding findings 2026-09-01 16:06:32 +00:00
wmantly aeb72cecfa Initial commit: Complete deployment scripts, power governor, systemd units, and architecture documentation for Turing multi-GPU LLM rig 2026-09-01 01:55:19 +00:00
Powered by Gitea Version: 1.27.3 Page: 25ms Template: 5ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API