#
mining-gpu
Here are 4 public repositories matching this topic...
What it costs to split an image generation model across GPUs - layer vs row split vs worker parallel, measured on P104-100 x4 (PCIe Gen1 x4) and V100 x4 (PCIe Gen3 x16)
flux pascal benchmark gpu cuda self-hosted image-generation quantization multi-gpu model-parallelism v100 diffusion-models tensor-parallelism stable-diffusion-cpp mining-gpu p104-100
-
Updated
Sep 16, 2026 - Shell
64 GB of HBM2e on a dead mining card: running Qwen3-Next-80B on an NVIDIA CMP 170HX with vLLM. Configs, gotchas, and measured benchmarks.
benchmark gpu nvidia homelab a100 llm-serving vllm local-llm llm-inference qwen cmp-170hx mining-gpu
-
Updated
Sep 13, 2026 - Python
Layer split vs tensor parallelism in llama.cpp - measured on two platforms 16x apart in interconnect bandwidth
pascal benchmark gpu cuda inference self-hosted quantization multi-gpu model-parallelism nccl v100 pipeline-parallelism tensor-parallelism llm llama-cpp qwen speculative-decoding mining-gpu p104-100
-
Updated
Sep 16, 2026 - Python
Add this topic to your repo
To associate your repository with the mining-gpu topic, visit your repo's landing page and select "manage topics."