Skip to content
View daimonionnn's full-sized avatar

Block or report daimonionnn

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. amd-vega-rocm-vulkan-llm-toolkit amd-vega-rocm-vulkan-llm-toolkit Public

    Run LLMs on AMD Vega8 / Vega10 APUs and GPUs (Ryzen 5700G / gfx90c and Vega 56/64): ROCm 7 gfx900 backport + Vulkan/RADV launchers, Docker images, benchmarks for llama.cpp and LM Studio

    Shell 14

  2. amd-r9700-vllm-and-tuning-toolkit amd-r9700-vllm-and-tuning-toolkit Public

    AMD Radeon 9700 AI PRO RDNA multi GPU Tools for llm/vLLM inference, tuning, overclocking, undervolting and LLM models benchmarking

    Shell 13 2

  3. hermes-installation-toolkit hermes-installation-toolkit Public

    Shell 1 1

  4. multi-gpu-llm-toolkit multi-gpu-llm-toolkit Public

    Run llama.cpp across two GPUs of different vendors at once (AMD ROCm/HIP or Vulkan + NVIDIA CUDA) in one llama-server process, no RPC server. Windows (PowerShell) and Linux (bash) implementations, …

    Shell 1

  5. ik-llama-toolkit ik-llama-toolkit Public

    A one-command inference server and benchmark harness for running large language models that don't fit in your GPU's VRAM. Optimized for RTX PRO 6000 and Deepseek v4 flash

    Shell

  6. rtxpro6000-llm-toolkit rtxpro6000-llm-toolkit Public

    Launch profiles, scripts and measurements for running LLMs on a single RTX PRO 6000 Blackwell 96 GB with SGLang, vLLM or ExLlamaV3

    Shell