Skip to content
View Yuepixel's full-sized avatar

Block or report Yuepixel

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Popular repositories Loading

  1. tensorfold-qwen38-exl3-lookup tensorfold-qwen38-exl3-lookup Public

    Unofficial fork of TensorFold v0.6.5 (Python engine line; upstream main is now Zig): EXL3 3.05 bpw + a prompt-lookup (suffix) drafter for Qwen3.8-Flash-Next on CUDA — ~2.4x EXL3 prefill, native vis…

    Python 3

  2. qwen38-flash-next-16gb qwen38-flash-next-16gb Public

    Batchfile 2

  3. freetoken-tp2-nvfp4 freetoken-tp2-nvfp4 Public

    FreeToken TP=2 across heterogeneous consumer CUDA GPUs (Ampere+Ada, no native FP4): NVFP4 quantized-MoE expert sharding, bit-exact vs single-GPU greedy. Port of issue #478 methodology.

    Python

  4. TensorFold TensorFold Public

    Forked from ashhart/TensorFold

    LLM Inference Engine for Metal, CUDA and Vulkan.

    Python