feat(qwen4_exp): load-time per-tensor FP8 dense projections (W8A8 via _scaled_mm) - #389
Draft
gdevenyi wants to merge 16 commits into
Draft
feat(qwen4_exp): load-time per-tensor FP8 dense projections (W8A8 via _scaled_mm)#389gdevenyi wants to merge 16 commits into
gdevenyi wants to merge 16 commits into
Enhance your code review process with GitHub Actions
GitHub Actions make it easy to automate all your software workflows, now with world-class CI/CD.
Build, test, and deploy your code right from GitHub. Learn more about GitHub Actions.