Skip to content

feat(qwen4_exp): load-time per-tensor FP8 dense projections (W8A8 via _scaled_mm) - #389

Draft
gdevenyi wants to merge 6 commits into
FlashML-org:mainfrom
gdevenyi:feat/qwen4-exp-fp8-dense
Draft

feat(qwen4_exp): load-time per-tensor FP8 dense projections (W8A8 via _scaled_mm)#389
gdevenyi wants to merge 6 commits into
FlashML-org:mainfrom
gdevenyi:feat/qwen4-exp-fp8-dense

Enhance your code review process with GitHub Actions

GitHub Actions make it easy to automate all your software workflows, now with world-class CI/CD.
Build, test, and deploy your code right from GitHub. Learn more about GitHub Actions.

Linux, macOS, Windows, and containers
Linux, macOS, Windows, and containers
Matrix builds
Matrix builds
Any language
Any language
Live logs
Live logs
Built-in secret store
Built-in secret store
Multi-container testing
Multi-container testing