Skip to content

[MODEL] add diffusion_gemma quantization support - #3082

Merged
Qubitium merged 3 commits into
mainfrom
feat/diffusion-gemma-quantization
Sep 15, 2026
Merged

Qubitium merged 3 commits into
mainfrom
feat/diffusion-gemma-quantization

Conversation

@ZX-ModelCloud

@ZX-ModelCloud ZX-ModelCloud commented Sep 14, 2026 •

Copy link
Copy Markdown
Collaborator

Summary

  • add native DiffusionGemma multimodal quantization support for diffusiongemma-26B-A4B-it
  • quantize the encoder capture path and restore shared QuantLinear/dense aliases on the diffusion decoder
  • support LazyTurtle checkpoint namespace remapping, packed MoE experts, attention-mask replay, and specialized generation config
  • replace stale broad tied-weight rules with exact storage-backed mappings to prevent Transformers tie_weights_keys failures
  • add image calibration/generation coverage and document the model in README
  • remove the unused defuser_auto_detect_moe model flag

Dependency

Validation

  • conda run -n gp_311 pytest -q tests/test_diffusion_gemma_support.py — 17 passed

@ZX-ModelCloud ZX-ModelCloud changed the title feat: add DiffusionGemma quantization support [MODEL] add diffusion_gemma quantization support Sep 14, 2026
@Qubitium
Qubitium merged commit 877f918 into main Sep 15, 2026
6 checks passed
@Qubitium
Qubitium deleted the feat/diffusion-gemma-quantization branch September 15, 2026 02:11
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants