A FLUX.2-dev model quantized to INT8 with ConvRot using a conservative quantization policy.
On my potato machine, in T2I without using distilled LoRA, int8-convrot-aggressive is 20 seconds faster than int8-convrot. Neither model produced images that deviated from bf16.
T2I Example:
Image Edit Example:

- Downloads last month
- 1,278
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support
Model tree for AX1Y2JP/FLUX.2-dev-INT8-ConvRot
Base model
black-forest-labs/FLUX.2-dev