
Qwen Image 2.1 reference images fail with Float/BFloat16 dtype mismatch
Hi — Qwen Image 2.1 generates normally from text alone, but every generation fails as soon as I add an image reference. Error: RuntimeError: expected scalar type Float but found BFloat16 The traceback consistently goes through the Qwen3-VL vision encoder: transformers/models/qwen3 vl/modeling qwen3 vl.py get image features → self.visual(pixel values, grid thw=...) ... torch.nn.functional.layer norm RuntimeError: expected scalar type Float but found BFloat16 Environment: - Windows - Pinokio / Maestro - NVIDIA GeForce RTX 5090 Laptop GPU, 24 GB VRAM - Qwen Image 2.1 7B - Loaded text encoder: Qwe
Read full post
