Floyo
Floyo
Workflows
API
Pricing
Floyo
Floyo
Workflows
API
Pricing

ComfyUI-QwenVL

Created by 1038lab

https://github.com/1038lab/ComfyUI-QwenVL

900

Last updated 2026-09-14

Last updated

2026-09-14

Run hundreds of ComfyUI nodes and workflows in your browser.

ComfyUI and My Files on Floyo

The ComfyUI-QwenVL tool integrates Alibaba Cloud's advanced Qwen-VL vision-language models into the ComfyUI framework, enabling sophisticated multimodal AI tasks like text generation, image comprehension, and video analysis. It supports various Qwen models, including Qwen2.5-VL and multiple versions of Qwen3-VL, while utilizing a high-performance dual-engine architecture for efficient processing.

  • Supports a wide range of Qwen models, providing flexibility for users to select the most suitable model for their projects.
  • Features intelligent video auto-scaling and token budget management, ensuring optimal performance without overwhelming system resources.
  • Offers both standard and advanced nodes, allowing users to choose between quick setups or detailed control over model parameters.

Context

The ComfyUI-QwenVL extension is designed to incorporate the Qwen-VL series of models into the ComfyUI environment, facilitating a seamless integration for users looking to leverage advanced AI capabilities. Its purpose is to enhance the existing functionalities of ComfyUI by providing powerful tools for processing and understanding both visual and textual data.

Key Features & Benefits

This extension boasts several practical features, including:

  • Multi-Model Support: Users can easily switch between different Qwen models, enabling tailored approaches for specific tasks.
  • Automatic Model Management: Models are automatically downloaded upon first use, streamlining the setup process for users.
  • Dynamic Video Scaling: The tool intelligently manages video context tokens, preserving quality while preventing memory overflow issues.

Advanced Functionalities

The QwenVL extension includes advanced capabilities such as:

  • SageAttention Support: This feature optimizes the attention mechanism for different GPU architectures, enhancing processing speed and efficiency.
  • Custom Model Expansion: Users can easily integrate their own models into the workflow, allowing for greater customization and flexibility.
  • Detailed Parameter Control: The advanced node provides comprehensive control over sampling, device selection, and other critical parameters, catering to users who require fine-tuned outputs.

Practical Benefits

Utilizing the ComfyUI-QwenVL tool significantly enhances workflow efficiency by providing robust error handling and memory management features. It allows users to maintain models in VRAM for quicker processing, thereby improving the overall quality of output generated from both images and videos. The streamlined integration of multimodal capabilities translates to a more effective and intuitive user experience within ComfyUI.

Credits/Acknowledgments

This tool is developed by the Qwen Team at Alibaba Cloud, with contributions from the ComfyUI community and other collaborators. It is licensed under the GPL-3.0 License, ensuring that it remains open-source and accessible for further development and enhancement by users.

Inner Nodes

AILab_HuggingFaceDownloader
AILab_QwenVL
AILab_QwenVL_Advanced
AILab_QwenVL_GGUF
AILab_QwenVL_GGUF_Advanced
AILab_QwenVL_GGUF_PromptEnhancer
AILab_QwenVL_PromptEnhancer

Discover most popular workflows

Hand-picked based on what hundreds of other artists looked at.

Created by 1038lab

Inner Nodes

AILab_HuggingFaceDownloader

AILab_QwenVL

AILab_QwenVL_Advanced

AILab_QwenVL_GGUF

AILab_QwenVL_GGUF_Advanced

AILab_QwenVL_GGUF_PromptEnhancer

AILab_QwenVL_PromptEnhancer