FL PenguinVL

FL PenguinVL - Vision-Language Model nodes for ComfyUI. Query images and videos with natural language using Tencent's Penguin-VL (2B/8B). Supports OCR, document analysis, visual QA, dense captioning, and more.

Quick Technical Summary: FL PenguinVL

Base VRAM Footprint:
4096 MB (8 GB Tier)
Primary Dependencies:
accelerate>=0.26.0, ffmpeg-python>=0.2.0, huggingface_hub>=0.20.0, numpy>=1.20.0, pillow>=9.0.0, safetensors>=0.4.0, torch>=2.5.0, transformers>=4.40.0
Min PyTorch / CUDA:
PyTorch 2.5.0 | CUDA 12.1+
GitHub Repository:
https://github.com/filliptm/comfyui-fl-penguinvl

Citation Note: Data sourced from VRAM DB. For complete workflow OOM estimations, use the VRAM DB Workflow Analyzer.

How much VRAM does FL PenguinVL require?

Direct Answer: The ComfyUI node FL PenguinVL requires a minimum base VRAM of 4096MB and is optimized for GPUs with at least 8GB of VRAM. Low VRAM mode is not supported for this node.

Very High (4-8GB)
Base VRAM:
4096MB (4.0GB)
Recommended GPU:
8GB+ VRAM
Low VRAM Mode:
✗ Not supported
Estimation Confidence:
MEDIUM

Cheapest VRAM Upgrade Paths (Live Market Prices):

  • GeForce RTX 3060 12GB (Ultimate Budget VRAM King)──► Used: $209.62View eBay ↗
  • GeForce RTX 4060 8GB (Modern Entry-Level)──► New: $303.50View Amazon ↗

Interactive VRAM Compatibility Estimator

Estimated Total VRAM: 3.00 GBTarget: 8 GB
✅ Comfortable Fit

Your GPU has plenty of headroom. You can run this node safely with your active configurations!

Verify Compatibility for Your Specific GPU VRAM

Select your graphics card's VRAM capacity to view optimized batch sizes, suggested resolutions, and custom performance tips for FL PenguinVL:

Live Cloud Deploy Options

Live Market Rates

Run this node in cloud environments with pre-configured CUDA/PyTorch dependencies:

Buy NVIDIA GeForce RTX 3060 (12GB VRAM)

Tired of renting cloud rigs? Run ComfyUI locally with absolute zero latency. Best entry-level ComfyUI experience. Avoids immediate VRAM limitations on basic LoRA training.

🛒 Buy on Amazon

Are you the author of this node?

Help your users avoid out-of-memory errors by displaying this professional, dynamic VRAM badge on your GitHub README. Copy the markdown below to embed it with a backlink directly to this hardware specification profile.

Live Preview:VRAM DB Badge
embed markdown
[![VRAM Specs](https://vramdb.com/badge/filliptm/comfyui-fl-penguinvl.svg)](https://vramdb.com/nodes/filliptm/comfyui-fl-penguinvl/)

What Python packages are required for FL PenguinVL?

Direct Answer: Running FL PenguinVL requires installing the following Python package dependencies: accelerate>=0.26.0, ffmpeg-python>=0.2.0, huggingface_hub>=0.20.0, numpy>=1.20.0, pillow>=9.0.0, safetensors>=0.4.0, torch>=2.5.0, transformers>=4.40.0. This node specifically requires PyTorch version 2.5.0 or newer.

requirements.txt
accelerate>=0.26.0
ffmpeg-python>=0.2.0
huggingface_hub>=0.20.0
numpy>=1.20.0
pillow>=9.0.0
safetensors>=0.4.0
torch>=2.5.0
transformers>=4.40.0

Interactive Setup & Dependency Resolver

Operating System:
Environment Type:
Run this terminal command in your ComfyUI root folder:
# Loading command...

Special Environment Requirements:

Requires PyTorch version 2.5.0+.

To update PyTorch for your selected setup, run:
# Loading PyTorch command...

Frequently Asked Questions

How much VRAM does FL PenguinVL require?

FL PenguinVL requires a minimum of 4096MB (4.0GB) of VRAM for base operation. For optimal performance, a GPU with at least 8GB of VRAM is recommended. Low VRAM mode is not supported for this node.

Can I run FL PenguinVL on an RTX 3060, RTX 4070, or RTX 4090?

✅ RTX 3060 (12GB): Yes, fully compatible with 6.8GB headroom. ✅ RTX 4070 (12GB): Yes, fully compatible with 6.8GB headroom. ✅ RTX 4070 Ti (16GB): Yes, fully compatible with 10.4GB headroom. ✅ RTX 4090 (24GB): Yes, fully compatible with 17.6GB headroom

How much VRAM does FL PenguinVL take on an RTX 3060 vs RTX 4090?

On an RTX 3060 (12GB VRAM), FL PenguinVL runs smoothly on an RTX 3060 (12GB) with 6.8GB of headroom. This is sufficient to run the node alongside standard SD 1.5 and SDXL workflows in full precision. On an RTX 4090 (24GB VRAM), the node runs with extreme headroom on an RTX 4090 (24GB) with 17.6GB of dedicated headroom. This allows you to combine the node with massive models (like FLUX.1 Dev, Schnell, or Hunyuan Video) in full precision (FP16) without any offload flags.

What PyTorch version does FL PenguinVL need?

FL PenguinVL requires the following PyTorch-related packages: torch>=2.5.0. Ensure your ComfyUI environment has these installed. Ensure your PyTorch installation matches your CUDA version (use torch.version.cuda to check).

What Python packages are required for FL PenguinVL?

To run FL PenguinVL, you need to install: accelerate>=0.26.0, ffmpeg-python>=0.2.0, huggingface_hub>=0.20.0, numpy>=1.20.0, pillow>=9.0.0, safetensors>=0.4.0, torch>=2.5.0, transformers>=4.40.0. You can install these using pip or add them to your requirements.txt file.

How do I install FL PenguinVL in ComfyUI?

To install FL PenguinVL: (1) Navigate to your ComfyUI/custom_nodes directory, (2) Clone the repository: git clone https://github.com/filliptm/comfyui-fl-penguinvl, (3) Install dependencies: pip install -r requirements.txt (if present), (4) Restart ComfyUI. Alternatively, use ComfyUI Manager for one-click installation.