F

Flux: flux-2-klein-4b:free

flux-2-klein-4b:free
VisionFiles
Released Jan 14, 2026Knowledge cutoff Sep 2025Updated Aug 27, 2026

FLUX.2 [klein]-4B is the ultra-compact, high-speed variant of the FLUX.2 family, featuring a 4-billion parameter Rectified Flow Transformer. Designed for sub-second inference on consumer hardware, it can run on GPUs with as little as 13GB VRAM (e.g., RTX 3090/4070). Despite its small size, it unifies text-to-image generation, single-image editing, and multi-reference consistency into a single architecture. Released under the Apache 2.0 license, it is the most accessible frontier model for real-time interactive apps, edge deployment, and local development, delivering a visual quality-to-latency ratio that outperforms much larger models. The price is calculated based on the unit price per cost point, where the cost value is returned in the response after the task is submitted. For example: `json { "id": "5aaeac12-5dbc-43ff-93f0-eb61cfe7274b", "polling_url": "https://api.us2.bfl.ai/v1/get_result?id=5aaeac12-5dbc-43ff-93f0-eb61cfe7274b", "cost": 1.7, "input_mp": 3.0, "output_mp": 1.0 } `

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Compatible endpoints image-generation, openaiVendor Flux

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Frequently asked questions

Similar models