π§
Qwen3 4b Dflash B16 model by unieai
β 39.1
π¬Technical Deep Dive
Full Specifications [+]
π Updated daily
Source summary: Based on Hugging Face metadata. Not a recommendation.
π‘οΈ Model Transparency Report
Technical metadata sourced from upstream repositories.
Open Metadata
π Identity & Source
- id
- hf-model--unieai--qwen3-4b-dflash-b16
- slug
- unieai--qwen3-4b-dflash-b16
- source
- huggingface
- author
- unieai
- license
- MIT
- tags
- transformers, safetensors, qwen3, feature-extraction, dflash, speculative-decoding, diffusion, efficiency, flash-decoding, qwen, diffusion-language-model, text-generation, custom_code, arxiv:2602.06036, license:mit, text-generation-inference, endpoints_compatible, region:us
βοΈ Technical Specs
- architecture
- DFlashDraftModel
- params billions
- 0.54
- context length
- 32,768
- pipeline tag
- text-generation
- vram gb
- 2.9
- vram is estimated
- true
- vram formula
- VRAM β (params * 0.75) + 2GB (KV) + 0.5GB (OS)
π Engagement & Metrics
- downloads
- 16
Data indexed from public sources. Updated daily.