Phi-3.5-vision-instruct

Try Try

Phi-3.5-vision-instruct has 4.15B parameters (3.72B active per token) and takes 8.29 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

microsoft/Phi-3.5-vision-instruct
Repo createdAug 16, 2024updated Dec 10, 2025
Model typeDenseFull attention (MHA)
InputsText + images326M vision encoder
Total parameters4.15B
Active per token3.72B90% of the model
ExpertsNone (dense)
Max context (from config)128K tokens
Layers32
On disk8.29 GB2 files
PrecisionBF16 (100%)
QuantizationNoneoriginal precision (BF16)
Fewest GPUs1× RTX 3060 12 GBa single 12 GB card · weights only
Licensemit
GitHubNot linked