surya-ocr-2-gguf

Try Try

surya-ocr-2-gguf has 632M parameters (565M active per token) and takes 1.27 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

datalab-to/surya-ocr-2-gguf
Repo createdMay 14, 2026updated May 27, 2026
Model typeDenseHybrid attention: 18 linear + 6 full (GQA)
InputsText + imagesvision encoder in a separate mmproj file, not counted here
Total parameters632M
Active per token565M89% of the model
ExpertsNone (dense)
Max context (from config)256K tokens
Layers2418 linear + 6 full attention
On disk1.27 GB1 file
PrecisionFP16 (100%), other (0.16%)
Quantizationsurya-2 (GGUF)original precision (FP16)
Fewest GPUs1× RTX 4060a single 8 GB card · weights only
Licenseopenrail
GitHubdatalab-to/surya