Skip to content
local-ai

What can I run?

Pick your hardware, or enter how much memory you have, and see which models in our catalogue fit, at the best quantisation each allows. Everything runs in your browser; nothing is sent anywhere.

Model type

15 models fit in 12GB.

Text generation

  • Best fitQ4_K_M
    Approx. VRAM~9GB / 12GB
    Tight — little room for long contextCommercial use permitted
  • Best fitQ8_0
    Approx. VRAM~8.5GB / 12GB
    Tight — little room for long contextCommercial use permitted
  • Best fitFP16
    Approx. VRAM~8GB / 12GB
    Tight — little room for long contextCommercial use permitted
  • Best fitFP16
    Approx. VRAM~6.5GB / 12GB
    Workable — some headroomPermitted with conditions

Code

Embedding

  • BGE-M3

    568M
    Best fitFP16
    Approx. VRAM~1.2GB / 12GB
    Comfortable — room for longer contextCommercial use permitted
  • Best fitFP16
    Approx. VRAM~1.2GB / 12GB
    Comfortable — room for longer contextCommercial use permitted
  • Best fitFP16
    Approx. VRAM~0.3GB / 12GB
    Comfortable — room for longer contextCommercial use permitted

Vision-language

  • Best fitQ8_0
    Approx. VRAM~9GB / 12GB
    Tight — little room for long contextCommercial use permitted
  • Best fitQ4_K_M
    Approx. VRAM~8GB / 12GB
    Tight — little room for long contextPermitted with conditions

Image generation

  • Best fitGGUF Q4
    Approx. VRAM~8GB / 12GB
    Tight — little room for long contextNon-commercial only
  • Best fitGGUF Q4
    Approx. VRAM~8GB / 12GB
    Tight — little room for long contextCommercial use permitted
  • Best fitFP16
    Approx. VRAM~7GB / 12GB
    Workable — some headroomPermitted with conditions

Speech to text

  • Best fitFP16
    Approx. VRAM~3GB / 12GB
    Workable — some headroomCommercial use permitted

Text to speech

  • Best fitFP16
    Approx. VRAM~0.3GB / 12GB
    Comfortable — room for longer contextCommercial use permitted