Skip to content
local-ai

NVIDIA GeForce RTX 4090

NVIDIA · Enthusiast

The high-end consumer GPU that has been the default enthusiast choice for local AI, with 24GB of fast GDDR6X memory and strong compute. Enough to run capable models at good quantisations, though 24GB sets a real ceiling on model size.

Specifications

VRAM 24GB
Memory bandwidth 1008 GB/s
Power draw 450W
Type gpu

Prices have moved around a lot with demand. Treat any figure as indicative and check current retail before buying.

Strengths

  • Fast memory bandwidth, which matters a great deal for inference speed
  • 24GB comfortably runs strong models up to roughly 32B at good quantisations
  • Mature CUDA support across every major inference engine

Weaknesses

  • 24GB rules out larger models without heavy quantisation or a second card
  • High power draw and heat under sustained load
  • Enthusiast pricing, and availability has been inconsistent
See what models fit in 24GB →

Entry last verified 15 January 2026. Specifications and especially prices change; verify before buying.