Community Forum

Notifications
Clear all

Is 16GB VRAM enough for DeepSeek 14B GGUF Q4_K_M?

3 Posts
2 Users
0 Reactions
245 Views
(@nvidiafan99)
New Member
Joined: 1 month ago
Posts: 0
Topic starter   [#4]

quick question guys before i upgrade my GPU... will a RTX 4060 Ti 16GB run DeepSeek R1 14B Q4 smoothly with 8k context or should i save up for a used 3090 24GB?


Moderation Report Detected
Type: Auto Moderation
Score: -
Action: Unapproved
Summary: Content unapproved: User requires manual approval (no "Can pass moderation" permission).
Auto Moderation

   
Quote
(@hardwareguru_mark)
New Member
Joined: 1 month ago
Posts: 0
 

16GB will fit 14B Q4 easily with around 45 tokens/sec on llama.cpp. but if u wanna use 16k context or 32B quant then definitely get the used 3090 24gb, VRAM is king for LLMs.



   
ReplyQuote
(@nvidiafan99)
New Member
Joined: 1 month ago
Posts: 0
Topic starter  

thanks Mark! yeah I might just look for a good 3090 deal to be future proof for 32b models.



   
ReplyQuote
Share: