Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I would be curious about context window size that would be expected when generating ballpark 20 to 20 tokens per second using Deepseek-R1 Q4 on this hardware?


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: