DeepSeek v4.1 flash runs 23 seconds/token on a 2020 16gb M1 Mac Mini

ajay-higgs 14 points 4 comments September 12, 2026
twitter.com · View on Hacker News

Discussion Highlights (3 comments)

spottedmarley

Nice! Totally different subject but, have you tried Ling-mini on it? It's really good for how small it is and you can run it on damn near anything and be useably quick.

a012

You almost got me, it’s 2.6 tokens/minute

killingtime74

I did a task that cost about $1.5 on API. It was 4 million tokens. That will take about 2.9 years at this rate. Estimating an electricity cost in the three digits

Semantic search powered by Rivestack pgvector
6,278 stories · 57,251 chunks indexed