Qwen3.8-2.4T

mmastrac 83 points 11 comments August 12, 2026
huggingface.co · View on Hacker News

Discussion Highlights (5 comments)

brrrrrm

this is Qwen3.8 max, right? https://qwen.ai/blog?id=qwen3.8

UncleOxidant

We're all really waiting for the 3.8-27B which is due out in 2 days.

WalterGR

Submitted 3 hours ago with 63 comments so far: https://news.ycombinator.com/item?id=49273478

xlayn

I wonder besides big labs and REALLY big corporations who can run the 2.6TB q8... I mean, the 1 bit one is 508GB. Assuming 256k context size as irrelevant at those sizes, I would need 22 AMD 7900XTX (24GB vram) to run this for the 1 bit one, and 113 for the 2.6TB 113 and assumming close to constant load and the gpus taking turns as it does on my machine with two gpus is around 113gpus*113W = 11KW... just to hold and run that one instance... and only god knows the cooling requirements... I guess just someone on Meta/Googly/Claudy will say... oh nice, let's download it and run it... How did the unsloth guy did to process a file that size? https://huggingface.co/unsloth/Qwen3.8-2.4T-A95B-GGUF

edhan235

I have been testing this for hark.news

Semantic search powered by Rivestack pgvector
4,128 stories · 37,281 chunks indexed