Laya on Mac M4 CoreML Offline
putna
148 points
30 comments
September 20, 2026
Related Discussions
Found 5 related stories in 89.9ms across 7,193 title embeddings via pgvector HNSW
- My local model setup on an M4 Pro Mac Mini raybb · 132 pts · September 01, 2026 · 56% similar
- Apple Silicon and macOS VMs: Faster LLM Inference with llama.cpp frabonacci · 289 pts · August 11, 2026 · 55% similar
- LFM2.5-DSpark: Up to 3.2x Faster Inference from H100 to MacB Alephinitesimal · 15 pts · August 21, 2026 · 52% similar
- I Benchmarked Local LLMs on the Laptop I Have konmam · 20 pts · August 10, 2026 · 48% similar
- DeepSeek v4.1 flash runs 23 seconds/token on a 2020 16gb M1 Mac Mini ajay-higgs · 14 pts · September 12, 2026 · 48% similar
Discussion Highlights (8 comments)
PaulRobinson
Local LLMs are the future, and one of the reasons I think the data centre furore is just going to end in a market crash. LLMs that can reliably be used for control problems are the future, and I think classic/deep RL has generally been overlooked for years for a whole host of problems by wider industry because it felt inaccessible. The first thing I thought of when I saw Jev (and then Laya), was "this might move the needle in a really, really interesting way". Local LLMs that can reliably be used for control problems smash through a lot of barriers I'm interested in, and this intrigues me a lot. Guess I'm about to become a big Laya fan if it can run on this kind of hardware to this performance.
tentacleuno
This looks like a local AI model playing Snake -- is that correct? The article offers no explanation.
frag
Great job! Did you finetune your own Laya for the snake game or what?
altano
How much memory does this use of the test machine's (M3 Max) 128 GB unified memory?
imranq
Based on my admittedly limited research, it seems like you should use Laya for much more deterministic tasks where you have some training data. It won't be as good as Jev for zero shot cases.
EgregiousCube
Isn't part of the Jev marketing that it has "terra-class intelligence"? I don't know how much it actually achieves that, but unless that's EXTREMELY wrong, it's hard to see how a 0.3B model could claim to be an OS Jev.
speedping
So cool. I've fired up pumas (energy monitor) and it seems to run almost fully on the neural engine and not the GPU so it plays really nicely with CoreML
brcmthrowaway
Why is Laya being shilled here? It doesn't have real intelligence backing it.