Introducing System One Models and Jev

albelfio 1080 points 326 comments September 15, 2026
typesafe.ai · View on Hacker News

Discussion Highlights (20 comments)

albelfio

https://x.com/completeskeptic/status/2099925682726002904?s=4... The doom demo is quite cool

dgellow

Side note: it took me more time than I would like to admit to realize that Diogo Almeida isn’t a satirical version of the name Dario Amodei

scottyah

Wild that it doesn't generate text. I wonder how its technology compares to Tesla's FSD stack.

jrickert

Signed up for the beta! :) would love to put this through some real-world shootouts against traditional LLMs to see where this type of model really excels. I’m guessing it might be able to replace maybe 40-70% of LLM calls for a given pipeline depending on the business task, cutting the API costs on those calls by an order of magnitude.

pennomi

> Extraordinary claims require extraordinary evidence so see below for the receipts. Yes, that’s the kind of attitude I want to see in these model releases

esafak

Looks like a great model for NLP.

hunterbrooks

um what is going on with the outfit changes in the launch video... https://x.com/CompleteSkeptic/status/2099925682726002904

vatsachak

It could be used for coding if you gave it an AST. If you work at TypeSafe please try this. Side note: This is probably how LLMs would perform with better encoders and next-latent prediction, so eventually those will beat this architecture out. Still amazing though.

sim04ful

This sort of stuff almost sends shivers down my spine, it's like i'm looking 5 years into the future.

erichocean

I could put this to use today. I think we'll see a bunch of different architectures over the next five years.

himata4113

They never show exactly how they use it? Only a bunch of animations of it 'working'. Would like to see the actual code used for the demos!

bfeynman

Super intrigued by this - large scale automation using LLMs is quite annoying due to deprecation cycles of models from frontier labs and cost of running your own being prohibitive when you have a blend of them.

mushufasa

I would love for things like this to be accessible via hubs like open router or AWS bedrock. It's hard to justify adding new model vendors directly with all the heightened concerns about privacy and security, but if bold new capabilities are added to a centralized already-vendor like AWS, technical people can adopt them without going through a whole compliance/purchasing/vendor review process. And an extra middleman tax is well worth it when the cost savings of the model itself can be one-two orders of magnitude.

ramon156

This sounds good but so far all claims just sound like marketing terms. I'd love to see real proof. e.g. "RLCD" and "parallel sampling" have nothing to back it up. also "70-500ms vs 3-329 seconds" are apples-to-oranges unless the LLM baseline is doing comparable work (e.g., long chain-of-thought). If Jev is skipping generation entirely for a narrow structured task, of course it's faster. Nonetheless i want this to be true, so I'm looking forward to Jev Edit: I really have to say that I like their manifesto https://typesafe.ai/manifesto

initsecret

> [others] Output tokens: ~5x more expensive than input tokens. > [them] Output tokens: FREE (too cheap to meter). I'm very confused by this.

whalesalad

What is it about the rendering of this page that is so... off? It almost looks like the entire thing is a <canvas> element. edit: looks like a framer export where there is a text stroke being applied :|

andai

Why did they pick the name System One? It's not really explained what "System One tasks" and "System One shaped queries" are. Things that need a fast response? Does this imply it's a very small model? I couldn't find anything about the model itself.

totallygeeky

Woof, that page is hard to read. I don't understand what they've done to the way text is rendering but it's not great for my eyes.

jceg

> We deliberately chose not to publish performance against public benchmarks. In fact, we plan to only have one-off evals when we make product updates. lol, I bet they would publish them if their score on those benchmarks were good.

larodi

"is this the real thing or is just fantasy"

Semantic search powered by Rivestack pgvector
6,718 stories · 61,457 chunks indexed