OpenAI has a LOT of work to do if they think Luna can compete with Jev

AnthusAI 12 points 4 comments October 01, 2026
anth.us · View on Hacker News

Discussion Highlights (3 comments)

verdverm

apples and oranges, also Clef and Kev (the middle path) https://blog.cloudflare.com/clef-decision-models/ https://github.com/jaredpalmer/kev/tree/main

rileymat2

Alan is young, round, and kind, but that doesn't mean he isn't also rough and cold at times, as well. … Young round people who are green are usually blue. … Kind people with rough skin are usually red because it's wind burn. If someone shows that they are red, then they are also showing that they are green. … Statement: Alan is not blue. A log-probability of −0.00182 is a probability of 99.82%. We asked for five alternatives and got none: Luna put essentially nothing on "true" or "false". And it's wrong. Alan is kind with rough skin, so he's red; red means green; young, round and green means blue. "Alan is not blue" is false, three steps in. —————- Can someone explain this I got unknown as well. The problem statement includes the word “usually” a few times.

itg

If I'm reading this right, they didn't actually use the Decisions API, they used GPT-6 Luna. Wouldn't call this a good comparison.

Semantic search powered by Rivestack pgvector
8,245 stories · 76,959 chunks indexed