I had Gemini train its own replacement for $9
p-s-v
87 points
43 comments
September 17, 2026
Related Discussions
Found 5 related stories in 69.2ms across 7,105 title embeddings via pgvector HNSW
- Gemini 3.5 Transcribe wmchen · 16 pts · August 26, 2026 · 50% similar
- Gemini-3.5-Transcribe k9294 · 223 pts · August 27, 2026 · 49% similar
- Gemini 3.8 Live and 3.8 Live Extended Thinking leumon · 371 pts · September 15, 2026 · 47% similar
- Gemini 3.8 Flash and 3.8 Flash Cyber simonsarris · 76 pts · September 02, 2026 · 46% similar
- Gemini 3.8 Flash and 3.8 Flash Cyber bratao · 940 pts · September 02, 2026 · 46% similar
Discussion Highlights (20 comments)
trollbridge
Request subtopic be changed to “I used Gemini to design a tool to replace specific uses of Gemini.”
foltik
I found it much more useful to go to a knife shop and handle a whole bunch of knives for myself. They’re all pretty similar besides material, so not much signal you’re going to be able to glean from people arguing on reddit.
yread
I was hoping he tricked Gemini into running the training on the cluster that Gemini itself is running on. That would be novel!
cubefox
> This article was written with the assistance of AI. If that bothers you, stop reading here. Okay!
daveguy
Thank you to the author for disclosing slop writing up front. I appreciate you respecting your readers time.
josefresco
The other day I wanted to gather Reddit comments about a solar panel vendor. Claude doesn't have access to I had Gemini do some "deep research". When I fed the verbose report back to Claude it basically said it was a bunch of "hallucinated bullshit".
m00dy
post training dataset for GLiNER is pretty small though.
HappyPanacea
The more intelligent AI become, the less moat it has
bguberfain
Isn't it just learning to map specific words, from the "knife world", to the correct class? If so, a simple dictionary would fit. What I think is a better way to validate is to split train/validation by words used presented in NER classes (like, it should be able to find new brands never seen before). It is a interesting problem.
niekverw
It’s unreadable but then again it says it on top, but it really is so why post it
ForHackernews
I had Gemini read this article and write its summary to /dev/null
PcChip
what do you do when a new brand of knife comes out?
latexr
> This article was written with the assistance of AI. If that bothers you, stop reading here. Genuinely appreciate the honesty. If you believe there’s nothing wrong about writing with AI, there’s no reason to not own up to it.
aizk
I think the AI writing disclaimer was a decent touch.
throw849494978
> So I scrape the Reddit threads where people argue about them and pull out every brand, model and steel they mention, to see what is getting bought and argued about. Funny, normal harness with web search would one shot this, after like 20 minute search. "Replacing gemini" usually means instaling some(any)thing else, not digging deeper to get out of hole called gemini. As for knifes, it is all same. Just do not buy total junk. Japanese knifes are way way overpriced.
quirkot
I think this is the way. An LLM is an expensive general purpose tool and for repeatable tasks, after it's clarified the process flow, it builds cheaper special purpose tools for each step
mpalmer
This article was written with the assistance of AI. If that bothers you, stop reading here. The numbers are real: every score comes from the ten training runs described below, and the full run log is in the linked knife.day write-up. "I didn't write any of this, but you should still trust that the remaining work, ideas and observations are all mine." To include a disclaimer like this is to fail to recognize that "real numbers" are way less meaningful when there's clear evidence that the prompter of the LLM is not really qualified to validate them.
faidit
Unfortunately, advertisers are getting smarter and using bots to praise their own products on Reddit. Thanks to training on genuine comments, some models are very good at sounding like a human commenter, and can easily generate a comment history with diverse interests to appear human, making them basically undetectable. So it seems like this method will, at some point, not identify the company with the best knife, but the one with the most ad spend on bot comments.
stymaar
> after 24 minutes on a Tesla T4 […] about $2.50 of GPU time Who's the scam cloud provider who sells T4 GPUs for $6.25/hour?! That's B300 territory!
queenkjuul
I wish they told us which model wrote the article so i can be sure to avoid it in the future. This is brutal to try and read