Building Food Metadata with LLM Juries
tie-in
25 points
7 comments
July 14, 2026
Related Discussions
Found 5 related stories in 60.3ms across 5,215 title embeddings via pgvector HNSW
- Humanising LLM Outputs Is Dumb kuberwastaken · 18 pts · August 10, 2026 · 48% similar
- The efficient frontier of LLM inference philipkiely · 83 pts · September 01, 2026 · 48% similar
- Can a MUD evaluate LLMs? A $99 proof of concept Davisb135 · 101 pts · July 22, 2026 · 47% similar
- The LLM Critics Are Right. I Use LLMs Anyway JeremyTheo · 209 pts · July 16, 2026 · 46% similar
- GenRec: Towards LLM-Native Recommendation at Netflix Anon84 · 32 pts · August 15, 2026 · 45% similar
Discussion Highlights (3 comments)
sigmar
>Evaluators validate each tag individually — for example, protein, preparation, or health, individually rather than judging the item as a whole. Am I reading this right that the jury is multiple LLMs each iterating through each tag and voting on each? Why wouldn't you tune one LLM to be really competent at a single tag? Like a single "spicy evaluator LLM" or "protein evaluator LLM"?
TeeWEE
Basically it’s AI on top of AI for metadata extraction. There are a lot of claims in the article but not a lot of hard data. In the end they still don’t know if the data is correct. Good luck with your glutes allergy. The weird thing for me is the prompt optimization loop? Why not fine tune the model instead of AI generating the prompt?
vector_spaces
I am sorry to be harsh but I find it amateurish that they would use an AI generated hero image for this and presumably fabricated LLM output -- fabricated by an AI image generator no less Whenever I create an image like this for the purpose of a demo, I make certain that it demonstrates either real input/output or at least is exemplary of real input/output because the whole point is to instill confidence in the tool. Sure, if the raw outputs aren't clean/comprehensible enough for presenting to stakeholders or others, fine, clean them up to make them comprehensible or add explainers, but there shouldn't be any need to fabricate the inputs. I feel obligated to respond to the hypothetical "But they don't want to tie it to a particular restaurant or brand" -- you don't have to! Doordash has taken generic food photos for this exact purpose.