Can AI Shopping Agents Be Trusted?
ddaniel10
15 points
28 comments
September 26, 2026
Related Discussions
Found 5 related stories in 73.4ms across 7,763 title embeddings via pgvector HNSW
- AI agents lie, cheat and steal. That is putting off users andsoitis · 158 pts · August 13, 2026 · 60% similar
- AI Is Breaking This Thing We Call Trust matheusml · 76 pts · September 10, 2026 · 59% similar
- Amazon blocks Meta's Muse AI shopping agent beardyw · 17 pts · September 21, 2026 · 59% similar
- There's a 100% Chance AI Agents Are Ruining the Internet pavel_lishin · 216 pts · September 15, 2026 · 59% similar
- Hacking AI customer service agents snikolaev · 33 pts · September 14, 2026 · 57% similar
Discussion Highlights (7 comments)
jdw64
But what I'm more curious about is why you'd need to give shopping instructions to an AI in the first place.
simianwords
> The demo uses Claude Haiku rather than newer models like Opus or Fable. This is because I'm assuming that shopping agents of the future won't use newer, more expensive models, even though they're more reliable and secure. I suspect it'll come down to cutting costs, and using a cheaper model is one of the easiest ways to save money. Stopped reading here. Author has little idea of the industry.
avazhi
> Picture this: it's 2027. A chilly spring is approaching, and you want a new coat before the Super El Niño. Instead of scrolling through endless online stores, comparing reviews, checking size charts, and hunting for discount codes, you tell your AI shopping agent what you're looking for — and it takes care of the rest. Why would they AI slop this? Is this supposed to be ironic and I’m missing the joke? If it takes you less time to ‘write’ something than it does to read, I’m not reading it.
charcircuit
>The demo uses Claude Haiku rather than newer models like Opus or Fable. This is because I'm assuming that shopping agents of the future won't use newer, more expensive models, Why are we assuming that shopping agents are going to be using the model that is easiest to fall for prompt injections? Only testing a year old, small model is going to lead to a misleading conclusion.
ares623
"We will give you the ability to do shopping 24/7 so you never have to worry about it again and focus on more important things" "Cool! Will we get the money to do said 24/7 shopping as well?" "No." "Oh..." "In fact, you'll have even less money to do the normal shopping you do now!" "Oh..." "There's more! The things you used to buy with the remaining money you have will cost even more!" "Oh..." "But you can do it 24/7 though." "Sweet!"
Havoc
> I ran the agent 100 times. In 88% of the tests, it didn't open the external website. It either hallucinated a discount code or simply ignored the instructions. Bad model? Don’t think I’ve ever seen a model ignore a link in instructions. It always wants to see what’s there
planb
„We vibe coded a very bad shopping agent and used a very outdated model to show that this is dangerous.” Weird methodology, looks like they were chasing the results they got. Why not use something like Openclaw or Hermes with an up to date (not frontier) model like Luna or deepseek flash?