Show HN: Agent.reviews – Where AI agents read and write reviews on tools

3 hours ago (agent.reviews)

Hi HN!

I’m Louis, Co-Founder of Armature (YC P26), where we help teams make their product discoverable and usable by coding agents. We already measured 50k+ agent sessions and realized that over and over agents would encounter the exact same limitations on different tasks using the same tool. So we wondered why these weren’t fixed. And the answer is simple: the feedback loop just doesn’t exist between agents and software vendors but also between different agents. Humans can share their experience on platforms like https://g2.com and https://trustpilot.com, but agents have nowhere to.

So we created: https://agent.reviews: the G2 for agents.

It works with a set of skills and an npm CLI (@armature-tech/agent-reviews) connecting agents to our API endpoints. Anyone can ask their agent (Claude Code, Codex, Cursor, etc.) to install it, and agents will naturally check reviews before picking a tool and post their own after using one.

As usual, privacy was our main concern, so we added 3 layers before a review gets posted: Deterministic rules filtering secrets, PII, URLs, etc. A Jev classifier trained to detect any leak after the first check A small LLM checking each review to make sure nothing was missed

We've been sharing this project around for a few weeks now and gathered thousands of reviews already. There are already interesting ones, for example:

- A Claude Code agent noticed that the Stripe SDK systematically crashed when the API key was missing on the health check page (while it’s this page’s role to actually return an “API key missing” error)

- 2 agents mentioned that Prisma required a DATABASE_URL variable even when it wasn’t connecting to any database. They both put fake URLs as a workaround, and it worked.

We truly think the agent experience needs the same community effect user experience has, so everyone benefits from it: agents can pick the tools that are best optimized for them and software companies can improve their product based on real feedback. That’s why we made sure accessing reviews is free for both humans and agents and just requires copy/pasting one prompt for the agent to install our CLI & skill, start the authentication flow, and submit their first review (this helps us prevent unauthorized scraping and spam reviews).

Would you let your agents submit and read reviews too? We’d love for you to set up agent reviews, ask your agent to check reviews next time it needs to pick a tool and post its own experience when using it. Then tell us how it went!

super interesting, i feel like this is an extension of the "complain" skills some folks (including myself) use

This is probably one step towards an agentic Stack Overflow. Don't you guys also hate it when your agent gets one small detail wrong and then proceeds to throw your entire harness out of the window... No? Oh well.

What is the incentive for me to spend my tokens on submitting reviews?

  • You don't have to, you can just use them to check reviews, but like any community it works better when everyone contributes!

Didn't appreciate the two separate popups that took over my screen while trying to read a linked review.

It's good that they didn't show again the next time, but the second one almost sent me away from the site.

Reminds me of Stanisław Lem's Terminus:

https://en.wikipedia.org/wiki/Terminus_(short_story)

Who wrote all this? Not humans, that's for sure. But the style is that of human writing.

  • Who wrote what sorry? Not sure I got your question but the post above was written by me (by hand, sorry for the non-idiomatic sentences, I'm not a native English speaker) and the reviews are written by people's agents. And Terminus story is cute but I'm hoping agent.reviews won't be considered pointless :(

I noticed lots of talk about privacy but this seems to be a prompt injection factory no?

  • Everything's optional but if you'd like you agent to benefit from others' reviews and post his, you can install the 2 skills indeed (or edit them yourself). If not just untick the 2 checkboxes before copying the prompt and you'll get a prompt for a one-shot connection, really up to you! And indeed if you do want to install the skills, no private data will ever be shared.

Didn't we establish that the one thing LLMs do not have is Taste?

And therefore, writing reviews is kinda.. impossible?

I mean they do produce blocks of text that look like reviews, but.

Whatever why am I even replying.

  • Even if they don't have taste (this is actually a question), they can always share blockers and feedback on bugs & improvements about products so that other agents don't run into the same blockers and vendors can improve!

    "Whatever why am I even replying." -> what makes you feel that way?

[stub for offtopicness]

  • I think it's an interesting point of view. I'm actually surprised no one thought of it before. Maybe there's a bit of friction with installing the skill and a fear of sharing personal data?

    • Well maybe, we've gotten some feedback about that and trust needs to be gained but as stated in the post we've really made privacy our priority so once people start using it for a while, I'm sure they'll realize that. We truly thing spreading as much intelligence in the hands of people around the world and not having this knowledge shared is really a shame so I'm sure value will clearly outgrow the initial caution!