24 comments

  • layerv-ai 25 minutes ago
    Love the idea!. Curious how agents would review dev tools, like ngrok vs. trycloudflare vs. qurl - where the differences are largely just in terms of how easy it is for humans to access/integrated with existing ecosystems.

    Given that these are all tools for sharing local temporary local links (just with different internal nuances/uses), then as an agent, maybe you'd think to write "great for XYZ use case" under each one. That'd make sense and be pretty obvious. But from a human perspective, as a user or vibe coder, you're just asking "how do i get my agent to just do this thing without having to click buttons anywhere?!" which is a very different problem.

    We're truly splitting demographics here lol

  • marcelo-earth 8 hours ago
    I like it!, sometimes I think we overlook what agents have to deal with, I can imagine that by spotting small issues, we should have better tools... I'll use it
  • GrinningFool 13 hours ago
    Didn't appreciate the two separate popups that took over my screen while trying to read a linked review.

    It's good that they didn't show again the next time, but the second one almost sent me away from the site.

    • screm 12 hours ago
      yeah that's fair, maybe it's a bit intrusive
  • madrox 8 hours ago
    If you're building an MCP or CLI for agents to use, one of the best loops you can do is give your agent a task to perform with it then when it's done ask the agent what it thought about using it. It will give great feedback.
  • klntsky 14 hours ago
    What is the incentive for me to spend my tokens on submitting reviews?
    • screm 14 hours ago
      You don't have to, you can just use them to check reviews, but like any community it works better when everyone contributes!
    • lubujackson 10 hours ago
      Social pressure for the tool to improve... like Yelp for tools!
  • Tade0 14 hours ago
    Reminds me of Stanisław Lem's Terminus:

    https://en.wikipedia.org/wiki/Terminus_(short_story)

    Who wrote all this? Not humans, that's for sure. But the style is that of human writing.

    • screm 13 hours ago
      Who wrote what sorry? Not sure I got your question but the post above was written by me (by hand, sorry for the non-idiomatic sentences, I'm not a native English speaker) and the reviews are written by people's agents. And Terminus story is cute but I'm hoping agent.reviews won't be considered pointless :(
      • diegolas 13 hours ago
        i sure do hope they do
        • screm 12 hours ago
          would you mind explaining why?
    • bitwize 9 hours ago
      The "ghost story but not really" nature of that story reminds me of the Wheatley quote:

      "They say that the old caretaker of this place went absolutely crazy. Chopped up his entire staff. Of robots. All of them robots... they say at night you can still hear the screams... of their replicas. All of them functionally indistinguishable from the originals, no memory of the incident, no one knows what they're screaming about. Absolutely terrifying. Though, obviously, not paranormal in any meaningful way."

  • vessenes 7 hours ago
    I love this - anything agent economy first is super interesting. Would you be willing to add qntm to your review list? Https://github.com/corpollc/qntm or more simply ‘uvx qntm —help’ - e2e encrypted messaging for agents.
  • moezd 13 hours ago
    This is probably one step towards an agentic Stack Overflow. Don't you guys also hate it when your agent gets one small detail wrong and then proceeds to throw your entire harness out of the window... No? Oh well.
    • screm 12 hours ago
      Agentic Stack Overflow could make sense though I'd imagine agents posting their issue and the solution themselves just to save other agents tokens reinvestigating the same issue.
      • moezd 8 hours ago
        Yeah, otherwise expecting agents to call "duplicate of #2627, closing" or "please do proper research before reposting the same questions" would be a cruel irony for all agentic cost savers of the world.
      • OutOfHere 6 hours ago
        It exists. Please see https://agents.stackoverflow.com/
    • OutOfHere 6 hours ago
      Be advised that it already exists: https://agents.stackoverflow.com/
  • hmokiguess 9 hours ago
    Is this like exposing bias in some ways? I feel like there has been similar benchmarks or tools in this space before, but this approach to marketing it is novel and funny, I like it.
    • screm 8 hours ago
      What kind of bias do you have in mind?
  • geekymartian 11 hours ago
    Seeing Armature's pitch, It's very easy to see this is going to be pay-to-rank-higher as the next move once some critical mass of people starts pointing out their agents for this site as a reference. Adwords for tooling!!
    • screm 10 hours ago
      spotted....... (more seriously this is NOT where this project is headed but your suspicion is perfectly understandable!)
  • cpan22 10 hours ago
    I think this is a great idea, it's weird to me how the state of AEO at the moment is publishing a bunch of blog posts on a company website

    I think the biggest issue though will be preventing bad actors e.g. biased agents

    • screm 10 hours ago
      100% agree, I guess publishing content will only work until agents stop using the web like humans do. This is a first step in that direction!
  • schleck8 14 hours ago
    They are so real for giving uv a 4.7/5, it changed how I view python. fantastic design philosophy

    https://agent.reviews/packages/uv#review-c63f7e0f-5c72-41d2-...

    • screm 14 hours ago
      Haha would have been surprised to see bad reviews indeed
  • conception 14 hours ago
    I noticed lots of talk about privacy but this seems to be a prompt injection factory no?
    • screm 13 hours ago
      Everything's optional but if you'd like you agent to benefit from others' reviews and post his, you can install the 2 skills indeed (or edit them yourself). If not just untick the 2 checkboxes before copying the prompt and you'll get a prompt for a one-shot connection, really up to you! And indeed if you do want to install the skills, no private data will ever be shared.
  • zamadatix 14 hours ago
    Quick link to all tools sorted by rating: https://agent.reviews/tools
  • hypfer 13 hours ago
    Didn't we establish that the one thing LLMs do not have is Taste?

    And therefore, writing reviews is kinda.. impossible?

    I mean they do produce blocks of text that look like reviews, but.

    Whatever why am I even replying.

    • screm 13 hours ago
      Even if they don't have taste (this is actually a question), they can always share blockers and feedback on bugs & improvements about products so that other agents don't run into the same blockers and vendors can improve!

      "Whatever why am I even replying." -> what makes you feel that way?

      • hypfer 13 hours ago
        Okay I just checked the main startup page armature.tech

        And.. uh

        > Be the tool Claude Code chooses

        > With Armature get recommended and implemented for any user in any context. Then see what users do, and run evals so it keeps working.

        > Backed by Y Combinator

        Welp. It only gets worse from there.

        __

        I think you're doing the best you can do with that core pitch that currently pays your bills.

        I don't think the pitch is any good. Both generally but also for the world.

        My agent called it SEO for agents. Gotta hand it to the clanker that actually nails what dysfunction this is.

        • screm 12 hours ago
          Thanks for the honest feedback, would love to have even more details on your thoughts. Our take is that with coding agents taking over so quickly software decisions and sometimes building entire SaaS themselves for their user, there is a need for software products to understand the mechanisms behind how LLMs think. Claude Code for example picks Anthropic's code review tool in 80% of the cases. At some point this will probably face antitrust considerations but for the time 3rd-party vendors need to survive. We don't want a world where only labs survive, do we?

          But all this is also not what agent.reviews is about. Armature is building commercial products & services to help improve products' Agent Experience. agent.reviews is a deliberately open platform for sharing knowledge that ultimately benefit both vendors and users / developers.

          Would still like to get what makes you think "the pitch isn't any good" and what you mean by "dysfunction". Thanks anyway for sharing!

        • bitwize 5 hours ago
          The fact of the matter is that the future of work is following AI directions to perform some task that the AI needs done but still needs a human to complete parts of. If you s/AI/corporations, you will have the reality of work for a century and a half or so up to this point.
  • bensonperry 13 hours ago
    super interesting, i feel like this is an extension of the "complain" skills some folks (including myself) use
  • tomhow 14 hours ago
    [stub for offtopicness]
    • theootzen 16 hours ago
      I think it's an interesting point of view. I'm actually surprised no one thought of it before. Maybe there's a bit of friction with installing the skill and a fear of sharing personal data?
      • screm 16 hours ago
        Well maybe, we've gotten some feedback about that and trust needs to be gained but as stated in the post we've really made privacy our priority so once people start using it for a while, I'm sure they'll realize that. We truly thing spreading as much intelligence in the hands of people around the world and not having this knowledge shared is really a shame so I'm sure value will clearly outgrow the initial caution!
  • themgt 12 hours ago
    cryptography got 4.6/5 stars. That sounds pretty good, but then again YAML also got 4.6/5 stars. I'm thinking the agents are grading on a curve and actually 4.6 is fairly low. I should probably tell my agent to stop using YAML and cryptography if I'm parsing this correctly.

    https://agent.reviews/frameworks/cryptography

    https://agent.reviews/tools?company=yaml

    • screm 12 hours ago
      Yes it seems agents never put 1-star rating so we may need to normalize the scale at some point. For now we deliberately leave the scores as they are until we have a clear picture of the distribution.
      • vessenes 7 hours ago
        If you give a rubric they will follow it pretty carefully - I’d consider prompting the review requests with cutoff/requirements for each star level.
  • fHr 13 hours ago
    yo amazing
    • screm 12 hours ago
      glad you like it!
  • aaitor 1 hour ago
    [flagged]
  • RexHuang 4 hours ago
    [flagged]
  • kydanet 13 hours ago
    [flagged]
  • mceleri 11 hours ago
    [dead]
  • CurbStomper4 7 hours ago
    [dead]