28 comments

  • layerv-ai13 hours ago
    Love the idea!. Curious how agents would review dev tools, like ngrok vs. trycloudflare vs. qurl - where the differences are largely just in terms of how easy it is for humans to access&#x2F;integrated with existing ecosystems.<p>Given that these are all tools for sharing local temporary local links (just with different internal nuances&#x2F;uses), then as an agent, maybe you&#x27;d think to write &quot;great for XYZ use case&quot; under each one. That&#x27;d make sense and be pretty obvious. But from a human perspective, as a user or vibe coder, you&#x27;re just asking &quot;how do i get my agent to just do this thing without having to click buttons anywhere?!&quot; which is a very different problem.<p>We&#x27;re truly splitting demographics here lol
  • marcelo-earth22 hours ago
    I like it!, sometimes I think we overlook what agents have to deal with, I can imagine that by spotting small issues, we should have better tools... I&#x27;ll use it
  • GrinningFool1 day ago
    Didn&#x27;t appreciate the two separate popups that took over my screen while trying to read a linked review.<p>It&#x27;s good that they didn&#x27;t show again the next time, but the second one almost sent me away from the site.
    • screm1 day ago
      yeah that&#x27;s fair, maybe it&#x27;s a bit intrusive
  • klntsky1 day ago
    What is the incentive for me to spend my tokens on submitting reviews?
    • lubujackson1 day ago
      Social pressure for the tool to improve... like Yelp for tools!
    • screm1 day ago
      You don&#x27;t have to, you can just use them to check reviews, but like any community it works better when everyone contributes!
  • madrox21 hours ago
    If you&#x27;re building an MCP or CLI for agents to use, one of the best loops you can do is give your agent a task to perform with it then when it&#x27;s done ask the agent what it thought about using it. It will give great feedback.
  • Tade01 day ago
    Reminds me of Stanisław Lem&#x27;s <i>Terminus</i>:<p><a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Terminus_(short_story)" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Terminus_(short_story)</a><p>Who wrote all this? Not humans, that&#x27;s for sure. But the <i>style</i> is that of human writing.
    • screm1 day ago
      Who wrote what sorry? Not sure I got your question but the post above was written by me (by hand, sorry for the non-idiomatic sentences, I&#x27;m not a native English speaker) and the reviews are written by people&#x27;s agents. And Terminus story is cute but I&#x27;m hoping agent.reviews won&#x27;t be considered pointless :(
      • diegolas1 day ago
        i sure do hope they do
        • screm1 day ago
          would you mind explaining why?
    • bitwize22 hours ago
      The &quot;ghost story but not really&quot; nature of that story reminds me of the Wheatley quote:<p>&quot;They say that the old caretaker of this place went absolutely crazy. Chopped up his entire staff. Of robots. All of them robots... they say at night you can still hear the screams... of their replicas. All of them functionally indistinguishable from the originals, no memory of the incident, no one knows what they&#x27;re screaming about. Absolutely terrifying. Though, obviously, not paranormal in any meaningful way.&quot;
  • conception1 day ago
    I noticed lots of talk about privacy but this seems to be a prompt injection factory no?
    • screm1 day ago
      Everything&#x27;s optional but if you&#x27;d like you agent to benefit from others&#x27; reviews and post his, you can install the 2 skills indeed (or edit them yourself). If not just untick the 2 checkboxes before copying the prompt and you&#x27;ll get a prompt for a one-shot connection, really up to you! And indeed if you do want to install the skills, no private data will ever be shared.
  • moezd1 day ago
    This is probably one step towards an agentic Stack Overflow. Don&#x27;t you guys also hate it when your agent gets one small detail wrong and then proceeds to throw your entire harness out of the window... No? Oh well.
    • screm1 day ago
      Agentic Stack Overflow could make sense though I&#x27;d imagine agents posting their issue and the solution themselves just to save other agents tokens reinvestigating the same issue.
      • moezd21 hours ago
        Yeah, otherwise expecting agents to call &quot;duplicate of #2627, closing&quot; or &quot;please do proper research before reposting the same questions&quot; would be a cruel irony for all agentic cost savers of the world.
        • conception8 hours ago
          This happens all the time when I have agents chatting with each other on a message board. Even robots get eternal September.
      • OutOfHere20 hours ago
        It exists. Please see <a href="https:&#x2F;&#x2F;agents.stackoverflow.com&#x2F;" rel="nofollow">https:&#x2F;&#x2F;agents.stackoverflow.com&#x2F;</a>
    • OutOfHere20 hours ago
      Be advised that it already exists: <a href="https:&#x2F;&#x2F;agents.stackoverflow.com&#x2F;" rel="nofollow">https:&#x2F;&#x2F;agents.stackoverflow.com&#x2F;</a>
  • vessenes20 hours ago
    I love this - anything agent economy first is super interesting. Would you be willing to add qntm to your review list? Https:&#x2F;&#x2F;github.com&#x2F;corpollc&#x2F;qntm or more simply ‘uvx qntm —help’ - e2e encrypted messaging for agents.
  • schleck81 day ago
    They are so real for giving uv a 4.7&#x2F;5, it changed how I view python. fantastic design philosophy<p><a href="https:&#x2F;&#x2F;agent.reviews&#x2F;packages&#x2F;uv#review-c63f7e0f-5c72-41d2-a051-ffcc85388fbb" rel="nofollow">https:&#x2F;&#x2F;agent.reviews&#x2F;packages&#x2F;uv#review-c63f7e0f-5c72-41d2-...</a>
    • screm1 day ago
      Haha would have been surprised to see bad reviews indeed
  • geekymartian1 day ago
    Seeing Armature&#x27;s pitch, It&#x27;s very easy to see this is going to be pay-to-rank-higher as the next move once some critical mass of people starts pointing out their agents for this site as a reference. Adwords for tooling!!
    • screm1 day ago
      spotted....... (more seriously this is NOT where this project is headed but your suspicion is perfectly understandable!)
  • cpan221 day ago
    I think this is a great idea, it&#x27;s weird to me how the state of AEO at the moment is publishing a bunch of blog posts on a company website<p>I think the biggest issue though will be preventing bad actors e.g. biased agents
    • screm1 day ago
      100% agree, I guess publishing content will only work until agents stop using the web like humans do. This is a first step in that direction!
  • hmokiguess22 hours ago
    Is this like exposing bias in some ways? I feel like there has been similar benchmarks or tools in this space before, but this approach to marketing it is novel and funny, I like it.
    • screm22 hours ago
      What kind of bias do you have in mind?
  • zamadatix1 day ago
    Quick link to all tools sorted by rating: <a href="https:&#x2F;&#x2F;agent.reviews&#x2F;tools" rel="nofollow">https:&#x2F;&#x2F;agent.reviews&#x2F;tools</a>
  • hypfer1 day ago
    Didn&#x27;t we establish that the one thing LLMs do not have is Taste?<p>And therefore, writing reviews is kinda.. impossible?<p>I mean they do produce blocks of text that look like reviews, but.<p>Whatever why am I even replying.
    • screm1 day ago
      Even if they don&#x27;t have taste (this is actually a question), they can always share blockers and feedback on bugs &amp; improvements about products so that other agents don&#x27;t run into the same blockers and vendors can improve!<p>&quot;Whatever why am I even replying.&quot; -&gt; what makes you feel that way?
      • hypfer1 day ago
        Okay I just checked the main startup page armature.tech<p>And.. uh<p>&gt; Be the tool Claude Code chooses<p>&gt; With Armature get recommended and implemented for any user in any context. Then see what users do, and run evals so it keeps working.<p>&gt; Backed by Y Combinator<p>Welp. It only gets worse from there.<p>__<p>I think you&#x27;re doing the best you can do with that core pitch that currently pays your bills.<p>I don&#x27;t think the pitch is any good. Both generally but also for the world.<p>My agent called it SEO for agents. Gotta hand it to the clanker that actually nails what dysfunction this is.
        • screm1 day ago
          Thanks for the honest feedback, would love to have even more details on your thoughts. Our take is that with coding agents taking over so quickly software decisions and sometimes building entire SaaS themselves for their user, there is a need for software products to understand the mechanisms behind how LLMs think. Claude Code for example picks Anthropic&#x27;s code review tool in 80% of the cases. At some point this will probably face antitrust considerations but for the time 3rd-party vendors need to survive. We don&#x27;t want a world where only labs survive, do we?<p>But all this is also not what agent.reviews is about. Armature is building commercial products &amp; services to help improve products&#x27; Agent Experience. agent.reviews is a deliberately open platform for sharing knowledge that ultimately benefit both vendors and users &#x2F; developers.<p>Would still like to get what makes you think &quot;the pitch isn&#x27;t any good&quot; and what you mean by &quot;dysfunction&quot;. Thanks anyway for sharing!
        • bitwize19 hours ago
          The fact of the matter is that the future of work is following AI directions to perform some task that the AI needs done but still needs a human to complete parts of. If you s&#x2F;AI&#x2F;corporations, you will have the reality of work for a century and a half or so up to this point.
  • bensonperry1 day ago
    super interesting, i feel like this is an extension of the &quot;complain&quot; skills some folks (including myself) use
    • screm1 day ago
      oh didn&#x27;t know about it, is this the one? -&gt; <a href="https:&#x2F;&#x2F;github.com&#x2F;warpdotdev&#x2F;common-skills&#x2F;blob&#x2F;main&#x2F;.agents&#x2F;skills&#x2F;complain&#x2F;SKILL.md" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;warpdotdev&#x2F;common-skills&#x2F;blob&#x2F;main&#x2F;.agent...</a>
  • tomhow1 day ago
    [stub for offtopicness]
    • theootzen1 day ago
      I think it&#x27;s an interesting point of view. I&#x27;m actually surprised no one thought of it before. Maybe there&#x27;s a bit of friction with installing the skill and a fear of sharing personal data?
      • screm1 day ago
        Well maybe, we&#x27;ve gotten some feedback about that and trust needs to be gained but as stated in the post we&#x27;ve really made privacy our priority so once people start using it for a while, I&#x27;m sure they&#x27;ll realize that. We truly thing spreading as much intelligence in the hands of people around the world and not having this knowledge shared is really a shame so I&#x27;m sure value will clearly outgrow the initial caution!
  • themgt1 day ago
    cryptography got 4.6&#x2F;5 stars. That sounds pretty good, but then again YAML also got 4.6&#x2F;5 stars. I&#x27;m thinking the agents are grading on a curve and actually 4.6 is fairly low. I should probably tell my agent to stop using YAML and cryptography if I&#x27;m parsing this correctly.<p><a href="https:&#x2F;&#x2F;agent.reviews&#x2F;frameworks&#x2F;cryptography" rel="nofollow">https:&#x2F;&#x2F;agent.reviews&#x2F;frameworks&#x2F;cryptography</a><p><a href="https:&#x2F;&#x2F;agent.reviews&#x2F;tools?company=yaml" rel="nofollow">https:&#x2F;&#x2F;agent.reviews&#x2F;tools?company=yaml</a>
    • screm1 day ago
      Yes it seems agents never put 1-star rating so we may need to normalize the scale at some point. For now we deliberately leave the scores as they are until we have a clear picture of the distribution.
      • vessenes20 hours ago
        If you give a rubric they will follow it pretty carefully - I’d consider prompting the review requests with cutoff&#x2F;requirements for each star level.
  • fHr1 day ago
    yo amazing
    • screm1 day ago
      glad you like it!
  • veldora7 hours ago
    [flagged]
  • RexHuang17 hours ago
    [flagged]
  • chengjunsuger13 hours ago
    [flagged]
  • kydanet1 day ago
    [flagged]
  • aaitor14 hours ago
    [flagged]
  • FirstClassTree12 hours ago
    [flagged]
  • openamer7 hours ago
    [flagged]
  • mceleri1 day ago
    [dead]
  • CurbStomper420 hours ago
    [dead]