4 comments

  • sean_pedersen38 minutes ago
    Good project but this one also exists <a href="https:&#x2F;&#x2F;huggingface.co&#x2F;spaces&#x2F;multimodalart&#x2F;jev-decision-index" rel="nofollow">https:&#x2F;&#x2F;huggingface.co&#x2F;spaces&#x2F;multimodalart&#x2F;jev-decision-ind...</a> and the results do not seem to add up and also model sets are different... still needs time to mature likely
  • swyx1 hour ago
    jev ceo on why he eschewed benchmarking: <a href="https:&#x2F;&#x2F;www.latent.space&#x2F;i&#x2F;216783460&#x2F;privacy-benchmarking-and-trusting-intelligence" rel="nofollow">https:&#x2F;&#x2F;www.latent.space&#x2F;i&#x2F;216783460&#x2F;privacy-benchmarking-an...</a>
    • arbot3609 minutes ago
      Many SaaS vendors forbid benchmarking, I find it crazy that such anti-competitive terms are standard across the industry but they are. Generally the goal of such terms is to &quot;control the narrative&quot; around the product, regardless of the truth of performance being better or worse than competitors.
    • jldugger37 minutes ago
      Interesting; was curious how this didn&#x27;t fall into trouble with ToS. Apparently the &quot;no benchmarks&quot; clause was intended for &quot;limited preview&quot; audiences and didn&#x27;t get removed at launch on accident.
  • nzoschke30 minutes ago
    <a href="https:&#x2F;&#x2F;is-it-ai-slop.app.mintapis.com&#x2F;" rel="nofollow">https:&#x2F;&#x2F;is-it-ai-slop.app.mintapis.com&#x2F;</a> is a fun tool. Is the source or methodology for that in the github repo? I couldn&#x27;t find it immediately.<p>We&#x27;ve been experimenting with Jev for classifying email, some thoughts here: <a href="https:&#x2F;&#x2F;housecat.com&#x2F;blog&#x2F;classifying-email" rel="nofollow">https:&#x2F;&#x2F;housecat.com&#x2F;blog&#x2F;classifying-email</a><p>Flagging AI written email is a much requested feature too.