3 comments

  • brcmthrowaway5 hours ago
    How about learning the internals of LLMs, is Sebastian Raschka's content still the best in 2026?
  • lazy_dev_1_to_947 minutes ago
    [flagged]
  • petcat3 hours ago
    &gt; Open-Source AI<p>There are no open source AI models, at least not useful ones (yet [1]). Open weight is not the same as open source. &quot;Open weight&quot; models are still just inscrutable binary blobs that you can (theoretically) run on your own computer instead of through a SAAS web app. The open weight model labs don&#x27;t even provide a high level catalog or any description whatsoever about what went into the training data.<p>This is not open source and we should stop conflating the two things.<p>[1] <a href="https:&#x2F;&#x2F;allenai.org&#x2F;" rel="nofollow">https:&#x2F;&#x2F;allenai.org&#x2F;</a>
    • vlyan2 hours ago
      &gt;The open weight model labs don&#x27;t even provide a high level catalog or any description whatsoever about what went into the training data.<p>because the training data is full of copyrighted works.<p>the answer to &quot;what went into the training data&quot; is &quot;everything we could get our hands on&quot;.
    • kennywinker53 minutes ago
      K2 horizon is also fully open source i believe - <a href="https:&#x2F;&#x2F;ifm.ai&#x2F;k2&#x2F;" rel="nofollow">https:&#x2F;&#x2F;ifm.ai&#x2F;k2&#x2F;</a><p>I understand your quibble with terminology, but i think the “inscrutable binary blob” thing is a bit off base. You can create finetunes and post train models using only their open weights. You can’t create derivative works like that from an inscrutable binary blob
    • Gracana3 hours ago
      I think you would be pleasantly surprised by the content of the linked article.
    • foopod2 hours ago
      Completely agree with this, I&#x27;m sick of people conflating the two. Open-weight models should be treated no more favourably than proprietary freeware.<p>Sure you can run tests and benchmarks on open-weight models, but that is the extent - there is no scrutiny, no auditing for bias or copyright contamination - just a black box that you rely on for &quot;intelligence&quot;. I&#x27;m still shocked the way people can hand over not just huge swathes of data, but also decisions of all shapes and sizes - to AI companies with no way of being able to assess how the sausage is made.
      • kennywinker59 minutes ago
        You can create derivative works from open weight models
    • markasoftware3 hours ago
      I thought Nemotron tried to be pretty open?
      • nessex1 hour ago
        Many of the Nemotron datasets are gated behind approval, and a license agreement.<p>The preamble on these datasets is: &quot;This repository is publicly accessible, but you have to accept the conditions to access its files and contents.&quot;<p>I don&#x27;t know what others have experienced, but I requested access to multiple Nemotron datasets and those requests were ignored for months before all but one request was rejected. There&#x27;s no explanation for why, nor anything I can see which would lead to a rejection. So it&#x27;s purely anecdotal and YMMV, but I don&#x27;t see these as being particularly open.