• Zetta@mander.xyz
      link
      fedilink
      English
      arrow-up
      7
      arrow-down
      6
      ·
      10 days ago

      I love open Chinese llms but this seems 100% plausible and and I think ur wring. This is in line with how Chinese companies need to work around the system since America tries to handicap them. This is actually a very smart way to get high quality trianjng data for their future models. Unfortunate about leaking private customer data to a us company though.

      Also the article and the evidence presented by Antropic all seems very plausible.

    • Womble@piefed.world
      link
      fedilink
      English
      arrow-up
      3
      arrow-down
      10
      ·
      10 days ago

      You almost certainly weren’t using a Deepseek model offline (unless you have 100s of GB of RAM/VRAM). You might have been using a small model distilled on Deepseek responses, but the smallest model made by Deepseek is about a 300B param model.

  • francisco_1844@discuss.online
    link
    fedilink
    English
    arrow-up
    10
    arrow-down
    1
    ·
    10 days ago

    What this really boils down to is that Anthropic and OpenAI hope to be the only ones that use someone else’s copyrighted information without any legal or financial complications.

  • rmrf@lemmy.ml
    link
    fedilink
    English
    arrow-up
    5
    ·
    9 days ago

    You’re telling me that is routing requests through a model 70x the price?