I usually don’t even understand the lingo they use. “Open-weighted” is the most recent one, then it usually goes down to specific “models” that everybody is supposed to know about.

These are my thoughts (I will stick to the vague “it” for now, but of course therein lies another question: “and how does all this apply to various specialised AIs”):

  • Is it really feasible to run it 100% locally? I know there’s plenty of people with very powerful rigs indeed, but still. Or are 99% of these people really saying “it would, in theory, be possible to run that locally, therefore your concerns are invalid”?
  • If yes to the previous: the software doesn’t come from nowhere and ultimately still relies on gas-turbine-powered datacenters and stolen IP and stolen personal data, no?

If what I wrote above is true, what exactly are people arguing when they say it’s still possible to use LLMs ethically or true to FOSS philosophy, because … ???


edit

Thanks to all who answered.

I guess it’s my fault for asking several questions in one, but this thread has attracted exactly the type of people I’m writing about; several even used the term “open-weighted models” without explaining it.

Asking to get arguments explained, I got more arguments instead.

  • leaky_shower_thought@feddit.nl
    link
    fedilink
    English
    arrow-up
    2
    ·
    9 hours ago

    Is it really feasible to run it 100% locally? I know there’s plenty of people with very powerful rigs indeed, but still. Or are 99% of these people really saying “it would, in theory, be possible to run that locally, therefore your concerns are invalid”?

    There’s usually a lot of caveats they don’t mention on this. Like saying it** is feasible to run it** locally! (** it depends on your rig, memory and storage constraints, model licensing constraints blah-di-dah, blah-di-dee)

    Yes, there are super small and simple models so this is technically true.

  • jet@hackertalks.com
    link
    fedilink
    English
    arrow-up
    3
    ·
    18 hours ago

    Computers are used to do math, hosted agents are doing big math for people from far away, but some people only need a little math they can do at home.

  • Pratai@piefed.ca
    link
    fedilink
    English
    arrow-up
    11
    ·
    2 days ago

    There is no argument to use AI ethically. Only bullshit they tell themselves so as to circumvent responsibility.

  • iceberg314@slrpnk.net
    link
    fedilink
    English
    arrow-up
    11
    ·
    2 days ago

    I think it’s at least way more ethical to use locally hosted open weight AI. Mostly because a lot of these models are usually fine tuned and distilled to run on a single GPU and that is the direction the technology needs to develop to be sustainable.

    Obviously for most things it’s way more ethical to just not use AI or LLMs.

    I think people should at least experiment with running local models on their own hardware. It really puts it into perspective how much compute they use. Like the power, speed, and performance ratios of different size LLM and how absurd anything involving generating video is. I have a pretty beefy gaming PC and it took it like 6 minutes to generate 10 seconds of 720p video. And then that time goes up exponentially past 10 seconds. Like 20 minutes for a 15 second video

    Running on a small model is basically like playing a modern full sized video game. When I measure my PC with a power meter

    • mojofrododojo@lemmy.world
      link
      fedilink
      English
      arrow-up
      2
      ·
      16 hours ago

      I think people should at least experiment with running local models on their own hardware.

      show me the one that’s built without theft and I might consider it.

    • hendrik@palaver.p3x.de
      link
      fedilink
      English
      arrow-up
      2
      ·
      1 day ago

      And I can give numbers for people without a graphics card as well. I’m not a gamer so I didn’t bother buying a GPU. It takes my computer 20-40 minutes to generate a single image. Trying to generate video will make Linux step in and kill the process with an Out of Memory error. Generating text works. It’ll “write” a bit slower than I can read. And it takes additional time upfront to ingest my question.

  • flamingo_pinyata@sopuli.xyz
    link
    fedilink
    English
    arrow-up
    12
    ·
    2 days ago

    Let me just answer this part

    Is it really feasible to run it 100% locally? I know there’s plenty of people with very powerful rigs indeed, but still. Or are 99% of these people really saying “it would, in theory, be possible to run that locally, therefore your concerns are invalid”?

    Yes absolutely, and you don’t even need an extremely powerful machine. Basic text generation will work on a macbook pro. If the hardware prices were at normal levels, buying a machine for 2-3k USD/EUR and self-hosting powerful models would be feasible. Right now the same hardware will run you about 10k, that why I don’t think it’s practical at this moment.

    However training is where the real cost lies. It’s pretty much impossible to train a model from 0 on anything resembling a personal machine. First you need all the data, measuring in millions of terabytes. And then you need to train the model on all of that.

    On a side note - I once tried training a GPT2 replica on my laptop with about 3GB of random text. The estimated time was 3 months. That’s the level of requirements we’re talking about.

  • Sunshine@piefed.ca
    link
    fedilink
    English
    arrow-up
    5
    ·
    1 day ago

    When it doesn’t use external disinformation causing data centres and runs from home generated solar power.

  • An llm is just a set of rules for cobbling together text. In and of itself, the technology is ethically neutral. Mind you, all ethics are both situational and relative, but what llms do really isn’t ethically any different from something like how your game handles enemy actions, except in scale.

    What makes them a nightmare is the wonton fuckery involved in how they’re trained. The open sourced ones lack the abusive and nigh maniacal plundering of human effort that the commercial ones do.

    So, when it comes to a locally hosted, open source option, there’s no fundamental difference ethically between an llm, your music player, your word processor, whatever.

    Even data centers for generative ai could conceivably be run ethically, given some massive assumptions regarding resource allocation and usage. Mainly in that the energy it takes, and the physical resources, are being allocated to those centres rather than something else. But if everyone agrees in the benefits, then using renewable resources carefully would make it no better or worse than any of the non ai data centers already in existence. I mean, most VPNs are renting either space or hardware from some company that has a data center. “The cloud” is just a giant computer owned and operated by some company somewhere, and that’s a data center.

    Yeah, generative systems are way more resource intensive, but we can’t pretend that AWS isn’t doing the same basic thing, just with a different use case. Again, there’s the caveat that that details aren’t the same, I’m talking about principles of ethics, nor the specific technology. We’ve decided that building spaces, or filling existing space, for computing is useful enough to be worth the problems it causes. The consensus hasn’t really been reached that generative ai is also worth it, but in ethical terms, the same framework applies.

    So, your concerns are absolutely valid. They are, however, partially met by sustainable implementation

  • hendrik@palaver.p3x.de
    link
    fedilink
    English
    arrow-up
    6
    ·
    edit-2
    1 day ago

    Yes. “open-weights” means they publish the AI model file to the public and people can download and run it themselves.

    And there’s very different AI models out there. Some are smaller, some bigger… Bigger usually means more “intelligent” and you also need more resources to run them. On average, if you own something like a beefy gaming computer, you can do a reasonably “clever” AI at reasonable speed. Something like ChatGPT needs a datacenter, and at the other end we also have some small models which run on some smartphones or average laptops. They surely won’t be able to do the same things ChatGPT does… And if your laptop is slow, it’ll output text very slowly. But maybe you’re okay with less “intelligent” AI.

    And no, they don’t come from nowhere. They’re mainly made by the AI companies. They’re trained the same way all large language models are trained with the same problematic procedure. Oftentimes they publish some more information like a scientific paper alongside. There might be information inside about the energy used / carbon footprint of the training steps. Sometimes/rarely they also give information on what data they used.

    And internet people have all sorts of weird opinions on (FOSS) philosophy and ethics. I think we’d need to be more specific than that. Dirty gas turbines and allowing companies to be exempt from copyright while other people get sued for copying a single movie isn’t really ethical by any means.

    • mojofrododojo@lemmy.world
      link
      fedilink
      English
      arrow-up
      2
      ·
      16 hours ago

      you don’t need to even get into the environmental impacts - what’s it trained on? was that theft?

      that’s where it needs to stop.

      • hendrik@palaver.p3x.de
        link
        fedilink
        English
        arrow-up
        1
        ·
        16 hours ago

        Sure. Though I kinda struggle with those kinds of arguments. Because the Chinese definitely think this kind of theft is alright. And the Americans are a-okay with any kind of ripoff, as long as it’s big nasty companies doing it.

        • mojofrododojo@lemmy.world
          link
          fedilink
          English
          arrow-up
          1
          ·
          2 hours ago

          Though I kinda struggle with those kinds of arguments. Because the Chinese

          decent human beings shouldn’t justify their use of the theft machine by the theft perpetrated by scum.

    • A_norny_mousse@piefed.socialOP
      link
      fedilink
      English
      arrow-up
      2
      ·
      1 day ago

      Thanks.

      Yes. “open-weights” means they publish the AI model file to the public and people can download and run it themselves.

      So “they” are the big corpos involved in AI? I have my hard time wrapping my head around this explanation (although somebody else said the same so it must be true /s). What’s open-weight supposed to mean here? And, usually the thing you get for free is inferior to the paid vrsion - how does that work?

      • hendrik@palaver.p3x.de
        link
        fedilink
        English
        arrow-up
        2
        ·
        edit-2
        1 day ago

        It’s big US companies like Meta (who own Facebook, WhatsApp etc), Google, Chinese companies like DeepSeek, Alibaba… Some other various startups and AI companies or spinoffs. And the occasional university or research institute. There’s a select few European ones. And Nvidia (who sell the graphics cards and hardware), they do research and publish stuff as well.

        I think “open-weights” has been coined because the companies will mislead people and advertise with “open-source”, as in open-source software (like Linux, Firefox etc). But their models rarely include the sources (so to speak). That’d be the training recipe and all the data that went in. They can’t publish the training data though, because they regularly get sued by the people they stole the books from. So… they only publish the resulting AI model.

        It’s a bit like an executable file on a computer or a purchased game which you can run at your own terms, on your computer, without any online stuff or anticheat attached. You just don’t own any of the development resources.

        That means we don’t know what went in, we can’t recreate it, and we might not be able to learn a lot. We can however run and use the model. (We can even modify it to some degree.)

        Of course that’s too easy. It’s not really an executable file. Those AI models are neural networks. All the information and “knowledge” is stored in the parameters of that network. That’d be the weights. Lots of numbers for all the nodes and edges which make up the graph/network.

        And not sure about the inferior/superior models… I mean it’s hard to compete with OpenAI or Anthropic and the huge pile of money they have. They hired a lot of talent… And their results will be a trade secret, only offered as a service… But then there’s this whole AI war going on between the US and China. And all the companies against each other. They’re constantly trying to outcompete each other. And usually it won’t take long until someone claims they made an open-weights model as good as (or better than) the current version of ChatGPT.