Executives working on AI at Microsoft and OpenAI admitted what its critics have been saying all along: Large language models are predatory pieces of technology that have been built on what a Microsoft executive called “an astonishing theft of unprecedented proportions,” and the “largest theft of labor in human history.” An internal Microsoft document said generative AI products have created a “doom loop” that is killing “the entire web.”

Those statements and a series of other mask-off moments feature heavily in an unredacted court filing that was unsealed Thursday in the behemoth New York Times vs OpenAI copyright lawsuit that has been winding its way through the court system for years. In a filing asking for summary judgment (basically, a filing with the court asking it to rule), lawyers for the New York Times laid out a series of admissions made by Microsoft and OpenAI executives in documents and depositions that until now had remained either sealed or redacted at the request of Microsoft and OpenAI.

It’s easy to see why the AI companies wanted to hide this from the public. The statements, taken together, are some of the most damning indictments of the ways LLMs were trained, how they worked, and the immediate threat they pose to human labor. It is a reminder that even as AI becomes more powerful and companies try to shift the narrative to the supposed existential risk of “superintelligent” AI, the tools they have already built were created by stealing from human creativity and labor and are by definition existential threats to the human labor market.


  • FoxAlive@lemmy.zip
    link
    fedilink
    English
    arrow-up
    57
    arrow-down
    1
    ·
    1 day ago

    I honestly don’t see a point where we can move past this without Sam altman, nadell, musk, etc all facing mandatory life time sentence without parole, work release etc.

    I would also accept the removal of their heads.

    • Alaknár@sopuli.xyz
      link
      fedilink
      English
      arrow-up
      30
      ·
      1 day ago

      I’d prefer if they were forced to pay royalties for all the work they stole to the people they stole it from. And, like, have someone actually force them to comply. They’d have to hire a Microsoft-sized compliance department just to figure this shit out and track payments.

      • Scrollone@feddit.it
        link
        fedilink
        English
        arrow-up
        8
        arrow-down
        1
        ·
        23 hours ago

        Yeah, but with which money? All those AI companies are not and will never be profitable (and they’ll be the reason for the next stock market bubble pop)

        • JustPlainDave@lemmy.zip
          link
          fedilink
          English
          arrow-up
          15
          ·
          1 day ago

          Could you imagine Elon Musk running a fryer or injection mold for 8 hours? Who am I kidding? he’d never pass the piss test required to run the injection mold.

          • FoxAlive@lemmy.zip
            link
            fedilink
            English
            arrow-up
            9
            ·
            1 day ago

            I’m sure many of them think they would just prove us wrong, and because they are so smart that they would just rebuild from the ground up on their own.

            My concern is they could convince another rich person of this, and get them to invest in their newest scam.

            • JustPlainDave@lemmy.zip
              link
              fedilink
              English
              arrow-up
              3
              ·
              1 day ago

              There’d have to be some sort of stipulation that made it so nobody could do that. I don’t know how you’d do that, but it could probably be done. Are there any laws that prevent people convicted of fraud from doing the same thing again? Maybe something like that.

              I would love to see them roofing houses or working in a foundry. Hell, even seeing them trying to be a janitor or something.

              • llama@lemmy.zip
                link
                fedilink
                English
                arrow-up
                1
                ·
                1 day ago

                They could still grift their way to a life of comfort through book deals with Ayn Rand or something

                • JustPlainDave@lemmy.zip
                  link
                  fedilink
                  English
                  arrow-up
                  1
                  ·
                  1 day ago

                  I doubt it. Nobody who cares about what they’d have to say actually reads books, and they wouldn’t be able to afford a ghost writer anyway.

  • Mrkawfee@lemmy.world
    link
    fedilink
    English
    arrow-up
    73
    arrow-down
    1
    ·
    1 day ago

    Internet search is terrible now. Every website I go to reads like it was generated by an LLM.

    • clif@lemmy.world
      link
      fedilink
      English
      arrow-up
      36
      arrow-down
      2
      ·
      1 day ago

      Because it was generated by a llm.

      I did a search for the torque spec on a castle nut last week and the second result was for a nut (as in, food nuts that you eat) website and the llm had gone all in on nuts and added a page about castle nuts. It even generated an image of a castle for the top because… castle nut.

      It proceeded to provide vague instructions and a torque recommendation of 2x the actual spec.

      Then there was the other one about a small 4 stroke engine where it stated “other sites will say you don’t need to mix oil with gas but you absolutely must!” (You don’t and shouldn’t) … I wonder how many people have fucked up their shit by trusting llm garbage without knowing better.

      • docandersonn@literature.cafe
        link
        fedilink
        English
        arrow-up
        20
        arrow-down
        1
        ·
        1 day ago

        Running into this same problem but with parenting. My wife is constantly asking me to “Google if it’s good/bad if our baby is doing” X. And I’ll find 20 slopposts from dozens of “parenting” websites that all give conflicting answers. This week, I asked my parents for their copy of Dr. Spock’s Baby and Childcare because we need some solid reference. Sorry kiddo, you’re going to get raised like it’s 1987 because 2026 sucks.

        • Doomsider@lemmy.world
          link
          fedilink
          English
          arrow-up
          2
          ·
          edit-2
          23 hours ago

          Try the 1940’s when the book was first published. Most of what he said, like talking to and listening to your kids and treating them like a human is really good. After a quick review of criticism it doesn’t appear there is much worthwhile other than some of his earlier views that did change later on (like male circumcision).

        • clif@lemmy.world
          link
          fedilink
          English
          arrow-up
          2
          ·
          1 day ago

          I can’t even fathom that level of frustration. At least I can only muck up a part or maybe an engine… not a little human : /

          Sorry bud, and good luck.

    • Doom@discuss.online
      link
      fedilink
      English
      arrow-up
      34
      ·
      1 day ago

      Glad I spent my teens and 20s getting stoned and reading the shit out of wikipedia before this shit happened. Can’t trust anything written online now.

    • XiELEd@piefed.social
      link
      fedilink
      English
      arrow-up
      6
      ·
      1 day ago

      I once looked for a tutorial on a process in a Minecraft mod and there was a site that gave out incorrect information that was obviously written by an LLM.

      • kuiskaaja@lemmy.ml
        link
        fedilink
        English
        arrow-up
        7
        ·
        1 day ago

        this website used to pop up in pretty much all web searches i made about retro gaming. looks like it’s blacklisted from qwant now, since i couldn’t find it anymore by searching and had to dig through my discord messages to find the link

        “To give you an idea of the size range, Kirby’s Adventure, the largest officially released DS game, takes up 768 KB, while smaller games like Tetris weigh in at around 8MiB. The maximum recorded storage capacity of a Nintendo DS cartridge is 64 MB.”

        and then right after that “Size Range: 8MiB to 512MB”, followed by:

        “According to various sources, the average DS game size is 32mb to 64mb, with some games being larger or smaller. The actual average would be hard to calculate, but it’s likely somewhere between 128 KB and 256 KB.”

        crazy stuff lol, it’s like it’s suffering from psychosis

    • grrgyle@slrpnk.net
      link
      fedilink
      English
      arrow-up
      3
      ·
      1 day ago

      Yeah I don’t even use ai so I can’t compare, but even paid search services are noticeably so much more frustrating.

  • RunawayFixer@lemmy.world
    link
    fedilink
    English
    arrow-up
    45
    arrow-down
    2
    ·
    1 day ago

    The best succinct description of llm that I’ve read was “plagiarism machine”, because that’s basically what they are: automated plagiarism. At best they create a collage of other works, at worst they copy one work verbatim, but they never create something original because they can’t.

    Copying something is a lot easier and thus cheaper than creating an original work from scratch, so genuine creators cannot hope to compete on cost. Which leads to less people being able to afford to earn a living from creating works, which leads to less original works being created.

    Short term that’s great for the neoliberal company executives: fire the expensive artists, create cheap slop with the plagiarism machine, and thus maximize profits now. That the plagiarism machine becomes stagnant because not enough new original works are being created is a future problem, by which time the current executives will have already jumped ship.

    But while it’s great for the neoliberal business model, for the rest of society it will suck. So now that we have confirmation that the AI executives know how bad their products are, and that they were just publicly lying about it in the style of tobacco executives, will anything be done about this “astonishing theft of unprecedented proportions”? Personally I doubt it, there’s too much regulatory capture.

    • Matty Roses@lemmy.today
      link
      fedilink
      English
      arrow-up
      1
      ·
      6 hours ago

      Which leads to less people being able to afford to earn a living from creating works, which leads to less original works being created.

      It’s why the only solution for AI is socialism.

      It breaks every profit model and the wage form otherwise.

  • velma@sh.itjust.works
    link
    fedilink
    English
    arrow-up
    226
    ·
    2 days ago

    “Our AI content strategy has started a ‘doom loop’ that will hurt the performance of our models and the entire web at the same time: It is highly unusual that an end-product threatens the economic foundations of its essential suppliers, but that is the situation we have created for our LLM business with respect to its ‘content supply chain,’” the document said.

    Microsoft executives, including CEO Satya Nadella, testified under oath that after ripping content from the New York Times and other news sites, clicks to those news sites fully cratered, falling by more than 90 percent on Bing.

    Documents obtained during the court proceedings found that OpenAI created “a hack to get around nytimes paywall,” to which OpenAI cofounder Greg Brockman said “ah, nice.” Microsoft executive Brent Hecht wrote that LLMs steal content “without ways of distributing economic value down the supply chain, [which] necessarily threatens the economic stability of those who create the content.”

    Fuck these guys.

    • FatCrab@slrpnk.net
      link
      fedilink
      English
      arrow-up
      2
      ·
      32 minutes ago

      So the admission of intentionally circumventing the NYT paywall is where this becomes a real actual IP infringement issue for them. From a legal perspective, that is literally as crazy a thing to come out in discovery as Anthropic’s torrenting an enormous chunk of their training corpus. These are IP infringements, totally irrespective their use in training GPTs. I cannot imagine having to represent these idiots.

        • Matty Roses@lemmy.today
          link
          fedilink
          English
          arrow-up
          1
          ·
          6 hours ago

          someone should write a book about how capitalism functions. Call it like The Capital or something.

      • boonhet@sopuli.xyz
        link
        fedilink
        English
        arrow-up
        7
        arrow-down
        3
        ·
        1 day ago

        That’s the goal of most technology. To allow more things to get done in less time with fewer employees.

        CNC machine, tractor, chainsaw, printing press for some examples.

    • tangeli@piefed.social
      link
      fedilink
      English
      arrow-up
      66
      arrow-down
      1
      ·
      2 days ago

      That’s because, thus far, they get away with choosing not to distribute any of their trillions of dollars to the suppliers of the information they consume - money has only gone to the suppliers of hardware and power, and to influencing politicians and rewarding investors. That’s their choice, and they should not be allowed to continue to make that choice. Good luck to the NYT.

      • Grail@multiverse.soulism.net
        link
        fedilink
        English
        arrow-up
        18
        arrow-down
        4
        ·
        2 days ago

        https://www.advocate.com/politics/national/new-york-times-transgender-controversy

        https://www.npr.org/2023/02/15/1157181127/nyt-letter-trans

        https://theintercept.com/2024/04/15/nyt-israel-gaza-genocide-palestine-coverage/

        NYT is transphobic and pro-genocide. I hope they go bankrupt, and papers with fewer conservative biases survive. Unfortunately, the opposite will happen. Billionaire-backed news will run at a loss to manufacture propaganda, while independent journalism dies.

      • Grandwolf319@sh.itjust.works
        link
        fedilink
        English
        arrow-up
        9
        ·
        2 days ago

        But would it be worth it if they had to pay the creative costs for it?

        They are only worth it now cause it doesn’t include that AND it’s subsidized by investor money.

        • Rothe@piefed.social
          link
          fedilink
          English
          arrow-up
          10
          ·
          1 day ago

          From an economic viewpoint AI companies are not worth is as it is now. OpenAI and Anthropic are hundreds of billions of dollars in debt and will never make a profit. Same goes for the AI subsidiaries of Microsoft, Meta, Google etc, except they are just leeching off of the mother companies and hiding their figures among their finances. Paying creators for their data would make very little difference in their overall finances.

          • kestrel7_7@lemmy.world
            link
            fedilink
            English
            arrow-up
            6
            ·
            1 day ago

            This is why I’ve been arguing to anyone who will listen for like four years now. If this tech is so amazing, someone will figure out a way to make money off of it. Right? The fact that it’s been 4+ years and no one has made a damn cent off of it should be making more people skeptical. The fact that 4 years ago it was widely celebrated even though no one had a plan to make a damn cent off of it should have made more people skeptical back then. It’s weird that I have to keep arguing this with folks.

        • tangeli@piefed.social
          link
          fedilink
          English
          arrow-up
          7
          ·
          2 days ago

          I think the way they are operating now is commonly called predatory pricing.

          I don’t know if there is a potential for a profitable business that covers all the costs. Possibly the direct costs, but I think not if they had to fairly compensate those who created the intellectual property they consumed to train their models, and much less likely if they had to pay the indirect costs. They might, some day, compensate creators a little, like Google shares a pittance of their ad revenue with some creators. But it is unlikely to be anything near a living wage.

        • bad1080@piefed.social
          link
          fedilink
          English
          arrow-up
          3
          ·
          1 day ago

          But would it be worth it if they had to pay the creative costs for it?

          no and they already admitted as much

        • BlaestEgnen@feddit.dk
          link
          fedilink
          English
          arrow-up
          4
          arrow-down
          1
          ·
          2 days ago

          AND it’s subsidised by too big a supply of compute power - Microsoft prints demand for Azure, by buying ownership stakes in openAI through the grant of Azure credits. There’s too much available compute.

    • TeaWithDani@lemmy.world
      link
      fedilink
      English
      arrow-up
      37
      ·
      2 days ago

      That’s kind of the funniest part in fact: these companies are destroying their own viable business segments. Bing was a huge growth driver for Microsoft. Less clicks is bad for them. They make more money on Bing ads than they do on LLMs. Same with Google.

      Reddit is getting crushed atm after it sold access to its data to train models. Chat bots make visiting these websites pointless, without replacing that traffic with anything they can meaningfully monetize.

      The more popular Gemini is, the less money Google will make. The market has already shown how much people are willing to pay for Ai subscriptions, and it isn’t all that much. None of these companies have found a way to make ads viable in LLMs either. They are beyond self sabotaging themselves at this point. It really is a doom loop.

    • belochka@lemmy.world
      link
      fedilink
      English
      arrow-up
      11
      ·
      2 days ago

      At least they are going to pick one between copyright and using all the Web as accumulated material for their answering machine.

  • CosmoNova@lemmy.world
    link
    fedilink
    English
    arrow-up
    65
    arrow-down
    1
    ·
    2 days ago

    They knew what they were doing every step of the way. They are criminals that need to be disarmed and locked away. And we need to create a new Internet from scratch somehow thanks to these donkeys.

  • radiofreebc@lemmy.world
    link
    fedilink
    English
    arrow-up
    3
    arrow-down
    1
    ·
    20 hours ago

    I was there when the internet started, and it was always doomed from the beginning. This was never going to last as a free, open resource.

  • Arancello@aussie.zone
    link
    fedilink
    English
    arrow-up
    95
    arrow-down
    3
    ·
    2 days ago

    Relax guys, the very stable high IQ president of the united states will protect you. No meed for guardrails or regulations.

    • AmyAye@nord.pub
      link
      fedilink
      English
      arrow-up
      70
      ·
      2 days ago

      God the guardrails things. These stupid companies are all hyping up “We need to slow down, we need guard rails!”

      Ok.

      No one is fucking stopping you. Just… Slow yourself down, guardrail yourself.

      Oh wait, it’s just an excuse to create regulatory capture.

      • Fluke@feddit.uk
        link
        fedilink
        English
        arrow-up
        19
        ·
        1 day ago

        It’s the theftbot manufacturers trying to have excuses in the public perception for why all their claims about their product don’t ever happen.

        “We had to slow down for safety. All those things we promised are coming, just one more round of funding bro.”

      • iamthetot@piefed.ca
        link
        fedilink
        English
        arrow-up
        5
        ·
        1 day ago

        Devil’s advocate, self regulation does not work. 9/10 companies can agree to cool it and if 1/10 presses on full steam ahead for whatever reason, then they can still do all the damage and they will be ahead of the curve.

        We need laws and regulations that hold all companies accountable.

      • CileTheSane@lemmy.ca
        link
        fedilink
        English
        arrow-up
        1
        ·
        1 day ago

        They want Trump to create the guardrails (the exact guardrails they tell him to create) before the midterms and the Republicans lose a bunch of seats.

    • Archangel1313@lemmy.ca
      link
      fedilink
      English
      arrow-up
      18
      ·
      2 days ago

      That clown is legalizing all kinds of white collar crime, already. This kind of shit is nothing to him.

  • rumba@lemmy.zip
    link
    fedilink
    English
    arrow-up
    12
    ·
    1 day ago

    NGL, I was kinda hoping something would destroy the internet and we could move to Reticulum/NomadNet/I2P/Tor

    Shit’s been going downhill since 1999

    • Matty Roses@lemmy.today
      link
      fedilink
      English
      arrow-up
      2
      ·
      6 hours ago

      Setting up a Reticulum node this week.

      And wondering if there’s a matrix client that can use it, or if that’s something that needs to be made . . .

      • rumba@lemmy.zip
        link
        fedilink
        English
        arrow-up
        1
        ·
        5 hours ago

        They’re very different.

        You could try to do some kind of bridge server side or just add straight up identity/send/receive client side. You need RNSD running somewhere to hook.

      • rumba@lemmy.zip
        link
        fedilink
        English
        arrow-up
        6
        ·
        1 day ago

        Transport isn’t the issue, but both those protocols provide low-bandwidth hosting requirements. NomadNet is straight out of the 90’s bbs territory, and I2P is slow enough on its own to preclude anything too fancy. It’s forcing us to go back to the small distributed hosting models, which I consider a good thing. They’re also resistant to censorship and, in some cases, reasonably privacy-conscious

  • gravitas_deficiency@sh.itjust.works
    link
    fedilink
    English
    arrow-up
    87
    ·
    2 days ago

    Microsoft, in a policy document, wrote that generative AI could “significantly disrupt the employment of the very people who generated the data on which the foundation model was trained […] LLMs are a product that destroys its supply chain.”

    If that’s not the very definition of a categorically unsustainable business model, I don’t know what the fuck is.

      • belochka@lemmy.world
        link
        fedilink
        English
        arrow-up
        6
        arrow-down
        1
        ·
        2 days ago

        That’s not quite the thing. They know that people in various societies in the course of history could create artifacts getting much less in return than they’d get today in a market without LLMs. They just want those people to become crops in their fields, so to say. To reap and not sow.

        There’s the obvious problem with this, that there still are alternative business models of closed circles and ordering artifacts, not accepting offers to buy them. And, of course, that either it’s IP violation of ultimate proportions, or if it’s not, then they are going to have too much competition to get anything out of it anyway.

        So it’s either oligopoly with regulatory capture, or too much competition to make this profitable.

        They apparently want to make some theft legal and some not, a bit like copyright protection entities in ex-Soviet states, which are usually controlled by former pirates legalized (with ex-pirate electronic libraries which are now legitimate stores, except I’m confident they didn’t get consent of most authors whose books are being sold, similarly with movies and such), hunting their “honest pirate” competition.

        So - no, at the point this is mainstream, the only regulation of this should be aimed at enforcing existing laws against them. Then one can think of some, but before existing laws are being firmly enforced - no.

        Except enforcing existing laws can mean fines bigger than their capitalization and some life sentences for all chief people involved, and these are all “too big to fail” companies, so I don’t even know.

        If these companies are, figuratively, sentenced to death for doing this, then it’ll be a process of the century. It’ll define future. A bit like with Standard Oil and United Fruit Company. Or so I think.

        Except it might go differently, in their favor, then it won’t be a very nice future.

        • ScoffingLizard@lemmy.dbzer0.com
          link
          fedilink
          English
          arrow-up
          6
          ·
          1 day ago

          Standard Oil’s monopoly wasn’t ever truly dealt with in my opinion. We’re still seeing the fallout of people who got way too rich off it. I’m sure some asshole Rockefeller has had a meeting on Venezuelan oil in the last year.

        • Valmond@lemmy.dbzer0.com
          link
          fedilink
          English
          arrow-up
          2
          ·
          2 days ago

          Or they will pop all by themselves?

          I mean there is not very much new data to train on, and only so much refining you can do (and you don’t need to be one of the big endebted companies to do that).

          • belochka@lemmy.world
            link
            fedilink
            English
            arrow-up
            2
            ·
            1 day ago

            “All by itself” usually doesn’t happen, “it” just stops looking like a problem for those not destroyed by “it”, and those destroyed by “it” stop existing.

    • CileTheSane@lemmy.ca
      link
      fedilink
      English
      arrow-up
      7
      ·
      1 day ago

      If that’s not the very definition of a categorically unsustainable business model, I don’t know what the fuck is.

      So is expecting infinite growth in a finite system and they’ve been doing that for years.

  • kablez@lemmy.world
    link
    fedilink
    English
    arrow-up
    50
    ·
    2 days ago

    Gonna get real weird soon when they run out of rich new training data and they begin to consume their own shit. When that happens their entire model will collapse and if the bubble hasn’t popped already that may be what causes it.

    • RepleteLocum@lemmy.blahaj.zone
      link
      fedilink
      English
      arrow-up
      33
      arrow-down
      1
      ·
      2 days ago

      They’re already doing it. They call it distilling when they take it from another llm. Pretty sure most content was already stolen in the early days and they now rely on distillation and stealing new content.

    • MalReynolds@slrpnk.net
      link
      fedilink
      English
      arrow-up
      10
      ·
      2 days ago

      Pretty sure they mostly use the pre-AI internet (that they scraped and kept) and synthetic data currently. Probably trying (and failing so far or we’d have heard) to adapt to using video as training material at the moment, but developments there will likely apply to robotics at some point. Here’s hoping the current chuds have crashed and burned before then and that some sanity has taken over from unfettered capitalist oligarchs dreams of computer slavery.

        • MalReynolds@slrpnk.net
          link
          fedilink
          English
          arrow-up
          3
          ·
          1 day ago

          Flock does pattern recognition, a quite old piece of machine learning. Nothing to do with training a large language model or other ‘AI’ model.

    • WorldsDumbestMan@lemmy.today
      link
      fedilink
      English
      arrow-up
      2
      ·
      1 day ago

      Or they can just keep a database of actual data on their servers, instead of getting new data every single time for some fucking reason.

    • Thorry@feddit.org
      link
      fedilink
      English
      arrow-up
      4
      arrow-down
      1
      ·
      2 days ago

      Why do you think there has been such an emphasis on hacking with LLMs lately (especially by OpenAI). They figured out all those vulnerability databases were an excellent source for training material. In the past they scraped those, but just for general language training. Now they’ve specifically trained the models on the information within. Some team figured out how to use that data to train a model and have testing scenarios automated so they could write a good reward function. It wasn’t that they figured out the models are good at hacking, they ran out of content and found a new source of good data.

      With all the books they’ve been scanning I wonder if the next thing is going to be a writing assistant or editor or something like that. Even though writing good books is an art form and the actual writing down of the words is the easiest part (still not easy tho).

      These companies are starving for content and they’ve not just poisoned buy absolutely destroyed the content well that is the internet. Given they were already hitting diminishing returns hard, it doesn’t matter too much to them probably. But more compute and storage has also been hitting diminishing returns hard and customers are complaining about the cost. So they are getting a bit desperate on how to improve these things at all.

        • DeadDigger@lemmy.zip
          link
          fedilink
          English
          arrow-up
          2
          ·
          1 day ago

          It’s either agi or model collapse. For every AI system actually. If you have a high enough adoption you start to muss original data so if your AI is not self sufficient in time it will collapse, because it will be trained on its own data, which just is an incentive loop

        • BilSabab@lemmy.world
          link
          fedilink
          English
          arrow-up
          1
          ·
          1 day ago

          in a manner of speaking - you always end up there. not by design though. models operate via continuous refinement and you can only optimize a model so much until it is a mess and you need to figure out where to roll back. so you either get shit like semantic drift or variance decay and you can whack a mole it to an extent but then you hit the rlhf wall when the model starts gaming its reinforcement framework and the fat lady sings.

  • grrgyle@slrpnk.net
    link
    fedilink
    English
    arrow-up
    17
    arrow-down
    1
    ·
    1 day ago

    Actually surprised how candid some of these statements are. Maybe points to more internal resistance than I would have assumed… like, they see the problem.

    Their focus on threats to media companies is probably borne out of fear of litigation, so maybe they only pay a lil lip service to how they also hurt “content creators” (people).

    Could also be that they know that when talking to the representatives of runaway financial automatons like multinational corporations, they have to appeal in terms the machine will understand.

  • k0e3@lemmy.ca
    link
    fedilink
    English
    arrow-up
    29
    arrow-down
    10
    ·
    1 day ago

    LLMs are NOT destroying the planet. It’s the humans running the companies making LLMs.

  • Grandwolf319@sh.itjust.works
    link
    fedilink
    English
    arrow-up
    32
    ·
    2 days ago

    “Our AI content strategy has started a ‘doom loop’ that will hurt the performance of our models and the entire web at the same time: It is highly unusual that an end-product threatens the economic foundations of its essential suppliers, but that is the situation we have created for our LLM business with respect to its ‘content supply chain,’” the document said.

    So basically the ouroboros

    • BilSabab@lemmy.world
      link
      fedilink
      English
      arrow-up
      5
      ·
      1 day ago

      funny thing is that this ouroboros thing been known ever since LLM concept became a thing. that’s why it took forever to take off beyond r&d experiments. basically, the only remotely adequate way to apply LLM is what NotebookLM does - which literally parsing an uploaded document and reiterating it in some preset format. that’s literally it.