• eicker@lemmy.worldOP
    link
    fedilink
    English
    arrow-up
    33
    arrow-down
    2
    ·
    1 month ago

    The interesting part is not whether Apple wins the biggest model race, but whether it changes the economics: If enough AI runs locally, every token avoided is cloud capacity nobody has to build. That is a very different business model from selling ever more cloud compute.

    • errer@lemmy.world
      link
      fedilink
      English
      arrow-up
      4
      arrow-down
      1
      ·
      1 month ago

      I’m pretty skeptical local models can hold a candle to the cloud-based ones, particularly the ones Apple trains.

      • eicker@lemmy.worldOP
        link
        fedilink
        English
        arrow-up
        28
        arrow-down
        1
        ·
        1 month ago

        Raw capability is only one metric: A local model probably will not beat the best cloud model any time soon, but it does not need to. If it handles 80 to 90% of everyday tasks instantly, privately and at near zero marginal cost, that is a huge win. Reserve the cloud for the genuinely hard requests, not every prompt.

      • thehermet@lemmy.ca
        link
        fedilink
        English
        arrow-up
        11
        ·
        1 month ago

        Most people don’t need deep agentic ai on their phones, they just need quick answers to questions, to add events to their calendars, answer emails, and remember things about their lives. These local llms actually perform better than the cloud ones for these tasks

  • Zwuzelmaus@feddit.org
    link
    fedilink
    English
    arrow-up
    27
    ·
    1 month ago

    He can see into the future.

    Future #1: The lifetime of the pure AI companies is limited.

    Future #2: Local LLM’s are trending, because “token” prices will go up like crazy.

    • eicker@lemmy.worldOP
      link
      fedilink
      English
      arrow-up
      8
      arrow-down
      2
      ·
      1 month ago

      The present: Open Weight AI, such as Kimi’s, is already almost exactly as good as ClosedAI from Anthropic and »OpenAI«.

      • makeshift0546@lemmy.today
        link
        fedilink
        English
        arrow-up
        2
        arrow-down
        1
        ·
        edit-2
        1 month ago

        How much compute do you think you need to run kimi locally at mythos fable levels? That ain’t going to work unless people start buying small data centers locally.

        • eicker@lemmy.worldOP
          link
          fedilink
          English
          arrow-up
          8
          arrow-down
          2
          ·
          edit-2
          1 month ago

          About 4+ maxed M4 Studios, I guess. But that‘s not the point: in 80%+ of cases, people won’t need that kind of AI model to solve their problems.

    • stealth_cookies@lemmy.ca
      link
      fedilink
      English
      arrow-up
      4
      ·
      1 month ago

      Yeah “See into the future”. They’ve done the business analysis and determined the outcomes you gave. Apple has the confidence as a company to hold back even if the market wants them to do something and know they can hold firm through the irrationality of the market and come out the other side.

      One has to wonder how companies like Google and Microsoft will fare when the financial engineering blows up in their faces. I’m guessing they think the government will bail them out.

      • eicker@lemmy.worldOP
        link
        fedilink
        English
        arrow-up
        2
        arrow-down
        2
        ·
        1 month ago

        Apple has always been unusually willing to sacrifice short term hype for long term positioning. That does not guarantee they are right, but it is a very different bet from spending hundreds of billions assuming demand will eventually justify the buildout. If AI demand disappoints, discipline suddenly looks a lot more valuable than scale.

    • weew@lemmy.ca
      link
      fedilink
      English
      arrow-up
      2
      ·
      edit-2
      1 month ago

      This isn’t seeing into the future. This is just taking a look at the present.

      Present #1: AI datacenters cost a fuck ton

      Present #2: No customer is willing to pay the costs of AI unless they sell at a loss

  • fartsparkles@lemmy.world
    link
    fedilink
    English
    arrow-up
    8
    ·
    1 month ago

    It seems to have been a plan for a long time, given their huge shift to unified memory architectures across most of their hardware.

    They’re pretty much the only vendor where you can cost-effectively deploy a foundational LLM locally.

    • eleitl@lemmy.zip
      link
      fedilink
      English
      arrow-up
      1
      ·
      1 month ago

      Still nobody build a unified memory model with HBM. I would buy Apple hardware, then. If I can put Linux on it.

      Perhaps I can buy a used Instict with generics science acceleration, once the smoke clears. And add a couple solar kWp to my capacity.

    • eicker@lemmy.worldOP
      link
      fedilink
      English
      arrow-up
      2
      arrow-down
      1
      ·
      1 month ago

      It would seem so. On the other hand, it is puzzling that they did not also allocate the necessary resources to the development of LLMs. 🤷

    • eicker@lemmy.worldOP
      link
      fedilink
      English
      arrow-up
      7
      arrow-down
      1
      ·
      1 month ago

      Cook’s biggest product might be expectation management. He rarely promises tomorrow’s miracle, which buys Apple room to ship when it suits them instead of when Wall Street gets impatient.

    • eicker@lemmy.worldOP
      link
      fedilink
      English
      arrow-up
      1
      arrow-down
      2
      ·
      1 month ago

      Absolutely. Looking forward to seeing the next generation of Macs.

      • Whostosay@sh.itjust.works
        link
        fedilink
        English
        arrow-up
        7
        ·
        1 month ago

        I couldn’t give a shit less about apple but man I’m I hoping this comes to fruition.

        I’d like to buy hardware again.

  • crystalmerchant@lemmy.world
    link
    fedilink
    English
    arrow-up
    4
    ·
    1 month ago

    And this will have huge ramifications for my field, energy, because a substantial portion of compute power usage will move out of data centers and into the edge (your iPhone)

    • eicker@lemmy.worldOP
      link
      fedilink
      English
      arrow-up
      6
      ·
      1 month ago

      The decentralised operation of LLMs would also be significantly simpler and cheaper for the use of decentralised renewable energy sources.

  • CompactFlax@discuss.tchncs.de
    link
    fedilink
    English
    arrow-up
    3
    ·
    1 month ago

    Apple’s AI strategy is leaving them behind

    Apple’s stock drops on failure to meet AI promises

    Etc.

    Now whose stock is dropping?

    • eicker@lemmy.worldOP
      link
      fedilink
      English
      arrow-up
      4
      ·
      1 month ago

      Stock prices aren’t proof of being right, but they do show investors can change their minds a lot faster than the narratives do.

  • unitedwithme@lemmy.today
    link
    fedilink
    English
    arrow-up
    3
    arrow-down
    1
    ·
    1 month ago

    Tim Cook will say anything to try and get ahead. First Apple’s talking about how good their AI is/will be, but now fears of over-hyped AI spending and data center issues, etc, now he’s claims oh and it’ll all run locally. Apples typical “we’re not the bad guys”

    While local LLMs are getting better, I still feel like overall, this will tank battery life, and through several million devices charging, will still use a lot of extra power overall. Plus, a lot of the data center issues falls to training and ingesting information, so, Apple let’s the other guys do the heavy lifting for them to seem less evil? Wasn’t Apple initially involved with OpenAI when they were a nonprofit and cofunded business? Idk, I still say Apple isn’t trustworthy.

    • eicker@lemmy.worldOP
      link
      fedilink
      English
      arrow-up
      5
      ·
      1 month ago

      Apple’s marketing deserves skepticism, but the technical argument is separate. Local inference does not eliminate giant training clusters, it mainly cuts inference costs, latency and improves privacy. Apple still uses cloud models when needed.

  • Sumocat@lemmy.world
    link
    fedilink
    English
    arrow-up
    0
    arrow-down
    1
    ·
    1 month ago

    I am practicing that strategy now. I recently upgraded to an iPad Pro M5 (the day price increases were announced, jumped on a deal immediately), upgraded several shortcuts with Apple Intelligence, and am refining them to run entirely on-device instead of in PCC (not using ChatGPT at all).

    • vollkorntomate@infosec.pub
      link
      fedilink
      English
      arrow-up
      1
      ·
      1 month ago

      Just a note that PCC is different from the ChatGPT integration. You might be using PCC resources without noticing, because the iPad won’t tell you in advance. Only way to be sure is to disconnect from the Internet and/or check your Apple Intelligence Report (in System Settings > Privacy) after the fact