this post was submitted on 06 Sep 2026
114 points (97.5% liked)

Technology

87853 readers
2415 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] tal@lemmy.today 2 points 3 hours ago* (last edited 3 hours ago)

Use a local model

I also think that local models are generally preferable, but everyone having a local model requires more hardware, since much of the time the local hardware to run the local model is probably going to be idle, whereas hardware in the cloud can be servicing User B when it's not servicing User A.

In 2026 and 2027, absent some very unexpected development, memory prices are going to be elevated.

And while some people could get local hardware, we couldn't build a comparable-to-what-someone-could-get-in-the-cloud AI compute box for everyone if everyone wanted to go local, not for years, as we don't have the memory production capacity today. If everyone tried to get a local box for AI compute at the same time, it'd just do what happened when the cloud companies did it, but at greater scale, because it'd require even more memory. There'd be a new memory shortage. Prices on memory would keep rising, pricing would-be buyers out, until the number of buyers still willing to buy was equal to what supply was present.