this post was submitted on 11 Feb 2025
18 points (80.0% liked)

Technology

63082 readers
3546 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 2 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] hendrik@palaver.p3x.de 1 points 1 week ago* (last edited 1 week ago)

Yeah, that just depends on what you're trying to achieve. Depending on what kind of AI workload you have, you can scale it across 4 GPUs. Or it'll become super slow if it needs to transfer a lot of data between these GPUs. And depending on what kinds of maths is involved, a Pascal generation GPU might be perfectly fine, or it'll lack support for some of the operations involved. So yes, of course you can build that rig. Whether it's going to be useful in your scenario is a different question. But I'd argue, if you need 96GB of VRAM for more than just the sake of it, you should be able to tell... I've seen people discuss these rigs with several P40 or similar, on Reddit and in some forums and Github discussions of the software involved. You might just have to do some research and find out if your AI inference framework and the model does well on specific hardware.