top 50 comments

sorted by: hot top controversial new old
[–] 93 points 1 month ago* (3 children)

The article talks a lot trash about AMD and ROCm but vulkan works fine too. In fact from a datacenter GPU standpoint there is an AMD option called the V620 available on US EBay that I was able to haggle to $350, with 32GB VRAM, 512GB/s bandwidth, and runs the same Qwen-3.6-27b at about 20t/s. I would argue that's even more cost effective.
It requires a few of the same fan shenanigans this guy did but there is no need to pull specific past software versions to make it usable in Linux

  • source
  • hideshow 3 child comments
  • [–] 8 points 1 month ago (1 child)

    Honestly even for the prices of around 500$ that I'm seeing it for it looks like a pretty good value to get 32gb of vram. I see it says 300w on AMD's product page for it does it have any way of power limiting the card to get more efficiency/less heat?

  • source
  • parent
  • hideshow 1 child comment
  • [–] 8 points 1 month ago* (last edited 1 month ago)

    I spent a lot of time researching and testing different methods for that, the only thing that worked was LACT in Linux. Using that I was able to undervolt 100mV and GPU power usage dropped about 10%. On my B450 ITX board with a Ryzen 2400GE CPU the entire system pulls 30w idle from the wall, and about 300w inferencing with VRAM filled. (330w before LACT).
    My fan solution ended up being to buy the 80mm 3d-printed shroud off ebay, the fan that came with it was super loud so I switched to an arctic p8 Max, and control it with the motherboard targeting a t-sensor header with the probe attached to the backplate.

  • source
  • parent
  • [–] 60 points 1 month ago (4 children)

    Here I was expecting a graphics demo to blow our collective minds but instead I got a story about a local LLM for cheap. It is . I should have known better.

    Were I the author / tech cobbler here, I'd be concerned that too much time with an LLM, local or otherwise, might erode or dull my apparently fairly sharp reasoning and tech skills. (Clarification: Not my sharpness, theirs. I'm a potato.)

    Other thoughts: For a minute I thought this whole thing was a tribute to, or a troll in the manner of, that one Redditor that always spun their stories around to being about their dad beating them with jumper cables.

    Also, my old PC developed an issue like the warm reboot problem, except with the network interface. I couldn't just restart, I had to power off and back on. I never did bother to find out whether it was early signs of hardware failure or whether it was an old hardware / newer kernel mismatch.

  • source
  • hideshow 4 child comments
  • [–] 5 points 1 month ago (1 child)
  • [–] 8 points 1 month ago

    The change of the meaning of the G in GPU from "graphics" to "general" is even less well documented and used than the "V" of DVD changing from "video" to "versatile".

    Indeed it only occurred to me what it must have changed to and to go looking to confirm after seeing your comment.

    And frankly they ought to have changed the name to something like "MPPU" if they wanted it to stick (massively parallel).

  • source
  • parent
  • load more comments (1 reply)
    [–] 25 points 1 month ago (5 children)

    Seems like an awful lot of trouble to save $100 not buying a 5060 Ti that also has 16GB.

  • source
  • hideshow 5 child comments
  • load more comments (1 reply)
    [–] 22 points 1 month ago (5 children)

    eBay has some rad Chinese mezzanine boards for these guys too. Nvlink works and everything lol 3566 file-QvfRnBhkoKQBxmqtmrGBLV

  • source
  • hideshow 5 child comments
  • [–] 21 points 1 month ago (5 children)

    Sure it's got a lot of VRAM, but the 4080 has five times the compute power.

  • source
  • hideshow 5 child comments
  • [–] 21 points 1 month ago (6 children)
  • [–] 17 points 1 month ago (3 children)

    The main thing that itches me with the V100 is the fact that given that pascal is about to be EOL, a 2017 card is probably soon next

  • source
  • hideshow 3 child comments
  • load more comments (2 replies)
    [–] 16 points 1 month ago (5 children)

    82db? Wild. I guess they don't really care that much about noise in data centers though.

  • source
  • hideshow 5 child comments
  • [–] 9 points 1 month ago

    They mostly don't, but this is also not how the datacenters cool them.

    A datacenter will either have an open water loop, or an all in one taking heat to a more advantagous place for a radiator to be, or at the very least better managed airflow with bigger fans and more specific air baffles.

    This thing has no such luxury and has a small area and unknown broader thermal context, so screaming it is to make up for the limitations of the scenario.

  • source
  • parent
  • [–] 15 points 1 month ago

    Yeah but the 0 point doing this unless you want to run AI models for some reason. These GPUs can't do video game graphics so this isn't a solution to the GPU shortage.

    This is a bit like me writing an article about NASCAR, now I can turn left whenever I want. But I haven't magically acquired a functional vehicle for a fraction of its value. I've purchased a second hand specialist product that is usually useless outside of that environment.

  • source
  • [–] 10 points 1 month ago (5 children)

    Not an AI guy, but I do like using niche hardware wrong to get results cheap. Can anyone tell me what this would be like for gaming or general computing? My 1660 super was a budget pick when I got it back in '18.

  • source
  • hideshow 5 child comments
  • load more comments (5 replies)
    load more comments
    view more: next ›