▲ 334 ▼ OpenAI Hoarding Tens Of Thousands Of Apple Mac mini And Mac Studio Devices, As ASUS And MSI Burn Through Their Entire First Batch Of NVIDIA RTX Spark Chip And Beg For More. (wccftech.com) submitted 1 month ago by eicker@lemmy.world to c/technology@lemmy.world 79 comments fedilink hide all child comments
[–] deleted@lemmy.world 35 points 1 month ago* (19 children) Local 27b models are good enough for most tasks. Can’t wait to buy one of these from Ebay for 10% of the price next year. permalink fedilink source hideshow 19 child comments replies: [–] gdog05@lemmy.world 24 points 1 month ago (2 children) And that's really why they're hoarding them. permalink fedilink source parent hideshow 2 child comments replies: [–] 4am@lemmy.zip 14 points 1 month ago (1 child) No, they’re hoarding them so you have to pay for cloud services they control from now on. With your little Fire tablet. No more pirating movies or political organizing for you, piggy permalink fedilink source parent hideshow 1 child comment replies: [+] lyrial@anarchist.nexus 1 point 1 month ago* (last edited 2 weeks ago) [deleted] permalink fedilink source parent [–] unexposedhazard@discuss.tchncs.de 21 points 1 month ago (1 child) Yeah no way they will allow any of this hardware to go back onto the market. Anything they dont use anymore will be destroyed. permalink fedilink source parent hideshow 1 child comment replies: [–] berty@feddit.org 5 points 1 month ago Buy it, destroy it. Just like buying bunch of old books, train their LLM's and burn it. Humanity has gone a long way to be that stupid. permalink fedilink source parent [–] Lydia_K@lemmy.world 7 points 1 month ago (3 children) https://github.com/AtomicBot-ai/atomic-llama-cpp-turboquant I'm running gwen 3.6 with 131k context window on a 3090, it's fast enough and about as good as pay to play Claude at work. permalink fedilink source parent hideshow 3 child comments replies: [–] ArchAengelus@lemmy.dbzer0.com 3 points 1 month ago (1 child) Upgrade that to 3.8 as soon as your hardware allows (and your use case makes sense). 3.8 is quite a bit more rational. permalink fedilink source parent hideshow 1 child comment replies: [–] Lydia_K@lemmy.world 1 point 1 month ago I plan to once there is a version with turboquant and MTP as that huge context window is key. permalink fedilink source parent [–] boonhet@sopuli.xyz 1 point 1 month ago A used 3090 is like 2-3k though, IF you can find one :| A month of claude is like 20 EUR. A month of opencode go is half that, but you get less usage. permalink fedilink source parent [–] Chee_Koala@lemmy.world 4 points 1 month ago* (last edited 1 month ago) (9 children) Any 27b Model you can currently recommend for a 16gb AMD ? Mostly coding tasks but not exclusively. permalink fedilink source parent hideshow 9 child comments replies: [–] abcdqfr@lemmy.world 6 points 1 month ago* (1 child) There is a way. There was a post yesterday on exactly this, let me find it... https://lemmy.world/post/51283416 permalink fedilink source parent hideshow 1 child comment replies: [–] Chee_Koala@lemmy.world 2 points 1 month ago Thx I'll give this a try! permalink fedilink source parent [–] mierdabird@lemmy.dbzer0.com 5 points 1 month ago (4 children) 9060xt 16gb is the most cost effective new GPU, but if you're going used look for a V620 on eBay. It's a 6800xt chip but in server form factor GPU with 32GB vram. Can be a bit of a pain to set up but by far the most cost effective option IMO permalink fedilink source parent hideshow 4 child comments replies: [–] Darkaga@lemmy.world 6 points 1 month ago (2 children) V620s were a good deal when you could get them for $350, now they’re $700+ and no longer a good deal. permalink fedilink source parent hideshow 2 child comments replies: [–] floofloof@lemmy.ca 5 points 1 month ago* (last edited 1 month ago) There just aren't any good deals any more. Prices for everything have gone crazy in the last few months. For coding LLMs the cloud services may now be the least worst value, by design, until they hike the prices. That said, I still just paid way too much for a used graphics card so I could do many things locally, because I just don't want to give the likes of Sam Altman a single penny. permalink fedilink source parent [–] mierdabird@lemmy.dbzer0.com 4 points 1 month ago Oh wow you aren't kidding. The dude I bought from on eBay @ $350 in February is sold out now. Guess I retract my statement. This AI pricing is wrecking every deal on the market lol permalink fedilink source parent [–] Chee_Koala@lemmy.world 2 points 1 month ago Thanks! actually have a 6800xt already and was asking for model tips. I saw my question was easily read as asking for GFX card tips, edited. permalink fedilink source parent [–] deleted@lemmy.world 2 points 1 month ago* (1 child) For your hardware, the VRam is not enough to run 27b but, I’d recommend Qwen 3.5 9b for image / text to text. And I’m planning to experiment with Qwen 3.8 9b for text to text. 4_k_m quantization is the sweet spot for performance and ram usage. Also, I find Llama cpp is better than Ollama in terms of performance. permalink fedilink source parent hideshow 1 child comment replies: [–] Chee_Koala@lemmy.world 1 point 1 month ago Thx for the tips! permalink fedilink source parent
[–] gdog05@lemmy.world 24 points 1 month ago (2 children) And that's really why they're hoarding them. permalink fedilink source parent hideshow 2 child comments replies: [–] 4am@lemmy.zip 14 points 1 month ago (1 child) No, they’re hoarding them so you have to pay for cloud services they control from now on. With your little Fire tablet. No more pirating movies or political organizing for you, piggy permalink fedilink source parent hideshow 1 child comment replies: [+] lyrial@anarchist.nexus 1 point 1 month ago* (last edited 2 weeks ago) [deleted] permalink fedilink source parent
[–] 4am@lemmy.zip 14 points 1 month ago (1 child) No, they’re hoarding them so you have to pay for cloud services they control from now on. With your little Fire tablet. No more pirating movies or political organizing for you, piggy permalink fedilink source parent hideshow 1 child comment replies: [+] lyrial@anarchist.nexus 1 point 1 month ago* (last edited 2 weeks ago) [deleted] permalink fedilink source parent
[+] lyrial@anarchist.nexus 1 point 1 month ago* (last edited 2 weeks ago) [deleted] permalink fedilink source parent
[–] unexposedhazard@discuss.tchncs.de 21 points 1 month ago (1 child) Yeah no way they will allow any of this hardware to go back onto the market. Anything they dont use anymore will be destroyed. permalink fedilink source parent hideshow 1 child comment replies: [–] berty@feddit.org 5 points 1 month ago Buy it, destroy it. Just like buying bunch of old books, train their LLM's and burn it. Humanity has gone a long way to be that stupid. permalink fedilink source parent
[–] berty@feddit.org 5 points 1 month ago Buy it, destroy it. Just like buying bunch of old books, train their LLM's and burn it. Humanity has gone a long way to be that stupid. permalink fedilink source parent
[–] Lydia_K@lemmy.world 7 points 1 month ago (3 children) https://github.com/AtomicBot-ai/atomic-llama-cpp-turboquant I'm running gwen 3.6 with 131k context window on a 3090, it's fast enough and about as good as pay to play Claude at work. permalink fedilink source parent hideshow 3 child comments replies: [–] ArchAengelus@lemmy.dbzer0.com 3 points 1 month ago (1 child) Upgrade that to 3.8 as soon as your hardware allows (and your use case makes sense). 3.8 is quite a bit more rational. permalink fedilink source parent hideshow 1 child comment replies: [–] Lydia_K@lemmy.world 1 point 1 month ago I plan to once there is a version with turboquant and MTP as that huge context window is key. permalink fedilink source parent [–] boonhet@sopuli.xyz 1 point 1 month ago A used 3090 is like 2-3k though, IF you can find one :| A month of claude is like 20 EUR. A month of opencode go is half that, but you get less usage. permalink fedilink source parent
[–] ArchAengelus@lemmy.dbzer0.com 3 points 1 month ago (1 child) Upgrade that to 3.8 as soon as your hardware allows (and your use case makes sense). 3.8 is quite a bit more rational. permalink fedilink source parent hideshow 1 child comment replies: [–] Lydia_K@lemmy.world 1 point 1 month ago I plan to once there is a version with turboquant and MTP as that huge context window is key. permalink fedilink source parent
[–] Lydia_K@lemmy.world 1 point 1 month ago I plan to once there is a version with turboquant and MTP as that huge context window is key. permalink fedilink source parent
[–] boonhet@sopuli.xyz 1 point 1 month ago A used 3090 is like 2-3k though, IF you can find one :| A month of claude is like 20 EUR. A month of opencode go is half that, but you get less usage. permalink fedilink source parent
[–] Chee_Koala@lemmy.world 4 points 1 month ago* (last edited 1 month ago) (9 children) Any 27b Model you can currently recommend for a 16gb AMD ? Mostly coding tasks but not exclusively. permalink fedilink source parent hideshow 9 child comments replies: [–] abcdqfr@lemmy.world 6 points 1 month ago* (1 child) There is a way. There was a post yesterday on exactly this, let me find it... https://lemmy.world/post/51283416 permalink fedilink source parent hideshow 1 child comment replies: [–] Chee_Koala@lemmy.world 2 points 1 month ago Thx I'll give this a try! permalink fedilink source parent [–] mierdabird@lemmy.dbzer0.com 5 points 1 month ago (4 children) 9060xt 16gb is the most cost effective new GPU, but if you're going used look for a V620 on eBay. It's a 6800xt chip but in server form factor GPU with 32GB vram. Can be a bit of a pain to set up but by far the most cost effective option IMO permalink fedilink source parent hideshow 4 child comments replies: [–] Darkaga@lemmy.world 6 points 1 month ago (2 children) V620s were a good deal when you could get them for $350, now they’re $700+ and no longer a good deal. permalink fedilink source parent hideshow 2 child comments replies: [–] floofloof@lemmy.ca 5 points 1 month ago* (last edited 1 month ago) There just aren't any good deals any more. Prices for everything have gone crazy in the last few months. For coding LLMs the cloud services may now be the least worst value, by design, until they hike the prices. That said, I still just paid way too much for a used graphics card so I could do many things locally, because I just don't want to give the likes of Sam Altman a single penny. permalink fedilink source parent [–] mierdabird@lemmy.dbzer0.com 4 points 1 month ago Oh wow you aren't kidding. The dude I bought from on eBay @ $350 in February is sold out now. Guess I retract my statement. This AI pricing is wrecking every deal on the market lol permalink fedilink source parent [–] Chee_Koala@lemmy.world 2 points 1 month ago Thanks! actually have a 6800xt already and was asking for model tips. I saw my question was easily read as asking for GFX card tips, edited. permalink fedilink source parent [–] deleted@lemmy.world 2 points 1 month ago* (1 child) For your hardware, the VRam is not enough to run 27b but, I’d recommend Qwen 3.5 9b for image / text to text. And I’m planning to experiment with Qwen 3.8 9b for text to text. 4_k_m quantization is the sweet spot for performance and ram usage. Also, I find Llama cpp is better than Ollama in terms of performance. permalink fedilink source parent hideshow 1 child comment replies: [–] Chee_Koala@lemmy.world 1 point 1 month ago Thx for the tips! permalink fedilink source parent
[–] abcdqfr@lemmy.world 6 points 1 month ago* (1 child) There is a way. There was a post yesterday on exactly this, let me find it... https://lemmy.world/post/51283416 permalink fedilink source parent hideshow 1 child comment replies: [–] Chee_Koala@lemmy.world 2 points 1 month ago Thx I'll give this a try! permalink fedilink source parent
[–] Chee_Koala@lemmy.world 2 points 1 month ago Thx I'll give this a try! permalink fedilink source parent
[–] mierdabird@lemmy.dbzer0.com 5 points 1 month ago (4 children) 9060xt 16gb is the most cost effective new GPU, but if you're going used look for a V620 on eBay. It's a 6800xt chip but in server form factor GPU with 32GB vram. Can be a bit of a pain to set up but by far the most cost effective option IMO permalink fedilink source parent hideshow 4 child comments replies: [–] Darkaga@lemmy.world 6 points 1 month ago (2 children) V620s were a good deal when you could get them for $350, now they’re $700+ and no longer a good deal. permalink fedilink source parent hideshow 2 child comments replies: [–] floofloof@lemmy.ca 5 points 1 month ago* (last edited 1 month ago) There just aren't any good deals any more. Prices for everything have gone crazy in the last few months. For coding LLMs the cloud services may now be the least worst value, by design, until they hike the prices. That said, I still just paid way too much for a used graphics card so I could do many things locally, because I just don't want to give the likes of Sam Altman a single penny. permalink fedilink source parent [–] mierdabird@lemmy.dbzer0.com 4 points 1 month ago Oh wow you aren't kidding. The dude I bought from on eBay @ $350 in February is sold out now. Guess I retract my statement. This AI pricing is wrecking every deal on the market lol permalink fedilink source parent [–] Chee_Koala@lemmy.world 2 points 1 month ago Thanks! actually have a 6800xt already and was asking for model tips. I saw my question was easily read as asking for GFX card tips, edited. permalink fedilink source parent
[–] Darkaga@lemmy.world 6 points 1 month ago (2 children) V620s were a good deal when you could get them for $350, now they’re $700+ and no longer a good deal. permalink fedilink source parent hideshow 2 child comments replies: [–] floofloof@lemmy.ca 5 points 1 month ago* (last edited 1 month ago) There just aren't any good deals any more. Prices for everything have gone crazy in the last few months. For coding LLMs the cloud services may now be the least worst value, by design, until they hike the prices. That said, I still just paid way too much for a used graphics card so I could do many things locally, because I just don't want to give the likes of Sam Altman a single penny. permalink fedilink source parent [–] mierdabird@lemmy.dbzer0.com 4 points 1 month ago Oh wow you aren't kidding. The dude I bought from on eBay @ $350 in February is sold out now. Guess I retract my statement. This AI pricing is wrecking every deal on the market lol permalink fedilink source parent
[–] floofloof@lemmy.ca 5 points 1 month ago* (last edited 1 month ago) There just aren't any good deals any more. Prices for everything have gone crazy in the last few months. For coding LLMs the cloud services may now be the least worst value, by design, until they hike the prices. That said, I still just paid way too much for a used graphics card so I could do many things locally, because I just don't want to give the likes of Sam Altman a single penny. permalink fedilink source parent
[–] mierdabird@lemmy.dbzer0.com 4 points 1 month ago Oh wow you aren't kidding. The dude I bought from on eBay @ $350 in February is sold out now. Guess I retract my statement. This AI pricing is wrecking every deal on the market lol permalink fedilink source parent
[–] Chee_Koala@lemmy.world 2 points 1 month ago Thanks! actually have a 6800xt already and was asking for model tips. I saw my question was easily read as asking for GFX card tips, edited. permalink fedilink source parent
[–] deleted@lemmy.world 2 points 1 month ago* (1 child) For your hardware, the VRam is not enough to run 27b but, I’d recommend Qwen 3.5 9b for image / text to text. And I’m planning to experiment with Qwen 3.8 9b for text to text. 4_k_m quantization is the sweet spot for performance and ram usage. Also, I find Llama cpp is better than Ollama in terms of performance. permalink fedilink source parent hideshow 1 child comment replies: [–] Chee_Koala@lemmy.world 1 point 1 month ago Thx for the tips! permalink fedilink source parent