eicker@lemmy.world to Technology@lemmy.worldEnglish · 24 days agoOpenAI Hoarding Tens Of Thousands Of Apple Mac mini And Mac Studio Devices, As ASUS And MSI Burn Through Their Entire First Batch Of NVIDIA RTX Spark Chip And Beg For More.wccftech.comexternal-linkmessage-square58linkfedilinkarrow-up11arrow-down10
arrow-up11arrow-down1external-linkOpenAI Hoarding Tens Of Thousands Of Apple Mac mini And Mac Studio Devices, As ASUS And MSI Burn Through Their Entire First Batch Of NVIDIA RTX Spark Chip And Beg For More.wccftech.comeicker@lemmy.world to Technology@lemmy.worldEnglish · 24 days agomessage-square58linkfedilink
minus-squaredeleted@lemmy.worldlinkfedilinkEnglisharrow-up0·24 days agoLocal 27b models are good enough for most tasks. Can’t wait to buy one of these from Ebay for 10% of the price next year.
minus-squaregdog05@lemmy.worldlinkfedilinkEnglisharrow-up0·24 days agoAnd that’s really why they’re hoarding them.
minus-square4am@lemmy.ziplinkfedilinkEnglisharrow-up0·23 days agoNo, they’re hoarding them so you have to pay for cloud services they control from now on. With your little Fire tablet. No more pirating movies or political organizing for you, piggy
minus-squareunexposedhazard@discuss.tchncs.delinkfedilinkEnglisharrow-up0·24 days agoYeah no way they will allow any of this hardware to go back onto the market. Anything they dont use anymore will be destroyed.
minus-squareberty@feddit.orglinkfedilinkEnglisharrow-up0·23 days agoBuy it, destroy it. Just like buying bunch of old books, train their LLM’s and burn it. Humanity has gone a long way to be that stupid.
minus-squareLydia_K@lemmy.worldlinkfedilinkEnglisharrow-up0·23 days agohttps://github.com/AtomicBot-ai/atomic-llama-cpp-turboquant I’m running gwen 3.6 with 131k context window on a 3090, it’s fast enough and about as good as pay to play Claude at work.
minus-squareArchAengelus@lemmy.dbzer0.comlinkfedilinkEnglisharrow-up0·23 days agoUpgrade that to 3.8 as soon as your hardware allows (and your use case makes sense). 3.8 is quite a bit more rational.
minus-squareLydia_K@lemmy.worldlinkfedilinkEnglisharrow-up0·23 days agoI plan to once there is a version with turboquant and MTP as that huge context window is key.
Local 27b models are good enough for most tasks.
Can’t wait to buy one of these from Ebay for 10% of the price next year.
And that’s really why they’re hoarding them.
No, they’re hoarding them so you have to pay for cloud services they control from now on. With your little Fire tablet. No more pirating movies or political organizing for you, piggy
Yeah no way they will allow any of this hardware to go back onto the market. Anything they dont use anymore will be destroyed.
Buy it, destroy it. Just like buying bunch of old books, train their LLM’s and burn it. Humanity has gone a long way to be that stupid.
https://github.com/AtomicBot-ai/atomic-llama-cpp-turboquant
I’m running gwen 3.6 with 131k context window on a 3090, it’s fast enough and about as good as pay to play Claude at work.
Upgrade that to 3.8 as soon as your hardware allows (and your use case makes sense). 3.8 is quite a bit more rational.
I plan to once there is a version with turboquant and MTP as that huge context window is key.