eicker@lemmy.world to Technology@lemmy.worldEnglish · 1 day agoOpenAI Hoarding Tens Of Thousands Of Apple Mac mini And Mac Studio Devices, As ASUS And MSI Burn Through Their Entire First Batch Of NVIDIA RTX Spark Chip And Beg For More.wccftech.comexternal-linkmessage-square52linkfedilinkarrow-up11arrow-down10
arrow-up11arrow-down1external-linkOpenAI Hoarding Tens Of Thousands Of Apple Mac mini And Mac Studio Devices, As ASUS And MSI Burn Through Their Entire First Batch Of NVIDIA RTX Spark Chip And Beg For More.wccftech.comeicker@lemmy.world to Technology@lemmy.worldEnglish · 1 day agomessage-square52linkfedilink
minus-squaredeleted@lemmy.worldlinkfedilinkEnglisharrow-up0·1 day agoLocal 27b models are good enough for most tasks. Can’t wait to buy one of these from Ebay for 10% of the price next year.
minus-squareLydia_K@lemmy.worldlinkfedilinkEnglisharrow-up0·23 hours agohttps://github.com/AtomicBot-ai/atomic-llama-cpp-turboquant I’m running gwen 3.6 with 131k context window on a 3090, it’s fast enough and about as good as pay to play Claude at work.
minus-squareArchAengelus@lemmy.dbzer0.comlinkfedilinkEnglisharrow-up0·22 hours agoUpgrade that to 3.8 as soon as your hardware allows (and your use case makes sense). 3.8 is quite a bit more rational.
minus-squareLydia_K@lemmy.worldlinkfedilinkEnglisharrow-up0·20 hours agoI plan to once there is a version with turboquant and MTP as that huge context window is key.
minus-squareunexposedhazard@discuss.tchncs.delinkfedilinkEnglisharrow-up0·1 day agoYeah no way they will allow any of this hardware to go back onto the market. Anything they dont use anymore will be destroyed.
minus-squaregdog05@lemmy.worldlinkfedilinkEnglisharrow-up0·1 day agoAnd that’s really why they’re hoarding them.
minus-square4am@lemmy.ziplinkfedilinkEnglisharrow-up0·1 day agoNo, they’re hoarding them so you have to pay for cloud services they control from now on. With your little Fire tablet. No more pirating movies or political organizing for you, piggy
Local 27b models are good enough for most tasks.
Can’t wait to buy one of these from Ebay for 10% of the price next year.
https://github.com/AtomicBot-ai/atomic-llama-cpp-turboquant
I’m running gwen 3.6 with 131k context window on a 3090, it’s fast enough and about as good as pay to play Claude at work.
Upgrade that to 3.8 as soon as your hardware allows (and your use case makes sense). 3.8 is quite a bit more rational.
I plan to once there is a version with turboquant and MTP as that huge context window is key.
Yeah no way they will allow any of this hardware to go back onto the market. Anything they dont use anymore will be destroyed.
And that’s really why they’re hoarding them.
No, they’re hoarding them so you have to pay for cloud services they control from now on. With your little Fire tablet. No more pirating movies or political organizing for you, piggy