SuspiciousCarrot78@aussie.zone to Selfhosted@lemmy.worldEnglish · 19 hours agoDo you host your own AI?message-squaremessage-square161fedilinkarrow-up1130arrow-down132file-text
arrow-up198arrow-down1message-squareDo you host your own AI?SuspiciousCarrot78@aussie.zone to Selfhosted@lemmy.worldEnglish · 19 hours agomessage-square161fedilinkfile-text
minus-squareSuspiciousCarrot78@aussie.zoneOPlinkfedilinkEnglisharrow-up6·edit-24 hours agoHa. You were doing inference on CPU on a haswell era. Been there, done that. OTOH…whisper.cpp is heavily optimised for it. Plus, you’re doing batch transcription, not real-time, so slow doesn’t actually matter. Fire Whisper small or medium overnight and wake up to searchable text. PS: if you want a good fast little llm, something like Qwen 3.6 2B will work well on the Xeon.
Ha. You were doing inference on CPU on a haswell era. Been there, done that.
OTOH…whisper.cpp is heavily optimised for it.
Plus, you’re doing batch transcription, not real-time, so slow doesn’t actually matter.
Fire Whisper small or medium overnight and wake up to searchable text.
PS: if you want a good fast little llm, something like Qwen 3.6 2B will work well on the Xeon.