SuspiciousCarrot78@aussie.zone to Selfhosted@lemmy.worldEnglish · 23 hours agoDo you host your own AI?message-squaremessage-square174fedilinkarrow-up1147file-text
arrow-up1147message-squareDo you host your own AI?SuspiciousCarrot78@aussie.zone to Selfhosted@lemmy.worldEnglish · 23 hours agomessage-square174fedilinkfile-text
minus-squareSuspiciousCarrot78@aussie.zoneOPlinkfedilinkEnglisharrow-up7·edit-28 hours agoHa. You were doing inference on CPU on a haswell era. Been there, done that. OTOH…whisper.cpp is heavily optimised for it. Plus, you’re doing batch transcription, not real-time, so slow doesn’t actually matter. Fire Whisper small or medium overnight and wake up to searchable text. PS: if you want a good fast little llm, something like Qwen 3.6 2B will work well on the Xeon.
Ha. You were doing inference on CPU on a haswell era. Been there, done that.
OTOH…whisper.cpp is heavily optimised for it.
Plus, you’re doing batch transcription, not real-time, so slow doesn’t actually matter.
Fire Whisper small or medium overnight and wake up to searchable text.
PS: if you want a good fast little llm, something like Qwen 3.6 2B will work well on the Xeon.