misk@sopuli.xyz to Technology@lemmy.worldEnglish · edit-25 days agoApple Intelligence summary botches a headline, causing jitters in BBC newsroomwww.theregister.comexternal-linkmessage-square68fedilinkarrow-up1248arrow-down14
arrow-up1244arrow-down1external-linkApple Intelligence summary botches a headline, causing jitters in BBC newsroomwww.theregister.commisk@sopuli.xyz to Technology@lemmy.worldEnglish · edit-25 days agomessage-square68fedilink
minus-squarebrucethemoose@lemmy.worldlinkfedilinkEnglisharrow-up4·5 days agoFor RAG data? It works. But its too slow for the weights. What generative models fundamentally do is run a full pass through the multi-gigabyte weights for every ‘word’ or diffusion step, so even 128-bit DDR5 like you find on desktop CPUs is too slow.
For RAG data? It works.
But its too slow for the weights. What generative models fundamentally do is run a full pass through the multi-gigabyte weights for every ‘word’ or diffusion step, so even 128-bit DDR5 like you find on desktop CPUs is too slow.