[untitled]1 points·by mwitiderrick·3 lata temu·0 comments1 commentsPost comment[–]mwitiderrick·3 lata temureply"What’s impressive is that the sparse fine-tuned LLM can achieve 7.7 tokens per second on a single core and 26.7 tokens per second on 4 cores of a cheap consumer AMD Ryzen CPU."