POCKET-35B agentic model on cpu 59 t/s
2.40T2 sourcer/LocalLLaMA
r/LocalLLaMAoriginal source ↗
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1v6zseq/pocket35b_agentic_model_on_cpu_59_ts/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryA Reddit post on r/LocalLLaMA reports a 'POCKET-35B' model intended for agentic use achieving 59 tokens per second on CPU. No further content available beyond the title.
Why it mattersA 35B-class agentic model running at 59 t/s on CPU is a concrete local-inference data point useful for sizing agent setups without a GPU.
Cited by
No citations on record.
