Discontinued Optane Local LLM Powers a Kimi K2.5 Desktop Run

A user on r/LocalLLaMA reported on May 12 that an Optane local LLM desktop build ran Moonshot’s Kimi K2.5 at about 4 tokens per second using discontinued Intel Optane Persistent Memory, a 12GB RTX 3060, and llama.cpp. The hardware, model family, and software path are all documentable from official sources. The 4 tokens per second figure is not vendor-confirmed — it comes from the builder’s own report. Intel Optane Persistent Memory powers a local trillion-parameter run The builder s...

Original Source

Read the full article at Dev →

KhanList aggregates and links to publicly available news content. We do not host full articles from third-party sources. Always verify important information with original sources.