Show HN: ExpertCache – Run the Full GPT-OSS 120B (63GB) on a 16GB M1 Pro
缓存技术让63GB的GPT-OSS 120B在16GB M1 Pro上完整运行,本地推理大模型不再是梦想。
Article URL: https://github.com/amos-labs/expertcache Comments URL: https://news.ycombinator.com/item?id=49226502 Points: 1 # Comments: 0