Turbovec – Google's TurboQuant for vector search in Rust
内存压缩8倍还比FAISS快,免训练免调参的Rust向量索引,值得一试
Article URL: https://github.com/RyanCodrai/turbovec Comments URL: https://news.ycombinator.com/item?id=49349898 Points: 289 # Comments: 34
内存压缩8倍还比FAISS快,免训练免调参的Rust向量索引,值得一试
Article URL: https://github.com/RyanCodrai/turbovec Comments URL: https://news.ycombinator.com/item?id=49349898 Points: 289 # Comments: 34
仅400行C++代码实现1-4bit嵌入向量量化,无需训练即可极速压缩,性能与精度兼得。
Article URL: https://github.com/RunEdgeAI/turboquant.cpp Comments URL: https://news.ycombinator.com/item?id=48544682 Points: 2 # Comments: 0
RTX 5090本地部署大模型新方案:450K上下文+多模态,基于llama.cpp fork与turboquant,实测效果出色
Hi folks, I found this setup on consummer hardware that seems to have great results on local hardware. - qwen 3.6 q6 - 450 K context using turboquant …
TurboQuant号称8倍速,实测CPU端到端慢2.2倍,Qwen准确率还降17个百分点,别被合成数据骗了。
Article URL: https://deemwar-products.github.io/llama-cpu-benchmarks/ Comments URL: https://news.ycombinator.com/item?id=48212222 Points: 1 # Comments…