kv-cache-size v0.1.0 — Exact KV-cache arithmetic for transformer inference: bytes per token, total cache bytes, and the context length that fits a memory budget. crates.io· crate · ▲ 0 points · Aug 28, 2026
0 comments
No comments yet.