Workstation on turbocharging LLMs: PagedAttention paging for KV cache, vLLM serving, Self-Debugging for agents, PowerInfer token rates, and EG-MLA memory cuts — with production trade-offs.
Balinder WaliaRead
2026-08-26
Rust Async Blocking, Rayon & Modern Applications
Workstation on async Rust blocking, Rayon for CPU-bound work, and why Rust fits modern APIs, edge, agents, and CI/CD — inspired by Alice Ryhl’s Rayon guidance.
Balinder WaliaRead
2026-05-16
Polyglot Benchmarks: Choosing the Right Tool for the Right Job
Stop default-stack bias. Polyglot Benchmarks compares njs, Lua, Python, Go, Rust, Bun, Java, and Kotlin on the same workloads — reproducible harness + live dashboard for architecture decisions.