Reading up on Daniel Lemire
2 deep · digging since aug 23
- Fast unicode (UTF-8) validation with autovectorization
A branchless, autovectorized UTF-8 validation algorithm using a 4-byte lookbehind window achieves 10-80 GB/s, far outperforming branchy and DFA validators.
- Fast and Hard Code
LLMs diminish the need to learn language specifics, letting developers choose languages like Rust and Zig for speed and size, and tackle complex low‑level projects without deep expertise.