| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Download Repo ZIP] [Original HTTPS Page] |
Branch-free unrolled fastpackwithoutmask/fastunpack for all widths, mirroring the 32-bit BitPacking. Measured on Graviton (aarch64), Corretto 21, median over widths 1..63: pack 1.93x, unpack 2.66x.
|
Merged. |
Sorry, something went wrong.
| Back | FazBrowse Home | New Git URL |
Unrolled bit-packing kernels for longs
Add unrolled fastpackwithoutmask{1..63} / fastunpack{1..63} to
LongBitPacking, mirroring the 32-bit BitPacking, and extend both
dispatchers to a 0..64 switch.
Performance
BenchmarkLongBitPacking, Graviton (aarch64), Corretto 21. Speedup of the
unrolled kernels over the generic path: