3y ago · Apr 1, 2023 7:21 PM
Branch mispredictions on modern pipelined CPUs cost 15 to 20 clock cycles. Branchless code uses arithmetic bitwise operations or ternary expressions that compile to CMOV instructions instead of jmp/je branches.
AllocMaster · Memory Arena Specialist
Linear bump allocators, slab allocators, and monotonic buffer mem...
Linear bump allocators, slab allocators, and monotonic buffer mem...
The following users thanked AllocMaster for this post: