Memory
Load, store, gather, scatter, and data layout patterns
Moving data between memory and SIMD registers efficiently. Data layout, arithmetic, and target-feature context all affect performance.
- Load & Store — Array loads/stores, length proofs, scalar tails
- Gather & Scatter — Checked access and the limits of the current API
- Interleaved Data —
deinterleave_4ch,interleave_4chfor RGBA and similar - Chunked Processing — Processing large arrays in SIMD-sized chunks, alignment, performance