Low Level Performance Archives - Page 2 of 5

Unexpected Ways Memory Subsystem Interacts with Branch Prediction

December 26, 2023December 30, 2023Ivica BogosavljevićLow Level Performance, Memory Subsystem Performance, Performance3 Replies

We investigate the unusual way memory subsystem interacts with branch prediction and how this interaction shapes software performance.

Multithreading and the Memory Subsystem

November 30, 2023December 3, 2023Ivica BogosavljevićLow Level Performance, Memory Subsystem Performance, PerformanceLeave a Reply

In this post we investigate how the memory subsystem behaves in an environment where several threads compete for memory subsystem resources. We also investigate techniques to improve the performance of multithreaded programs – programs that split the workload onto several CPU cores so that they finish faster.

Speeding Up Translation of Virtual To Physical Memory Addresses: TLB and Huge Pages

October 29, 2023October 30, 2023Ivica BogosavljevićLow Level Performance, Memory Subsystem Performance, PerformanceLeave a Reply

In this post we explore how to speed up our memory intensive programs by decreasing the number of TLB cache misses

Faster hash maps, binary trees etc. through data layout modification

September 30, 2023October 27, 2023Ivica BogosavljevićLow Level Performance, Memory Subsystem Performance, Performance1 Reply

We investigate how to make faster hash maps, trees, linked lists and vector of pointers by changing their data layout.

Performance Through Memory Layout

August 31, 2023February 1, 2024Ivica BogosavljevićLow Level Performance, Memory Subsystem Performance, Performance2 Replies

In this post we investigate how we can improve the performance of our memory-intensive codes through changing the memory layout of our performance-critical data structures.

Measuring Memory Subsystem Performance

July 31, 2023July 31, 2023Ivica BogosavljevićLow Level Performance, Memory Subsystem Performance, PerformanceLeave a Reply

In this post we introduce a few most common tools used for memory subsystem performance debugging.

Hiding Memory Latency With In-Order CPU Cores OR How Compilers Optimize Your Code

June 26, 2023July 31, 2023Ivica BogosavljevićMemory Subsystem Performance, PerformanceLeave a Reply

We investigate techniques for hiding memory latency on in-order CPU cores. The same techniques that the compilers employ.

Software Performance and Class Layout

May 28, 2023September 28, 2023Ivica BogosavljevićMemory Subsystem Performance4 Replies

We investigate the secret connection between class layout and software performance. And of course, how to make your software faster.

Decreasing the Number of Memory Accesses: The Compiler’s Secret Life 2/2

March 30, 2023March 30, 2023Ivica BogosavljevićHelp the Compiler, Memory Subsystem PerformanceLeave a Reply

We investigate memory loads and stores that the compiler inserts for us without our knowledge: “the compiler’s secret life”. We show that these loads and stores, although necessary for the compiler are not necessary for the correct functioning of our program. And finally, we explain how you can improve the performance of your program by removing them.