Performance Archives - Page 3 of 8 - Johnny's Software Lab

Measuring Memory Subsystem Performance

July 31, 2023July 31, 2023Ivica BogosavljevićLow Level Performance, Memory Subsystem Performance, PerformanceLeave a Reply

In this post we introduce a few most common tools used for memory subsystem performance debugging.

Hiding Memory Latency With In-Order CPU Cores OR How Compilers Optimize Your Code

June 26, 2023July 31, 2023Ivica BogosavljevićMemory Subsystem Performance, PerformanceLeave a Reply

We investigate techniques for hiding memory latency on in-order CPU cores. The same techniques that the compilers employ.

Software Performance and Class Layout

May 28, 2023September 28, 2023Ivica BogosavljevićMemory Subsystem Performance4 Replies

We investigate the secret connection between class layout and software performance. And of course, how to make your software faster.

Horrible Code, Clean Performance

April 14, 2023July 31, 2023Ivica BogosavljevićHelp the Compiler14 Replies

A short tale of how horrible code yields clean performance.

Decreasing the Number of Memory Accesses: The Compiler’s Secret Life 2/2

March 30, 2023March 30, 2023Ivica BogosavljevićHelp the Compiler, Memory Subsystem PerformanceLeave a Reply

We investigate memory loads and stores that the compiler inserts for us without our knowledge: “the compiler’s secret life”. We show that these loads and stores, although necessary for the compiler are not necessary for the correct functioning of our program. And finally, we explain how you can improve the performance of your program by removing them.

Decreasing the Number of Memory Accesses 1/2

February 25, 2023April 15, 2023Ivica BogosavljevićMemory Subsystem Performance6 Replies

In this post, we are investigating a few common ways to decrease the number of memory accesses in your program.

Frugal Programming: Saving Memory Subsystem Bandwidth

January 30, 2023February 3, 2023Ivica BogosavljevićMemory Subsystem PerformanceLeave a Reply

We investigate techniques of frugal programming: how to program so you don’t waste the limited memory resources in your computer system.

Loop Optimizations: interpreting the compiler optimization report

December 15, 2022January 11, 2025Ivica BogosavljevićHelp the Compiler, Toolchain and Performance6 Replies

We introduce compiler optimization report, a useful tool if you wish to speed up your program by looking at what the compiler failed to optimize.

For Software Performance, the Way Data is Accessed Matters!

November 12, 2022March 3, 2023Ivica BogosavljevićLow Level Performance, Memory Subsystem Performance, Performance2 Replies

In our experiments with the memory access pattern, we have seen that good data locality is a key to good software performance. Accessing memory sequentially and splitting the data set into small-sized pieces which are processed individually improves data locality and software speed. In this post, we will present a few techniques to improve the…

Read

What is faster: vec.emplace_back(x) or vec[x] ?

October 24, 2022October 24, 2022Ivica BogosavljevićC++ Performance, Performance5 Replies

When we need to fill std::vector with values and the size of vector is known in advance, there are two possibilities: using emplace_back() or using operator[]. For the emplace_back() we should reserve the necessary amount of space with reserve() before emplacing into vector. This will avoid unnecessary vector regrow and benefit performance. Alternatively, if we…

Read