linux

mirror of https://github.com/hardkernel/linux.git synced 2026-06-07 03:15:31 +09:00

Go to file

Marco Elver 0370dc314d perf/hw_breakpoint: Optimize list of per-task breakpoints

On a machine with 256 CPUs, running the recently added perf breakpoint
benchmark results in:

 | $> perf bench -r 30 breakpoint thread -b 4 -p 64 -t 64
 | # Running 'breakpoint/thread' benchmark:
 | # Created/joined 30 threads with 4 breakpoints and 64 parallelism
 |      Total time: 236.418 [sec]
 |
 |   123134.794271 usecs/op
 |  7880626.833333 usecs/op/cpu

The benchmark tests inherited breakpoint perf events across many
threads.

Looking at a perf profile, we can see that the majority of the time is
spent in various hw_breakpoint.c functions, which execute within the
'nr_bp_mutex' critical sections which then results in contention on that
mutex as well:

    37.27%  [kernel]       [k] osq_lock
    34.92%  [kernel]       [k] mutex_spin_on_owner
    12.15%  [kernel]       [k] toggle_bp_slot
    11.90%  [kernel]       [k] __reserve_bp_slot

The culprit here is task_bp_pinned(), which has a runtime complexity of
O(#tasks) due to storing all task breakpoints in the same list and
iterating through that list looking for a matching task. Clearly, this
does not scale to thousands of tasks.

Instead, make use of the "rhashtable" variant "rhltable" which stores
multiple items with the same key in a list. This results in average
runtime complexity of O(1) for task_bp_pinned().

With the optimization, the benchmark shows:

 | $> perf bench -r 30 breakpoint thread -b 4 -p 64 -t 64
 | # Running 'breakpoint/thread' benchmark:
 | # Created/joined 30 threads with 4 breakpoints and 64 parallelism
 |      Total time: 0.208 [sec]
 |
 |      108.422396 usecs/op
 |     6939.033333 usecs/op/cpu

On this particular setup that's a speedup of ~1135x.

While one option would be to make task_struct a breakpoint list node,
this would only further bloat task_struct for infrequently used data.
Furthermore, after all optimizations in this series, there's no evidence
it would result in better performance: later optimizations make the time
spent looking up entries in the hash table negligible (we'll reach the
theoretical ideal performance i.e. no constraints).

Signed-off-by: Marco Elver <elver@google.com>
Signed-off-by: Peter Zijlstra (Intel) <peterz@infradead.org>
Reviewed-by: Dmitry Vyukov <dvyukov@google.com>
Acked-by: Ian Rogers <irogers@google.com>
Link: https://lore.kernel.org/r/20220829124719.675715-5-elver@google.com

2022-08-30 10:56:21 +02:00

arch

perf: Add system error and not in transaction branch types

2022-08-29 09:42:41 +02:00

block

Merge tag 'block-6.0-2022-08-19' of git://git.kernel.dk/linux-block

2022-08-20 10:17:05 -07:00

certs

Merge tag 'kbuild-v5.20' of git://git.kernel.org/pub/scm/linux/kernel/git/masahiroy/linux-kbuild

2022-08-10 10:40:41 -07:00

crypto

crypto: blake2b: effectively disable frame size warning

2022-08-10 17:59:11 -07:00

Documentation

asm goto: eradicate CC_HAS_ASM_GOTO

2022-08-21 10:06:28 -07:00

drivers

Merge tag 'irq-urgent-2022-08-21' of git://git.kernel.org/pub/scm/linux/kernel/git/tip/tip

2022-08-21 15:09:55 -07:00

Merge tag '6.0-rc1-smb3-client-fixes' of git://git.samba.org/sfrench/cifs-2.6

2022-08-21 10:21:16 -07:00

include

perf/hw_breakpoint: Optimize list of per-task breakpoints

2022-08-30 10:56:21 +02:00

init

asm goto: eradicate CC_HAS_ASM_GOTO

2022-08-21 10:06:28 -07:00

io_uring

io_uring/net: use right helpers for async_data

2022-08-18 07:27:20 -06:00

ipc

Merge tag 'mm-nonmm-stable-2022-08-06-2' of git://git.kernel.org/pub/scm/linux/kernel/git/akpm/mm

2022-08-07 10:03:24 -07:00

kernel

perf/hw_breakpoint: Optimize list of per-task breakpoints

2022-08-30 10:56:21 +02:00

lib

perf/hw_breakpoint: Add KUnit test for constraints accounting

2022-08-30 10:56:20 +02:00

LICENSES

…

Merge tag 'mm-stable-2022-08-09' of git://git.kernel.org/pub/scm/linux/kernel/git/akpm/mm

2022-08-10 11:18:00 -07:00

net

tcp: handle pure FIN case correctly

2022-08-18 11:04:56 -07:00

samples

Merge tag 'trace-v6.0' of git://git.kernel.org/pub/scm/linux/kernel/git/rostedt/linux-trace

2022-08-05 09:41:12 -07:00

scripts

asm goto: eradicate CC_HAS_ASM_GOTO

2022-08-21 10:06:28 -07:00

security

Merge tag 'hardening-v6.0-rc2' of git://git.kernel.org/pub/scm/linux/kernel/git/kees/linux

2022-08-19 13:56:14 -07:00

sound

Merge tag 'sound-6.0-rc2' of git://git.kernel.org/pub/scm/linux/kernel/git/tiwai/sound

2022-08-19 09:46:11 -07:00

tools

asm goto: eradicate CC_HAS_ASM_GOTO

2022-08-21 10:06:28 -07:00

usr

Merge tag 'mm-nonmm-stable-2022-05-26' of git://git.kernel.org/pub/scm/linux/kernel/git/akpm/mm

2022-05-27 11:22:03 -07:00

virt

KVM: Drop unnecessary initialization of "ops" in kvm_ioctl_create_device()

2022-08-19 04:05:43 -04:00

.clang-format

PCI/DOE: Add DOE mailbox support functions

2022-07-19 15:38:04 -07:00

.cocciconfig

…

.get_maintainer.ignore

…

.gitattributes

…

.gitignore

kbuild: split the second line of *.mod into *.usyms

2022-05-08 03:16:59 +09:00

.mailmap

Merge tag 'mm-nonmm-stable-2022-08-06-2' of git://git.kernel.org/pub/scm/linux/kernel/git/akpm/mm

2022-08-07 10:03:24 -07:00

COPYING

…

CREDITS

Merge tag 'drm-next-2022-08-03' of git://anongit.freedesktop.org/drm/drm

2022-08-03 19:52:08 -07:00

Kbuild

…

Kconfig

…

MAINTAINERS

Merge tag '6.0-rc1-smb3-client-fixes' of git://git.samba.org/sfrench/cifs-2.6

2022-08-21 10:21:16 -07:00

Makefile

Linux 6.0-rc2

2022-08-21 17:32:54 -07:00

README

…

README

Linux kernel
============

There are several guides for kernel developers and users. These guides can
be rendered in a number of formats, like HTML and PDF. Please read
Documentation/admin-guide/README.rst first.

In order to build the documentation, use ``make htmldocs`` or
``make pdfdocs``.  The formatted documentation can also be read online at:

    https://www.kernel.org/doc/html/latest/

There are various text files in the Documentation/ subdirectory,
several of them using the Restructured Text markup notation.

Please read the Documentation/process/changes.rst file, as it contains the
requirements for building and running the kernel, and information about
the problems which may result by upgrading your kernel.