From: weilin.wang@intel.com
To: weilin.wang@intel.com, Ian Rogers <irogers@google.com>,
Peter Zijlstra <peterz@infradead.org>,
Ingo Molnar <mingo@redhat.com>,
Arnaldo Carvalho de Melo <acme@kernel.org>,
Alexander Shishkin <alexander.shishkin@linux.intel.com>,
Jiri Olsa <jolsa@kernel.org>, Namhyung Kim <namhyung@kernel.org>,
Adrian Hunter <adrian.hunter@intel.com>,
Kan Liang <kan.liang@linux.intel.com>
Cc: linux-perf-users@vger.kernel.org, linux-kernel@vger.kernel.org,
Perry Taylor <perry.taylor@intel.com>,
Samantha Alt <samantha.alt@intel.com>,
Caleb Biggers <caleb.biggers@intel.com>,
Mark Rutland <mark.rutland@arm.com>,
Yang Jihong <yangjihong1@huawei.com>
Subject: [RFC PATCH v2 00/17] Perf stat metric grouping with hardware information
Date: Fri, 13 Oct 2023 18:51:45 -0700 [thread overview]
Message-ID: <20231014015202.1175377-1-weilin.wang@intel.com> (raw)
From: Weilin Wang <weilin.wang@intel.com>
Chanages in v2:
- Updated the ordering of patches and removed bug fixing patches [Kan]
- Move code into dedicated patches or more related patches [Kan]
- Removed unnecessary code [Kan, Yang]
Grouping algorithm details:
- Currently, this grouping method groups a combined list of events, which is a
unique list of events for the required metrics.
- The grouping process traverses the event list to build up a event_info_list
that includes counter restriction rules of each event. It also creates a
pmu_info_list for the counter info - the number of gp counters and fixed
counters available in each PMU in current system. When grouping is in progress,
there are two bitmaps used to represent usages of gp counters and fixed
counters in each group. These two bitmaps are initialized based on the two
counter sizes of the corresponding PMU.
Assumption: this method reads counter availablility from data described in JSON
files and builds up groups based on this info.
- The grouping step traverses the event_info_list. For each event, it checks
bitmaps to get counter availability in current group. It sets the bit and
inserts event into the group if found a vacant counter that fits the
restrictions of the event. Otherwise, it creates a new group, places event in
it, and then inserts it into the group list.
v1: https://lore.kernel.org/all/20230925061824.3818631-1-weilin.wang@intel.com/
Weilin Wang (17):
perf stat: Add hardware-grouping cmd option to perf stat
perf stat: Add basic functions for the hardware-grouping stat cmd
option
perf pmu-events: Add functions in jevent.py to parse counter and event
info for hardware aware grouping
perf pmu-events: Add counter info into JSON files for SapphireRapids
perf pmu-events: Add event counter data for Cascadelakex
perf pmu-events: Add event counter data for Icelakex
perf stat: Add functions to set counter bitmaps for hardware-grouping
method
perf stat: Add functions to get counter info
perf stat: Add functions to create new group and assign events into
groups for hardware-grouping method
perf stat: Add build string function and topdown events handling in
hardware-grouping
perf stat: Add function to handle special events in hardware-grouping
perf stat: Add function to combine metrics for hardware-grouping
perf stat: Handle taken alone in hardware-grouping
perf stat: Handle NMI in hardware-grouping
perf stat: Code refactoring in hardware-grouping
perf stat: Add tool events support in hardware-grouping
perf pmu-events: Add event counter data for Tigerlake
tools/lib/bitmap.c | 20 +
tools/perf/builtin-stat.c | 7 +
.../arch/x86/cascadelakex/cache.json | 1237 ++++++++++++
.../arch/x86/cascadelakex/counter.json | 17 +
.../arch/x86/cascadelakex/floating-point.json | 16 +
.../arch/x86/cascadelakex/frontend.json | 68 +
.../arch/x86/cascadelakex/memory.json | 751 ++++++++
.../arch/x86/cascadelakex/other.json | 168 ++
.../arch/x86/cascadelakex/pipeline.json | 102 +
.../arch/x86/cascadelakex/uncore-cache.json | 1138 +++++++++++
.../x86/cascadelakex/uncore-interconnect.json | 1272 +++++++++++++
.../arch/x86/cascadelakex/uncore-io.json | 394 ++++
.../arch/x86/cascadelakex/uncore-memory.json | 509 +++++
.../arch/x86/cascadelakex/uncore-power.json | 25 +
.../arch/x86/cascadelakex/virtual-memory.json | 28 +
.../pmu-events/arch/x86/icelakex/cache.json | 98 +
.../pmu-events/arch/x86/icelakex/counter.json | 17 +
.../arch/x86/icelakex/floating-point.json | 13 +
.../arch/x86/icelakex/frontend.json | 55 +
.../pmu-events/arch/x86/icelakex/memory.json | 53 +
.../pmu-events/arch/x86/icelakex/other.json | 52 +
.../arch/x86/icelakex/pipeline.json | 92 +
.../arch/x86/icelakex/uncore-cache.json | 965 ++++++++++
.../x86/icelakex/uncore-interconnect.json | 1667 +++++++++++++++++
.../arch/x86/icelakex/uncore-io.json | 966 ++++++++++
.../arch/x86/icelakex/uncore-memory.json | 186 ++
.../arch/x86/icelakex/uncore-power.json | 26 +
.../arch/x86/icelakex/virtual-memory.json | 22 +
.../arch/x86/sapphirerapids/cache.json | 104 +
.../arch/x86/sapphirerapids/counter.json | 17 +
.../x86/sapphirerapids/floating-point.json | 25 +
.../arch/x86/sapphirerapids/frontend.json | 98 +-
.../arch/x86/sapphirerapids/memory.json | 44 +
.../arch/x86/sapphirerapids/other.json | 40 +
.../arch/x86/sapphirerapids/pipeline.json | 118 ++
.../arch/x86/sapphirerapids/uncore-cache.json | 534 +++++-
.../arch/x86/sapphirerapids/uncore-cxl.json | 56 +
.../sapphirerapids/uncore-interconnect.json | 476 +++++
.../arch/x86/sapphirerapids/uncore-io.json | 373 ++++
.../x86/sapphirerapids/uncore-memory.json | 391 ++++
.../arch/x86/sapphirerapids/uncore-power.json | 24 +
.../x86/sapphirerapids/virtual-memory.json | 20 +
.../pmu-events/arch/x86/tigerlake/cache.json | 65 +
.../arch/x86/tigerlake/counter.json | 7 +
.../arch/x86/tigerlake/floating-point.json | 13 +
.../arch/x86/tigerlake/frontend.json | 56 +
.../pmu-events/arch/x86/tigerlake/memory.json | 31 +
.../pmu-events/arch/x86/tigerlake/other.json | 4 +
.../arch/x86/tigerlake/pipeline.json | 96 +
.../x86/tigerlake/uncore-interconnect.json | 11 +
.../arch/x86/tigerlake/uncore-memory.json | 6 +
.../arch/x86/tigerlake/uncore-other.json | 1 +
.../arch/x86/tigerlake/virtual-memory.json | 20 +
tools/perf/pmu-events/jevents.py | 177 +-
tools/perf/pmu-events/pmu-events.h | 26 +-
tools/perf/util/metricgroup.c | 947 ++++++++++
tools/perf/util/metricgroup.h | 106 ++
tools/perf/util/pmu.c | 5 +
tools/perf/util/pmu.h | 1 +
tools/perf/util/stat.h | 1 +
60 files changed, 13832 insertions(+), 25 deletions(-)
create mode 100644 tools/perf/pmu-events/arch/x86/cascadelakex/counter.json
create mode 100644 tools/perf/pmu-events/arch/x86/icelakex/counter.json
create mode 100644 tools/perf/pmu-events/arch/x86/sapphirerapids/counter.json
create mode 100644 tools/perf/pmu-events/arch/x86/tigerlake/counter.json
--
2.39.3
next reply other threads:[~2023-10-14 1:52 UTC|newest]
Thread overview: 22+ messages / expand[flat|nested] mbox.gz Atom feed top
2023-10-14 1:51 weilin.wang [this message]
2023-10-14 1:51 ` [RFC PATCH v2 01/17] perf stat: Add hardware-grouping cmd option to perf stat weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 02/17] perf stat: Add basic functions for the hardware-grouping stat cmd option weilin.wang
2023-10-30 15:07 ` Liang, Kan
2023-10-14 1:51 ` [RFC PATCH v2 03/17] perf pmu-events: Add functions in jevent.py to parse counter and event info for hardware aware grouping weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 04/17] perf pmu-events: Add counter info into JSON files for SapphireRapids weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 05/17] perf pmu-events: Add event counter data for Cascadelakex weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 06/17] perf pmu-events: Add event counter data for Icelakex weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 07/17] perf stat: Add functions to set counter bitmaps for hardware-grouping method weilin.wang
2023-10-30 15:32 ` Liang, Kan
2023-10-30 17:46 ` Wang, Weilin
2023-10-30 18:30 ` Liang, Kan
2023-10-14 1:51 ` [RFC PATCH v2 08/17] perf stat: Add functions to get counter info weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 09/17] perf stat: Add functions to create new group and assign events into groups for hardware-grouping method weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 10/17] perf stat: Add build string function and topdown events handling in hardware-grouping weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 11/17] perf stat: Add function to handle special events " weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 12/17] perf stat: Add function to combine metrics for hardware-grouping weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 13/17] perf stat: Handle taken alone in hardware-grouping weilin.wang
2023-10-14 1:51 ` [RFC PATCH v2 14/17] perf stat: Handle NMI " weilin.wang
2023-10-14 1:52 ` [RFC PATCH v2 15/17] perf stat: Code refactoring " weilin.wang
2023-10-14 1:52 ` [RFC PATCH v2 16/17] perf stat: Add tool events support " weilin.wang
2023-10-14 1:52 ` [RFC PATCH v2 17/17] perf pmu-events: Add event counter data for Tigerlake weilin.wang
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=20231014015202.1175377-1-weilin.wang@intel.com \
--to=weilin.wang@intel.com \
--cc=acme@kernel.org \
--cc=adrian.hunter@intel.com \
--cc=alexander.shishkin@linux.intel.com \
--cc=caleb.biggers@intel.com \
--cc=irogers@google.com \
--cc=jolsa@kernel.org \
--cc=kan.liang@linux.intel.com \
--cc=linux-kernel@vger.kernel.org \
--cc=linux-perf-users@vger.kernel.org \
--cc=mark.rutland@arm.com \
--cc=mingo@redhat.com \
--cc=namhyung@kernel.org \
--cc=perry.taylor@intel.com \
--cc=peterz@infradead.org \
--cc=samantha.alt@intel.com \
--cc=yangjihong1@huawei.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is an external index of several public inboxes,
see mirroring instructions on how to clone and mirror
all data and code used by this external index.