All of lore.kernel.org
 help / color / mirror / Atom feed
From: "Dr. David Alan Gilbert" <dgilbert@redhat.com>
To: Richard Henderson <rth@twiddle.net>
Cc: peter.maydell@linaro.org, vijay.kilari@gmail.com,
	liang.z.li@intel.com, qemu-devel@nongnu.org, qemu-arm@nongnu.org,
	pbonzini@redhat.com
Subject: Re: [Qemu-arm] [Qemu-devel] [PATCH 0/7] Improve buffer_is_zero
Date: Wed, 24 Aug 2016 09:34:57 +0100	[thread overview]
Message-ID: <20160824083457.GA2032@work-vm> (raw)
In-Reply-To: <1472012279-20581-1-git-send-email-rth@twiddle.net>


cc'ing in Liang Li who did the original avx2 code.

Dave


* Richard Henderson (rth@twiddle.net) wrote:
> Patches 1-3 remove the use of ifunc from the implementation.
> 
> Patch 5 adjusts the x86 implementation a bit more to take
> advantage of ptest (in sse4.1) and unaligned accesses (in avx1).
> 
> Patches 2 and 6 are the result of my conversation with Vijaya
> Kumar with respect to ThunderX.
> 
> Patch 7 is the result of seeing some really really horrible code
> produced for ppc64le (gcc 4.9 and mainline).
> 
> This has had limited testing.  What I don't know is the best way
> to benchmark this -- the only way I know to trigger this is via
> the console, by hand, which doesn't make for reasonable timing.
> 
> 
> r~
> 
> 
> Richard Henderson (7):
>   cutils: Remove SPLAT macro
>   cutils: Export only buffer_is_zero
>   cutils: Rearrange buffer_is_zero acceleration
>   cutils: Add generic prefetch
>   cutils: Rewrite x86 buffer zero checking
>   cutils: Rewrite aarch64 buffer zero checking
>   cutils: Rewrite ppc buffer zero checking
> 
>  configure             |  21 +-
>  include/qemu/cutils.h |   2 -
>  migration/ram.c       |   2 +-
>  migration/rdma.c      |   5 +-
>  util/cutils.c         | 526 +++++++++++++++++++++++++++++++++-----------------
>  5 files changed, 352 insertions(+), 204 deletions(-)
> 
> -- 
> 2.7.4
> 
> 
--
Dr. David Alan Gilbert / dgilbert@redhat.com / Manchester, UK

WARNING: multiple messages have this Message-ID (diff)
From: "Dr. David Alan Gilbert" <dgilbert@redhat.com>
To: Richard Henderson <rth@twiddle.net>
Cc: qemu-devel@nongnu.org, pbonzini@redhat.com, qemu-arm@nongnu.org,
	vijay.kilari@gmail.com, peter.maydell@linaro.org,
	liang.z.li@intel.com
Subject: Re: [Qemu-devel] [PATCH 0/7] Improve buffer_is_zero
Date: Wed, 24 Aug 2016 09:34:57 +0100	[thread overview]
Message-ID: <20160824083457.GA2032@work-vm> (raw)
In-Reply-To: <1472012279-20581-1-git-send-email-rth@twiddle.net>


cc'ing in Liang Li who did the original avx2 code.

Dave


* Richard Henderson (rth@twiddle.net) wrote:
> Patches 1-3 remove the use of ifunc from the implementation.
> 
> Patch 5 adjusts the x86 implementation a bit more to take
> advantage of ptest (in sse4.1) and unaligned accesses (in avx1).
> 
> Patches 2 and 6 are the result of my conversation with Vijaya
> Kumar with respect to ThunderX.
> 
> Patch 7 is the result of seeing some really really horrible code
> produced for ppc64le (gcc 4.9 and mainline).
> 
> This has had limited testing.  What I don't know is the best way
> to benchmark this -- the only way I know to trigger this is via
> the console, by hand, which doesn't make for reasonable timing.
> 
> 
> r~
> 
> 
> Richard Henderson (7):
>   cutils: Remove SPLAT macro
>   cutils: Export only buffer_is_zero
>   cutils: Rearrange buffer_is_zero acceleration
>   cutils: Add generic prefetch
>   cutils: Rewrite x86 buffer zero checking
>   cutils: Rewrite aarch64 buffer zero checking
>   cutils: Rewrite ppc buffer zero checking
> 
>  configure             |  21 +-
>  include/qemu/cutils.h |   2 -
>  migration/ram.c       |   2 +-
>  migration/rdma.c      |   5 +-
>  util/cutils.c         | 526 +++++++++++++++++++++++++++++++++-----------------
>  5 files changed, 352 insertions(+), 204 deletions(-)
> 
> -- 
> 2.7.4
> 
> 
--
Dr. David Alan Gilbert / dgilbert@redhat.com / Manchester, UK

  parent reply	other threads:[~2016-08-24  8:35 UTC|newest]

Thread overview: 40+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2016-08-24  4:17 [Qemu-arm] [PATCH 0/7] Improve buffer_is_zero Richard Henderson
2016-08-24  4:17 ` [Qemu-devel] " Richard Henderson
2016-08-24  4:17 ` [Qemu-devel] [PATCH 1/7] cutils: Remove SPLAT macro Richard Henderson
2016-08-24  4:17   ` Richard Henderson
2016-08-24  4:17 ` [Qemu-devel] [PATCH 2/7] cutils: Export only buffer_is_zero Richard Henderson
2016-08-24  4:17   ` Richard Henderson
2016-08-24  8:37   ` [Qemu-arm] " Dr. David Alan Gilbert
2016-08-24  8:37     ` Dr. David Alan Gilbert
2016-08-24  4:17 ` [Qemu-devel] [PATCH 3/7] cutils: Rearrange buffer_is_zero acceleration Richard Henderson
2016-08-24  4:17   ` Richard Henderson
2016-08-24  4:17 ` [Qemu-devel] [PATCH 4/7] cutils: Add generic prefetch Richard Henderson
2016-08-24  4:17   ` Richard Henderson
2016-08-24  4:17 ` [Qemu-arm] [PATCH 5/7] cutils: Rewrite x86 buffer zero checking Richard Henderson
2016-08-24  4:17   ` [Qemu-devel] " Richard Henderson
2016-08-24  4:17 ` [Qemu-arm] [PATCH 6/7] cutils: Rewrite aarch64 " Richard Henderson
2016-08-24  4:17   ` [Qemu-devel] " Richard Henderson
2016-08-24  4:17 ` [Qemu-arm] [PATCH 7/7] cutils: Rewrite ppc " Richard Henderson
2016-08-24  4:17   ` [Qemu-devel] " Richard Henderson
2016-08-24  4:30 ` [Qemu-arm] [Qemu-devel] [PATCH 0/7] Improve buffer_is_zero no-reply
2016-08-24  4:30   ` no-reply
2016-08-24  4:38   ` [Qemu-arm] " Paolo Bonzini
2016-08-24  4:38     ` [Qemu-devel] " Paolo Bonzini
2016-08-24 14:53     ` [Qemu-arm] " Richard Henderson
2016-08-24 14:53       ` Richard Henderson
2016-08-24 14:59       ` [Qemu-arm] " Paolo Bonzini
2016-08-24 14:59         ` Paolo Bonzini
2016-08-24  8:34 ` Dr. David Alan Gilbert [this message]
2016-08-24  8:34   ` Dr. David Alan Gilbert
2016-08-24 10:26   ` Adam Richter
2016-08-24 10:26     ` Adam Richter
2016-08-24 10:52     ` [Qemu-arm] " Peter Maydell
2016-08-24 10:52       ` Peter Maydell
2016-08-24 11:45       ` [Qemu-arm] " Paolo Bonzini
2016-08-24 11:45         ` Paolo Bonzini
2016-08-24 12:22         ` [Qemu-arm] " Peter Maydell
2016-08-24 12:22           ` Peter Maydell
2016-08-25  6:37 ` [Qemu-arm] " Vijay Kilari
2016-08-25  6:37   ` [Qemu-devel] " Vijay Kilari
2016-08-25  8:04   ` [Qemu-arm] " Vijay Kilari
2016-08-25  8:04     ` [Qemu-devel] " Vijay Kilari

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=20160824083457.GA2032@work-vm \
    --to=dgilbert@redhat.com \
    --cc=liang.z.li@intel.com \
    --cc=pbonzini@redhat.com \
    --cc=peter.maydell@linaro.org \
    --cc=qemu-arm@nongnu.org \
    --cc=qemu-devel@nongnu.org \
    --cc=rth@twiddle.net \
    --cc=vijay.kilari@gmail.com \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is an external index of several public inboxes,
see mirroring instructions on how to clone and mirror
all data and code used by this external index.