From: Matthew Brost <matthew.brost@intel.com>
To: Francois Dugast <francois.dugast@intel.com>
Cc: intel-xe@lists.freedesktop.org,
"Thomas Hellström" <thomas.hellstrom@linux.intel.com>,
"Himal Prasad Ghimiray" <himal.prasad.ghimiray@intel.com>,
Copilot <223556219+Copilot@users.noreply.github.com>
Subject: Re: [PATCH v8 03/12] drm/xe: Thread prefetch of SVM ranges
Date: Mon, 27 Jul 2026 12:11:51 -0700 [thread overview]
Message-ID: <ametd1SyOn1/EtJC@gsse-cloud1.jf.intel.com> (raw)
In-Reply-To: <ameMGby_tJBIi0Lv@fdugast-desk>
On Mon, Jul 27, 2026 at 06:49:29PM +0200, Francois Dugast wrote:
> On Fri, Jul 24, 2026 at 04:25:52PM -0700, Matthew Brost wrote:
> > The migrate_vma_* functions are very CPU-intensive; as a result,
> > prefetching SVM ranges is limited by CPU performance rather than paging
> > copy engine bandwidth. To accelerate SVM range prefetching, the step
> > that calls migrate_vma_* is now threaded. A dedicated prefetch
> > workqueue is used for threading so prefetch work is never mixed with
> > page fault or garbage collector work.
> >
> > Running xe_exec_system_allocator --r prefetch-benchmark, which tests
> > 64MB prefetches, shows an increase from ~4.35 GB/s to 12.25 GB/s with
> > this patch on drm-tip. Enabling high SLPC further increases throughput
> > to ~15.25 GB/s, and combining SLPC with ULLS raises it to ~16 GB/s. Both
> > of these optimizations are upcoming.
> >
> > Since the dedicated prefetch workqueue is not shared with page fault
> > or SVM garbage collector work, page fault servicing and garbage
> > collection can keep using a plain down_read() on vm->lock: there is no
> > risk of a blocked reader starving a worker that a pending writer is
> > waiting to flush, because that flushing is now confined to the
> > separate prefetch workqueue.
> >
> > v2:
> > - Use dedicated prefetch workqueue
> > - Pick dedicated prefetch thread count based on profiling
> > - Skip threaded prefetch for only 1 range or if prefetching to SRAM
> > - Fully tested
> > v3:
> > - Use page fault work queue
> > v4:
> > - Go back to a dedicated prefetch workqueue (usm.prefetch_wq) rather
> > than reusing the page fault workqueue (usm.pagefault_wq), so
> > threaded prefetches and page fault / garbage collector work no
> > longer contend for the same workqueue. This removes the need for
> > down_read_trylock() based lock avoidance in the page fault and
> > garbage collector paths.
> >
> > Cc: Thomas Hellström <thomas.hellstrom@linux.intel.com>
> > Cc: Himal Prasad Ghimiray <himal.prasad.ghimiray@intel.com>
> > Signed-off-by: Matthew Brost <matthew.brost@intel.com>
> > Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>
>
> I believe we should go with "Assisted-by" instead of "Co-authored-by".
>
> This also applies to patch #8 in the series.
>
Yes, that was an oversight. I fixed somethings on the tip of the series
and had copilot squash the fixes into the correct patches and this got
added. Will adjust.
Matt
> Francois
>
> > ---
> > drivers/gpu/drm/xe/xe_device_types.h | 6 +-
> > drivers/gpu/drm/xe/xe_pagefault.c | 29 ++++--
> > drivers/gpu/drm/xe/xe_svm.c | 8 +-
> > drivers/gpu/drm/xe/xe_svm.h | 6 +-
> > drivers/gpu/drm/xe/xe_vm.c | 147 ++++++++++++++++++++-------
> > drivers/gpu/drm/xe/xe_vm_types.h | 15 +--
> > 6 files changed, 154 insertions(+), 57 deletions(-)
> >
> > diff --git a/drivers/gpu/drm/xe/xe_device_types.h b/drivers/gpu/drm/xe/xe_device_types.h
> > index b8d1726c0513..159ee16ee6d5 100644
> > --- a/drivers/gpu/drm/xe/xe_device_types.h
> > +++ b/drivers/gpu/drm/xe/xe_device_types.h
> > @@ -307,8 +307,10 @@ struct xe_device {
> > u32 current_pf_queue;
> > /** @usm.lock: protects UM state */
> > struct rw_semaphore lock;
> > - /** @usm.pf_wq: page fault work queue, unbound, high priority */
> > - struct workqueue_struct *pf_wq;
> > + /** @usm.pagefault_wq: page fault work queue, unbound, high priority */
> > + struct workqueue_struct *pagefault_wq;
> > + /** @usm.prefetch_wq: threaded prefetch work queue, unbound */
> > + struct workqueue_struct *prefetch_wq;
> > /*
> > * We pick 4 here because, in the current implementation, it
> > * yields the best bandwidth utilization of the kernel paging
> > diff --git a/drivers/gpu/drm/xe/xe_pagefault.c b/drivers/gpu/drm/xe/xe_pagefault.c
> > index 80196e874e06..fd7ef0718153 100644
> > --- a/drivers/gpu/drm/xe/xe_pagefault.c
> > +++ b/drivers/gpu/drm/xe/xe_pagefault.c
> > @@ -303,8 +303,8 @@ static void xe_pagefault_queue_work(struct work_struct *w)
> >
> > err = xe_pagefault_service(&pf);
> > if (err) {
> > - xe_pagefault_save_to_vm(gt_to_xe(pf.gt), &pf);
> > if (!(pf.consumer.access_type & XE_PAGEFAULT_ACCESS_PREFETCH)) {
> > + xe_pagefault_save_to_vm(gt_to_xe(pf.gt), &pf);
> > xe_pagefault_print(&pf);
> > xe_gt_info(pf.gt, "Fault response: Unsuccessful %pe\n",
> > ERR_PTR(err));
> > @@ -318,7 +318,7 @@ static void xe_pagefault_queue_work(struct work_struct *w)
> > pf.producer.ops->ack_fault(&pf, err);
> >
> > if (time_after(jiffies, threshold)) {
> > - queue_work(gt_to_xe(pf.gt)->usm.pf_wq, w);
> > + queue_work(gt_to_xe(pf.gt)->usm.pagefault_wq, w);
> > break;
> > }
> > }
> > @@ -376,7 +376,8 @@ static void xe_pagefault_fini(void *arg)
> > {
> > struct xe_device *xe = arg;
> >
> > - destroy_workqueue(xe->usm.pf_wq);
> > + destroy_workqueue(xe->usm.prefetch_wq);
> > + destroy_workqueue(xe->usm.pagefault_wq);
> > }
> >
> > /**
> > @@ -394,12 +395,20 @@ int xe_pagefault_init(struct xe_device *xe)
> > if (!xe->info.has_usm)
> > return 0;
> >
> > - xe->usm.pf_wq = alloc_workqueue("xe_page_fault_work_queue",
> > - WQ_UNBOUND | WQ_HIGHPRI,
> > - XE_PAGEFAULT_QUEUE_COUNT);
> > - if (!xe->usm.pf_wq)
> > + xe->usm.pagefault_wq = alloc_workqueue("xe_page_fault_work_queue",
> > + WQ_UNBOUND | WQ_HIGHPRI,
> > + XE_PAGEFAULT_QUEUE_COUNT);
> > + if (!xe->usm.pagefault_wq)
> > return -ENOMEM;
> >
> > + xe->usm.prefetch_wq = alloc_workqueue("xe_prefetch_work_queue",
> > + WQ_UNBOUND,
> > + XE_PAGEFAULT_QUEUE_COUNT);
> > + if (!xe->usm.prefetch_wq) {
> > + err = -ENOMEM;
> > + goto err_pagefault_wq;
> > + }
> > +
> > for (i = 0; i < XE_PAGEFAULT_QUEUE_COUNT; ++i) {
> > err = xe_pagefault_queue_init(xe, xe->usm.pf_queue + i);
> > if (err)
> > @@ -409,7 +418,9 @@ int xe_pagefault_init(struct xe_device *xe)
> > return devm_add_action_or_reset(xe->drm.dev, xe_pagefault_fini, xe);
> >
> > err_out:
> > - destroy_workqueue(xe->usm.pf_wq);
> > + destroy_workqueue(xe->usm.prefetch_wq);
> > +err_pagefault_wq:
> > + destroy_workqueue(xe->usm.pagefault_wq);
> > return err;
> > }
> >
> > @@ -495,7 +506,7 @@ int xe_pagefault_handler(struct xe_device *xe, struct xe_pagefault *pf)
> > memcpy(pf_queue->data + pf_queue->head, pf, sizeof(*pf));
> > pf_queue->head = (pf_queue->head + xe_pagefault_entry_size()) %
> > pf_queue->size;
> > - queue_work(xe->usm.pf_wq, &pf_queue->worker);
> > + queue_work(xe->usm.pagefault_wq, &pf_queue->worker);
> > } else {
> > drm_warn(&xe->drm,
> > "PageFault Queue (%d) full, shouldn't be possible\n",
> > diff --git a/drivers/gpu/drm/xe/xe_svm.c b/drivers/gpu/drm/xe/xe_svm.c
> > index cc36addb4f4f..6a470a02fee7 100644
> > --- a/drivers/gpu/drm/xe/xe_svm.c
> > +++ b/drivers/gpu/drm/xe/xe_svm.c
> > @@ -148,7 +148,7 @@ xe_svm_garbage_collector_add_range(struct xe_vm *vm, struct xe_svm_range *range,
> > &vm->svm.garbage_collector.range_list);
> > spin_unlock(&vm->svm.garbage_collector.list_lock);
> >
> > - queue_work(xe->usm.pf_wq, &vm->svm.garbage_collector.work);
> > + queue_work(xe->usm.pagefault_wq, &vm->svm.garbage_collector.work);
> > }
> >
> > static void xe_svm_tlb_inval_count_stats_incr(struct xe_gt *gt)
> > @@ -1051,6 +1051,7 @@ void xe_svm_range_migrate_to_smem(struct xe_vm *vm, struct xe_svm_range *range)
> > * @tile_mask: Mask representing the tiles to be checked
> > * @dpagemap: if !%NULL, the range is expected to be present
> > * in device memory identified by this parameter.
> > + * @valid_pages: Pages are valid, result written back to caller
> > *
> > * The xe_svm_range_validate() function checks if a range is
> > * valid and located in the desired memory region.
> > @@ -1059,7 +1060,8 @@ void xe_svm_range_migrate_to_smem(struct xe_vm *vm, struct xe_svm_range *range)
> > */
> > bool xe_svm_range_validate(struct xe_vm *vm,
> > struct xe_svm_range *range,
> > - u8 tile_mask, const struct drm_pagemap *dpagemap)
> > + u8 tile_mask, const struct drm_pagemap *dpagemap,
> > + bool *valid_pages)
> > {
> > bool ret;
> >
> > @@ -1071,6 +1073,8 @@ bool xe_svm_range_validate(struct xe_vm *vm,
> > else
> > ret = ret && !range->pages.dpagemap;
> >
> > + *valid_pages = xe_svm_range_pages_valid(range);
> > +
> > xe_svm_notifier_unlock(vm);
> >
> > return ret;
> > diff --git a/drivers/gpu/drm/xe/xe_svm.h b/drivers/gpu/drm/xe/xe_svm.h
> > index 0d1f1107af5f..46be2e5c6f7f 100644
> > --- a/drivers/gpu/drm/xe/xe_svm.h
> > +++ b/drivers/gpu/drm/xe/xe_svm.h
> > @@ -134,7 +134,8 @@ void xe_svm_range_migrate_to_smem(struct xe_vm *vm, struct xe_svm_range *range);
> >
> > bool xe_svm_range_validate(struct xe_vm *vm,
> > struct xe_svm_range *range,
> > - u8 tile_mask, const struct drm_pagemap *dpagemap);
> > + u8 tile_mask, const struct drm_pagemap *dpagemap,
> > + bool *valid_pages);
> >
> > u64 xe_svm_find_vma_start(struct xe_vm *vm, u64 addr, u64 end, struct xe_vma *vma);
> >
> > @@ -376,7 +377,8 @@ void xe_svm_range_migrate_to_smem(struct xe_vm *vm, struct xe_svm_range *range)
> > static inline
> > bool xe_svm_range_validate(struct xe_vm *vm,
> > struct xe_svm_range *range,
> > - u8 tile_mask, bool devmem_preferred)
> > + u8 tile_mask, const struct drm_pagemap *dpagemap,
> > + bool *valid_pages)
> > {
> > return false;
> > }
> > diff --git a/drivers/gpu/drm/xe/xe_vm.c b/drivers/gpu/drm/xe/xe_vm.c
> > index d7e6644df4bc..7eed38b78e5f 100644
> > --- a/drivers/gpu/drm/xe/xe_vm.c
> > +++ b/drivers/gpu/drm/xe/xe_vm.c
> > @@ -2525,6 +2525,7 @@ vm_bind_ioctl_ops_create(struct xe_vm *vm, struct xe_vma_ops *vops,
> > struct drm_pagemap *dpagemap = NULL;
> > u8 id, tile_mask = 0;
> > u32 i;
> > + bool valid_pages;
> >
> > if (xe_vma_is_userptr(vma))
> > vops->flags |= XE_VMA_OPS_FLAG_MODIFIES_GPUVA;
> > @@ -2569,9 +2570,11 @@ vm_bind_ioctl_ops_create(struct xe_vm *vm, struct xe_vma_ops *vops,
> > goto unwind_prefetch_ops;
> > }
> >
> > - if (xe_svm_range_validate(vm, svm_range, tile_mask, dpagemap)) {
> > + if (xe_svm_range_validate(vm, svm_range, tile_mask,
> > + dpagemap, &valid_pages)) {
> > xe_svm_range_debug(svm_range, "PREFETCH - RANGE IS VALID");
> > xe_svm_range_put(svm_range);
> > + xe_assert(vm->xe, valid_pages);
> > goto check_next_range;
> > }
> >
> > @@ -2586,6 +2589,8 @@ vm_bind_ioctl_ops_create(struct xe_vm *vm, struct xe_vma_ops *vops,
> >
> > op->prefetch_range.ranges_count++;
> > vops->flags |= XE_VMA_OPS_FLAG_HAS_SVM_PREFETCH;
> > + if (valid_pages)
> > + vops->flags |= XE_VMA_OPS_FLAG_HAS_SVM_VALID_RANGE;
> > xe_svm_range_debug(svm_range, "PREFETCH - RANGE CREATED");
> > check_next_range:
> > if (range_end > xe_svm_range_end(svm_range) &&
> > @@ -3151,16 +3156,80 @@ static int check_ufence(struct xe_vma *vma)
> > return 0;
> > }
> >
> > -static int prefetch_ranges(struct xe_vm *vm, struct xe_vma_op *op)
> > +struct prefetch_thread {
> > + struct work_struct work;
> > + struct drm_gpusvm_ctx *ctx;
> > + struct xe_vma *vma;
> > + struct xe_svm_range *svm_range;
> > + struct drm_pagemap *dpagemap;
> > + int err;
> > +};
> > +
> > +static void prefetch_thread_func(struct prefetch_thread *thread)
> > +{
> > + struct xe_vma *vma = thread->vma;
> > + struct xe_vm *vm = xe_vma_vm(vma);
> > + struct xe_svm_range *svm_range = thread->svm_range;
> > + struct drm_pagemap *dpagemap = thread->dpagemap;
> > + int err = 0;
> > +
> > + guard(mutex)(&svm_range->lock);
> > +
> > + if (xe_svm_range_is_removed(svm_range))
> > + return;
> > +
> > + if (!dpagemap)
> > + xe_svm_range_migrate_to_smem(vm, svm_range);
> > +
> > + if (IS_ENABLED(CONFIG_DRM_XE_DEBUG_VM)) {
> > + drm_dbg(&vm->xe->drm,
> > + "Prefetch pagemap is %s start 0x%016lx end 0x%016lx\n",
> > + dpagemap ? dpagemap->drm->unique : "system",
> > + xe_svm_range_start(svm_range), xe_svm_range_end(svm_range));
> > + }
> > +
> > + if (xe_svm_range_needs_migrate_to_vram(svm_range, vma, dpagemap)) {
> > + err = xe_svm_alloc_vram(svm_range, thread->ctx, dpagemap);
> > + if (err) {
> > + drm_dbg(&vm->xe->drm, "VRAM allocation failed, retry from userspace, asid=%u, gpusvm=%p, errno=%pe\n",
> > + vm->usm.asid, &vm->svm.gpusvm, ERR_PTR(err));
> > + return;
> > + }
> > + xe_svm_range_debug(svm_range, "PREFETCH - RANGE MIGRATED TO VRAM");
> > + }
> > +
> > + err = xe_svm_range_get_pages(vm, svm_range, thread->ctx);
> > + if (err) {
> > + drm_dbg(&vm->xe->drm, "Get pages failed, asid=%u, gpusvm=%p, errno=%pe\n",
> > + vm->usm.asid, &vm->svm.gpusvm, ERR_PTR(err));
> > + if (err == -EOPNOTSUPP || err == -EFAULT || err == -EPERM)
> > + err = 0;
> > + thread->err = err;
> > + return;
> > + }
> > + xe_svm_range_debug(svm_range, "PREFETCH - RANGE GET PAGES DONE");
> > +}
> > +
> > +static void prefetch_work_func(struct work_struct *w)
> > +{
> > + struct prefetch_thread *thread =
> > + container_of(w, struct prefetch_thread, work);
> > +
> > + prefetch_thread_func(thread);
> > +}
> > +
> > +static int prefetch_ranges(struct xe_vm *vm, struct xe_vma_ops *vops,
> > + struct xe_vma_op *op)
> > {
> > bool devmem_possible = IS_DGFX(vm->xe) && IS_ENABLED(CONFIG_DRM_XE_PAGEMAP);
> > struct xe_vma *vma = gpuva_to_vma(op->base.prefetch.va);
> > struct drm_pagemap *dpagemap = op->prefetch_range.dpagemap;
> > - int err = 0;
> > -
> > struct xe_svm_range *svm_range;
> > struct drm_gpusvm_ctx ctx = {};
> > + struct prefetch_thread stack_thread, *thread, *prefetches;
> > unsigned long i;
> > + int err = 0, idx = 0;
> > + bool skip_threads;
> >
> > if (!xe_vma_is_cpu_addr_mirror(vma))
> > return 0;
> > @@ -3170,42 +3239,49 @@ static int prefetch_ranges(struct xe_vm *vm, struct xe_vma_op *op)
> > ctx.check_pages_threshold = devmem_possible ? SZ_64K : 0;
> > ctx.device_private_page_owner = xe_svm_private_page_owner(vm, !dpagemap);
> >
> > - /* TODO: Threading the migration */
> > - xa_for_each(&op->prefetch_range.range, i, svm_range) {
> > - guard(mutex)(&svm_range->lock);
> > -
> > - if (xe_svm_range_is_removed(svm_range))
> > - continue;
> > + skip_threads = op->prefetch_range.ranges_count == 1 ||
> > + (!dpagemap && !(vops->flags &
> > + XE_VMA_OPS_FLAG_HAS_SVM_VALID_RANGE)) ||
> > + !(vops->flags & XE_VMA_OPS_FLAG_DOWNGRADE_LOCK);
> > + thread = skip_threads ? &stack_thread : NULL;
> >
> > - if (!dpagemap)
> > - xe_svm_range_migrate_to_smem(vm, svm_range);
> > + if (!skip_threads) {
> > + prefetches = kvmalloc_array(op->prefetch_range.ranges_count,
> > + sizeof(*prefetches), GFP_KERNEL);
> > + if (!prefetches)
> > + return -ENOMEM;
> > + }
> >
> > - if (IS_ENABLED(CONFIG_DRM_XE_DEBUG_VM)) {
> > - drm_dbg(&vm->xe->drm,
> > - "Prefetch pagemap is %s start 0x%016lx end 0x%016lx\n",
> > - dpagemap ? dpagemap->drm->unique : "system",
> > - xe_svm_range_start(svm_range), xe_svm_range_end(svm_range));
> > + xa_for_each(&op->prefetch_range.range, i, svm_range) {
> > + if (!skip_threads) {
> > + thread = prefetches + idx++;
> > + INIT_WORK(&thread->work, prefetch_work_func);
> > }
> >
> > - if (xe_svm_range_needs_migrate_to_vram(svm_range, vma, dpagemap)) {
> > - err = xe_svm_alloc_vram(svm_range, &ctx, dpagemap);
> > - if (err) {
> > - drm_dbg(&vm->xe->drm, "VRAM allocation failed, retry from userspace, asid=%u, gpusvm=%p, errno=%pe\n",
> > - vm->usm.asid, &vm->svm.gpusvm, ERR_PTR(err));
> > - return -ENODATA;
> > - }
> > - xe_svm_range_debug(svm_range, "PREFETCH - RANGE MIGRATED TO VRAM");
> > + thread->ctx = &ctx;
> > + thread->vma = vma;
> > + thread->svm_range = svm_range;
> > + thread->dpagemap = dpagemap;
> > + thread->err = 0;
> > +
> > + if (skip_threads) {
> > + prefetch_thread_func(thread);
> > + if (thread->err)
> > + return thread->err;
> > + } else {
> > + queue_work(vm->xe->usm.prefetch_wq, &thread->work);
> > }
> > + }
> >
> > - err = xe_svm_range_get_pages(vm, svm_range, &ctx);
> > - if (err) {
> > - drm_dbg(&vm->xe->drm, "Get pages failed, asid=%u, gpusvm=%p, errno=%pe\n",
> > - vm->usm.asid, &vm->svm.gpusvm, ERR_PTR(err));
> > - if (err == -EOPNOTSUPP || err == -EFAULT || err == -EPERM)
> > - err = -ENODATA;
> > - return err;
> > + if (!skip_threads) {
> > + for (i = 0; i < idx; ++i) {
> > + thread = prefetches + i;
> > +
> > + flush_work(&thread->work);
> > + if (thread->err && !err)
> > + err = thread->err;
> > }
> > - xe_svm_range_debug(svm_range, "PREFETCH - RANGE GET PAGES DONE");
> > + kvfree(prefetches);
> > }
> >
> > return err;
> > @@ -3336,7 +3412,8 @@ static int op_lock_and_prep(struct drm_exec *exec, struct xe_vm *vm,
> > return err;
> > }
> >
> > -static int vm_bind_ioctl_ops_prefetch_ranges(struct xe_vm *vm, struct xe_vma_ops *vops)
> > +static int vm_bind_ioctl_ops_prefetch_ranges(struct xe_vm *vm,
> > + struct xe_vma_ops *vops)
> > {
> > struct xe_vma_op *op;
> > int err;
> > @@ -3346,7 +3423,7 @@ static int vm_bind_ioctl_ops_prefetch_ranges(struct xe_vm *vm, struct xe_vma_ops
> >
> > list_for_each_entry(op, &vops->list, link) {
> > if (op->base.op == DRM_GPUVA_OP_PREFETCH) {
> > - err = prefetch_ranges(vm, op);
> > + err = prefetch_ranges(vm, vops, op);
> > if (err)
> > return err;
> > }
> > diff --git a/drivers/gpu/drm/xe/xe_vm_types.h b/drivers/gpu/drm/xe/xe_vm_types.h
> > index 2f5f74fed9d2..68588b624212 100644
> > --- a/drivers/gpu/drm/xe/xe_vm_types.h
> > +++ b/drivers/gpu/drm/xe/xe_vm_types.h
> > @@ -556,13 +556,14 @@ struct xe_vma_ops {
> > /** @pt_update_ops: page table update operations */
> > struct xe_vm_pgtable_update_ops pt_update_ops[XE_MAX_TILES_PER_DEVICE];
> > /** @flag: signify the properties within xe_vma_ops*/
> > -#define XE_VMA_OPS_FLAG_HAS_SVM_PREFETCH BIT(0)
> > -#define XE_VMA_OPS_FLAG_MADVISE BIT(1)
> > -#define XE_VMA_OPS_ARRAY_OF_BINDS BIT(2)
> > -#define XE_VMA_OPS_FLAG_SKIP_TLB_WAIT BIT(3)
> > -#define XE_VMA_OPS_FLAG_ALLOW_SVM_UNMAP BIT(4)
> > -#define XE_VMA_OPS_FLAG_MODIFIES_GPUVA BIT(5)
> > -#define XE_VMA_OPS_FLAG_DOWNGRADE_LOCK BIT(6)
> > +#define XE_VMA_OPS_FLAG_HAS_SVM_PREFETCH BIT(0)
> > +#define XE_VMA_OPS_FLAG_MADVISE BIT(1)
> > +#define XE_VMA_OPS_ARRAY_OF_BINDS BIT(2)
> > +#define XE_VMA_OPS_FLAG_SKIP_TLB_WAIT BIT(3)
> > +#define XE_VMA_OPS_FLAG_ALLOW_SVM_UNMAP BIT(4)
> > +#define XE_VMA_OPS_FLAG_MODIFIES_GPUVA BIT(5)
> > +#define XE_VMA_OPS_FLAG_DOWNGRADE_LOCK BIT(6)
> > +#define XE_VMA_OPS_FLAG_HAS_SVM_VALID_RANGE BIT(7)
> > u32 flags;
> > #ifdef TEST_VM_OPS_ERROR
> > /** @inject_error: inject error to test error handling */
> > --
> > 2.34.1
> >
next prev parent reply other threads:[~2026-07-27 19:11 UTC|newest]
Thread overview: 24+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-07-24 23:25 [PATCH v8 00/12] Fine grained fault locking, threaded prefetch, storm cache Matthew Brost
2026-07-24 23:25 ` [PATCH v8 01/12] drm/xe: Fine grained page fault locking Matthew Brost
2026-07-24 23:25 ` [PATCH v8 02/12] drm/xe: Allow prefetch-only VM bind IOCTLs to use VM read lock Matthew Brost
2026-07-27 14:33 ` Francois Dugast
2026-07-24 23:25 ` [PATCH v8 03/12] drm/xe: Thread prefetch of SVM ranges Matthew Brost
2026-07-27 16:41 ` Francois Dugast
2026-07-27 19:21 ` Matthew Brost
2026-07-27 19:46 ` Matthew Brost
2026-07-27 16:49 ` Francois Dugast
2026-07-27 19:11 ` Matthew Brost [this message]
2026-07-24 23:25 ` [PATCH v8 04/12] drm/xe: Use a single page-fault queue with multiple workers Matthew Brost
2026-07-24 23:25 ` [PATCH v8 05/12] drm/xe: Add num_pf_work modparam Matthew Brost
2026-07-24 23:25 ` [PATCH v8 06/12] drm/xe: Engine class and instance into a u8 Matthew Brost
2026-07-24 23:25 ` [PATCH v8 07/12] drm/xe: Track pagefault worker runtime Matthew Brost
2026-07-24 23:25 ` [PATCH v8 08/12] drm/xe: Chain page faults via queue-resident cache to avoid fault storms Matthew Brost
2026-07-29 12:31 ` Francois Dugast
2026-07-29 18:20 ` Matthew Brost
2026-07-24 23:25 ` [PATCH v8 09/12] drm/xe: Add pagefault chaining stats Matthew Brost
2026-07-24 23:25 ` [PATCH v8 10/12] drm/xe: Add debugfs pagefault_info Matthew Brost
2026-07-24 23:26 ` [PATCH v8 11/12] drm/xe: batch CT pagefault acks with periodic flush Matthew Brost
2026-07-24 23:26 ` [PATCH v8 12/12] drm/xe: Track parallel page fault activity in GT stats Matthew Brost
2026-07-24 23:32 ` ✗ CI.checkpatch: warning for Fine grained fault locking, threaded prefetch, storm cache (rev8) Patchwork
2026-07-24 23:33 ` ✓ CI.KUnit: success " Patchwork
2026-07-25 0:17 ` ✓ Xe.CI.BAT: " Patchwork
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=ametd1SyOn1/EtJC@gsse-cloud1.jf.intel.com \
--to=matthew.brost@intel.com \
--cc=223556219+Copilot@users.noreply.github.com \
--cc=francois.dugast@intel.com \
--cc=himal.prasad.ghimiray@intel.com \
--cc=intel-xe@lists.freedesktop.org \
--cc=thomas.hellstrom@linux.intel.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox