Commits · 71f4e45a4ed3807aaed0d1ab3ef472a121753546 · Kirill Smelkov / linux

20 Feb, 2019 1 commit

Merge branch 'linux-5.1' of git://github.com/skeggsb/linux into drm-next · 71f4e45a

Dave Airlie authored Feb 20, 2019

Various fixes/cleanups, along with initial support for SVM features
utilising HMM address-space mirroring and device memory migration.
There's a lot more work to do in these areas, both in terms of
features and efficiency, but these can slowly trickle in later down
the track.
Signed-off-by: Dave Airlie <airlied@redhat.com>
From: Ben Skeggs <skeggsb@gmail.com>
Link: https://patchwork.freedesktop.org/patch/msgid/CACAvsv5bsB4rRY1Gqa_Bp_KAd-v_q1rGZ4nYmOAQhceL0Nr-Xg@mail.gmail.com

71f4e45a

19 Feb, 2019 39 commits

drm/nouveau/dmem: use dma addresses during migration copies · a788ade4
Ben Skeggs authored Feb 15, 2019
```
Removes the need for temporary VMM mappings.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>
```
a788ade4
drm/nouveau/dmem: use physical vram addresses during migration copies · fd5e9856
Ben Skeggs authored Feb 15, 2019
```
Removes the need for temporary VMM mappings.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>
```
fd5e9856
drm/nouveau/dmem: extend copy function to allow direct use of physical addresses · 6c762d1b
Ben Skeggs authored Feb 15, 2019
```
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>
```
6c762d1b

drm/nouveau/svm: new ioctl to migrate process memory to GPU memory · f180bf12

Jérôme Glisse authored Aug 07, 2018

This add an ioctl to migrate a range of process address space to the
device memory. On platform without cache coherent bus (x86, ARM, ...)
this means that CPU can not access that range directly, instead CPU
will fault which will migrate the memory back to system memory.

This is behind a staging flag so that we can evolve the API.
Signed-off-by: Jérôme Glisse <jglisse@redhat.com>

f180bf12

drm/nouveau/dmem: device memory helpers for SVM · 5be73b69

Jérôme Glisse authored Jul 26, 2018

Device memory can be use in SVM, in which case we do not have any of
the existing buffer object. This commit add infrastructure to allow
use of device memory without nouveau_bo. Again this is a temporary
solution until a rework of GPU memory management.
Signed-off-by: Jérôme Glisse <jglisse@redhat.com>

5be73b69

drm/nouveau/svm: initial support for shared virtual memory · eeaf06ac

Ben Skeggs authored Jul 05, 2018

This uses HMM to mirror a process' CPU page tables into a channel's page
tables, and keep them synchronised so that both the CPU and GPU are able
to access the same memory at the same virtual address.

While this code also supports Volta/Turing, it's only enabled for Pascal
GPUs currently due to channel recovery being unreliable right now on the
later GPUs.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

eeaf06ac

drm/nouveau: prepare for enabling svm with existing userspace interfaces · bfe91afa

Ben Skeggs authored Feb 19, 2019

For a channel to make use of SVM features, it requires a different GPU MMU
configuration than we would normally use, which is not desirable to switch
to unless a client is actively going to use SVM.

In order to supporting SVM without more extensive changes to the userspace
interfaces, the SVM_INIT ioctl needs to replace the previous configuration
safely.

The only way we can currently do this safely, accounting for some unlikely
failure conditions, is to allocate the new VMM without destroying the last
one, and prioritising the SVM-enabled configuration in the code that cares.

This will get cleaned up again further down the track.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

bfe91afa

drm/nouveau/fault/gv100-: expose VoltaFaultBufferA · a261a20c

Ben Skeggs authored May 08, 2018

This nvclass exposes the replayable fault buffer, which will be used
by SVM to manage GPU page faults.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

a261a20c

drm/nouveau/fault/gp100: expose MaxwellFaultBufferA · 13e95729

Ben Skeggs authored May 08, 2018

This nvclass exposes the replayable fault buffer, which will be used
by SVM to manage GPU page faults.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

13e95729

drm/nouveau/mmu/gp100-: support vmms with gcc/tex replayable faults enabled · ab2ee9ff

Ben Skeggs authored May 08, 2018

Some GPU units are capable of supporting "replayable" page faults, where
the execution unit will wait for SW to fixup GPU page tables rather than
triggering a channel-fatal fault.

This feature isn't useful (it's harmful, even) unless something like HMM
is being used to manage events appearing in the replayable fault buffer,
so, it's disabled by default.

This commit allows a client to request it be enabled.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

ab2ee9ff

drm/nouveau/mmu/gp100-: add privileged methods for fault replay/cancel · 71871aa6

Ben Skeggs authored Jul 09, 2018

Host methods exist to do at least some of what we need, but we are not
currently pushing replay/cancels through a channel like UVM does as it's
not clear whether it's necessary in our case (UVM also updates PTEs with
the GPU).

UVM also pushes a software method for fault cancels on Pascal, seemingly
because the host methods don't appear to be sufficient.  If/when we want
to push the replay/cancel on the GPU, we can re-purpose the cancellation
code here to implement that swmthd.

Keep it simple for now, until we figure out exactly what we need here.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

71871aa6

drm/nouveau/mmu: add a privileged method to directly manage PTEs · a5ff307f

Ben Skeggs authored Jul 07, 2018

This provides a somewhat more direct method of manipulating the GPU page
tables, which will be required to support SVM.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

a5ff307f

drm/nouveau/mmu: store mapped flag separately from memory pointer · 8e68271d

Ben Skeggs authored Jul 07, 2018

This will be used to support a privileged client providing PTEs directly,
without a memory object to use as a reference.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

8e68271d

drm/nouveau/mmu: support initialisation of client-managed address-spaces · 2606f291

Ben Skeggs authored Jun 13, 2018

NVKM is currently responsible for managing the allocation of a client's
GPU address-space, but there's various use-cases (ie. HMM address-space
mirroring) where giving a client more direct control is desirable.

This commit allows for a VMM to be created where the area allocated for
NVKM is limited to a client-specified window, the remainder of address-
space is controlled directly by the client.

Leaving a window is necessary to support various internal requirements,
but also to support existing allocation interfaces as not all of the HW
is capable of working with a HMM allocation.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

2606f291

drm/nouveau/gr/gf100-: expose method to determine current context · ae5ea7f6
Ben Skeggs authored Feb 05, 2019
```
MMU will need access to this info.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>
```
ae5ea7f6

drm/nouveau/gr/gf100-: expose fecs methods for pausing ctxsw · 169f30b3

Ben Skeggs authored Feb 01, 2019

MMU will need access to these.

v2. Apply fix from Rhys Kidd to send correct FECS method for STOP_CTXSW.
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>

169f30b3

drm/nouveau/falcon: fix a few indentation issues · 8e083686

Colin Ian King authored Feb 12, 2019

There are a few statements that are indented incorrectly. Fix these.
Signed-off-by: Colin Ian King <colin.king@canonical.com>
Signed-off-by: Ben Skeggs <bskeggs@redhat.com>