From: "Benjamin Coddington" <bcodding@redhat.com>
To: "Trond Myklebust" <trondmy@primarydata.com>
Cc: "hch@infradead.org" <hch@infradead.org>,
"List Linux NFS Mailing" <linux-nfs@vger.kernel.org>
Subject: Re: [PATCH v4 24/28] Getattr doesn't require data sync semantics
Date: Wed, 27 Jul 2016 12:14:32 -0400 [thread overview]
Message-ID: <CFD74F0F-5375-4C05-8918-516F730FDBA7@redhat.com> (raw)
In-Reply-To: <06BBFA63-02D9-4BDE-A61D-14FE9F4E9E5F@primarydata.com>
On 27 Jul 2016, at 8:31, Trond Myklebust wrote:
>> On Jul 27, 2016, at 08:15, Trond Myklebust <trondmy@primarydata.com>
>> wrote:
>>
>>
>>> On Jul 27, 2016, at 07:55, Benjamin Coddington <bcodding@redhat.com>
>>> wrote:
>>>
>>> After adding more debugging, I see that all of that is working
>>> correctly,
>>> but the first LAYOUTCOMMIT is taking the size back down to 4096 from
>>> the
>>> last nfs_writeback_done(), and the second LAYOUTCOMMIT never brings
>>> it back
>>> up again.
>>>
>>
>> Excellent! Thanks for debugging that.
>>
>>> Now I see that we should be marking the block extents as written
>>> atomically with
>>> setting LAYOUTCOMMIT and nfsi->layout->plh_lwb, otherwise a
>>> LAYOUTCOMMIT can
>>> collect extents just added from the next bl_write_cleanup(). Then,
>>> the next
>>> LAYOUTCOMMIT fails, and all we're left with is the size from the
>>> first
>>> LAYOUTCOMMIT. Not sure if that particular problem is the whole fix,
>>> but
>>> that's something to work on.
>>>
>>> I see ways to fix that:
>>>
>>> - make a new pnfs_set_layoutcommit_locked() that can be used to
>>> call
>>> ext_tree_mark_written() inside the i_lock
>>>
>>> - make another pnfs_layoutdriver_type operation to be used within
>>> pnfs_set_layoutcommit (mark_layoutcommit? set_layoutcommit?),
>>> and call
>>> ext_tree_mark_written() within that..
>>>
>>> - have .prepare_layoutcommit return a new positive plh_lwb that
>>> would
>>> extend the current LAYOUTCOMMIT
>>>
>>> - make ext_tree_prepare_commit only encode up to plh_lwb
>>
>> I see no reason why ext_tree_prepare_commit() shouldn’t be allowed
>> to extend the args->lastbytewritten. This is a metadata operation
>> that is owned by the pNFS layout driver.
>> The only thing I’d note is you should then rewrite the failure case
>> in pnfs_layoutcommit_inode() so that it doesn’t rely on the saved
>> “end_pos”, but uses args->lastbytewritten instead (with a comment
>> to the effect why)…
>
> In fact, given the potential for races here, I think the right thing
> to do is to have ext_tree_prepare_commit() always set the correct
> value for args->lastbytewritten.
OK, that has cleared up that common failure case that was getting in the
way, but now it can still fail like this:
nfs_writeback_update_inode sets size 4096 w/ NFS_INO_INVALID_ATTR set,
and sets NFS_INO_LAYOUTCOMMIT
1st nfs_getattr -> pnfs_layoutcommit_inode starts, clears layoutcommit
flag sets NFS_INO_LAYOUTCOMMITING
nfs_writeback_update_inode sets size 8192 w/ NFS_INO_INVALID_ATTR set,
and sets NFS_INO_LAYOUTCOMMIT
1st nfs_getattr -> nfs4_layoutcommit_release sets size 4096,
NFS_INO_INVALID_ATTR set, clears NFS_INO_LAYOUTCOMMITTING
1st nfs_getattr -> __revalidate_inode sets size 4096,
NFS_INO_INVALID_ATTR not set.. cache is valid
2nd nfs_getattr immediately returns 4096 even though
NFS_INO_LAYOUTCOMMIT
Ben
next prev parent reply other threads:[~2016-07-27 16:12 UTC|newest]
Thread overview: 69+ messages / expand[flat|nested] mbox.gz Atom feed top
2016-07-06 22:29 [PATCH v4 00/28] NFS writeback performance patches for v4.8 Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 01/28] NFS: Don't flush caches for a getattr that races with writeback Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 02/28] NFS: Cache access checks more aggressively Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 03/28] NFS: Cache aggressively when file is open for writing Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 04/28] NFS: Kill NFS_INO_NFS_INO_FLUSHING: it is a performance killer Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 05/28] NFS: writepage of a single page should not be synchronous Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 06/28] NFS: Don't hold the inode lock across fsync() Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 07/28] NFS: Don't call COMMIT in ->releasepage() Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 08/28] pNFS/files: Fix layoutcommit after a commit to DS Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 09/28] pNFS/flexfiles: " Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 10/28] pNFS/flexfiles: Clean up calls to pnfs_set_layoutcommit() Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 11/28] pNFS: Files and flexfiles always need to commit before layoutcommit Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 12/28] pNFS: Ensure we layoutcommit before revalidating attributes Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 13/28] pNFS: pnfs_layoutcommit_outstanding() is no longer used when !CONFIG_NFS_V4_1 Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 14/28] NFS: Fix O_DIRECT verifier problems Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 15/28] NFS: Ensure we reset the write verifier 'committed' value on resend Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 16/28] NFS: Remove racy size manipulations in O_DIRECT Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 17/28] NFS Cleanup: move call to generic_write_checks() into fs/nfs/direct.c Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 18/28] NFS: Move buffered I/O locking into nfs_file_write() Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 19/28] NFS: Do not serialise O_DIRECT reads and writes Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 20/28] NFS: Cleanup nfs_direct_complete() Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 21/28] NFS: Remove redundant waits for O_DIRECT in fsync() and write_begin() Trond Myklebust
2016-07-06 22:29 ` [PATCH v4 22/28] NFS: Remove unused function nfs_revalidate_mapping_protected() Trond Myklebust
2016-07-06 22:30 ` [PATCH v4 23/28] NFS: Do not aggressively cache file attributes in the case of O_DIRECT Trond Myklebust
2016-07-06 22:30 ` [PATCH v4 24/28] NFS: Getattr doesn't require data sync semantics Trond Myklebust
2016-07-06 22:30 ` [PATCH v4 25/28] NFSv4.2: Fix a race in nfs42_proc_deallocate() Trond Myklebust
2016-07-06 22:30 ` [PATCH v4 26/28] NFSv4.2: Fix writeback races in nfs4_copy_file_range Trond Myklebust
2016-07-06 22:30 ` [PATCH v4 27/28] NFSv4.2: llseek(SEEK_HOLE) and llseek(SEEK_DATA) don't require data sync Trond Myklebust
2016-07-06 22:30 ` [PATCH v4 28/28] NFS nfs_vm_page_mkwrite: Don't freeze me, Bro Trond Myklebust
2016-07-18 3:48 ` [PATCH v4 24/28] NFS: Getattr doesn't require data sync semantics Christoph Hellwig
2016-07-18 4:32 ` Trond Myklebust
2016-07-18 4:59 ` Trond Myklebust
2016-07-19 3:58 ` hch
2016-07-19 20:00 ` [PATCH v4 24/28] " Benjamin Coddington
2016-07-19 20:06 ` Trond Myklebust
2016-07-20 15:03 ` Benjamin Coddington
2016-07-21 8:22 ` hch
2016-07-21 8:32 ` Benjamin Coddington
2016-07-21 9:10 ` Benjamin Coddington
2016-07-21 9:52 ` Benjamin Coddington
2016-07-21 12:46 ` Trond Myklebust
2016-07-21 13:05 ` Benjamin Coddington
2016-07-21 13:20 ` Trond Myklebust
2016-07-21 14:00 ` Trond Myklebust
2016-07-21 14:02 ` Benjamin Coddington
2016-07-25 16:26 ` Benjamin Coddington
2016-07-25 16:39 ` Trond Myklebust
2016-07-25 18:26 ` Benjamin Coddington
2016-07-25 18:34 ` Trond Myklebust
2016-07-25 18:41 ` Benjamin Coddington
2016-07-26 16:32 ` Benjamin Coddington
2016-07-26 16:35 ` Trond Myklebust
2016-07-26 17:57 ` Benjamin Coddington
2016-07-26 18:07 ` Trond Myklebust
2016-07-27 11:55 ` Benjamin Coddington
2016-07-27 12:15 ` Trond Myklebust
2016-07-27 12:31 ` Trond Myklebust
2016-07-27 16:14 ` Benjamin Coddington [this message]
2016-07-27 18:05 ` Trond Myklebust
2016-07-28 9:47 ` Benjamin Coddington
2016-07-28 12:31 ` Trond Myklebust
2016-07-28 14:04 ` Trond Myklebust
2016-07-28 15:38 ` Benjamin Coddington
2016-07-28 15:39 ` Trond Myklebust
2016-07-28 15:33 ` Benjamin Coddington
2016-07-28 15:36 ` Trond Myklebust
2016-07-28 16:40 ` Benjamin Coddington
2016-07-28 16:41 ` Trond Myklebust
2016-07-19 20:09 ` Benjamin Coddington
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=CFD74F0F-5375-4C05-8918-516F730FDBA7@redhat.com \
--to=bcodding@redhat.com \
--cc=hch@infradead.org \
--cc=linux-nfs@vger.kernel.org \
--cc=trondmy@primarydata.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox;
as well as URLs for NNTP newsgroup(s).