vfs: fix data corruption when blocksize < pagesize for mmaped data

->page_mkwrite() is used by filesystems to allocate blocks under a page which is becoming writeably mmapped in some process' address space. This allows a filesystem to return a page fault if there is not enough space available, user exceeds quota or similar problem happens, rather than silently discarding data later when writepage is called. However VFS fails to call ->page_mkwrite() in all the cases where filesystems need it when blocksize < pagesize. For example when blocksize = 1024, pagesize = 4096 the following is problematic: ftruncate(fd, 0); pwrite(fd, buf, 1024, 0); map = mmap(NULL, 1024, PROT_WRITE, MAP_SHARED, fd, 0); map[0] = 'a'; ----> page_mkwrite() for index 0 is called ftruncate(fd, 10000); /* or even pwrite(fd, buf, 1, 10000) */ mremap(map, 1024, 10000, 0); map[4095] = 'a'; ----> no page_mkwrite() called At the moment ->page_mkwrite() is called, filesystem can allocate only one block for the page because i_size == 1024. Otherwise it would create blocks beyond i_size which is generally undesirable. But later at ->writepage() time, we also need to store data at offset 4095 but we don't have block allocated for it. This patch introduces a helper function filesystems can use to have ->page_mkwrite() called at all the necessary moments. Signed-off-by: Jan Kara <jack@suse.cz> Signed-off-by: Theodore Ts'o <tytso@mit.edu> Cc: stable@vger.kernel.org
author: Jan Kara <jack@suse.cz> 2014-10-01 21:49:18 -0400
committer: Theodore Ts'o <tytso@mit.edu> 2014-10-01 21:49:18 -0400
commit: 90a8020278c1598fafd071736a0846b38510309c (patch)
tree: 2ab461b549a2b5f6b933895b1e61eb98627bba94 /fs/buffer.c
parent: f6e63f90809946d410c42045577cb159fedabf8c (diff)
download: lwn-90a8020278c1598fafd071736a0846b38510309c.tar.gz
lwn-90a8020278c1598fafd071736a0846b38510309c.zip
1 files changed, 3 insertions, 0 deletions
diff --git a/fs/buffer.c b/fs/buffer.c
index 9a6029e0dd71..6dc1475dcb2d 100644
--- a/fs/buffer.c
+++ b/fs/buffer.c
@@ -2087,6 +2087,7 @@ int generic_write_end(struct file *file, struct address_space *mapping,
 			struct page *page, void *fsdata)
 {
 	struct inode *inode = mapping->host;
+	loff_t old_size = inode->i_size;
 	int i_size_changed = 0;
 
 	copied = block_write_end(file, mapping, pos, len, copied, page, fsdata);
@@ -2106,6 +2107,8 @@ int generic_write_end(struct file *file, struct address_space *mapping,
 	unlock_page(page);
 	page_cache_release(page);
 
+	if (old_size < pos)
+		pagecache_isize_extended(inode, old_size, pos);
 	/*
 	 * Don't mark the inode dirty under page lock. First, it unnecessarily
 	 * makes the holding time of page lock longer. Second, it forces lock
author	Jan Kara <jack@suse.cz>	2014-10-01 21:49:18 -0400
committer	Theodore Ts'o <tytso@mit.edu>	2014-10-01 21:49:18 -0400
commit	90a8020278c1598fafd071736a0846b38510309c (patch)
tree	2ab461b549a2b5f6b933895b1e61eb98627bba94 /fs/buffer.c
parent	f6e63f90809946d410c42045577cb159fedabf8c (diff)
download	lwn-90a8020278c1598fafd071736a0846b38510309c.tar.gz lwn-90a8020278c1598fafd071736a0846b38510309c.zip