linux - linux

	Commit message (Collapse)	Author	Age	Files	Lines
*	ceph: make sure syncfs flushes all cap snaps	Yan, Zheng	2015-06-25	1	-24/+62
\| \| \| \|	Signed-off-by: Yan, Zheng <zyan@redhat.com>
*	ceph: don't trim auth cap when there are cap snaps	Yan, Zheng	2015-06-25	1	-1/+2
\| \| \| \|	Signed-off-by: Yan, Zheng <zyan@redhat.com>
*	ceph: check OSD caps before read/write	Yan, Zheng	2015-06-25	1	-0/+4
\| \| \| \|	Signed-off-by: Yan, Zheng <zyan@redhat.com>
*	Merge branch 'for-linus' of ↵	Linus Torvalds	2015-04-27	1	-12/+12
\|\ \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	git://git.kernel.org/pub/scm/linux/kernel/git/viro/vfs Pull fourth vfs update from Al Viro: "d_inode() annotations from David Howells (sat in for-next since before the beginning of merge window) + four assorted fixes" * 'for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/viro/vfs: RCU pathwalk breakage when running into a symlink overmounting something fix I_DIO_WAKEUP definition direct-io: only inc/dec inode->i_dio_count for file systems fs/9p: fix readdir() VFS: assorted d_backing_inode() annotations VFS: fs/inode.c helpers: d_inode() annotations VFS: fs/cachefiles: d_backing_inode() annotations VFS: fs library helpers: d_inode() annotations VFS: assorted weird filesystems: d_inode() annotations VFS: normal filesystems (and lustre): d_inode() annotations VFS: security/: d_inode() annotations VFS: security/: d_backing_inode() annotations VFS: net/: d_inode() annotations VFS: net/unix: d_backing_inode() annotations VFS: kernel/: d_inode() annotations VFS: audit: d_backing_inode() annotations VFS: Fix up some ->d_inode accesses in the chelsio driver VFS: Cachefiles should perform fs modifications on the top layer only VFS: AF_UNIX sockets should call mknod on the top layer only
\| *	VFS: normal filesystems (and lustre): d_inode() annotations	David Howells	2015-04-15	1	-12/+12
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	that's the bulk of filesystem drivers dealing with inodes of their own Signed-off-by: David Howells <dhowells@redhat.com> Signed-off-by: Al Viro <viro@zeniv.linux.org.uk>
* \|	ceph: fix null pointer dereference in send_mds_reconnect()	Yan, Zheng	2015-04-22	1	-1/+2
\| \| \| \| \| \| \| \| \| \| \| \|	sb->s_root can be null when umounting Signed-off-by: Yan, Zheng <zyan@redhat.com>
* \|	ceph: cleanup unsafe requests when reconnecting is denied	Yan, Zheng	2015-04-20	1	-0/+28
\| \| \| \| \| \| \| \|	Signed-off-by: Yan, Zheng <zyan@redhat.com>
* \|	ceph: don't zero i_wrbuffer_ref when reconnecting is denied	Yan, Zheng	2015-04-20	1	-7/+0
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	remove_session_caps_cb() does not truncate dirty data in page cache, but zeros i_wrbuffer_ref/i_wrbuffer_ref_head. This will result negtive i_wrbuffer_ref/i_wrbuffer_ref_head Signed-off-by: Yan, Zheng <zyan@redhat.com>
* \|	ceph: don't mark dirty caps when there is no auth cap	Yan, Zheng	2015-04-20	1	-1/+1
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	No i_auth_cap means reconnecting to MDS was denied. So don't add new dirty caps. Signed-off-by: Yan, Zheng <zyan@redhat.com>
* \|	ceph: use msecs_to_jiffies for time conversion	Nicholas Mc Guire	2015-04-20	1	-1/+1
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This is only an API consolidation and should make things more readable it replaces var * HZ / 1000 by msecs_to_jiffies(var). Signed-off-by: Nicholas Mc Guire <hofrat@osadl.org> Signed-off-by: Yan, Zheng <zyan@redhat.com>
* \|	ceph: drop cap releases in requests composed before cap reconnect	Yan, Zheng	2015-04-20	1	-6/+13
\|/ \| \| \| \| \| \| \| \| \|	These cap releases are stale because MDS will re-establish client caps according to the cap reconnect messages. Note: MDS can detect stale cap messages, so these stale cap releases are harmless even we don't drop them. Signed-off-by: Yan, Zheng <zyan@redhat.com>
*	Merge branch 'for-linus' of ↵	Linus Torvalds	2015-02-19	1	-34/+93
\|\ \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	git://git.kernel.org/pub/scm/linux/kernel/git/sage/ceph-client Pull Ceph changes from Sage Weil: "On the RBD side, there is a conversion to blk-mq from Christoph, several long-standing bug fixes from Ilya, and some cleanup from Rickard Strandqvist. On the CephFS side there is a long list of fixes from Zheng, including improved session handling, a few IO path fixes, some dcache management correctness fixes, and several blocking while !TASK_RUNNING fixes. The core code gets a few cleanups and Chaitanya has added support for TCP_NODELAY (which has been used on the server side for ages but we somehow missed on the kernel client). There is also an update to MAINTAINERS to fix up some email addresses and reflect that Ilya and Zheng are doing most of the maintenance for RBD and CephFS these days. Do not be surprised to see a pull request come from one of them in the future if I am unavailable for some reason" * 'for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/sage/ceph-client: (27 commits) MAINTAINERS: update Ceph and RBD maintainers libceph: kfree() in put_osd() shouldn't depend on authorizer libceph: fix double __remove_osd() problem rbd: convert to blk-mq ceph: return error for traceless reply race ceph: fix dentry leaks ceph: re-send requests when MDS enters reconnecting stage ceph: show nocephx_require_signatures and notcp_nodelay options libceph: tcp_nodelay support rbd: do not treat standalone as flatten ceph: fix atomic_open snapdir ceph: properly mark empty directory as complete client: include kernel version in client metadata ceph: provide seperate {inode,file}_operations for snapdir ceph: fix request time stamp encoding ceph: fix reading inline data when i_size > PAGE_SIZE ceph: avoid block operation when !TASK_RUNNING (ceph_mdsc_close_sessions) ceph: avoid block operation when !TASK_RUNNING (ceph_get_caps) ceph: avoid block operation when !TASK_RUNNING (ceph_mdsc_sync) rbd: fix error paths in rbd_dev_refresh() ...
\| *	ceph: re-send requests when MDS enters reconnecting stage	Yan, Zheng	2015-02-19	1	-3/+26
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	So that MDS can check if any request is already completed and process completed requests in clientreplay stage. When completed requests are processed in clientreplay stage, MDS can avoid sending traceless replies. Signed-off-by: Yan, Zheng <zyan@redhat.com>
\| *	client: include kernel version in client metadata	Yan, Zheng	2015-02-19	1	-1/+2
\| \| \| \| \| \| \| \|	Signed-off-by: Yan, Zheng <zyan@redhat.com>
\| *	ceph: fix request time stamp encoding	Yan, Zheng	2015-02-19	1	-2/+10
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	struct timespec uses 'long' to present second and nanosecond. 'long' is 64 bits on 64bits machine. ceph MDS expects time stamp to be encoded as struct ceph_timespec, which uses 'u32' to present second and nanosecond. Signed-off-by: Yan, Zheng <zyan@redhat.com>
\| *	ceph: avoid block operation when !TASK_RUNNING (ceph_mdsc_close_sessions)	Yan, Zheng	2015-02-19	1	-9/+4
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	use an atomic variable to track number of sessions, this can avoid block operation inside wait loops. Signed-off-by: Yan, Zheng <zyan@redhat.com>
\| *	ceph: avoid block operation when !TASK_RUNNING (ceph_mdsc_sync)	Yan, Zheng	2015-02-19	1	-17/+34
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	check_cap_flush() calls mutex_lock(), which may block. So we can't use it as condition check function for wait_event(); Signed-off-by: Yan, Zheng <zyan@redhat.com>
\| *	ceph: improve reference tracking for snaprealm	Yan, Zheng	2015-02-19	1	-2/+7
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	When snaprealm is created, its initial reference count is zero. But in some rare cases, the newly created snaprealm is not referenced by anyone. This causes snaprealm with zero reference count not freed. The fix is set reference count of newly snaprealm to 1. The reference is return the function who requests to create the snaprealm. When the function finishes its job, it releases the reference. Signed-off-by: Yan, Zheng <zyan@redhat.com>
\| *	ceph: handle SESSION_FORCE_RO message	Yan, Zheng	2015-02-19	1	-0/+10
\| \| \| \| \| \| \| \| \| \| \| \|	mark session as readonly and wake up all cap waiters. Signed-off-by: Yan, Zheng <zyan@redhat.com>
* \|	ceph: move spinlocking into ceph_encode_locks_to_buffer and ceph_count_locks	Jeff Layton	2015-01-16	1	-4/+0
\|/ \| \| \| \| \| \| \| \|	There is only a single call site for each of these functions, and the caller takes the i_lock prior to calling them and drops it just afterward. Move the spinlocking into the functions instead. Signed-off-by: Jeff Layton <jlayton@primarydata.com> Acked-by: Christoph Hellwig <hch@lst.de>
*	ceph: parse inline data in MClientReply and MClientCaps	Yan, Zheng	2014-12-17	1	-0/+10
\| \| \| \|	Signed-off-by: Yan, Zheng <zyan@redhat.com>
*	ceph: message versioning fixes	John Spray	2014-12-17	1	-2/+5
\| \| \| \| \| \| \| \| \| \| \| \| \|	There were two places we were assigning version in host byte order instead of network byte order. Also in MSG_CLIENT_SESSION we weren't setting compat_version in the header to reflect continued compatability with older MDSs. Fixes: http://tracker.ceph.com/issues/9945 Signed-off-by: John Spray <john.spray@redhat.com> Reviewed-by: Sage Weil <sage@redhat.com>
*	libceph: message signature support	Yan, Zheng	2014-12-17	1	-0/+16
\| \| \| \|	Signed-off-by: Yan, Zheng <zyan@redhat.com>
*	ceph, rbd: delete unnecessary checks before two function calls	SF Markus Elfring	2014-12-17	1	-4/+2
\| \| \| \| \| \| \| \| \| \| \| \|	The functions ceph_put_snap_context() and iput() test whether their argument is NULL and then return immediately. Thus the test around the call is not needed. This issue was detected by using the Coccinelle software. Signed-off-by: Markus Elfring <elfring@users.sourceforge.net> [idryomov@redhat.com: squashed rbd.c hunk, changelog] Signed-off-by: Ilya Dryomov <idryomov@redhat.com>
*	ceph: fix file lock interruption	Yan, Zheng	2014-12-17	1	-0/+2
\| \| \| \| \| \| \| \| \| \| \| \| \|	When a lock operation is interrupted, current code sends a unlock request to MDS to undo the lock operation. This method does not work as expected because the unlock request can drop locks that have already been acquired. The fix is use the newly introduced CEPH_LOCK_FCNTL_INTR/CEPH_LOCK_FLOCK_INTR requests to interrupt blocked file lock request. These requests do not drop locks that have alread been acquired, they only interrupt blocked file lock request. Signed-off-by: Yan, Zheng <zyan@redhat.com>
*	ceph: export ceph_session_state_name function	John Spray	2014-10-14	1	-7/+7
\| \| \| \| \| \| \|	...so that it can be used from the ceph debugfs code when dumping session info. Signed-off-by: John Spray <john.spray@redhat.com>
*	ceph: use pagelist to present MDS request data	Yan, Zheng	2014-10-14	1	-5/+9
\| \| \| \| \| \| \| \| \| \| \| \| \|	Current code uses page array to present MDS request data. Pages in the array are allocated/freed by caller of ceph_mdsc_do_request(). If request is interrupted, the pages can be freed while they are still being used by the request message. The fix is use pagelist to present MDS request data. Pagelist is reference counted. Signed-off-by: Yan, Zheng <zyan@redhat.com> Reviewed-by: Sage Weil <sage@redhat.com>
*	libceph: reference counting pagelist	Yan, Zheng	2014-10-14	1	-1/+0
\| \| \| \| \| \| \|	this allow pagelist to present data that may be sent multiple times. Signed-off-by: Yan, Zheng <zyan@redhat.com> Reviewed-by: Sage Weil <sage@redhat.com>
*	ceph: send client metadata to MDS	John Spray	2014-10-14	1	-1/+70
\| \| \| \| \| \| \| \| \| \|	Implement version 2 of CEPH_MSG_CLIENT_SESSION syntax, which includes additional client metadata to allow the MDS to report on clients by user-sensible names like hostname. Signed-off-by: John Spray <john.spray@redhat.com> Reviewed-by: Yan, Zheng <zyan@redhat.com>
*	ceph: move ceph_find_inode() outside the s_mutex	Yan, Zheng	2014-10-14	1	-3/+4
\| \| \| \| \| \| \| \| \|	ceph_find_inode() may wait on freeing inode, using it inside the s_mutex may cause deadlock. (the freeing inode is waiting for OSD read reply, but dispatch thread is blocked by the s_mutex) Signed-off-by: Yan, Zheng <zyan@redhat.com> Reviewed-by: Sage Weil <sage@redhat.com>
*	ceph: make sure request isn't in any waiting list when kicking request.	Yan, Zheng	2014-10-14	1	-0/+1
\| \| \| \| \| \| \|	we may corrupt waiting list if a request in the waiting list is kicked. Signed-off-by: Yan, Zheng <zyan@redhat.com> Reviewed-by: Sage Weil <sage@redhat.com>
*	ceph: protect kick_requests() with mdsc->mutex	Yan, Zheng	2014-10-14	1	-2/+3
\| \| \| \| \|	Signed-off-by: Yan, Zheng <zyan@redhat.com> Reviewed-by: Sage Weil <sage@redhat.com>
*	ceph: trim unused inodes before reconnecting to recovering MDS	Yan, Zheng	2014-10-14	1	-10/+13
\| \| \| \| \| \| \|	So the recovering MDS does not need to fetch these ununsed inodes during cache rejoin. This may reduce MDS recovery time. Signed-off-by: Yan, Zheng <zyan@redhat.com>
*	ceph: fix kick_requests()	Yan, Zheng	2014-08-07	1	-2/+3
\| \| \| \| \| \| \|	__do_request() may unregister the request. So we should update iterator 'p' before calling __do_request() Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	ceph: reset r_resend_mds after receiving -ESTALE	Yan, Zheng	2014-07-14	1	-0/+1
\| \| \| \| \| \|	this makes __choose_mds() choose mds according caps Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	ceph: include time stamp in replayed MDS requests	Yan, Zheng	2014-07-08	1	-2/+8
\| \| \| \|	Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	Merge branch 'for-linus' of ↵	Linus Torvalds	2014-06-13	1	-1/+8
\|\ \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	git://git.kernel.org/pub/scm/linux/kernel/git/sage/ceph-client Pull Ceph updates from Sage Weil: "This has a mix of bug fixes and cleanups. Alex's patch fixes a rare race in RBD. Ilya's patches fix an ENOENT check when a second rbd image is mapped and a couple memory leaks. Zheng fixes several issues with fragmented directories and multiple MDSs. Josh fixes a spin/sleep issue, and Josh and Guangliang's patches fix setting and unsetting RBD images read-only. Naturally there are several other cleanups mixed in for good measure" * 'for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/sage/ceph-client: (23 commits) rbd: only set disk to read-only once rbd: move calls that may sleep out of spin lock range rbd: add ioctl for rbd ceph: use truncate_pagecache() instead of truncate_inode_pages() ceph: include time stamp in every MDS request rbd: fix ida/idr memory leak rbd: use reference counts for image requests rbd: fix osd_request memory leak in __rbd_dev_header_watch_sync() rbd: make sure we have latest osdmap on 'rbd map' libceph: add ceph_monc_wait_osdmap() libceph: mon_get_version request infrastructure libceph: recognize poolop requests in debugfs ceph: refactor readpage_nounlock() to make the logic clearer mds: check cap ID when handling cap export message ceph: remember subtree root dirfrag's auth MDS ceph: introduce ceph_fill_fragtree() ceph: handle cap import atomically ceph: pre-allocate ceph_cap struct for ceph_add_cap() ceph: update inode fields according to issued caps rbd: replace IS_ERR and PTR_ERR with PTR_ERR_OR_ZERO ...
\| *	ceph: include time stamp in every MDS request	Sage Weil	2014-06-06	1	-1/+8
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	We recently modified the client/MDS protocol to include a timestamp in the client request. This allows ctime updates to follow the client's clock in most cases, which avoids subtle problems when clocks are out of sync and timestamps are updated sometimes by the MDS clock (for most requests) and sometimes by the client clock (for cap writeback). Signed-off-by: Sage Weil <sage@inktank.com>
* \|	fs/ceph: replace pr_warning by pr_warn	Fabian Frederick	2014-06-07	1	-3/+3
\|/ \| \| \| \| \| \| \| \|	Update the last pr_warning callsites in fs branch Signed-off-by: Fabian Frederick <fabf@skynet.be> Cc: Sage Weil <sage@inktank.com> Signed-off-by: Andrew Morton <akpm@linux-foundation.org> Signed-off-by: Linus Torvalds <torvalds@linux-foundation.org>
*	ceph: flush cap release queue when trimming session caps	Yan, Zheng	2014-04-05	1	-0/+3
\| \| \| \|	Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	ceph: preallocate buffer for readdir reply	Yan, Zheng	2014-04-05	1	-15/+51
\| \| \| \| \| \| \|	Preallocate buffer for readdir reply. Limit number of entries in readdir reply according to the buffer size. Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	ceph: fix null pointer dereference in discard_cap_releases()	Yan, Zheng	2014-04-05	1	-9/+12
\| \| \| \| \| \| \| \|	send_mds_reconnect() may call discard_cap_releases() after all release messages have been dropped by cleanup_cap_releases() Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com> Reviewed-by: Sage Weil <sage@inktank.com>
*	ceph: do not assume r_old_dentry[_dir] always set together	Sage Weil	2014-04-03	1	-3/+4
\| \| \| \| \| \| \| \| \|	Do not assume that r_old_dentry implies that r_old_dentry_dir is also true. Separate out the ref cleanup and make the debugs dump behave when it is NULL. Signed-off-by: Sage Weil <sage@inktank.com> Reviewed-by: Yan, Zheng <zheng.z.yan@intel.com>
*	ceph: add open export target session helper	Yan, Zheng	2014-01-21	1	-15/+36
\| \| \| \|	Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	ceph: handle session flush message	Yan, Zheng	2014-01-21	1	-0/+19
\| \| \| \|	Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	ceph: handle -ESTALE reply	Yan, Zheng	2014-01-21	1	-20/+11
\| \| \| \| \| \| \| \| \|	Send requests that operate on path to directory's auth MDS if mode == USE_AUTH_MDS. Always retry using the auth MDS if got -ESTALE reply from non-auth MDS. Also clean up the code that handles auth MDS change. Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	ceph: fix trim caps	Yan, Zheng	2014-01-21	1	-6/+11
\| \| \| \| \| \| \| \|	- don't trim auth cap if there are flusing caps - don't trim auth cap if any 'write' cap is wanted - allow trimming non-auth cap even if the inode is dirty Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com>
*	libceph: all features fields must be u64	Ilya Dryomov	2013-12-31	1	-7/+7
\| \| \| \| \| \| \| \| \|	In preparation for ceph_features.h update, change all features fields from unsigned int/u32 to u64. (ceph.git has ~40 feature bits at this point.) Signed-off-by: Ilya Dryomov <ilya.dryomov@inktank.com> Reviewed-by: Sage Weil <sage@inktank.com>
*	ceph: wake up 'safe' waiters when unregistering request	Yan, Zheng	2013-11-23	1	-1/+2
\| \| \| \| \| \| \| \|	We also need to wake up 'safe' waiters if error occurs or request aborted. Otherwise sync(2)/fsync(2) may hang forever. Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com> Signed-off-by: Sage Weil <sage@inktank.com>
*	ceph: cleanup aborted requests when re-sending requests.	Yan, Zheng	2013-11-23	1	-1/+4
\| \| \| \| \| \| \| \| \| \|	Aborted requests usually get cleared when the reply is received. If MDS crashes, no reply will be received. So we need to cleanup aborted requests when re-sending requests. Signed-off-by: Yan, Zheng <zheng.z.yan@intel.com> Reviewed-by: Greg Farnum <greg@inktank.com> Signed-off-by: Sage Weil <sage@inktank.com>