linux - Linux Kernel (branches are rebased on master from time to time)

Age	Commit message (Collapse)	Author	Files	Lines
2006-09-24	ocfs2: Teach ocfs2_drop_lock() to use ->set_lvb() callback	Mark Fasheh	1	-31/+28
	With this, we don't need to pass an additional struct with function pointer. Now that the callbacks are fully used, comment the remaining API. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Remove ->unblock lockres operation	Mark Fasheh	1	-146/+6
	Have ocfs2_process_blocked_lock() call ocfs2_generic_unblock_lock(), which gets to be ocfs2_unblock_lock() now that it's the only possible unblock function. Remove the ->unblock() callback from the structure, and all lock type specific unblock functions. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: move downconvert worker to lockres ops	Mark Fasheh	1	-18/+32
	This way lock types don't have to manually pass it to ocfs2_generic_unblock_lock(). Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Remove unused dlmglue functions	Mark Fasheh	1	-103/+0
	The meta data unblocking code no longer needs ocfs2_do_unblock_meta() or ocfs2_can_downconvert_meta_lock(), so remove them. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Have the metadata lock use generic dlmglue functions	Mark Fasheh	1	-1/+31
	Fill in the ->check_downconvert and ->set_lvb callbacks with meta data specific operations and switch ocfs2_unblock_meta() to call ocfs2_generic_unblock_lock() Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Add ->set_lvb callback in dlmglue	Mark Fasheh	1	-2/+29
	This allows a lock type to set the value block before downconvert. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Add ->check_downconvert callback in dlmglue	Mark Fasheh	1	-1/+18
	This will allow lock types to force a requeue of a lock downconvert. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Check for refreshing locks in generic unblock function	Mark Fasheh	1	-12/+19
	Tidy up the exit path a bit too. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: don't unconditionally pass LVB flags	Mark Fasheh	1	-3/+15
	Allow a lock type to specifiy whether it makes use of the LVB. The only type which does this right now is the meta data lock. This should save us some space on network messages since they won't have to needlessly transmit value blocks. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: combine inode and generic blocking AST functions	Mark Fasheh	1	-112/+11
	There is extremely little difference between the two now. We can remove the callback from ocfs2_lock_res_ops as well. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Add ->get_osb() dlmglue locking operation	Mark Fasheh	1	-0/+33
	Will be used to find the ocfs2_super structure from a given lockres. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: remove ->unlock_ast() callback from ocfs2_lock_res_ops	Mark Fasheh	1	-13/+3
	This was always defined to the same function in all locks, so clean things up by removing and passing ocfs2_unlock_ast() directly to the DLM. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: combine inode and generic AST functions	Mark Fasheh	1	-110/+10
	There is extremely little difference between the two now. We can remove the callback from ocfs2_lock_res_ops as well. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Clean up lock resource refresh flags	Mark Fasheh	1	-14/+35
	Use of the refresh mechanism is lock-type wide, so move knowledge of that to the ocfs2_lock_res_ops structure. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Remove i_generation from inode lock names	Mark Fasheh	10	-53/+170
	OCFS2 puts inode meta data in the "lock value block" provided by the DLM. Typically, i_generation is encoded in the lock name so that a deleted inode on and a new one in the same block don't share the same lvb. Unfortunately, that scheme means that the read in ocfs2_read_locked_inode() is potentially thrown away as soon as the meta data lock is taken - we cannot encode the lock name without first knowing i_generation, which requires a disk read. This patch encodes i_generation in the inode meta data lvb, and removes the value from the inode meta data lock name. This way, the read can be covered by a lock, and at the same time we can distinguish between an up to date and a stale LVB. This will help cold-cache stat(2) performance in particular. Since this patch changes the protocol version, we take the opportunity to do a minor re-organization of two of the LVB fields. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Encode i_generation in the meta data lvb	Mark Fasheh	2	-7/+12
	When i_generation is removed from the lockname, this will help us determine whether a meta data lvb has information that is in sync with the local struct inode. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Free up some space in the lvb	Mark Fasheh	2	-4/+6
	lvb_version doesn't need to be a whole 32 bits. Make it an 8 bit field to free up some space. This should be backwards compatible until we use one of the fields, in which case we'd bump the lvb version anyway. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Remove special casing for inode creation in ocfs2_dentry_attach_lock()	Mark Fasheh	3	-37/+14
	We can't use LKM_LOCAL for new dentry locks because an unlink and subsequent re-create of a name/inode pair may result in the lock still being mastered somewhere in the cluster. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: manually d_move() during ocfs2_rename()	Mark Fasheh	2	-2/+5
	Make use of FS_RENAME_DOES_D_MOVE to avoid a race condition that can occur during ->rename() if we d_move() outside of the parent directory cluster locks, and another node discovers the new name (created during the rename) and unlinks it. d_move() will unconditionally rehash a dentry - which will leave stale data in the system. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	[PATCH] Allow file systems to manually d_move() inside of ->rename()	Mark Fasheh	3	-10/+9
	Some file systems want to manually d_move() the dentries involved in a rename. We can do this by making use of the FS_ODD_RENAME flag if we just have nfs_rename() unconditionally do the d_move(). While there, we rename the flag to be more descriptive. OCFS2 uses this to protect that part of the rename operation with a cluster lock. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com> Cc: Trond Myklebust <trond.myklebust@fys.uio.no> Cc: Al Viro <viro@zeniv.linux.org.uk> Cc: Christoph Hellwig <hch@lst.de> Signed-off-by: Andrew Morton <akpm@osdl.org>
2006-09-24	ocfs2: Remove the dentry vote	Mark Fasheh	2	-183/+2
	This is unused now. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Hook rest of the file system into dentry locking API	Mark Fasheh	4	-41/+94
	Actually replace the vote calls with the new dentry operations. Make any necessary adjustments to get the scheme to work. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Add dentry tracking API	Mark Fasheh	3	-32/+369
	Replace the dentry vote mechanism with a cluster lock which covers a set of dentries. This allows us to force d_delete() only on nodes which actually care about an unlink. Every node that does a ->lookup() gets a read only lock on the dentry, until an unlink during which the unlinking node, will request an exclusive lock, forcing the other nodes who care about that dentry to d_delete() it. The effect is that we retain a very lightweight ->d_revalidate(), and at the same time get to make large improvements to the average case performance of the ocfs2 unlink and rename operations. This patch adds the higher level API and the dentry manipulation code. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Add new cluster lock type	Mark Fasheh	5	-104/+436
	Replace the dentry vote mechanism with a cluster lock which covers a set of dentries. This allows us to force d_delete() only on nodes which actually care about an unlink. Every node that does a ->lookup() gets a read only lock on the dentry, until an unlink during which the unlinking node, will request an exclusive lock, forcing the other nodes who care about that dentry to d_delete() it. The effect is that we retain a very lightweight ->d_revalidate(), and at the same time get to make large improvements to the average case performance of the ocfs2 unlink and rename operations. This patch adds the cluster lock type which OCFS2 can attach to dentries. A small number of fs/ocfs2/dcache.c functions are stubbed out so that this change can compile. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Update dlmglue for new dlmlock() API	Mark Fasheh	1	-0/+3
	File system lock names are very regular right now, so we really only need to pass an extra parameter to dlmlock(). Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Update dlmfs for new dlmlock() API	Mark Fasheh	2	-51/+31
	We just need to add a namelen field to the user_lock_res structure, and update a few debug prints. Instead of updating all debug prints, I took the opportunity to remove a few that are likely unnecessary these days. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Allow binary names in the DLM	Mark Fasheh	5	-8/+11
	The OCFS2 DLM uses strlen() to determine lock name length, which excludes the possibility of putting binary values in the name string. Fix this by requiring that string length be passed in as a parameter. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	ocfs2: Silence dlm error print	Mark Fasheh	1	-3/+3
	An AST can be delivered via the network after a lock has been removed, so no need to print an error when we see that. Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>
2006-09-24	Move several *_SUPER_MAGIC symbols to include/linux/magic.h.	Jeff Garzik	9	-7/+5
	Signed-off-by: Jeff Garzik <jeff@garzik.org>
2006-09-22	NFS: unmark NFS direct I/O as experimental	Chuck Lever	1	-2/+2
	Remove the EXPERIMENTAL flag from the NFS_DIRECTIO option. Test plan: Unset the EXPERIMENTAL kernel build option and check to see that the NFS direct I/O option is still available. Signed-off-by: Chuck Lever <chuck.lever@oracle.com> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: add comments clarifying the use of nfs_post_op_update()	Chuck Lever	2	-1/+13
	Comments-only change to clarify a detail of the NFS protocol and how it is implemented in Linux. Test plan: None. Signed-off-by: Chuck Lever <chuck.lever@oracle.com> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: Use SEEK_END instead of hardcoded value	Josef 'Jeff' Sipek	1	-1/+1
	Signed-off-by: Josef 'Jeff' Sipek <jeffpc@josefsipek.net> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFSv4: When mounting with a port=0 argument, substitute port=2049	Trond Myklebust	1	-0/+3
	RFC3530 states that the registered port 2049 for the NFS protocol should be the default configuration in order to allow clients not to use the RPC binding protocols. If the mount program sends us a port=0, we therefore substitute port=2049. Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFSv4: Poll more aggressively when handling NFS4ERR_DELAY	Trond Myklebust	1	-1/+1
	Change the initial retry delay from 1s to 0.1s (and then back off exponentially). Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFSv4: Handle the condition NFS4ERR_FILE_OPEN	Trond Myklebust	1	-0/+1
	Retry a few times before we give up: the error is usually due to ordering issues with asynchronous RPC calls. Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFSv4: Retry lease recovery if it failed during a synchronous operation.	Trond Myklebust	1	-2/+9
	Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: Don't invalidate the symlink we just stuffed into the cache	Trond Myklebust	1	-7/+5
	And slight optimisation of nfs_end_data_update(): directories never have delegations anyway. Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: Make read() return an ESTALE if the file has been deleted	Trond Myklebust	1	-3/+16
	Currently, a read() request will return EIO even if the file has been deleted on the server, simply because that is what the VM will return if the call to readpage() fails to update the page. Ensure that readpage() marks the inode as stale if it receives an ESTALE. Then return that error to userland. Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFSv4: It's perfectly legal for clp to be NULL here....	J. Bruce Fields	1	-1/+1
	Signed-off-by: J. Bruce Fields <bfields@citi.umich.edu> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: nfs_lookup - don't hash dentry when optimising away the lookup	Trond Myklebust	1	-3/+11
	If the open intents tell us that a given lookup is going to result in a, exclusive create, we currently optimize away the lookup call itself. The reason is that the lookup would not be atomic with the create RPC call, so why do it in the first place? A problem occurs, however, if the VFS aborts the exclusive create operation after the lookup, but before the call to create the file/directory: in this case we will end up with a hashed negative dentry in the dcache that has never been looked up. Fix this by only actually hashing the dentry once the create operation has been successfully completed. Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	Fix a referral error Oops	andros@citi.umich.edu	1	-0/+2
	Fix an oops when the referral server is not responding. Check the error return from nfs4_set_client() in nfs4_create_referral_server. Signed-off-by: Andy Adamson <andros@citi.umich.edu> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: NFS_ROOT should use the new rpc_create API	Chuck Lever	1	-16/+13
	Teach NFS_ROOT to use the new rpc_create API instead of the old two-call API for creating an RPC transport. Test plan: Compile the kernel with the NFS client build-in, and set CONFIG_NFS_ROOT. Signed-off-by: Chuck Lever <chuck.lever@oracle.com> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: Fix up compiler warnings on 64-bit platforms in client.c	David Howells	1	-6/+14
	Fix up warnings from compiling on ppc64. Signed-Off-By: David Howells <dhowells@redhat.com> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	SUNRPC: Make rpc_mkpipe() take the parent dentry as an argument	Trond Myklebust	1	-5/+1
	Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFSv4: Fix a use-after-free issue with the nfs server.	Trond Myklebust	3	-18/+27
	Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	Add a real API for dealing with blk_congestion_wait()	Trond Myklebust	1	-0/+1
	Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: Use cached page as buffer for NFS symlink requests	Chuck Lever	7	-35/+49
	Now that we have a copy of the symlink path in the page cache, we can pass a struct page down to the XDR routines instead of a string buffer. Test plan: Connectathon, all NFS versions. Signed-off-by: Chuck Lever <chuck.lever@oracle.com> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: copy symlinks into page cache before sending NFS SYMLINK request	Chuck Lever	1	-18/+68
	Currently the NFS client does not cache symlinks it creates. They get cached only when the NFS client reads them back from the server. Copy the symlink into the page cache before sending it. Test plan: Connectathon, all NFS versions. Signed-off-by: Chuck Lever <chuck.lever@oracle.com> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: Fix double d_drop in nfs_instantiate() error path	Chuck Lever	4	-45/+57
	If the LOOKUP or GETATTR in nfs_instantiate fail, nfs_instantiate will do a d_drop before returning. But some callers already do a d_drop in the case of an error return. Make certain we do only one d_drop in all error paths. This issue was introduced because over time, the symlink proc API diverged slightly from the create/mkdir/mknod proc API. To prevent other coding mistakes of this type, change the symlink proc API to be more like create/mkdir/mknod and move the nfs_instantiate call into the symlink proc routines so it is used in exactly the same way for create, mkdir, mknod, and symlink. Test plan: Connectathon, all versions of NFS. Signed-off-by: Chuck Lever <chuck.lever@oracle.com> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>
2006-09-22	NFS: remove a no-longer-needed error check in nfs_symlink()	Chuck Lever	1	-6/+2
	In the early days of NFS, there was no duplicate reply cache on the server. Thus retransmitted non-idempotent requests often found that the request had already completed on the server. To avoid passing an unanticipated return code to unsuspecting applications, NFS clients would often shunt error codes that implied the request had been retried but already completed. Thanks to NFS over TCP, duplicate reply caches on the server, and network performance and reliability improvements, it is safe to remove such checks. Test plan: None. Signed-off-by: Chuck Lever <chuck.lever@oracle.com> Signed-off-by: Trond Myklebust <Trond.Myklebust@netapp.com>