freebsd-dev

Author	SHA1	Message	Date
Alan Cox	5285558ac2	- Change uma_zone_set_obj() to call kmem_alloc_nofault() instead of kmem_alloc_pageable(). The difference between these is that an errant memory access to the zone will be detected sooner with kmem_alloc_nofault(). The following changes serve to eliminate the following lock-order reversal reported by witness: 1st 0xc1a3c084 vm object (vm object) @ vm/swap_pager.c:1311 2nd 0xc07acb00 swap_pager swhash (swap_pager swhash) @ vm/swap_pager.c:1797 3rd 0xc1804bdc vm object (vm object) @ vm/uma_core.c:931 There is no potential deadlock in this case. However, witness is unable to recognize this because vm objects used by UMA have the same type as ordinary vm objects. To remedy this, we make the following changes: - Add a mutex type argument to VM_OBJECT_LOCK_INIT(). - Use the mutex type argument to assign distinct types to special vm objects such as the kernel object, kmem object, and UMA objects. - Define a static swap zone object for use by UMA. (Only static objects are assigned a special mutex type.)	2004-07-22 19:44:49 +00:00
Bruce M Simpson	9bd86a9861	Properly brucify a string by outdenting it.	2004-07-06 02:27:30 +00:00
Bruce M Simpson	0e3fe6e3e6	In swap_pager_getpages(), bp->b_dev can be NULL, particularly for the case of NFS mounted swap, so do not try to dereference it. While we're here, brucify the printf() call which happens when we time out on acquisition of vm_page_queue_mtx. PR: kern/67898 Submitted by: bde (style)	2004-06-23 15:15:07 +00:00
Poul-Henning Kamp	f3732fd15b	Second half of the dev_t cleanup. The big lines are: NODEV -> NULL NOUDEV -> NODEV udev_t -> dev_t udev2dev() -> findcdev() Various minor adjustments including handling of userland access to kernel space struct cdev etc.	2004-06-17 17:16:53 +00:00
Poul-Henning Kamp	89c9c53da0	Do the dreaded s/dev_t/struct cdev */ Bump __FreeBSD_version accordingly.	2004-06-16 09:47:26 +00:00
Alan Cox	5a32489377	Make vm_page's PG_ZERO flag immutable between the time of the page's allocation and deallocation. This flag's principal use is shortly after allocation. For such cases, clearing the flag is pointless. The only unusual use of PG_ZERO is in vfs_bio_clrbuf(). However, allocbuf() never requests a prezeroed page. So, vfs_bio_clrbuf() never sees a prezeroed page. Reviewed by: tegge@	2004-05-06 05:03:23 +00:00
Alan Cox	2c840b1f65	- Substitute bdone() and bwait() from vfs_bio.c for swap_pager_putpages()'s buffer completion code. Note: the only difference between swp_pager_sync_iodone() and bdone(), aside from the locking in the latter, was the unnecessary clearing of B_ASYNC. - Remove an unnecessary pmap_page_protect() from swp_pager_async_iodone(). Reviewed by: tegge	2004-02-23 03:15:13 +00:00
Poul-Henning Kamp	d2bae332d6	Remove the absolute count g_access_abs() function since experience has shown that it is not useful. Rename the relative count g_access_rel() function to g_access(), only the name has changed. Change all g_access_rel() calls in our CVS tree to call g_access() instead. Add an #ifndef BURN_BRIDGES #define of g_access_rel() for source code compatibility.	2004-02-12 22:42:11 +00:00
Alan Cox	c5aebf380c	swp_pager_async_iodone() no longer requires Giant. Modify bufdone() and swapgeom_done() to perform swp_pager_async_iodone() without Giant. Reviewed by: tegge	2004-02-07 08:54:50 +00:00
Poul-Henning Kamp	3e5b686160	Check error return from g_clone_bio(). (netchild@) Add XXX comment about why this is still not optimal. (phk@) Submitted by: netchild@	2004-02-02 13:08:03 +00:00
Alan Cox	7dea2c2e3b	1. Statically initialize swap_pager_full and swap_pager_almost_full to the full state. (When swap is added their state will change appropriately.) 2. Set swap_pager_full and swap_pager_almost_full to the full state when the last swap device is removed. Combined these changes eliminate nonsense messages from the kernel on swap- less machines. Item 2 submitted by: Divacky Roman <xdivac02@stud.fit.vutbr.cz> Prodding by: phk	2004-01-24 21:31:06 +00:00
Alan Cox	2f7af3db57	Simplify the various pager allocation routines by computing the desired object size once and assigning that value to a local variable.	2004-01-04 20:55:15 +00:00
Alan Cox	e793b7797d	Reduce the scope of Giant in swap_pager_alloc().	2004-01-03 20:02:17 +00:00
Alan Cox	bd228075c7	Remove swap_pager_un_object_list; it is unused.	2003-12-29 04:21:44 +00:00
Alan Cox	c7c8dd7e80	- Modify swap_pager_copy() and its callers such that the source and destination objects are locked on entry and exit. Add comments to the callers noting that the locks can be released by swap_pager_copy(). - Remove several instances of GIANT_REQUIRED.	2003-11-01 08:57:26 +00:00
Alan Cox	2928cef7e1	- Synchronize access to the swdevt's sw_flags with sw_dev_mtx. - Remove several instances of GIANT_REQUIRED.	2003-10-31 05:18:45 +00:00
Alan Cox	7645e88596	- Synchronize access to the swdevt's sw_blist with sw_dev_mtx. - Remove several instances of GIANT_REQUIRED.	2003-10-30 09:12:43 +00:00
Alan Cox	d05bc12976	- Synchronize access to swdevhd using sw_dev_mtx. - Use swp_sizecheck() rather than assignment to swap_pager_full in swaponsomething().	2003-10-30 07:11:06 +00:00
Alan Cox	0676a140b2	- Synchronize updates to nswapdev using sw_dev_mtx.	2003-10-29 07:51:41 +00:00
Alan Cox	2d9974c1e8	- Avoid a race in swaponsomething(): Calculate the new swdevt's first and end swblk and insert this new swdevt into the list of swap devices in the same critical section.	2003-10-29 05:42:28 +00:00
Alan Cox	d536c58f53	- Complete the synchronization of accesses to the swblock hash table.	2003-10-27 05:58:15 +00:00
Alan Cox	7827d9b0fe	- Introduce and use a mutex synchronizing access to the swblock hash table.	2003-10-26 19:55:35 +00:00
Alan Cox	ee3dc7d7fe	- Add some of the required vm object locking, including assertions where the vm object lock is required and already held.	2003-10-25 23:42:17 +00:00
Alan Cox	2e3b314d3a	- Push down Giant from vm_pageout() to vm_pageout_scan(), freeing vm_pageout_page_stats() from Giant. - Modify vm_pager_put_pages() and vm_pager_page_unswapped() to expect the vm object to be locked on entry. (All of the pager routines now expect this.)	2003-10-24 06:43:04 +00:00
Poul-Henning Kamp	2c18019f14	DuH! bp->b_iooffset (the spot on the disk), not bp->b_offset (the offset in the file)	2003-10-18 14:10:28 +00:00
Poul-Henning Kamp	9fbf91c0dd	Initialize bp->b_offset before calling VOP_[SPEC]STRATEGY(). Remove stale comment about B_PHYS.	2003-10-18 11:11:05 +00:00
Poul-Henning Kamp	afeb65e61d	Don't open with exclusive bit, swapon(8) wants to trash our swapdev. Add XXX comment with a rating of this concept.	2003-09-02 05:53:44 +00:00
Poul-Henning Kamp	dee34ca4fc	Add a close() method to a swapdev. Add a GEOM based backend. Remove the device/VOP_SPECSTRATEGY() based backend.	2003-08-30 16:44:26 +00:00
Poul-Henning Kamp	20da9c2eaf	Protect the swapdevice tailq with a mutex. Store the udev_t we will report to userland in the swdevt.	2003-08-30 16:10:28 +00:00
Poul-Henning Kamp	59efee01a3	Continue the objectification of the swapdev backends: Remove the vnode and dev_t fields and replace them with a void *. Introduce separate strategy functions for devices and regular (NFS) vnodes. For devices we don't need the vnode v_numoutput stuff. Add a generic swaponsomething() function to add a swapdevice and split the remainder of swaponvp() into swaponvp() and swapondev() which calls this backend.	2003-08-30 11:33:25 +00:00
Poul-Henning Kamp	4b03903a46	Make the strategy function a method of the individual swapdev.	2003-08-30 09:42:00 +00:00
Poul-Henning Kamp	2f249180f5	Consistent use modern function definitions	2003-08-30 08:32:42 +00:00
Poul-Henning Kamp	395714feb7	Eliminate unnecessary udev_t variable: we can derive it from the dev_t when we need it.	2003-08-15 13:14:25 +00:00
Poul-Henning Kamp	89dc784fa3	Make swaponvp() static to the swap_pager.	2003-08-15 12:04:29 +00:00
Poul-Henning Kamp	ef3c5abdba	Make the first two pages magic to protect the BSD labels rather than only one.	2003-08-06 14:13:38 +00:00
Poul-Henning Kamp	751221fd32	Staticize swap_pager_putpages() Eliminate a lot of checkes to make sure requests are not cross-device which is unnecessary with the new layout. We know a sequential request cannot possibly be cross-device because there is a reserved page between the devices. Remove a couple of comments which no longer are relevant.	2003-08-06 12:08:27 +00:00
Poul-Henning Kamp	5e04322a6e	Explicitly set B_PAGING	2003-08-06 09:22:47 +00:00
Poul-Henning Kamp	c37a77ee86	Rip out the totally bogos vnode swapdev_vp with extreeme prejudice. Don't mark buffers with B_KEEPGIANT, we don't drop giant in strategy at this point in time.	2003-08-06 06:53:31 +00:00
Poul-Henning Kamp	e04e4bacf6	Use sparse struct initialization for struct pagerops. Mark our buffers B_KEEPGIANT before sending them downstream. Remove swap_pager_strategy implementation.	2003-08-05 06:54:56 +00:00
Poul-Henning Kamp	665c0caf03	Put an uncovered page between the swap devices, that way we can be sure to not get any cross-device I/O requests. (The unallocated first page protecting BSD labels already gave us this, but that hack may go away at some point in time). Remove the check for cross-device I/O requests in swap_pager_strategy. Move the repeated statistics updating into flushchainbuf().	2003-08-04 08:22:49 +00:00
Poul-Henning Kamp	12692209a6	Name swap_pager_find_dev() more correctly swp_pager_finde_dev(). Use ->bio_children to count child buffers, rather than abuse the bio_caller1 pointer. Expand the relevant bits of waitchainbuf() inline, this clarifies the code a little bit.	2003-08-03 21:22:42 +00:00
Poul-Henning Kamp	5ff0108d21	I accidentally hit undo before committing, fix the resulting off-by-one.	2003-08-03 14:53:52 +00:00
Poul-Henning Kamp	8f60c087e6	Change the layout policy of the swap_pager from a hardcoded width striping to a per device round-robin algorithm. Because of the policy of not attempting to retain previous swap allocation on page-out, this means that a newly added swap device almost instantly takes its 1/N share of the I/O load but it takes somewhat longer for it to assume it's 1/N share of the pages if there is plenty of space on the other devices. Change the 8G total swapspace limitation to 8G per device instead by using a per device blist rather than one global blist. This reduces the memory footprint by 75% (typically a couple hundred kilobytes) for the common case with one swapdevice but NSWAPDEV=4. Remove the compile time constant limit of number of swap devices, there is no limit now. Instead of a fixed size array, store the per swapdev structure in a TAILQ. Total swap space is still addressed by a 32 bit page number and therefore the upper limit is now 2^42 bytes = 16TB (for i386). We still do not allocate the first page of each device in order to give some amount of protection to any bsdlabel at the start of the device. A new device is appended after the existing devices in the swap space, no attempt is made to fill in holes left behind by swapoff (this can trivially be changed should it ever become a problem). The sysctl vm.nswapdev now reflects the number of currently configured swap devices. Rename vm_swap_size to swap_pager_avail for consistency with other exported names. Change argument type for vm_proc_swapin_all() and swap_pager_isswapped() to be a struct swdevt pointer rather than an index. Not changed: we are still using blists to manage the free space, but since the swapspace is no longer fragmented by the striping different resource managers might fare better.	2003-08-03 13:35:31 +00:00
Poul-Henning Kamp	8d677ef93f	Remove unused stuff. Move used stuff to swap_pager.c where it belongs. This file no longer exports anything to userland.	2003-07-31 22:19:28 +00:00
Poul-Henning Kamp	a8d43c90af	Add a "int fd" argument to VOP_OPEN() which in the future will contain the filedescriptor number on opens from userland. The index is used rather than a "struct file " since it conveys a bit more information, which may be useful to in particular fdescfs and /dev/fd/ For now pass -1 all over the place.	2003-07-26 07:32:23 +00:00
Poul-Henning Kamp	a5edd34afe	Remove all but one of the inlines here, this reduces the code size by 2032 bytes and has no measurable impact on performance.	2003-07-22 20:54:26 +00:00
Peter Wemm	da5fd14534	swp_pager_hash() was called before it was instantiated inline. This made gcc (quite rightly) unhappy. Move it earlier.	2003-07-22 06:55:48 +00:00
Poul-Henning Kamp	85fdafb98d	Fix a printf format warning I introduced. Use the macro max number of swap devices rather than cache the constant in a variable. Avoid a (now) pointless variable.	2003-07-18 22:11:17 +00:00
Poul-Henning Kamp	d3dd89ab11	If a proposed swap device exceeds the 8G artificial limit which out radix-tree code imposes, truncate the device instead of rejecting it.	2003-07-18 11:01:23 +00:00
Poul-Henning Kamp	ec38b344cb	Move the implementation of the vmspace_swap_count() (used only in the "toss the largest process" emergency handling) from vm_map.c to swap_pager.c. The quantity calculated depends strongly on the internals of the swap_pager and by moving it, we no longer need to expose the internal metrics of the swap_pager to the world.	2003-07-18 10:47:58 +00:00

1 2 3 4 5 ...

254 Commits