freebsd-nq

Author	SHA1	Message	Date
Jake Burkholder	9bc8610af5	Don't demap the requested page from the tlb in pmap_kenter or pmap_kremove, even on the local cpu. These are no longer used unsafely in MI code, and the MD code has been adjusted to compensate.	2002-03-17 01:53:51 +00:00
Jake Burkholder	bfd501b637	Fix a problem where kernel text could become unmapped when clearing out all the user mappings from the tlb due to the context numbers rolling over. The store to the internal mmu register must be followed by a membar #Sync before much else happens to "avoid data corruption", so we use special inlines which both disable interrupts and ensure that the compiler will not insert extra instructions between the two. Also, load the tte tag and check if the context is nucleus context, rather than relying on the priviledged bit which doesn't actually serve any purpose in our design, and check the lock bit too for sanity.	2002-03-17 01:51:32 +00:00
Jake Burkholder	2f3e2b8795	Use the tlb data access register to map the kernel tsb, rather than the data in register. The latter uses the random replacment algorithm to pick the slot, we want a specific slot.	2002-03-17 01:45:29 +00:00
Dag-Erling Smørgrav	a2e0658045	Move the definition of PT_[GS]ET{,DB,FP}REGS from the MD ptrace.h to the MI ptrace.h, since all platforms define them. Keep the MD ptrace.h around for FIX_SSTEP (which is currently only needed on Alpha).	2002-03-16 00:25:53 +00:00
Jake Burkholder	9ad1d32f3d	Fix ifdef LOCORE protection.	2002-03-13 06:04:36 +00:00
Jake Burkholder	76cf1369d7	Add a DEBUGGER_ON_POWERFAIL option. This makes the power button on ultra 10s work like an NMI button.	2002-03-13 05:58:45 +00:00
Jake Burkholder	064d9e8af4	Fix braino.	2002-03-13 05:54:00 +00:00
Jake Burkholder	5ce88ec47e	Add support for starting and stopping cpus with ipis. Stop the other cpus when shutting down or entering the debugger. Submitted by: tmm	2002-03-13 04:59:01 +00:00
Jake Burkholder	63a33ce158	Use intr_disable/intr_restore instead of doing it manually. Submitted by: tmm	2002-03-13 04:43:45 +00:00
Jake Burkholder	4968137eb9	Add support for driving the clocks on secondary cpus. Submitted by: tmm	2002-03-13 04:38:33 +00:00
Jake Burkholder	8f5eafcdd1	Fix a bug where the wrong number of windows were copied for a failed fill on return to user mode. We may not have frame pointers setup for more than 1 on return from exec.	2002-03-13 04:02:27 +00:00
Jake Burkholder	37453e292e	White space.	2002-03-13 03:55:28 +00:00
Jake Burkholder	1bc589c745	Make IPI_WAIT use a bit mask of the cpus that a pmap is active on and only wait for those cpus, instead of all of them by using a count. Oops. Make the pointer to the mask that the primary cpu spins on volatile, so gcc doesn't optimize out an important load. Oops again. Activate tlb shootdown ipi synchronization now that it works. We have all involved cpus wait until all the others are done. This may not be necessary, it is mostly for sanity. Make the trigger level interrupt ipi handler work. Submitted by: tmm	2002-03-13 03:43:00 +00:00
Jake Burkholder	453e54056e	Add an ATOMIC_CLEAR_INT macro. Submitted by: tmm	2002-03-13 03:28:47 +00:00
Thomas Moestl	9390de81a1	Fix the type of some constants, and make some macros safer by casting the argument.	2002-03-11 03:04:28 +00:00
Thomas Moestl	25aab485b6	Add convenience macros to extract the cc0 and cc1 from format 2 and 3 instructions.	2002-03-11 03:03:35 +00:00
Jake Burkholder	a46877abfc	Increase VM_KMEM_SIZE to 16 megs from 12. Define VM_KMEM_SIZE_SCALE so that the number of physical pages per KVA page allocated scales properly with memory size. This fixes problems with kmem_map being too small. Noticed by: mike, wollman Submitted by: tmm	2002-03-09 23:35:50 +00:00
Thomas Moestl	0c530eb321	Add a driver for the mem and kmem devices, based off the i386 version.	2002-03-09 22:33:16 +00:00
Thomas Moestl	d4f1dcab55	Set the interrupt map type accordingly if we need to fall back to using the PCI bus interrupt map.	2002-03-09 22:02:02 +00:00
Thomas Moestl	db20923f96	Fix a warning by adding a missing include.	2002-03-09 22:00:30 +00:00
Mike Barcroft	d846855da8	o Don't require long long support in bswap64() functions. o In i386's <machine/endian.h>, macros have some advantages over inlines, so change some inlines to macros. o In i386's <machine/endian.h>, ungarbage collect word_swap_int() (previously __uint16_swap_uint32), it has some uses on i386's with PDP endianness. Submitted by: bde o Move a comment up in <machine/endian.h> that was accidentially moved down a few revisions ago. o Reenable userland's use of optimized inline-asm versions of byteorder(3) functions. o Fix ordering of prototypes vs. redefinition of byteorder(3) functions, so that the non-GCC (libc asm) case has proper prototypes. o Add proper prototypes for byteorder(3) functions in <sys/param.h>. o Prevent redundant duplicate prototypes by making use of the _BYTEORDER_PROTOTYPED define. o Move the bswap16(), bswap32(), bswap64() C functions into MD space for platforms in which asm versions don't exist. This significantly reduces the complexity of some things at the cost of duplicate code. Reviewed by: bde	2002-03-09 21:02:16 +00:00
Jake Burkholder	4f91e3efb2	Implement delivery of tlb shootdown ipis. This is currently more fine grained than the other implementations; we have complete control over the tlb, so we only demap specific pages. We take advantage of the ranged tlb flush api to send one ipi for a range of pages, and due to the pm_active optimization we rarely send ipis for demaps from user pmaps. Remove now unused routines to load the tlb; this is only done once outside of the tlb fault handlers. Minor cleanups to the smp startup code. This boots multi user with both cpus active on a dual ultra 60 and on a dual ultra 2.	2002-03-07 06:01:40 +00:00
Jake Burkholder	39028e8396	Modify the tlb demap API to take a pmap instead of a tlb context number. Due to allocating tlb contexts on the fly, we only ever need to demap the primary context, non-primary contexts have already been implicitly flushed by context switching. All we really need to tell is if its a kernel demap or not, and its easier just to compare against the kernel_pmap which is a constant.	2002-03-07 05:25:15 +00:00
Jake Burkholder	bc9b764621	Implement kthread context stealing. This is a bit of a misnomer because the context is not actually stolen, as it would be for i386. Instead of deactivating a user vmspace immediately when switching out, and recycling its tlb context, wait until the next context switch to a different user vmspace. In this way we can switch from a user process to any number of kernel threads and back to the same user process again, without losing any of its mappings in the tlb that would not already be knocked by the automatic replacement algorithm. This is not expected to have a measurable performance improvement on the machines we currently run on, but it sounds cool and makes the sparc64 port SMPng buzz word compliant.	2002-03-07 05:15:43 +00:00
Jake Burkholder	eb5b9c0be3	Add support for starting secondary cpus in kernel, as opposed to relying on the loader to do it. Improve smp startup code to be less racy and to defer certain things until the right time. This almost boots single user on my dual ultra 60, it is still very fragile: SMP: AP CPU #1 Launched! Enter full pathname of shell or RETURN for /bin/sh: # ls Debugger("trapsig") Stopped at Debugger+0x1c: ta %xcc, 1 db> heh No such command db>	2002-03-04 07:12:36 +00:00
Jake Burkholder	907164660a	Dig the information about which tlb slots were used to map the kernel out of the metadata passed by the loader.	2002-03-04 07:07:10 +00:00
Jake Burkholder	5573db3f0b	Allocate tlb contexts on the fly in cpu_switch, instead of statically 1 to 1 with pmaps. When the context numbers wrap around we flush all user mappings from the tlb. This makes use of the array indexed by cpuid to allow a pmap to have a different context number on a different cpu. If the context numbers are then divided evenly among cpus such that none are shared, we can avoid sending tlb shootdown ipis in an smp system for non-shared pmaps. This also removes a limit of 8192 processes (pmaps) that could be active at any given time due to running out of tlb contexts. Inspired by: the brown book Crucial bugfix from: tmm	2002-03-04 05:20:29 +00:00
Jake Burkholder	34aef253aa	Fix obscure problems with vfork where part of the parent's stack could be clobbered by the child. This is more complicated than usual because the window that could get clobbered is pushed in kernel mode, so a lot of registers would have to be saved in other registers in userland and we don't have enough. What we do have is space in the pcb to temporarily store user windows that were spilled in kernel mode, but could not be immediately stored to the user stack. So we copy in the parent's topmost window and save it in the pcb, and arrange for it to be copied back out when the child is done frobbing the stack. Reviewed by: tmm	2002-03-04 05:07:22 +00:00
Jake Burkholder	e45a21d07a	We don't need KTR_COMPILE in assym.s, its already in opt_global.h. Add assyms for more ktr trace classes.	2002-03-01 16:22:06 +00:00
Jake Burkholder	b8c926a9ab	Use a better trace class for ktr traces in the tlb fault handlers, which are rather loud.	2002-03-01 16:17:50 +00:00
Andrew R. Reiter	66c862bc1b	- Move a comment from being on the same line as a #ifdef to the line following it. This should have gone in the previous commit, but misviewed Bruce's patch. Requested by: bde	2002-02-28 21:52:08 +00:00
Andrew R. Reiter	216ae18217	- Fix panic() message and a couple style nits that snuck in from the recent diagnostics commit (rev. 1.84).	2002-02-28 08:28:14 +00:00
Mike Silbersack	f4e18c9afd	Fix a minor swap leak. Previously, the UPAGES/KSTACK area of processes/threads would leak memory at the time that a previously swapped process was terminated. Lukcily, the leak was only 12K/proc, so it was unlikely to be a major problem unless you had an undersized swap partition. Submitted by: dillon Reviewed by: silby MFC after: 1 week	2002-02-28 07:41:12 +00:00
Mike Silbersack	7f3a40933b	Fix a horribly suboptimal algorithm in the vm_daemon. In order to determine what to page out, the vm_daemon checks reference bits on all pages belonging to all processes. Unfortunately, the algorithm used reacted badly with shared pages; each shared page would be checked once per process sharing it; this caused an O(N^2) growth of tlb invalidations. The algorithm has been changed so that each page will be checked only 16 times. Prior to this change, a fork/sleepbomb of 1300 processes could cause the vm_daemon to take over 60 seconds to complete, effectively freezing the system for that time period. With this change in place, the vm_daemon completes in less than a second. Any system with hundreds of processes sharing pages should benefit from this change. Note that the vm_daemon is only run when the system is under extreme memory pressure. It is likely that many people with loaded systems saw no symptoms of this problem until they reached the point where swapping began. Special thanks go to dillon, peter, and Chuck Cranor, who helped me get up to speed with vm internals. PR: 33542, 20393 Reviewed by: dillon MFC after: 1 week	2002-02-27 18:03:02 +00:00
Thomas Moestl	90ce56c287	Add the following functions/macros to support byte order conversions and device drivers for bus system with other endinesses than the CPU (using interfaces compatible to NetBSD): - bwap16() and bswap32(). These have optimized implementations on some architectures; for those that don't, there exist generic implementations. - macros to convert from a certain byte order to host byte order and vice versa, using a naming scheme like le16toh(), htole16(). These are implemented using the bswap functions. - stream bus space access functions, which do not perform a byte order conversion (while the normal access functions would if the bus endianess differs from the CPU endianess). htons(), htonl(), ntohs() and ntohl() are implemented using the new functions above for kernel usage. None of the above interfaces is currently exported to user land. Make use of the new functions in a few places where local implementations of the same functionality existed. Reviewed by: mike, bde Tested on alpha by: mike	2002-02-27 17:16:18 +00:00
Jake Burkholder	df38f87be1	Minimal testing has shown that a 4 page tsb is a nice sweet spot for current work loads. It tapers off after that as gcc's working set generally just fits. compiling bin/csh: TSB_PAGES = 2 213.33 real 77.59 user 110.01 sys TSB_PAGES = 4 116.43 real 75.78 user 19.16 sys TSB_PAGES = 8 119.27 real 76.38 user 18.12 sys Testing by: tmm	2002-02-27 06:18:02 +00:00
Jake Burkholder	95a44511f3	Parameterize the number of pages to allocate for the per-cpu area on PCPU_PAGES.	2002-02-27 06:08:13 +00:00
Jake Burkholder	62ad058292	Make cpu_identify take the value of the ver register and cpuid as arguments so we can print nice things about non-current cpus.	2002-02-27 06:05:50 +00:00
Jake Burkholder	ad414cb452	Minor cleanup.	2002-02-27 00:31:31 +00:00
Jake Burkholder	d689965cb4	Wrap long lines.	2002-02-27 00:28:35 +00:00
Jake Burkholder	9d16662aa3	Use pcpu.pc_cpumask instead of computing 1 << cpuid.	2002-02-27 00:27:05 +00:00
Jake Burkholder	5d60dc233a	Add a macro for shift of an integer (1 << shift == sizeof). Move the pointer define to live alongside it. For kicks assert at compile time that they are correct. Use these instead of magic numbers.	2002-02-27 00:21:04 +00:00
Jake Burkholder	7a8ee66881	Wrap long lines.	2002-02-27 00:03:01 +00:00
David E. O'Brien	db3442f54d	Define basic macros required by GDB.	2002-02-26 21:49:46 +00:00
Jake Burkholder	c7f9e1fdbf	Apparently gcc3.1 is now using deprcated v8 instructions in v9 code due to them being faster in certain cases. Therefore we need to save and restore the v8 %y register around traps in kernel mode as well as traps in usermode. Tested by: obrien, tmm	2002-02-26 17:09:24 +00:00
Jake Burkholder	de3fee8992	Convert pmap.pm_context to an array of contexts indexed by cpuid. This doesn't make sense for SMP right now, but it is a means to an end.	2002-02-26 06:57:30 +00:00
Jake Burkholder	200e6309c9	Pu back a call to pmap_context_destroy which was accidentily removed in the previous commit. Spotted by: tmm	2002-02-26 06:39:38 +00:00
Jake Burkholder	6e5f3d0f0f	Allow the user tsb to span multiple pages. Make the default 2 pages for now until we do some testing to see what's best. This gives a massive reduction in system time for processes with a relatively large working set. The size of the tsb directly affects the rss size that a user process can keep mapped. When it starts to get full replacements occur and the process takes a lot of soft vm faults. Increasing the default from 1 page to 2 gives the following before and after numbers for compiling vfs_bio.c: before: 14.27 real 6.56 user 5.69 sys after: 8.57 real 6.11 user 1.62 sys This should make self hosted builds more tolerable.	2002-02-26 02:37:43 +00:00
Jake Burkholder	07d99740b6	Remove code to lock the user tsb into the tlb. We can handle faults on it now, as we do for normal wired kernel memory.	2002-02-25 22:58:41 +00:00
David E. O'Brien	65955377e3	I was able to boot this kernel using the latest WIP kernel sources. I don't believe anyone is quite using the sparc64 kernel sources in CVS yet -- things aren't just quite ready (but almost). So this commit should be OK to make.	2002-02-25 22:13:44 +00:00

1 2 3 4 5 ...

385 Commits