minix

Author	SHA1	Message	Date
Cristiano Giuffrida	cb176df60f	New RS and new signal handling for system processes. UPDATING INFO: 20100317: /usr/src/etc/system.conf updated to ignore default kernel calls: copy it (or merge it) to /etc/system.conf. The hello driver (/dev/hello) added to the distribution: # cd /usr/src/commands/scripts && make clean install # cd /dev && MAKEDEV hello KERNEL CHANGES: - Generic signal handling support. The kernel no longer assumes PM as a signal manager for every process. The signal manager of a given process can now be specified in its privilege slot. When a signal has to be delivered, the kernel performs the lookup and forwards the signal to the appropriate signal manager. PM is the default signal manager for user processes, RS is the default signal manager for system processes. To enable ptrace()ing for system processes, it is sufficient to change the default signal manager to PM. This will temporarily disable crash recovery, though. - sys_exit() is now split into sys_exit() (i.e. exit() for system processes, which generates a self-termination signal), and sys_clear() (i.e. used by PM to ask the kernel to clear a process slot when a process exits). - Added a new kernel call (i.e. sys_update()) to swap two process slots and implement live update. PM CHANGES: - Posix signal handling is no longer allowed for system processes. System signals are split into two fixed categories: termination and non-termination signals. When a non-termination signaled is processed, PM transforms the signal into an IPC message and delivers the message to the system process. When a termination signal is processed, PM terminates the process. - PM no longer assumes itself as the signal manager for system processes. It now makes sure that every system signal goes through the kernel before being actually processes. The kernel will then dispatch the signal to the appropriate signal manager which may or may not be PM. SYSLIB CHANGES: - Simplified SEF init and LU callbacks. - Added additional predefined SEF callbacks to debug crash recovery and live update. - Fixed a temporary ack in the SEF init protocol. SEF init reply is now completely synchronous. - Added SEF signal event type to provide a uniform interface for system processes to deal with signals. A sef_cb_signal_handler() callback is available for system processes to handle every received signal. A sef_cb_signal_manager() callback is used by signal managers to process system signals on behalf of the kernel. - Fixed a few bugs with memory mapping and DS. VM CHANGES: - Page faults and memory requests coming from the kernel are now implemented using signals. - Added a new VM call to swap two process slots and implement live update. - The call is used by RS at update time and in turn invokes the kernel call sys_update(). RS CHANGES: - RS has been reworked with a better functional decomposition. - Better kernel call masks. com.h now defines the set of very basic kernel calls every system service is allowed to use. This makes system.conf simpler and easier to maintain. In addition, this guarantees a higher level of isolation for system libraries that use one or more kernel calls internally (e.g. printf). - RS is the default signal manager for system processes. By default, RS intercepts every signal delivered to every system process. This makes crash recovery possible before bringing PM and friends in the loop. - RS now supports fast rollback when something goes wrong while initializing the new version during a live update. - Live update is now implemented by keeping the two versions side-by-side and swapping the process slots when the old version is ready to update. - Crash recovery is now implemented by keeping the two versions side-by-side and cleaning up the old version only when the recovery process is complete. DS CHANGES: - Fixed a bug when the process doing ds_publish() or ds_delete() is not known by DS. - Fixed the completely broken support for strings. String publishing is now implemented in the system library and simply wraps publishing of memory ranges. Ideally, we should adopt a similar approach for other data types as well. - Test suite fixed. DRIVER CHANGES: - The hello driver has been added to the Minix distribution to demonstrate basic live update and crash recovery functionalities. - Other drivers have been adapted to conform the new SEF interface.	2010-03-17 01:15:29 +00:00
Ben Gras	0937d6c367	re-establish kernel assert()s. use the regular <assert.h> assert() instead of vmassert() in kernel. throw out some #if 0 code. fix a few assert() conditions. enable by default.	2010-03-10 13:00:05 +00:00
Arun Thomas	1f9ce647cf	Move archtypes.h, fpu.h, and stackframe.h Move archtypes.h to include/ dir, since several servers require it. Move fpu.h and stackframe.h to arch-specific header directory. Make source files and makefiles aware of the new header locations.	2010-03-09 09:41:14 +00:00
Arun Thomas	2a8fabf4ad	Include directory reorg and makefile updates. -Convert the include directory over to using bsdmake syntax -Update/add mkfiles -Modify install(1) so that it can create symlinks -Update makefiles to use new install(1) options -Rename /usr/include/ibm to /usr/include/i386 -Create /usr/include/machine symlink to arch header files -Move vm_i386.h to its new home in the /usr/include/i386 -Update source files to #include the header files at their new homes. -Add new gnu-includes target for building GCC headers	2010-03-08 11:04:59 +00:00
Tomas Hruby	ecf1a36d48	Fix for FPU broken by r6131 - cycles accounting must be called earlier, firstly not to clobber the %ebx register, secondly to be correctly called in both branches.	2010-03-05 22:23:03 +00:00
Ben Gras	35a108b911	panic() cleanup. this change - makes panic() variadic, doing full printf() formatting - no more NO_NUM, and no more separate printf() statements needed to print extra info (or something in hex) before panicing - unifies panic() - same panic() name and usage for everyone - vm, kernel and rest have different names/syntax currently in order to implement their own luxuries, but no longer - throws out the 1st argument, to make source less noisy. the panic() in syslib retrieves the server name from the kernel so it should be clear enough who is panicing; e.g. panic("sigaction failed: %d", errno); looks like: at_wini(73130): panic: sigaction failed: 0 syslib:panic.c: stacktrace: 0x74dc 0x2025 0x100a - throws out report() - printf() is more convenient and powerful - harmonizes/fixes the use of panic() - there were a few places that used printf-style formatting (didn't work) and newlines (messes up the formatting) in panic() - throws out a few per-server panic() functions - cleans up a tie-in of tty with panic() merging printf() and panic() statements to be done incrementally.	2010-03-05 15:05:11 +00:00
Ben Gras	e6cb76a2e2	no more kprintf - kernel uses libsys printf now, only kputc is special to the kernel.	2010-03-03 15:45:01 +00:00
Ben Gras	18924ea563	New P_BLOCKEDON for kernel - a macro that encodes the "who is this process waiting for" logic, which is duplicated a few times in the kernel. (For a new feature for top.) Introducing it and throwing out ESRCDIED and EDSTDIED (replaced by EDEADSRCDST - so we don't have to care which part of the blocking is failing in system.c) simplifies some code in the kernel and callers that check for E{DEADSRCDST,ESRCDIED,EDSTDIED}, but don't care about the difference, a fair bit, and more significantly doesn't duplicate the 'blocked-on' logic.	2010-03-03 15:32:26 +00:00
Arun Thomas	cbd276e4ce	Convert library asm files to GAS syntax	2010-03-03 14:27:30 +00:00
Kees van Reeuwijk	bf7397b64e	More correctly use cp_grant_id_t. More correctly use vir_bytes. More correctly use endpoint_t.	2010-03-02 23:12:13 +00:00
Kees van Reeuwijk	1ba0936619	Fix some uses of uninitialized variables.	2010-02-19 10:41:02 +00:00
Kees van Reeuwijk	97c169b93a	Remove some unused #include. Remove some unused variables and computations on them.	2010-02-17 20:24:42 +00:00
Arun Thomas	b706112487	Incorporate bsdmake into buildsystem and reorganize libs	2010-02-16 14:41:33 +00:00
David van Moolenbroek	e306663455	fix the somehow newly introduced warnings	2010-02-14 18:39:47 +00:00
Kees van Reeuwijk	df60646f98	Undo the use of #include <...> because it caused some errors.	2010-02-12 14:43:18 +00:00
Tomas Hruby	1b56fdb33c	Time accounting based on TSC - as thre are still KERNEL and IDLE entries, time accounting for kernel and idle time works the same as for any other process - everytime we stop accounting for the currently running process, kernel or idle, we read the TSC counter and increment the p_cycles entry. - the process cycles inherently include some of the kernel cycles as we can stop accounting for the process only after we save its context and we start accounting just before we restore its context - this assumes that the system does not scale the CPU frequency which will be true for ... long time ;-)	2010-02-10 15:36:54 +00:00
Tomas Hruby	c6fec6866f	No locking in kernel code - No locking in RTS_(UN)SET macros - No lock_notify() - Removed unused lock_send() - No lock/unlock macros anymore	2010-02-09 15:26:58 +00:00
Tomas Hruby	391fd926ff	TASK_PRIVILEGE and level0() removed - there are no tasks running, we don't need TASK_PRIVILEGE priviledge anymore - as there is no ring 1 anymore, there is no need for level0() to call sensitive code from ring 1 in ring 0 - 286 related macros removed as clean up	2010-02-09 15:23:31 +00:00
Tomas Hruby	728f0f0c49	Removal of the system task * Userspace change to use the new kernel calls - _taskcall(SYSTASK...) changed to _kernel_call(...) - int 32 reused for the kernel calls - _do_kernel_call() to make the trap to kernel - kernel_call() to make the actuall kernel call from C using _do_kernel_call() - unlike ipc call the kernel call always succeeds as kernel is always available, however, kernel may return an error * Kernel side implementation of kernel calls - the SYSTEm task does not run, only the proc table entry is preserved - every data_copy(SYSTEM is no data_copy(KERNEL - "locking" is an empty operation now as everything runs in kernel - sys_task() is replaced by kernel_call() which copies the message into kernel, dispatches the call to its handler and finishes by either copying the results back to userspace (if need be) or by suspending the process because of VM - suspended processes are later made runnable once the memory issue is resolved, picked up by the scheduler and only at this time the call is resumed (in fact restarted) which does not need to copy the message from userspace as the message is already saved in the process structure. - no ned for the vmrestart queue, the scheduler will restart the system calls - no special case in do_vmctl(), all requests remove the RTS_VMREQUEST flag	2010-02-09 15:20:09 +00:00
Tomas Hruby	5e57818431	copy_msg_from_user() and copy_msg_to_user() - copies a mesage from/to userspace without need of translating addresses - the assumption is that the address space is installed, i.e. ldt and cr3 are loaded correctly - if a pagefault or a general protection occurs while copying from userland to kernel (or vice versa) and error is returned which gives the caller a chance to respond in a proper way - error happens _only_ because of a wrong user pointer if the function is used correctly - if the prerequisites of the function do no hold, the function will most likely fail as the user address becomes random	2010-02-09 15:15:45 +00:00
Tomas Hruby	ad9ba944d1	Early address space switch - switch_address_space() implements a switch of the user address space for the destination process - this makes memory of this process easily accessible, e.g. a pointer valid in the userspace can be used with a little complexity to access the process's memory - the switch does not happed only just before we return to userspace, however, it happens right after we know which process we are going to schedule. This happens before we start processing the misc flags of this process so its memory is available - if the process becomes not runnable while processing the mics flags we pick a new process and we switch the address space again which introduces possibly a little bit more overhead, however, it is hopefully hidden by reducing the overheads when we actually access the memory	2010-02-09 15:13:52 +00:00
Tomas Hruby	b14a86ca5c	Sys calls are called ipc calls now - the syscalls are pretty much just ipc calls, however, sendrec() is used to implement system task (sys) calls - sendrec() won't be used anymore for this, therefore ipc calls will become pure ipc calls	2010-02-09 15:13:07 +00:00
Tomas Hruby	cca24d06d8	This patch removes the global variables who_p and who_e from the kernel (sys task). The main reason is that these would have to become cpu local variables on SMP. Once the system task is not a task but a genuine part of the kernel there is even less reason to have these extra variables as proc_ptr will already contain all neccessary information. In addition converting who_e to the process pointer and back again all the time will be avoided. Although proc_ptr will contain all important information, accessing it as a cpu local variable will be fairly expensive, hence the value would be assigned to some on stack local variable. Therefore it is better to add the 'caller' argument to the syscall handlers to pass the value on stack anyway. It also clearly denotes on who's behalf is the syscall being executed. This patch also ANSIfies the syscall function headers. Last but not least, it also fixes a potential bug in virtual_copy_f() in case the check is disabled. So far the function in case of a failure could possible reuse an old who_p in case this function had not been called from the system task. virtual_copy_f() takes the caller as a parameter too. In case the checking is disabled, the caller must be NULL and non NULL if it is enabled as we must be able to suspend the caller.	2010-02-03 09:04:48 +00:00
Kees van Reeuwijk	477b616fe8	Fixed a number of complaints about missing return statements. Some cases were fixed by declaring the function void, others were fixed by adding a return <value> statement, thereby avoiding potentially incorrect behavior (usually in error handling). Some enum correctness in boot.c.	2010-01-28 13:17:07 +00:00
Kees van Reeuwijk	c8a11b5453	Fixed some type inconsistencies in the kernel.	2010-01-26 12:26:06 +00:00
Kees van Reeuwijk	b67f788eea	Removed a number of useless #includes	2010-01-26 10:59:01 +00:00
Kees van Reeuwijk	a701e290f7	Removed unused symbols. Made some functions PRIVATE, including ones that aren't used anywhere.	2010-01-25 18:13:48 +00:00
Kees van Reeuwijk	a7cee5bec4	Removed unused symbols. Minor cleanups.	2010-01-22 22:01:08 +00:00
Kees van Reeuwijk	f30c82b430	Restored idt_reload() prototype.	2010-01-21 11:40:22 +00:00
Kees van Reeuwijk	d6383bef47	Removed some unused tests.	2010-01-20 17:55:14 +00:00
Tomas Hruby	7d51b0cce1	Fixed warnings in watchdog.c	2010-01-19 14:47:25 +00:00
Tomas Hruby	5efa92f754	NMI watchdog is an awesome feature for debugging locked up kernels. There is not that much use for it on a single CPU, however, deadlock between kernel and system task can be delected. Or a runaway loop. If a kernel gets locked up the timer interrupts don't occure (as all interrupts are disabled in kernel mode). The only chance is to interrupt the kernel by a non-maskable interrupt. This patch generates NMIs using performance counters. It uses the most widely available performace counters. As the performance counters are highly model-specific this patch is not guaranteed to work on every machine. Unfortunately this is also true for KVM :-/ On the other hand adding this feature for other models is not extremely difficult and the framework makes it hopefully easy enough. Depending on the frequency of the CPU an NMI is generated at most about every 0.5s If the cpu's speed is less then 2Ghz it is generated at most every 1s. In general an NMI is generated much less often as the performance counter counts down only if the cpu is not idle. Therefore the overhead of this feature is fairly minimal even if the load is high. Uppon detecting that the kernel is locked up the kernel dumps the state of the kernel registers and panics. Local APIC must be enabled for the watchdog to work. The code is _always_ compiled in, however, it is only enabled if watchdog=<non-zero> is set in the boot monitor. One corner case is serial console debugging. As dumping a lot of stuff to the serial link may take a lot of time, the watchdog does not detect lockups during this time!!! as it would result in too many false positives. 10 nmi have to be handled before the lockup is detected. This means something between ~5s to 10s. Another corner case is that the watchdog is enabled only after the paging is enabled as it would be pure madness to try to get it right.	2010-01-16 20:53:55 +00:00
Cristiano Giuffrida	c5b309ff07	Merge of Wu's GSOC 09 branch (src.20090525.r4372.wu) Main changes: - COW optimization for safecopy. - safemap, a grant-based interface for sharing memory regions between processes. - Integration with safemap and complete rework of DS, supporting new data types natively (labels, memory ranges, memory mapped ranges). - For further information: http://wiki.minix3.org/en/SummerOfCode2009/MemoryGrants Additional changes not included in the original Wu's branch: - Fixed unhandled case in VM when using COW optimization for safecopy in case of a block that has already been shared as SMAP. - Better interface and naming scheme for sys_saferevmap and ds_retrieve_map calls. - Better input checking in syslib: check for page alignment when creating memory mapping grants. - DS notifies subscribers when an entry is deleted. - Documented the behavior of indirect grants in case of memory mapping. - Test suite in /usr/src/test/safeperf\|safecopy\|safemap\|ds/* reworked and extended. - Minor fixes and general cleanup. - TO-DO: Grant ids should be generated and managed the way endpoints are to make sure grant slots are never misreused.	2010-01-14 15:24:16 +00:00
Kees van Reeuwijk	da3b64d8bc	Fixed a bug in do_sdevio() that broke I/O size computations. Removed redundant size computations. Cleaned up code.	2010-01-14 14:51:23 +00:00
Tomas Hruby	98563a4afa	Killing Minix by typing Q on serial console - if debugging on serial console is enabled typing Q kills the system. It is handy if the system gets locked up and the timer interrupts still work. Good for remote debugging. - NOT_REACHABLE reintroduced and fixed. It should be used for marking code which is not reachable because the previous code _should_ not return. Such places are not always obvious	2010-01-14 09:46:16 +00:00
Tomas Hruby	8a2a4f97fc	Fixed redundant typecast in lapic write/read macros	2010-01-13 18:23:58 +00:00
Tomas Hruby	42c13951a7	APIC disabled if CPU lacks TSC - we cannot calibrate local APIC timer in such a case - fixes possible uninitialized variable problem during calibration if no TSC	2010-01-13 18:22:41 +00:00
Kees van Reeuwijk	ad4c0ff698	Fixed a bug in apic.c that broke lapic_stop_timer(). Fixed bugs in liveupdate.c that rendered load_state_info() meaningless. More informative error message in do_config() in service.c.	2010-01-13 14:44:19 +00:00
Erik van der Kouwe	38ed5b2685	Fix brackets in kernel/arch/i386/include/archconst.h	2010-01-06 08:46:33 +00:00
Kees van Reeuwijk	d8f3af3672	Fixed a typing bug. More explicit type conversion from virual to physical bytes. Bracket negative #defines for extra paranoia. Added a forgotten 'void' to a function.	2010-01-06 08:23:14 +00:00
David van Moolenbroek	bac0e91705	typo (Bug#376, reported by Kees van Reeuwijk)	2010-01-04 12:29:51 +00:00
David van Moolenbroek	ac9a5829a2	suppress kernel/VM memory debugging information	2009-12-29 21:35:12 +00:00
David van Moolenbroek	e423c86009	ptrace(2) modifications: - add T_GETRANGE/T_SETRANGE to get/set ranges of values - change EIO error code to EFAULT - move common-I&D text-to-data translation to umap_local	2009-12-29 21:32:15 +00:00
Tomas Hruby	51065a1b47	Cooments to warn not to use certains instructions - gas2ack cannot handle all variants of some instructions. Until this issues is addressed, this patch places a big warning where appropriate. This code is not supposed to change frequently.	2009-12-07 12:01:05 +00:00
David van Moolenbroek	fe982ca684	FPU: fix field names, compiler warning, long lines	2009-12-02 23:12:46 +00:00
Ben Gras	bd42705433	FPU context switching support by Evgeniy Ivanov.	2009-12-02 13:01:48 +00:00
David van Moolenbroek	fce9fd4b4e	Add 'getidle' CPU utilization measurement infrastructure	2009-12-02 11:52:26 +00:00
Tomas Hruby	8a44a44cb9	Local APIC - local APIC timer used as the source of time - PIC is still used as the hw interrupt controller as we don't have enough info without ACPI or MPS to set up IO APICs - remapping of APIC when switching paging on, uses the new mechanism to tell VM what phys areas to map in kernel's virtual space - one more step to SMP based on code by Arun C.	2009-11-16 21:41:44 +00:00
Tomas Hruby	9e62bd5241	.align replaced by .balign in mpx386.S	2009-11-13 09:30:45 +00:00
Tomas Hruby	ad4dcaab71	Idle task never runs - idle task becomes a pseudo task which is never scheduled. It is never put on any run queue and never enters userspace. An entry for this task still remains in the process table for time accounting - Instead of panicing if there is not process to schedule, pick_proc() returns NULL which is a signal to put the cpu in an idle state and set everything in such a way that after receiving and interrupt it looks like idle task was preempted - idle task is set non-preemptible to avoid handling in the timer interrupt code which make userspace scheduling simpler as idle task does not need to be handled as a special case.	2009-11-12 08:42:18 +00:00

1 2 3

104 commits