linux_dsm_epyc7002

mirror of https://github.com/AuxXxilium/linux_dsm_epyc7002.git synced 2024-12-03 07:56:42 +07:00

History

Paul E. McKenney 26861faf89 rcu: Protect __rcu_read_unlock() against scheduler-using irq handlers This commit ports commit #10f39bb1b2 (rcu: protect __rcu_read_unlock() against scheduler-using irq handlers) from TREE_PREEMPT_RCU to TINY_PREEMPT_RCU. The following is a corresponding port of that commit message. The addition of RCU read-side critical sections within runqueue and priority-inheritance critical sections introduced some deadlocks, for example, involving interrupts from __rcu_read_unlock() where the interrupt handlers call wake_up(). This situation can cause the instance of __rcu_read_unlock() invoked from interrupt to do some of the processing that would otherwise have been carried out by the task-level instance of __rcu_read_unlock(). When the interrupt-level instance of __rcu_read_unlock() is called with a scheduler lock held from interrupt-entry/exit situations where in_irq() returns false, deadlock can result. Of course, in a UP kernel, there are not really any deadlocks, but the upper-level critical section can still be be fatally confused by the lower-level critical section changing things out from under it. This commit resolves these deadlocks by using negative values of the per-task ->rcu_read_lock_nesting counter to indicate that an instance of __rcu_read_unlock() is in flight, which in turn prevents instances from interrupt handlers from doing any special processing. Note that nested rcu_read_lock()/rcu_read_unlock() pairs are still permitted, but they will never see ->rcu_read_lock_nesting go to zero, and will therefore never invoke rcu_read_unlock_special(), thus preventing them from seeing the RCU_READ_UNLOCK_BLOCKED bit should it be set in ->rcu_read_unlock_special. This patch also adds a check for ->rcu_read_unlock_special being negative in rcu_check_callbacks(), thus preventing the RCU_READ_UNLOCK_NEED_QS bit from being set should a scheduling-clock interrupt occur while __rcu_read_unlock() is exiting from an outermost RCU read-side critical section. Of course, __rcu_read_unlock() can be preempted during the time that ->rcu_read_lock_nesting is negative. This could result in the setting of the RCU_READ_UNLOCK_BLOCKED bit after __rcu_read_unlock() checks it, and would also result it this task being queued on the corresponding rcu_node structure's blkd_tasks list. Therefore, some later RCU read-side critical section would enter rcu_read_unlock_special() to clean up -- which could result in deadlock (OK, OK, fatal confusion) if that RCU read-side critical section happened to be in the scheduler where the runqueue or priority-inheritance locks were held. To prevent the possibility of fatal confusion that might result from preemption during the time that ->rcu_read_lock_nesting is negative, this commit also makes rcu_preempt_note_context_switch() check for negative ->rcu_read_lock_nesting, thus refraining from queuing the task (and from setting RCU_READ_UNLOCK_BLOCKED) if we are already exiting from the outermost RCU read-side critical section (in other words, we really are no longer actually in that RCU read-side critical section). In addition, rcu_preempt_note_context_switch() invokes rcu_read_unlock_special() to carry out the cleanup in this case, which clears out the ->rcu_read_unlock_special bits and dequeues the task (if necessary), in turn avoiding needless delay of the current RCU grace period and needless RCU priority boosting. It is still illegal to call rcu_read_unlock() while holding a scheduler lock if the prior RCU read-side critical section has ever had both preemption and irqs enabled. However, the common use case is legal, namely where then entire RCU read-side critical section executes with irqs disabled, for example, when the scheduler lock is held across the entire lifetime of the RCU read-side critical section. Signed-off-by: Paul E. McKenney <paul.mckenney@linaro.org> Signed-off-by: Paul E. McKenney <paulmck@linux.vnet.ibm.com>		2012-02-21 09:03:39 -08:00
..
debug	module: struct module_ref should contains long fields	2012-01-13 09:32:14 +10:30
events	perf: Fix double start/stop in x86_pmu_start()	2012-02-07 16:58:56 +01:00
gcov	gcov: disable CONSTRUCTORS for UML	2011-07-26 16:49:45 -07:00
irq	module_param: make bool parameters really bool (core code)	2012-01-13 09:32:18 +10:30
power	PM / Freezer: Thaw only kernel threads if freezing of kernel threads fails	2012-02-04 22:23:05 +01:00
sched	sched/rt: Fix task stack corruption under __ARCH_WANT_INTERRUPTS_ON_CTXSW	2012-01-27 12:49:41 +01:00
time	Merge branch 'rcu/fixes-for-v3.2' into rcu/urgent	2012-01-16 09:41:18 -08:00
trace	Merge branch 'perf-core-for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/tip/tip	2012-01-15 11:26:35 -08:00
.gitignore
acct.c	Merge branch 'for-linus2' of git://git.kernel.org/pub/scm/linux/kernel/git/viro/vfs	2012-01-08 12:19:57 -08:00
async.c	kernel/async: remove redundant declaration.	2012-01-13 09:32:18 +10:30
audit_tree.c	audit_tree,rcu: Convert call_rcu(__put_tree) to kfree_rcu()	2011-07-20 14:10:11 -07:00
audit_watch.c	kill path_lookup()	2011-03-14 09:15:23 -04:00
audit.c	Merge branch 'for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/viro/audit	2012-01-17 16:41:31 -08:00
audit.h	audit: remove AUDIT_SETUP_CONTEXT as it isn't used	2012-01-17 16:16:57 -05:00
auditfilter.c	audit: allow interfield comparison in audit rules	2012-01-17 16:17:01 -05:00
auditsc.c	kernel-doc: fix new warnings in auditsc.c	2012-01-23 08:44:53 -08:00
backtracetest.c
bounds.c	memcg: remove direct page_cgroup-to-page pointer	2011-03-23 19:46:28 -07:00
capability.c	Revert "capabitlies: ns_capable can use the cap helpers rather than lsm call"	2012-01-17 10:19:41 -08:00
cgroup_freezer.c	Merge branch 'for-3.3' of git://git.kernel.org/pub/scm/linux/kernel/git/tj/cgroup	2012-01-09 12:59:24 -08:00
cgroup.c	Merge branch 'for-3.3' of git://git.kernel.org/pub/scm/linux/kernel/git/tj/cgroup	2012-01-09 12:59:24 -08:00
compat.c	kernel: Fix files explicitly needing EXPORT_SYMBOL infrastructure	2011-10-31 19:30:05 -04:00
configs.c	kernel/configs.c: include MODULE_*() when CONFIG_IKCONFIG_PROC=n	2011-07-25 20:57:15 -07:00
cpu_pm.c	cpu_pm: call notifiers during suspend	2011-09-23 12:05:29 +05:30
cpu.c	Merge branch 'pm-for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/rafael/linux-pm	2012-01-08 13:10:57 -08:00
cpuset.c	Merge branch 'for-3.3' of git://git.kernel.org/pub/scm/linux/kernel/git/tj/cgroup	2012-01-09 12:59:24 -08:00
crash_dump.c	Merge branch 'modsplit-Oct31_2011' of git://git.kernel.org/pub/scm/linux/kernel/git/paulg/linux	2011-11-06 19:44:47 -08:00
cred.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
delayacct.c	KVM: Steal time implementation	2011-07-14 12:59:14 +03:00
dma.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
elfcore.c
exec_domain.c
exit.c	sched: Fix ancient race in do_exit()	2012-01-27 11:55:36 +01:00
extable.c	extable, core_kernel_data(): Make sure all archs define _sdata	2011-05-20 08:56:56 +02:00
fork.c	Merge branch 'for-linus' of git://git.kernel.dk/linux-block	2012-02-11 10:07:11 -08:00
freezer.c	freezer: kill unused set_freezable_with_signal()	2011-11-23 09:28:17 -08:00
futex_compat.c	userns: user namespaces: convert several capable() calls	2011-03-23 19:47:08 -07:00
futex.c	futex: Fix uninterruptible loop due to gate_area	2011-12-31 11:48:28 -08:00
groups.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
hrtimer.c	Merge branch 'timers-urgent-for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/tip/tip	2011-11-28 08:43:52 -08:00
hung_task.c	hung_task: fix false positive during vfork	2012-01-03 16:14:32 -08:00
irq_work.c	kernel: fix two implicit header assumptions in irq_work.c	2011-10-31 09:20:12 -04:00
itimer.c	[S390] cputime: add sparse checking and cleanup	2011-12-15 14:56:19 +01:00
jump_label.c	Merge remote-tracking branch 'tip/perf/core' into kvm-updates/3.3	2011-12-27 11:22:24 +02:00
kallsyms.c	Merge branch 'core-fixes-for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/tip/linux-2.6-tip	2011-03-25 17:52:22 -07:00
Kconfig.freezer
Kconfig.hz
Kconfig.locks	arch:Kconfig.locks Remove unused config option.	2011-04-10 17:01:05 +02:00
Kconfig.preempt	sched: Isolate preempt counting in its own config option	2011-06-10 15:15:40 +02:00
kexec.c	kdump: crashk_res init check for /sys/kernel/kexec_crash_size	2012-01-12 20:13:11 -08:00
kfifo.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
kmod.c	Merge branch 'pm-sleep' into pm-for-linus	2011-12-25 23:42:20 +01:00
kprobes.c	kprobes: fix a memory leak in function pre_handler_kretprobe()	2012-02-03 16:16:41 -08:00
ksysfs.c	kernel: ksysfs.c is implicitly using stat.h	2011-10-31 09:20:13 -04:00
kthread.c	freezer: kill unused set_freezable_with_signal()	2011-11-23 09:28:17 -08:00
latencytop.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
lockdep_internals.h
lockdep_proc.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
lockdep_states.h
lockdep.c	Merge branch 'perf-core-for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/tip/tip	2012-01-06 08:02:58 -08:00
Makefile	PM: Make sysrq-o be available for CONFIG_PM unset	2012-01-14 00:33:03 +01:00
module.c	error: implicit declaration of function 'module_flags_taint'	2012-01-15 16:21:07 -08:00
mutex-debug.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
mutex-debug.h	mutex: Use p->on_cpu for the adaptive spin	2011-04-14 08:52:33 +02:00
mutex.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
mutex.h	mutex: Use p->on_cpu for the adaptive spin	2011-04-14 08:52:33 +02:00
notifier.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
nsproxy.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
padata.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
panic.c	panic: don't print redundant backtraces on oops	2012-01-12 20:13:11 -08:00
params.c	module: make module param bint handle nul value	2012-02-14 11:02:15 +10:30
pid_namespace.c	sysctl: add the kernel.ns_last_pid control	2012-01-12 20:13:11 -08:00
pid.c	sysctl: add the kernel.ns_last_pid control	2012-01-12 20:13:11 -08:00
posix-cpu-timers.c	[S390] cputime: add sparse checking and cleanup	2011-12-15 14:56:19 +01:00
posix-timers.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
printk.c	module_param: make bool parameters really bool (core code)	2012-01-13 09:32:18 +10:30
profile.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
ptrace.c	Merge branch 'for-linus' of git://selinuxproject.org/~jmorris/linux-security	2012-01-14 18:36:33 -08:00
range.c	range: fix bogus misuse of module.h to get printk()	2011-10-31 09:20:11 -04:00
rcu.h	rcu: Avoid waking up CPUs having only kfree_rcu() callbacks	2012-02-21 09:03:25 -08:00
rcupdate.c	rcu: Detect illegal rcu dereference in extended quiescent state	2011-12-11 10:31:30 -08:00
rcutiny_plugin.h	rcu: Protect __rcu_read_unlock() against scheduler-using irq handlers	2012-02-21 09:03:39 -08:00
rcutiny.c	rcu: Avoid waking up CPUs having only kfree_rcu() callbacks	2012-02-21 09:03:25 -08:00
rcutorture.c	rcu: Make rcutorture flag online/offline failures	2012-02-21 09:03:35 -08:00
rcutree_plugin.h	rcu: Limit lazy-callback duration	2012-02-21 09:03:36 -08:00
rcutree_trace.c	rcu: Avoid waking up CPUs having only kfree_rcu() callbacks	2012-02-21 09:03:25 -08:00
rcutree.c	rcu: Remove single-rcu_node optimization in rcu_start_gp()	2012-02-21 09:03:38 -08:00
rcutree.h	rcu: Simplify offline processing	2012-02-21 09:03:34 -08:00
relay.c	relay: prevent integer overflow in relay_open()	2012-02-10 09:04:49 +01:00
res_counter.c	net: introduce res_counter_charge_nofail() for socket allocations	2012-01-22 15:08:46 -05:00
resource.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
rtmutex_common.h	rtmutex: Simplify PI algorithm and make highest prio task get lock	2011-01-27 21:13:51 -05:00
rtmutex-debug.c	lockdep, rtmutex, bug: Show taint flags on error	2011-12-06 08:16:49 +01:00
rtmutex-debug.h
rtmutex-tester.c	rtmutex-tester: convert sysdev_class to a regular subsystem	2011-12-14 14:54:22 -08:00
rtmutex.c	Revert "rcu: Permit rt_mutex_unlock() with irqs disabled"	2011-12-11 10:33:18 -08:00
rtmutex.h
rwsem.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
seccomp.c	seccomp: audit abnormal end to a process due to seccomp	2012-01-17 16:16:55 -05:00
semaphore.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
signal.c	user namespace: make signal.c respect user namespaces	2012-01-10 16:30:54 -08:00
smp.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
softirq.c	rcu: Fix early call to rcu_idle_enter()	2011-12-11 10:31:38 -08:00
spinlock.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
srcu.c	rcu: Add lockdep-RCU checks for simple self-deadlock	2012-02-21 09:03:23 -08:00
stacktrace.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
stop_machine.c	Merge branch 'modsplit-Oct31_2011' of git://git.kernel.org/pub/scm/linux/kernel/git/paulg/linux	2011-11-06 19:44:47 -08:00
sys_ni.c	Cross Memory Attach	2011-10-31 17:30:44 -07:00
sys.c	c/r: prctl: add PR_SET_MM codes to set up mm_struct entries	2012-01-12 20:13:13 -08:00
sysctl_binary.c	binary_sysctl(): fix memory leak	2011-12-20 10:25:04 -08:00
sysctl_check.c	xfs: remove subdirectories	2011-08-12 16:21:35 -05:00
sysctl.c	x86: Panic on detection of stack overflow	2011-12-05 11:37:47 +01:00
taskstats.c	Make TASKSTATS require root access	2011-09-19 17:04:37 -07:00
test_kprobes.c	kprobes: Fix selftest to clear flags field for reusing probes	2010-10-14 08:55:27 +02:00
time.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
timeconst.pl
timer.c	Merge branch 'core-debugobjects-for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/tip/tip	2012-01-06 07:53:34 -08:00
tracepoint.c	tracepoints/module: Fix disabling tracepoints with taint CRAP or OOT	2012-01-16 11:35:57 -05:00
tsacct.c	[S390] cputime: add sparse checking and cleanup	2011-12-15 14:56:19 +01:00
uid16.c	userns: user namespaces: convert several capable() calls	2011-03-23 19:47:08 -07:00
up.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
user_namespace.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
user-return-notifier.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
user.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
utsname_sysctl.c	Merge branch 'modsplit-Oct31_2011' of git://git.kernel.org/pub/scm/linux/kernel/git/paulg/linux	2011-11-06 19:44:47 -08:00
utsname.c	kernel: Map most files to use export.h instead of module.h	2011-10-31 09:20:12 -04:00
wait.c	lockdep/waitqueues: Add better annotation	2011-12-21 10:07:39 +01:00
watchdog.c	bugs, x86: Fix printk levels for panic, softlockups and stack dumps	2012-01-26 21:28:45 +01:00
workqueue_sched.h
workqueue.c	workqueue: make alloc_workqueue() take printf fmt and args for name	2012-01-10 16:30:54 -08:00