Re: [PATCH 03/12] task_isolation: userspace hard isolation from kernel

[PATCH 00/12] "Task_isolation" mode · Alex Belits <hidden> · 2020-03-04
[PATCH 01/12] task_isolation: vmstat: add quiet_vmstat_sync function · Alex Belits <hidden> · 2020-03-04
[PATCH 02/12] task_isolation: vmstat: add vmstat_idle function · Alex Belits <hidden> · 2020-03-04
[PATCH 03/12] task_isolation: userspace hard isolation from kernel · Alex Belits <hidden> · 2020-03-04
Re: [PATCH 03/12] task_isolation: userspace hard isolation from kernel · Frederic Weisbecker <frederic@kernel.org> · 2020-03-05
Re: [EXT] Re: [PATCH 03/12] task_isolation: userspace hard isolation from kernel · Alex Belits <hidden> · 2020-03-08
Re: [PATCH 03/12] task_isolation: userspace hard isolation from kernel · Marcelo Tosatti <hidden> · 2020-04-28
Re: [PATCH 03/12] task_isolation: userspace hard isolation from kernel · Frederic Weisbecker <frederic@kernel.org> · 2020-03-06
Re: [EXT] Re: [PATCH 03/12] task_isolation: userspace hard isolation from kernel · Alex Belits <hidden> · 2020-03-08
Re: [PATCH 03/12] task_isolation: userspace hard isolation from kernel · Frederic Weisbecker <frederic@kernel.org> · 2020-03-06
Re: [EXT] Re: [PATCH 03/12] task_isolation: userspace hard isolation from kernel · Alex Belits <hidden> · 2020-03-08
[PATCH 04/12] task_isolation: Add task isolation hooks to arch-independent code · Alex Belits <hidden> · 2020-03-04
[PATCH 05/12] task_isolation: arch/x86: enable task isolation functionality · Alex Belits <hidden> · 2020-03-04
[PATCH 06/12] task_isolation: arch/arm64: enable task isolation functionality · Alex Belits <hidden> · 2020-03-04
Re: [PATCH 06/12] task_isolation: arch/arm64: enable task isolation functionality · Mark Rutland <mark.rutland@arm.com> · 2020-03-04
Re: [EXT] Re: [PATCH 06/12] task_isolation: arch/arm64: enable task isolation functionality · Alex Belits <hidden> · 2020-03-08
[PATCH 07/12] task_isolation: arch/arm: enable task isolation functionality · Alex Belits <hidden> · 2020-03-04
[PATCH 08/12] task_isolation: don't interrupt CPUs with tick_nohz_full_kick_cpu() · Alex Belits <hidden> · 2020-03-04
Re: [PATCH 08/12] task_isolation: don't interrupt CPUs with tick_nohz_full_kick_cpu() · Frederic Weisbecker <frederic@kernel.org> · 2020-03-06
Re: [EXT] Re: [PATCH 08/12] task_isolation: don't interrupt CPUs with tick_nohz_full_kick_cpu() · Alex Belits <hidden> · 2020-03-08
Re: [EXT] Re: [PATCH 08/12] task_isolation: don't interrupt CPUs with tick_nohz_full_kick_cpu() · Frederic Weisbecker <frederic@kernel.org> · 2020-03-09
[PATCH 09/12] task_isolation: net: don't flush backlog on CPUs running isolated tasks · Alex Belits <hidden> · 2020-03-04
[PATCH 10/12] task_isolation: ringbuffer: don't interrupt CPUs running isolated tasks on buffer resize · Alex Belits <hidden> · 2020-03-04
[PATCH 11/12] task_isolation: kick_all_cpus_sync: don't kick isolated cpus · Alex Belits <hidden> · 2020-03-04
Re: [PATCH 11/12] task_isolation: kick_all_cpus_sync: don't kick isolated cpus · Frederic Weisbecker <frederic@kernel.org> · 2020-03-06
Re: [EXT] Re: [PATCH 11/12] task_isolation: kick_all_cpus_sync: don't kick isolated cpus · Alex Belits <hidden> · 2020-03-08
Re: [EXT] Re: [PATCH 11/12] task_isolation: kick_all_cpus_sync: don't kick isolated cpus · Frederic Weisbecker <frederic@kernel.org> · 2020-03-09
[PATCH 12/12] task_isolation: CONFIG_TASK_ISOLATION prevents distribution of jobs to non-housekeeping CPUs · Alex Belits <hidden> · 2020-03-04
[PATCH v2 00/12] "Task_isolation" mode · Alex Belits <hidden> · 2020-03-08
[PATCH v2 01/12] task_isolation: vmstat: add quiet_vmstat_sync function · Alex Belits <hidden> · 2020-03-08
[PATCH v2 02/12] task_isolation: vmstat: add vmstat_idle function · Alex Belits <hidden> · 2020-03-08
[PATCH v2 03/12] task_isolation: userspace hard isolation from kernel · Alex Belits <hidden> · 2020-03-08
Re: [PATCH v2 03/12] task_isolation: userspace hard isolation from kernel · Marta Rybczynska <hidden> · 2020-03-27
Re: [PATCH v2 03/12] task_isolation: userspace hard isolation from kernel · Kevyn-Alexandre Paré <hidden> · 2020-04-06
Re: [PATCH v2 03/12] task_isolation: userspace hard isolation from kernel · Kevyn-Alexandre Paré <hidden> · 2020-04-06
[PATCH v2 04/12] task_isolation: Add task isolation hooks to arch-independent code · Alex Belits <hidden> · 2020-03-08
[PATCH v2 05/12] task_isolation: arch/x86: enable task isolation functionality · Alex Belits <hidden> · 2020-03-08
[PATCH v2 06/12] task_isolation: arch/arm64: enable task isolation functionality · Alex Belits <hidden> · 2020-03-08
Re: [PATCH v2 06/12] task_isolation: arch/arm64: enable task isolation functionality · Mark Rutland <mark.rutland@arm.com> · 2020-03-09
[PATCH v2 07/12] task_isolation: arch/arm: enable task isolation functionality · Alex Belits <hidden> · 2020-03-08
[PATCH v2 08/12] task_isolation: don't interrupt CPUs with tick_nohz_full_kick_cpu() · Alex Belits <hidden> · 2020-03-08
[PATCH v2 09/12] task_isolation: net: don't flush backlog on CPUs running isolated tasks · Alex Belits <hidden> · 2020-03-08
[PATCH v2 10/12] task_isolation: ringbuffer: don't interrupt CPUs running isolated tasks on buffer resize · Alex Belits <hidden> · 2020-03-08
Re: [PATCH v2 10/12] task_isolation: ringbuffer: don't interrupt CPUs running isolated tasks on buffer resize · Kevyn-Alexandre Paré <hidden> · 2020-04-06
[PATCH v2 11/12] task_isolation: kick_all_cpus_sync: don't kick isolated cpus · Alex Belits <hidden> · 2020-03-08
[PATCH v2 12/12] task_isolation: CONFIG_TASK_ISOLATION prevents distribution of jobs to non-housekeeping CPUs · Alex Belits <hidden> · 2020-03-08
Re: [PATCH v3 00/13] "Task_isolation" mode · Alex Belits <hidden> · 2020-04-09
[PATCH 01/13] task_isolation: vmstat: add quiet_vmstat_sync function · Alex Belits <hidden> · 2020-04-09
[PATCH 02/13] task_isolation: vmstat: add vmstat_idle function · Alex Belits <hidden> · 2020-04-09
[PATCH v3 03/13] task_isolation: add instruction synchronization memory barrier · Alex Belits <hidden> · 2020-04-09
Re: [PATCH v3 03/13] task_isolation: add instruction synchronization memory barrier · Mark Rutland <mark.rutland@arm.com> · 2020-04-15
Re: [EXT] Re: [PATCH v3 03/13] task_isolation: add instruction synchronization memory barrier · Alex Belits <hidden> · 2020-04-19
Re: [EXT] Re: [PATCH v3 03/13] task_isolation: add instruction synchronization memory barrier · Will Deacon <will@kernel.org> · 2020-04-20
Re: [EXT] Re: [PATCH v3 03/13] task_isolation: add instruction synchronization memory barrier · Mark Rutland <mark.rutland@arm.com> · 2020-04-20
Re: [EXT] Re: [PATCH v3 03/13] task_isolation: add instruction synchronization memory barrier · Will Deacon <will@kernel.org> · 2020-04-20
Re: [EXT] Re: [PATCH v3 03/13] task_isolation: add instruction synchronization memory barrier · Will Deacon <will@kernel.org> · 2020-04-21
Re: [EXT] Re: [PATCH v3 03/13] task_isolation: add instruction synchronization memory barrier · Mark Rutland <mark.rutland@arm.com> · 2020-04-20
[PATCH v3 04/13] task_isolation: userspace hard isolation from kernel · Alex Belits <hidden> · 2020-04-09
Re: [PATCH v3 04/13] task_isolation: userspace hard isolation from kernel · Andy Lutomirski <luto@amacapital.net> · 2020-04-09
Re: [PATCH v3 04/13] task_isolation: userspace hard isolation from kernel · Alex Belits <hidden> · 2020-04-19
[PATCH 05/13] task_isolation: Add task isolation hooks to arch-independent code · Alex Belits <hidden> · 2020-04-09
[PATCH 06/13] task_isolation: arch/x86: enable task isolation functionality · Alex Belits <hidden> · 2020-04-09
[PATCH v3 07/13] task_isolation: arch/arm64: enable task isolation functionality · Alex Belits <hidden> · 2020-04-09
Re: [PATCH v3 07/13] task_isolation: arch/arm64: enable task isolation functionality · Catalin Marinas <catalin.marinas@arm.com> · 2020-04-22
[PATCH v3 08/13] task_isolation: arch/arm: enable task isolation functionality · Alex Belits <hidden> · 2020-04-09
[PATCH v3 09/13] task_isolation: don't interrupt CPUs with tick_nohz_full_kick_cpu() · Alex Belits <hidden> · 2020-04-09
[PATCH v3 10/13] task_isolation: net: don't flush backlog on CPUs running isolated tasks · Alex Belits <hidden> · 2020-04-09
[PATCH v3 11/13] task_isolation: ringbuffer: don't interrupt CPUs running isolated tasks on buffer resize · Alex Belits <hidden> · 2020-04-09
[PATCH v3 12/13] task_isolation: kick_all_cpus_sync: don't kick isolated cpus · Alex Belits <hidden> · 2020-04-09
[PATCH v3 13/13] task_isolation: CONFIG_TASK_ISOLATION prevents distribution of jobs to non-housekeeping CPUs · Alex Belits <hidden> · 2020-04-09

From: Frederic Weisbecker <frederic@kernel.org>
Date: 2020-03-06 15:26:37
Also in: linux-arch, lkml

On Wed, Mar 04, 2020 at 04:07:12PM +0000, Alex Belits wrote:

+
+/*
+ * Print message prefixed with the description of the current (or
+ * last) isolated task on a given CPU. Intended for isolation breaking
+ * messages that include target task for the user's convenience.
+ *
+ * Messages produced with this function may have obsolete task
+ * information if isolated tasks managed to exit, start and enter
+ * isolation multiple times, or multiple tasks tried to enter
+ * isolation on the same CPU at once. For those unusual cases it would
+ * contain a valid description of the cause for isolation breaking and
+ * target CPU number, just not the correct description of which task
+ * ended up losing isolation.
+ */
+int task_isolation_message(int cpu, int level, bool supp, const char *fmt, ...)
+{
+	struct isol_task_desc *desc;
+	struct task_struct *task;
+	va_list args;
+	char buf_prefix[TASK_COMM_LEN + 20 + 3 * 20];
+	char buf[200];
+	int curr_cpu, ind_counter, ind_counter_old, ind;
+
+	curr_cpu = get_cpu();
+	desc = &per_cpu(isol_task_descs, cpu);
+	ind_counter = atomic_read(&desc->curr_index);
+
+	if (curr_cpu == cpu) {
+		/*
+		 * Message is for the current CPU so current
+		 * task_struct should be used instead of cached
+		 * information.
+		 *
+		 * Like in other diagnostic messages, if issued from
+		 * interrupt context, current will be the interrupted
+		 * task. Unlike other diagnostic messages, this is
+		 * always relevant because the message is about
+		 * interrupting a task.
+		 */
+		ind = ind_counter & 1;
+		if (supp && desc->warned[ind]) {
+			/*
+			 * If supp is true, skip the message if the
+			 * same task was mentioned in the message
+			 * originated on remote CPU, and it did not
+			 * re-enter isolated state since then (warned
+			 * is true). Only local messages following
+			 * remote messages, likely about the same
+			 * isolation breaking event, are skipped to
+			 * avoid duplication. If remote cause is
+			 * immediately followed by a local one before
+			 * isolation is broken, local cause is skipped
+			 * from messages.
+			 */
+			put_cpu();
+			return 0;
+		}
+		task = current;
+		snprintf(buf_prefix, sizeof(buf_prefix),
+			 "isolation %s/%d/%d (cpu %d)",
+			 task->comm, task->tgid, task->pid, cpu);
+		put_cpu();
+	} else {
+		/*
+		 * Message is for remote CPU, use cached information.
+		 */
+		put_cpu();
+		/*
+		 * Make sure, index remained unchanged while data was
+		 * copied. If it changed, data that was copied may be
+		 * inconsistent because two updates in a sequence could
+		 * overwrite the data while it was being read.
+		 */
+		do {
+			/* Make sure we are reading up to date values */
+			smp_mb();
+			ind = ind_counter & 1;
+			snprintf(buf_prefix, sizeof(buf_prefix),
+				 "isolation %s/%d/%d (cpu %d)",
+				 desc->comm[ind], desc->tgid[ind],
+				 desc->pid[ind], cpu);
+			desc->warned[ind] = true;
+			ind_counter_old = ind_counter;
+			/* Record the warned flag, then re-read descriptor */
+			smp_mb();
+			ind_counter = atomic_read(&desc->curr_index);
+			/*
+			 * If the counter changed, something was updated, so
+			 * repeat everything to get the current data
+			 */
+		} while (ind_counter != ind_counter_old);
+	}

So the need to log the fact we are sending an event to a remote CPU that *may be*
running an isolated task makes things very complicated and even racy.

How bad would it be to only log those interruptions once they land on the target?

Thanks.

`h`	back out one level
`j`	next message in thread
`k`	previous message in thread
`l`	drill in
`Esc`	close help / fold thread tree
`?`	toggle this help