Re: [patch 1/12] sched: ratelimit nohz

Previous message: [thread] [date] [author]
Next message: [thread] [date] [author]
From: Mike Galbraith
Date: Thursday, March 11, 2010 - 2:50 am

sched: ratelimit nohz

Entering nohz code on every micro-idle is costing ~10% throughput for netperf
TCP_RR when scheduling cross-cpu.  Rate limiting entry fixes this, but raises
ticks a bit.  On my Q6600, an idle box goes from ~85 interrupts/sec to 128.

The higher the context switch rate, the more nohz entry costs.  With this patch
and some cycle recovery patches in my tree, max cross cpu context switch rate is
improved by ~16%, a large portion of which of which is this ratelimiting.

Signed-off-by: Mike Galbraith <efault@gmx.de>
Cc: Ingo Molnar <mingo@elte.hu>
Cc: Peter Zijlstra <a.p.zijlstra@chello.nl>
LKML-Reference: <new-submission>

---
 include/linux/sched.h    |    6 ++++++
 kernel/sched.c           |   12 ++++++++++++
 kernel/time/tick-sched.c |    3 +++
 3 files changed, 21 insertions(+)

Index: linux-2.6/include/linux/sched.h
===================================================================
--- linux-2.6.orig/include/linux/sched.h
+++ linux-2.6/include/linux/sched.h
@@ -271,11 +271,17 @@ extern cpumask_var_t nohz_cpu_mask;
 #if defined(CONFIG_SMP) && defined(CONFIG_NO_HZ)
 extern int select_nohz_load_balancer(int cpu);
 extern int get_nohz_load_balancer(void);
+extern int nohz_ratelimit(int cpu);
 #else
 static inline int select_nohz_load_balancer(int cpu)
 {
 	return 0;
 }
+
+static inline int nohz_ratelimit(int cpu)
+{
+	return 0;
+}
 #endif
 
 /*
Index: linux-2.6/kernel/sched.c
===================================================================
--- linux-2.6.orig/kernel/sched.c
+++ linux-2.6/kernel/sched.c
@@ -492,6 +492,7 @@ struct rq {
 	#define CPU_LOAD_IDX_MAX 5
 	unsigned long cpu_load[CPU_LOAD_IDX_MAX];
 #ifdef CONFIG_NO_HZ
+	u64 nohz_stamp;
 	unsigned char in_nohz_recently;
 #endif
 	/* capture load from *all* tasks on this cpu: */
@@ -1228,6 +1229,17 @@ void wake_up_idle_cpu(int cpu)
 	if (!tsk_is_polling(rq->idle))
 		smp_send_reschedule(cpu);
 }
+
+int nohz_ratelimit(int cpu)
+{
+	struct rq *rq = cpu_rq(cpu);
+	u64 diff = rq->clock - rq->nohz_stamp;
+
+	rq->nohz_stamp = rq->clock;
+
+	return diff < (NSEC_PER_SEC / HZ) >> 1;
+}
+
 #endif /* CONFIG_NO_HZ */
 
 static u64 sched_avg_period(void)
Index: linux-2.6/kernel/time/tick-sched.c
===================================================================
--- linux-2.6.orig/kernel/time/tick-sched.c
+++ linux-2.6/kernel/time/tick-sched.c
@@ -262,6 +262,9 @@ void tick_nohz_stop_sched_tick(int inidl
 		goto end;
 	}
 
+	if (nohz_ratelimit(cpu))
+		goto end;
+
 	ts->idle_calls++;
 	/* Read jiffies and the time when jiffies were updated last */
 	do {


--
Previous message: [thread] [date] [author]
Next message: [thread] [date] [author]

Messages in current thread:
[patch 0/12] sched: fastpath cycle recovery, Mike Galbraith, (Thu Mar 11, 2:49 am)
Re: [patch 1/12] sched: ratelimit nohz, Mike Galbraith, (Thu Mar 11, 2:50 am)
Re: [patch 2/12] sched: remove avg_wakeup, Mike Galbraith, (Thu Mar 11, 2:51 am)
Re: [patch 3/12] sched: remove avg_overlap, Mike Galbraith, (Thu Mar 11, 2:52 am)
Re: [patch 4/12] sched: cleanup/optimize clock updates, Mike Galbraith, (Thu Mar 11, 2:53 am)
Re: [patch 6/12] sched: fix select_idle_sibling(), Mike Galbraith, (Thu Mar 11, 2:56 am)
Re: [patch 7/12] sched: remove NORMALIZED_SLEEPER, Mike Galbraith, (Thu Mar 11, 2:57 am)
Re: [patch 8/12] sched: remove FAIR_SLEEPERS feature, Mike Galbraith, (Thu Mar 11, 2:58 am)
Re: [patch 9/12] sched: remove WAKEUP_SYNC feature, Mike Galbraith, (Thu Mar 11, 2:59 am)
Re: [patch 11/12] sched: remove ASYM_GRAN feature, Mike Galbraith, (Thu Mar 11, 3:01 am)
Re: [patch 10/12] sched: remove SYNC_WAKEUPS feature, Mike Galbraith, (Thu Mar 11, 3:03 am)
Re: [patch 12/12] sched: remove AFFINE_WAKEUPS feature, Mike Galbraith, (Thu Mar 11, 3:04 am)
[tip:sched/core] sched: Rate-limit nohz, tip-bot for Mike Gal ..., (Thu Mar 11, 11:30 am)
[tip:sched/core] sched: Remove avg_wakeup, tip-bot for Mike Gal ..., (Thu Mar 11, 11:30 am)
[tip:sched/core] sched: Remove avg_overlap, tip-bot for Mike Gal ..., (Thu Mar 11, 11:31 am)
[tip:sched/core] sched: Cleanup/optimize clock updates, tip-bot for Mike Gal ..., (Thu Mar 11, 11:31 am)
[tip:sched/core] sched: Tweak sched_latency and min_granul ..., tip-bot for Mike Gal ..., (Thu Mar 11, 11:31 am)
[tip:sched/core] sched: Fix select_idle_sibling(), tip-bot for Mike Gal ..., (Thu Mar 11, 11:32 am)
[tip:sched/core] sched: Remove NORMALIZED_SLEEPER, tip-bot for Mike Gal ..., (Thu Mar 11, 11:32 am)
[tip:sched/core] sched: Remove FAIR_SLEEPERS feature, tip-bot for Mike Gal ..., (Thu Mar 11, 11:32 am)
[tip:sched/core] sched: Remove WAKEUP_SYNC feature, tip-bot for Mike Gal ..., (Thu Mar 11, 11:32 am)
[tip:sched/core] sched: Remove SYNC_WAKEUPS feature, tip-bot for Mike Gal ..., (Thu Mar 11, 11:33 am)
[tip:sched/core] sched: Remove ASYM_GRAN feature, tip-bot for Mike Gal ..., (Thu Mar 11, 11:33 am)
[tip:sched/core] sched: Remove AFFINE_WAKEUPS feature, tip-bot for Mike Gal ..., (Thu Mar 11, 11:33 am)
Re: [tip:sched/core] sched: Remove AFFINE_WAKEUPS feature, Mike Galbraith, (Thu Mar 11, 9:37 pm)