提交 · 8446f1d391f3d27e6bf9c43d4cbcdac0ca720417 · openeuler / raspberrypi-kernel

08 9月, 2005 12 次提交

由 Ingo Molnar 提交于 9月 06, 2005

This patch adds a new kernel debug feature: CONFIG_DETECT_SOFTLOCKUP.

When enabled then per-CPU watchdog threads are started, which try to run
once per second.  If they get delayed for more than 10 seconds then a
callback from the timer interrupt detects this condition and prints out a
warning message and a stack dump (once per lockup incident).  The feature
is otherwise non-intrusive, it doesnt try to unlock the box in any way, it
only gets the debug info out, automatically, and on all CPUs affected by
the lockup.
Signed-off-by: NIngo Molnar <mingo@elte.hu>
Signed-off-by: NNishanth Aravamudan <nacc@us.ibm.com>
Signed-Off-By: NMatthias Urlichs <smurf@smurf.noris.de>
Signed-off-by: NRichard Purdie <rpurdie@rpsys.net>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

8446f1d3

[PATCH] FUTEX_WAKE_OP: pthread_cond_signal() speedup · 4732efbe

由 Jakub Jelinek 提交于 9月 06, 2005

ATM pthread_cond_signal is unnecessarily slow, because it wakes one waiter
(which at least on UP usually means an immediate context switch to one of
the waiter threads).  This waiter wakes up and after a few instructions it
attempts to acquire the cv internal lock, but that lock is still held by
the thread calling pthread_cond_signal.  So it goes to sleep and eventually
the signalling thread is scheduled in, unlocks the internal lock and wakes
the waiter again.

Now, before 2003-09-21 NPTL was using FUTEX_REQUEUE in pthread_cond_signal
to avoid this performance issue, but it was removed when locks were
redesigned to the 3 state scheme (unlocked, locked uncontended, locked
contended).

Following scenario shows why simply using FUTEX_REQUEUE in
pthread_cond_signal together with using lll_mutex_unlock_force in place of
lll_mutex_unlock is not enough and probably why it has been disabled at
that time:

The number is value in cv->__data.__lock.
        thr1            thr2            thr3
0       pthread_cond_wait
1       lll_mutex_lock (cv->__data.__lock)
0       lll_mutex_unlock (cv->__data.__lock)
0       lll_futex_wait (&cv->__data.__futex, futexval)
0                       pthread_cond_signal
1                       lll_mutex_lock (cv->__data.__lock)
1                                       pthread_cond_signal
2                                       lll_mutex_lock (cv->__data.__lock)
2                                         lll_futex_wait (&cv->__data.__lock, 2)
2                       lll_futex_requeue (&cv->__data.__futex, 0, 1, &cv->__data.__lock)
                          # FUTEX_REQUEUE, not FUTEX_CMP_REQUEUE
2                       lll_mutex_unlock_force (cv->__data.__lock)
0                         cv->__data.__lock = 0
0                         lll_futex_wake (&cv->__data.__lock, 1)
1       lll_mutex_lock (cv->__data.__lock)
0       lll_mutex_unlock (cv->__data.__lock)
          # Here, lll_mutex_unlock doesn't know there are threads waiting
          # on the internal cv's lock

Now, I believe it is possible to use FUTEX_REQUEUE in pthread_cond_signal,
but it will cost us not one, but 2 extra syscalls and, what's worse, one of
these extra syscalls will be done for every single waiting loop in
pthread_cond_*wait.

We would need to use lll_mutex_unlock_force in pthread_cond_signal after
requeue and lll_mutex_cond_lock in pthread_cond_*wait after lll_futex_wait.

Another alternative is to do the unlocking pthread_cond_signal needs to do
(the lock can't be unlocked before lll_futex_wake, as that is racy) in the
kernel.

I have implemented both variants, futex-requeue-glibc.patch is the first
one and futex-wake_op{,-glibc}.patch is the unlocking inside of the kernel.
 The kernel interface allows userland to specify how exactly an unlocking
operation should look like (some atomic arithmetic operation with optional
constant argument and comparison of the previous futex value with another
constant).

It has been implemented just for ppc*, x86_64 and i?86, for other
architectures I'm including just a stub header which can be used as a
starting point by maintainers to write support for their arches and ATM
will just return -ENOSYS for FUTEX_WAKE_OP.  The requeue patch has been
(lightly) tested just on x86_64, the wake_op patch on ppc64 kernel running
32-bit and 64-bit NPTL and x86_64 kernel running 32-bit and 64-bit NPTL.

With the following benchmark on UP x86-64 I get:

for i in nptl-orig nptl-requeue nptl-wake_op; do echo time elf/ld.so --library-path .:$i /tmp/bench; \
for j in 1 2; do echo ( time elf/ld.so --library-path .:$i /tmp/bench ) 2>&1; done; done
time elf/ld.so --library-path .:nptl-orig /tmp/bench
real 0m0.655s user 0m0.253s sys 0m0.403s
real 0m0.657s user 0m0.269s sys 0m0.388s
time elf/ld.so --library-path .:nptl-requeue /tmp/bench
real 0m0.496s user 0m0.225s sys 0m0.271s
real 0m0.531s user 0m0.242s sys 0m0.288s
time elf/ld.so --library-path .:nptl-wake_op /tmp/bench
real 0m0.380s user 0m0.176s sys 0m0.204s
real 0m0.382s user 0m0.175s sys 0m0.207s

The benchmark is at:
http://sourceware.org/ml/libc-alpha/2005-03/txt00001.txt
Older futex-requeue-glibc.patch version is at:
http://sourceware.org/ml/libc-alpha/2005-03/txt00002.txt
Older futex-wake_op-glibc.patch version is at:
http://sourceware.org/ml/libc-alpha/2005-03/txt00003.txt
Will post a new version (just x86-64 fixes so that the patch
applies against pthread_cond_signal.S) to libc-hacker ml soon.

Attached is the kernel FUTEX_WAKE_OP patch as well as a simple-minded
testcase that will not test the atomicity of the operation, but at least
check if the threads that should have been woken up are woken up and
whether the arithmetic operation in the kernel gave the expected results.
Acked-by: NIngo Molnar <mingo@redhat.com>
Cc: Ulrich Drepper <drepper@redhat.com>
Cc: Jamie Lokier <jamie@shareable.org>
Cc: Rusty Russell <rusty@rustcorp.com.au>
Signed-off-by: NYoichi Yuasa <yuasa@hh.iij4u.or.jp>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

4732efbe

[PATCH] 3c59x PM fixes · 5b039e68

由 Rafael J. Wysocki 提交于 9月 06, 2005

This patch adds some missing pci-related calls to the suspend and resume
routines of the 3c59x driver. It also makes the driver free/request IRQ on
suspend/resume, in accordance with the proposal at:
http://lists.osdl.org/pipermail/linux-pm/2005-May/000955.htmlSigned-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

5b039e68

[PATCH] swsusp: update documentation · d7ae79c7

由 Pavel Machek 提交于 9月 06, 2005

This updates documentation a bit (mostly removing obsolete stuff), and
marks swsusp as no longer experimental in config.
Signed-off-by: NPavel Machek <pavel@suse.cz>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

d7ae79c7

[PATCH] x86_64: Fix off by one in e820_mapped · 48c8b113

由 Eric W. Biederman 提交于 9月 06, 2005

This allows a valid iommu placed immediately after memory to work, to be
recognized as after the last byte of memory and not overlapping it.
Signed-off-by: NEric W. Biederman <ebiederm@xmission.com>
Acked-by: NAndi Kleen <ak@suse.de>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

48c8b113

[PATCH] x86_64: create sysfs entries for cpu only for present cpus · a888cebe

由 Ashok Raj 提交于 9月 06, 2005

Need to create sysfs only for cpus that are present.  Without which we see
NR_CPUS entries created when we have CONFIG_HOTPLUG and CONFIG_HOTPLUG_CPU
enabled.
Signed-off-by: NAshok Raj <ashok.raj@intel.com>
Acked-by: NAndi Kleen <ak@muc.de>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

a888cebe

[PATCH] x86_64: Fix cluster mode send_IPI_allbutself to use get_cpu()/put_cpu() · 0c2b9d5c

由 Ashok Raj 提交于 9月 06, 2005

Need to ensure we dont get prempted when we clear ourself from mask when using
clustered mode genapic code.
Signed-off-by: NAshok Raj <ashok.raj@intel.com>
Acked-by: NAndi Kleen <ak@muc.de>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

0c2b9d5c

[PATCH] x86_64: prefetchw() can fall back to prefetch() if !3DNOW · 19aaabb5

由 Eric Dumazet 提交于 9月 06, 2005

This is a multi-part message in MIME format.  If the cpu lacks 3DNOW
feature, we can use a normal prefetcht0 instruction instead of NOP5.
"prefetchw (%rxx)" and "prefetcht0 (%rxx)" have the same length, ranging
from 3 to 5 bytes depending on the register.  So this patch even helps
AMD64, shortening the length of the code.
Signed-off-by: NEric Dumazet <dada1@cosmosbay.com>
Acked-by: NAndi Kleen <ak@muc.de>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

19aaabb5

[PATCH] x86_64: print processor number in show_regs · c078d326

由 Zwane Mwaikambo 提交于 9月 06, 2005

Up to date I've been using the GS value to determine the processor number
in dumps from show_regs, however this can be cumbersome to do if you don't
have the vmlinux to verify with the address of cpu_pda, how about the
following?  I considered using hard_smp_processor_id for robustness but we
already dereference current so we're already relying on MSR_GS_BASE being
sane.
Signed-off-by: NZwane Mwaikambo <zwane@arm.linux.org.uk>
Acked-by: NAndi Kleen <ak@muc.de>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

c078d326

[PATCH] x86/x86_64: deferred handling of writes to /proc/irqxx/smp_affinity · 54d5d424

由 Ashok Raj 提交于 9月 06, 2005

When handling writes to /proc/irq, current code is re-programming rte
entries directly. This is not recommended and could potentially cause
chipset's to lockup, or cause missing interrupts.

CONFIG_IRQ_BALANCE does this correctly, where it re-programs only when the
interrupt is pending. The same needs to be done for /proc/irq handling as well.
Otherwise user space irq balancers are really not doing the right thing.

- Changed pending_irq_balance_cpumask to pending_irq_migrate_cpumask for
  lack of a generic name.
- added move_irq out of IRQ_BALANCE, and added this same to X86_64
- Added new proc handler for write, so we can do deferred write at irq
  handling time.
- Display of /proc/irq/XX/smp_affinity used to display CPU_MASKALL, instead
  it now shows only active cpu masks, or exactly what was set.
- Provided a common move_irq implementation, instead of duplicating
  when using generic irq framework.

Tested on i386/x86_64 and ia64 with CONFIG_PCI_MSI turned on and off.
Tested UP builds as well.

MSI testing: tbd: I have cards, need to look for a x-over cable, although I
did test an earlier version of this patch.  Will test in a couple days.
Signed-off-by: NAshok Raj <ashok.raj@intel.com>
Acked-by: NZwane Mwaikambo <zwane@holomorphy.com>
Grudgingly-acked-by: NAndi Kleen <ak@muc.de>
Signed-off-by: NCoywolf Qi Hunt <coywolf@lovecn.org>
Signed-off-by: NAshok Raj <ashok.raj@intel.com>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

54d5d424

[PATCH] ppc32: add missing sysfs node for ocp_func_emac_data.phy_feat_exc · f63ed39c

由 Eugene Surovegin 提交于 9月 06, 2005

Add sysfs node for ocp_func_emac_data.phy_feat_exc field.
Signed-off-by: NEugene Surovegin <ebs@ebshome.net>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

f63ed39c

[PATCH] ppc32: fix ocp_device_suspend to use pm_message_t instead of u32 · 842363ff

由 Eugene Surovegin 提交于 9月 06, 2005

Recent "u32 -> pm_message_t" change triggered hidden bug in
ocp_device_suspend.  Fix it to correctly use pm_message_t instead of u32.
Signed-off-by: NEugene Surovegin <ebs@ebshome.net>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

842363ff

06 9月, 2005 28 次提交

L

Merge master.kernel.org:/pub/scm/linux/kernel/git/paulus/ppc64-2.6 · 4706df3d
由 Linus Torvalds 提交于 9月 06, 2005

4706df3d
L

Merge branch 'upstream' of master.kernel.org:/pub/scm/linux/kernel/git/jgarzik/netdev-2.6 · 5bcaa155
由 Linus Torvalds 提交于 9月 06, 2005

5bcaa155
L

Merge master.kernel.org:/home/rmk/linux-2.6-arm · 1e231efe
由 Linus Torvalds 提交于 9月 06, 2005

1e231efe
L

Merge master.kernel.org:/pub/scm/linux/kernel/git/sam/kbuild · ef88b7db
由 Linus Torvalds 提交于 9月 06, 2005

ef88b7db
L

Merge master.kernel.org:/pub/scm/linux/kernel/git/gregkh/driver-2.6 · f65e7769
由 Linus Torvalds 提交于 9月 06, 2005

f65e7769
L

Merge master.kernel.org:/pub/scm/linux/kernel/git/gregkh/i2c-2.6 · 8566cfc9
由 Linus Torvalds 提交于 9月 06, 2005

8566cfc9
L

Merge master.kernel.org:/pub/scm/linux/kernel/git/davem/sparc-2.6 · 7bdb2b6a
由 Linus Torvalds 提交于 9月 06, 2005

7bdb2b6a

[PATCH] remove linux/version.h include from arch/ppc64 · cebb2b15

由 Olaf Hering 提交于 7月 10, 2005

Changing CONFIG_LOCALVERSION rebuilds too much, for no apparent reason.

Use system_utsname for progress and debug header.
Signed-off-by: NOlaf Hering <olh@suse.de>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

cebb2b15

[PATCH] Invert sense of SLB class bit · 14b34661

由 David Gibson 提交于 9月 06, 2005

Currently, we set the class bit in kernel SLB entries, and clear it on
user SLB entries.  On POWER5, ERAT entries created in real mode have
the class bit clear.  So to avoid flushing kernel ERAT entries on each
context switch, this patch inverts our usage of the class bit, setting
it on user SLB entries and clearing it on kernel SLB entries.

Booted on POWER5 and G5.
Signed-off-by: NDavid Gibson <dwg@au1.ibm.com>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

14b34661

[SPARC64]: Kconfig fix (GEN_RTC dependencies) · c0f2f761

由 Al Viro 提交于 9月 05, 2005

Yet another architecture not coverd by GEN_RTC - sparc64 never picked
it until now and it doesn't have asm/rtc.h to go with it, so it
wouldn't compile anyway (or have these ioctls in the user-visible
headers, for that matter).

FWIW, I'm very tempted to introduce ARCH_HAS_GEN_RTC and have it set
in arch/*/Kconfig for architectures that know what to do with this
stuff - for something supposedly generic the list of architectures
where it doesn't work is getting too long...
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

c0f2f761

[SUNSU]: Compile fixes. · 3d9c9948

由 Al Viro 提交于 9月 05, 2005

sunsu had been broken by ->stop_tx/->start_tx API changes.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

3d9c9948

[SPARC64]: Don't include drivers/firmware/Kconfig · e5e25946

由 David S. Miller 提交于 9月 05, 2005

It's really not relevant for this platform in any
way, after all.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

e5e25946

D
[RTC]: Use SA_SHIRQ in sparc specific code. · 53d0fc27
由 David S. Miller 提交于 9月 05, 2005
```
Based upon a report from Jason Wever.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>
```
53d0fc27

[MOXA]: Fix this driver properly. · 1d25240f

由 Al Viro 提交于 9月 05, 2005

Actually, proper fix of that breakage is embarrassingly simple - it's yet
another gratitious leftover include of asm/segment.h, so incremental to the
previos would be removal of that BROKEN and removal of bogus include from
mxser.c itself.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

1d25240f

D
[IEEE80211]: Use correct size_t printf format string in ieee80211_rx.c · 4c2cac89
由 David S. Miller 提交于 9月 05, 2005
```
Signed-off-by: NDavid S. Miller <davem@davemloft.net>
```
4c2cac89

[PATCH] ppc64: Fix build with oprofile disabled · 0fdf0b86

由 Anton Blanchard 提交于 9月 06, 2005

Fix build with oprofile disabled.
Signed-off-by: NAnton Blanchard <anton@samba.org>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

0fdf0b86

[PATCH] ppc64: Move oprofile_model into cpu feature struct · 8fef0306

由 Anton Blanchard 提交于 9月 06, 2005

Move oprofile_model into cpu feature struct.
Signed-off-by: NAnton Blanchard <anton@samba.org>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

8fef0306

[PATCH] ppc64: Move oprofile_impl.h into include/asm-ppc64 · dca85932

由 Anton Blanchard 提交于 9月 06, 2005

Move oprofile_impl.h into include/asm-ppc64 in preparation for moving
oprofile_model into cpu feature struct.
Signed-off-by: NAnton Blanchard <anton@samba.org>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

dca85932

[PATCH] ppc64: Add oprofile cpu_type to cpu feature struct · 1a410d88

由 Anton Blanchard 提交于 9月 06, 2005

Add oprofile cpu_type to cpu feature struct.
Signed-off-by: NAnton Blanchard <anton@samba.org>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

1a410d88

[PATCH] ppc64: Use num_pmcs in oprofile code · a6908cd0

由 Anton Blanchard 提交于 9月 06, 2005

Change oprofile to use num_pmcs from the cpu feature struct.
Signed-off-by: NAnton Blanchard <anton@samba.org>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

a6908cd0

[PATCH] ppc64: remove CPU_FTR_PMC8 · 8530935d

由 Anton Blanchard 提交于 9月 06, 2005

Remove the CPU_FTR_PMC8 feature now we encode the number of PMCs
directly.
Signed-off-by: NAnton Blanchard <anton@samba.org>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

8530935d

[PATCH] ppc64: add number of PMCs to cputable · fd5b4377

由 Anton Blanchard 提交于 9月 06, 2005

Add a field in the cputable struct to store the number of PMCs.
Signed-off-by: NAnton Blanchard <anton@samba.org>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

fd5b4377

D
[IPW2200]: ipw2200.h needs linux/dma-mapping.h · 3da54c5b
由 David S. Miller 提交于 9月 05, 2005
```
Signed-off-by: NDavid S. Miller <davem@davemloft.net>
```
3da54c5b

[PATCH] ppc64: Allow world readable /proc/ppc64/lparcfg · 71839267

由 Wim Coekaerts 提交于 9月 05, 2005

I would like to be able to read the lparcfg data from any user so we
can make "intelligent" decisions based on underlying attributes when
running in lpars.  Yes there's software that likes to do this :) and
runs as non-root.

It's very similar to say VM where you can get CP to provide feedback
of the real hardware inside a VM guest.
Signed-off-by: NWim Coekaerts <wim.coekaerts@oracle.com>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

71839267

[PATCH] ppc64: remove use of asm/segment.h · fa2259b0

由 Kumar Gala 提交于 8月 24, 2005

Removed PPC64 architecture specific users of asm/segment.h.
Signed-off-by: NKumar Gala <kumar.gala@freescale.com>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

fa2259b0

[PATCH] ppc/ppc64: Merge more include files · 6b9269ab

由 Jon Loeliger 提交于 9月 01, 2005

This patch merges several include files from
asm-ppc and asm-ppc64 into the new asm-powerpc.
Signed-off-by: NJon Loeliger <jdl@freescale.com>
Signed-off-by: NKumar Gala <kumar.gala@freescale.com>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

6b9269ab

[PATCH] Move 3 more headers to asm-powerpc · ad6571a7

由 Becky Bruce 提交于 9月 03, 2005

Merged several nearly-identical header files from asm-ppc and asm-ppc64
into asm-powerpc.
Signed-off-by: NKumar Gala <kumar.gala@freescale.com>
Signed-off-by: NBecky Bruce <becky.bruce@freescale.com>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

ad6571a7

[PATCH] ppc64: speedup cmpxchg · b2c0ab17

由 Anton Blanchard 提交于 9月 06, 2005

cmpxchg has the following code:

__typeof__(*(ptr)) _o_ = (o);
__typeof__(*(ptr)) _n_ = (n);

Unfortunately it makes gcc 4.0 store and load the variables to the stack.
Eg in atomic_dec_and_test we get:

  stw     r10,112(r1)
  stw     r9,116(r1)
  lwz     r9,112(r1)
  lwz     r0,116(r1)

x86 is just casting the values so do that instead. Also change __xchg*
and __cmpxchg* to take unsigned values, removing a few sign extensions.
Signed-off-by: NAnton Blanchard <anton@samba.org>
Signed-off-by: NPaul Mackerras <paulus@samba.org>

b2c0ab17