提交 · 25a4f8b05917f8137bfff8a3f8c6c8c1ac561208 · openanolis / cloud-kernel

07 1月, 2011 2 次提交

EDAC, MCE: Add F15h DC MCE decoder · 25a4f8b0

由 Borislav Petkov 提交于 9月 17, 2010

Add a decoder for F15h DC MCEs to support the new types of DC MCEs
introduced by the BD microarchitecture.
Signed-off-by: NBorislav Petkov <borislav.petkov@amd.com>

25a4f8b0

EDAC, MCE: Select extended error code mask · 2be64bfa

由 Borislav Petkov 提交于 9月 17, 2010

F15h enlarges the extended error code of an MCE to a 5-bit field
(MCi_STATUS[20:16]). Add a mask variable which default 0xf is overridden
on F15h.
Signed-off-by: NBorislav Petkov <borislav.petkov@amd.com>

2be64bfa

09 12月, 2010 3 次提交

amd64_edac: Fix interleaving check · e726f3c3

由 Borislav Petkov 提交于 12月 06, 2010

When matching error address to the range contained by one memory node,
we're in valid range when node interleaving

1. is disabled, or
2. enabled and when the address bits we interleave on match the
interleave selector on this node (see the "Node Interleaving" section in
the BKDG for an enlightening example).

Thus, when we early-exit, we need to reverse the compound logic
statement properly.

Cc: <stable@kernel.org>
Signed-off-by: NBorislav Petkov <borislav.petkov@amd.com>

e726f3c3

EDAC: Correct MiB_TO_PAGES() macro · 76f04f25

由 Andrei Konovalov 提交于 12月 07, 2010

This corrects the misprint introduced when moving '#if
PAGE_SHIFT' from i7core_edac.c to edac_core.h (commit
e9144601)

Cc: Mauro Carvalho Chehab <mchehab@redhat.com>
Signed-off-by: NAndrei Konovalov <akonovalov@mvista.com>
Signed-off-by: NBorislav Petkov <borislav.petkov@amd.com>

76f04f25

EDAC: Fix workqueue-related crashes · bb31b312

由 Borislav Petkov 提交于 12月 02, 2010

00740c58 changed edac_core to
un-/register a workqueue item only if a lowlevel driver supplies a
polling routine. Normally, when we remove a polling low-level driver, we
go and cancel all the queued work. However, the workqueue unreg happens
based on the ->op_state setting, and edac_mc_del_mc() sets this to
OP_OFFLINE _before_ we cancel the work item, leading to NULL ptr oops on
the workqueue list.

Fix it by putting the unreg stuff in proper order.

Cc: <stable@kernel.org> #36.x
Reported-and-tested-by: NTobias Karnat <tobias.karnat@googlemail.com>
LKML-Reference: <1291201307.3029.21.camel@Tobias-Karnat>
Signed-off-by: NBorislav Petkov <borislav.petkov@amd.com>

bb31b312

22 11月, 2010 2 次提交

EDAC, MCE: Fix edac_init_mce_inject error handling · df4b2a30

由 Axel Lin 提交于 11月 18, 2010

Otherwise, variable i will be -1 inside the latest iteration of the
while loop.
Signed-off-by: NAxel Lin <axel.lin@gmail.com>
Signed-off-by: NBorislav Petkov <borislav.petkov@amd.com>

df4b2a30

EDAC: Remove deprecated kbuild goal definitions · f570e1dd

由 Tracey Dent 提交于 11月 07, 2010

Change EDAC's Makefile to use <modules>-y instead of
<modules>-objs because -objs is deprecated and not mentioned in
Documentation/kbuild/makefiles.txt.

 [bp: Fixup commit message]
 [bp: Fixup indentation]
Signed-off-by: NTracey Dent <tdent48227@gmail.com>
Signed-off-by: NBorislav Petkov <borislav.petkov@amd.com>

f570e1dd

24 10月, 2010 33 次提交

i7core_edac: return -ENODEV when devices were already probed · 76a7bd81

由 Mauro Carvalho Chehab 提交于 10月 24, 2010

Due to the nature of i7core, we need to probe and attach all PCI
devices used by this driver during the first time probe is called.
However, PCI core will call the probe routine one time for each CPU
socket. If we return -EINVAL to those calls, it would seem that the
driver fails, when, in fact, there's no more devices left to initialize.

Changing the return code to -ENODEV solves this issue.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

76a7bd81

i7core_edac: properly terminate pci_dev_table · 3c52cc57

由 Mauro Carvalho Chehab 提交于 10月 24, 2010

At pci_xeon_fixup(), it waits for a null-terminated table, while at
i7core_get_all_devices, it just do a for 0..ARRAY_SIZE. As other tables
are zero-terminated, change it to be terminate with 0 as well, and fixes
a bug where it may be running out of the table elements.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

3c52cc57

i7core_edac: Avoid PCI refcount to reach zero on successive load/reload · a3e15416

由 Mauro Carvalho Chehab 提交于 8月 21, 2010

That's a nasty bug that took me a lot of time to track, and whose
solution took just one line to solve. The best fragrances and the worse
poisons are shipped on the smalest bottles.

The drivers/pci/quick.c implements the pci_get_device function. The normal
behavior is that you call it, the function returns you a pdev pointer
and increment pdev->kobj.kref.refcount of the pci device. However,
if you want to keep searching an object, you need to pass the previous
pdev function to the search.

When you use a not null pointer to pdev "from" field, pci_get_device
will decrement pdev->kobj.kref.refcount, assuming that the driver won't
be using the previous pdev.

The solution is simple: we just need to call pci_dev_get() manually,
for the pdev's that the driver will actually use.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

a3e15416

i7core_edac: Fix refcount error at PCI devices · 79daef20

由 Mauro Carvalho Chehab 提交于 8月 21, 2010

Probably due to a bug or some testing logic at PCI level, device
refcount for <bus>:00.0 device is decremented at the end of the
pci_get_device, made by i7core_get_all_devices(). The fact is that
the first versions of the driver relied on those devices to probe
for Nehalem, but the current versions don't use it at all.

So, let's just remove those devices from the driver, making it simpler
and fixing the bug.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

79daef20

i7core_edac: it is safe to i7core_unregister_mci() when mci=NULL · 88ef5ea9

由 Mauro Carvalho Chehab 提交于 8月 20, 2010

i7core_unregister_mci() checks internally when mci=NULL. There's no
need to test it outside.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

88ef5ea9

i7core_edac: Fix an oops at i7core probe · 6d37d240

由 Mauro Carvalho Chehab 提交于 8月 20, 2010

changeset c91d57ba9ce5b5c93a7077e2f72510eb1f9131c4 moved the init
of the priv pointer to the end of the probe routine. However, we need
them before that, otherwise, we hit an OOPS:

[   67.743453] EDAC DEBUG: mci_bind_devs: Associated fn 0.0, dev = ffff88011b46e000, socket 0
[   67.751861] BUG: unable to handle kernel NULL pointer dereference at 0000000000000010
[   67.759685] IP: [<ffffffffa017e484>] i7core_probe+0x979/0x130c [i7core_edac]
[   67.766721] PGD 10bd38067 PUD 10bd37067 PMD 0
[   67.771178] Oops: 0000 [#1] SMP
[   67.774414] last sysfs file: /sys/devices/system/cpu/cpu1/cache/index2/shared_cpu_map
[   67.782213] CPU 1
[   67.784042] Modules linked in: i7core_edac(+) edac_core cpufreq_ondemand binfmt_misc dm_multipath video output pci_slot snd_hda_codd
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

6d37d240

i7core_edac: Remove unused member channels in i7core_pvt · 21b6806a

由 Hidetoshi Seto 提交于 8月 20, 2010

Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

21b6806a

i7core_edac: Remove unused arg csrow from get_dimm_config · 2e5185f7

由 Hidetoshi Seto 提交于 8月 20, 2010

A local is enough.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

2e5185f7

i7core_edac: Reduce args of i7core_register_mci · aace4283

由 Hidetoshi Seto 提交于 8月 20, 2010

We can check the number of channels in i7core_register_mci.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

aace4283

i7core_edac: Introduce i7core_unregister_mci · 1c6edbbe

由 Hidetoshi Seto 提交于 8月 20, 2010

In i7core_probe, when setup of mci for 2nd or later socket failed,
we should cleanup prepared mci for 1st socket or so before "put" of
all devices.

So let have i7core_unregister_mci that can be shared between here
and i7core_remove.

While here fix a typo "hanler".
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

1c6edbbe

i7core_edac: Use saved pointers · 73589c80

由 Hidetoshi Seto 提交于 8月 20, 2010

We already have saved pointers.  Use shorter ones.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

73589c80

i7core_edac: Check probe counter in i7core_remove · 71fe0170

由 Hidetoshi Seto 提交于 8月 20, 2010

Prevent i7core_remove from running multiple times.
Otherwise value proved will be negative and something will be wrong.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

71fe0170

i7core_edac: Call pci_dev_put() when alloc_i7core_dev() failed · 2896637b

由 Hidetoshi Seto 提交于 8月 20, 2010

Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

2896637b

i7core_edac: Fix error path of i7core_register_mci · 628c5ddf

由 Hidetoshi Seto 提交于 8月 20, 2010

Release resources properly.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

628c5ddf

i7core_edac: Fix order of lines in i7core_register_mci · 5939813b

由 Hidetoshi Seto 提交于 8月 20, 2010

The flag is_registered is not initialized until mci_bind_devs()
is called.  Refer it properly.

The mci->dev and mci->edac_check is required in edac_mc_add_mc(),
so prepare them just before the call.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

5939813b

i7core_edac: Always do get/put for all devices · 64c10f6e

由 Hidetoshi Seto 提交于 8月 20, 2010

We already do 'get' for all sockets at once. So do 'put' in the
same way.

And let args of the 'get' function to void since it handles
only the single, static and known size table pci_dev_table[].
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

64c10f6e

i7core_edac: Introduce i7core_pci_ctl_create/release · a3aa0a4a

由 Hidetoshi Seto 提交于 8月 20, 2010

Have a couple of method.
while here sort out lines in the i7core_register_mci() a bit.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

a3aa0a4a

i7core_edac: Introduce free_i7core_dev · 2aa9be44

由 Hidetoshi Seto 提交于 8月 20, 2010

Have a method to make a couple with alloc_i7core_dev() previously
introduced. Using in pair will help proper resource handling.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

2aa9be44

i7core_edac: Introduce alloc_i7core_dev · 848b2f7e

由 Hidetoshi Seto 提交于 8月 20, 2010

It's nice to have a method for a single purpose.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

848b2f7e

i7core_edac: Reduce args of i7core_get_onedevice · b197cba0

由 Hidetoshi Seto 提交于 8月 20, 2010

Since we need to pass the index of the entry, pass the table itself
instead of passing individual members of the table.

While here make it static.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

b197cba0

i7core_edac: Fix the logic in i7core_remove() · 45b7c981

由 Hidetoshi Seto 提交于 8月 20, 2010

commit 47251b4d960bdfa648b0d06dbc6d445f41cb3906 have changed
the logic for unexplained reasons.  It looks strange that it
can release i7core_dev without calling i7core_put_devices()
that releases i7core_dev->pdev.

Fix the part.
Signed-off-by: NHidetoshi Seto <seto.hidetoshi@jp.fujitsu.com>
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

45b7c981

i7core_edac: Don't do the legacy PCI probe by default · 54a08ab1

由 Mauro Carvalho Chehab 提交于 8月 19, 2010

The legacy PCI probe sometimes cause hangs. Better to have it
disabled by default, and have a parameter to enable it.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

54a08ab1

i7core_edac: don't use a freed mci struct · accf74ff

由 Mauro Carvalho Chehab 提交于 8月 16, 2010

This is a nasty bug. Since kobject count will be reduced by zero by
edac_mc_del_mc(), and this triggers the kobj release method, the
mci memory will be freed automatically. So, all we have left is ctl_name,
as shown by enabling debug:

[   80.822186] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 1020: edac_remove_sysfs_mci_device()  remove_link
[   80.832590] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 1024: edac_remove_sysfs_mci_device()  remove_mci_instance
[   80.843776] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 640: edac_mci_control_release() mci instance idx=0 releasing
[   80.855163] EDAC MC: Removed device 0 for i7core_edac.c i7 core #0: DEV 0000:3f:03.0
[   80.862936] EDAC DEBUG: in drivers/edac/i7core_edac.c, line at 2089: (null): free structs
[   80.871134] EDAC DEBUG: in drivers/edac/edac_mc.c, line at 238: edac_mc_free()
[   80.878379] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 726: edac_mc_unregister_sysfs_main_kobj()
[   80.888043] EDAC DEBUG: in drivers/edac/i7core_edac.c, line at 1232: drivers/edac/i7core_edac.c: i7core_put_devices()

Also, kfree(mci) shouldn't happen at the kobj.release, as it happens
when edac_remove_sysfs_mci_device() is called, but the logic is:
	edac_remove_sysfs_mci_device(mci);
	edac_printk(KERN_INFO, EDAC_MC,
		"Removed device %d for %s %s: DEV %s\n", mci->mc_idx,
		mci->mod_name, mci->ctl_name, edac_dev_name(mci));
So, as the edac_printk() needs the mci struct, this generates an OOPS.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

accf74ff

edac_core: Print debug messages at release calls · bbc560ae

由 Mauro Carvalho Chehab 提交于 8月 16, 2010

This is important to track a nasty bug at the free logic.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

bbc560ae

edac_core: Don't let free(mci) happen while using it · ac99768c

由 Mauro Carvalho Chehab 提交于 8月 12, 2010

A very nasty bug were happening on edac core, due to the way mci objects are
freed. mci memory is freed when kobject count reaches zero, by
edac_mci_control_release(). However, from the logs, this is clearly happening
before the final usage of mci struct:

[15799.607454] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 640: edac_mci_control_release() mci instance idx=0 releasing
[15799.618773] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 769: edac_inst_grp_release()
[15799.627326] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 894: edac_remove_mci_instance_attributes() end of seeking for group all_channel_counts
[15799.640887] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 877: edac_remove_mci_instance_attributes() sysfs_attrib = ffffffffa01d7240
[15799.653412] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 1020: edac_remove_sysfs_mci_device()  remove_link
[15799.663753] EDAC DEBUG: in drivers/edac/edac_mc_sysfs.c, line at 1024: edac_remove_sysfs_mci_device()  remove_mci_instance
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

ac99768c

M
edac_core: Do a better job with node removal · 6fe1108f
由 Mauro Carvalho Chehab 提交于 8月 12, 2010
```
Make sure we remove groups at the right order
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>
```
6fe1108f
M
i7core_edac: explicitly remove PCI devices from the devices list · 39300e71
由 Mauro Carvalho Chehab 提交于 8月 11, 2010
```
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>
```
39300e71

i7core_edac: MCE NMI handling should stop first · 41ba6c10

由 Mauro Carvalho Chehab 提交于 8月 11, 2010

Otherwise, a NMI may happen causing a race condition and a panic.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

41ba6c10

M
i7core_edac: Initialize all priv vars before start polling · 6ee7dd50
由 Mauro Carvalho Chehab 提交于 8月 10, 2010
```
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>
```
6ee7dd50
M
i7core_edac: Improve debug to seek for register/remove errors · 3cfd0146
由 Mauro Carvalho Chehab 提交于 8月 10, 2010
```
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>
```
3cfd0146
M
i7core_edac: move #if PAGE_SHIFT to edac_core.h · e9144601
由 Mauro Carvalho Chehab 提交于 8月 10, 2010
```
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>
```
e9144601

i7core_edac: Properly mark const static vars as such · 1288c18f

由 Mauro Carvalho Chehab 提交于 8月 10, 2010

There are two groups of sysfs attributes: one for rdimm and another
for udimm. Instead of changing dynamically the unique static struct
for handling udimm's, declare two vars and make them constant.

This avoids the risk of having two or more memory controllers, each
needing a different set of attributes.

While here, use const on all places where it is applicable.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

edac_core: use const for constant sysfs arguments
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>

1288c18f

M
i7core_edac: move static vars to the beginning of the file · 18c29002
由 Mauro Carvalho Chehab 提交于 8月 10, 2010
```
While here, don't initialize probed with 0.
Signed-off-by: NMauro Carvalho Chehab <mchehab@redhat.com>
```
18c29002

openanolis / cloud-kernel 大约 1 年 前同步成功

openanolis / cloud-kernel
大约 1 年前同步成功