提交 · 66a8cb95ed04025664d1db4e952155ee1dccd048 · openeuler / Kernel

01 4月, 2010 1 次提交

ring-buffer: Add place holder recording of dropped events · 66a8cb95

由 Steven Rostedt 提交于 14年前

Currently, when the ring buffer drops events, it does not record
the fact that it did so. It does inform the writer that the event
was dropped by returning a NULL event, but it does not put in any
place holder where the event was dropped.

This is not a trivial thing to add because the ring buffer mostly
runs in overwrite (flight recorder) mode. That is, when the ring
buffer is full, new data will overwrite old data.

In a produce/consumer mode, where new data is simply dropped when
the ring buffer is full, it is trivial to add the placeholder
for dropped events. When there's more room to write new data, then
a special event can be added to notify the reader about the dropped
events.

But in overwrite mode, any new write can overwrite events. A place
holder can not be inserted into the ring buffer since there never
may be room. A reader could also come in at anytime and miss the
placeholder.

Luckily, the way the ring buffer works, the read side can find out
if events were lost or not, and how many events. Everytime a write
takes place, if it overwrites the header page (the next read) it
updates a "overrun" variable that keeps track of the number of
lost events. When a reader swaps out a page from the ring buffer,
it can record this number, perfom the swap, and then check to
see if the number changed, and take the diff if it has, which would be
the number of events dropped. This can be stored by the reader
and returned to callers of the reader.

Since the reader page swap will fail if the writer moved the head
page since the time the reader page set up the swap, this gives room
to record the overruns without worrying about races. If the reader
sets up the pages, records the overrun, than performs the swap,
if the swap succeeds, then the overrun variable has not been
updated since the setup before the swap.

For binary readers of the ring buffer, a flag is set in the header
of each sub page (sub buffer) of the ring buffer. This flag is embedded
in the size field of the data on the sub buffer, in the 31st bit (the size
can be 32 or 64 bits depending on the architecture), but only 27
bits needs to be used for the actual size (less actually).

We could add a new field in the sub buffer header to also record the
number of events dropped since the last read, but this will change the
format of the binary ring buffer a bit too much. Perhaps this change can
be made if the information on the number of events dropped is considered
important enough.

Note, the notification of dropped events is only used by consuming reads
or peeking at the ring buffer. Iterating over the ring buffer does not
keep this information because the necessary data is only available when
a page swap is made, and the iterator does not swap out pages.

Cc: Robert Richter <robert.richter@amd.com>
Cc: Andi Kleen <andi@firstfloor.org>
Cc: Li Zefan <lizf@cn.fujitsu.com>
Cc: Arnaldo Carvalho de Melo <acme@redhat.com>
Cc: "Luis Claudio R. Goncalves" <lclaudio@uudg.org>
Cc: Frederic Weisbecker <fweisbec@gmail.com>
Signed-off-by: NSteven Rostedt <rostedt@goodmis.org>

66a8cb95

29 10月, 2009 1 次提交

percpu: make percpu symbols in oprofile unique · b3e9f672

由 Tejun Heo 提交于 15年前

This patch updates percpu related symbols in oprofile such that percpu
symbols are unique and don't clash with local symbols.  This serves
two purposes of decreasing the possibility of global percpu symbol
collision and allowing dropping per_cpu__ prefix from percpu symbols.

* drivers/oprofile/cpu_buffer.c: s/cpu_buffer/op_cpu_buffer/

Partly based on Rusty Russell's "alloc_percpu: rename percpu vars
which cause name clashes" patch.
Signed-off-by: NTejun Heo <tj@kernel.org>
Acked-by: NRobert Richter <robert.richter@amd.com>
Cc: Rusty Russell <rusty@rustcorp.com.au>

b3e9f672

10 10月, 2009 2 次提交

oprofile: warn on freeing event buffer too early · c0868934

由 Robert Richter 提交于 15年前

A race shouldn't happen since all workqueues or handlers are canceled
or flushed before the event buffer is freed. A warning is triggered
now if the buffer is freed too early.

Also, this patch adds some comments about event buffer protection,
reworks some code and adds code to clear buffer_pos during alloc and
free of the event buffer.

Cc: David Rientjes <rientjes@google.com>
Cc: Stephane Eranian <eranian@google.com>
Signed-off-by: NRobert Richter <robert.richter@amd.com>

c0868934

oprofile: fix race condition in event_buffer free · 066b3aa8

由 David Rientjes 提交于 15年前

Looking at the 2.6.31-rc9 code, it appears there is a race condition
in the event_buffer cleanup code path (shutdown). This could lead to
kernel panic as some CPUs may be operating on the event buffer AFTER
it has been freed. The attached patch solves the problem and makes
sure CPUs check if the buffer is not NULL before they access it as
some may have been spinning on the mutex while the buffer was being
freed.

The race may happen if the buffer is freed during pending reads. But
it is not clear why there are races in add_event_entry() since all
workqueues or handlers are canceled or flushed before the event buffer
is freed.
Signed-off-by: NDavid Rientjes <rientjes@google.com>
Signed-off-by: NStephane Eranian <eranian@google.com>
Signed-off-by: NRobert Richter <robert.richter@amd.com>

066b3aa8

24 9月, 2009 1 次提交

cpumask: use zalloc_cpumask_var() where possible · 79f55997

由 Li Zefan 提交于 15年前

Remove open-coded zalloc_cpumask_var() and zalloc_cpumask_var_node().
Signed-off-by: NLi Zefan <lizf@cn.fujitsu.com>
Signed-off-by: NRusty Russell <rusty@rustcorp.com.au>

79f55997

22 9月, 2009 1 次提交

const: mark remaining super_operations const · b87221de

由 Alexey Dobriyan 提交于 15年前

Signed-off-by: NAlexey Dobriyan <adobriyan@gmail.com>
Signed-off-by: NAndrew Morton <akpm@linux-foundation.org>
Signed-off-by: NLinus Torvalds <torvalds@linux-foundation.org>

b87221de

20 7月, 2009 6 次提交

oprofile: Adding switch counter to oprofile statistic variables · 1b294f59

由 Robert Richter 提交于 15年前

This patch moves the multiplexing switch counter from x86 code to
common oprofile statistic variables. Now the value will be available
and usable for all architectures. The initialization and
incrementation also moved to common code.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

1b294f59

oprofile: Grouping multiplexing code in oprof.c · a5659d17

由 Robert Richter 提交于 15年前

This patch moves multiplexing code to a single section of code. This
reduces the use of #ifdefs especially within functions.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

a5659d17

oprofile: Remove oprofile_multiplexing_init() · 16422a6e

由 Robert Richter 提交于 15年前

oprofile_multiplexing_init() can be removed when moving the
initialization of oprofile_time_slice to oprofile_create_files().
Signed-off-by: NRobert Richter <robert.richter@amd.com>

16422a6e

oprofile: Rename variable timeout_jiffies and move to oprofile_files.c · afe1b50f

由 Robert Richter 提交于 15年前

This patch renames timeout_jiffies into an oprofile specific name. The
macro MULTIPLEXING_TIMER_DEFAULT is changed too.

Also, since this variable is controlled using oprofilefs, its
definition is moved to oprofile_files.c.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

afe1b50f

oprofile: oprofile_set_timeout(), return with error for invalid args · 2051cade

由 Robert Richter 提交于 15年前

Return with -EINVAL for invalid parameters instead of setting the
default value in oprofile_set_timeout().
Signed-off-by: NRobert Richter <robert.richter@amd.com>

2051cade

oprofile: Implement performance counter multiplexing · 4d4036e0

由 Jason Yeh 提交于 15年前

The number of hardware counters is limited. The multiplexing feature
enables OProfile to gather more events than counters are provided by
the hardware. This is realized by switching between events at an user
specified time interval.

A new file (/dev/oprofile/time_slice) is added for the user to specify
the timer interval in ms. If the number of events to profile is higher
than the number of hardware counters available, the patch will
schedule a work queue that switches the event counter and re-writes
the different sets of values into it. The switching mechanism needs to
be implemented for each architecture to support multiplexing. This
patch only implements AMD CPU support, but multiplexing can be easily
extended for other models and architectures.

There are follow-on patches that rework parts of this patch.
Signed-off-by: NJason Yeh <jason.yeh@amd.com>
Signed-off-by: NRobert Richter <robert.richter@amd.com>

4d4036e0

10 7月, 2009 1 次提交

oprofile: reset bt_lost_no_mapping with other stats · 2b8777ca

由 Maynard Johnson 提交于 15年前

The bt_lost_no_mapping is not getting reset at the start of a
profiling run, thus the oprofiled.log shows erroneous values for this
statistic. The attached patch fixes this problem.
Signed-off-by: NMaynard Johnson <maynardj@us.ibm.com>
Signed-off-by: NRobert Richter <robert.richter@amd.com>
Signed-off-by: NIngo Molnar <mingo@elte.hu>

2b8777ca

12 6月, 2009 2 次提交

oprofile: reset bt_lost_no_mapping with other stats · 1cc4ce6f

由 Maynard Johnson 提交于 15年前

The bt_lost_no_mapping is not getting reset at the start of a
profiling run, thus the oprofiled.log shows erroneous values for this
statistic. The attached patch fixes this problem.
Signed-off-by: NMaynard Johnson <maynardj@us.ibm.com>
Signed-off-by: NRobert Richter <robert.richter@amd.com>

1cc4ce6f

x86/oprofile: introduce oprofile_add_data64() · 51563a0e

由 Robert Richter 提交于 15年前

The IBS implemention writes 64 bit register values to the cpu buffer
by writing two 32 values using oprofile_add_data(). This patch
introduces oprofile_add_data64() to write a single 64 bit value to the
buffer.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

51563a0e

11 6月, 2009 1 次提交

oprofile: remove obselete include headers · fecfe632

由 Robert Richter 提交于 15年前

This became obsolete with this commit:

 6dad828b oprofile: port to the new ring_buffer
Signed-off-by: NRobert Richter <robert.richter@amd.com>

fecfe632

07 5月, 2009 1 次提交

oprofile: fix cpu buffer size · 54f2c841

由 Robert Richter 提交于 15年前

The unit of oprofile_cpu_buffer_size is in samples, but was allocated
in bytes. This led to the allocation of too small cpu buffers. This
patch recalculates the buffer size in bytes taking also the
ring_buffer_event header size into account.
Reported-by: NSuravee Suthikulpanit <suravee.suthikulpanit@amd.com>
Signed-off-by: NRobert Richter <robert.richter@amd.com>

54f2c841

30 3月, 2009 1 次提交

oprofile: Thou shalt not call __exit functions from __init functions · bb75efdd

由 Russell King 提交于 15年前

Impact: fix ref to discarded function

`buffer_sync_cleanup' referenced in section `.init.text' of arch/arm/oprofile/built-in.o: defined in discarded section `.exit.text' of arch/arm/oprofile/built-in.o
Signed-off-by: NRussell King <rmk+kernel@arm.linux.org.uk>
Signed-off-by: NRusty Russell <rusty@rustcorp.com.au>

bb75efdd

06 2月, 2009 1 次提交
- I
  ring_buffer: remove unused flags parameter, fix · 304cc6ae
  由 Ingo Molnar 提交于 16年前
```
Oprofile's ring-buffer use was not considered.
Signed-off-by: NIngo Molnar <mingo@elte.hu>
```
  304cc6ae
22 1月, 2009 1 次提交

cpumask: modifiy oprofile initialization · 4c50d9ea

由 Robert Richter 提交于 16年前

Delta patch to f7df8ed1 for
tip/cpus4096.

Moved initialization to sync_start()/sync_stop(). No changes needed in
buffer_sync.h and oprof.c anymore.
Signed-off-by: NRobert Richter <robert.richter@amd.com>
Signed-off-by: NIngo Molnar <mingo@elte.hu>

4c50d9ea

18 1月, 2009 1 次提交

oprofile: fix uninitialized use of struct op_entry · fdb6a8f4

由 Robert Richter 提交于 16年前

Impact: fix crash

In case of losing samples struct op_entry could have been used
uninitialized causing e.g. a wrong preemption count or NULL pointer
access. This patch fixes this.
Signed-off-by: NRobert Richter <robert.richter@amd.com>
Signed-off-by: NIngo Molnar <mingo@elte.hu>

fdb6a8f4

12 1月, 2009 1 次提交

cpumask: convert misc driver functions · f7df8ed1

由 Rusty Russell 提交于 16年前

Impact: use new cpumask API.

Convert misc driver functions to use struct cpumask.

To Do:
  - Convert iucv_buffer_cpumask to cpumask_var_t.
Signed-off-by: NRusty Russell <rusty@rustcorp.com.au>
Signed-off-by: NMike Travis <travis@sgi.com>
Acked-by: NDean Nelson <dcn@sgi.com>
Cc: Robert Richter <robert.richter@amd.com>
Cc: oprofile-list@lists.sf.net
Cc: Jeremy Fitzhardinge <jeremy@xensource.com>
Cc: Chris Wright <chrisw@sous-sol.org>
Cc: virtualization@lists.osdl.org
Cc: xen-devel@lists.xensource.com
Cc: Ursula Braun <ursula.braun@de.ibm.com>
Cc: linux390@de.ibm.com
Cc: linux-s390@vger.kernel.org

f7df8ed1

08 1月, 2009 15 次提交

oprofile: make new cpu buffer functions part of the api · 14f0ca8e

由 Robert Richter 提交于 16年前

This patch creates the new functions

 oprofile_write_reserve()
 oprofile_add_data()
 oprofile_write_commit()

and makes them part of the oprofile api.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

14f0ca8e

oprofile: remove #ifdef CONFIG_OPROFILE_IBS in non-ibs code · ebf8d974

由 Robert Richter 提交于 16年前

The ifdefs can be removed since the code is no longer ibs specific and
can be used for other purposes as well. IBS specific code is only in
op_model_amd.c.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

ebf8d974

oprofile: use new data sample format for ibs · 1acda878

由 Robert Richter 提交于 16年前

The new ring buffer implementation allows the storage of samples with
different size. This patch implements the usage of the new sample
format to store ibs samples in the cpu buffer. Until now, writing to
the cpu buffer could lead to incomplete sampling sequences since IBS
samples were transfered in multiple samples. Due to a full buffer,
data could be lost at any time. This can't happen any more since the
complete data is reserved in advance and then stored in a single
sample.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

1acda878

oprofile: add op_cpu_buffer_get_data() · bd7dc46f

由 Robert Richter 提交于 16年前

This function provides access to attached data of a sample. It returns
the size of data including the current value. Also,
op_cpu_buffer_get_size() is available to check if there is data
attached.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

bd7dc46f

oprofile: add op_cpu_buffer_add_data() · d9928c25

由 Robert Richter 提交于 16年前

This function can be used to attach data to a sample. It returns the
remaining free buffer size that has been reserved with
op_cpu_buffer_write_reserve().
Signed-off-by: NRobert Richter <robert.richter@amd.com>

d9928c25

oprofile: rework implementation of cpu buffer events · ae735e99

由 Robert Richter 提交于 16年前

Special events such as task or context switches are marked with an
escape code in the cpu buffer followed by an event code or a task
identifier. There is one escape code per event. To make escape
sequences also available for data samples the internal cpu buffer
format must be changed. The current implementation does not allow the
extension of event codes since this would lead to collisions with the
task identifiers. To avoid this, this patch introduces an event mask
that allows the storage of multiple events with one escape code. Now,
task identifiers are stored in the data section of the sample. The
implementation also allows the usage of custom data in a sample. As a
side effect the new code is much more readable and easier to
understand.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

ae735e99

oprofile: modify op_cpu_buffer_read_entry() · 2d87b14c

由 Robert Richter 提交于 16年前

This implements the support of samples with attached data.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

2d87b14c

oprofile: add op_cpu_buffer_write_reserve() · 2cc28b9f

由 Robert Richter 提交于 16年前

This function prepares the cpu buffer to write a sample.

Struct op_entry is used during operations on the ring buffer while
struct op_sample contains the data that is stored in the ring
buffer. Struct entry can be uninitialized. The function reserves a
data array that is specified by size. Use op_cpu_buffer_write_commit()
after preparing the sample. In case of errors a null pointer is
returned, otherwise the pointer to the sample.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

2cc28b9f

oprofile: rename variables in add_ibs_begin() · d358e75f

由 Robert Richter 提交于 16年前

This unifies usage of variable names within oprofile.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

d358e75f

oprofile: rename add_sample() in cpu_buffer.c · d0e23384

由 Robert Richter 提交于 16年前

Rename the fucntion to op_add_sample() since there is a collision with
another one with the same name in buffer_sync.c.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

d0e23384

R
oprofile: making add_sample_entry() inline · 6368a1f4
由 Robert Richter 提交于 16年前
```
Signed-off-by: NRobert Richter <robert.richter@amd.com>
```
6368a1f4

oprofile: remove backtrace code for ibs · 8350c787

由 Robert Richter 提交于 16年前

This code is broken since a TRACE_BEGIN_CODE is never sent to the
daemon. The data becomes corrupt since the backtrace is interpreted as
ibs sample.
Signed-off-by: NRobert Richter <robert.richter@amd.com>

8350c787

R
oprofile: remove unused ibs macro · f4ff2364
由 Robert Richter 提交于 16年前
```
Signed-off-by: NRobert Richter <robert.richter@amd.com>
```
f4ff2364
R
oprofile: remove unused components in struct oprofile_cpu_buffer · 8d15df84
由 Robert Richter 提交于 16年前
```
Signed-off-by: NRobert Richter <robert.richter@amd.com>
```
8d15df84
R
oprofile: simplify add_ibs_begin() · dbe6e283
由 Robert Richter 提交于 16年前
```
Signed-off-by: NRobert Richter <robert.richter@amd.com>
```
dbe6e283

06 1月, 2009 1 次提交

zero i_uid/i_gid on inode allocation · 56ff5efa

由 Al Viro 提交于 16年前

... and don't bother in callers.  Don't bother with zeroing i_blocks,
while we are at it - it's already been zeroed.

i_mode is not worth the effort; it has no common default value.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

56ff5efa

01 1月, 2009 1 次提交

shrink struct dentry · c2452f32

由 Nick Piggin 提交于 16年前

struct dentry is one of the most critical structures in the kernel. So it's
sad to see it going neglected.

With CONFIG_PROFILING turned on (which is probably the common case at least
for distros and kernel developers), sizeof(struct dcache) == 208 here
(64-bit). This gives 19 objects per slab.

I packed d_mounted into a hole, and took another 4 bytes off the inline
name length to take the padding out from the end of the structure. This
shinks it to 200 bytes. I could have gone the other way and increased the
length to 40, but I'm aiming for a magic number, read on...

I then got rid of the d_cookie pointer. This shrinks it to 192 bytes. Rant:
why was this ever a good idea? The cookie system should increase its hash
size or use a tree or something if lookups are a problem. Also the "fast
dcookie lookups" in oprofile should be moved into the dcookie code -- how
can oprofile possibly care about the dcookie_mutex? It gets dropped after
get_dcookie() returns so it can't be providing any sort of protection.

At 192 bytes, 21 objects fit into a 4K page, saving about 3MB on my system
with ~140 000 entries allocated. 192 is also a multiple of 64, so we get
nice cacheline alignment on 64 and 32 byte line systems -- any given dentry
will now require 3 cachelines to touch all fields wheras previously it
would require 4.

I know the inline name size was chosen quite carefully, however with the
reduction in cacheline footprint, it should actually be just about as fast
to do a name lookup for a 36 character name as it was before the patch (and
faster for other sizes). The memory footprint savings for names which are
<= 32 or > 36 bytes long should more than make up for the memory cost for
33-36 byte names.

Performance is a feature...
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

c2452f32

30 12月, 2008 1 次提交
- R
  oprofile: simplify add_sample() in cpu_buffer.c · 3967e93e
  由 Robert Richter 提交于 16年前
```
Signed-off-by: NRobert Richter <robert.richter@amd.com>
```
  3967e93e

openeuler / Kernel 1 年多 前同步成功

openeuler / Kernel
1 年多前同步成功