提交 · b03efcfb2180289718991bb984044ce6c5b7d1b0 · openeuler / raspberrypi-kernel

09 7月, 2005 1 次提交

[NET]: Transform skb_queue_len() binary tests into skb_queue_empty() · b03efcfb

由 David S. Miller 提交于 7月 08, 2005

This is part of the grand scheme to eliminate the qlen
member of skb_queue_head, and subsequently remove the
'list' member of sk_buff.

Most users of skb_queue_len() want to know if the queue is
empty or not, and that's trivially done with skb_queue_empty()
which doesn't use the skb_queue_head->qlen member and instead
uses the queue list emptyness as the test.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

b03efcfb

08 7月, 2005 1 次提交

[PATCH] coverity: sunrpc/xprt task null check · 7e8d7e3c

由 KAMBAROV, ZAUR 提交于 7月 07, 2005

In __xprt_lock_write() we check to see if `task' is NULL, but in other places
we just go and dereference it.

`task' shouldn't be NULL anyway, so remove this test.

This defect was found automatically by Coverity Prevent, a static analysis
tool.
Signed-off-by: NZaur Kambarov <zkambarov@coverity.com>
Acked-by: NTrond Myklebust <trond.myklebust@fys.uio.no>
Cc: Neil Brown <neilb@cse.unsw.edu.au>
Signed-off-by: NAndrew Morton <akpm@osdl.org>
Signed-off-by: NLinus Torvalds <torvalds@osdl.org>

7e8d7e3c

06 7月, 2005 33 次提交

[TCP]: Never TSO defer under periods of congestion. · 908a75c1

由 David S. Miller 提交于 7月 05, 2005

Congestion window recover after loss depends upon the fact
that if we have a full MSS sized frame at the head of the
send queue, we will send it.  TSO deferral can defeat the
ACK clocking necessary to exit cleanly from recovery.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

908a75c1

[PKT_SCHED]: Blackhole queueing discipline · 63d886c9

由 Thomas Graf 提交于 7月 05, 2005

Useful in combination with classful qdiscs to drop or
temporary disable certain flows, e.g. one could block
specific ds flows with dsmark.

Unlike the noop qdisc it can be controlled by the user and
statistic accounting is done.
Signed-off-by: NThomas Graf <tgraf@suug.ch>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

63d886c9

[TCP]: Move to new TSO segmenting scheme. · c1b4a7e6

由 David S. Miller 提交于 7月 05, 2005

Make TSO segment transmit size decisions at send time not earlier.

The basic scheme is that we try to build as large a TSO frame as
possible when pulling in the user data, but the size of the TSO frame
output to the card is determined at transmit time.

This is guided by tp->xmit_size_goal. It is always set to a multiple
of MSS and tells sendmsg/sendpage how large an SKB to try and build.

Later, tcp_write_xmit() and tcp_push_one() chop up the packet if
necessary and conditions warrant. These routines can also decide to
"defer" in order to wait for more ACKs to arrive and thus allow larger
TSO frames to be emitted.

A general observation is that TSO elongates the pipe, thus requiring a
larger congestion window and larger buffering especially at the sender
side. Therefore, it is important that applications 1) get a large
enough socket send buffer (this is accomplished by our dynamic send
buffer expansion code) 2) do large enough writes.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

c1b4a7e6

[TCP]: Break out send buffer expansion test. · 0d9901df

由 David S. Miller 提交于 7月 05, 2005

This makes it easier to understand, and allows easier
tweaking of the heuristic later on.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

0d9901df

[TCP]: Do not call tcp_tso_acked() if no work to do. · cb83199a

由 David S. Miller 提交于 7月 05, 2005

In tcp_clean_rtx_queue(), if the TSO packet is not even partially
acked, do not waste time calling tcp_tso_acked().
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

cb83199a

[TCP]: Kill bogus comment above tcp_tso_acked(). · a5647696

由 David S. Miller 提交于 7月 05, 2005

Everything stated there is out of data.  tcp_trim_skb()
does adjust the available socket send buffer space and
skb->truesize now.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

a5647696

[TCP]: Fix send-side cpu utiliziation regression. · b4e26f5e

由 David S. Miller 提交于 7月 05, 2005

Only put user data purely to pages when doing TSO.

The extra page allocations cause two problems:

1) Add the overhead of the page allocations themselves.
2) Make us do small user copies when we get to the end
   of the TCP socket cache page.

It is still beneficial to purely use pages for TSO,
so we will do it for that case.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

b4e26f5e

[TCP]: Eliminate redundant computations in tcp_write_xmit(). · aa93466b

由 David S. Miller 提交于 7月 05, 2005

tcp_snd_test() is run for every packet output by a single
call to tcp_write_xmit(), but this is not necessary.

For one, the congestion window space needs to only be
calculated one time, then used throughout the duration
of the loop.

This cleanup also makes experimenting with different TSO
packetization schemes much easier.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

aa93466b

[TCP]: Break out tcp_snd_test() into it's constituent parts. · 7f4dd0a9

由 David S. Miller 提交于 7月 05, 2005

tcp_snd_test() does several different things, use inline
functions to express this more clearly.

1) It initializes the TSO count of SKB, if necessary.
2) It performs the Nagle test.
3) It makes sure the congestion window is adhered to.
4) It makes sure SKB fits into the send window.

This cleanup also sets things up so that things like the
available packets in the congestion window does not need
to be calculated multiple times by packet sending loops
such as tcp_write_xmit().
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

7f4dd0a9

[TCP]: Fix __tcp_push_pending_frames() 'nonagle' handling. · 55c97f3e

由 David S. Miller 提交于 7月 05, 2005

'nonagle' should be passed to the tcp_snd_test() function
as 'TCP_NAGLE_PUSH' if we are checking an SKB not at the
tail of the write_queue.  This is because Nagle does not
apply to such frames since we cannot possibly tack more
data onto them.

However, while doing this __tcp_push_pending_frames() makes
all of the packets in the write_queue use this modified
'nonagle' value.

Fix the bug and simplify this function by just calling
tcp_write_xmit() directly if sk_send_head is non-NULL.

As a result, we can now make tcp_data_snd_check() just call
tcp_push_pending_frames() instead of the specialized
__tcp_data_snd_check().
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

55c97f3e

[TCP]: Fix redundant calculations of tcp_current_mss() · a2e2a59c

由 David S. Miller 提交于 7月 05, 2005

tcp_write_xmit() uses tcp_current_mss(), but some of it's callers,
namely __tcp_push_pending_frames(), already has this value available
already.

While we're here, fix the "cur_mss" argument to be "unsigned int"
instead of plain "unsigned".
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

a2e2a59c

[TCP]: tcp_write_xmit() tabbing cleanup · 92df7b51

由 David S. Miller 提交于 7月 05, 2005

Put the main basic block of work at the top-level of
tabbing, and mark the TCP_CLOSE test with unlikely().
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

92df7b51

[TCP]: Kill extra cwnd validate in __tcp_push_pending_frames(). · a762a980

由 David S. Miller 提交于 7月 05, 2005

The tcp_cwnd_validate() function should only be invoked
if we actually send some frames, yet __tcp_push_pending_frames()
will always invoke it.  tcp_write_xmit() does the call for us,
so the call here can simply be removed.

Also, tcp_write_xmit() can be marked static.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

a762a980

[TCP]: Add missing skb_header_release() call to tcp_fragment(). · f44b5271

由 David S. Miller 提交于 7月 05, 2005

When we add any new packet to the TCP socket write queue,
we must call skb_header_release() on it in order for the
TSO sharing checks in the drivers to work.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

f44b5271

[TCP]: Move __tcp_data_snd_check into tcp_output.c · 84d3e7b9

由 David S. Miller 提交于 7月 05, 2005

It reimplements portions of tcp_snd_check(), so it
we move it to tcp_output.c we can consolidate it's
logic much easier in a later change.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

84d3e7b9

[TCP]: Move send test logic out of net/tcp.h · f6302d1d

由 David S. Miller 提交于 7月 05, 2005

This just moves the code into tcp_output.c, no code logic changes are
made by this patch.

Using this as a baseline, we can begin to untangle the mess of
comparisons for the Nagle test et al.  We will also be able to reduce
all of the redundant computation that occurs when outputting data
packets.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

f6302d1d

[TCP]: Fix quick-ack decrementing with TSO. · fc6415bc

由 David S. Miller 提交于 7月 05, 2005

On each packet output, we call tcp_dec_quickack_mode()
if the ACK flag is set.  It drops tp->ack.quick until
it hits zero, at which time we deflate the ATO value.

When doing TSO, we are emitting multiple packets with
ACK set, so we should decrement tp->ack.quick that many
segments.

Note that, unlike this case, tcp_enter_cwr() should not
take the tcp_skb_pcount(skb) into consideration.  That
function, one time, readjusts tp->snd_cwnd and moves
into TCP_CA_CWR state.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

fc6415bc

[TCP]: Simplify SKB data portion allocation with NETIF_F_SG. · c65f7f00

由 David S. Miller 提交于 7月 05, 2005

The ideal and most optimal layout for an SKB when doing
scatter-gather is to put all the headers at skb->data, and
all the user data in the page array.

This makes SKB splitting and combining extremely simple,
especially before a packet goes onto the wire the first
time.

So, when sk_stream_alloc_pskb() is given a zero size, make
sure there is no skb_tailroom().  This is achieved by applying
SKB_DATA_ALIGN() to the header length used here.

Next, make select_size() in TCP output segmentation use a
length of zero when NETIF_F_SG is true on the outgoing
interface.
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

c65f7f00

[NET]: improve readability of dev_set_promiscuity() in net/core/dev.c · 52609c0b

由 David Chau 提交于 7月 05, 2005

A trivial patch to improve the readability of dev_set_promiscuity()
in net/core/dev.c. New code does exactly the same thing as original
code.
Signed-off-by: NDavid Chau <ddcc@mit.edu>
Signed-off-by: NDomen Puncer <domen@coderock.org>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

52609c0b

[IPV4]: More broken memory allocation fixes for fib_trie · 2f36895a

由 Robert Olsson 提交于 7月 05, 2005

Below a patch to preallocate memory when doing resize of trie (inflate halve)
If preallocations fails it just skips the resize of this tnode for this time.

The oops we got when killing bgpd (with full routing) is now gone. 
Patrick memory patch is also used.
Signed-off-by: NRobert Olsson <robert.olsson@its.uu.se>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

2f36895a

T
[DECNET]: Fix memset overflow on 64bit archs while dumping decnet routing rules · db1322b8
由 Thomas Graf 提交于 7月 05, 2005
```
Signed-off-by: NThomas Graf <tgraf@suug.ch>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>
```
db1322b8

[IPV4]: Bug fix in rt_check_expire() · bb1d23b0

由 Eric Dumazet 提交于 7月 05, 2005

- rt_check_expire() fixes (an overflow occured if size of the hash
  was >= 65536)

reminder of the bugfix:

The rt_check_expire() has a serious problem on machines with large
route caches, and a standard HZ value of 1000.

With default values, ie ip_rt_gc_interval = 60*HZ = 60000 ;

the loop count :

     for (t = ip_rt_gc_interval << rt_hash_log; t >= 0;


overflows (t is a 31 bit value) as soon rt_hash_log is >= 16  (65536
slots in route cache hash table).

In this case, rt_check_expire() does nothing at all
Signed-off-by: NEric Dumazet <dada1@cosmosbay.com>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

bb1d23b0

[IPV4]: Use the fancy alloc_large_system_hash() function for route hash table · 424c4b70

由 Eric Dumazet 提交于 7月 05, 2005

- rt hash table allocated using alloc_large_system_hash() function.
Signed-off-by: NEric Dumazet <dada1@cosmosbay.com>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

424c4b70

[NET]: Hashed spinlocks in net/ipv4/route.c · 22c047cc

由 Eric Dumazet 提交于 7月 05, 2005

- Locking abstraction
- Spinlocks moved out of rt hash table : Less memory (50%) used by rt 
  hash table. it's a win even on UP.
- Sizing of spinlocks table depends on NR_CPUS
Signed-off-by: NEric Dumazet <dada1@cosmosbay.com>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

22c047cc

[IPV4]: Handle large allocations in fib_trie · f0e36f8c

由 Patrick McHardy 提交于 7月 05, 2005

Inflating a node a couple of times makes it exceed the 128k kmalloc limit.
Use __get_free_pages for allocations > PAGE_SIZE, as in fib_hash.
Signed-off-by: NPatrick McHardy <kaber@trash.net>
Acked-by: NRobert Olsson <Robert.Olsson@data.slu.se>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

f0e36f8c

H
[IPV6]: Makes IPv6 rcv registration happen last during initialisation. · e2ed4052
由 Herbert Xu 提交于 7月 05, 2005
```
Signed-off-by: NHerbert Xu <herbert@gondor.apana.org.au>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>
```
e2ed4052

[IPV4]: Fix crash in ip_rcv while booting related to netconsole · 30e224d7

由 Herbert Xu 提交于 7月 05, 2005

Makes IPv4 ip_rcv registration happen last in af_inet.
Signed-off-by: NHerbert Xu <herbert@gondor.apana.org.au>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

30e224d7

[PKT_SCHED]: Report rate estimator configuration errors during qdisc allocation · 023e09a7

由 Thomas Graf 提交于 7月 05, 2005

Current behaviour is to not report an error if a rate
estimator is created together with a qdisc and the
configuration of the rate estimator is bogus. This leads
to unexpected behaviour because the user is not notified.

New behaviour is to report the error and let the whole
qdisc creation operation fail so the user is able to fix
his mistake.
Signed-off-by: NThomas Graf <tgraf@suug.ch>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

023e09a7

[PKT_SCHED]: Cleanup qdisc creation and alignment macros · 3d54b82f

由 Thomas Graf 提交于 7月 05, 2005

Adds qdisc_alloc() to share code between qdisc_create()
and qdisc_create_dflt(). Hides the qdisc alignment behind
macros and makes use of them.
Signed-off-by: NThomas Graf <tgraf@suug.ch>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

3d54b82f

T
[NET]: Remove unused security member in sk_buff · e176fe89
由 Thomas Graf 提交于 7月 05, 2005
```
Signed-off-by: NThomas Graf <tgraf@suug.ch>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>
```
e176fe89

[NET]: net/core/filter.c: make len cover the entire packet · 3154e540

由 Patrick McHardy 提交于 7月 05, 2005

As suggested by Herbert Xu:

Since we don't require anything to be in the linear packet range
anymore make len cover the entire packet.
Signed-off-by: NPatrick McHardy <kaber@trash.net>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

3154e540

P
[NET]: Consolidate common code in net/core/filter.c · 0b05b2a4
由 Patrick McHardy 提交于 7月 05, 2005
```
Signed-off-by: NPatrick McHardy <kaber@trash.net>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>
```
0b05b2a4

[NET]: Remove redundant code in net/core/filter.c · 6935d46c

由 Patrick McHardy 提交于 7月 05, 2005

skb_header_pointer handles linear and non-linear data, no need to handle
linear data again.
Signed-off-by: NPatrick McHardy <kaber@trash.net>
Acked-by: NHerbert Xu <herbert@gondor.apana.org.au>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

6935d46c

29 6月, 2005 5 次提交

[NETFILTER]: Fix connection tracking bug in 2.6.12 · 9666dae5

由 Patrick McHardy 提交于 6月 28, 2005

In 2.6.12 we started dropping the conntrack reference when a packet
leaves the IP layer. This broke connection tracking on a bridge,
because bridge-netfilter defers calling some NF_IP_* hooks to the bridge
layer for locally generated packets going out a bridge, where the
conntrack reference is no longer available. This patch keeps the
reference in this case as a temporary solution, long term we will
remove the defered hook calling. No attempt is made to drop the
reference in the bridge-code when it is no longer needed, tc actions
could already have sent the packet anywhere.
Signed-off-by: NPatrick McHardy <kaber@trash.net>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

9666dae5

[NET]: Micro optimization in eth_header() · ff593c59

由 Denis Vlasenko 提交于 6月 28, 2005

Signed-off-by: NDenis Vlasenko <vda@ilport.com.ua>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

ff593c59

[IPV6]: remove more unused IPV6_AUTHHDR things. · 7fe40f73

由 YOSHIFUJI Hideaki 提交于 6月 28, 2005

Remove two more unused IPV6_AUTHHDR option things, 
which I failed to remove them last time,
plus, mark IPV6_AUTHHDR obsolete.
Signed-off-by: NYOSHIFUJI Hideaki <yoshfuji@linux-ipv6.org>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

7fe40f73

[IPVS]: Close race conditions on ip_vs_conn_tab list modification · fb3d8949

由 Neil Horman 提交于 6月 28, 2005

In an smp system, it is possible for an connection timer to expire, calling
ip_vs_conn_expire while the connection table is being flushed, before
ct_write_lock_bh is acquired.

Since the list iterator loop in ip_vs_con_flush releases and re-acquires the
spinlock (even though it doesn't re-enable softirqs), it is possible for the
expiration function to modify the connection list, while it is being traversed
in ip_vs_conn_flush.

The result is that the next pointer gets set to NULL, and subsequently
dereferenced, resulting in an oops.
Signed-off-by: NNeil Horman <nhorman@redhat.com>
Acked-by: JulianAnastasov
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

fb3d8949

[IPV4]: Broken memory allocation in fib_trie · f835e471

由 Robert Olsson 提交于 6月 28, 2005

This should help up the insertion... but the resize is more crucial.
and complex and needs some thinking. 
Signed-off-by: NRobert Olsson <robert.olsson@its.uu.se>
Signed-off-by: NDavid S. Miller <davem@davemloft.net>

f835e471