提交 · 597b0d21626da4e6f09f132442caf0cc2b0eb47c · openanolis / cloud-kernel

31 12月, 2008 40 次提交

KVM: MMU: handle large host sptes on invlpg/resync · 87917239

由 Marcelo Tosatti 提交于 12月 22, 2008

The invlpg and sync walkers lack knowledge of large host sptes,
descending to non-existant pagetable level.

Stop at directory level in such case.

Fixes SMP Windows XP with hugepages.
Signed-off-by: NMarcelo Tosatti <mtosatti@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

87917239

KVM: Add locking to virtual i8259 interrupt controller · 3f353858

由 Avi Kivity 提交于 12月 21, 2008

While most accesses to the i8259 are with the kvm mutex taken, the call
to kvm_pic_read_irq() is not. We can't easily take the kvm mutex there
since the function is called with interrupts disabled.

Fix by adding a spinlock to the virtual interrupt controller. Since we
can't send an IPI under the spinlock (we also take the same spinlock in
an irq disabled context), we defer the IPI until the spinlock is released.
Similarly, we defer irq ack notifications until after spinlock release to
avoid lock recursion.
Signed-off-by: NAvi Kivity <avi@redhat.com>

3f353858

KVM: MMU: Don't treat a global pte as such if cr4.pge is cleared · 25e23432

由 Avi Kivity 提交于 12月 21, 2008

The pte.g bit is meaningless if global pages are disabled; deferring
mmu page synchronization on these ptes will lead to the guest using stale
shadow ptes.

Fixes Vista x86 smp bootloader failure.
Signed-off-by: NAvi Kivity <avi@redhat.com>

25e23432

KVM: x86: Rework user space NMI injection as KVM_CAP_USER_NMI · 4531220b

由 Jan Kiszka 提交于 12月 11, 2008

There is no point in doing the ready_for_nmi_injection/
request_nmi_window dance with user space. First, we don't do this for
in-kernel irqchip anyway, while the code path is the same as for user
space irqchip mode. And second, there is nothing to loose if a pending
NMI is overwritten by another one (in contrast to IRQs where we have to
save the number). Actually, there is even the risk of raising spurious
NMIs this way because the reason for the held-back NMI might already be
handled while processing the first one.

Therefore this patch creates a simplified user space NMI injection
interface, exporting it under KVM_CAP_USER_NMI and dropping the old
KVM_CAP_NMI capability. And this time we also take care to provide the
interface only on archs supporting NMIs via KVM (right now only x86).
Signed-off-by: NJan Kiszka <jan.kiszka@siemens.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

4531220b

KVM: VMX: Fix pending NMI-vs.-IRQ race for user space irqchip · 264ff01d

由 Jan Kiszka 提交于 11月 24, 2008

As with the kernel irqchip, don't allow an NMI to stomp over an already
injected IRQ; instead wait for the IRQ injection to be completed.
Signed-off-by: NJan Kiszka <jan.kiszka@siemens.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

264ff01d

KVM: MMU: check for present pdptr shadow page in walk_shadow · eb64f1e8

由 Marcelo Tosatti 提交于 12月 09, 2008

walk_shadow assumes the caller verified validity of the pdptr pointer in
question, which is not the case for the invlpg handler.

Fixes oops during Solaris 10 install.
Signed-off-by: NMarcelo Tosatti <mtosatti@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

eb64f1e8

A
KVM: Consolidate userspace memory capability reporting into common code · ca9edaee
由 Avi Kivity 提交于 12月 08, 2008
```
Signed-off-by: NAvi Kivity <avi@redhat.com>
```
ca9edaee

KVM: MMU: prepopulate the shadow on invlpg · ad218f85

由 Marcelo Tosatti 提交于 12月 01, 2008

If the guest executes invlpg, peek into the pagetable and attempt to
prepopulate the shadow entry.

Also stop dirty fault updates from interfering with the fork detector.

2% improvement on RHEL3/AIM7.
Signed-off-by: NMarcelo Tosatti <mtosatti@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

ad218f85

KVM: MMU: skip global pgtables on sync due to cr3 switch · 6cffe8ca

由 Marcelo Tosatti 提交于 12月 01, 2008

Skip syncing global pages on cr3 switch (but not on cr4/cr0). This is
important for Linux 32-bit guests with PAE, where the kmap page is
marked as global.
Signed-off-by: NMarcelo Tosatti <mtosatti@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

6cffe8ca

KVM: MMU: collapse remote TLB flushes on root sync · b1a36821

由 Marcelo Tosatti 提交于 12月 01, 2008

Collapse remote TLB flushes on root sync.

kernbench is 2.7% faster on 4-way guest. Improvements have been seen
with other loads such as AIM7.
Signed-off-by: NMarcelo Tosatti <mtosatti@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

b1a36821

KVM: MMU: use page array in unsync walk · 60c8aec6

由 Marcelo Tosatti 提交于 12月 01, 2008

Instead of invoking the handler directly collect pages into
an array so the caller can work with it.

Simplifies TLB flush collapsing.
Signed-off-by: NMarcelo Tosatti <mtosatti@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

60c8aec6

KVM: x86 emulator: Fix handling of VMMCALL instruction · fbce554e

由 Amit Shah 提交于 12月 04, 2008

The VMMCALL instruction doesn't get recognised and isn't processed
by the emulator.

This is seen on an Intel host that tries to execute the VMMCALL
instruction after a guest live migrates from an AMD host.
Signed-off-by: NAmit Shah <amit.shah@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

fbce554e

KVM: x86 emulator: add the emulation of shld and shrd instructions · 9bf8ea42

由 Guillaume Thouvenin 提交于 12月 04, 2008

Add emulation of shld and shrd instructions
Signed-off-by: NGuillaume Thouvenin <guillaume.thouvenin@ext.bull.net>
Signed-off-by: NAvi Kivity <avi@redhat.com>

9bf8ea42

KVM: x86 emulator: add the assembler code for three operands · d175226a

由 Guillaume Thouvenin 提交于 12月 04, 2008

Add the assembler code for instruction with three operands and one
operand is stored in ECX register
Signed-off-by: NGuillaume Thouvenin <guillaume.thouvenin@ext.bull.net>
Signed-off-by: NAvi Kivity <avi@redhat.com>

d175226a

KVM: x86 emulator: add a new "implied 1" Src decode type · bfcadf83

由 Guillaume Thouvenin 提交于 12月 04, 2008

Add SrcOne operand type when we need to decode an implied '1' like with
regular shift instruction
Signed-off-by: NGuillaume Thouvenin <guillaume.thouvenin@ext.bull.net>
Signed-off-by: NAvi Kivity <avi@redhat.com>

bfcadf83

KVM: x86 emulator: add Src2 decode set · 0dc8d10f

由 Guillaume Thouvenin 提交于 12月 04, 2008

Instruction like shld has three operands, so we need to add a Src2
decode set. We start with Src2None, Src2CL, and Src2ImmByte, Src2One to
support shld/shrd and we will expand it later.
Signed-off-by: NGuillaume Thouvenin <guillaume.thouvenin@ext.bull.net>
Signed-off-by: NAvi Kivity <avi@redhat.com>

0dc8d10f

KVM: x86 emulator: Extend the opcode descriptor · 45ed60b3

由 Guillaume Thouvenin 提交于 12月 04, 2008

Extend the opcode descriptor to 32 bits. This is needed by the
introduction of a new Src2 operand type.
Signed-off-by: NGuillaume Thouvenin <guillaume.thouvenin@ext.bull.net>
Signed-off-by: NAvi Kivity <avi@redhat.com>

45ed60b3

KVM: VMX: fix sparse warning · efff9e53

由 Hannes Eder 提交于 11月 28, 2008

Impact: make global function static

  arch/x86/kvm/vmx.c:134:3: warning: symbol 'vmx_capability' was not declared. Should it be static?
Signed-off-by: NHannes Eder <hannes@hanneseder.net>
Signed-off-by: NAvi Kivity <avi@redhat.com>

efff9e53

A
KVM: Remove extraneous semicolon after do/while · f3fd92fb
由 Avi Kivity 提交于 11月 29, 2008
```
Notices by Guillaume Thouvenin.
Signed-off-by: NAvi Kivity <avi@redhat.com>
```
f3fd92fb

KVM: x86 emulator: fix popf emulation · 2b48cc75

由 Avi Kivity 提交于 11月 29, 2008

Set operand type and size to get correct writeback behavior.
Signed-off-by: NAvi Kivity <avi@redhat.com>

2b48cc75

KVM: x86 emulator: fix ret emulation · cf5de4f8

由 Avi Kivity 提交于 11月 28, 2008

'ret' did not set the operand type or size for the destination, so
writeback ignored it.
Signed-off-by: NAvi Kivity <avi@redhat.com>

cf5de4f8

A
KVM: x86 emulator: switch 'pop reg' instruction to emulate_pop() · 8a09b687
由 Avi Kivity 提交于 11月 27, 2008
```
Signed-off-by: NAvi Kivity <avi@redhat.com>
```
8a09b687
A
KVM: x86 emulator: allow pop from mmio · 781d0edc
由 Avi Kivity 提交于 11月 27, 2008
```
Signed-off-by: NAvi Kivity <avi@redhat.com>
```
781d0edc
A
KVM: x86 emulator: Extract 'pop' sequence into a function · faa5a3ae
由 Avi Kivity 提交于 11月 27, 2008
```
Switch 'pop r/m' instruction to use the new function.
Signed-off-by: NAvi Kivity <avi@redhat.com>
```
faa5a3ae
A
KVM: x86 emulator: consolidate emulation of two operand instructions · 6b7ad61f
由 Avi Kivity 提交于 11月 26, 2008
```
No need to repeat the same assembly block over and over.
Signed-off-by: NAvi Kivity <avi@redhat.com>
```
6b7ad61f
A
KVM: x86 emulator: reduce duplication in one operand emulation thunks · dda96d8f
由 Avi Kivity 提交于 11月 26, 2008
```
Signed-off-by: NAvi Kivity <avi@redhat.com>
```
dda96d8f

KVM: MMU: optimize set_spte for page sync · ecc5589f

由 Marcelo Tosatti 提交于 11月 25, 2008

The write protect verification in set_spte is unnecessary for page sync.

Its guaranteed that, if the unsync spte was writable, the target page
does not have a write protected shadow (if it had, the spte would have
been write protected under mmu_lock by rmap_write_protect before).

Same reasoning applies to mark_page_dirty: the gfn has been marked as
dirty via the pagefault path.

The cost of hash table and memslot lookups are quite significant if the
workload is pagetable write intensive resulting in increased mmu_lock
contention.
Signed-off-by: NMarcelo Tosatti <mtosatti@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

ecc5589f

KVM: VMX: Conditionally request interrupt window after injecting irq · df203ec9

由 Avi Kivity 提交于 11月 23, 2008

If we're injecting an interrupt, and another one is pending, request
an interrupt window notification so we don't have excess latency on the
second interrupt.

This shouldn't happen in practice since an EOI will be issued, giving a second
chance to request an interrupt window, but...
Signed-off-by: NAvi Kivity <avi@redhat.com>

df203ec9

KVM: SVM: move svm_hardware_disable() code to asm/virtext.h · 2c8dceeb

由 Eduardo Habkost 提交于 11月 17, 2008

Create cpu_svm_disable() function.
Signed-off-by: NEduardo Habkost <ehabkost@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

2c8dceeb

KVM: SVM: move has_svm() code to asm/virtext.h · 63d1142f

由 Eduardo Habkost 提交于 11月 17, 2008

Use a trick to keep the printk()s on has_svm() working as before. gcc
will take care of not generating code for the 'msg' stuff when the
function is called with a NULL msg argument.
Signed-off-by: NEduardo Habkost <ehabkost@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

63d1142f

KVM: VMX: extract kvm_cpu_vmxoff() from hardware_disable() · 710ff4a8

由 Eduardo Habkost 提交于 11月 17, 2008

Along with some comments on why it is different from the core cpu_vmxoff()
function.
Signed-off-by: NEduardo Habkost <ehabkost@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

710ff4a8

KVM: VMX: move cpu_has_kvm_support() to an inline on asm/virtext.h · 6210e37b

由 Eduardo Habkost 提交于 11月 17, 2008

It will be used by core code on kdump and reboot, to disable
vmx if needed.
Signed-off-by: NEduardo Habkost <ehabkost@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

6210e37b

KVM: SVM: move svm.h to include/asm · c2cedf7b

由 Eduardo Habkost 提交于 11月 17, 2008

svm.h will be used by core code that is independent of KVM, so I am
moving it outside the arch/x86/kvm directory.
Signed-off-by: NEduardo Habkost <ehabkost@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

c2cedf7b

KVM: VMX: move vmx.h to include/asm · 13673a90

由 Eduardo Habkost 提交于 11月 17, 2008

vmx.h will be used by core code that is independent of KVM, so I am
moving it outside the arch/x86/kvm directory.
Signed-off-by: NEduardo Habkost <ehabkost@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

13673a90

KVM: Fix cpuid iteration on multiple leaves per eac · 0fdf8e59

由 Nitin A Kamble 提交于 11月 05, 2008

The code to traverse the cpuid data array list for counting type of leaves is
currently broken.

This patches fixes the 2 things in it.

 1. Set the 1st counting entry's flag KVM_CPUID_FLAG_STATE_READ_NEXT. Without
    it the code will never find a valid entry.

 2. Also the stop condition in the for loop while looking for the next unflaged
    entry is broken. It needs to stop when it find one matching entry;
    and in the case of count of 1, it will be the same entry found in this
    iteration.
Signed-Off-By: NNitin A Kamble <nitin.a.kamble@intel.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

0fdf8e59

KVM: Fix cpuid leaf 0xb loop termination · 0853d2c1

由 Nitin A Kamble 提交于 11月 05, 2008

For cpuid leaf 0xb the bits 8-15 in ECX register define the end of counting
leaf.      The previous code was using bits 0-7 for this purpose, which is
a bug.
Signed-off-by: NNitin A Kamble <nitin.a.kamble@intel.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

0853d2c1

KVM: MMU: Fix aliased gfns treated as unaliased · 2843099f

由 Izik Eidus 提交于 10月 03, 2008

Some areas of kvm x86 mmu are using gfn offset inside a slot without
unaliasing the gfn first.  This patch makes sure that the gfn will be
unaliased and add gfn_to_memslot_unaliased() to save the calculating
of the gfn unaliasing in case we have it unaliased already.
Signed-off-by: NIzik Eidus <ieidus@redhat.com>
Acked-by: NMarcelo Tosatti <mtosatti@redhat.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

2843099f

KVM: Enable Function Level Reset for assigned device · 6eb55818

由 Sheng Yang 提交于 10月 31, 2008

Ideally, every assigned device should in a clear condition before and after
assignment, so that the former state of device won't affect later work.
Some devices provide a mechanism named Function Level Reset, which is
defined in PCI/PCI-e document. We should execute it before and after device
assignment.

(But sadly, the feature is new, and most device on the market now don't
support it. We are considering using D0/D3hot transmit to emulate it later,
but not that elegant and reliable as FLR itself.)

[Update: Reminded by Xiantao, execute FLR after we ensure that the device can
be assigned to the guest.]
Signed-off-by: NSheng Yang <sheng@linux.intel.com>
Signed-off-by: NAvi Kivity <avi@redhat.com>

6eb55818

KVM: VMX: Handle mmio emulation when guest state is invalid · 1d5a4d9b

由 Guillaume Thouvenin 提交于 10月 29, 2008

If emulate_invalid_guest_state is enabled, the emulator is called
when guest state is invalid.  Until now, we reported an mmio failure
when emulate_instruction() returned EMULATE_DO_MMIO.  This patch adds
the case where emulate_instruction() failed and an MMIO emulation
is needed.
Signed-off-by: NGuillaume Thouvenin <guillaume.thouvenin@ext.bull.net>
Signed-off-by: NAvi Kivity <avi@redhat.com>

1d5a4d9b

KVM: allow emulator to adjust rip for emulated pio instructions · e93f36bc

由 Guillaume Thouvenin 提交于 10月 28, 2008

If we call the emulator we shouldn't call skip_emulated_instruction()
in the first place, since the emulator already computes the next rip
for us. Thus we move ->skip_emulated_instruction() out of
kvm_emulate_pio() and into handle_io() (and the svm equivalent). We
also replaced "return 0" by "break" in the "do_io:" case because now
the shadow register state needs to be committed. Otherwise eip will never
be updated.
Signed-off-by: NGuillaume Thouvenin <guillaume.thouvenin@ext.bull.net>
Signed-off-by: NAvi Kivity <avi@redhat.com>

e93f36bc

openanolis / cloud-kernel 接近 2 年 前同步成功

openanolis / cloud-kernel
接近 2 年前同步成功