• J
    [PATCH] i386: PARAVIRT: Hooks to set up initial pagetable · b239fb25
    Jeremy Fitzhardinge 提交于
    This patch introduces paravirt_ops hooks to control how the kernel's
    initial pagetable is set up.
    
    In the case of a native boot, the very early bootstrap code creates a
    simple non-PAE pagetable to map the kernel and physical memory.  When
    the VM subsystem is initialized, it creates a proper pagetable which
    respects the PAE mode, large pages, etc.
    
    When booting under a hypervisor, there are many possibilities for what
    paging environment the hypervisor establishes for the guest kernel, so
    the constructon of the kernel's pagetable depends on the hypervisor.
    
    In the case of Xen, the hypervisor boots the kernel with a fully
    constructed pagetable, which is already using PAE if necessary.  Also,
    Xen requires particular care when constructing pagetables to make sure
    all pagetables are always mapped read-only.
    
    In order to make this easier, kernel's initial pagetable construction
    has been changed to only allocate and initialize a pagetable page if
    there's no page already present in the pagetable.  This allows the Xen
    paravirt backend to make a copy of the hypervisor-provided pagetable,
    allowing the kernel to establish any more mappings it needs while
    keeping the existing ones.
    
    A slightly subtle point which is worth highlighting here is that Xen
    requires all kernel mappings to share the same pte_t pages between all
    pagetables, so that updating a kernel page's mapping in one pagetable
    is reflected in all other pagetables.  This makes it possible to
    allocate a page and attach it to a pagetable without having to
    explicitly enumerate that page's mapping in all pagetables.
    
    And:
    
    +From: "Eric W. Biederman" <ebiederm@xmission.com>
    
    If we don't set the leaf page table entries it is quite possible that
    will inherit and incorrect page table entry from the initial boot
    page table setup in head.S.  So we need to redo the effort here,
    so we pick up PSE, PGE and the like.
    
    Hypervisors like Xen require that their page tables be read-only,
    which is slightly incompatible with our low identity mappings, however
    I discussed this with Jeremy he has modified the Xen early set_pte
    function to avoid problems in this area.
    Signed-off-by: NEric W. Biederman <ebiederm@xmission.com>
    Signed-off-by: NJeremy Fitzhardinge <jeremy@xensource.com>
    Signed-off-by: NAndi Kleen <ak@suse.de>
    Acked-by: NWilliam Irwin <bill.irwin@oracle.com>
    Cc: Ingo Molnar <mingo@elte.hu>
    b239fb25
init.c 21.5 KB