[RFC,v5,078/104] KVM: TDX: Implement interrupt injection

Message ID	776d48b5c88ebf189ffac1eb94ef190bfc7210da.1646422845.git.isaku.yamahata@intel.com (mailing list archive)
State	New, archived
Headers	show Return-Path: <kvm-owner@kernel.org> From: isaku.yamahata@intel.com To: kvm@vger.kernel.org, linux-kernel@vger.kernel.org Cc: isaku.yamahata@intel.com, isaku.yamahata@gmail.com, Paolo Bonzini <pbonzini@redhat.com>, Jim Mattson <jmattson@google.com>, erdemaktas@google.com, Connor Kuehl <ckuehl@redhat.com>, Sean Christopherson <seanjc@google.com> Subject: [RFC PATCH v5 078/104] KVM: TDX: Implement interrupt injection Date: Fri, 4 Mar 2022 11:49:34 -0800 Message-Id: <776d48b5c88ebf189ffac1eb94ef190bfc7210da.1646422845.git.isaku.yamahata@intel.com> In-Reply-To: <cover.1646422845.git.isaku.yamahata@intel.com> References: <cover.1646422845.git.isaku.yamahata@intel.com> MIME-Version: 1.0 Content-Transfer-Encoding: 8bit Precedence: bulk
Series	KVM TDX basic feature support \| expand [RFC,v5,000/104] KVM TDX basic feature support [RFC,v5,001/104] KVM: VMX: Move out vmx_x86_ops to 'main.c' to wrap VMX and TDX [RFC,v5,002/104] x86/virt/tdx: export platform_has_tdx [RFC,v5,003/104] KVM: TDX: Detect CPU feature on kernel module initialization [RFC,v5,004/104] KVM: Enable hardware before doing arch VM initialization [RFC,v5,005/104] KVM: x86: Refactor KVM VMX module init/exit functions [RFC,v5,006/104] KVM: TDX: Add placeholders for TDX VM/vcpu structure [RFC,v5,007/104] x86/virt/tdx: Add a helper function to return system wide info about TDX module [RFC,v5,008/104] KVM: TDX: Add a function to initialize TDX module [RFC,v5,009/104] KVM: x86: Introduce vm_type to differentiate default VMs from confidential VMs [RFC,v5,010/104] KVM: TDX: Make TDX VM type supported [RFC,v5,011/104,MARKER] The start of TDX KVM patch series: TDX architectural definitions [RFC,v5,012/104] KVM: TDX: Define TDX architectural definitions [RFC,v5,013/104] KVM: TDX: Add TDX "architectural" error codes [RFC,v5,014/104] KVM: TDX: Add a function for KVM to invoke SEAMCALL [RFC,v5,015/104] KVM: TDX: add a helper function for KVM to issue SEAMCALL [RFC,v5,016/104] KVM: TDX: Add C wrapper functions for SEAMCALLs to the TDX module [RFC,v5,017/104] KVM: TDX: Add helper functions to print TDX SEAMCALL error [RFC,v5,018/104,MARKER] The start of TDX KVM patch series: TD VM creation/destruction [RFC,v5,019/104] KVM: TDX: Stub in tdx.h with structs, accessors, and VMCS helpers [RFC,v5,020/104] KVM: TDX: allocate per-package mutex [RFC,v5,021/104] KVM: x86: Introduce hooks to free VM callback prezap and vm_free [RFC,v5,022/104] KVM: Add max_vcpus field in common 'struct kvm' [RFC,v5,023/104] x86/cpu: Add helper functions to allocate/free MKTME keyid [RFC,v5,024/104] KVM: TDX: create/destroy VM structure [RFC,v5,025/104] KVM: TDX: Add place holder for TDX VM specific mem_enc_op ioctl [RFC,v5,026/104] KVM: TDX: x86: Add vm ioctl to get TDX systemwide parameters [RFC,v5,027/104] KVM: TDX: initialize VM with TDX specific parameters [RFC,v5,028/104,MARKER] The start of TDX KVM patch series: TD vcpu creation/destruction [RFC,v5,029/104] KVM: TDX: allocate/free TDX vcpu structure [RFC,v5,030/104] KVM: TDX: Do TDX specific vcpu initialization [RFC,v5,031/104,MARKER] The start of TDX KVM patch series: KVM MMU GPA stolen bits [RFC,v5,032/104] KVM: x86/mmu: introduce config for PRIVATE KVM MMU [RFC,v5,033/104] KVM: x86: Add infrastructure for stolen GPA bits [RFC,v5,034/104,MARKER] The start of TDX KVM patch series: KVM TDP refactoring for TDX [RFC,v5,035/104] KVM: x86/mmu: Disallow dirty logging for x86 TDX [RFC,v5,036/104] KVM: x86/mmu: Explicitly check for MMIO spte in fast page fault [RFC,v5,037/104] KVM: x86/mmu: Allow non-zero init value for shadow PTE [RFC,v5,038/104] KVM: x86/mmu: Allow per-VM override of the TDP max page level [RFC,v5,039/104] KVM: x86/mmu: Disallow fast page fault on private GPA [RFC,v5,040/104] KVM: VMX: Split out guts of EPT violation to common/exposed function [RFC,v5,041/104] KVM: VMX: Move setting of EPT MMU masks to common VT-x code [RFC,v5,042/104] KVM: x86/mmu: Track shadow MMIO value/mask on a per-VM basis [RFC,v5,043/104] KVM: TDX: Add load_mmu_pgd method for TDX [RFC,v5,044/104,MARKER] The start of TDX KVM patch series: KVM TDP MMU hooks [RFC,v5,045/104] KVM: x86/tdp_mmu: make REMOVED_SPTE include shadow_initial value [RFC,v5,046/104] KVM: x86/tdp_mmu: refactor kvm_tdp_mmu_map() [RFC,v5,047/104] KVM: x86/mmu: add a private pointer to struct kvm_mmu_page [RFC,v5,048/104] KVM: x86/tdp_mmu: Support TDX private mapping for TDP MMU [RFC,v5,049/104] KVM: x86/tdp_mmu: Ignore unsupported mmu operation on private GFNs [RFC,v5,050/104,MARKER] The start of TDX KVM patch series: TDX EPT violation [RFC,v5,051/104] KVM: TDX: TDP MMU TDX support [RFC,v5,052/104,MARKER] The start of TDX KVM patch series: KVM TDP MMU MapGPA [RFC,v5,053/104] KVM: x86/mmu: steal software usable bit for EPT to represent shared page [RFC,v5,054/104] KVM: x86/tdp_mmu: Keep PRIVATE_PROHIBIT bit when zapping [RFC,v5,055/104] KVM: x86/tdp_mmu: prevent private/shared map based on PRIVATE_PROHIBIT [RFC,v5,056/104] KVM: x86/tdp_mmu: implement MapGPA hypercall for TDX [RFC,v5,057/104] KVM: x86/mmu: Introduce kvm_mmu_map_tdp_page() for use by TDX [RFC,v5,058/104] KVM: x86/mmu: Focibly use TDP MMU for TDX [RFC,v5,059/104,MARKER] The start of TDX KVM patch series: TD finalization [RFC,v5,060/104] KVM: TDX: Create initial guest memory [RFC,v5,061/104] KVM: TDX: Finalize VM initialization [RFC,v5,062/104,MARKER] The start of TDX KVM patch series: TD vcpu enter/exit [RFC,v5,063/104] KVM: TDX: Add helper assembly function to TDX vcpu [RFC,v5,064/104] KVM: TDX: Implement TDX vcpu enter/exit path [RFC,v5,065/104] KVM: TDX: vcpu_run: save/restore host state(host kernel gs) [RFC,v5,066/104] KVM: TDX: restore host xsave state when exit from the guest TD [RFC,v5,067/104] KVM: x86: Allow to update cached values in kvm_user_return_msrs w/o wrmsr [RFC,v5,068/104] KVM: TDX: restore user ret MSRs [RFC,v5,069/104,MARKER] The start of TDX KVM patch series: TD vcpu exits/interrupts/hypercalls [RFC,v5,070/104] KVM: TDX: complete interrupts after tdexit [RFC,v5,071/104] KVM: TDX: restore debug store when TD exit [RFC,v5,072/104] KVM: TDX: handle vcpu migration over logical processor [RFC,v5,073/104] KVM: TDX: track LP tdx vcpu run and teardown vcpus on descroing the guest TD [RFC,v5,074/104] KVM: x86: Add a switch_db_regs flag to handle TDX's auto-switched behavior [RFC,v5,075/104] KVM: x86: Check for pending APICv interrupt in kvm_vcpu_has_events() [RFC,v5,076/104] KVM: x86: Add option to force LAPIC expiration wait [RFC,v5,077/104] KVM: TDX: Use vcpu_to_pi_desc() uniformly in posted_intr.c [RFC,v5,078/104] KVM: TDX: Implement interrupt injection [RFC,v5,079/104] KVM: TDX: Implements vcpu request_immediate_exit [RFC,v5,080/104] KVM: TDX: Implement methods to inject NMI [RFC,v5,081/104] KVM: VMX: Modify NMI and INTR handlers to take intr_info as function argument [RFC,v5,082/104] KVM: VMX: Move NMI/exception handler to common helper [RFC,v5,083/104] KVM: x86: Split core of hypercall emulation to helper function [RFC,v5,084/104] KVM: TDX: Add a place holder to handle TDX VM exit [RFC,v5,085/104] KVM: TDX: handle EXIT_REASON_OTHER_SMI [RFC,v5,086/104] KVM: TDX: handle ept violation/misconfig exit [RFC,v5,087/104] KVM: TDX: handle EXCEPTION_NMI and EXTERNAL_INTERRUPT [RFC,v5,088/104] KVM: TDX: Add TDG.VP.VMCALL accessors to access guest vcpu registers [RFC,v5,089/104] KVM: TDX: Add a placeholder for handler of TDX hypercalls (TDG.VP.VMCALL) [RFC,v5,090/104] KVM: TDX: handle KVM hypercall with TDG.VP.VMCALL [RFC,v5,091/104] KVM: TDX: Handle TDX PV CPUID hypercall [RFC,v5,092/104] KVM: TDX: Handle TDX PV HLT hypercall [RFC,v5,093/104] KVM: TDX: Handle TDX PV port io hypercall [RFC,v5,094/104] KVM: TDX: Handle TDX PV MMIO hypercall [RFC,v5,095/104] KVM: TDX: Implement callbacks for MSR operations for TDX [RFC,v5,096/104] KVM: TDX: Handle TDX PV rdmsr hypercall [RFC,v5,097/104] KVM: TDX: Handle TDX PV wrmsr hypercall [RFC,v5,098/104] KVM: TDX: Handle TDX PV report fatal error hypercall [RFC,v5,099/104] KVM: TDX: Handle TDX PV map_gpa hypercall [RFC,v5,100/104] KVM: TDX: Silently discard SMI request [RFC,v5,101/104] KVM: TDX: Silently ignore INIT/SIPI [RFC,v5,102/104] KVM: TDX: Add methods to ignore accesses to CPU state [RFC,v5,103/104] Documentation/virtual/kvm: Document on Trust Domain Extensions(TDX) [RFC,v5,104/104] KVM: x86: design documentation on TDX support of x86 KVM TDP MMU

diff --git a/arch/x86/kvm/vmx/common.h b/arch/x86/kvm/vmx/common.h index 1052b3c93eb8..79a4517e43d1 100644 --- a/arch/x86/kvm/vmx/common.h +++ b/arch/x86/kvm/vmx/common.h @@ -4,6 +4,7 @@ #include <linux/kvm_host.h> +#include "posted_intr.h" #include "mmu.h" static inline int __vmx_handle_ept_violation(struct kvm_vcpu *vcpu, gpa_t gpa, @@ -32,4 +33,73 @@ static inline int __vmx_handle_ept_violation(struct kvm_vcpu *vcpu, gpa_t gpa, return kvm_mmu_page_fault(vcpu, gpa, error_code, NULL, 0); } +static inline void kvm_vcpu_trigger_posted_interrupt(struct kvm_vcpu *vcpu, + int pi_vec) +{ +#ifdef CONFIG_SMP + if (vcpu->mode == IN_GUEST_MODE) { + /* + * The vector of interrupt to be delivered to vcpu had + * been set in PIR before this function. + * + * Following cases will be reached in this block, and + * we always send a notification event in all cases as + * explained below. + * + * Case 1: vcpu keeps in non-root mode. Sending a + * notification event posts the interrupt to vcpu. + * + * Case 2: vcpu exits to root mode and is still + * runnable. PIR will be synced to vIRR before the + * next vcpu entry. Sending a notification event in + * this case has no effect, as vcpu is not in root + * mode. + * + * Case 3: vcpu exits to root mode and is blocked. + * vcpu_block() has already synced PIR to vIRR and + * never blocks vcpu if vIRR is not cleared. Therefore, + * a blocked vcpu here does not wait for any requested + * interrupts in PIR, and sending a notification event + * which has no effect is safe here. + */ + + apic->send_IPI_mask(get_cpu_mask(vcpu->cpu), pi_vec); + return; + } +#endif + /* + * The vCPU isn't in the guest; wake the vCPU in case it is blocking, + * otherwise do nothing as KVM will grab the highest priority pending + * IRQ via ->sync_pir_to_irr() in vcpu_enter_guest(). + */ + kvm_vcpu_wake_up(vcpu); +} + +/* + * Send interrupt to vcpu via posted interrupt way. + * 1. If target vcpu is running(non-root mode), send posted interrupt + * notification to vcpu and hardware will sync PIR to vIRR atomically. + * 2. If target vcpu isn't running(root mode), kick it to pick up the + * interrupt from PIR in next vmentry. + */ +static inline void __vmx_deliver_posted_interrupt( + struct kvm_vcpu *vcpu, struct pi_desc *pi_desc, int vector) +{ + if (pi_test_and_set_pir(vector, pi_desc)) + return; + + /* If a previous notification has sent the IPI, nothing to do. */ + if (pi_test_and_set_on(pi_desc)) + return; + + /* + * The implied barrier in pi_test_and_set_on() pairs with the smp_mb_*() + * after setting vcpu->mode in vcpu_enter_guest(), thus the vCPU is + * guaranteed to see PID.ON=1 and sync the PIR to IRR if triggering a + * posted interrupt "fails" because vcpu->mode != IN_GUEST_MODE. + */ + kvm_vcpu_trigger_posted_interrupt(vcpu, POSTED_INTR_VECTOR); +} + + #endif /* __KVM_X86_VMX_COMMON_H */ diff --git a/arch/x86/kvm/vmx/main.c b/arch/x86/kvm/vmx/main.c index d75caf0d6861..a0bcc4dca678 100644 --- a/arch/x86/kvm/vmx/main.c +++ b/arch/x86/kvm/vmx/main.c @@ -148,6 +148,34 @@ static void vt_vcpu_load(struct kvm_vcpu *vcpu, int cpu) return vmx_vcpu_load(vcpu, cpu); } +static void vt_apicv_post_state_restore(struct kvm_vcpu *vcpu) +{ + if (is_td_vcpu(vcpu)) + return tdx_apicv_post_state_restore(vcpu); + + return vmx_apicv_post_state_restore(vcpu); +} + +static int vt_sync_pir_to_irr(struct kvm_vcpu *vcpu) +{ + if (is_td_vcpu(vcpu)) + return -1; + + return vmx_sync_pir_to_irr(vcpu); +} + +static void vt_deliver_interrupt(struct kvm_lapic *apic, int delivery_mode, + int trig_mode, int vector) +{ + if (is_td_vcpu(apic->vcpu)) { + tdx_deliver_interrupt(apic, delivery_mode, trig_mode, + vector); + return; + } + + vmx_deliver_interrupt(apic, delivery_mode, trig_mode, vector); +} + static bool vt_apicv_has_pending_interrupt(struct kvm_vcpu *vcpu) { if (is_td_vcpu(vcpu)) @@ -205,6 +233,53 @@ static void vt_sched_in(struct kvm_vcpu *vcpu, int cpu) vmx_sched_in(vcpu, cpu); } +static void vt_set_interrupt_shadow(struct kvm_vcpu *vcpu, int mask) +{ + if (is_td_vcpu(vcpu)) + return; + vmx_set_interrupt_shadow(vcpu, mask); +} + +static u32 vt_get_interrupt_shadow(struct kvm_vcpu *vcpu) +{ + if (is_td_vcpu(vcpu)) + return 0; + + return vmx_get_interrupt_shadow(vcpu); +} + +static void vt_inject_irq(struct kvm_vcpu *vcpu) +{ + if (is_td_vcpu(vcpu)) + return; + + vmx_inject_irq(vcpu); +} + +static void vt_cancel_injection(struct kvm_vcpu *vcpu) +{ + if (is_td_vcpu(vcpu)) + return; + + vmx_cancel_injection(vcpu); +} + +static int vt_interrupt_allowed(struct kvm_vcpu *vcpu, bool for_injection) +{ + if (is_td_vcpu(vcpu)) + return true; + + return vmx_interrupt_allowed(vcpu, for_injection); +} + +static void vt_enable_irq_window(struct kvm_vcpu *vcpu) +{ + if (is_td_vcpu(vcpu)) + return; + + vmx_enable_irq_window(vcpu); +} + static int vt_mem_enc_op(struct kvm *kvm, void __user *argp) { if (!is_td(kvm)) @@ -279,31 +354,31 @@ struct kvm_x86_ops vt_x86_ops __initdata = { .handle_exit = vmx_handle_exit, .skip_emulated_instruction = vmx_skip_emulated_instruction, .update_emulated_instruction = vmx_update_emulated_instruction, - .set_interrupt_shadow = vmx_set_interrupt_shadow, - .get_interrupt_shadow = vmx_get_interrupt_shadow, + .set_interrupt_shadow = vt_set_interrupt_shadow, + .get_interrupt_shadow = vt_get_interrupt_shadow, .patch_hypercall = vmx_patch_hypercall, - .set_irq = vmx_inject_irq, + .set_irq = vt_inject_irq, .set_nmi = vmx_inject_nmi, .queue_exception = vmx_queue_exception, - .cancel_injection = vmx_cancel_injection, - .interrupt_allowed = vmx_interrupt_allowed, + .cancel_injection = vt_cancel_injection, + .interrupt_allowed = vt_interrupt_allowed, .nmi_allowed = vmx_nmi_allowed, .get_nmi_mask = vmx_get_nmi_mask, .set_nmi_mask = vmx_set_nmi_mask, .enable_nmi_window = vmx_enable_nmi_window, - .enable_irq_window = vmx_enable_irq_window, + .enable_irq_window = vt_enable_irq_window, .update_cr8_intercept = vmx_update_cr8_intercept, .set_virtual_apic_mode = vmx_set_virtual_apic_mode, .set_apic_access_page_addr = vmx_set_apic_access_page_addr, .refresh_apicv_exec_ctrl = vmx_refresh_apicv_exec_ctrl, .load_eoi_exitmap = vmx_load_eoi_exitmap, - .apicv_post_state_restore = vmx_apicv_post_state_restore, + .apicv_post_state_restore = vt_apicv_post_state_restore, .check_apicv_inhibit_reasons = vmx_check_apicv_inhibit_reasons, .hwapic_irr_update = vmx_hwapic_irr_update, .hwapic_isr_update = vmx_hwapic_isr_update, .guest_apic_has_interrupt = vmx_guest_apic_has_interrupt, - .sync_pir_to_irr = vmx_sync_pir_to_irr, - .deliver_interrupt = vmx_deliver_interrupt, + .sync_pir_to_irr = vt_sync_pir_to_irr, + .deliver_interrupt = vt_deliver_interrupt, .dy_apicv_has_pending_interrupt = pi_has_pending_interrupt, .apicv_has_pending_interrupt = vt_apicv_has_pending_interrupt, diff --git a/arch/x86/kvm/vmx/posted_intr.c b/arch/x86/kvm/vmx/posted_intr.c index c8a81c916eed..e22c3015f064 100644 --- a/arch/x86/kvm/vmx/posted_intr.c +++ b/arch/x86/kvm/vmx/posted_intr.c @@ -7,6 +7,7 @@ #include "lapic.h" #include "irq.h" #include "posted_intr.h" +#include "tdx.h" #include "trace.h" #include "vmx.h" @@ -31,6 +32,11 @@ static DEFINE_PER_CPU(raw_spinlock_t, wakeup_vcpus_on_cpu_lock); static inline struct pi_desc *vcpu_to_pi_desc(struct kvm_vcpu *vcpu) { +#ifdef CONFIG_INTEL_TDX_HOST + if (is_td_vcpu(vcpu)) + return &(to_tdx(vcpu)->pi_desc); +#endif + return &(to_vmx(vcpu)->pi_desc); } diff --git a/arch/x86/kvm/vmx/tdx.c b/arch/x86/kvm/vmx/tdx.c index 3a0e826fbe0c..bdc658ca9e4f 100644 --- a/arch/x86/kvm/vmx/tdx.c +++ b/arch/x86/kvm/vmx/tdx.c @@ -7,6 +7,7 @@ #include "capabilities.h" #include "x86_ops.h" +#include "common.h" #include "mmu.h" #include "tdx.h" #include "vmx.h" @@ -494,6 +495,9 @@ int tdx_vcpu_create(struct kvm_vcpu *vcpu) vcpu->arch.guest_state_protected = !(to_kvm_tdx(vcpu->kvm)->attributes & TDX_TD_ATTRIBUTE_DEBUG); + tdx->pi_desc.nv = POSTED_INTR_VECTOR; + tdx->pi_desc.sn = 1; + tdx->host_state_need_save = true; tdx->host_state_need_restore = false; @@ -514,6 +518,7 @@ void tdx_vcpu_load(struct kvm_vcpu *vcpu, int cpu) { struct vcpu_tdx *tdx = to_tdx(vcpu); + vmx_vcpu_pi_load(vcpu, cpu); if (vcpu->cpu == cpu) return; @@ -735,6 +740,12 @@ fastpath_t tdx_vcpu_run(struct kvm_vcpu *vcpu) trace_kvm_entry(vcpu); + if (pi_test_on(&tdx->pi_desc)) { + apic->send_IPI_self(POSTED_INTR_VECTOR); + + kvm_wait_lapic_expire(vcpu, true); + } + tdx_vcpu_enter_exit(vcpu, tdx); tdx_user_return_update_cache(); @@ -1008,6 +1019,24 @@ static void tdx_handle_changed_private_spte( } } +void tdx_apicv_post_state_restore(struct kvm_vcpu *vcpu) +{ + struct vcpu_tdx *tdx = to_tdx(vcpu); + + pi_clear_on(&tdx->pi_desc); + memset(tdx->pi_desc.pir, 0, sizeof(tdx->pi_desc.pir)); +} + +void tdx_deliver_interrupt(struct kvm_lapic *apic, int delivery_mode, + int trig_mode, int vector) +{ + struct kvm_vcpu *vcpu = apic->vcpu; + struct vcpu_tdx *tdx = to_tdx(vcpu); + + /* TDX supports only posted interrupt. No lapic emulation. */ + __vmx_deliver_posted_interrupt(vcpu, &tdx->pi_desc, vector); +} + static int tdx_capabilities(struct kvm *kvm, struct kvm_tdx_cmd *cmd) { struct kvm_tdx_capabilities __user *user_caps; @@ -1425,6 +1454,10 @@ int tdx_vcpu_ioctl(struct kvm_vcpu *vcpu, void __user *argp) return -EIO; } + td_vmcs_write16(tdx, POSTED_INTR_NV, POSTED_INTR_VECTOR); + td_vmcs_write64(tdx, POSTED_INTR_DESC_ADDR, __pa(&tdx->pi_desc)); + td_vmcs_setbit32(tdx, PIN_BASED_VM_EXEC_CONTROL, PIN_BASED_POSTED_INTR); + tdx->initialized = true; return 0; } diff --git a/arch/x86/kvm/vmx/tdx.h b/arch/x86/kvm/vmx/tdx.h index 180360a65545..7cd81780f3fa 100644 --- a/arch/x86/kvm/vmx/tdx.h +++ b/arch/x86/kvm/vmx/tdx.h @@ -83,6 +83,9 @@ struct vcpu_tdx { struct list_head cpu_list; + /* Posted interrupt descriptor */ + struct pi_desc pi_desc; + union tdx_exit_reason exit_reason; bool initialized; diff --git a/arch/x86/kvm/vmx/vmx.c b/arch/x86/kvm/vmx/vmx.c index 9b7bd52d19a9..4bd1e61b8d45 100644 --- a/arch/x86/kvm/vmx/vmx.c +++ b/arch/x86/kvm/vmx/vmx.c @@ -3931,48 +3931,6 @@ void vmx_msr_filter_changed(struct kvm_vcpu *vcpu) pt_update_intercept_for_msr(vcpu); } -static inline void kvm_vcpu_trigger_posted_interrupt(struct kvm_vcpu *vcpu, - int pi_vec) -{ -#ifdef CONFIG_SMP - if (vcpu->mode == IN_GUEST_MODE) { - /* - * The vector of interrupt to be delivered to vcpu had - * been set in PIR before this function. - * - * Following cases will be reached in this block, and - * we always send a notification event in all cases as - * explained below. - * - * Case 1: vcpu keeps in non-root mode. Sending a - * notification event posts the interrupt to vcpu. - * - * Case 2: vcpu exits to root mode and is still - * runnable. PIR will be synced to vIRR before the - * next vcpu entry. Sending a notification event in - * this case has no effect, as vcpu is not in root - * mode. - * - * Case 3: vcpu exits to root mode and is blocked. - * vcpu_block() has already synced PIR to vIRR and - * never blocks vcpu if vIRR is not cleared. Therefore, - * a blocked vcpu here does not wait for any requested - * interrupts in PIR, and sending a notification event - * which has no effect is safe here. - */ - - apic->send_IPI_mask(get_cpu_mask(vcpu->cpu), pi_vec); - return; - } -#endif - /* - * The vCPU isn't in the guest; wake the vCPU in case it is blocking, - * otherwise do nothing as KVM will grab the highest priority pending - * IRQ via ->sync_pir_to_irr() in vcpu_enter_guest(). - */ - kvm_vcpu_wake_up(vcpu); -} - static int vmx_deliver_nested_posted_interrupt(struct kvm_vcpu *vcpu, int vector) { @@ -4024,20 +3982,7 @@ static int vmx_deliver_posted_interrupt(struct kvm_vcpu *vcpu, int vector) if (!vcpu->arch.apicv_active) return -1; - if (pi_test_and_set_pir(vector, &vmx->pi_desc)) - return 0; - - /* If a previous notification has sent the IPI, nothing to do. */ - if (pi_test_and_set_on(&vmx->pi_desc)) - return 0; - - /* - * The implied barrier in pi_test_and_set_on() pairs with the smp_mb_*() - * after setting vcpu->mode in vcpu_enter_guest(), thus the vCPU is - * guaranteed to see PID.ON=1 and sync the PIR to IRR if triggering a - * posted interrupt "fails" because vcpu->mode != IN_GUEST_MODE. - */ - kvm_vcpu_trigger_posted_interrupt(vcpu, POSTED_INTR_VECTOR); + __vmx_deliver_posted_interrupt(vcpu, &vmx->pi_desc, vector); return 0; } diff --git a/arch/x86/kvm/vmx/x86_ops.h b/arch/x86/kvm/vmx/x86_ops.h index 0f1a28f67e60..c3768a20347f 100644 --- a/arch/x86/kvm/vmx/x86_ops.h +++ b/arch/x86/kvm/vmx/x86_ops.h @@ -148,7 +148,8 @@ void tdx_vcpu_put(struct kvm_vcpu *vcpu); void tdx_vcpu_load(struct kvm_vcpu *vcpu, int cpu); void tdx_apicv_post_state_restore(struct kvm_vcpu *vcpu); -int tdx_deliver_posted_interrupt(struct kvm_vcpu *vcpu, int vector); +void tdx_deliver_interrupt(struct kvm_lapic *apic, int delivery_mode, + int trig_mode, int vector); int tdx_vm_ioctl(struct kvm *kvm, void __user *argp); int tdx_vcpu_ioctl(struct kvm_vcpu *vcpu, void __user *argp); @@ -176,6 +177,10 @@ static inline void tdx_prepare_switch_to_guest(struct kvm_vcpu *vcpu) {} static inline void tdx_vcpu_put(struct kvm_vcpu *vcpu) {} static inline void tdx_vcpu_load(struct kvm_vcpu *vcpu, int cpu) {} +static inline void tdx_apicv_post_state_restore(struct kvm_vcpu *vcpu) {} +static inline void tdx_deliver_interrupt( + struct kvm_lapic *apic, int delivery_mode, int trig_mode, int vector) {} + static inline int tdx_vm_ioctl(struct kvm *kvm, void __user *argp) { return -EOPNOTSUPP; } static inline int tdx_vcpu_ioctl(struct kvm_vcpu *vcpu, void __user *argp) { return -EOPNOTSUPP; }

[RFC,v5,078/104] KVM: TDX: Implement interrupt injection

Commit Message

Comments

Patch