cloud-hypervisor

mirror of https://github.com/cloud-hypervisor/cloud-hypervisor.git synced 2024-11-05 11:31:14 +00:00

Author	SHA1	Message	Date
Rob Bradford	21db6f53c8	vmm: memory_manager: Write all guest region to disk As a mirror of `bdbea19e23` which ensured that GuestMemoryMmap::read_exact_from() was used to read all the file to the region ensure that all the guest memory region is written to disk. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-10-27 12:11:31 -07:00
Rob Bradford	dfd21cbfc5	vmm: Use thiserror/anyhow for vmm::Error This gives a nicer user experience and this error can now be used as the source for other errors based off this. See: #1910 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-10-27 13:27:23 +00:00
dependabot-preview[bot]	f0d0d8ccaf	build(deps): bump libc from 0.2.79 to 0.2.80 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.79 to 0.2.80. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.79...0.2.80) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-10-26 09:31:03 +00:00
Sebastien Boeuf	7e127df415	vmm: memory_manager: Replace 'ext_region' by 'saved_region' Any occurrence of of a variable containing `ext_region` is replaced with the less confusing name `saved_region`. The point is to clearly identify the memory regions that might have been saved during a snapshot, while the `ext` standing for `external` was pretty unclear. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-23 21:59:52 +02:00
Sebastien Boeuf	c0e8e5b53f	vmm: memory_manager: Replace 'backing_file' variable names In the context of saving the memory regions content through snapshot, using the term "backing file" brings confusion with the actual backing file that might back the memory mapping. To avoid such conflicting naming, the 'backing_file' field from the MemoryRegion structure gets replaced with 'content', as this is designating the potential file containing the memory region data. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-23 21:59:52 +02:00
Rob Bradford	bdbea19e23	vmm: memory_manager: Completely fill guest ram from snapshot Use GuestRegionMmap::read_exact_from() to ensure that all of the file is read into the guest. This addresses an issue where GuestRegionMmap::read_from() was only copying the first 2GiB of the memory and so lead to snapshot-restore was failing when the guest RAM was 2GiB or greater. This change also propagates any error from the copying upwards. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-10-23 17:56:19 +01:00
Rob Bradford	a60b437f89	vmm: memory_manager: Always copy anonymous RAM regions from disk When restoring if a region of RAM is backed by anonymous memory i.e from memfd_create() then copy the contents of the ram from the file that has been saved to disk. Previously the code would map the memory from that file into the guest using a MAP_PRIVATE mapping. This has the effect of minimising the restore time but provides an issue where the restored VM does not have the same structure as the snapshotted VM, in particular memory is backed by files in the restored VM that were anonymously backed in the original. This creates two problems: * The snapshot data is mapped from files for the pages of the guest which prevents the storage from being reclaimed. * When snapshotting again the guest memory will not be correctly saved as it will have looked like it was backed by a file so it will not be written to disk but as it is a MAP_PRIVATE mapping the changes will never be written to the disk again. This results in incorrect behaviour. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-10-23 12:34:32 +02:00
Sebastien Boeuf	f4e391922f	vmm: Remove balloon options from --memory parameter The standalone `--balloon` parameter being fully functional at this point, we can get rid of the balloon options from the --memory parameter. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-22 16:33:16 +02:00
Sebastien Boeuf	3594685279	vmm: Move balloon code from MemoryManager to DeviceManager Now that we have a new dedicated way of asking for a balloon through the CLI and the REST API, we can move all the balloon code to the device manager. This allows us to simplify the memory manager, which is already quite complex. It also simplifies the behavior of the balloon resizing command. Instead of providing the expected size for the RAM, which is complex when memory zones are involved, it now expects the balloon size. This is a much more straightforward behavior as it really resizes the balloon to the desired size. Additionally to the simplication, the benefit of this approach is that it does not need to be tied to the memory manager at all. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-22 16:33:16 +02:00
Sebastien Boeuf	1d479e5e08	vmm: Introduce new --balloon parameter This introduces a new way of defining the virtio-balloon device. Instead of going through the --memory parameter, the idea is to consider balloon as a standalone virtio device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-22 16:33:16 +02:00
Sebastien Boeuf	28e12e9f3a	vmm, hypervisor: Fix snapshot/restore for Windows guest The snasphot/restore feature is not working because some CPU states are not properly saved, which means they can't be restored later on. First thing, we ensure the CPUID is stored so that it can be properly restored later. The code is simplified and pushed down to the hypervisor crate. Second thing, we identify for each vCPU if the Hyper-V SynIC device is emulated or not. In case it is, that means some specific MSRs will be set by the guest. These MSRs must be saved in order to properly restore the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-21 19:11:03 +01:00
Rob Bradford	885ee9567b	vmm: Add support for creating virtio-watchdog The watchdog device is created through the "--watchdog" parameter. At most a single watchdog can be created per VM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-10-21 16:02:39 +01:00
Michael Zhao	0b0596ef30	arch: Simplify PCI space address handling in AArch64 FDT Before Virtio-mmio was removed, we passed an optional PCI space address parameter to AArch64 code for generating FDT. The address is none if the transport is MMIO. Now Virtio-PCI is the only option, the parameter is mandatory. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-10-21 12:20:30 +01:00
Michael Zhao	2f2e10ea35	arch: Remove GICv2 Virtio-mmio is removed, now virtio-pci is the only option for virtio transport layer. We use MSI for PCI device interrupt. While GICv2, the legacy interrupt controller, doesn't support MSI. So GICv2 is not very practical for Cloud-hypervisor, we can remove it. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-10-19 14:58:48 +01:00
Sebastien Boeuf	cc8b553e86	virtio-devices: Remove mmio and pci differentiation Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-19 14:58:48 +01:00
Sebastien Boeuf	af3c6c34c3	vmm: Remove mmio and pci differentiation Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-19 14:58:48 +01:00
Sebastien Boeuf	7cbd47a71a	vmm: Prevent KVM device fd from being unusable from VfioContainer When shutting down a VM using VFIO, the following error has been detected: vfio-ioctls/src/vfio_device.rs:312 -- Could not delete VFIO group: KvmSetDeviceAttr(Error(9)) After some investigation, it appears the KVM device file descriptor used for removing a VFIO group was already closed. This is coming from the Rust sequence of Drop, from the DeviceManager all the way down to VfioDevice. Because the DeviceManager owns passthrough_device, which is effectively a KVM device file descriptor, when the DeviceManager is dropped, the passthrough_device follows, with the effect of closing the KVM device file descriptor. Problem is, VfioDevice has not been dropped yet and it still needs a valid KVM device file descriptor. That's why the simple way to fix this issue coming from Rust dropping all resources is to make Linux accountable for it by duplicating the file descriptor. This way, even when the passthrough_device is dropped, the KVM file descriptor is closed, but a duplicated instance is still valid and owned by the VfioContainer. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-15 19:15:28 +02:00
Wei Liu	d667ed0c70	vmm: don't call notify_guest_clock_paused when Hyper-V emulation is on We turn on that emulation for Windows. Windows does not have KVM's PV clock, so calling notify_guest_clock_paused results in an error. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-10-15 19:14:25 +02:00
Sebastien Boeuf	1b9890b807	vmm: cpu: Set CPU physical bits based on user input If the user specified a maximum physical bits value through the `max_phys_bits` option from `--cpus` parameter, the guest CPUID will be patched accordingly to ensure the guest will find the right amount of physical bits. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-13 18:58:36 +02:00
Sebastien Boeuf	aec88e20d7	vmm: memory_manager: Rely on physical bits for address space size If the user provided a maximum physical bits value for the vCPUs, the memory manager will adapt the guest physical address space accordingly so that devices are not placed further than the specified value. It's important to note that if the number exceed what is available on the host, the smaller number will be picked. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-13 18:58:36 +02:00
Sebastien Boeuf	52ad78886c	vmm: Introduce new CPU option to set maximum physical bits In order to let the user choose maximum address space size, this patch introduces a new option `max_phys_bits` to the `--cpus` parameter. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-13 18:58:36 +02:00
Bo Chen	e9738a4a49	vmm: Replace the use of 'unchecked_add' with 'checked_add' The 'GuestAddress::unchecked_add' function has undefined behavior when an overflow occurs. Its alternative 'checked_add' requires use to handle the overflow explicitly. Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-10-13 12:09:22 +02:00
Bo Chen	9ab2a34b40	vmm: Remove reserved 256M gaps for hotplugging memory with ACPI We are now reserving a 256M gap in the guest address space each time when hotplugging memory with ACPI, which prevents users from hotplugging memory to the maximum size they requested. We confirm that there is no need to reserve this gap. This patch removes the 'reserved gaps'. It also refactors the 'MemoryManager::start_addr' so that it is rounding-up to 128M alignment when hotplugged memory is allowed with ACPI. Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-10-13 12:09:22 +02:00
Bo Chen	10f380f95b	vmm: Report no error when resizing to current memory size with ACPI We now try to create a ram region of size 0 when the requested memory size is the same as current memory size. It results in an error of `GuestMemoryRegion(Mmap(Os { code: 22, kind: InvalidInput, message: "Invalid argument" }))`. This error is not meaningful to users and we should not report it. Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-10-12 08:46:38 +02:00
Bo Chen	789ee7b3e4	vmm: Support resizing memory up to and including hotplug size The start address after the hottplugged memory can be the start address of device area. Fixes: #1803 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-10-10 09:51:32 +02:00
Rob Bradford	bb5b9584d2	pci, ch-remote, vmm: Replace simple match blocks with matches! This is a new clippy check introduced in 1.47 which requires the use of the matches!() macro for simple match blocks that return a boolean. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-10-09 10:49:54 +02:00
Wei Liu	ed1fdd1f7d	hypervisor, arch: rename "OneRegister" and relevant code The OneRegister literally means "one (arbitrary) register". Just call it "Register" instead. There is no need to inherit KVM's naming scheme in the hypervisor agnostic code. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-10-08 08:55:10 +02:00
Sebastien Boeuf	1e3a6cb450	vmm: Simplify some of the io_uring code Small patch creating a dedicated `block_io_uring_is_supported()` function for the non-io_uring case, so that we can simplify the code in the DeviceManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-07 14:26:49 +02:00
Sebastien Boeuf	c02a02edfc	vmm: Allow unlink syscall for vCPU threads Without the unlink(2) syscall being allowed, Cloud-Hypervisor crashes when we remove a virtio-vsock device that has been previously added. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-06 16:05:59 +01:00
Sebastien Boeuf	6aa5e21212	vmm: device_manager: Fix PCI device unplug issues Because of the PCI refactoring that happened in the previous commit `d793cc4da3`, the ability to fully remove a PCI device was altered. The refactoring was correct, but the usage of a generic function to pass the same reference for both BusDevice, PciDevice and Any + Send + Sync causes the Arc::ptr_eq() function to behave differently than expected, as it does not match the references later in the code. That means we were not able to remove the device reference from the MMIO and/or PIO buses, which was leading to some bus range overlapping error once we were trying to add a device again to the previous range that should have been removed. Fixes #1802 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-06 12:56:17 +02:00
dependabot-preview[bot]	c2cc26fc82	build(deps): bump libc from 0.2.78 to 0.2.79 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.78 to 0.2.79. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.78...0.2.79) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-10-05 07:02:05 +00:00
Praveen Paladugu	71c435ce91	hypervisor, vmm: Introduce VmmOps trait Run loop in hypervisor needs a callback mechanism to access resources like guest memory, mmio, pio etc. VmmOps trait is introduced here, which is implemented by vmm module. While handling vcpuexits in run loop, this trait allows hypervisor module access to the above mentioned resources via callbacks. Signed-off-by: Praveen Paladugu <prapal@microsoft.com> Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-10-02 16:42:55 +01:00
Rob Bradford	6a9934d933	build: Fix vm-memory bump build error A new version of vm-memory was released upstream which resulted in some components pulling in that new version. Update the version number used to point to the latest version but continue to use our patched version due to the fix for #1258 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-10-02 16:38:02 +01:00
Rob Bradford	2d457ab974	vmm: device_manager: Make PMEM "discard_writes" mode true CoW The PMEM support has an option called "discard_writes" which when true will prevent changes to the device from hitting the backing file. This is trying to be the equivalent of "readonly" support of the block device. Previously the memory of the device was marked as KVM_READONLY. This resulted in a trap when the guest attempted to write to it resulting a VM exit (and recently a warning). This has a very detrimental effect on the performance so instead make "discard_writes" truly CoW by mapping the memory as `PROT_READ \| PROT_WRITE` and using `MAP_PRIVATE` to establish the CoW mapping. Fixes: #1795 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-10-02 14:26:15 +02:00
Hui Zhu	c75f8b2f89	virtio-balloon: Add memory_actual_size to vm.info to show memory actual size The virtio-balloon change the memory size is asynchronous. VirtioBalloonConfig.actual of balloon device show current balloon size. This commit add memory_actual_size to vm.info to show memory actual size. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-10-01 17:46:30 +02:00
dependabot-preview[bot]	76c3230e08	build(deps): bump libc from 0.2.77 to 0.2.78 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.77 to 0.2.78. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.77...0.2.78) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-10-01 05:40:02 +00:00
Rob Bradford	664c3ceda6	vmm: device_manager: Warn that vhost-user self spawning is deprecated See #1724 for details. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-30 18:32:50 +02:00
Rob Bradford	0a4be7ddf5	vmm: "Cleanly" shutdown on SIGTERM Write to the exit_evt EventFD which will trigger all the devices and vCPUs to exit. This is slightly cleaner than just exiting the process as any temporary files will be removed. Fixes: #1242 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-30 18:32:16 +02:00
Bo Chen	6d30fe05e4	vmm: openapi: Add the 'iommu' and 'id' option to 'VmAddDevice' This patch adds the missing the `iommu` and `id` option for `VmAddDevice` in the openApi yaml to respect the internal data structure in the code base. Also, setting the `id` explicitly for VFIO device hotplug is required for VFIO device unplug through openAPI calls. Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-09-30 08:17:44 +01:00
Julio Montes	668c563dac	vmm: openapi: fix integers format According to openAPI specification [1], the format for `integer` types can be only `int32` or `int64`, unsigned and 8-bits integers are not supported. This patch replaces `uint64` with `int64`, `uint32` with `int32` and `uint8` with `int32`. [1]: https://swagger.io/specification/#data-types Signed-off-by: Julio Montes <julio.montes@intel.com>	2020-09-29 12:55:40 -07:00
Rob Bradford	5a0d3277c8	vmm: vm: Replace \n newline character with \r This allows the CMD prompt under SAC to be used without affecting getty on Linux. Fixes: #1770 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-29 16:10:12 +02:00
Wei Liu	4ef97d8ddb	vmm: interrupts: clearly separate MsiInterruptGroup and InterruptRoute MsiInterruptGroup doesn't need to know the internal field names of InterruptRoute. Introduce two helper functions to eliminate references to irq_fd. This is done similarly to the enable and disable helper functions. Also drop the pub keyword from InterruptRoute fields. It is not needed anymore. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-09-29 13:51:35 +02:00
Praveen Paladugu	f10872e706	vmm: fix clippy warnings Signed-off-by: Praveen Paladugu <prapal@microsoft.com>	2020-09-26 14:07:12 +01:00
Praveen Paladugu	4b32252028	hypervisor, vmm: fix clippy warnings Signed-off-by: Praveen Paladugu <prapal@microsoft.com>	2020-09-26 14:07:12 +01:00
Julio Montes	c54452c08a	vmm: openapi: fix integers format According to openAPI specification[1], the format for `integer` types can be only `int32` or `int64`, unsigned integers are not supported. This patch replaces `uint64` with `int64`. [1]: https://swagger.io/specification/#data-types Signed-off-by: Julio Montes <julio.montes@intel.com>	2020-09-26 14:05:51 +01:00
Wei Liu	7e130a65ba	vmm: interrupts: adjust set_gsi_routes There is no point in manually dropping the lock for gsi_msi_routes then instantly grabbing it again in set_gsi_routes. Make set_gsi_routes take a reference to the routing hashmap instead. No functional change intended. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-09-25 17:17:35 +02:00
Sebastien Boeuf	c85e396ce5	vmm: cpu: x86: Enable MTRR feature in CPUID The MTRR feature was missing from the CPUID, which is causing the guest to ignore the MTRR settings exposed through dedicated MSRs. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-25 15:03:52 +02:00
Sebastien Boeuf	2eaf1c70c0	vmm: acpi: Advertise the correct PCI bus range Since Cloud-Hypervisor currently support one single PCI bus, we must reflect this through the MCFG table, as it advertises the first bus and the last bus available. In this case both are bus 0. This patch saves quite some time during guest kernel boot, as it prevents from checking each bus for available devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-23 19:03:19 +02:00
Henry Wang	961c5f2cb2	vmm: AArch64: enable VM states save/restore for AArch64 The states of GIC should be part of the VM states. This commit enables the AArch64 VM states save/restore by adding save/restore of GIC states. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	3ea4a0797d	vmm: seccomp: unify AArch64 and x86_64 FTRUNCATE syscall The definition of libc::SYS_ftruncate on AArch64 is different from that on x86_64. This commit unifies the previously hard-coded syscall number for AArch64. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	48544e4e82	vmm: seccomp: whitelist `KVM_GET_REG_LIST` in seccomp `KVM_GET_REG_LIST` ioctl is needed in save/restore AArch64 vCPU. Therefore we whitelist this ioctl in seccomp. Also this commit unifies the `SYS_FTRUNCATE` syscall for x86_64 and AArch64. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	c6b47d39e0	vmm: refactor vCPU save/restore code in restoring VM Similarly as the VM booting process, on AArch64 systems, the vCPUs should be created before the creation of GIC. This commit refactors the vCPU save/restore code to achieve the above-mentioned restoring order. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	970a5a410d	vmm: decouple vCPU init from `configure_vcpus` Since calling `KVM_GET_ONE_REG` before `KVM_VCPU_INIT` will result in an error: Exec format error (os error 8). This commit decouples the vCPU init process from `configure_vcpus`. Therefore in the process of restoring the vCPUs, these vCPUs can be initialized separately before started. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	47e65cd341	vmm: AArch64: add methods to get saved vCPU states The construction of `GICR_TYPER` register will need vCPU states. Therefore this commit adds methods to extract saved vCPU states from the cpu manager. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	381d0b4372	devices: remove the migration traits for the `Gic` struct Unlike x86_64, the "interrupt_controller" in the device manager for AArch64 is only a `Gic` object that implements the `InterruptController` to provide the interrupt delivery service. This is not the real GIC device so that we do not need to save its states. Also, we do not need to insert it to the device_tree. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	7ddcad1d8b	arch: AArch64: add a field `gicr_typers` for GIC implementations The value of GIC register `GICR_TYPER` is needed in restoring the GIC states. This commit adds a field in the GIC device struct and a method to construct its value. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	dcf6d9d731	device_manager: AArch64: add a field to set/get GIC device entity In AArch64 systems, the state of GIC device can only be retrieved from `KVM_GET_DEVICE_ATTR` ioctl. Therefore to implement saving/restoring the GIC states, we need to make sure that the GIC object (either the file descriptor or the device itself) can be extracted after the VM is started. This commit refactors the code of GIC creation by adding a new field `gic_device_entity` in device manager and methods to set/get this field. The GIC object can be therefore saved in the device manager after calling `arch::configure_system`. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	e7acbcc184	arch: AArch64: support saving RDIST pending tables into guest RAM This commit adds a function which allows to save RDIST pending tables to the guest RAM, as well as unit test case for it. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	29ce3076c2	tests: AArch64: Add unit test cases for accessing GIC registers This commit adds the unit test cases for getting/setting the GIC distributor, redistributor and ICC registers. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	9dd188a8e8	tests: AArch64: Add unit test cases for vCPU save/restore Adds 3 more unit test cases for AArch64: save_restore_core_regs save_restore_system_regs *get_set_mpstate Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Henry Wang	e3d45be6f7	AArch64: Preparation for vCPU save/restore This commit ports code from firecracker and refactors the existing AArch64 code as the preparation for implementing save/restore AArch64 vCPU, including: 1. Modification of `arm64_core_reg` macro to retrive the index of arm64 core register and implemention of a helper to determine if a register is a system register. 2. Move some macros and helpers in `arch` crate to the `hypervisor` crate. 3. Added related unit tests for above functions and macros. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2020-09-23 12:37:25 +01:00
Josh Soref	5c3f4dbe6f	ch: Fix various misspelled words Misspellings were identified by https://github.com/marketplace/actions/check-spelling * Initial corrections suggested by Google Sheets * Additional corrections by Google Chrome auto-suggest * Some manual corrections Signed-off-by: Josh Soref <jsoref@users.noreply.github.com>	2020-09-23 08:59:31 +01:00
Jiangbo Wu	22a2a99e5f	acpi: Add hotplug numa node virtio-mem device would use 'VIRTIO_MEM_F_ACPI_PXM' to add memory to NUMA node, which MUST be existed, otherwise it will be assigned to node id 0, even if user specify different node id. According ACPI spec about Memory Affinity Structure, system hardware supports hot-add memory region using 'Hot Pluggable \| Enabled' flags. Signed-off-by: Jiangbo Wu <jiangbo.wu@intel.com>	2020-09-22 13:11:39 +02:00
Jiangbo Wu	223189c063	mm: Apply zone's property instread of global config Apply memory zone's property for associated virtio-mem regions. Signed-off-by: Jiangbo Wu <jiangbo.wu@intel.com>	2020-09-22 09:56:37 +02:00
Jiangbo Wu	80be8ac0dc	mm: Apply memory policy for virtio-mem region Use zone.host_numa_node to create memory zone, so that memory zone can apply memory policy in according with host numa node ID Signed-off-by: Jiangbo Wu <jiangbo.wu@intel.com>	2020-09-22 09:56:37 +02:00
Sebastien Boeuf	7c346c3844	vmm: Kill vhost-user self-spawned process on failure If after the creation of the self-spawned backend, the VMM cannot create the corresponding vhost-user frontend, the VMM must kill the freshly spawned process in order to ensure the error propagation can happen. In case the child process would still be around, the VMM cannot return the error as it waits onto the child to terminate. This should help us identify when self-spawned failures are caused by a connection being refused between the VMM and the backend. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-18 17:26:25 +01:00
Sebastien Boeuf	555c5c5d9c	vmm: Add missing syscalls to signal thread When the VMM is terminated by receiving a SIGTERM signal, the signal handler thread must be able to invoke ioctl(TCGETS) and ioctl(TCSETS) without error. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-18 13:40:10 +01:00
Rob Bradford	41a9b1adef	vmm: Add missing syscall to vCPU thread Fixes: #1717 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-18 13:40:10 +01:00
Sebastien Boeuf	1e1a50ef70	vmm: Update memory configuration upon virtio-mem resizing Based on all the preparatory work achieved through previous commits, this patch updates the 'hotplugged_size' field for both MemoryConfig and MemoryZoneConfig structures when either the whole memory is resized, or simply when a memory zone is resized. This fixes the lack of support for rebooting a VM with the right amount of memory plugged in. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	de2b917f55	vmm: Add hotplugged_size to VirtioMemZone Adding a new field to VirtioMemZone structure, as it lets us associate with a particular virtio-mem region the amount of memory that should be plugged in at boot. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	3faf8605f3	vmm: Group virtio-mem fields under a dedicated structure This patch simplifies the code as we have one single Option for the VirtioMemZone. This also prepares for storing additional information related to the virtio-mem region. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	4e1b78e1ff	vmm: Add 'hotplugged_size' to memory parameters Add the new option 'hotplugged_size' to both --memory-zone and --memory parameters so that we can let the user specify a certain amount of memory being plugged at boot. This is also part of making sure we can store the virtio-mem size over a reboot of the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Hui Zhu	33a1e37c35	virtio-devices: mem: Allow for an initial size This commit gives the possibility to create a virtio-mem device with some memory already plugged into it. This is preliminary work to be able to reboot a VM with the virtio-mem region being already resized. Signed-off-by: Hui Zhu <teawater@antfin.com> Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	8b5202aa5a	vmm: Always add virtio-mem region upon VM creation Now that e820 tables are created from the 'boot_guest_memory', we can simplify the memory manager code by adding the virtio-mem regions when they are created. There's no need to wait for the first hotplug to insert these regions. This also anticipates the need for starting a VM with some memory already plugged into the virtio-mem region. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	66fc557015	vmm: Store boot guest memory and use it for boot sequence In order to differentiate the 'boot' memory regions from the virtio-mem regions, we store what we call 'boot_guest_memory'. This is useful to provide the adequate list of regions to the configure_system() function as it expects only the list of regions that should be exposed through the e820 table. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	1798ed8194	vmm: virtio-mem: Enforce alignment and size requirements The virtio-mem driver is generating some warnings regarding both size and alignment of the virtio-mem region if not based on 128MiB: The alignment of the physical start address can make some memory unusable. The alignment of the physical end address can make some memory unusable. For these reasons, the current patch enforces virtio-mem regions to be 128MiB aligned and checks the size provided by the user is a multiple of 128MiB. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	eb7b923e22	vmm: Create virtio-mem device with appropriate NUMA node Now that virtio-mem device accept a guest NUMA node as parameter, we retrieve this information from the list of NUMA nodes. Based on the memory zone associated with the virtio-mem device, we obtain the NUMA node identifier, which we provide to the virtio-mem device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	dcedd4cded	virtio-devices: virtio-mem: Add NUMA support Implement support for associating a virtio-mem device with a specific guest NUMA node, based on the ACPI proximity domain identifier. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	0658559880	vmm: memory_manager: Rename 'use_zones' with 'user_provided_zones' This brings more clarity on the meaning of this boolean. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	775f3346e3	vmm: Rename 'virtiomem' to 'virtio_mem' For more consistency and help reading the code better, this commit renames all 'virtiomem' variables into 'virtio_mem'. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	015c78411e	vmm: Add a 'resize-zone' action to the API actions Implement a new VM action called 'resize-zone' allowing the user to resize one specific memory zone at a time. This relies on all the preliminary work from the previous commits to resize each virtio-mem device independently from each others. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	141df701dd	vmm: memory_manager: Make virtiomem_resize function generic By adding a new parameter 'id' to the virtiomem_resize() function, we prepare this function to be usable for both global memory resizing and memory zone resizing. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	34331d3e72	vmm: memory_manager: Fix virtio-mem resize It's important to return the region covered by virtio-mem the first time it is inserted as the device manager must update all devices with this information. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	adc59a6f15	vmm: memory_manager: Create one virtio-mem per memory zone Based on the previous code changes, we can now update the MemoryManager code to create one virtio-mem region and resizing handler per memory zone. This will naturally create one virtio-mem device per memory zone from the DeviceManager's code which has been previously updated as well. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	c645a72c17	vmm: Add 'hotplug_size' to memory zones In anticipation for resizing support of an individual memory zone, this commit introduces a new option 'hotplug_size' to '--memory-zone' parameter. This defines the amount of memory that can be added through each specific memory zone. Because memory zone resize is tied to virtio-mem, make sure the user selects 'virtio-mem' hotplug method, otherwise return an error. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	30ff7e108f	vmm: Prepare code to accept multiple virtio-mem devices Both MemoryManager and DeviceManager are updated through this commit to handle the creation of multiple virtio-mem devices if needed. For now, only the framework is in place, but the behavior remains the same, which means only the memory zone created from '--memory' generates a virtio-mem region that can be used for resize. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Sebastien Boeuf	b173b6c5b4	vmm: Create a MemoryZone structure In order to anticipate the need for storing memory regions along with virtio-mem information for each memory zone, we create a new structure MemoryZone that will replace Vec<Arc<GuestRegionMmap>> in the hash map MemoryZones. This makes thing more logical as MemoryZones becomes a list of MemoryZone sorted by their identifier. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 19:20:04 +02:00
Rob Bradford	27c28fa3b0	vmm, arch: Enable KVM HyperV support Inject CPUID leaves for advertising KVM HyperV support when the "kvm_hyperv" toggle is enabled. Currently we only enable a selection of features required to boot. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-16 16:08:01 +01:00
Rob Bradford	da642fcf7f	hypervisor: Add "HyperV" exit to list of KVM exits Currently we don't need to do anything to service these exits but when the synthetic interrupt controller is active an exit will be triggered to notify the VMM of details of the synthetic interrupt page. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-16 16:08:01 +01:00
Rob Bradford	5495ab7415	vmm: Add "kvm_hyperv" toggle to "--cpus" This turns on the KVM HyperV emulation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-16 16:08:01 +01:00
Sebastien Boeuf	b3435d51d9	vmm: cpu: Add missing io_uring syscalls to vCPU threads Some of the io_uring setup happens upon activation of the virtio-blk device, which is initially triggered through an MMIO VM exit. That's why the vCPU threads must authorize io_uring related syscalls. This commit ensures the virtio-blk io_uring implementation can be used along with the seccomp filters enabled. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-16 11:59:47 +02:00
Bo Chen	9682d74763	vmm: seccomp: Add seccomp filters for signal_handler worker thread This patch covers the last worker thread with dedicated secomp filters. Fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-09-11 07:42:31 +02:00
Bo Chen	2612a6df29	vmm: seccomp: Add seccomp filters for the vcpu worker thread Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-09-11 07:42:31 +02:00
Rob Bradford	d793cc4da3	vmm: device_manager: Extract common PCI code Extract common code for adding devices to the PCI bus into its own function from the VFIO and VIRTIO code paths. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-11 07:33:18 +02:00
Rob Bradford	15025d71b1	devices, vm-device: Move BusDevice and Bus into vm-device This removes the dependency of the pci crate on the devices crate which now only contains the device implementations themselves. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-09-10 09:35:38 +01:00
dependabot-preview[bot]	f24a12913a	build(deps): bump libc from 0.2.76 to 0.2.77 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.76 to 0.2.77. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.76...0.2.77) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-09-10 06:45:09 +00:00
Bo Chen	3c923f0727	virtio-devices: seccomp: Add seccomp filters for virtio_vsock thread This patch enables the seccomp filters for the virtio_vsock worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-09-09 17:04:39 +01:00
Bo Chen	1175fa2bc7	virtio-devices: seccomp: Add seccomp filters for blk_io_uring thread This patch enables the seccomp filters for the block_io_uring worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-09-09 17:04:39 +01:00
Sebastien Boeuf	e15dba2925	vmm: Rename NUMA option 'id' into 'guest_numa_id' The goal of this commit is to rename the existing NUMA option 'id' with 'guest_numa_id'. This is done without any modification to the way this option behaves. The reason for the rename is caused by the observation that all other parameters with an option called 'id' expect a string to be provided. Because in this particular case we expect a u32 representing a proximity domain from the ACPI specification, it's better to name it with a more explicit name. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-07 07:37:14 +02:00
Sebastien Boeuf	1970ee89da	main, vmm: Remove guest_numa_node option from memory zones The way to describe guest NUMA nodes has been updated through previous commits, letting the user describe the full NUMA topology through the --numa parameter (or NumaConfig). That's why we can remove the deprecated and unused 'guest_numa_node' option. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-07 07:37:14 +02:00
Sebastien Boeuf	f21c04166a	vmm: Move NUMA node list creation to Vm structure Based on the previous changes introducing new options for both memory zones and NUMA configuration, this patch changes the behavior of the NUMA node definition. Instead of relying on the memory zones to define the guest NUMA nodes, everything goes through the --numa parameter. This allows for defining NUMA nodes without associating any particular memory range to it. And in case one wants to associate one or multiple memory ranges to it, the expectation is to describe a list of memory zone through the --numa parameter. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-07 07:37:14 +02:00
Sebastien Boeuf	dc42324351	vmm: Add 'memory_zones' option to NumaConfig This new option provides a new way to describe the memory associated with a NUMA node. This is the first step before we can remove the 'guest_numa_node' option from the --memory-zone parameter. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-07 07:37:14 +02:00
Sebastien Boeuf	5d7215915f	vmm: memory_manager: Store a list of memory zones Now that we have an identifier per memory zone, and in order to keep track of the memory regions associated with the memory zones, we create and store a map referencing list of memory regions per memory zone ID. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-07 07:37:14 +02:00
Sebastien Boeuf	3ff82b4b65	main, vmm: Add mandatory id to memory zones In anticipation for allowing memory zones to be removed, but also in anticipation for refactoring NUMA parameter, we introduce a mandatory 'id' option to the --memory-zone parameter. This forces the user to provide a unique identifier for each memory zone so that we can refer to these. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-07 07:37:14 +02:00
Samuel Ortiz	e5ce6dc43c	vmm: cpu: Warn if the guest is trying to access unregistered IO ranges Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-09-04 14:39:58 +02:00
Sebastien Boeuf	c0d0d23932	vmm: acpi: Introduce SLIT for NUMA nodes distances By introducing the SLIT (System Locality Distance Information Table), we provide the guest with the distance between each node. This lets the user describe the NUMA topology with a lot of details so that slower memory backing the VM can be exposed as being further away from other nodes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 18:09:01 +02:00
Sebastien Boeuf	9548e7e857	vmm: Update NUMA node distances internally Based on the NumaConfig which now provides distance information, we can internally update the list of NUMA nodes with the exact distances they should be located from other nodes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 18:09:01 +02:00
Sebastien Boeuf	a5a29134ca	vmm: Extend --numa parameter with NUMA node distances By introducing 'distances' option, we let the user describe a list of destination NUMA nodes with their associated distances compared to the current node (defined through 'id'). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 18:09:01 +02:00
Sebastien Boeuf	629befdb4a	vmm: acpi: Add CPUs to NUMA nodes Based on the list of CPUs related to each NUMA node, Processor Local x2APIC Affinity structures are created and included into the SRAT table. This describes which CPUs are part of each node. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 15:25:00 +02:00
Sebastien Boeuf	db28db8567	vmm: Update NUMA nodes based on NumaConfig Relying on the list of CPUs defined through the NumaConfig, this patch will update the internal list of CPUs attached to each NUMA node. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 15:25:00 +02:00
Sebastien Boeuf	42f963d6f2	main, vmm: Add new --numa parameter Through this new parameter, we give users the opportunity to specify a set of CPUs attached to a NUMA node that has been previously created from the --memory-zone parameter. This parameter will be extended in the future to describe the distance between multiple nodes. For instance, if a user wants to attach CPUs 0, 1, 2 and 6 to a NUMA node, here are two different ways of doing so: Either ./cloud-hypervisor ... --numa id=0,cpus=0-2:6 Or ./cloud-hypervisor ... --numa id=0,cpus=0:1:2:6 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 15:25:00 +02:00
Sebastien Boeuf	65a23c6fc6	vmm: acpi: Create the SRAT table The SRAT table (System Resource Affinity Table) is needed to describe NUMA nodes and how memory ranges and CPUs are attached to them. For now it simply attaches a list of Memory Affinity structures based on the list of NUMA nodes created from the VMM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 14:11:49 +02:00
Sebastien Boeuf	cf81254a8d	vmm: memory_manager: Create a NUMA node list Based on the 'guest_numa_node' option, we create and store a list of NUMA nodes in the MemoryManager. The point being to associate a list of memory regions to each node, so that we can later create the ACPI tables with the proper memory range information. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 14:11:49 +02:00
Sebastien Boeuf	768dbd1fb0	vmm: Add 'guest_numa_node' option to 'memory-zone' With the introduction of this new option, the user will be able to describe if a particular memory zone should belong to a specific NUMA node from a guest perspective. For instance, using '--memory-zone size=1G,guest_numa_node=2' would let the user describe that a memory zone of 1G in the guest should be exposed as being associated with the NUMA node 2. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 14:11:49 +02:00
Sebastien Boeuf	274c001eab	vmm: Use u32 instead of u64 for host_numa_node option Given that ACPI uses u32 as the type for the Proximity Domain, we can use u32 instead of u64 as the type for 'host_numa_node' option. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-09-01 13:29:42 +02:00
Michael Zhao	a95b6bbd8b	vmm: Add seccomp rules for starting vhost-user-net backend on AArch64 Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-08-31 08:19:23 +02:00
Hui Zhu	f7b3581645	cloud-hypervisor.yaml: MemoryConfig: Add balloon_size "struct MemoryConfig" has balloon_size but not in MemoryConfig of cloud-hypervisor.yaml. This commit adds it. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-08-28 09:58:39 +02:00
Sebastien Boeuf	a8a9e61c3d	vmm: memory_manager: Allow host NUMA for RAM backed files Let's narrow down the limitation related to mbind() by allowing shared mappings backed by a file backed by RAM. This leaves the restriction on only for mappings backed by a regular file. With this patch, host NUMA node can be specified even if using vhost-user devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-27 08:39:38 -07:00
Sebastien Boeuf	1b4591aecc	vmm: memory_manager: Apply NUMA policy to memory zones Relying on the new option 'host_numa_node' from the 'memory-zone' parameter, the user can now define which NUMA node from the host should be used to back the current memory zone. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-27 08:39:38 -07:00
Sebastien Boeuf	e6f585a31c	vmm: Add 'host_numa_nodes' option to memory zones Since memory zones have been introduced, it is now possible for a user to specify multiple backends for the guest RAM. By adding a new option 'host_numa_node' to the 'memory-zone' parameter, we allow the guest RAM to be backed by memory that might come from a specific NUMA node on the host. The option expects a node identifier, specifying which NUMA node should be used to allocate the memory associated with a specific memory zone. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-27 08:39:38 -07:00
Sebastien Boeuf	ad5d0e4713	vmm: Remove 'mergeable' from memory zones The flag 'mergeable' should only apply to the entire guest RAM, which is why it is removed from the MemoryZoneConfig as it is defined as a global parameter at the MemoryConfig level. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-27 07:26:49 +02:00
Sebastien Boeuf	89e7774b96	vmm: openapi: Don't expect cmdline to always be there The 'cmdline' parameter should not be required as it is not needed when the 'kernel' parameter is the rust-hypervisor-fw, which means the kernel and the associated command line will be found from the EFI partition. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:49:05 +02:00
Sebastien Boeuf	e8149380b7	vmm: memory_manager: Factorize memory regions creation Factorize the codepath between simple memory and multiple memory zones. This simplifies the way regions are memory mapped, as everything relies on the same codepath. This is performed by creating a memory zone on the fly for the specific use case where --memory is used with size being different from 0. Internally, the code can rely on memory zones to create the memory regions forming the guest memory. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	c58dd761f4	vmm: Remove 'file' option from MemoryConfig After the introduction of user defined memory zones, we can now remove the deprecated 'file' option from --memory parameter. This makes this parameter simpler, letting more advanced users define their own custom memory zones through the dedicated parameter. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	5bf7113768	vmm: memory_manager: Remove restrictions about snapshot/restore User defined memory regions can now support being snapshot and restored, therefore this commit removes the restrictions that were applied through earlier commit. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	2583d572fc	vmm: memory_manager: Simplify how to restore memory regions By factorizing a lot of code into create_ram_region(), this commit achieves the simplification of the restore codepath. Additionally, it makes user defined memory zones compatible with snapshot/restore. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	b14c861c6f	vmm: memory_manager: Store memory regions content only when necessary First thing, this patch introduces a new function to identify if a file descriptor is linked to any hard link on the system. This can let the VMM know if the file can be accessed by the user, or if the file will be destroyed as soon as the VMM releases the file descriptor. Based on this information, and associated with the knowledge about the region being MAP_SHARED or not, the VMM can now decide to skip the copy of the memory region content. If the user has access to the file from the filesystem, and if the file has been mapped as MAP_SHARED, we can consider the guest memory region content to be present in this file at any point in time. That's why in this specific case, there's no need for performing the copy of the memory region content into a dedicated file. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	d1ce52f3a8	vmm: memory_manager: Make backing file from snapshot optional Let's not assume that a backing file is going to be the result from a snapshot for each memory region. These regions might be backed by a file on the host filesystem (not a temporary file in host RAM), which means they don't need to be copied and stored into dedicated files. That's why this commit prepares for further changes by introducing an optional PathBuf associated with the snapshot of each memory region. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	871138d5cc	vm-migration: Make snapshot() mutable There will be some cases where the implementation of the snapshot() function from the Snapshottable trait will require to modify some internal data, therefore we make this possible by updating the trait definition with snapshot(&mut self). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	c13721fdbd	vmm: memory_manager: Handle user defined memory zones In case the memory size is 0, this means the user defined memory zones are used as a way to specify how to back the guest memory. This is the first step in supporting complex use cases where the user can define exactly which type of memory from the host should back the memory from the guest. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	7cd3867e2c	vmm: memory_manager: Provide file offset through create_ram_region() In anticipation for the need to map part of a file with the function create_ram_region(), it is extended to accept a file offset as argument. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	59d4a56ab7	vmm: memory_manager: Don't truncate backing file In case the provided backing file is an actual file and not a directory, we should not truncate it, as we expect the file to already be the right size. This change will be important once we try to map the same file through multiple memory mappings. We can't let the file be truncated as the second mapping wouldn't work properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	be475ddc22	main, vmm: Let the user define distincts memory zones Introducing a new CLI option --memory-zone letting the user specify custom memory zones. When this option is present, the --memory size must be explicitly set to 0. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Sebastien Boeuf	d25ec66bb6	vmm: memory_manager: Simplify start_addr() Small simplification for the function calculating the start address. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-25 16:43:10 +02:00
Anatol Belski	12212d2966	pci: device_manager: Remove hardcoded I/O port assignment It is otherwise seems to be able to cause resource conflicts with Windows APCI_HAL. The OS might do a better job on assigning resources to this device, withouth them to be requested explicitly. 0xcf8 and 0xcfc are only what is certainly needed for the PCI device enumeration. Signed-off-by: Anatol Belski <anatol.belski@microsoft.com>	2020-08-25 09:00:06 +02:00
Michael Zhao	afc98a5ec9	vmm: Fix AArch64 clippy warnings of vmm and other crates Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-08-24 10:59:08 +02:00
Muminul Islam	92b4499c1e	vmm, hypervisor: Add vmstate to snapshot and restore path Signed-off-by: Muminul Islam <muislam@microsoft.com>	2020-08-24 08:48:15 +02:00
dependabot-preview[bot]	57ff608be9	build(deps): bump libc from 0.2.74 to 0.2.76 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.74 to 0.2.76. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.74...0.2.76) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-08-21 07:08:35 +00:00
Bo Chen	02d87833f0	virtio-devices: seccomp: Add seccomp filters for vhost_blk thread This patch enables the seccomp filters for the vhost_blk worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-19 08:33:58 +02:00
Bo Chen	896b9a1d4b	virtio-devices: seccomp: Add seccomp filter for vhost_net_ctl thread This patch enables the seccomp filters for the vhost_net_ctl worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-19 08:33:58 +02:00
Bo Chen	02d63149fe	virtio-devices: seccomp: Add seccomp filters for vhost_fs thread This patch enables the seccomp filters for the vhost_fs worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-19 08:33:58 +02:00
Bo Chen	c82ded8afa	virtio-devices: seccomp: Add seccomp filters for balloon thread This patch enables the seccomp filters for the balloon worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-19 08:33:58 +02:00
Bo Chen	c460178723	virtio-devices: seccomp: Add seccomp filters for mem thread This patch enables the seccomp filters for the mem worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-19 08:33:58 +02:00
Bo Chen	4539236690	virtio-devices: seccomp: Add seccomp filters for iommu thread This patch enables the seccomp filters for the iommu worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-17 21:08:49 +02:00
Anatol Belski	eba42c392f	devices: acpi: Add UID to devices with common HID Some OS might check for duplicates and bail out, if it can't create a distinct mapping. According to ACPI 5.0 section 6.1.12, while _UID is optional, it becomes required when there are multiple devices with the same _HID. Signed-off-by: Anatol Belski <ab@php.net>	2020-08-14 08:52:02 +02:00
dependabot-preview[bot]	ebe61de0d1	build(deps): bump clap from 2.33.2 to 2.33.3 Bumps [clap](https://github.com/clap-rs/clap) from 2.33.2 to 2.33.3. - [Release notes](https://github.com/clap-rs/clap/releases) - [Changelog](https://github.com/clap-rs/clap/blob/master/CHANGELOG.md) - [Commits](https://github.com/clap-rs/clap/commits) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-08-14 06:18:55 +00:00
Sebastien Boeuf	bdef54ead6	vmm: Add brk syscall to the API thread The brk syscall is not always called as the system might not need it. But when it's needed from the API thread, this causes the thread to terminate as it is not part of the authorized list of syscalls. This should fix some sporadic failures on the CI with the musl build. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-11 15:04:21 +01:00
dependabot-preview[bot]	7529a9ac05	build(deps): bump seccomp from v0.21.2 to v0.22.0 Bumps [seccomp](https://github.com/firecracker-microvm/firecracker) from v0.21.2 to v0.22.0. - [Release notes](https://github.com/firecracker-microvm/firecracker/releases) - [Changelog](`cc5387637c/CHANGELOG.md`) - [Commits](`a06d358b2e...cc5387637c`) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-08-06 07:25:30 +00:00
dependabot-preview[bot]	8e8ec74b2a	build(deps): bump clap from 2.33.1 to 2.33.2 Bumps [clap](https://github.com/clap-rs/clap) from 2.33.1 to 2.33.2. - [Release notes](https://github.com/clap-rs/clap/releases) - [Changelog](https://github.com/clap-rs/clap/blob/master/CHANGELOG.md) - [Commits](https://github.com/clap-rs/clap/commits) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-08-05 20:36:12 +00:00
Jose Carlos Venegas Munoz	90acb01bad	vmm: seccomp: add mprotect to API thread filter Add mprotect to API thread rules. Prevent the VMM is killed when it is used. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2020-08-05 21:35:21 +01:00
dependabot-preview[bot]	ec9de259ba	build(deps): bump seccomp from v0.21.1 to v0.21.2 Bumps [seccomp](https://github.com/firecracker-microvm/firecracker) from v0.21.1 to v0.21.2. - [Release notes](https://github.com/firecracker-microvm/firecracker/releases) - [Changelog](`a06d358b2e/CHANGELOG.md`) - [Commits](`047a379eb0...a06d358b2e`) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-08-05 07:34:44 +00:00
Bo Chen	dc71d2765a	virtio-devices: seccomp: Add seccomp filters for pmem thread This patch enables the seccomp filters for the pmem worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-05 08:13:31 +01:00
Bo Chen	d77977536d	virtio-devices: seccomp: Add seccomp filters for net thread This patch enables the seccomp filters for the net worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-05 08:13:31 +01:00
Bo Chen	276df6b71c	virtio-devices: seccomp: Add seccomp filters for console thread This patch enables the seccomp filters for the console worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-05 08:13:31 +01:00
Bo Chen	a426221167	virtio-devices: seccomp: Add seccomp filters for rng thread This patch enables the seccomp filters for the rng worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-05 08:13:31 +01:00
Bo Chen	704edd544c	virtio-devices: seccomp: Add seccomp_filter module This patch added the seccomp_filter module to the virtio-devices crate by taking reference code from the vmm crate. This patch also adds allowed-list for the virtio-block worker thread. Partially fixes: #925 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-04 11:40:49 +02:00
Bo Chen	ff7ed8f628	vmm: Propagate the SeccompAction value to the Vm struct constructor This patch propagates the SeccompAction value from main to the Vm struct constructor (i.e. Vm::new_from_memory_manager), so that we can use it to construct the DeviceManager and CpuManager struct for controlling the behavior of the seccomp filters for vcpu/virtio-device worker threads. Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-04 11:40:49 +02:00
Bo Chen	8e74637ebb	main, vmm: seccomp: Add the '--seccomp log' option This patch extends the CLI option '--seccomp' to accept the 'log' parameter in addition 'true/false'. It also refactors the vmm::seccomp_filters module to support both "SeccompAction::Trap" and "SeccompAction::Log". Fixes: #1180 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-04 11:40:49 +02:00
Bo Chen	b41884a406	main, vmm: seccomp: Use SeccompAction instead of SeccompLevel This patch replaces the usage of 'SeccompLevel' with 'SeccompAction', which is the first step to support the 'log' action over system calls that are not on the allowed list of seccomp filters. Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-08-04 11:40:49 +02:00
Sebastien Boeuf	8f0bf82648	io_uring: Add new feature gate By adding a new io_uring feature gate, we let the user the possibility to choose if he wants to enable the io_uring improvements or not. Since the io_uring feature depends on the availability on recent host kernels, it's better if we leave it off for now. As soon as our CI will have support for a kernel 5.6 with all the features needed from io_uring, we'll enable this feature gate permanently. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-03 14:15:01 +01:00
Sebastien Boeuf	917027c55b	vmm: Rely on virtio-blk io_uring when possible In case the host supports io_uring and the specific io_uring options needed, the VMM will choose the asynchronous version of virtio-blk. This will enable better I/O performances compared to the default synchronous version. This is also important to note the VMM won't be able to use the asynchronous version if the backend image is in QCOW format. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-08-03 14:15:01 +01:00
Praveen Paladugu	afa8ecc90c	vmm: add validation for network parameters Signed-off-by: Praveen Paladugu <prapal@microsoft.com>	2020-07-31 09:07:12 +02:00
Wei Liu	a52b614a61	vmm: device_manager: console input should be only consumed by one device Cloud Hypervisor allows either the serial or virtio console to output to TTY, but TTY input is pushed to both. This is not correct. When Linux guest is configured to spawn TTYs on both ttyS0 and hvc0, the user effectively issues the same commands twice in different TTYs. Fix this by only direct input to the one choice that is using host side TTY. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-30 18:05:01 +02:00
Wei Liu	5ed794a44c	vmm: device_manager: rename console_input to virtio_console_input Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-30 18:05:01 +02:00
Wei Liu	3e68867bb7	vmm: device_manager: eliminate KvmMsiInterruptManager from the new function The logic to create an MSI interrupt manager is applicable to Hyper-V as well. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-30 08:00:33 +02:00
dependabot-preview[bot]	12c5b7668a	build(deps): bump libc from 0.2.73 to 0.2.74 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.73 to 0.2.74. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.73...0.2.74) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-07-28 20:46:37 +00:00
Wei Liu	218ec563fc	vmm: fix warnings when KVM is not enabled Some imports are only used by KVM. Some variables and code become dead or unused when KVM is not enabled. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-28 21:08:39 +01:00
Jianyong Wu	d24b110519	seccomp: AArch64: Add SYS_unlinkat to seccomp whitelist This commit fixes an "Bad syscall" error when shutting down the VM on AArch64 by adding the SYS_unlinkat syscall to the seccomp whitelist. Signed-off-by: Jianyong Wu <jianyong.wu@arm.com>	2020-07-27 07:25:07 +00:00
Rob Bradford	9ae44aeada	vmm: acpi_tables: Fix PM timer I/O port width Ensure that the width of the I/O port is correctly set to 32-bits in the generic address used for the X_PM_TMR_BLK. Do this by type parameterising GenericAddress::io_port_address() fuction. TEST=Boot with clocksource=acpi_pm and observe no errors in the dmesg. Fixes: #1496 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-07-23 17:48:22 +02:00
Rob Bradford	aae5d988e1	devices: vmm: Add ACPI PM timer This is a counter exposed via an I/O port that runs at 3.579545MHz. Here we use a hardcoded I/O and expose the details through the FADT table. TEST=Boot Linux kernel and see the following in dmesg: [ 0.506198] clocksource: acpi_pm: mask: 0xffffff max_cycles: 0xffffff, max_idle_ns: 2085701024 ns Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-07-23 13:10:21 +01:00
Wei Liu	f03afea0d6	device_manager: document unsafe block in add_vfio_device It is not immediately obvious why the conversion is safe. Document the safety guarantee. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-21 17:13:10 +01:00
Samuel Ortiz	be51ea250d	device_manager: Simplify the passthrough internal API We store the device passthrough handler, so we should use it through our internal API and only carry the passed through device configuration. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-07-21 17:20:25 +02:00
Michael Zhao	ddf1b76906	hypervisor: Refactor create_passthrough_device() for generic type Changed the return type of create_passthrough_device() to generic type hypervisor::Device. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-07-21 16:22:02 +02:00
Michael Zhao	e3e771727a	arch: Refactor GIC code to seperate KVM specific code Shrink GICDevice trait to contain hypervisor agnostic API's only, which are used in generating FDT. Move all KVM specific logic into KvmGICDevice trait. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-07-21 16:22:02 +02:00
Michael Zhao	3e051e7b2c	arch, vmm: Enable initramfs on AArch64 Ported Firecracker commit 144b6c. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-07-20 14:20:53 +01:00
Wei Liu	e1af251c9f	vmm, hypervisor: adjust set_gsi_routing / set_gsi_routes Make set_gsi_routing take a list of IrqRoutingEntry. The construction of hypervisor specific structure is left to set_gsi_routing. Now set_gsi_routes, which is part of the interrupt module, is only responsible for constructing a list of routing entries. This further splits hypervisor specific code from hypervisor agnostic code. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-20 07:32:32 +02:00
dependabot-preview[bot]	12b37ef13b	build(deps): bump libc from 0.2.72 to 0.2.73 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.72 to 0.2.73. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.72...0.2.73) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-07-20 05:15:24 +00:00
Wei Liu	d484a3383c	vmm: device_manager: introduce add_passthrough_device It calls add_vfio_device on KVM or returns an error when not running on KVM. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-17 20:21:39 +02:00
Wei Liu	821892419c	vmm: device_manager: use generic names for passthrough device Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-17 20:21:39 +02:00
Wei Liu	ff8d7bfe83	hypervisor: add create_passthrough_device call to Vm trait That function is going to return a handle for passthrough related operations. Move create_kvm_device code there. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-17 20:21:39 +02:00
Wei Liu	c08d2b2c70	device_manager: avoid manipulating MemoryRegion fields directly Hyper-V may have different field names. Use make_user_memory_region instead. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-16 15:56:03 +02:00
dependabot-preview[bot]	cc57467d10	build(deps): bump log from 0.4.8 to 0.4.11 Bumps [log](https://github.com/rust-lang/log) from 0.4.8 to 0.4.11. - [Release notes](https://github.com/rust-lang/log/releases) - [Changelog](https://github.com/rust-lang/log/blob/master/CHANGELOG.md) - [Commits](https://github.com/rust-lang/log/compare/0.4.8...0.4.11) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-07-16 05:33:44 +00:00
Wei Liu	d80e383dbb	arch: move test cases to vmm crate This saves us from adding a "kvm" feature to arch crate merely for the purpose of running tests. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-15 17:21:07 +02:00
Wei Liu	598eaf9f86	vmm: use hypervisor::new in test_vm Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-15 17:21:07 +02:00
Sebastien Boeuf	a5c4f0fc6f	arch, vmm: Add e820 entry related to SGX EPC region SGX expects the EPC region to be reported as "reserved" from the e820 table. This patch adds a new entry to the table if SGX is enabled. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-07-15 15:08:56 +02:00
Sebastien Boeuf	e10d9b13d4	arch, hypervisor, vmm: Patch CPUID subleaves to expose EPC sections The support for SGX is exposed to the guest through CPUID 0x12. KVM passes static subleaves 0 and 1 from the host to the guest, without needing any modification from the VMM itself. But SGX also relies on dynamic subleaves 2 through N, used for describing each EPC section. This is not handled by KVM, which means the VMM is in charge of setting each subleaf starting from index 2 up to index N, depending on the number of EPC sections. These subleaves 2 through N are not listed as part of the supported CPUID entries from KVM. But it's important to set them as long as index 0 and 1 are present and indicate that SGX is supported. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-07-15 15:08:56 +02:00
Sebastien Boeuf	1603786374	vmm: Pass MemoryManager through CpuManager creation Instead of passing the GuestMemoryMmap directly to the CpuManager upon its creation, it's better to pass a reference to the MemoryManager. This way we will be able to know if SGX EPC region along with one or multiple sections are present. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-07-15 15:08:56 +02:00
Sebastien Boeuf	2b06ce0ed4	vmm: Add EPC device to ACPI tables The SGX EPC region must be exposed through the ACPI tables so that the guest can detect its presence. The guest only get the full range from ACPI, as the specific EPC sections are directly described through the CPUID of each vCPU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-07-15 15:08:56 +02:00
Sebastien Boeuf	84cf12d86a	arch, vmm: Create SGX virtual EPC sections from MemoryManager Based on the presence of one or multiple SGX EPC sections from the VM configuration, the MemoryManager will allocate a contiguous block of guest address space to hold the entire EPC region. Within this EPC region, each EPC section is memory mapped. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-07-15 15:08:56 +02:00
Sebastien Boeuf	d9244e9f4c	vmm: Add option for enabling SGX EPC regions Introducing the new CLI option --sgx-epc along with the OpenAPI structure SgxEpcConfig, so that a user can now enable one or multiple SGX Enclave Page Cache sections within a contiguous region from the guest address space. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-07-15 15:08:56 +02:00
Michael Zhao	cce6237536	pci: Enable GSI routing (MSI type) for AArch64 In this commit we saved the BDF of a PCI device and set it to "devid" in GSI routing entry, because this field is mandatory for GICv3-ITS. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-07-14 14:34:54 +01:00
Michael Zhao	f2e484750a	arch: aarch64: Add PCIe node in FDT for AArch64 Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-07-14 14:34:54 +01:00
Michael Zhao	17057a0dd9	vmm: Fix build errors with "pci" feature on AArch64 Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-07-14 14:34:54 +01:00
Rob Bradford	4963e37dc8	qcow, virtio-devices: Break cyclic dependency Move the definition of RawFile from virtio-devices crate into qcow crate. All the code that consumes RawFile also already depends on the qcow crate for image file type detection so this change breaks the need for the qcow crate to depend on the very large virtio-devices crate. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-07-10 17:47:31 +02:00
Wei Liu	5bfac796b3	build: add a default feature KVM It gets bubbled all the way up from hypervsior crate to top-level Cargo.toml. Cloud Hypervisor can't function without KVM at this point, so make it a default feature. Fix all scripts that use --no-default-features. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-08 11:07:15 +01:00
dependabot-preview[bot]	861337cc6f	build(deps): bump libc from 0.2.71 to 0.2.72 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.71 to 0.2.72. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.71...0.2.72) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-07-08 05:11:25 +00:00
Hui Zhu	800220acbb	virtio-balloon: Store the balloon size to support reboot This commit store balloon size to MemoryConfig. After reboot, virtio-balloon can use this size to inflate back to the size before reboot. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-07-07 17:25:13 +01:00
Hui Zhu	8ffbc3d031	vmm: api: ch-remote: Add balloon to VmResizeData Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-07-07 17:25:13 +01:00
Hui Zhu	f729b25a10	openapi: Add MemoryConfig balloon Add MemoryConfig balloon to vmm/src/api/openapi/cloud-hypervisor.yaml. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-07-07 17:25:13 +01:00
Hui Zhu	8b6b97b86f	vmm: Add virtio-balloon support This commit adds new option balloon to memory config. Set it to on will open the balloon function. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-07-07 17:25:13 +01:00
Rob Bradford	b69f6d4f6c	vhost_user_net, vhost_user_block, option_parser: Remove vmm dependency Remove the vmm dependency from vhost_user_block and vhost_user_net where it was existing to use config::OptionParser. By moving the OptionParser to its own crate at the top-level we can remove the very heavy dependency that these vhost-user backends had. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-07-06 18:33:29 +01:00
Michael Zhao	726e45e0ce	vmm: Divide Seccomp KVM IOCTL rules by architecture Refactored the construction of KVM IOCTL rules for Seccomp. Separating the rules by architecture can reduce the risk of bugs and attacks. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-07-06 13:40:38 +01:00
Wei Liu	a4f484bc5e	hypervisor: Define a VM-Exit abstraction In order to move the hypervisor specific parts of the VM exit handling path, we're defining a generic, hypervisor agnostic VM exit enum. This is what the hypervisor's Vcpu run() call should return when the VM exit can not be completely handled through the hypervisor specific bits. For KVM based hypervisors, this means directly forwarding the IO related exits back to the VMM itself. For other hypervisors that e.g. rely on the VMM to decode and emulate instructions, this means the decoding itself would happen in the hypervisor crate exclusively, and the rest of the VM exit handling would be handled through the VMM device model implementation. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Fix test_vm unit test by using the new abstraction and dropping some dead code. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-06 12:59:43 +01:00
Wei Liu	cfa758fbb1	vmm, hypervisor: introduce and use make_user_memory_region This removes the last KVM-ism from memory_manager. Also make use of that method in other places. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-06 12:31:19 +02:00
Wei Liu	8d97d628c3	vmm: drop "kvm" from memory slot code The code is purely for maintaining an internal counter. It is not really tied to KVM. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-07-06 12:31:19 +02:00
Samuel Ortiz	8186a8eee6	vmm: interrupt: Rename vm_fd The _fd suffix is KVM specific. But since it now point to an hypervisor agnostic hypervisor::Vm implementation, we should just rename it vm. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-07-06 09:35:30 +01:00
Samuel Ortiz	4cc8853fe4	vmm: device_manager: Rename vm_fd The _fd suffix is KVM specific. But since it now point to an hypervisor agnostic hypervisor::Vm implementation, we should just rename it vm. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-07-06 09:35:30 +01:00
Samuel Ortiz	2012287611	vmm: memory_manager: Rename fd variable into something more meaningful The fd naming is quite KVM specific. Since we're now using the hypervisor crate abstractions, we can rename those into something more readable and meaningful. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-07-06 09:35:30 +01:00
Samuel Ortiz	acfe5eb94f	vmm: vm: Rename fd variable into something more meaningful The fd naming is quite KVM specific. Since we're now using the hypervisor crate abstractions, we can rename those into something more readable and meaningful. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-07-06 09:35:30 +01:00
Samuel Ortiz	3db4c003a3	vmm: cpu: Rename fd variable into something more meaningful The fd naming is quite KVM specific. Since we're now using the hypervisor crate abstractions, we can rename those into something more readable and meaningful. Like e.g. vcpu or vm. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-07-06 09:35:30 +01:00
Samuel Ortiz	618722cdca	hypervisor: cpu: Rename state getter and setter vcpu.{set_}cpu_state() is a stutter. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-07-06 09:35:30 +01:00
Rob Bradford	2a6eb31d5b	vm-virtio, virtio-devices: Split device implementation from virt queues Split the generic virtio code (queues and device type) from the VirtioDevice trait, transport and device implementations. This also simplifies the feature handling in vhost_user_backend as the vm-virtio crate is no longer has any features. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-07-02 17:09:28 +01:00
Michael Zhao	8820e9e133	vmm: Fix Seccomp filter for AArch64 Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-07-02 08:46:24 +01:00
Sebastien Boeuf	e35d4c5b28	hypervisor: Store all supported MSRs On x86 architecture, we need to save a list of MSRs as part of the vCPU state. By providing the full list of MSRs supported by KVM, this patch fixes the remaining snapshot/restore issues, as the vCPU is restored with all its previous states. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-30 14:03:03 +01:00
Sebastien Boeuf	e2b5c78dc5	hypervisor: Re-order vCPU state for storing and restoring Some vCPU states such as MP_STATE can be modified while retrieving other states. For this reason, it's important to follow a specific order that will ensure a state won't be modified after it has been saved. Comments about ordering requirements have been copied over from Firecracker commit 57f4c7ca14a31c5536f188cacb669d2cad32b9ca. This patch also set the previously saved VCPU_EVENTS, as this was missing from the restore codepath. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-30 14:03:03 +01:00
Wei Liu	2b8accf49a	vmm: interrupt: put KVM code into a kvm module Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	c31e747005	vmm: interrupt: generify impl InterruptManager for MsiInterruptManager The logic can be shared among hypervisor implementations. The 'static bound is used such that we don't need to deal with extra lifetime parameter everywhere. It should be okay because we know the entry type E doesn't contain any reference. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	ade904e356	vmm: interrupt: generify impl InterruptSourceGroup for MsiInterruptGroup At this point we can use the same logic for all hypervisor implementations. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	2b466ed80c	vmm: interrupt: provide MsiInterruptGroupOps trait Currently it only contains a function named set_gsi_routes. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	b2abead65b	vmm: interrupt: provide and use extension trait RoutingEntryExt This trait contains a function which produces a interrupt routing entry. Implement that trait for KvmRoutingEntry and rewrite the update function. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	4dbca81b86	vmm: interrupt: rename set_kvm_gsi_routes to set_gsi_routes This function will be used to commit routing information to the hypervisor. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	fd7b42e54d	vmm: interrupt: inline mask_kvm_entry The logic for looking up the correct interrupt can be shared among hypervisors. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	0ec39da90c	vmm: interrupt: generify KvmMsiInterruptManager The observation is only the route entry is hypervisor dependent. Keep a definition of KvmMsiInterruptManager to avoid too much code churn. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	d5149e95cb	vmm: interrupt: generify KvmRoutingEntry and KvmMsiInterruptGroup The observation is that only the route field is hypervisor specific. Provide a new function in blanket implementation. Also redefine KvmRoutingEntry with RoutingEntry to avoid code churn. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	637f58bcd9	vmm: interrupt: drop Kvm prefix from KvmLegacyUserspaceInterruptManager This data structure doesn't contain KVM specific stuff. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
Wei Liu	574cab6990	vmm: interrupt: create GSI hashmap directly The observation is that the GSI hashmap remains untouched before getting passed into the MSI interrupt manager. We can create that hashmap directly in the interrupt manager's new function. The drops one import from the interrupt module. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-30 12:09:42 +01:00
dependabot-preview[bot]	f3c8f827cc	build(deps): bump linux-loader from `2a62f21` to `ec930d7` Bumps [linux-loader](https://github.com/rust-vmm/linux-loader) from `2a62f21` to `ec930d7`. - [Release notes](https://github.com/rust-vmm/linux-loader/releases) - [Commits](`2a62f21b44...ec930d700f`) Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-06-30 07:05:06 +00:00
Rob Bradford	522d8c8412	vmm: openapi: Add the /vm.counters API entry point This is a hash table of string to hash tables of u64s. In JSON these hash tables are object types. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-27 00:07:47 +02:00
Sebastien Boeuf	86377127df	vmm: Resume devices after vCPUs have been resumed Because we don't want the guest to miss any event triggered by the emulation of devices, it is important to resume all vCPUs before we can resume the DeviceManager with all its associated devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-25 12:01:34 +02:00
Sebastien Boeuf	f6eeba781b	vmm: Save and restore vCPU states during pause/resume operations We need consistency between pause/resume and snapshot/restore operations. The symmetrical behavior of pausing/snapshotting and restoring/resuming has been introduced recently, and we must now ensure that no matter if we're using pause/resume or snapshot/restore features, the resulting VM should be running in the exact same way. That's why the vCPU state is now stored upon VM pausing. The snapshot operation being a simple serialization of the previously saved state. The same way, the vCPU state is now restored upon VM resuming. The restore operation being a simple deserialization of the previously restored state. It's interesting to note that this patch ensures time consistency from a guest perspective, no matter which clocksource is being used. From a previous patch, the KVM clock was saved/restored upon VM pause/resume. We now have the same behavior for TSC, as the TSC from the vCPUs are saved/restored upon VM pause/resume too. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-25 12:01:34 +02:00
Sebastien Boeuf	18e7d7a1f7	vmm: cpu: Resume before shutdown in a specific way Instead of calling the resume() function from the CpuManager, which involves more than what is needed from the shutdown codepath, and potentially ends up with a deadlock, we replace it with a subset. The full resume operation is reserved for a VM that has been paused. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-25 12:01:34 +02:00
Sebastien Boeuf	65132fb99d	vmm: Implement Pausable trait for Vcpu We want each Vcpu to store the vCPU state upon VM pausing. This is the reason why we need to explicitly implement the Pausable trait for the Vcpu structure. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-25 12:01:34 +02:00
Wei Liu	1741af74ed	hypervisor: add safety statement in set_user_memory_region When set_user_memory_region was moved to hypervisor crate, it was turned into a safe function that wrapped around an unsafe call. All but one call site had the safety statements removed. But safety statement was not moved inside the wrapper function. Add the safety statement back to help reasoning in the future. Also remove that one last instance where the safety statement is not needed . No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-25 10:25:13 +02:00
Wei Liu	b27439b6ed	arch, hypervisor, vmm: KvmHyperVisor -> KvmHypervisor "Hypervisor" is one word. The "v" shouldn't be capitalised. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-25 10:25:13 +02:00
Wei Liu	b00171e17d	vmm: use MemoryRegion where applicable That removes one more KVM-ism in VMM crate. Note that there are more KVM specific code in those files to be split out, but we're not at that stage yet. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-25 10:25:13 +02:00
Rob Bradford	d983c0a680	vmm: Expose counters from virtio devices to API Collate the virtio device counters in DeviceManager for each device that exposes any and expose it through the recently added HTTP API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-25 07:02:44 +02:00
Rob Bradford	bca8a19244	vmm: Implement HTTP API for obtaining counters The counters are a hash of device name to hash of counter name to u64 value. Currently the API is only implemented with a stub that returns an empty set of counters. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-25 07:02:44 +02:00
Rob Bradford	fd4aba8eae	vmm: api: Implement support for GET handlers EndpointHandler This can be used for simple API requests which return data but do not require any input. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-25 07:02:44 +02:00
Rob Bradford	80be393b16	vmm: api: Order HTTP entry points in alphabetical order Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-25 07:02:44 +02:00
Wei Liu	4cc37d7b9a	vmm: interrupt: drop a few pub keywords Those items are not used elsewhere. Restrict their scope. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-24 12:39:42 +02:00
Wei Liu	1661adbbaf	vmm: interrupt: add "Kvm" prefix to MsiInterruptGroup The structure is tightly coupled with KVM. It uses KVM specific structures and calls. Add Kvm prefix to it. Microsoft hypervisor will implement its own interrupt group(s) later. No functional change intended. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-24 12:39:42 +02:00
Sebastien Boeuf	9f4714c32a	vmm: Extend seccomp filters with KVM_KVMCLOCK_CTRL Now that the VMM uses KVM_KVMCLOCK_CTRL from the KVM API, it must be added to the seccomp filters list. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-24 12:38:56 +02:00
Sebastien Boeuf	4a81d65f79	vmm: Notify the guest about vCPUs being paused Through the newly added API notify_guest_clock_paused(), this patch improves the vCPU pause operation by letting the guest know that each vCPU is being paused. This is important to avoid soft lockups detection from the guest that could happen because the VM has been paused for more than 20 seconds. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-24 12:38:56 +02:00
Sebastien Boeuf	9fa8438063	vmm: Fill CpuManager's vCPU list on restore path It's important that on restore path, the CpuManager's vCPU gets filled with each new vCPU that is being created. In order to cover both boot and restore paths, the list is being filled from the common function create_vcpu(). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-24 12:38:56 +02:00
Sebastien Boeuf	f5150aa261	vmm: Extend seccomp filters with KVM_GET_CLOCK and KVM_SET_CLOCK Now that the VMM uses both KVM_GET_CLOCK and KVM_SET_CLOCK from the KVM API, they must be added to the seccomp filters list. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 14:36:01 +01:00
Sebastien Boeuf	8038161861	vmm: Get and set clock during pause and resume operations In order to maintain correct time when doing pause/resume and snapshot/restore operations, this patch stores the clock value on pause, and restore it on resume. Because snapshot/restore expects a VM to be paused before the snapshot and paused after the restore, this covers the migration use case too. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 14:36:01 +01:00
Rob Bradford	4b64f2a027	vmm: cpu: Reuse already allocated vCPUs if available When a request is made to increase the number of vCPUs in the VM attempt to reuse any previously removed (and hence inactive) vCPUs before creating new ones. This ensures that the APIC ID is not reused for a different KVM vCPU (which is not allowed) and that the APIC IDs are also sequential. The two key changes to support this are: * Clearing the "kill" bit on the old vCPU state so that it does not immediately exit upon thread recreation. * Using the length of the vcpus vector (the number of allocated vcpus) rather than the number of active vCPUs (.present_vcpus()) to determine how many should be created. This change also introduced some new info!() debugging on the vCPU creation/removal path to aid further development in the future. TEST=Expanded test_cpu_hotplug test. Fixes: #1338 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-23 14:11:14 +01:00
Rob Bradford	9dcd0c37f3	vmm: cpu: Clear the "kill" flag on vCPU to support reuse After the vCPU has been ejected and the thread shutdown it is useful to clear the "kill" flag so that if the vCPU is reused it does not immediately exit upon thread recreation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-23 14:11:14 +01:00
Rob Bradford	b107bfcf2c	vmm: cpu: Add info!() level debugging to vCPU handling These messages are intended to be useful to support debugging related to vCPU hotplug/unplug issues. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-23 14:11:14 +01:00
Sebastien Boeuf	e382dc6657	vmm, vm-virtio: Restore DeviceManager's devices in a paused state The same way the VM and the vCPUs are restored in a paused state, all devices associated with the device manager must be restored in the same paused state. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 10:15:03 +02:00
Sebastien Boeuf	8a165b5314	vmm: Restore the VM in "paused" state Because we need to pause the VM before it is snapshot, it should be restored in a paused state to keep the sequence symmetrical. That's the reason why the state machine regarding the valid VM's state transition needed to be updated accordingly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 10:15:03 +02:00
Sebastien Boeuf	a16414dc87	vmm: Restore vCPUs in "paused" state To follow a symmetrical model, and avoid potential race conditions, it's important to restore a previously snapshot VM in a "paused" state. The snapshot operation being valid only if the VM has been previously paused. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 10:15:03 +02:00
Wei Liu	7552f4db61	vmm: device_manager: restore error handling When the hypervisor crate was introduced, a few places that handled errors were commented out in favor of unwrap, but that's bad practice. Restore proper error handling in those places in this patch. We cannot use from_raw_os_error anymore because it is wrapped deep under hypervisor crate. Create new custom errors instead. Fixes: `e4dee57e81` ("arch, pci, vmm: Initial switch to the hypervisor crate") Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-22 22:02:21 +01:00
Muminul Islam	cca59bc52f	hypervisor, arch: Fix warnings introduced in hypervisor crate This commit fixes some warnings introduced in the previous hyperviosr crate PR.Removed some unused variables from arch/aarch64 module. Signed-off-by: Muminul Islam <muislam@microsoft.com>	2020-06-22 21:58:45 +01:00
Rob Bradford	d714efe6d4	vmm: cpu: Import CpuTopology conditionally on x86_64 only The aarch64 build has no use for this structure at the moment. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-22 15:00:27 +01:00
Sebastien Boeuf	a998e89375	build(deps): bump signal-hook from 0.1.15 to 0.1.16 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.1.15 to 0.1.16. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](vorner/signal-hook@v0.1.15...v0.1.16) Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-22 14:09:11 +01:00
Muminul Islam	e4dee57e81	arch, pci, vmm: Initial switch to the hypervisor crate Start moving the vmm, arch and pci crates to being hypervisor agnostic by using the hypervisor trait and abstractions. This is not a complete switch and there are still some remaining KVM dependencies. Signed-off-by: Muminul Islam <muislam@microsoft.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-06-22 15:03:15 +02:00
Rob Bradford	a74c6fc14f	vmm, arch: x86_64: Fill the CPUID leaves with the topology There are two CPUID leaves for handling CPU topology, 0xb and 0x1f. The difference between the two is that the 0x1f leaf (Extended Topology Leaf) supports exposing multiple die packages. Fixes: #1284 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-17 12:18:09 +02:00
Rob Bradford	e19079782d	vmm, arch: x86_64: Set the APIC ID on the 0x1f CPUID leaf The extended topology leaf (0x1f) also needs to have the APIC ID (which is the KVM cpu ID) set. This mirrors the APIC ID set on the 0xb topology leaf Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-17 12:18:09 +02:00
Rob Bradford	b81bc77390	vmm: cpu: Save CpusConfig into CpuManager Rather than saving the individual parts into the CpuManager save the full struct as it now also contains the topology data. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-17 12:18:09 +02:00
Rob Bradford	4a0439a993	vmm: config: Extend CpusConfig to add the topology This allows the user to optionally specify the desired CPU topology. All parts of the topology must be specified and the product of all parts must match the maximum vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-17 12:18:09 +02:00
Wei Liu	103cd61bd2	vmm: device_tree: make available remove function unconditionally Its test case calls remove unconditionally. Instead of making the test code call remove conditionally, removing the pci_support dependency simplifies things -- that function is just a wrapper around HashMap's remove function anyway. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-15 11:41:34 +02:00
Wei Liu	fb461c820f	vmm: vm: enable test_vm test case Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-12 14:46:58 +01:00
Wei Liu	b99b5777bb	vmm: vm: move some imports into test_vm They are only needed there. Not moving them causes rustc to complain about unused imports. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-12 14:46:58 +01:00
Sebastien Boeuf	b62d5d22ff	vmm: openapi: Update the OpenAPI definition Now that PCI device hotplug returns a response, the OpenAPI definition must reflect it, describing what is expected to be received. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	4fe7347fb9	vmm: Manually implement Serialize for PciDeviceInfo In order to provide a more comprehensive b/d/f to the user, the serialization of PciDeviceInfo is implemented manually to control the formatting. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	83cd9969df	vmm: Enable HTTP response for PCI device hotplug This patch completes the series by connecting the dots between the HTTP frontend and the device manager backend. Any request to hotplug a VFIO, disk, fs, pmem, net, or vsock device will now return a response including the device name and the place of the device in the PCI topology. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	3316348d4c	vmm: vm: Carry information from hotplugged PCI device Pass from the device manager to the calling code the information about the PCI device that has just been hotplugged. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	f08e9b6a73	vmm: device_manager: Return PciDeviceInfo from a hotplugged device In order to provide the device name and PCI b/d/f associated with a freshly hotplugged device, the hotplugging functions from the device manager return a new structure called PciDeviceInfo. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	0bc2b08d3a	vmm: api: Return an optional response from vm_action() Any action that relies on vm_action() can now return a response body. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	038180269e	vmm: api: Allow HTTP PUT request to return a response Adding the codepath to return a response from a PUT request. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Wei Liu	5ebd02a572	vmm: vm: fix test_vm test case We should break out from the loop after getting the HLT exit, otherwise the VM hangs forever. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-12 08:38:07 +02:00
Michael Zhao	97a1e5e1d2	vmm: Exit VMM event loop after guest shutdown for AArch64 X86 and AArch64 work in different ways to shutdown a VM. X86 exit VMM event loop through ACPI device; AArch64 need to exit from CPU loop of a SystemEvent. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	5cd1730bc4	vmm: Configure VM on AArch64 Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	917219fa92	vmm: Enable VCPU for AArch64 Added MPIDR which is needed in system configuration. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	b5f1c912d6	vmm: Enable memory manager for AArch64 Screened IO space as it is not available on AArch64. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	eeeb45bbb9	vmm: Enable device manager for AArch64 Screened IO bus because it is not for AArch64. Enabled Serial, RTC and Virtio devices with MMIO transport option. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	e9488846f1	vm-allocator: Enable vm-allocator for AArch64 Implemented GSI allocator and system allocator for AArch64. Renamed some layout definitions to align more code between architectures. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Anatol Belski	abd6204d27	source: Fix file permissions Rust sources and some data files should not be executable. The perms are set to 644. Signed-off-by: Anatol Belski <ab@php.net>	2020-06-10 18:47:27 +01:00
Sebastien Boeuf	653087d7a3	vmm: Reduce MMIO address space by 4KiB In order to workaround a Linux bug that happens when we place devices at the end of the physical address space on recent hardware (52 bits limit) we reduce the MMIO address space by one 4k page. This way, nothing gets allocated in the last 4k of the address space, which is negligible given the amount of space in the address space. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-09 18:08:09 +01:00
Bo Chen	625bab69bd	vmm: api: Allow to delete non-booted VMs The action of "vm.delete" should not report errors on non-booted VMs. This patch also revised the "docs/api.md" to reflect the right 'Prerequisites' of different API actions, e.g. on "vm.delete" and "vm.boot". Fixes: #1110 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-06-09 05:58:32 +01:00
Rob Bradford	9b71ba20ac	vmm, vm-virtio: Stop always autogenerating a host MAC address This removes the need to use CAP_NET_ADMIN privileges and instead the host MAC addres is either provided by the user or alternatively it is retrieved from the kernel. TEST=Run cloud-hypervisor without CAP_NET_ADMIN permission and a preconfigured tap device: sudo ip tuntap add name tap0 mode tap sudo ifconfig tap0 192.168.249.1 netmask 255.255.255.0 up cargo clean cargo build target/debug/cloud-hypervisor --serial tty --console off --kernel ~/src/rust-hypervisor-firmware/target/target/release/hypervisor-fw --disk path=~/workloads/clear-33190-kvm.img --net tap=tap0 VM was also rebooted to check that works correctly. Fixes: #1274 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-08 17:56:10 +02:00
Rob Bradford	929d70bc7f	net_util: Only try and enable the TAP device if it not already enabled This allows an existing TAP interface to be used without needing CAP_NET_ADMIN permissions on the Cloud Hypervisor binary as the ioctl to bring up the interface is avoided. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-08 17:56:10 +02:00
Bo Chen	a8cdf2f070	tests,vm-virtio,vmm: Use 'socket' for all CLI/API parameters This patch unifies the inconsistent uses of 'socket' and 'sock' from our CLI/API parameters. Fixes: #1091 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-06-08 17:41:12 +02:00
Samuel Ortiz	3336e80192	vfio: Switch to the vfio-ioctls crate ch branch Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-06-04 08:48:55 +02:00
Samuel Ortiz	d24aa72d3e	vfio: Rename to vfio-ioctls Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-06-04 08:48:55 +02:00
Samuel Ortiz	53ce529875	vfio: Move the PCI implementation to the PCI crate There is a much stronger PCI dependency from vfio_pci.rs than a VFIO one from pci/src/vfio.rs. It seems more natural to have the PCI specific VFIO implementation in the PCI crate rather than the other way around. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-06-04 08:48:55 +02:00
Michael Zhao	8f7dc73562	vmm: Move Vcpu::configure() to arch crate Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-03 11:27:29 +02:00
Michael Zhao	969e5e0b51	vmm: Split configure_system() from load_kernel() for x86_64 Now the flow of both architectures are aligned to: 1. load kernel 2. create VCPU's 3. configure system 4. start VCPU's Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-03 11:27:29 +02:00
Michael Zhao	20cf21cd9d	vmm: Change booting process to cover AArch64 requirements Between X86 and AArch64, there is some difference in booting a VM: - X86_64 can setup IOAPIC before creating any VCPU. - AArch64 have to create VCPU's before creating GIC. The old process is: 1. load_kernel() load kernel binary configure system 2. activate_vcpus() create & start VCPU's So we need to separate "activate_vcpus" into "create_vcpus" and "activate_vcpus" (to start vcpus only). Setup GIC and create FDT between the 2 steps. The new procedure is: 1. load_kernel() load kernel binary (X86_64) configure system 2. create VCPU's 3. (AArch64) setup GIC 4. (AArch64) configure system 5. start VCPU's Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-03 11:27:29 +02:00
dependabot-preview[bot]	aac87196d6	build(deps): bump vm-memory from 0.2.0 to 0.2.1 Bumps [vm-memory](https://github.com/rust-vmm/vm-memory) from 0.2.0 to 0.2.1. - [Release notes](https://github.com/rust-vmm/vm-memory/releases) - [Changelog](https://github.com/rust-vmm/vm-memory/blob/v0.2.1/CHANGELOG.md) - [Commits](https://github.com/rust-vmm/vm-memory/compare/v0.2.0...v0.2.1) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-05-28 17:06:48 +01:00
Rob Bradford	c31ad72ee9	build: Address issues found by 1.43.0 clippy These are mostly due to use of "bare use" statements and unnecessary vector creation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-27 19:32:12 +02:00
Bo Chen	fbd1a6c5f1	vmm: api: Return complete error responses in handle_http_request() Instead of responding only headers with error code, we now return complete error responses to HTTP requests with errors (e.g. undefined endpoints and InternalSeverError). Fixes: #472 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-05-27 18:29:52 +01:00
Rob Bradford	0728bece0c	vmm: seccomp: Ensure that umask() can be reprogrammed When doing self spawning the child will attempt to set the umask() again. Let it through the seccomp rules so long as it the safe mask again. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-27 16:46:51 +01:00
dependabot-preview[bot]	a4bb96d45c	build(deps): bump libc from 0.2.70 to 0.2.71 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.70 to 0.2.71. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.70...0.2.71) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-05-27 09:02:13 +02:00
Michael Zhao	8f1f9d9e6b	devices: Implement InterruptController on AArch64 This commit only implements the InterruptController crate on AArch64. The device specific part for GIC is to be added. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-05-26 11:09:19 +02:00
Michael Zhao	b32d3025f3	devices: Refactor IOAPIC to cover other architectures IOAPIC, a X86 specific interrupt controller, is referenced by device manager and CPU manager. To work with more architectures, a common type for all architectures is needed. This commit introduces trait InterruptController to provide architecture agnostic functions. Device manager and CPU manager can use it without caring what the underlying device is. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-05-26 11:09:19 +02:00
Michael Zhao	1befae872d	build: Fixed build errors and warnings on AArch64 This is a preparing commit to build and test CH on AArch64. All building issues were fixed, but no functionality was introduced. For X86, the logic of code was not changed at all. For ARM, the architecture specific part is still empty. And we applied some tricks to workaround lint warnings. But such code will be replaced later by other commits with real functionality. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-05-21 11:56:26 +01:00
Rob Bradford	af8292b623	vmm, config, vhost_user_blk: remove "wce" parameter This config option provided very little value and instead we now enable this feature (which then lets the guest control the cache mode) unconditionally. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-21 08:40:43 +02:00
Bo Chen	7c3e19c65a	vhost_user_backend, vmm: Close leaked file descriptors Explicit call to 'close()' is required on file descriptors allocated from 'epoll::create()', which is missing for the 'EpollContext' and 'VringWorker'. This patch enforces to close the file descriptors by reusing the Drop trait of the 'File' struct. Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-05-19 09:22:09 +02:00
Rob Bradford	1b8b5ac179	vhost-user_net, vm-virtio, vmm: Permit host MAC address setting Add a new "host_mac" parameter to "--net" and "--net-backend" and use this to set the MAC address on the tap interface. If no address is given one is randomly assigned and is stored in the config. Support for vhost-user-net self spawning was also included. Fixes: #1177 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-15 11:45:09 +01:00
Rob Bradford	11049401ce	vmm: seccomp: Add ioctl() commands interface hardware address This is necessary to support setting the MAC address on the tap interface on the host. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-15 11:45:09 +01:00
Sebastien Boeuf	68fc432978	vmm: Update seccomp filters with clock_nanosleep The clock_nanosleep system call needs to be whitelisted since the commit `12e00c0f45` introduced the use of a sleep() function. Without this patch, we can see an error when the VM is paused or killed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-15 12:34:53 +02:00
Rob Bradford	6aa29bdb24	vmm: api: Use a common handler for data actions too Like the actions that don't take data such as "pause" or "resume" use a common handler implementation to remove duplicated code for handling simple endpoints like the hotplug ones. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	0fe223f00e	vmm: api: Extend VmAction to reduce code duplication Many of the API requests take a similar form with a single data item (i.e. config for a device hotplug) expand the VmAction enum to handle those actions and a single function to dispatch those API events. For now port the existing helper functions to use this new API. In the future the HTTP layer can create the VmAction directly avoiding the extra layer of indirection. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	6ec605a7fb	vmm: api: Refactor generic action handler Rather than save the save a function pointer and use that instead the underlying action. This is useful for two reasons: 1. We can ensure that we generate HttpErrors in the same way as the other endpoints where API error variant should be determined by the request being made not the underlying error. 2. It can be extended to handle other generic actions where the function prototype differs slightly. As result of this refactoring it was found that the "vm.delete" endpoint was not connected so address that issue. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	c652625beb	vmm: api: Add a default implementation for simple PUT requests Extend the EndpointHandler trait to include automatic support for handling PUT requests. This will allow the removal of lots of duplicated code in the following commit from the API handling code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	a3e8bea03c	vmm: api: Move HttpError enum to http module Minor rearrangement of code to make it easier to implement refactoring. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	9ccc7daa83	build, vmm: Update to latest kvm-ioctls The ch branch has been rebased to incorporate the latest upstream code requiring a small change to the unit tests. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-13 17:14:49 +02:00
Rob Bradford	88ec93d075	vmm: config: Add missing "id" from FsConfig parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-13 09:11:50 +01:00
dependabot-preview[bot]	2991fd2a48	build(deps): bump libc from 0.2.69 to 0.2.70 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.69 to 0.2.70. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.69...0.2.70) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-05-12 20:26:43 +02:00
Sebastien Boeuf	c37da600e8	vmm: Update DeviceTree upon PCI BAR reprogramming By passing a reference of the DeviceTree to the AddressManager, we can now update the DeviceTree whenever a PCI BAR is reprogrammed. This is mandatory to maintain the correct resources information related to each virtio-pci device, which will ensure correct information will be stored upon VM snapshot. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Sebastien Boeuf	d0ae9d7ce6	vmm: Share the DeviceTree across threads We want to be able to share the same DeviceTree across multiple threads, particularly to handle the use case where PCI BAR reprogramming might need to update the tree while from another thread a new device is being added to the tree. That's why this patch moves the DeviceTree instance into an Arc<Mutex<>> so that we can later share a reference of the same mutable tree with the AddressManager responsible for handling PCI BAR reprogramming. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Sebastien Boeuf	5e9d254564	vmm: Store and restore virtio-pci BAR resources By using the vector of resources provided by the DeviceNode, the device manager can store the information related to PCI BARs from a virtio-pci device. Based on this, and upon VM restoration, the device manager can restore the BARs in the expected location in the guest address space. One thing to note is that we only need to provide the VirtioPciDevice with the configuration BAR (BAR 0) since the SHaredMemory BAR info comes from the virtio device directly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Sebastien Boeuf	8a826ae24c	vmm: Store and restore virtio-pci device on right PCI slot Based on the new field "pci_bdf", a virtio-pci device can be restored at the same place on the PCI bus it was located before the VM snapshot. This ensures consistent placement on the PCI bus, based on the stored information related to each device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Sebastien Boeuf	98dac352b8	vmm: Add optional PCI b/d/f to each DeviceNode We need a way to store the information about where a PCI device was placed on the PCI bus before the VM was snapshotted. The way to do this is by adding an extra field to the DeviceNode structure. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Rob Bradford	b9ba81c30d	arch, vmm: Don't build mptable when using ACPI Use the ACPI feature to control whether to build the mptable. This is necessary as the mptable and ACPI RSDP table can easily overwrite each other leading to it failing to boot. TEST=Compile with default features and see that --cpus boot=48 now works, try with --no-default-features --features "pci" and observe the --cpus boot=48 also continues to work. Fixes: #1132 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-11 19:34:34 +01:00
dependabot-preview[bot]	1c44e917f9	build(deps): bump clap from 2.33.0 to 2.33.1 Bumps [clap](https://github.com/clap-rs/clap) from 2.33.0 to 2.33.1. - [Release notes](https://github.com/clap-rs/clap/releases) - [Changelog](https://github.com/clap-rs/clap/blob/v2.33.1/CHANGELOG.md) - [Commits](https://github.com/clap-rs/clap/compare/v2.33.0...v2.33.1) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-05-11 20:15:13 +02:00
dependabot-preview[bot]	4cd2eccf2f	build(deps): bump signal-hook from 0.1.14 to 0.1.15 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.1.14 to 0.1.15. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/compare/v0.1.14...v0.1.15) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-05-11 20:15:03 +02:00
Rob Bradford	5016fcf8d5	vhost_user_block: Use config::OptionParser to simplify block backend parsing Switch to using the recently added OptionParser in the code that parses the block backend. Fixes: #1092 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-11 09:40:40 +02:00
Rob Bradford	592de97fbd	vhost_user_net: Use config::OptionParser to simplify net backend parsing Switch to using the recently added OptionParser in the code that parses the network backend. Whilst doing this also update the net-backend syntax to use "sock" rather than socket. Fixes: #1092 Partially fixes: #1091 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-11 09:40:40 +02:00
Rob Bradford	12e00c0f45	vmm: cpu: Retry sending signals if necessary To avoid a race condition where the signal might "miss" the KVM_RUN ioctl() instead reapeatedly try sending a signal until the vCPU run is interrupted (as indicated by setting a new per vCPU atomic.) It important to also clear this atomic when coming out of a paused state. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Rob Bradford	31bde4f5da	vmm: Unpark the DeviceManager threads in shutdown To ensure that the DeviceManager threads (such as those used for virtio devices) are cleaned up it is necessary to unpark them so that they get cleanly terminated as part of the shutdown. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Rob Bradford	801e72ac6d	vmm: cpu: Unpause vCPU threads After setting the kill signal flag for the vCPU thread release the pause flag and unpark the threads. This ensures that that the vCPU thread will wake up and check the kill signal flag if the VM is paused. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Rob Bradford	91a4a2581e	vmm: cpu: When coming out of the pause event check for a kill signal Rather than immediately entering the vCPU run() code check if the kill signal is set. This allows paused VMs to be shutdown. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Rob Bradford	cd60de8f7f	Revert "vmm: vm: Unpark the threads before shutdown when the current state is paused" This reverts commit `e1a07ce3c4`. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Sebastien Boeuf	f6a71bec36	vmm: Add unit tests for DeviceTree Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	64e01684f9	vmm: Create new module device_tree This module will be dedicated to DeviceNode and DeviceTree definitions along with some dedicated unit tests. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	3b77be903d	vmm: Add device_node!() macro to improve code readability Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	83ec716ec4	vmm: Create breadth-first search iterator for the DeviceTree This iterator will let the VMM enumerate the resources associated with the DeviceManager, allowing for introspection. Moreover, by implementing a double ended iterator, we can get the hierarchy from the leaves to the root of the tree, which is very helpful in the context of restoring the devices in the right order. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	b91ab1e3a5	vmm: Remove the list of migratable devices Now that the device tree fully replaced the need for a dedicated list of migratable devices, this commit cleans up the codebase by removing it from the DeviceManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	1be7037229	vmm: Don't use migratable_devices for restore This commit switches from migratable_devices to device_tree in order to restore devices exclusively based on the device tree. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	bc6084390f	vmm: Add migratable field to the DeviceNode This commit adds an extra field to the DeviceNode so that the structure can hold a Migratable device. The long term plan is to be able to remove the dedicated table of migratable devices, but instead rely only on the device tree. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	7fec020f53	vmm: Create a dedicated DeviceTree structure In order to hide the complexity chosen for the device tree stored in the DeviceManager, we introduce a new DeviceTree structure. For now, this structure is a simple passthrough of a HashMap, but it can be extended to handle some DeviceTree specific operations. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	14b379dec5	vmm: Add an identifier field to DeviceNode structure Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	0805d458c4	vmm: Add support for multiple children per DeviceNode Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	daaeba5142	vmm: Change Node into DeviceNode Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	5c7df03efe	vmm: Store and restore virtio-pmem resources This device has a dedicated memory region in the guest address space, which means in case of snapshot/restore, it must be restored in the exact same location it was during the snapshot. That's through the resources that we can describe the location of this extra memory region, allowing the device for correct restoring. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	2e6895d911	vmm: Store and restore virtio-fs resources This device has a dedicated memory region in the guest address space, which means in case of snapshot/restore, it must be restored in the exact same location it was during the snapshot. That's through the resources that we can describe the location of this extra memory region, allowing the device for correct restoring. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	987f82152e	vmm: Store and restore virtio-mmio resources Based on the device tree, retrieve the resources associated with a virtio-mmio device to restore it at the right location in guest address space. Also, the IRQ number is correctly restored. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	9cb1e1cc6b	vmm: Perform MMIO allocation from virtio-mmio device creation Instead of splitting the MMIO allocation and the device creation into separate functions for virtio-mmio devices, it's is easier to move everything into the same function as we'll be able to gather resources in the same place for the same device. These resources will be stored in the device tree in a follow up patch. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	adf297066d	vmm: Create devices in different path if restoring the VM In case the VM is created from scratch, the devices should be created after the DeviceManager has been created. But this should not affect the restore codepath, as in this case the devices should be created as part of the restore() function. It's necessary to perform this differentiation as the restore must go through the following steps: - Create the DeviceManager - Restore the DeviceManager with the right state - Create the devices based on the restored DeviceManager's device tree - Restore each device based on the restored DeviceManager's device tree That's why this patch leverages the recent split of the DeviceManager's creation to achieve what's needed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	d39f91de02	vmm: Reorganize DeviceManager creation This commit performs the split of the DeviceManager's creation into two separate functions by moving anything related to device's creation after the DeviceManager structure has been initialized. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	89c2a5868c	vmm: Restore devices following the device tree Based on the device tree, we now ensure the restore can be done in the right order, as it will respect the dependencies between nodes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	52c80cfcf5	vmm: Snapshot and restore DeviceManager state The DeviceManager itself must be snapshotted in order to store the information regarding the devices associated with it, which effectively means we need to store the device tree. The mechanics to snapshot and restore the DeviceManagerState are added to the existing snapshot() and restore() implementations. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	5b408eec66	vmm: Create a device tree The DeviceManager now creates a tree of devices in order to store the resources associated with each device, but also to track dependencies between devices. This is a key part for proper introspection, but also to support snapshot and restore correctly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Rob Bradford	fec97e0586	vm-virtio, vmm: Delete unix socket on shutdown It's not possible to call UnixListener::Bind() on an existing file so unlink the created socket when shutting down the Vsock device. This will allow the VM to be rebooted with a vsock device. Fixes: #1083 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-05 13:01:38 +02:00
Rob Bradford	5109f914eb	vmm: config: Reject attempts to use VFIO or IOMMU without PCI Generate an error during validation if an attempt it made to place a device behind an IOMMU or using a VFIO device when not using PCI. Fixes: #751 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-05 11:20:52 +01:00
dependabot-preview[bot]	5571c6af2d	build(deps): bump signal-hook from 0.1.13 to 0.1.14 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.1.13 to 0.1.14. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/compare/v0.1.13...v0.1.14) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-05-02 19:01:33 +01:00
Rob Bradford	5115ad6e56	vmm: config: Support on/off/true/false for all booleans Migrate missing boolean controls over to the Toggle to handle all values. Fixes: #936 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-30 15:21:09 +02:00
Rob Bradford	d5bfa2dfc8	vmm, vhost_user_block: Make parameter names match --disk Make the --block-backend parameters match the --disk parameters. Fixes: #898 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-30 15:20:55 +02:00
Sebastien Boeuf	2f0bc06bec	vmm: Update default devices names as "internal" Let's put an underscore "_" in front of each device name to identify when it has been set internally. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	aaba6e777f	vmm: Add virtio-console to the list of Migratable devices The virtio-console was not added to the list of Migratable devices, which is fixed from this patch. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9ab4bb1ae2	devices: serial: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	06487131f9	vm-virtio: pci: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. It is based off the name from the virtio device attached to this transport layer. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	eeb7e10d1f	vm-virtio: mmio: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. It is based off the name from the virtio device attached to this transport layer. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9d84ef5073	vmm: Make the virtio identifier mandatory Because we know we will need every virtio device to be identified with a unique id, we can simplify the code by making the identifier mandatory. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	14350f5de4	devices: ioapic: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	556871570e	vm-virtio: iommu: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	052eff1ca7	vm-virtio: console: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	354c2a4b3d	vm-virtio: vhost-user-net: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	46e0b3ff75	vm-virtio: vhost-user-blk: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	bb7fa71fcb	vm-virtio: vhost-user-fs: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	ec5ff395cf	vm-virtio: vsock: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9b53044aae	vm-virtio: mem: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	1592a9292f	vm-virtio: pmem: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	2e91b73881	vm-virtio: rng: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9eb7413fab	vm-virtio: net: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	be946caf4b	vm-virtio: blk: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	ff9c8b847f	vmm: Always generate the next device name Even in the context of "mmio" feature, we need the next device name to be generated as we need to identify virtio-mmio devices to support snapshot and restore functionalities. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	8183141399	vmm: Add an identifier to the ioapic device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	e4386c8bb7	vmm: Add an identifier to the virtio-iommu device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	75ddd2a244	vmm: Add an identifier to the --console device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	eac350c454	vmm: Add an identifier to the virtio-mem device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	6802ef5406	vmm: Add an identifier to the --rng device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	d71d52e9b0	vmm: Fix virtio-console creation with virtual IOMMU If the virtio-console device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	b08fde5928	vmm: Fix virtio-rng creation with virtual IOMMU If the virtio-rng device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	8031ac33c3	vmm: Fix virtio-vsock creation with virtual IOMMU If the virtio-vsock device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Rob Bradford	8cef35745b	vmm: seccomp: Add fork, gettid and pipe2 syscalls to permitted list This is needed for self spawning with the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Rob Bradford	ce7678f29f	vmm: seccomp: Add tkill syscall to permitted list This is needed for rebooting on the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Rob Bradford	12758d7fad	vmm: seccomp: Add epoll_pwait syscall to permitted list This is needed for basic operation on the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Samuel Ortiz	86fcd19b8a	build: Initial musl support Fix all build failures and add musl to the gihub workflows. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-29 17:57:01 +01:00
Sebastien Boeuf	a5de49558e	vmm: Only allow removal of specific types of virtio device Now that all virtio devices are assigned with identifiers, they could all be removed from the VM. This is not something that we want to allow because it does not make sense for some devices. That's why based on the device type, we remove the device or we return an error to the user. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 13:33:19 +01:00
Sebastien Boeuf	9ed880d74e	vmm: Add an identifier to the --fs device By giving the devices ids this effectively enables the removal of the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 13:33:19 +01:00
Sebastien Boeuf	7e0ab6b56d	vmm: Fix pmem device creation The parameters regarding the attachment to the virtio-iommu device was not propagated correclty, and any modification to the configuration was not stored back into it. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 13:33:19 +01:00
Rob Bradford	8de7448d44	vmm: api: Add "add-vsock" API entry point This allows the hotplugging of vsock devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	bf09a1e695	openapi: Add "id" field to VsockConfig Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	a76cf0865f	vmm: vm: Remove vsock device from config When doing device unplug remove the vsock device from the configuration if present. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	99422324a7	vmm: vm: Add "add_vsock()" Add the vsock device to the device manager and patch the config to add the new vsock device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	1d61c476a1	vmm: device_manager: Add support for hotplugging virtio-vsock devices Create a new VirtioVsock device and add it to the PCI bus upon hotplug. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	f8501a3bd3	vmm: config: Move --vsock syntax to VsockConfig This means it can be reused with ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Sebastien Boeuf	6e049e0da1	vmm: Add an identifier to the --vsock device It's possible to have multiple vsock devices so in preparation for hotplug/unplug it is important to be able to have a unique identifier for each device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	10348f73e4	vmm, main: Support only zero or one vsock devices The Linux kernel does not support multiple virtio-vsock devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-28 20:07:18 +02:00
Rob Bradford	9d1f95a3cc	openapi: Add missing "id" field NetConfig/DiskConfig/PmemConfig/FsConfig were all missing the id field in the API yaml file. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-28 18:27:45 +02:00
Muminul Islam	e1a07ce3c4	vmm: vm: Unpark the threads before shutdown when the current state is paused If the current state is paused that means most of the handles got killed by pthread_kill We need to unpark those threads to make the shutdown worked. Otherwise The shutdown API hangs and the API is not responding afterwards. So before the shutdown call we need to resume the VM make it succeed. Fixes: #817 Signed-off-by: Muminul Islam <muislam@microsoft.com>	2020-04-27 09:09:12 +02:00
Rob Bradford	1df38daf74	vmm, tests: Make specifying a size optional for virtio-pmem If a size is specified use it (in particular this is required if the destination is a directory) otherwise seek in the file to get the size of the file. Add a new check that the size is a multiple of 2MiB otherwise the kernel will reject it. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-24 18:30:05 +01:00
Rob Bradford	7481e4d959	vmm: config: Validate that shared memory is enabled if using vhost-user Check that if any device using vhost-user (net & disk with vhost_user=true) or virtio-fs is enabled then check shared memory is also enabled. Fixes: #848 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-24 16:01:49 +01:00
Bo Chen	2ac6971a8b	vmm: MemoryManager: Cleanup the usage of std::ffi/io/result Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-04-23 21:39:51 +02:00
Bo Chen	3f42f86d81	vmm: Add the 'shared' and 'hugepages' controls to MemoryConfig The new 'shared' and 'hugepages' controls aim to replace the 'file' option in MemoryConfig. This patch also updated all related integration tests to use the new controls (instead of providing explicit paths to "/dev/shm" or "/dev/hugepages"). Fixes: #1011 Signed-off-by: Rob Bradford <robert.bradford@intel.com> Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-04-23 21:39:51 +02:00
Martin Xu	5a380a6918	vmm: memory_manager: Support non-power-of-2 block sizes Replace alignment calculation of start address with functionally equivalent version that does not assume that the block size is a power of two. Signed-off-by: Martin Xu <martin.xu@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-22 09:11:51 +02:00
Sebastien Boeuf	c22fd39170	vmm: Remove virtio device's userspace mapping on hot-unplug When a virtio device is dynamically removed from the VM through the hot-unplug mechanism, every mapping associated with it must be properly removed. Based on the previous patches letting a VirtioDevice expose the list of userspace mappings associated with it, this patch can now remove all the KVM userspace memory regions through the MemoryManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	0a97c25464	vmm: Extend MemoryManager to remove userspace mappings The same way we added a helper for creating userspace memory mappings from the MemoryManager, this patch adds a new helper to remove some previously added mappings. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	fbcf3a7a7a	vm-virtio: Implement userspace_mappings() for virtio-pmem When hot-unplugging the virtio-pmem from the VM, we don't remove the associated userspace mapping. This patch will let us fix this in a following patch. For now, it simply adapts the code so that the Pmem device knows about the mapping associated with it. By knowing about it, it can expose it to the caller through the new userspace_mappings() function. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	18f7789a81	vmm: Add hotplugged virtio devices to the DeviceManager list The hotplugged virtio devices were not added to the list of virtio devices from the DeviceManager. This patch fixes it, as it was causing hotplugged virtio-fs devices from not supporting memory hotplug, since they were never getting the update as they were not part of the list of virtio devices held by the DeviceManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-20 20:36:26 +02:00
Dean Sheather	c2abadc293	vmm: Add ability to add virtio-fs device post-boot Adds DeviceManager method `make_virtio_fs_device` which creates a single device, and modifies `make_virtio_fs_devices` to use this method. Implements the new `vm.add-fs route`. Signed-off-by: Dean Sheather <dean@coder.com>	2020-04-20 20:36:26 +02:00
Dean Sheather	bb2139a408	vmm/api: Add vm.add-fs route Currently unimplemented. Once implemented, this API will allow for creating virtio-fs devices in the VM after it has booted. Signed-off-by: Dean Sheather <dean@coder.com>	2020-04-20 20:36:26 +02:00
Sebastien Boeuf	d35e775ed9	vmm: Update KVM userspace mapping when PCI BAR remapping In the context of the shared memory region used by virtio-fs in order to support DAX feature, the shared region is exposed as a dedicated PCI BAR, and it is backed by a KVM userspace mapping. Upon BAR remapping, the BAR is moved to a different location in the guest address space, and the KVM mapping must be updated accordingly. Additionally, we need the VirtioDevice to report the updated guest address through the shared memory region returned by get_shm_regions(). That's why a new setter is added to the VirtioDevice trait, so that after the mapping has been updated for KVM, we can tell the VirtioDevice the new guest address the shared region is located at. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-20 16:01:25 +02:00
Sebastien Boeuf	ac7178ef2a	vmm: Keep migratable devices list as a Vec The order the elements are pushed into the list is important to restore them in the right order. This is particularly important for MmioDevice (or VirtioPciDevice) and their VirtioDevice counterpart. A device must be fully ready before its associated transport layer management can trigger its restoration, which will end up activating the device in most cases. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-17 19:29:41 +02:00
Rob Bradford	e7e0e8ac38	vmm, devices: Add firmware debug port device OVMF and other standard firmwares use I/O port 0x402 as a simple debug port by writing ASCII characters to it. This is gated under a feature that is not enabled by default. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-17 12:54:00 +02:00
Rob Bradford	f9a0445c3d	vmm: vm: Remove device from configuration after unplug This ensures that a device that is removed will not reappear after a reboot. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	444e5c2a04	vmm: device_manager: Generalise NoAvailableVfioDeviceName We now support assigning device ids for VFIO and virtio-pci devices so this error can be generalised. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	5bab9c3894	vmm: device_manager: Assign ids to pmem/net/disk devices if absent If the id has not been provided by the user generate an incrementing id. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	514491a051	vmm: device_manager: Support unplugging virtio-pci devices Extend the eject_device() method on DeviceManager to also support virtio-pci devices being unplugged. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	476e4ce24f	vmm: device_manager: Add virtio-pci devices into id to BDF map In order to support hotplugging there is a map of human readable device id to PCI BDF map. As the device id is part of the specific device configuration (e.g. NetConfig) it is necessary to return the id through from the helper functions that create the devices through to the functions that add those devices to the bus. This necessitates changing a great deal of function prototypes but otherwise has little impact. Currently only if an id is supplied by the user as part of the device configuration is it populated into this map. A later commit will populate with an autogenerated name where none is supplied by the user. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	b38470df4b	vmm: config: Add "id" parameter to {Net, Disk, Pmem}Config This id will be used to unplug the device if the user has chosen an id. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	1beb62ed2d	vmm: vm: Don't panic on kernel load error Rather than panic()ing when we get a kernel loading error populate the error upwards. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	72fdfff15d	vmm: device_manager: Remove unused "_mmap_regions" member Now that ownership of the memory regions used for the virtio-pmem and vhost-user-fs devices have been moved into those devices it is no longer necessary to track them inside DeviceManager. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Rob Bradford	70ecd6bab4	vmm, virtio: fs: Move freeing of mappped region into device Move the release of the managed memory region from the DeviceManager to the vhost-user-fs device. This ensures that the memory will be freed when the device is unplugged which will lead to it being Drop()ed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Rob Bradford	0c6706a510	vmm, virtio: pmem: Move freeing of mappped region into device Move the release of the managed memory region from the DeviceManager to the virtio-pmem device. This ensures that the memory will be freed when the device is unplugged which will lead to it being Drop()ed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Sebastien Boeuf	b1554642e4	vmm: seccomp: Add missing mremap() syscall While testing self spawned vhost-user backends, it appeared that the backend was aborting due to a missing system call in the seccomp filters. mremap() was the culprit and this patch simply adds it to the whitelist. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-14 14:11:41 +02:00
dependabot-preview[bot]	886c0f9093	build(deps): bump libc from 0.2.68 to 0.2.69 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.68 to 0.2.69. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.68...0.2.69) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-04-14 09:27:04 +01:00
Rob Bradford	28abfa9de5	vmm: openapi: Mark "initramfs" field nullable This should make it a pointer in the Go generated code so that it will be ommitted and thus not populated with an unhelpful default value. Fixes: #1015 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-09 23:25:18 +02:00
Rob Bradford	c260640fd5	vmm: config: Use Default::default() value for initramfs field This ensures that the field is filled with None when it is not specified as part of the deserialisation step. Fixes: #1015 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-09 17:28:45 +02:00
Alejandro Jimenez	7134f3129f	vmm: Allow PVH boot with initramfs We can now allow guests that specify an initramfs to boot using the PVH boot protocol. Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-04-09 17:28:03 +02:00
Rob Bradford	2d3f518c72	vmm: config: Error if both socket and path are specified for a disk This allows the validation of this requirement for both command line booted VMs and those booted via the API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	eeb7e2529d	vmm: config: Move max vCPUs > boot vCPUs check to validate() This allows the validation of this requirement for both command line booted VMs and those booted via the API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	12edb24678	vmm: config: Validate that serial/console file mode has a path Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	aaf382eee2	vmm: Move kernel check to VmConfig::validate() method Replace the existing VmConfig::valid() check with a call into .validate() as part of earlier config setup or boot API checks. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	3b0da2d895	vmm: vm: Validate configuration on API boot When performing an API boot validate the configuration. For now only some very basic validation is performed but in subsequent commits the validation will be extended. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	99b2ada4d0	vmm: Start splitting configuration parsing and validation The configuration comes from a variety of places (commandline, REST API and restore) however some validation was only happening on the command line parsing path. Therefore introduce a new ability to validate the configuration before proceeding so that this can be used for commandline and API boots. For now move just the console and serial output mode validation under the new validation API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Sebastien Boeuf	0ea706faf5	vmm: openapi: Update OpenAPI definition with RestoreConfig Making sure the OpenAPI definition is up to date with newly added structure and parameters to support VM restoration. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	8d9d22436a	vmm: Add "prefault" option when restoring Now that the restore path uses RestoreConfig structure, we add a new parameter called "prefault" to it. This will give the user the ability to populate the pages corresponding to the mapped regions backed by the snapshotted memory files. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	a517ca23a0	vmm: Move restore parameters into common RestoreConfig structure The goal here is to move the restore parameters into a dedicated structure that can be reused from the entire codebase, making the addition or removal of a parameter easier. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	6712958f23	vmm: memory: Add prefault option when creating region When CoW can be used, the VM restoration time is reduced, but the pages are not populated. This can lead to some slowness from the guest when accessing these pages. Depending on the use case, we might prefer a slower boot time for better performances from guest runtime. The way to achieve this is to prefault the pages in this case, using the MAP_POPULATE flag along with CoW. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	b2cdee80b6	vmm: memory: Restore with Copy-on-Write when possible This patch extends the previous behavior on the restore codepath. Instead of copying the memory regions content from the snapshot files into the new memory regions, the VMM will use the snapshot region files as the backing files behind each mapped region. This is done in order to reduce the time for the VM to be restored. When the source VM has been initially started with a backing file, this means it has been mapped with the MAP_SHARED flag. For this case, we cannot use the CoW trick to speed up the VM restore path and we simply fallback onto the copy of the memory regions content. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	d771223b2f	vmm: memory: Extend new() to support external backing files Whenever a MemoryManager is restored from a snapshot, the memory regions associated with it might need to directly back the mapped memory for increased performances. If that's the case, a list of external regions is provided and the MemoryManager should simply ignore what's coming from the MemoryConfig. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	ee5a041a0f	vmm: memory: Add Copy-on-Write parameter when creating region Now that we can choose specific mmap flags for the guest RAM, we create a new parameter "copy_on_write" meaning that the memory mappings backed by a file should be performed with MAP_PRIVATE instead of MAP_SHARED. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	be4e1e8712	vmm: memory: Use fine grained mmap wrapper In order to anticipate the need for special mmap flags when memory mapping the guest RAM, we need to switch from from_file() wrapper to build() wrapper. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	b9f9f01fcc	vmm: Extend seccomp filters to allow snapshot/restore A few KVM ioctls were missing in order to perform both snapshot and restore while keeping seccomp enabled. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	6eb721301c	vmm: Enable restore feature This connects the dots together, making the request from the user reach the actual implementation for restoring the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	53613319cc	vmm: Enable snapshot feature This connects the dots together, making the request from the user reach the actual implementation for snapshotting the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	2cd0bc0a2c	vmm: Create initial VM from its snapshot The MemoryManager is somehow a special case, as its restore() function was not implemented as part of the Snapshottable trait. Instead, and because restoring memory regions rely both on vm.json and every memory region snapshot file, the memory manager is restored at creation time. This makes the restore path slightly different from CpuManager, Vcpu, DeviceManager and Vm, but achieve the correct restoration of the MemoryManager along with its memory regions filled with the correct content. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	b55b83c6e8	vmm: vm: Implement the Transportable trait This is only implementing the send() function in order to store all Vm states into a file. This needs to be extended for live migration, by adding more transport methods, and also the recv() function must be implemented. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	1ed357cf34	vmm: vm: Implement the Snapshottable trait By aggregating snapshots from the CpuManager, the MemoryManager and the DeviceManager, Vm implements the snapshot() function from the Snapshottable trait. And by restoring snapshots from the CpuManager, the MemoryManager and the DeviceManager, Vm implements the restore() function from the Snapshottable trait. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	20ba271b6c	vmm: memory_manager: Implement the Transportable trait This implements the send() function of the Transportable trait, so that the guest memory regions can be saved into one file per region. This will need to be extended for live migration, as it will require other transport methods and the recv() function will need to be implemented too. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	e606112cef	vmm: memory_manager: Implement the Snapshottable trait In order to snapshot the content of the guest RAM, the MemoryManager must implement the Snapshottable trait. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	50b3f008d1	vmm: cpu: Implement the Snapshottable trait Implement the Snapshottable trait for Vcpu, and then implements it for CpuManager. Note that CpuManager goes through the Snapshottable implementation of Vcpu for every vCPU in order to implement the Snapshottable trait for itself. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	f787c409c4	vmm: cpu: Factorize vcpu starting code Anticipating the need for a slightly different function for restoring vCPUs, this patch factorizes most of the vCPU creation, so that it can be reused for migration purposes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Cathy Zhang	722f9b6628	vmm: cpu: Get and set KVM vCPU state These two new helpers will be useful to capture a vCPU state and being able to restore it at a later time. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Cathy Zhang	13756490b5	vmm: cpu: Track all Vcpus through CpuManager In anticipation for the CpuManager to aggregate all Vcpu snapshots together, this change makes sure the CpuManager has a handle onto every vCPU. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	a0d5dbce6c	vmm: device_manager: Implement the Snapshottable trait Based on the list of Migratable devices stored by the DeviceManager, the DeviceManager can implement the Snapshottable trait by aggregating all devices snapshots together. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	93d3abfd6e	vmm: device_manager: Make serial and ioapic devices migratable Serial and Ioapic both implement the Migratable trait, hence the DeviceManager can store them in the list of Migratable devices. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	12b036a824	Cargo: Update dependencies for the KVM serialization work We need the project to rely on kvm-bindings and kvm-ioctls branches which include the serde derive to be able to serialize and deserialize some KVM structures. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Rob Bradford	c7dfbd8a84	vmm: config: Implement fmt::Display for error Fixes: #367 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	d8119fda13	vmm: config: Remove unused error entries These entries are not currently used. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	1a10f16ad0	vmm: config: Consolidate size parsing code The parse_size helper function can now be consolidated into the ByteSized FromStr implementation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	f449486b9b	vmm: config: Make toggle parsing more tolerant Support "true" and "false" as well as well as capitalised forms. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	a4e0ce58c7	vmm: config: Consolidate on/off parsing Now all parsing code makes use of the Toggle and it's FromStr support move the helper function into the from_str() implementation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	c731a943d4	vmm: config: Port vsock to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	37264cf21b	vmm: config: Add unit testing for vsock Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	8665898ff3	vmm: config: Port device parsing to OptionParser Also make the "path" option required and generate an error if it is not provided. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	a85e2fa735	vmm: config: Add unit test for VFIO device parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	bed282b801	vmm: config: Add "valueless" options to OptionParser Valueless options are those like "off" or "tty" as used by the console options. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	2ae3392d32	vmm: config: Port console parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	143d63c88e	vmm: config: Add unit test for console parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	5ab58e743a	vmm: config: Port pmem option to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	233ad78b3a	vmm: config: Add parsing test for pmem Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	13dc637350	vmm: config: Port filesystem parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	7a071c28db	vmm: config: Implement unit testing for virtio-fs parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	e4cd3072d4	vmm: config: Port RNG options to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	708dbb973a	vmm: config: Add RNG parsing unit test Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	057e71d266	vmm: config: Accept empty value strings The integration tests and documentation make use of empty value strings like "--net tap=" accept them but return None so that the default value will be used as expected. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	218c780f67	vmm: config: Port network parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	8754720e2d	vmm: config: Add unit test for net parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	224e3ddef4	vmm: config: Switch disk parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e10244716	vmm: config: Add unit test for disk parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	e40ae6274b	vmm: config: Port memory option parsing to OptionParser This simplifies the parsing of the option by using OptionParser along with its automatic conversion behaviour. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	be32065aa4	vmm: config: Add "ByteSized" type for simplifying parsing of byte sizes Byte sizes are quantities ending in "K", "M", "G" and by implementing this type with a FromStr implementation the values can be converted using .parse(). Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	f01bd7d56d	vmm: config: Implement FromStr for HotplugMethod This allows the use of .parse() to automatically convert the string to the enum. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	746138039d	vmm: config: Add a Toggle type for "on/off" strings Some of the config parameters take an "on" or "off". Add a way to neatly parse that. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	929142bc2e	vmm: config: Add memory parsing unit test Before porting over to OptionParser add a unit test to validate the current memory parsing code. This showed up a bug where the "size=" was always required. Temporarily resolve this by assigning the string a default value which will later be replaced when the code is refactored. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	68203ea414	vmm: config: Port CPU parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e6a2825ba	vmm: config: Add unit test for CPU parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e7231cd69	vmm: config: Introduce basic OptionParser This will be used to simplify and consolidate much of the parsing code used for command line parameters. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Samuel Ortiz	447af8e702	vmm: vm: Factorize the device and cpu managers creation routine Into a new_from_memory_manager() routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	c73c9b112c	vmm: vm: Open kernel and initramfs once all managers are created Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	0646a90626	vmm: cpu: Pass CpusConfig to simplify the new() prototype Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	b584ec3fb3	vmm: memory_manager: Own the system allocator Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	ef2b11ee6c	vmm: memory_manager: Pass MemoryConfig to simplify the new() prototype Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	622f3f8fb6	vmm: vm: Avoid ioapic variable creation For a more readable VM creation routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	164e810069	vmm: cpu: Move CPUID patching to CpuManager Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	1a2c1f9751	vmm: vm: Factorize the KVM setup code Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	7a50646c02	vmm: device_manager: Convert migratable_devices to a map We must be able to map a migratable component id to its device. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	8f300bed83	vmm: api: Add a /api/v1/vm.restore endpoint Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	92c73c3b78	vmm: Add a VmRestore command Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	39d4f817f0	vmm: http: Add a /api/v1/vm.snapshot endpoint Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	cf8f8ce93a	vmm: api: Add a Snapshot command Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-02 13:24:25 +01:00
Sebastien Boeuf	452475c280	vmm: Add migration helpers Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	1b1a2175ca	vm-migration: Define the Snapshottable and Transportable traits A Snapshottable component can snapshot itself and provide a MigrationSnapshot payload as a result. A MigrationSnapshot payload is a map of component IDs to a list of migration sections (MigrationSection). As component can be made of several Migratable sub-components (e.g. the DeviceManager and its device objects), a migration snapshot can be made of multiple snapshot itself. A snapshot is a list of migration sections, each section being a component state snapshot. Having multiple sections allows for easier and backward compatible migration payload extensions. Once created, a migratable component snapshot may be transported and this is what the Transportable trait defines, through 2 methods: send and recv. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-02 13:24:25 +01:00
Sebastien Boeuf	2d17f4384a	vmm: seccomp: Add missing open() syscall On some systems, the open() system call is used by Cloud-Hypervisor, that's why it should be part of the seccomp filters whitelist. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-02 09:56:48 +02:00
Sebastien Boeuf	e4ea8b0bef	vmm: Add missing syscalls to the seccomp filters Both clock_gettime and gettimeofday syscalls where missing when running Cloud-Hypervisor on a Linux host without vDSO enabled. On a system with vDSO enabled, the syscalls performed by vDSO were not filtered, that's why we didn't have to whitelist them. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 16:50:52 +00:00
Sebastien Boeuf	9e18177654	vmm: Add memory hotplug support to VFIO PCI devices Extend the update_memory() method from DeviceManager so that VFIO PCI devices can update their DMA mappings to the physical IOMMU, after a memory hotplug has been performed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 09:35:39 +01:00
Sebastien Boeuf	cc67131ecc	vmm: Retrieve new memory region when memory is extended Whenever the memory is resized, it's important to retrieve the new region to pass it down to the device manager, this way it can decide what to do with it. Also, there's no need to use a boolean as we can instead use an Option to carry the information about the region. In case of virtio-mem, there will be no region since the whole memory has been reserved up front by the VMM at boot. This means only the ACPI hotplug will return a region and is the only method that requires the memory to be updated from the device manager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 09:35:39 +01:00
Samuel Ortiz	8fc7bf2953	vmm: Move to the latest linux-loader Commit 2adddce2 reorganized the crate for a cleaner multi architecture (x86_64 and aarch64) support. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-27 08:48:20 +01:00
Sebastien Boeuf	785812d976	vmm: Fallback to legacy boot if PVH is enabled along with initramfs For now, the codebase does not support booting from initramfs with PVH boot protocol, therefore we need to fallback to the legacy boot. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	6cce7b9560	arch: load initramfs and populate zero page * load the initramfs File into the guest memory, aligned to page size * finally setup the initramfs address and its size into the boot params (in configure_64bit_boot) Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	1f9bc68c54	openapi: Add initramfs support added InitramfsConfig property to the REST API spec Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	4db252b418	main, vmm: add --initramfs cli option currently unused, the initramfs argument is added to the cli, and stored in vmm::config:VmConfig as an Option(InitramfsConfig(PathBuf)) Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Rob Bradford	6244beb9d5	openapi: Add "vm.add-net" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	57c3fa4b1e	vmm: Add "add-net" to the API Add the HTTP and internal API entry points for adding a network device at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	f664cddec9	vmm: Add support for adding network devices to the VM The persistent memory will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	8f323e61d8	vmm: Add support to DeviceManager for hotplugging network devices Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	42a9896fe4	vmm: device_manager: Refactor make_virtio_net_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	9df601a1df	bin, vmm: Centralise the net syntax This will allow the syntax to be reused with cloud-hypervsor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Samuel Ortiz	41d7b3a387	vmm: memory_manager: Only send the GED notification for the ACPI method Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	15d9ec0149	openapit: Add hotplug_method to MemoryConfig Add hotplug_method to MemoryConfig in cloud-hypervisor.yaml. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	e63f98182a	vmm: device: Add make_virtio_mem_devices Add make_virtio_mem_devices to add virtio-mem to vmm. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	e6b934a56a	vmm: Add support for virtio-mem This commit adds new option hotplug_method to memory config. It can set the hotplug method to "acpi" or "virtio-mem". Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Rob Bradford	75878dd90a	openapi: Add "vm.add-pmem" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	f6f4c68fb4	vmm: Add "add-pmem" to the API Add the HTTP and internal API entry points for adding persistent memory at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	15de30f141	vmm: Add support for adding pmem devices to the VM The persistent memory will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	f7def621dd	vmm: Add support to DeviceManager for hotplugging pmem devices Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	8c3ea8cd76	vmm: device_manager: Refactor make_virtio_pmem_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	a7296bbb52	bin, vmm: Centralise the pmem syntax This will allow the syntax to be reused with cloud-hypervisor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	4c9d15d44c	vmm: Fix copy and paste error message vm_remove_device was copied from vm_add_device but the error message wasn't correctly updated. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	82cad99c0b	openapi: Add "vm.add-disk" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	f2151b2734	vmm: Add "add-disk" to the API Add the HTTP and internal API entry points for adding disks at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	164ec2b8e6	vmm: Add support for adding disks to the VM The disk will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	b3082c1984	vmm: Add support to DeviceManager for hotplugging disks Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	2be703ca92	vmm: device_manager: Refactor make_virtio_block_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	66da29d8dd	bin, vmm: Centralise the disk syntax This will allow the syntax to be reused with cloud-hypervsor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Sebastien Boeuf	e54f8ec8a5	vmm: Update memory through DeviceManager Whenever the VM memory is resized, DeviceManager needs to be notified so that it can subsequently notify each virtio devices about it. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 19:01:15 +00:00
Sebastien Boeuf	feb8d7ae90	vmm: Separate seccomp filters between VMM and API threads This separates the filters used between the VMM and API threads, so that we can apply different rules for each thread. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	f1a23d712f	vmm: api: Add seccomp to the HTTP API thread Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	db62cb3f4d	vmm: Add seccomp filter to the VMM thread This commit introduces the application of the seccomp filter to the VMM thread. The filter is empty for now (SeccompLevel::None). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	cb98d90097	vmm: Create new seccomp_filter module Based on the seccomp crate, we create a new vmm module responsible for creating a seccomp filter that will be applied to the VMM main thread. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	708f02dc26	vmm: Pull seccomp crate from Firecracker The seccomp crate from Firecracker is nicely implemented, documented and tested, which is a good reason for relying on it to create and apply seccomp filters. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Rob Bradford	8acc15a63c	build: Bump vm-memory and linux-loader dependencies linux-loader depends on vm-memory so must be updated at the same time. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-23 14:27:41 +00:00
Rob Bradford	f7197e8415	vmm: Add a "discard_writes=" to --pmem This opens the backing file read-only, makes the pages in the mmap() read-only and also makes the KVM mapping read-only. The file is also mapped with MAP_PRIVATE to make the changes local to this process only. This is functional alternative to having support for making a virtio-pmem device readonly. Unfortunately there is no concept of readonly virtio-pmem (or any type of NVDIMM/PMEM) in the Linux kernel so to be able to have a block device that is appears readonly in the guest requires significant specification and kernel changes. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Rob Bradford	d11a67b0fe	vmm: Use more generic MmapRegion constructor Switch to MmapRegion::build() and fill in the fields appropriately. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Rob Bradford	7257e890ef	vmm: Add "readonly" parameter MemoryManager::create_userspace_mapping Use this boolean to turn on the KVM_MEM_READONLY flag to indicate that this memory mapping should not be writable by the VM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Qiu Wenbo	c503118d16	vmm: fix a corrupted stack caused by get_win_size According to `asm-generic/termios.h`, the `struct winsize` should be: struct winsize { unsigned short ws_row; unsigned short ws_col; unsigned short ws_xpixel; unsigned short ws_ypixel; }; The ioctl of TIOCGWINSZ will trigger a segfault on aarch64. Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2020-03-20 07:30:06 +01:00
Rob Bradford	0788600702	build: Remove "pvh_boot" feature flag This feature is stable and there is no need for this to be behind a flag. This will also reduce the time needed to run the integration test as we will not be running them all again under the flag. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-19 13:05:44 +00:00
Rob Bradford	477bc17f18	bin: Share VFIO device syntax between cloud-hypervisor and ch-remote Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-18 23:38:55 +00:00
Jose Carlos Venegas Munoz	a31ffef085	openapi: Add hotplug_size for memory hotplug Add hotplug_size, needed to be defined when hotplug is used. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2020-03-18 19:06:07 +00:00
Rob Bradford	87990f9e67	vmm: Add virtio-pci device to B/D/F hash table This table currently contains only all the VFIO devices and it should really contain all the PCI devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-18 19:05:58 +00:00
Rob Bradford	fb185fa839	vmm: Always return PCI B/D/F from add_virtio_pci_device Previously this was only returned if the device had an IOMMU mapping and whether the device should be added to the virtio-iommu. This was already captured earlier as part of creating the device so use that information instead. Always returning the B/D/F is helpful as it facilitates virtio PCI device hotplug. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-18 19:05:58 +00:00
Samuel Ortiz	63eeed29cc	vm: Comment on the VM config update from memory hotplug I spent a few minutes trying to understand why we were unconditionally updating the VM config memory size, even if the guest memory resizing did not happen. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-18 12:48:40 +01:00
dependabot-preview[bot]	51f51ea17d	build(deps): bump libc from 0.2.67 to 0.2.68 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.67 to 0.2.68. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.67...0.2.68) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-03-17 21:36:38 +00:00
Rob Bradford	28a5f9dc19	vmm: acpi: Remove unused IORT related structures The IORT table for virtio-iommu use was removed and replaced with a purely virtio based solution. Although the table construction was removed these structures were left behind. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-17 12:46:26 +00:00
Alejandro Jimenez	9e247c4e06	pvh: Introduce "pvh_boot" feature Use a new feature called "pvh_boot" to enable using the PVH boot protocol if the guest kernel supports it. The feature can be enabled by building with: cargo build [--release] --features "pvh_boot" Once performance has been evaluated, this can be made part of the default set of features so that any guest that supports it boots using PVH as the preferred option as is the case in QEMU. Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-03-13 18:29:44 +01:00
Alejandro Jimenez	a22bc3559f	pvh: Write start_info structure to guest memory Fill the hvm_start_info and related memory map structures as specified in the PVH boot protocol. Write the data structures to guest memory at the GPA that will be stored in %rbx when the guest starts. Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-03-13 18:29:44 +01:00
Alejandro Jimenez	840a9a97ff	pvh: Initialize vCPU regs/sregs for PVH boot Set the initial values of the KVM vCPU registers as specified in the PVH boot ABI: https://xenbits.xen.org/docs/unstable/misc/pvh.html Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-03-13 18:29:44 +01:00
Alejandro Jimenez	24f0e42e6a	pvh: Introduce EntryPoint struct In order to properly initialize the kvm regs/sregs structs for the guest, the load_kernel() return type must specify which boot protocol to use with the entry point address it returns. Make load_kernel() return an EntryPoint struct containing the required information. This structure will later be used in the vCPU configuration methods to setup the appropriate initial conditions for the guest. Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-03-13 18:29:44 +01:00
Rob Bradford	4579afa091	vmm: For --disk error if socket and path is specified This is an error as the path should be specfied by the unmanaged backend. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-13 11:41:52 +00:00
Rob Bradford	7e599b4450	vmm: Make disk path optional When using "--disk" with a vhost socket and not using self spawning then it is not necessary or helpful to specify the path. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-13 11:41:52 +00:00
Sebastien Boeuf	8d785bbd5f	pci: Fix the PciBus using HashMap instead of Vec By using a Vec to hold the list of devices on the PciBus, there's a problem when we use unplug. Indeed, the vector of devices gets reduced and if the unplugged device was not the last one from the list, every other device after this one is shifted on the bus. To solve this problem, a HashMap is used. This allows to keep track of the exact place where each device stands on the bus. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-13 10:54:34 +01:00
Jose Carlos Venegas Munoz	40b38a4222	openapi: Make desired_ram int64 format The option desired_ram is in byte, make larger the amount of memory to add. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2020-03-12 23:17:56 +01:00
Sebastien Boeuf	efba48dddb	vmm: Don't put a VFIO device behind the vIOMMU by default With some of the factorization that happened to be able to support VFIO hotplug, one mistake was made. In case a vIOMMU is created through a virtio-iommu device, and no matter the "iommu" option value from the VFIO device parameter, the VFIO device was always placed behind the virtual IOMMU. This commit fixes this wrong behavior by making sure the device configuration is taken into account to decide if it should be attached or not to the virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 19:50:31 +01:00
Sebastien Boeuf	34412c9b41	vmm: Add id option to VFIO hotplug Add a new id option to the VFIO hotplug command so that it matches the VFIO coldplug semantic. This is done by refactoring the existing code for VFIO hotplug, where VmAddDeviceData structure is replaced by DeviceConfig. This structure is the one used whenever a VFIO device is coldplugged, which is why it makes sense to reuse it for the hotplug codepath. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 19:50:31 +01:00
Samuel Ortiz	18dc916380	vmm: Switch to the micro-http package It's been extracted from the Firecracker code base. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-11 17:38:01 +01:00
Sebastien Boeuf	9023444ad3	vmm: Add id field to --device through CLI Add the ability to specify the "id" associated with a device, by adding an extra option to the parameter --device. This new option is not mandatory, and by default, the VMM will take care of finding a unique identifier. If the identifier provided by the user through this new option is not unique, an error will be thrown and the VM won't be started. Fixes #881 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 13:10:57 +00:00
Sebastien Boeuf	f4a956a60a	vmm: Remove 32 bits MMIO range from correct address space The 32 bits MMIO address space is handled separately from the 64 bits one. For this reason, we need to invoke the appropriate freeing function to remove a range from this address space. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 13:10:30 +00:00
Sebastien Boeuf	432eb5b70a	vmm: Free PCI BARs when unplugging PCI device Now that PciDevice trait has a dedicated function to remove the bars, the DeviceManager can invoke this function whenever a PCI device is unplugged from the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 13:10:30 +00:00
Sebastien Boeuf	b50cbe5064	pci: Give PCI device ID back when removing a device Upon removal of a PCI device, make sure we don't hold onto the device ID as it could be reused for another device later. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	df71aaee3f	pci: Make the device ID allocation smarter In order to handle the case where devices are very often plugged and unplugged from a VM, we need to handle the PCI device ID allocation better. Any PCI device could be removed, which means we cannot simply rely on the vector size to give the next available PCI device ID. That's why this patch stores in memory the information about the 32 slots availability. Based on this information, whenever a new slot is needed, the code can correctly provide an available ID, or simply return an error because all slots are taken. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	e514b124ed	vmm: Update VmConfig when removing VFIO device This commit ensures that when a VFIO device is hot-unplugged from the VM, it is also removed from the VmConfig. This prevents a potential reboot from creating the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	81173bf4ab	vmm: Add id field to DeviceConfig structure Add a new field to the DeviceConfig, allowing the VMM to allocate a name to the VFIO devices. By identifying a VFIO device with a unique name, we can make sure a user can properly unplug it at any time. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	6cbdb9aa47	vmm: api: Introduce new "remove-device" HTTP endpoint This commit introduces the new command "remove-device" that will let a user hot-unplug a VFIO PCI device from an already running VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	991f3bb5da	vmm: Remove VFIO device from everywhere it is referenced This commit implements the eject function so that a VFIO device will be removed from any bus it might sit on, and from any list it might be stored in. The idea is to reach a point where there is no reference of the device anywhere in the code, so that the Drop implementation will be invoked and so that the device will be fully removed from the VMM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	6adebbc6a0	vmm: Detect when guest notifies about ejecting PCI device When the guest OS is done removing a PCI device, it will invoke the _EJ0 method from ACPI, associated with the device. This will trigger a port IO write to a region known by the VMM. Upon this writing, the VMM will trap the VM exit and retrieve the written value. Based on the value, the VMM will invoke its eject_device() method to finalize the removal of the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	08604ac6a8	vmm: Store PCI devices as Any devices from DeviceManager As we try to keep track of every PCI device related to the VM, we don't want to have separate lists depending on the concrete type associated with the PciDevice trait. Also, we want to be able to cast the actual type into any trait or concrete type. The most efficient way to solve all these issues is to store every device as an Arc<dyn Any + Send + Sync>. This gives the ability to downcast into the appropriate concrete type, and then to cast back into any trait that we might need. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	0f99d3f7cc	vmm: Store VFIO device's name and its PCI b/d/f Add a new list storing the device names across the entire codebase. VFIO devices are added to the list whenever a new one is created. By default, each VFIO device is given a name "vfioX" where X is the first available integer. Along with this new list of names, another list is created, grouping PCI device's name with its associated b/d/f. This will be useful to keep track of the created devices so that we can implement unplug functionality. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Rob Bradford	f0a3e7c4a1	build: Bump linux-loader and vm-memory dependencies linux-loader now uses the released vm-memory so we must move to that version at the same time. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-05 11:01:30 +01:00
Sebastien Boeuf	09829c44b2	vmm: Remove IO bus strong reference from Vm The Vm structure was used to store a strong reference to the IO bus. This is not needed anymore since the AddressManager is logically the one holding this strong reference. This has been made possible by the introduction of Weak references on the Bus structure itself. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	2dbb376175	vmm: Remove all Weak references from DeviceManager Now that the BusDevice devices are stored as Weak references by the IO and MMIO buses, there's no need to use Weak references from the DeviceManager anymore. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	9e915a0284	vmm: Remove all Weak references from CpuManager Now that the BusDevice devices are stored as Weak references by the IO and MMIO buses, there's no need to use Weak references from the CpuManager anymore. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	49268bff3b	pci: Remove all Weak references from PciBus Now that the BusDevice devices are stored as Weak references by the IO and MMIO buses, there's no need to use Weak references from the PciBus anymore. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	7773812f58	vmm: Store the list of BusDevice devices from DeviceManager The point is to make sure the DeviceManager holds a strong reference of each BusDevice inserted on the IO and MMIO buses. This will allow these buses to hold Weak references onto the BusDevice devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	d0820cc026	vmm: Make add_vfio_device mutable The method add_vfio_device() from the DeviceManager needs to be mutable if we want later to be able to update some internal fields from the DeviceManager from this same function. This commit simply takes care of making the necessary changes to change this function as mutable. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	948f808da6	vm: Rename DeviceManager field in Vm structure It's more logical to name the field referring to the DeviceManager as "device_manager" instead of "devices". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	d47f733e51	vmm: Break the cyclic dependency between DeviceManager and IO bus By inserting the DeviceManager on the IO bus, we introduced some cyclic dependency: DeviceManager ---> AddressManager ---> Bus ---> BusDevice ^ \| \| \| +---------------------------------------------+ This cycle needs to be broken by inserting a Weak reference instead of an Arc (considered as a strong reference). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	c1af13efeb	vmm: Update VmConfig when adding new device Ensures the configuration is updated after a new device has been hotplugged. In the event of a reboot, this means the new VM will be started with the new device that had been previously hotplugged. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	a86f4369a7	vmm: Add VFIO PCI device hotplug support This commit finalizes the VFIO PCI hotplug support, based on all the previous commits preparing for it. One thing to notice, this does not support vIOMMU yet. This means we can hotplug VFIO PCI devices, but we cannot attach them to an existing or a new virtio-iommu device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	320fea0eaf	vmm: Factorize VFIO PCI device creation This factorization is very important as it will allow both the standard codepath and the VFIO PCI hotplug codepath to rely on the same function to perform the addition of a new VFIO PCI device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	00716f90a0	vmm: Store virtio-iommu device from DeviceManager Helps with future refactoring of VFIO device creation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	5902dfa403	vmm: Store VFIO KVM device from DeviceManager Helps with future refactoring of VFIO device creation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	d9c1b4396e	vmm: Store MSI InterruptManager from DeviceManager Helps with future refactoring of VFIO device creation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	02adc4061a	vmm: Store PciBus from DeviceManager Helps with future refactoring of VFIO device creation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	d0218e94a3	vmm: Trigger hotplug notification to the guest Whenever the user wants to hotplug a new VFIO PCI device, the VMM will have to trigger a hotplug notification through the GED device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	0e58741a09	vmm: api: Introduce new "add-device" HTTP endpoint This commit introduces the new command "add-device" that will let a user hotplug a VFIO PCI device to an already running VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	0f1396acef	vmm: Insert PCI device hotplug operation region on IO bus Through the BusDevice implementation from the DeviceManager, and by inserting the DeviceManager on the IO bus for a specific IO port range, the VMM now has the ability to handle PCI device hotplug. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	65774e8a78	vmm: Implement BusDevice for DeviceManager In anticipation of inserting the DeviceManager on the IO/MMIO buses, the DeviceManager must implement the BusDevice trait. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	8dbc84318c	vmm: acpi: Add PCNT method to invoke DVNT Create a small method that will perform both hotplug of all the devices identified by PCIU bitmap, and then perform the hotunplug of all the devices identified by the PCID bitmap. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	c62db97a81	vmm: acpi: Add _EJ0 to each PCI device slot The _EJ0 method provides the guest OS a way to notify the VMM that the device has been properly ejected from the guest OS. Only after this point, the VMM can fully remove the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	4dc2a39f3a	vmm: acpi: Create PHPR container This new PHPR device in the DSDT table introduces some specific operation regions and the associated fields. PCIU stands for "PCI up", which identifies PCI devices that must be added. PCID stands for "PCI down", which identifies PCI devices that must be removed. B0EJ stands for "Bus 0 eject", which identifies which device on the bus has been ejected by the guest OS. Thanks to these fields, the VMM and the guest OS can communicate while performing hotplug/hotunplug operations. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	c3a0685e2d	vmm: acpi: Add notification method for PCI device slots Adds the DVNT method to the PCI0 device in the DSDT table. This new method is responsible for checking each slot and notify the guest OS if one of the slots is supposed to be added or removed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	5a68d5b6a7	vmm: acpi: Create PCI device slots This commit introduces the ACPI support for describing the 32 device slots attached to the main PCI host bridge. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Bin Liu	d6e6901957	vmm/api: Fix vm.info response definition Update cloud-hypervisor.yaml with latest code. Fixes: #841 Signed-off-by: liubin <liubin0329@gmail.com>	2020-03-03 09:34:25 +01:00
Sebastien Boeuf	8142c823ed	vmm: Move DeviceManager into an Arc<Mutex<>> In anticipation of the support for device hotplug, this commit moves the DeviceManager object into an Arc<Mutex<>> when the DeviceManager is being created. The reason is, we need the DeviceManager to implement the BusDevice trait and then provide it to the IO bus, so that IO accesses related to device hotplug can be handled correctly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-27 11:12:31 +01:00
Qiu Wenbo	9de3ace8c7	devices: implement Aml trait for GED device Fixes: #657 Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2020-02-25 08:32:16 +00:00
Sebastien Boeuf	b77fdeba2d	msi/msi-x: Prevent from losing masked interrupts We want to prevent from losing interrupts while they are masked. The way they can be lost is due to the internals of how they are connected through KVM. An eventfd is registered to a specific GSI, and then a route is associated with this same GSI. The current code adds/removes a route whenever a mask/unmask action happens. Problem with this approach, KVM will consume the eventfd but it won't be able to find an associated route and eventually it won't be able to deliver the interrupt. That's why this patch introduces a different way of masking/unmasking the interrupts, simply by registering/unregistering the eventfd with the GSI. This way, when the vector is masked, the eventfd is going to be written but nothing will happen because KVM won't consume the event. Whenever the unmask happens, the eventfd will be registered with a specific GSI, and if there's some pending events, KVM will trigger them, based on the route associated with the GSI. Suggested-by: Liu Jiang <gerry@linux.alibaba.com> Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-25 08:31:14 +00:00
Rob Bradford	bba5ef3a59	vmm: Remove deprecated CPU syntax Remove the old way of specifying the number of vCPUs to use. Fixes: #678 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-24 07:26:31 +01:00
Rob Bradford	374ac77c63	main, vmm: Remove deprecated --vhost-user-net This has been superseded by using --net with vhost_user=true and socket=<socket> Fixes: #678 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-24 07:26:31 +01:00
Rob Bradford	ffd816ebfa	main, vmm: Remove deprecated --vhost-user-blk This has been superseded by using --disk with vhost_user=true and socket=<socket> Fixes: #678 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-24 07:26:31 +01:00
dependabot-preview[bot]	f190cb05b5	build(deps): bump libc from 0.2.66 to 0.2.67 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.66 to 0.2.67. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.66...0.2.67) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-02-21 08:03:30 +00:00
Sergio Lopez	d2f1749edb	vmm: config: Add poll_queue property to DiskConfig Recently, vhost_user_block gained the ability of actively polling the queue, a feature that can be disabled with the poll_queue property. This change adds this property to DiskConfig, so it can be used through the "disk" argument. For the moment, it can only be used when vhost_user=true, but this will change once virtio-block gets the poll_queue feature too. Fixes: #787 Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-02-20 18:06:54 +01:00
Sergio Lopez	378dd81204	vmm: openapi: Add missing "direct" knob to DiskConfig Add missing "direct" knob that should be exposed through the REST API. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-02-20 18:06:54 +01:00
Sergio Lopez	056f5481ac	vmm: openapi: Fix "readonly" and "wce" defaults in DiskConfig Fix "readonly" and "wce" defaults in cloud-hypervisor.yaml to match their respective defaults in config.rs:DiskConfig. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-02-20 18:06:54 +01:00
Samuel Ortiz	c49e31a6d9	vmm: api: Return a resize error when resize fails And not a VmCreate one. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-20 12:26:12 +01:00
Samuel Ortiz	ebc6391bea	vmm: api: Fix resize command typos Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-20 12:26:12 +01:00
Samuel Ortiz	9de755334d	vmm: openapi: Update DiskConfig It's missing a few knobs (readonly, vhost, wce) that should be exposed through the rest API. Fixes: #790 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-20 12:17:50 +01:00
Rob Bradford	ed1e7817cc	vmm: Workaround double reboot triggered by the kernel The kernel does not adhere to the ACPI specification (probably to work around broken hardware) and rather than busy looping after requesting an ACPI reset it will attempt to reset by other mechanisms (such as i8042 reset.) In order to trigger a reset the devices write to an EventFd (called reset_evt.) This is used by the VMM to identify if a reset is requested and make the VM reboot. As the reset_evt is part of the VMM and reused for both the old and new VM it is possible for the newly booted VM to immediately get reset as there is an old event sitting in the EventFd. The simplest solution is to "drain" the reset_evt EventFd on reboot to make sure that there is no spurious events in the EventFd. Fixes: #783 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-19 18:51:14 +01:00
Sebastien Boeuf	793d4e7b8d	vmm: Move codebase to GuestMemoryAtomic from vm-memory Relying on the latest vm-memory version, including the freshly introduced structure GuestMemoryAtomic, this patch replaces every occurrence of Arc<ArcSwap<GuestMemoryMmap> with GuestMemoryAtomic<GuestMemoryMmap>. The point is to rely on the common RCU-like implementation from vm-memory so that we don't have to do it from Cloud-Hypervisor. Fixes #735 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-19 13:48:19 +00:00
Rob Bradford	1f6cbad01a	vmm: Add support for spawning vhost-user-block backend If no socket is supplied when enabling "vhost_user=true" on "--disk" follow the "exe" path in the /proc entry for this process and launch the network backend (via the vmm_path field.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-18 08:43:47 +00:00
Sebastien Boeuf	3edc2bd6ab	vmm: Prevent memory overcommitment through virtio-fs shared regions When a virtio-fs device is created with a dedicated shared region, by default the region should be mapped as PROT_NONE so that no pages can be faulted in. It's only when the guest performs the mount of the virtiofs filesystem that we can expect the VMM, on behalf of the backend, to perform some new mappings in the reserved shared window, using PROT_READ and/or PROT_WRITE. Fixes #763 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-17 15:03:47 +01:00
Rob Bradford	bc75c1b4e1	vmm: Add support for spawning vhost-user-net backend If no socket is supplied when enabling "vhost_user=true" on "--net" follow the "exe" path in the /proc entry for this process and launch the network backend (via the vmm_path field.) Currently this only supports creating a new tap interface as the network backend also only supports that. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-14 17:32:49 +00:00
Rob Bradford	b04eb4770b	vmm: Follow the "exe" symlink from the PID directory in /proc It is necessary to do this at the start of the VMM execution rather than later as it must be done in the main thread in order to satisfy the checks required by PTRACE_MODE_READ_FSCREDS (see proc(5) and ptrace(2)) The alternative is to run as CAP_SYS_PTRACE but that has its disadvantages. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-14 17:32:49 +00:00
Rob Bradford	7c9e8b103f	vmm: device_manager: Shutdown all virtio devices When the DeviceManager is dropped explicitly shutdown() all virtio devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-14 17:32:49 +00:00
Sebastien Boeuf	3447e226d9	dependencies: bump vm-memory from `4237db3` to `f3d1c27` This commit updates Cloud-Hypervisor to rely on the latest version of the vm-memory crate. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-06 11:40:45 +01:00
Sebastien Boeuf	62ccccc303	vmm: Make sure to retry creating the VM on EINTR If the ioctl syscall KVM_CREATE_VM gets interrupted while creating the VM, it is expected that we should retry since EINTR should not be considered a standard error. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-05 12:06:21 +01:00
Samuel Ortiz	da2b3c92d3	vm-device: interrupt: Remove InterruptType dependencies and definitions Having the InterruptManager trait depend on an InterruptType forces implementations into supporting potentially very different kind of interrupts from the same code base. What we're defining through the current, interrupt type based create_group() method is a need for having different interrupt managers for different kind of interrupts. By associating the InterruptManager trait to an interrupt group configuration type, we create a cleaner design to support that need as we're basically saying that one interrupt manager should have the single responsibility of supporting one kind of interrupt (defined through its configuration). Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-04 19:32:45 +01:00
Samuel Ortiz	84fc807bc6	interrupt: Interrupt manager split We create 2 different interrupt managers for separately handling creation of legacy and MSI interrupt groups. Doing so allows us to have a cleaner interrupt manager and IOAPIC initialization path. It also prepares for an InterruptManager trait design improvement where we remove the interrupt source type dependency by associating an interrupt configuration type to the trait. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-04 19:32:45 +01:00
Rob Bradford	880a57c920	vmm: Remove VmInfo struct After refactoring the VmInfo struct is no longer needed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	07bc292fa5	vmm: device_manager: Get VmFd from AddressManager A reference to the VmFd is stored on the AddressManager so it is not necessary to pass in the VmInfo into all methods that need it as it can be obtained from the AddressManager. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	6411c3ae42	vmm: device_manager: Use MemoryManager to get guest memory The DeviceManager has a reference to the MemoryManager so use that to get the GuestMemoryMmap rather than the version stored in the VmInfo struct. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	066fc6c0d1	vmm: device_manager: Get VM config from the struct member Remove the use of vm_info in methods to get the config and instead use the config stored on the DeviceManager itself. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	77ae3de4f3	vmm: device_manager: Make legacy device addition a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	599275b610	vmm: device_manager: Make ACPI device creation a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	b8c1b2e174	vmm: device_manager: Make console creation a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	b5440e2d0a	vmm: device_manager: Make virtio device creation functions methods Remove some in/out parameters and instead rely on them as members of the &mut self parameter. This prepares the way to more easily store state on the DeviceManager. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	e90c6f3c44	vmm: device_manager: Make make_virtio_devices a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. A follow-up commit will change the callee functions that create the devices themselves. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	dbc09ad0ef	vmm: device_manager: Make add_vfio_devices a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	d9e1c2cd22	vmm: device_manager: Make add_virtio_pci_device a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	aaa5e2e9ea	vmm: device_manager: Make add_virtio_mmio_device a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	2987476e0a	vmm: device_manager: Make add_pci_devices and add_mmio_devices methods Modify these functions to take an &mut self and become methods on DeviceManager. This allows the removal of some in/out parameters and leads the way to further refactoring and simplification. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	3dbae423bb	vmm: device_manager: Only add MemoryManager to I/O bus on ACPI builds The MemoryManager should only be included on the I/O bus when doing ACPI builds as that is the only time it will be interrogated. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	68fa97eb0e	vmm: device_manager: Always embed MemoryManager in the struct Currently the MemoryManager is only used on the ACPI code paths after the DeviceManager has been created. This will change in a future commit as part of the refactoring so for now always include it but name it with underscore prefix to indicate it might not always be used. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Sebastien Boeuf	ac01ceddbb	vmm: Cleanup list of PCI IDs related to virtual IOMMU Now that devices attached to the virtual IOMMU are described through virtio configuration, there is no need for the DeviceManager to store the list of IDs for all these devices. Instead, things are handled locally when PCI devices are being added. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-30 10:37:40 +01:00
Sebastien Boeuf	097cff2d85	vmm: Use virtio topology for virtio-iommu Instead of relying on the ACPI tables to describe the devices attached to the virtual IOMMU, let's use the virtio topology, as the ACPI support is getting deprecated. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-30 10:37:40 +01:00
dependabot-preview[bot]	1651cc3953	build(deps): bump kvm-ioctls from 0.4.0 to 0.5.0 Bumps [kvm-ioctls](https://github.com/rust-vmm/kvm-ioctls) from 0.4.0 to 0.5.0. - [Release notes](https://github.com/rust-vmm/kvm-ioctls/releases) - [Changelog](https://github.com/rust-vmm/kvm-ioctls/blob/master/CHANGELOG.md) - [Commits](https://github.com/rust-vmm/kvm-ioctls/compare/v0.4.0...v0.5.0) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-01-29 10:22:51 +00:00
Rob Bradford	75e6762897	vmm: Give deprecation warning for "--vhost-user-blk" syntax This will be removed in a future release. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-29 08:06:37 +00:00
Rob Bradford	969b5ee4e8	vmm: config: Add warning about specifying "wce" without "vhost-user" Currently configuring WCE is only supported when using vhost-user. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-29 08:06:37 +00:00
Rob Bradford	aeeae661fc	vmm: Support vhost-user-block via "--disks" Add a socket and vhost_user parameter to this option so that the same configuration option can be used for both virtio-block and vhost-user-block. For now it is necessary to specify both vhost_user and socket parameters as auto activation is not yet implemented. The wce parameter for supporting "Write Cache Enabling" is also added to the disk configuration. The original command line parameter is still supported for now and will be removed in a future release. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-29 08:06:37 +00:00
Rob Bradford	2c6f528c23	vmm: Give deprecation warning for "--vhost-user-net" syntax This will be removed in a future release. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-28 12:39:26 +00:00
Rob Bradford	a831aa214c	vmm: Support vhost-user-net via "--net" Add a socket and vhost_user parameter to this option so that the same configuration option can be used for both virtio-net and vhost-user-net. For now it is necessary to specify both vhost_user and socket parameters as auto activation is not yet implemented. The original command line parameter is still supported for now. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-28 12:39:26 +00:00
Sebastien Boeuf	f5b53ae4be	vm-virtio: Implement multiqueue/multithread support for virtio-blk This commit improves the existing virtio-blk implementation, allowing for better I/O performance. The cost for the end user is to accept allocating more vCPUs to the virtual machine, so that multiple I/O threads can run in parallel. One thing to notice, the amount of vCPUs must be egal or superior to the amount of queues dedicated to the virtio-blk device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-28 09:26:53 +01:00
Sebastien Boeuf	08e47ebd4b	vmm: Add num_queues and queue_size parameters to virtio-blk The number of queues and the size of each queue were not configurable. In anticipation for adding multiqueue support, this commit introduces some new parameters to let the user decide about the number of queues and the queue size. Note that the default values for each of these parameters are identical to the default values used for vhost-user-blk, that is 1 for the number of queues and 128 for the queue size. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-28 09:26:53 +01:00
dependabot-preview[bot]	16af54e583	build(deps): bump signal-hook from 0.1.12 to 0.1.13 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.1.12 to 0.1.13. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/compare/v0.1.12...v0.1.13) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2020-01-26 13:29:47 +00:00
Sebastien Boeuf	0fa1e2c241	vmm: Handle mapping from devices regions through vm-memory Devices like virtio-pmem and virtio-fs require some dedicated memory region to be mapped. The memory mapping from the DeviceManager is being replaced by the usage of MmapRegion from the vm-memory crate. The unmap will happen automatically when the MmapRegion will be dropped, which should happen when the DeviceManager gets dropped. Fixes #240 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-24 17:56:49 +01:00
Sebastien Boeuf	148a9ed5ce	vmm: Fix map_err losing the inner error Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-24 12:42:09 +01:00
Sebastien Boeuf	06396593c9	net_util: Fix map_err losing the inner error Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-24 12:42:09 +01:00
Rob Bradford	a34893a402	Revert "vmm: Move MemoryManager from I/O ports to MMIO region" This reverts commit `03108fb88b`.	2020-01-24 12:08:31 +01:00
Rob Bradford	57ed006992	Revert "devices, vmm: Move GED device to MMIO region" This reverts commit `5e3c62dc6a`.	2020-01-24 12:08:31 +01:00
Rob Bradford	6120d0fb1b	Revert "vmm: Move CpuManager device to MMIO region" This reverts commit `980e03fa0a`.	2020-01-24 12:08:31 +01:00
Rob Bradford	980e03fa0a	vmm: Move CpuManager device to MMIO region Move the CpuManager device from the I/O bus to living in an MMIO region. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-23 16:04:58 +00:00
Rob Bradford	5e3c62dc6a	devices, vmm: Move GED device to MMIO region Move GED device reporting of required device type to scan into an MMIO region rather than an I/O port. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-23 16:04:58 +00:00
Rob Bradford	03108fb88b	vmm: Move MemoryManager from I/O ports to MMIO region Rather than have the MemoryManager device sit on the I/O bus allocate space for MMIO and add it to the MMIO bus. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-23 16:04:58 +00:00
Sebastien Boeuf	0042f1de75	ioapic: Rely fully on the InterruptSourceGroup to manage interrupts This commit relies on the interrupt manager and the resulting interrupt source group to abstract the knowledge about KVM and how interrupts are updated and delivered. This allows the entire "devices" crate to be freed from kvm_ioctls and kvm_bindings dependencies. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-23 11:20:08 +00:00
Sebastien Boeuf	2dca959084	ioapic: Create the InterruptSourceGroup from InterruptManager The interrupt manager is passed to the IOAPIC creation, and the IOAPIC now creates an InterruptSourceGroup for MSI interrupts based on it. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-23 11:20:08 +00:00
Sebastien Boeuf	52800a871a	vmm: Create an InterruptManager dedicated to IOAPIC By introducing a new InterruptManager dedicated to the IOAPIC, we don't have to solve the chicken and eggs problem about which of the InterruptManager or the Ioapic should be created first. It's also totally fine to have two interrupt manager instances as they both share the same list of GSI routes and the same allocator. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-23 11:20:08 +00:00
Qiu Wenbo	2034fc2d84	vmm: Fix LENGTH_OFFSET_HIGH of MemoryManager Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2020-01-22 12:33:38 +00:00
Sergio Lopez	925c862f98	vmm: device_manager: Add 'direct' support for virtio-blk vhost_user_blk already has it, so it's only fair to give it to virtio-blk too. Extend DiskConfig with a 'direct' property, honor it while opening the file backing the disk image, and pass it to vm_virtio::RawFile. Fixes #631 Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-01-21 13:39:45 +00:00
Sergio Lopez	fb79e75afc	vmm: device_manager: Add read-only support for virtio-blk vhost_user_blk already has it, so it's only fair to give it to virtio-blk too. Extend DiskConfig with a 'readonly' properly, and pass it to vm_virtio::Block. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-01-21 13:39:45 +00:00
Sebastien Boeuf	9ac06bf613	ci: Run clippy for each specific feature The build is run against "--all-features", "pci,acpi", "pci" and "mmio" separately. The clippy validation must be run against the same set of features in order to validate the code is correct. Because of these new checks, this commit includes multiple fixes related to the errors generated when manually running the checks. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 11:44:40 +01:00
Sebastien Boeuf	99f39291fd	pci: Simplify PciDevice trait There's no need for assign_irq() or assign_msix() functions from the PciDevice trait, as we can see it's never used anywhere in the codebase. That's why it's better to remove these methods from the trait, and slightly adapt the existing code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	a20b383be8	vmm: Always use a reference for InterruptManager Since the InterruptManager is never stored into any structure, it should be passed as a reference instead of being cloned. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	bb8cd9eb24	vmm: Use LegacyUserspaceInterruptGroup for acpi device This commit replaces the way legacy interrupts were handled with the brand new implementation of the legacy InterruptSourceGroup for KVM. Additionally, since it removes the last bit relying on the Interrupt trait, the trait and its implementation can be removed from the codebase. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	75e22ff34e	vmm: Use LegacyUserspaceInterruptGroup for serial device This commit replaces the way legacy interrupts were handled with the brand new implementation of the legacy InterruptSourceGroup for KVM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	8d7c4ea334	vmm: Use LegacyUserspaceInterruptGroup for mmio devices This commit replaces the way legacy interrupts were handled with the brand new implementation of the legacy InterruptSourceGroup for KVM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	12657ef59f	vmm: Fully implement LegacyUserspaceInterruptGroup Relying on the previous commits, the legacy interrupt implementation can be completed. The IOAPIC handler is used to deliver the interrupt that will be triggered through the trigger() method. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	f70c9937fb	vmm: Add ioapic to KvmInterruptManager By having a reference to the IOAPIC, the KvmInterruptManager is going to be able to initialize properly the legacy interrupt source group. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	c9ea235a0e	vmm: Add LegacyUserspaceInterruptGroup skeleton for legacy interrupts In order to be able to use the InterruptManager abstraction with virtio-mmio devices, this commit introduces InterruptSourceGroup's skeleton for legacy interrupts. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	2aabf58bf5	vmm: Move irq_routes creation to specific MSI use case When KvmInterruptManager initializes a new InterruptSourceGroup, it's only for PCI_MSI_IRQ case that it needs to allocate the GSI and create a new InterruptRoute. That's why this commit moves the general code into the specific use case. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	d34f31fe7b	vmm: Fix KvmInterruptManager when base is different from 0 When the base InterruptIndex is different from 0, the loop allocating GSI and HashMap entries won't work as expected. The for loop needs to start from base, but the limit must be base+count so that we allocate a number of "count" entries. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	e73cb1ff80	vmm: Initialize InterruptManager sooner In order to let the InterruptManager be shared across both PCI and MMIO devices, this commit moves the initialization earlier in the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Rob Bradford	3901a1dd7d	vmm: Log an error if VM resize fails As well as returing an error to the API caller. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-17 23:44:21 +01:00
Rob Bradford	76d9bf2792	vmm: Start memory slots at zero After refactoring a common function is used to setup these slots and that function takes care of allocating a new slot so it is not necessary to reserve the initial region slots. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-17 23:44:21 +01:00
Rob Bradford	0ab22fea2c	vmm: Only generate GED event when new DIMM added Avoid the ACPI scan in the guest OS when no new DIMM is hotplugged. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-17 23:44:21 +01:00
Rob Bradford	211786ab42	vmm: Only generate GED interrupt when the number of vCPUs has changed Avoid activity in the the guest OS if the number of vCPUs has not changed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-17 23:44:21 +01:00
Sebastien Boeuf	4bb12a2d8d	interrupt: Reorganize all interrupt management with InterruptManager Based on all the previous changes, we can at this point replace the entire interrupt management with the implementation of InterruptManager and InterruptSourceGroup traits. By using KvmInterruptManager from the DeviceManager, we can provide both VirtioPciDevice and VfioPciDevice a way to pick the kind of InterruptSourceGroup they want to create. Because they choose the type of interrupt to be MSI/MSI-X, they will be given a MsiInterruptGroup. Both MsixConfig and MsiConfig are responsible for the update of the GSI routes, which is why, by passing the MsiInterruptGroup to them, they can still perform the GSI route management without knowing implementation details. That's where the InterruptSourceGroup is powerful, as it provides a generic way to manage interrupt, no matter the type of interrupt and no matter which hypervisor might be in use. Once the full replacement has been achieved, both SystemAllocator and KVM specific dependencies can be removed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	92082ad439	vmm: Fully implement interrupt traits After the skeleton of InterruptManager and InterruptSourceGroup traits have been implemented, this new commit takes care of fully implementing the content of KvmInterruptManager (InterruptManager trait) and MsiInterruptGroup (InterruptSourceGroup). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	0f727127d5	vmm: Implement InterruptSourceGroup and InterruptManager skeleton This commit introduces an empty implementation of both InterruptManager and InterruptSourceGroup traits, as a proper basis for further implementation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	c396baca46	vm-virtio: Modify VirtioInterrupt callback into a trait Callbacks are not the most Rust idiomatic way of programming. The right way is to use a Trait to provide multiple implementation of the same interface. Additionally, a Trait will allow for multiple functions to be defined while using callbacks means that a new callback must be introduced for each new function we want to add. For these two reasons, the current commit modifies the existing VirtioInterrupt callback into a Trait of the same name. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	2381f32ae0	msix: Add gsi_msi_routes to MsixConfig Because MsixConfig will be responsible for updating KVM GSI routes at some point, it is necessary that it can access the list of routes contained by gsi_msi_routes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	9b60fcdc39	msix: Add VmFd to MsixConfig Because MsixConfig will be responsible for updating the KVM GSI routes at some point, it must have access to the VmFd to invoke the KVM ioctl KVM_SET_GSI_ROUTING. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	86c760a0d9	msix: Add SystemAllocator to MsixConfig The point here is to let MsixConfig take care of the GSI allocation, which means the SystemAllocator must be passed from the vmm crate all the way down to the pci crate. Once this is done, the GSI allocation and irq_fd creation is performed by MsixConfig directly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	f5704d32b3	vmm: Move gsi_msi_routes creation to be shared across all PCI devices Because we will need to share the same list of GSI routes across multiple PCI devices (virtio-pci, VFIO), this commit moves the creation of such list to a higher level location in the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sergio Lopez	a14aee9213	qcow: Use RawFile as backend instead of File Use RawFile as backend instead of File. This allows us to abstract the access to the actual image with a specialized layer, so we have a place where we can deal with the low-level peculiarities. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-01-17 17:28:44 +00:00
Sergio Lopez	c5a656c9dc	vm-virtio: block: Add support for alignment restrictions Doing I/O on an image opened with O_DIRECT requires to adhere to certain restrictions, requiring the following elements to be aligned: - Address of the source/destination memory buffer. - File offset. - Length of the data to be read/written. The actual alignment value depends on various elements, and according to open(2) "(...) there is currently no filesystem-independent interface for an application to discover these restrictions (...)". To discover such value, we iterate through a list of alignments (currently, 512 and 4096) calling pread() with each one and checking if the operation succeeded. We also extend RawFile so it can be used as a backend for QcowFile, so the later can be easily adapted to support O_DIRECT too. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-01-17 17:28:44 +00:00
Cathy Zhang	652e7b9b8a	vm-virtio: Implement multiple queue support for net devices Update the common part in net_util.rs under vm-virtio to add mq support, meanwhile enable mq for virtio-net device, vhost-user-net device and vhost-user-net backend. Multiple threads will be created, one thread will be responsible to handle one queue pair separately. To gain the better performance, it requires to have the same amount of vcpus as queue pair numbers defined for the net device, due to the cpu affinity. Multiple thread support is not added for vhost-user-net backend currently, it will be added in future. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2020-01-17 12:06:19 +01:00
Cathy Zhang	404316eea1	vmm: Add multiple queue option and update config for virtio-net device Add num_queues and queue_size for virtio-net device to make them configurable, while add the associated options in command line. Update cloud-hypervisor.yaml with the new options for NetConfig. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2020-01-17 12:06:19 +01:00
Cathy Zhang	4ab88a8173	net_util: Add multiple queue support for tap Add support to allow VMMs to open the same tap device many times, it will create multiple file descriptors meanwhile. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2020-01-17 12:06:19 +01:00
Cathy Zhang	1ae7deb393	vm-virtio: Implement refactor for net devices and backend Since the common parts are put into net_util.rs under vm-virtio, refactoring code for virtio-net device, vhost-user-net device and backend to shrink the code size and improve readability meanwhile. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2020-01-17 12:06:19 +01:00
Rob Bradford	8b500d7873	deps: Bump vm-memory and linux-loader version The function GuestMemory::end_addr() has been renamed to last_addr() Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	7310ab6fa7	devices, vmm: Use a bit field for ACPI GED interrupt type Use independent bits for storing whether there is a CPU or memory device changed when reporting changes via ACPI GED interrupt. This prevents a later notification squashing an earlier one and ensure that hotplugging both CPU and memory at the same time succeeds. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	28c6652e57	vmm: Upon VmResize attempt to hotplug the memory If a new amount of RAM is requested in the VmResize command try and hotplug if it an increase (MemoryManager::Resize() silently ignores decreases.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	4e414f0d84	vmm: device_manager: Scan memory devices upon GED interrupt If there is a GED interrupt and the field indicates that the memory device has changed triggers a scan of the memory devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	284d5e011a	vmm: Add memory hotplug ACPI entries to DSDT Generate and expose the DSDT table entries required to support memory hotplug. The AML methods call into the MemoryManager via I/O ports exposed as fields. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	8ecf736982	vmm: device_manager: Add the MemoryManager to the I/O bus Now that the MemoryManager has I/O port functionality it needs to be exposed on the I/O bus. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	1218765df2	vmm: memory_manager: Expose the slots details via an I/O port Expose the details of hotplug RAM slots via an I/O port. This will be consumed by the ACPI DSDT tables to report the hotplug memory details to the guest. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	9880a2aba9	vmm: memory_manger: Add support for adding new memory to the VM Add a "resize()" method on MemoryManager which will create a new memory allocation based on the difference between the desired RAM amount and the amount already in use. After allocating the added RAM using the same backing method as the boot RAM store the details in a vector and update the KVM map and create a new GuestMemoryMmap and replace all the users. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	82fce5a4e2	vmm: Add support for resizing the memory used by the VM For now the new memory size is only used after a reboot but support for hotplugging memory will be added in a later commit. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	78dcb1862c	vmm: device_manager: Store the type of notification in a local value When the value is read from the I/O port via the ACPI AML functions to determine what has been triggered the notifiction value is reset preventing a second read from exposing the value. If we need support multiple types of GED notification (such as memory hotplug) then we should avoid reading the value multiple times. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	f5137e84bb	vmm, main: Add optional "hotplug_size" to --mem This specifies how much address space should be reserved for hotplugging of RAM. This space is reserved by adding move the start of the device area by the desired amount. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	f1b6657833	vmm: Make desired vCPUs optional in resize command In order to be able to support resizing either vCPUs or memory or both make the fields in the resize command optional. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	72b9e920a3	vmm: memory_manager: Further refactor memory region allocation This allows the memory regions to be allocated later which is necessary for hotplug memory. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	1af11a7c92	vmm: memory_manager: Refactor GuestMemoryMmap construction Make the GuestMemoryMmap from a Vec<Arc<GuestRegionMmap>> by using this method we can persist a set of regions in the MemoryManager and then extend this set with a newly created region. Ultimately that will allow the hotplug of memory. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Samuel Ortiz	5788d36583	vmm: Do not create virtio devices when missing a transport If neither PCI or MMIO are built in, we should not bother creating any virtio devices at all. When building a minimal VMM made of a kernel with an initramfs and a serial console, the RNG virtio device is still created even though there is no way it can ever get probed. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-01-14 07:42:09 +01:00
Sebastien Boeuf	ae6f27277b	acpi: Introduce VIOT to support latest virtio-iommu implementation Because virtio-iommu is still evolving (as it's only partly upstream), some pieces like the ACPI declaration of the different nodes and devices attached to the virtual IOMMU are changing. This patch introduces a new ACPI table called VIOT, standing as the high level table overseeing the IORT table and associated subtables. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-08 09:27:07 +01:00
Rob Bradford	b2589d4f3f	vm-virtio, vmm, vfio: Store GuestMemoryMmap in an Arc<ArcSwap<T>> This allows us to change the memory map that is being used by the devices via an atomic swap (by replacing the map with another one). The ArcSwap provides the mechanism for atomically swapping from to another whilst still giving good read performace. It is inside an Arc so that we can use a single ArcSwap for all users. Not covered by this change is replacing the GuestMemoryMmap itself. This change also removes some vertical whitespace from use blocks in the files that this commit also changed. Vertical whitespace was being used inconsistently and broke rustfmt's behaviour of ordering the imports as it would only do it within the block. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-02 13:20:11 +00:00
Rob Bradford	a551398135	vmm: device_manager: Use MemoryManager to create KVM mapping Use the newly exported funtionality to reduce the amount of duplicated code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	962dec2913	vmm: memory_manager: Refactor KVM userspace mapping creation This function will be useful for other parts of the VMM that also estabilish their own mappings. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	7df88793a0	vmm: device_manager: Get device range from MemoryManager This removes the duplication of these values. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	61cfe3e72d	vmm: Obtain sequential KVM memory slot numbers from MemoryManager This removes the need to handle a mutable integer and also centralises the allocation of these slot numbers. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	260cebb8cf	vmm: Introduce MemoryManager The memory manager is responsible for setting up the guest memory and in the long term will also handle addition of guest memory. In this commit move code for creating the backing memory and populating the allocator into the new implementation trying to make as minimal changes to other code as possible. Follow on commits will further reduce some of the duplicated code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	d5682cd306	vmm: device_manager: Rewrite if chain using match To reflect updated clippy rules: error: `if` chain can be rewritten with `match` --> vmm/src/device_manager.rs:1508:25 \| 1508 \| / if ret > 0 { 1509 \| \| debug!("MSI message successfully delivered"); 1510 \| \| } else if ret == 0 { 1511 \| \| warn!("failed to deliver MSI message, blocked by guest"); 1512 \| \| } \| \|_________________________^ \| = note: `-D clippy::comparison-chain` implied by `-D warnings` = help: Consider rewriting the `if` chain to use `cmp` and `match`. = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#comparison_chain Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	21b88c3ea0	vmm: cpu: Rewrite if chain using match Address updated clippy error: error: `if` chain can be rewritten with `match` --> vmm/src/cpu.rs:668:9 \| 668 \| / if desired_vcpus > self.present_vcpus() { 669 \| \| self.activate_vcpus(desired_vcpus, None)?; 670 \| \| } else if desired_vcpus < self.present_vcpus() { 671 \| \| self.mark_vcpus_for_removal(desired_vcpus)?; 672 \| \| } \| \|_________^ \| = note: `-D clippy::comparison-chain` implied by `-D warnings` = help: Consider rewriting the `if` chain to use `cmp` and `match`. = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#comparison_chain Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	e25a47b32c	vmm: device_manager: Remove redundant clones Address updated clippy errors: error: redundant clone --> vmm/src/device_manager.rs:699:32 \| 699 \| .insert(acpi_device.clone(), 0x3c0, 0x4) \| ^^^^^^^^ help: remove this \| = note: `-D clippy::redundant-clone` implied by `-D warnings` note: this value is dropped without further use --> vmm/src/device_manager.rs:699:21 \| 699 \| .insert(acpi_device.clone(), 0x3c0, 0x4) \| ^^^^^^^^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone error: redundant clone --> vmm/src/device_manager.rs:737:26 \| 737 \| .insert(i8042.clone(), 0x61, 0x4) \| ^^^^^^^^ help: remove this \| note: this value is dropped without further use --> vmm/src/device_manager.rs:737:21 \| 737 \| .insert(i8042.clone(), 0x61, 0x4) \| ^^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone error: redundant clone --> vmm/src/device_manager.rs:754:29 \| 754 \| .insert(cmos.clone(), 0x70, 0x2) \| ^^^^^^^^ help: remove this \| note: this value is dropped without further use --> vmm/src/device_manager.rs:754:25 \| 754 \| .insert(cmos.clone(), 0x70, 0x2) \| ^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	a6878accd5	vmm: cpu: Implement CPU removal When the running OS has been told that a CPU should be removed it will shutdown the CPU and then signal to the hypervisor via the "_EJ0" method on the device that ultimately writes into an I/O port than the vCPU should be shutdown. Upon notification the hypervisor signals to the individual thread that it should shutdown and waits for that thread to end. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	7b3fc72aea	vmm: cpu: Notify guest OS that it should offline vCPUs Allow the resizing of the number of vCPUs to less than the current active vCPUs. This does not currently remove them from the system but the kernel will take them offline. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	7e81b0ded7	vmm: cpu: Create vCPU state for all possible vCPUs This will make it more straightforward when we attempt to remove vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	156ea392a2	vmm: cpu: Only do ACPI notify on newly added vCPUs When we add a vCPU set an "inserting" boolean that is exposed as an ACPI field that will be checked for and reset when the ACPI GED notification for CPU devices happens. This change is a precursor for CPU unplug. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-16 23:57:14 +01:00
Rob Bradford	e8313e3e69	vmm: acpi: Refactor ACPI CPU notification Continue to notify on all vCPUs but instead separate the notification functionality into two methods, CSCN that walks through all the CPUs and CTFY which notifies based on the numerical CPU id. This is an interim step towards only notifying on changed CPUs and ultimately CPU removal. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-16 23:57:14 +01:00
Sebastien Boeuf	d1390906c8	vmm: config: Derive Debug and PartialEq for configuration structures In anticipation for the writing of unit tests comparing two VmConfig structures, this commit derives the PartialEq trait for VmConfig and all embedded structures. This patch also derives the Debug trait for the same set of structures so that we can print them to facilitate debugging. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Sebastien Boeuf	93f5f6ed45	vmm: config: Provide a default empty command line through OpenAPI The OpenAPI should not have to provide a command line since the CLI considers the command line as an empty string if nothing is provided. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Sebastien Boeuf	43bd0e53c4	main: Move VmParams creation into a dedicated function This brings more modularity to the code, which will be helpful when we will later test the CLI and OpenAPI generate the same VmConfig output. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Samuel Ortiz	f0b7412495	vmm: device_manager: Add all virtio devices to the migratable list We want to track all migratable devices through the DeviceManager. Fixes: #341 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	37557c8b35	vmm: vm: Implement the Pausable trait Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	9756fc2dd0	vmm: cpu_manager: Implement the Pausable trait Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	35dd1523c9	vmm: device_manager: Implement the Pausable trait Since the Snapshotable placeholder and Migratable traits are provided as well, the DeviceManager object and all its objects are now Migratable. All Migratable devices are tracked as Arc<Mutex<dyn Migratable>> references. Keeping track of all migratable devices allows for implementing the Migratable trait for the DeviceManager structure, making the whole device model potentially migratable. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	35d7721683	vmm: Convert virtio devices to Arc<Mutex<T>> Migratable devices can be virtio or legacy devices. In any case, they can potentially be tracked through one of the IO bus as an Arc<Mutex<dyn BusDevice>>. In order for the DeviceManager to also keep track of such devices as Migratable trait objects, they must be shared as mutable atomic references, i.e. Arc<Mutex<T>>. That forces all Migratable objects to be tracked as Arc<Mutex<dyn Migratable>>. Virtio devices are typically migratable, and thus for them to be referenced by the DeviceManager, they now should be built as Arc<Mutex<VirtioDevice>>. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Sebastien Boeuf	64c5e3d8cb	vmm: api: Adjust FsConfig for OpenAPI The FsConfig structure has been recently adjusted so that the default value matches between OpenAPI and CLI. Unfortunately, with the current description, there is no way from the OpenAPI to describe a cache_size value "None", so that DAX does not get enabled. Usually, using a Rust "Option" works because the default value is None. But in this case, the default value is Some(8G), which means we cannot describe a None. This commit tackles the problem, introducing an explicit parameter "dax", and leaving "cache_size" as a simple u64 integer. This way, the default value is dax=true and cache_size=8G, but it lets the opportunity to disable DAX entirely with dax=false, which will simply ignore the cache_size value. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	4bfd51cc42	vmm: api: Match VhostUserBlkConfig defaults between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the VhostUserBlkConfig structure, this patch defines some default values for num_queues, queue_size and wce. num_queues is 1, queue_size is 128 and wce is true. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	1c2587f8cb	vmm: api: Match VhostUserNetConfig defaults between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the VhostUserNetConfig structure, this patch defines some default values for num_queues, queue_size and mac. num_queues is 2 since that's a pair of TX/RX queues, queue_size is 256 and mac is a randomly generated value. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	5e0bbf9c3b	vmm: Don't factorize vhost-user configurations We want to set different default configurations for vhost-user-net and vhost-user-blk, which is the reason why the common part corresponding to the number of queues and the queue size cannot be embedded. This prepares for the following commit, matching API and CLI behaviors. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	793327cff8	vmm: api: Make ConsoleConfig default match between CLI and HTTP API A simple patch making sure the field "file" is provisioned with the same default value through CLI and OpenAPI. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	cc08c44cb9	vmm: api: Make MemoryConfig default match between CLI and HTTP API Just making sure we have a serde default for the field "file" since it is not a required field in the OpenAPI definition. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	5a72225856	vmm: api: Update CpuConfig name to match the internal name All structures match between the OpenAPI definition and the internal configuration code, that's why CpuConfig is being renamed into CpusConfig. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Rob Bradford	c61104df47	vmm: Port to latest vmm-sys-util The signal handling for vCPU signals has changed in the latest release so switch to the new API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-11 14:11:11 +00:00
Sebastien Boeuf	ee528ae808	vmm: api: Make FsConfig defaults match between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the FsConfig structure, this patch defines some default values for num_queues, queue_size and the cache_size. num_queues is set to 1, queue_size is set to 1024, and cache_size is set to Some(8G) which means that DAX is enabled by default with a shared region of 8GiB. Fixes #508 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-09 23:42:23 -08:00
Sebastien Boeuf	befd342da4	vmm: api: Make NetConfig defaults match between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the NetConfig structure, this patch defines some default values for tap, ip, mask, mac and iommu. tap is None, ip is 192.168.249.1, mask is 255.255.255.0, mac is a randomly generated value, and iommu is false. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-09 23:19:24 -08:00
Jose Carlos Venegas Munoz	99e608c240	openapi: Fix schema Fix openapi schema to be a valid yaml. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-12-09 14:30:15 -08:00
Rob Bradford	f994665610	vmm: Reduce the minimum IRQ constant Now that the GED device does not use a hardcoded IRQ number the starting IRQ number can be restored (needed for the hardcoded serial port IRQ.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-09 16:58:00 +00:00
Rob Bradford	ba59c62044	vmm, devices: Remove hardcoded IRQ number for GED device Remove the previously hardcoded IRQ number used for the GED device. Instead allocate the IRQ using the allocator and use that value in the definition in the ACPI device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-09 16:58:00 +00:00
Sebastien Boeuf	aa94e9b8f3	Revert "vmm: api: Modify FsConfig to be OpenAPI friendly" This reverts commit `defc5dcd9c`.	2019-12-06 18:08:10 +00:00
Rob Bradford	9b1ba14f2d	vmm: Delegate device related ACPI DSDT table work to DeviceManager Move the code for handling the creation of the DSDT entries for devices into the DeviceManager. This will make it easier to handle device hotplug and also in the future remove some hardcoded ACPI constants. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 17:44:00 +00:00
Rob Bradford	60e6609011	vmm: Delegate CPU related ACPI tables to CpuManager Move the code for generating the MADT (APIC) table and the DSDT generation for CPU related functionality into the CpuManager. There is no functional change just code rearrangement. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 17:44:00 +00:00
Sebastien Boeuf	defc5dcd9c	vmm: api: Modify FsConfig to be OpenAPI friendly When consumer of the HTTP API try to interact with cloud-hypervisor, they have to provide the equivalent of the config structure related to each component they need. Problem is, the Rust enum type "Option" cannot be obtained from the OpenAPI YAML definition. This patch intends to fix this inconsistency between what is possible through the CLI and what's possible through the HTTP API by using simple types bool and int64 instead of Option<u64>. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-06 06:38:48 -08:00
Rob Bradford	59d01712ad	vmm: Remove kernel based IOAPIC handling from the device manager Previously the device setup code assumed that if no IOAPIC was passed in then the device should be added to the kernel irqchip. As an earlier change meant that there was always a userspace IOAPIC this kernel based code can be removed. The accessor still returns an Option type to leave scope for implementing a situation without an IOAPIC (no serial or GED device). This change does not add support no-IOAPIC mode as the original code did not either. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00

... 13 14 15 16 17 ...

1752 Commits