cloud-hypervisor

mirror of https://github.com/cloud-hypervisor/cloud-hypervisor.git synced 2024-12-29 17:15:19 +00:00

Author	SHA1	Message	Date
Sebastien Boeuf	aaba6e777f	vmm: Add virtio-console to the list of Migratable devices The virtio-console was not added to the list of Migratable devices, which is fixed from this patch. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9ab4bb1ae2	devices: serial: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	06487131f9	vm-virtio: pci: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. It is based off the name from the virtio device attached to this transport layer. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	eeb7e10d1f	vm-virtio: mmio: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. It is based off the name from the virtio device attached to this transport layer. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9d84ef5073	vmm: Make the virtio identifier mandatory Because we know we will need every virtio device to be identified with a unique id, we can simplify the code by making the identifier mandatory. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	14350f5de4	devices: ioapic: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	556871570e	vm-virtio: iommu: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	052eff1ca7	vm-virtio: console: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	354c2a4b3d	vm-virtio: vhost-user-net: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	46e0b3ff75	vm-virtio: vhost-user-blk: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	bb7fa71fcb	vm-virtio: vhost-user-fs: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	ec5ff395cf	vm-virtio: vsock: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9b53044aae	vm-virtio: mem: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	1592a9292f	vm-virtio: pmem: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	2e91b73881	vm-virtio: rng: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9eb7413fab	vm-virtio: net: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	be946caf4b	vm-virtio: blk: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	ff9c8b847f	vmm: Always generate the next device name Even in the context of "mmio" feature, we need the next device name to be generated as we need to identify virtio-mmio devices to support snapshot and restore functionalities. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	8183141399	vmm: Add an identifier to the ioapic device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	e4386c8bb7	vmm: Add an identifier to the virtio-iommu device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	75ddd2a244	vmm: Add an identifier to the --console device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	eac350c454	vmm: Add an identifier to the virtio-mem device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	6802ef5406	vmm: Add an identifier to the --rng device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	d71d52e9b0	vmm: Fix virtio-console creation with virtual IOMMU If the virtio-console device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	b08fde5928	vmm: Fix virtio-rng creation with virtual IOMMU If the virtio-rng device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	8031ac33c3	vmm: Fix virtio-vsock creation with virtual IOMMU If the virtio-vsock device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Rob Bradford	8cef35745b	vmm: seccomp: Add fork, gettid and pipe2 syscalls to permitted list This is needed for self spawning with the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Rob Bradford	ce7678f29f	vmm: seccomp: Add tkill syscall to permitted list This is needed for rebooting on the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Rob Bradford	12758d7fad	vmm: seccomp: Add epoll_pwait syscall to permitted list This is needed for basic operation on the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Samuel Ortiz	86fcd19b8a	build: Initial musl support Fix all build failures and add musl to the gihub workflows. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-29 17:57:01 +01:00
Sebastien Boeuf	a5de49558e	vmm: Only allow removal of specific types of virtio device Now that all virtio devices are assigned with identifiers, they could all be removed from the VM. This is not something that we want to allow because it does not make sense for some devices. That's why based on the device type, we remove the device or we return an error to the user. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 13:33:19 +01:00
Sebastien Boeuf	9ed880d74e	vmm: Add an identifier to the --fs device By giving the devices ids this effectively enables the removal of the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 13:33:19 +01:00
Sebastien Boeuf	7e0ab6b56d	vmm: Fix pmem device creation The parameters regarding the attachment to the virtio-iommu device was not propagated correclty, and any modification to the configuration was not stored back into it. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 13:33:19 +01:00
Rob Bradford	8de7448d44	vmm: api: Add "add-vsock" API entry point This allows the hotplugging of vsock devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	bf09a1e695	openapi: Add "id" field to VsockConfig Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	a76cf0865f	vmm: vm: Remove vsock device from config When doing device unplug remove the vsock device from the configuration if present. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	99422324a7	vmm: vm: Add "add_vsock()" Add the vsock device to the device manager and patch the config to add the new vsock device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	1d61c476a1	vmm: device_manager: Add support for hotplugging virtio-vsock devices Create a new VirtioVsock device and add it to the PCI bus upon hotplug. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	f8501a3bd3	vmm: config: Move --vsock syntax to VsockConfig This means it can be reused with ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Sebastien Boeuf	6e049e0da1	vmm: Add an identifier to the --vsock device It's possible to have multiple vsock devices so in preparation for hotplug/unplug it is important to be able to have a unique identifier for each device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	10348f73e4	vmm, main: Support only zero or one vsock devices The Linux kernel does not support multiple virtio-vsock devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-28 20:07:18 +02:00
Rob Bradford	9d1f95a3cc	openapi: Add missing "id" field NetConfig/DiskConfig/PmemConfig/FsConfig were all missing the id field in the API yaml file. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-28 18:27:45 +02:00
Muminul Islam	e1a07ce3c4	vmm: vm: Unpark the threads before shutdown when the current state is paused If the current state is paused that means most of the handles got killed by pthread_kill We need to unpark those threads to make the shutdown worked. Otherwise The shutdown API hangs and the API is not responding afterwards. So before the shutdown call we need to resume the VM make it succeed. Fixes: #817 Signed-off-by: Muminul Islam <muislam@microsoft.com>	2020-04-27 09:09:12 +02:00
Rob Bradford	1df38daf74	vmm, tests: Make specifying a size optional for virtio-pmem If a size is specified use it (in particular this is required if the destination is a directory) otherwise seek in the file to get the size of the file. Add a new check that the size is a multiple of 2MiB otherwise the kernel will reject it. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-24 18:30:05 +01:00
Rob Bradford	7481e4d959	vmm: config: Validate that shared memory is enabled if using vhost-user Check that if any device using vhost-user (net & disk with vhost_user=true) or virtio-fs is enabled then check shared memory is also enabled. Fixes: #848 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-24 16:01:49 +01:00
Bo Chen	2ac6971a8b	vmm: MemoryManager: Cleanup the usage of std::ffi/io/result Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-04-23 21:39:51 +02:00
Bo Chen	3f42f86d81	vmm: Add the 'shared' and 'hugepages' controls to MemoryConfig The new 'shared' and 'hugepages' controls aim to replace the 'file' option in MemoryConfig. This patch also updated all related integration tests to use the new controls (instead of providing explicit paths to "/dev/shm" or "/dev/hugepages"). Fixes: #1011 Signed-off-by: Rob Bradford <robert.bradford@intel.com> Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-04-23 21:39:51 +02:00
Martin Xu	5a380a6918	vmm: memory_manager: Support non-power-of-2 block sizes Replace alignment calculation of start address with functionally equivalent version that does not assume that the block size is a power of two. Signed-off-by: Martin Xu <martin.xu@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-22 09:11:51 +02:00
Sebastien Boeuf	c22fd39170	vmm: Remove virtio device's userspace mapping on hot-unplug When a virtio device is dynamically removed from the VM through the hot-unplug mechanism, every mapping associated with it must be properly removed. Based on the previous patches letting a VirtioDevice expose the list of userspace mappings associated with it, this patch can now remove all the KVM userspace memory regions through the MemoryManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	0a97c25464	vmm: Extend MemoryManager to remove userspace mappings The same way we added a helper for creating userspace memory mappings from the MemoryManager, this patch adds a new helper to remove some previously added mappings. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	fbcf3a7a7a	vm-virtio: Implement userspace_mappings() for virtio-pmem When hot-unplugging the virtio-pmem from the VM, we don't remove the associated userspace mapping. This patch will let us fix this in a following patch. For now, it simply adapts the code so that the Pmem device knows about the mapping associated with it. By knowing about it, it can expose it to the caller through the new userspace_mappings() function. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	18f7789a81	vmm: Add hotplugged virtio devices to the DeviceManager list The hotplugged virtio devices were not added to the list of virtio devices from the DeviceManager. This patch fixes it, as it was causing hotplugged virtio-fs devices from not supporting memory hotplug, since they were never getting the update as they were not part of the list of virtio devices held by the DeviceManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-20 20:36:26 +02:00
Dean Sheather	c2abadc293	vmm: Add ability to add virtio-fs device post-boot Adds DeviceManager method `make_virtio_fs_device` which creates a single device, and modifies `make_virtio_fs_devices` to use this method. Implements the new `vm.add-fs route`. Signed-off-by: Dean Sheather <dean@coder.com>	2020-04-20 20:36:26 +02:00
Dean Sheather	bb2139a408	vmm/api: Add vm.add-fs route Currently unimplemented. Once implemented, this API will allow for creating virtio-fs devices in the VM after it has booted. Signed-off-by: Dean Sheather <dean@coder.com>	2020-04-20 20:36:26 +02:00
Sebastien Boeuf	d35e775ed9	vmm: Update KVM userspace mapping when PCI BAR remapping In the context of the shared memory region used by virtio-fs in order to support DAX feature, the shared region is exposed as a dedicated PCI BAR, and it is backed by a KVM userspace mapping. Upon BAR remapping, the BAR is moved to a different location in the guest address space, and the KVM mapping must be updated accordingly. Additionally, we need the VirtioDevice to report the updated guest address through the shared memory region returned by get_shm_regions(). That's why a new setter is added to the VirtioDevice trait, so that after the mapping has been updated for KVM, we can tell the VirtioDevice the new guest address the shared region is located at. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-20 16:01:25 +02:00
Sebastien Boeuf	ac7178ef2a	vmm: Keep migratable devices list as a Vec The order the elements are pushed into the list is important to restore them in the right order. This is particularly important for MmioDevice (or VirtioPciDevice) and their VirtioDevice counterpart. A device must be fully ready before its associated transport layer management can trigger its restoration, which will end up activating the device in most cases. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-17 19:29:41 +02:00
Rob Bradford	e7e0e8ac38	vmm, devices: Add firmware debug port device OVMF and other standard firmwares use I/O port 0x402 as a simple debug port by writing ASCII characters to it. This is gated under a feature that is not enabled by default. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-17 12:54:00 +02:00
Rob Bradford	f9a0445c3d	vmm: vm: Remove device from configuration after unplug This ensures that a device that is removed will not reappear after a reboot. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	444e5c2a04	vmm: device_manager: Generalise NoAvailableVfioDeviceName We now support assigning device ids for VFIO and virtio-pci devices so this error can be generalised. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	5bab9c3894	vmm: device_manager: Assign ids to pmem/net/disk devices if absent If the id has not been provided by the user generate an incrementing id. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	514491a051	vmm: device_manager: Support unplugging virtio-pci devices Extend the eject_device() method on DeviceManager to also support virtio-pci devices being unplugged. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	476e4ce24f	vmm: device_manager: Add virtio-pci devices into id to BDF map In order to support hotplugging there is a map of human readable device id to PCI BDF map. As the device id is part of the specific device configuration (e.g. NetConfig) it is necessary to return the id through from the helper functions that create the devices through to the functions that add those devices to the bus. This necessitates changing a great deal of function prototypes but otherwise has little impact. Currently only if an id is supplied by the user as part of the device configuration is it populated into this map. A later commit will populate with an autogenerated name where none is supplied by the user. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	b38470df4b	vmm: config: Add "id" parameter to {Net, Disk, Pmem}Config This id will be used to unplug the device if the user has chosen an id. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	1beb62ed2d	vmm: vm: Don't panic on kernel load error Rather than panic()ing when we get a kernel loading error populate the error upwards. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	72fdfff15d	vmm: device_manager: Remove unused "_mmap_regions" member Now that ownership of the memory regions used for the virtio-pmem and vhost-user-fs devices have been moved into those devices it is no longer necessary to track them inside DeviceManager. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Rob Bradford	70ecd6bab4	vmm, virtio: fs: Move freeing of mappped region into device Move the release of the managed memory region from the DeviceManager to the vhost-user-fs device. This ensures that the memory will be freed when the device is unplugged which will lead to it being Drop()ed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Rob Bradford	0c6706a510	vmm, virtio: pmem: Move freeing of mappped region into device Move the release of the managed memory region from the DeviceManager to the virtio-pmem device. This ensures that the memory will be freed when the device is unplugged which will lead to it being Drop()ed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Sebastien Boeuf	b1554642e4	vmm: seccomp: Add missing mremap() syscall While testing self spawned vhost-user backends, it appeared that the backend was aborting due to a missing system call in the seccomp filters. mremap() was the culprit and this patch simply adds it to the whitelist. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-14 14:11:41 +02:00
Rob Bradford	28abfa9de5	vmm: openapi: Mark "initramfs" field nullable This should make it a pointer in the Go generated code so that it will be ommitted and thus not populated with an unhelpful default value. Fixes: #1015 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-09 23:25:18 +02:00
Rob Bradford	c260640fd5	vmm: config: Use Default::default() value for initramfs field This ensures that the field is filled with None when it is not specified as part of the deserialisation step. Fixes: #1015 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-09 17:28:45 +02:00
Alejandro Jimenez	7134f3129f	vmm: Allow PVH boot with initramfs We can now allow guests that specify an initramfs to boot using the PVH boot protocol. Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-04-09 17:28:03 +02:00
Rob Bradford	2d3f518c72	vmm: config: Error if both socket and path are specified for a disk This allows the validation of this requirement for both command line booted VMs and those booted via the API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	eeb7e2529d	vmm: config: Move max vCPUs > boot vCPUs check to validate() This allows the validation of this requirement for both command line booted VMs and those booted via the API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	12edb24678	vmm: config: Validate that serial/console file mode has a path Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	aaf382eee2	vmm: Move kernel check to VmConfig::validate() method Replace the existing VmConfig::valid() check with a call into .validate() as part of earlier config setup or boot API checks. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	3b0da2d895	vmm: vm: Validate configuration on API boot When performing an API boot validate the configuration. For now only some very basic validation is performed but in subsequent commits the validation will be extended. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	99b2ada4d0	vmm: Start splitting configuration parsing and validation The configuration comes from a variety of places (commandline, REST API and restore) however some validation was only happening on the command line parsing path. Therefore introduce a new ability to validate the configuration before proceeding so that this can be used for commandline and API boots. For now move just the console and serial output mode validation under the new validation API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Sebastien Boeuf	0ea706faf5	vmm: openapi: Update OpenAPI definition with RestoreConfig Making sure the OpenAPI definition is up to date with newly added structure and parameters to support VM restoration. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	8d9d22436a	vmm: Add "prefault" option when restoring Now that the restore path uses RestoreConfig structure, we add a new parameter called "prefault" to it. This will give the user the ability to populate the pages corresponding to the mapped regions backed by the snapshotted memory files. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	a517ca23a0	vmm: Move restore parameters into common RestoreConfig structure The goal here is to move the restore parameters into a dedicated structure that can be reused from the entire codebase, making the addition or removal of a parameter easier. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	6712958f23	vmm: memory: Add prefault option when creating region When CoW can be used, the VM restoration time is reduced, but the pages are not populated. This can lead to some slowness from the guest when accessing these pages. Depending on the use case, we might prefer a slower boot time for better performances from guest runtime. The way to achieve this is to prefault the pages in this case, using the MAP_POPULATE flag along with CoW. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	b2cdee80b6	vmm: memory: Restore with Copy-on-Write when possible This patch extends the previous behavior on the restore codepath. Instead of copying the memory regions content from the snapshot files into the new memory regions, the VMM will use the snapshot region files as the backing files behind each mapped region. This is done in order to reduce the time for the VM to be restored. When the source VM has been initially started with a backing file, this means it has been mapped with the MAP_SHARED flag. For this case, we cannot use the CoW trick to speed up the VM restore path and we simply fallback onto the copy of the memory regions content. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	d771223b2f	vmm: memory: Extend new() to support external backing files Whenever a MemoryManager is restored from a snapshot, the memory regions associated with it might need to directly back the mapped memory for increased performances. If that's the case, a list of external regions is provided and the MemoryManager should simply ignore what's coming from the MemoryConfig. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	ee5a041a0f	vmm: memory: Add Copy-on-Write parameter when creating region Now that we can choose specific mmap flags for the guest RAM, we create a new parameter "copy_on_write" meaning that the memory mappings backed by a file should be performed with MAP_PRIVATE instead of MAP_SHARED. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	be4e1e8712	vmm: memory: Use fine grained mmap wrapper In order to anticipate the need for special mmap flags when memory mapping the guest RAM, we need to switch from from_file() wrapper to build() wrapper. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	b9f9f01fcc	vmm: Extend seccomp filters to allow snapshot/restore A few KVM ioctls were missing in order to perform both snapshot and restore while keeping seccomp enabled. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	6eb721301c	vmm: Enable restore feature This connects the dots together, making the request from the user reach the actual implementation for restoring the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	53613319cc	vmm: Enable snapshot feature This connects the dots together, making the request from the user reach the actual implementation for snapshotting the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	2cd0bc0a2c	vmm: Create initial VM from its snapshot The MemoryManager is somehow a special case, as its restore() function was not implemented as part of the Snapshottable trait. Instead, and because restoring memory regions rely both on vm.json and every memory region snapshot file, the memory manager is restored at creation time. This makes the restore path slightly different from CpuManager, Vcpu, DeviceManager and Vm, but achieve the correct restoration of the MemoryManager along with its memory regions filled with the correct content. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	b55b83c6e8	vmm: vm: Implement the Transportable trait This is only implementing the send() function in order to store all Vm states into a file. This needs to be extended for live migration, by adding more transport methods, and also the recv() function must be implemented. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	1ed357cf34	vmm: vm: Implement the Snapshottable trait By aggregating snapshots from the CpuManager, the MemoryManager and the DeviceManager, Vm implements the snapshot() function from the Snapshottable trait. And by restoring snapshots from the CpuManager, the MemoryManager and the DeviceManager, Vm implements the restore() function from the Snapshottable trait. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	20ba271b6c	vmm: memory_manager: Implement the Transportable trait This implements the send() function of the Transportable trait, so that the guest memory regions can be saved into one file per region. This will need to be extended for live migration, as it will require other transport methods and the recv() function will need to be implemented too. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	e606112cef	vmm: memory_manager: Implement the Snapshottable trait In order to snapshot the content of the guest RAM, the MemoryManager must implement the Snapshottable trait. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	50b3f008d1	vmm: cpu: Implement the Snapshottable trait Implement the Snapshottable trait for Vcpu, and then implements it for CpuManager. Note that CpuManager goes through the Snapshottable implementation of Vcpu for every vCPU in order to implement the Snapshottable trait for itself. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	f787c409c4	vmm: cpu: Factorize vcpu starting code Anticipating the need for a slightly different function for restoring vCPUs, this patch factorizes most of the vCPU creation, so that it can be reused for migration purposes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Cathy Zhang	722f9b6628	vmm: cpu: Get and set KVM vCPU state These two new helpers will be useful to capture a vCPU state and being able to restore it at a later time. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Cathy Zhang	13756490b5	vmm: cpu: Track all Vcpus through CpuManager In anticipation for the CpuManager to aggregate all Vcpu snapshots together, this change makes sure the CpuManager has a handle onto every vCPU. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	a0d5dbce6c	vmm: device_manager: Implement the Snapshottable trait Based on the list of Migratable devices stored by the DeviceManager, the DeviceManager can implement the Snapshottable trait by aggregating all devices snapshots together. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	93d3abfd6e	vmm: device_manager: Make serial and ioapic devices migratable Serial and Ioapic both implement the Migratable trait, hence the DeviceManager can store them in the list of Migratable devices. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Rob Bradford	c7dfbd8a84	vmm: config: Implement fmt::Display for error Fixes: #367 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	d8119fda13	vmm: config: Remove unused error entries These entries are not currently used. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	1a10f16ad0	vmm: config: Consolidate size parsing code The parse_size helper function can now be consolidated into the ByteSized FromStr implementation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	f449486b9b	vmm: config: Make toggle parsing more tolerant Support "true" and "false" as well as well as capitalised forms. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	a4e0ce58c7	vmm: config: Consolidate on/off parsing Now all parsing code makes use of the Toggle and it's FromStr support move the helper function into the from_str() implementation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	c731a943d4	vmm: config: Port vsock to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	37264cf21b	vmm: config: Add unit testing for vsock Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	8665898ff3	vmm: config: Port device parsing to OptionParser Also make the "path" option required and generate an error if it is not provided. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	a85e2fa735	vmm: config: Add unit test for VFIO device parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	bed282b801	vmm: config: Add "valueless" options to OptionParser Valueless options are those like "off" or "tty" as used by the console options. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	2ae3392d32	vmm: config: Port console parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	143d63c88e	vmm: config: Add unit test for console parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	5ab58e743a	vmm: config: Port pmem option to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	233ad78b3a	vmm: config: Add parsing test for pmem Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	13dc637350	vmm: config: Port filesystem parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	7a071c28db	vmm: config: Implement unit testing for virtio-fs parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	e4cd3072d4	vmm: config: Port RNG options to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	708dbb973a	vmm: config: Add RNG parsing unit test Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	057e71d266	vmm: config: Accept empty value strings The integration tests and documentation make use of empty value strings like "--net tap=" accept them but return None so that the default value will be used as expected. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	218c780f67	vmm: config: Port network parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	8754720e2d	vmm: config: Add unit test for net parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	224e3ddef4	vmm: config: Switch disk parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e10244716	vmm: config: Add unit test for disk parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	e40ae6274b	vmm: config: Port memory option parsing to OptionParser This simplifies the parsing of the option by using OptionParser along with its automatic conversion behaviour. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	be32065aa4	vmm: config: Add "ByteSized" type for simplifying parsing of byte sizes Byte sizes are quantities ending in "K", "M", "G" and by implementing this type with a FromStr implementation the values can be converted using .parse(). Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	f01bd7d56d	vmm: config: Implement FromStr for HotplugMethod This allows the use of .parse() to automatically convert the string to the enum. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	746138039d	vmm: config: Add a Toggle type for "on/off" strings Some of the config parameters take an "on" or "off". Add a way to neatly parse that. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	929142bc2e	vmm: config: Add memory parsing unit test Before porting over to OptionParser add a unit test to validate the current memory parsing code. This showed up a bug where the "size=" was always required. Temporarily resolve this by assigning the string a default value which will later be replaced when the code is refactored. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	68203ea414	vmm: config: Port CPU parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e6a2825ba	vmm: config: Add unit test for CPU parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e7231cd69	vmm: config: Introduce basic OptionParser This will be used to simplify and consolidate much of the parsing code used for command line parameters. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Samuel Ortiz	447af8e702	vmm: vm: Factorize the device and cpu managers creation routine Into a new_from_memory_manager() routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	c73c9b112c	vmm: vm: Open kernel and initramfs once all managers are created Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	0646a90626	vmm: cpu: Pass CpusConfig to simplify the new() prototype Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	b584ec3fb3	vmm: memory_manager: Own the system allocator Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	ef2b11ee6c	vmm: memory_manager: Pass MemoryConfig to simplify the new() prototype Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	622f3f8fb6	vmm: vm: Avoid ioapic variable creation For a more readable VM creation routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	164e810069	vmm: cpu: Move CPUID patching to CpuManager Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	1a2c1f9751	vmm: vm: Factorize the KVM setup code Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	7a50646c02	vmm: device_manager: Convert migratable_devices to a map We must be able to map a migratable component id to its device. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	8f300bed83	vmm: api: Add a /api/v1/vm.restore endpoint Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	92c73c3b78	vmm: Add a VmRestore command Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	39d4f817f0	vmm: http: Add a /api/v1/vm.snapshot endpoint Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	cf8f8ce93a	vmm: api: Add a Snapshot command Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-02 13:24:25 +01:00
Sebastien Boeuf	452475c280	vmm: Add migration helpers Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	1b1a2175ca	vm-migration: Define the Snapshottable and Transportable traits A Snapshottable component can snapshot itself and provide a MigrationSnapshot payload as a result. A MigrationSnapshot payload is a map of component IDs to a list of migration sections (MigrationSection). As component can be made of several Migratable sub-components (e.g. the DeviceManager and its device objects), a migration snapshot can be made of multiple snapshot itself. A snapshot is a list of migration sections, each section being a component state snapshot. Having multiple sections allows for easier and backward compatible migration payload extensions. Once created, a migratable component snapshot may be transported and this is what the Transportable trait defines, through 2 methods: send and recv. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-02 13:24:25 +01:00
Sebastien Boeuf	2d17f4384a	vmm: seccomp: Add missing open() syscall On some systems, the open() system call is used by Cloud-Hypervisor, that's why it should be part of the seccomp filters whitelist. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-02 09:56:48 +02:00
Sebastien Boeuf	e4ea8b0bef	vmm: Add missing syscalls to the seccomp filters Both clock_gettime and gettimeofday syscalls where missing when running Cloud-Hypervisor on a Linux host without vDSO enabled. On a system with vDSO enabled, the syscalls performed by vDSO were not filtered, that's why we didn't have to whitelist them. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 16:50:52 +00:00
Sebastien Boeuf	9e18177654	vmm: Add memory hotplug support to VFIO PCI devices Extend the update_memory() method from DeviceManager so that VFIO PCI devices can update their DMA mappings to the physical IOMMU, after a memory hotplug has been performed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 09:35:39 +01:00
Sebastien Boeuf	cc67131ecc	vmm: Retrieve new memory region when memory is extended Whenever the memory is resized, it's important to retrieve the new region to pass it down to the device manager, this way it can decide what to do with it. Also, there's no need to use a boolean as we can instead use an Option to carry the information about the region. In case of virtio-mem, there will be no region since the whole memory has been reserved up front by the VMM at boot. This means only the ACPI hotplug will return a region and is the only method that requires the memory to be updated from the device manager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 09:35:39 +01:00
Samuel Ortiz	8fc7bf2953	vmm: Move to the latest linux-loader Commit 2adddce2 reorganized the crate for a cleaner multi architecture (x86_64 and aarch64) support. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-27 08:48:20 +01:00
Sebastien Boeuf	785812d976	vmm: Fallback to legacy boot if PVH is enabled along with initramfs For now, the codebase does not support booting from initramfs with PVH boot protocol, therefore we need to fallback to the legacy boot. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	6cce7b9560	arch: load initramfs and populate zero page * load the initramfs File into the guest memory, aligned to page size * finally setup the initramfs address and its size into the boot params (in configure_64bit_boot) Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	1f9bc68c54	openapi: Add initramfs support added InitramfsConfig property to the REST API spec Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	4db252b418	main, vmm: add --initramfs cli option currently unused, the initramfs argument is added to the cli, and stored in vmm::config:VmConfig as an Option(InitramfsConfig(PathBuf)) Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Rob Bradford	6244beb9d5	openapi: Add "vm.add-net" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	57c3fa4b1e	vmm: Add "add-net" to the API Add the HTTP and internal API entry points for adding a network device at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	f664cddec9	vmm: Add support for adding network devices to the VM The persistent memory will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	8f323e61d8	vmm: Add support to DeviceManager for hotplugging network devices Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	42a9896fe4	vmm: device_manager: Refactor make_virtio_net_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	9df601a1df	bin, vmm: Centralise the net syntax This will allow the syntax to be reused with cloud-hypervsor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Samuel Ortiz	41d7b3a387	vmm: memory_manager: Only send the GED notification for the ACPI method Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	15d9ec0149	openapit: Add hotplug_method to MemoryConfig Add hotplug_method to MemoryConfig in cloud-hypervisor.yaml. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	e63f98182a	vmm: device: Add make_virtio_mem_devices Add make_virtio_mem_devices to add virtio-mem to vmm. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	e6b934a56a	vmm: Add support for virtio-mem This commit adds new option hotplug_method to memory config. It can set the hotplug method to "acpi" or "virtio-mem". Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Rob Bradford	75878dd90a	openapi: Add "vm.add-pmem" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	f6f4c68fb4	vmm: Add "add-pmem" to the API Add the HTTP and internal API entry points for adding persistent memory at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	15de30f141	vmm: Add support for adding pmem devices to the VM The persistent memory will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	f7def621dd	vmm: Add support to DeviceManager for hotplugging pmem devices Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	8c3ea8cd76	vmm: device_manager: Refactor make_virtio_pmem_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	a7296bbb52	bin, vmm: Centralise the pmem syntax This will allow the syntax to be reused with cloud-hypervisor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	4c9d15d44c	vmm: Fix copy and paste error message vm_remove_device was copied from vm_add_device but the error message wasn't correctly updated. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	82cad99c0b	openapi: Add "vm.add-disk" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	f2151b2734	vmm: Add "add-disk" to the API Add the HTTP and internal API entry points for adding disks at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	164ec2b8e6	vmm: Add support for adding disks to the VM The disk will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	b3082c1984	vmm: Add support to DeviceManager for hotplugging disks Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	2be703ca92	vmm: device_manager: Refactor make_virtio_block_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	66da29d8dd	bin, vmm: Centralise the disk syntax This will allow the syntax to be reused with cloud-hypervsor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Sebastien Boeuf	e54f8ec8a5	vmm: Update memory through DeviceManager Whenever the VM memory is resized, DeviceManager needs to be notified so that it can subsequently notify each virtio devices about it. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 19:01:15 +00:00
Sebastien Boeuf	feb8d7ae90	vmm: Separate seccomp filters between VMM and API threads This separates the filters used between the VMM and API threads, so that we can apply different rules for each thread. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	f1a23d712f	vmm: api: Add seccomp to the HTTP API thread Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	db62cb3f4d	vmm: Add seccomp filter to the VMM thread This commit introduces the application of the seccomp filter to the VMM thread. The filter is empty for now (SeccompLevel::None). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	cb98d90097	vmm: Create new seccomp_filter module Based on the seccomp crate, we create a new vmm module responsible for creating a seccomp filter that will be applied to the VMM main thread. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Rob Bradford	f7197e8415	vmm: Add a "discard_writes=" to --pmem This opens the backing file read-only, makes the pages in the mmap() read-only and also makes the KVM mapping read-only. The file is also mapped with MAP_PRIVATE to make the changes local to this process only. This is functional alternative to having support for making a virtio-pmem device readonly. Unfortunately there is no concept of readonly virtio-pmem (or any type of NVDIMM/PMEM) in the Linux kernel so to be able to have a block device that is appears readonly in the guest requires significant specification and kernel changes. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Rob Bradford	d11a67b0fe	vmm: Use more generic MmapRegion constructor Switch to MmapRegion::build() and fill in the fields appropriately. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Rob Bradford	7257e890ef	vmm: Add "readonly" parameter MemoryManager::create_userspace_mapping Use this boolean to turn on the KVM_MEM_READONLY flag to indicate that this memory mapping should not be writable by the VM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Qiu Wenbo	c503118d16	vmm: fix a corrupted stack caused by get_win_size According to `asm-generic/termios.h`, the `struct winsize` should be: struct winsize { unsigned short ws_row; unsigned short ws_col; unsigned short ws_xpixel; unsigned short ws_ypixel; }; The ioctl of TIOCGWINSZ will trigger a segfault on aarch64. Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2020-03-20 07:30:06 +01:00
Rob Bradford	0788600702	build: Remove "pvh_boot" feature flag This feature is stable and there is no need for this to be behind a flag. This will also reduce the time needed to run the integration test as we will not be running them all again under the flag. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-19 13:05:44 +00:00
Rob Bradford	477bc17f18	bin: Share VFIO device syntax between cloud-hypervisor and ch-remote Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-18 23:38:55 +00:00
Jose Carlos Venegas Munoz	a31ffef085	openapi: Add hotplug_size for memory hotplug Add hotplug_size, needed to be defined when hotplug is used. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2020-03-18 19:06:07 +00:00
Rob Bradford	87990f9e67	vmm: Add virtio-pci device to B/D/F hash table This table currently contains only all the VFIO devices and it should really contain all the PCI devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-18 19:05:58 +00:00
Rob Bradford	fb185fa839	vmm: Always return PCI B/D/F from add_virtio_pci_device Previously this was only returned if the device had an IOMMU mapping and whether the device should be added to the virtio-iommu. This was already captured earlier as part of creating the device so use that information instead. Always returning the B/D/F is helpful as it facilitates virtio PCI device hotplug. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-18 19:05:58 +00:00
Samuel Ortiz	63eeed29cc	vm: Comment on the VM config update from memory hotplug I spent a few minutes trying to understand why we were unconditionally updating the VM config memory size, even if the guest memory resizing did not happen. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-18 12:48:40 +01:00
Rob Bradford	28a5f9dc19	vmm: acpi: Remove unused IORT related structures The IORT table for virtio-iommu use was removed and replaced with a purely virtio based solution. Although the table construction was removed these structures were left behind. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-17 12:46:26 +00:00
Alejandro Jimenez	a22bc3559f	pvh: Write start_info structure to guest memory Fill the hvm_start_info and related memory map structures as specified in the PVH boot protocol. Write the data structures to guest memory at the GPA that will be stored in %rbx when the guest starts. Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-03-13 18:29:44 +01:00
Alejandro Jimenez	840a9a97ff	pvh: Initialize vCPU regs/sregs for PVH boot Set the initial values of the KVM vCPU registers as specified in the PVH boot ABI: https://xenbits.xen.org/docs/unstable/misc/pvh.html Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-03-13 18:29:44 +01:00
Alejandro Jimenez	24f0e42e6a	pvh: Introduce EntryPoint struct In order to properly initialize the kvm regs/sregs structs for the guest, the load_kernel() return type must specify which boot protocol to use with the entry point address it returns. Make load_kernel() return an EntryPoint struct containing the required information. This structure will later be used in the vCPU configuration methods to setup the appropriate initial conditions for the guest. Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-03-13 18:29:44 +01:00
Rob Bradford	4579afa091	vmm: For --disk error if socket and path is specified This is an error as the path should be specfied by the unmanaged backend. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-13 11:41:52 +00:00
Rob Bradford	7e599b4450	vmm: Make disk path optional When using "--disk" with a vhost socket and not using self spawning then it is not necessary or helpful to specify the path. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-13 11:41:52 +00:00
Sebastien Boeuf	8d785bbd5f	pci: Fix the PciBus using HashMap instead of Vec By using a Vec to hold the list of devices on the PciBus, there's a problem when we use unplug. Indeed, the vector of devices gets reduced and if the unplugged device was not the last one from the list, every other device after this one is shifted on the bus. To solve this problem, a HashMap is used. This allows to keep track of the exact place where each device stands on the bus. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-13 10:54:34 +01:00
Jose Carlos Venegas Munoz	40b38a4222	openapi: Make desired_ram int64 format The option desired_ram is in byte, make larger the amount of memory to add. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2020-03-12 23:17:56 +01:00
Sebastien Boeuf	efba48dddb	vmm: Don't put a VFIO device behind the vIOMMU by default With some of the factorization that happened to be able to support VFIO hotplug, one mistake was made. In case a vIOMMU is created through a virtio-iommu device, and no matter the "iommu" option value from the VFIO device parameter, the VFIO device was always placed behind the virtual IOMMU. This commit fixes this wrong behavior by making sure the device configuration is taken into account to decide if it should be attached or not to the virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 19:50:31 +01:00
Sebastien Boeuf	34412c9b41	vmm: Add id option to VFIO hotplug Add a new id option to the VFIO hotplug command so that it matches the VFIO coldplug semantic. This is done by refactoring the existing code for VFIO hotplug, where VmAddDeviceData structure is replaced by DeviceConfig. This structure is the one used whenever a VFIO device is coldplugged, which is why it makes sense to reuse it for the hotplug codepath. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 19:50:31 +01:00
Sebastien Boeuf	9023444ad3	vmm: Add id field to --device through CLI Add the ability to specify the "id" associated with a device, by adding an extra option to the parameter --device. This new option is not mandatory, and by default, the VMM will take care of finding a unique identifier. If the identifier provided by the user through this new option is not unique, an error will be thrown and the VM won't be started. Fixes #881 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 13:10:57 +00:00
Sebastien Boeuf	f4a956a60a	vmm: Remove 32 bits MMIO range from correct address space The 32 bits MMIO address space is handled separately from the 64 bits one. For this reason, we need to invoke the appropriate freeing function to remove a range from this address space. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 13:10:30 +00:00
Sebastien Boeuf	432eb5b70a	vmm: Free PCI BARs when unplugging PCI device Now that PciDevice trait has a dedicated function to remove the bars, the DeviceManager can invoke this function whenever a PCI device is unplugged from the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-11 13:10:30 +00:00
Sebastien Boeuf	b50cbe5064	pci: Give PCI device ID back when removing a device Upon removal of a PCI device, make sure we don't hold onto the device ID as it could be reused for another device later. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	df71aaee3f	pci: Make the device ID allocation smarter In order to handle the case where devices are very often plugged and unplugged from a VM, we need to handle the PCI device ID allocation better. Any PCI device could be removed, which means we cannot simply rely on the vector size to give the next available PCI device ID. That's why this patch stores in memory the information about the 32 slots availability. Based on this information, whenever a new slot is needed, the code can correctly provide an available ID, or simply return an error because all slots are taken. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	e514b124ed	vmm: Update VmConfig when removing VFIO device This commit ensures that when a VFIO device is hot-unplugged from the VM, it is also removed from the VmConfig. This prevents a potential reboot from creating the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	81173bf4ab	vmm: Add id field to DeviceConfig structure Add a new field to the DeviceConfig, allowing the VMM to allocate a name to the VFIO devices. By identifying a VFIO device with a unique name, we can make sure a user can properly unplug it at any time. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	6cbdb9aa47	vmm: api: Introduce new "remove-device" HTTP endpoint This commit introduces the new command "remove-device" that will let a user hot-unplug a VFIO PCI device from an already running VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	991f3bb5da	vmm: Remove VFIO device from everywhere it is referenced This commit implements the eject function so that a VFIO device will be removed from any bus it might sit on, and from any list it might be stored in. The idea is to reach a point where there is no reference of the device anywhere in the code, so that the Drop implementation will be invoked and so that the device will be fully removed from the VMM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	6adebbc6a0	vmm: Detect when guest notifies about ejecting PCI device When the guest OS is done removing a PCI device, it will invoke the _EJ0 method from ACPI, associated with the device. This will trigger a port IO write to a region known by the VMM. Upon this writing, the VMM will trap the VM exit and retrieve the written value. Based on the value, the VMM will invoke its eject_device() method to finalize the removal of the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	08604ac6a8	vmm: Store PCI devices as Any devices from DeviceManager As we try to keep track of every PCI device related to the VM, we don't want to have separate lists depending on the concrete type associated with the PciDevice trait. Also, we want to be able to cast the actual type into any trait or concrete type. The most efficient way to solve all these issues is to store every device as an Arc<dyn Any + Send + Sync>. This gives the ability to downcast into the appropriate concrete type, and then to cast back into any trait that we might need. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	0f99d3f7cc	vmm: Store VFIO device's name and its PCI b/d/f Add a new list storing the device names across the entire codebase. VFIO devices are added to the list whenever a new one is created. By default, each VFIO device is given a name "vfioX" where X is the first available integer. Along with this new list of names, another list is created, grouping PCI device's name with its associated b/d/f. This will be useful to keep track of the created devices so that we can implement unplug functionality. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-10 17:05:06 +00:00
Sebastien Boeuf	09829c44b2	vmm: Remove IO bus strong reference from Vm The Vm structure was used to store a strong reference to the IO bus. This is not needed anymore since the AddressManager is logically the one holding this strong reference. This has been made possible by the introduction of Weak references on the Bus structure itself. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	2dbb376175	vmm: Remove all Weak references from DeviceManager Now that the BusDevice devices are stored as Weak references by the IO and MMIO buses, there's no need to use Weak references from the DeviceManager anymore. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	9e915a0284	vmm: Remove all Weak references from CpuManager Now that the BusDevice devices are stored as Weak references by the IO and MMIO buses, there's no need to use Weak references from the CpuManager anymore. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	49268bff3b	pci: Remove all Weak references from PciBus Now that the BusDevice devices are stored as Weak references by the IO and MMIO buses, there's no need to use Weak references from the PciBus anymore. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	7773812f58	vmm: Store the list of BusDevice devices from DeviceManager The point is to make sure the DeviceManager holds a strong reference of each BusDevice inserted on the IO and MMIO buses. This will allow these buses to hold Weak references onto the BusDevice devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	d0820cc026	vmm: Make add_vfio_device mutable The method add_vfio_device() from the DeviceManager needs to be mutable if we want later to be able to update some internal fields from the DeviceManager from this same function. This commit simply takes care of making the necessary changes to change this function as mutable. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	948f808da6	vm: Rename DeviceManager field in Vm structure It's more logical to name the field referring to the DeviceManager as "device_manager" instead of "devices". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 18:46:44 +01:00
Sebastien Boeuf	d47f733e51	vmm: Break the cyclic dependency between DeviceManager and IO bus By inserting the DeviceManager on the IO bus, we introduced some cyclic dependency: DeviceManager ---> AddressManager ---> Bus ---> BusDevice ^ \| \| \| +---------------------------------------------+ This cycle needs to be broken by inserting a Weak reference instead of an Arc (considered as a strong reference). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	c1af13efeb	vmm: Update VmConfig when adding new device Ensures the configuration is updated after a new device has been hotplugged. In the event of a reboot, this means the new VM will be started with the new device that had been previously hotplugged. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	a86f4369a7	vmm: Add VFIO PCI device hotplug support This commit finalizes the VFIO PCI hotplug support, based on all the previous commits preparing for it. One thing to notice, this does not support vIOMMU yet. This means we can hotplug VFIO PCI devices, but we cannot attach them to an existing or a new virtio-iommu device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	320fea0eaf	vmm: Factorize VFIO PCI device creation This factorization is very important as it will allow both the standard codepath and the VFIO PCI hotplug codepath to rely on the same function to perform the addition of a new VFIO PCI device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	00716f90a0	vmm: Store virtio-iommu device from DeviceManager Helps with future refactoring of VFIO device creation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	5902dfa403	vmm: Store VFIO KVM device from DeviceManager Helps with future refactoring of VFIO device creation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	d9c1b4396e	vmm: Store MSI InterruptManager from DeviceManager Helps with future refactoring of VFIO device creation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	02adc4061a	vmm: Store PciBus from DeviceManager Helps with future refactoring of VFIO device creation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	d0218e94a3	vmm: Trigger hotplug notification to the guest Whenever the user wants to hotplug a new VFIO PCI device, the VMM will have to trigger a hotplug notification through the GED device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	0e58741a09	vmm: api: Introduce new "add-device" HTTP endpoint This commit introduces the new command "add-device" that will let a user hotplug a VFIO PCI device to an already running VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	0f1396acef	vmm: Insert PCI device hotplug operation region on IO bus Through the BusDevice implementation from the DeviceManager, and by inserting the DeviceManager on the IO bus for a specific IO port range, the VMM now has the ability to handle PCI device hotplug. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	65774e8a78	vmm: Implement BusDevice for DeviceManager In anticipation of inserting the DeviceManager on the IO/MMIO buses, the DeviceManager must implement the BusDevice trait. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	8dbc84318c	vmm: acpi: Add PCNT method to invoke DVNT Create a small method that will perform both hotplug of all the devices identified by PCIU bitmap, and then perform the hotunplug of all the devices identified by the PCID bitmap. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	c62db97a81	vmm: acpi: Add _EJ0 to each PCI device slot The _EJ0 method provides the guest OS a way to notify the VMM that the device has been properly ejected from the guest OS. Only after this point, the VMM can fully remove the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	4dc2a39f3a	vmm: acpi: Create PHPR container This new PHPR device in the DSDT table introduces some specific operation regions and the associated fields. PCIU stands for "PCI up", which identifies PCI devices that must be added. PCID stands for "PCI down", which identifies PCI devices that must be removed. B0EJ stands for "Bus 0 eject", which identifies which device on the bus has been ejected by the guest OS. Thanks to these fields, the VMM and the guest OS can communicate while performing hotplug/hotunplug operations. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	c3a0685e2d	vmm: acpi: Add notification method for PCI device slots Adds the DVNT method to the PCI0 device in the DSDT table. This new method is responsible for checking each slot and notify the guest OS if one of the slots is supposed to be added or removed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Sebastien Boeuf	5a68d5b6a7	vmm: acpi: Create PCI device slots This commit introduces the ACPI support for describing the 32 device slots attached to the main PCI host bridge. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-04 12:06:02 +00:00
Bin Liu	d6e6901957	vmm/api: Fix vm.info response definition Update cloud-hypervisor.yaml with latest code. Fixes: #841 Signed-off-by: liubin <liubin0329@gmail.com>	2020-03-03 09:34:25 +01:00
Sebastien Boeuf	8142c823ed	vmm: Move DeviceManager into an Arc<Mutex<>> In anticipation of the support for device hotplug, this commit moves the DeviceManager object into an Arc<Mutex<>> when the DeviceManager is being created. The reason is, we need the DeviceManager to implement the BusDevice trait and then provide it to the IO bus, so that IO accesses related to device hotplug can be handled correctly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-27 11:12:31 +01:00
Qiu Wenbo	9de3ace8c7	devices: implement Aml trait for GED device Fixes: #657 Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2020-02-25 08:32:16 +00:00
Sebastien Boeuf	b77fdeba2d	msi/msi-x: Prevent from losing masked interrupts We want to prevent from losing interrupts while they are masked. The way they can be lost is due to the internals of how they are connected through KVM. An eventfd is registered to a specific GSI, and then a route is associated with this same GSI. The current code adds/removes a route whenever a mask/unmask action happens. Problem with this approach, KVM will consume the eventfd but it won't be able to find an associated route and eventually it won't be able to deliver the interrupt. That's why this patch introduces a different way of masking/unmasking the interrupts, simply by registering/unregistering the eventfd with the GSI. This way, when the vector is masked, the eventfd is going to be written but nothing will happen because KVM won't consume the event. Whenever the unmask happens, the eventfd will be registered with a specific GSI, and if there's some pending events, KVM will trigger them, based on the route associated with the GSI. Suggested-by: Liu Jiang <gerry@linux.alibaba.com> Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-25 08:31:14 +00:00
Rob Bradford	bba5ef3a59	vmm: Remove deprecated CPU syntax Remove the old way of specifying the number of vCPUs to use. Fixes: #678 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-24 07:26:31 +01:00
Rob Bradford	374ac77c63	main, vmm: Remove deprecated --vhost-user-net This has been superseded by using --net with vhost_user=true and socket=<socket> Fixes: #678 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-24 07:26:31 +01:00
Rob Bradford	ffd816ebfa	main, vmm: Remove deprecated --vhost-user-blk This has been superseded by using --disk with vhost_user=true and socket=<socket> Fixes: #678 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-24 07:26:31 +01:00
Sergio Lopez	d2f1749edb	vmm: config: Add poll_queue property to DiskConfig Recently, vhost_user_block gained the ability of actively polling the queue, a feature that can be disabled with the poll_queue property. This change adds this property to DiskConfig, so it can be used through the "disk" argument. For the moment, it can only be used when vhost_user=true, but this will change once virtio-block gets the poll_queue feature too. Fixes: #787 Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-02-20 18:06:54 +01:00
Sergio Lopez	378dd81204	vmm: openapi: Add missing "direct" knob to DiskConfig Add missing "direct" knob that should be exposed through the REST API. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-02-20 18:06:54 +01:00
Sergio Lopez	056f5481ac	vmm: openapi: Fix "readonly" and "wce" defaults in DiskConfig Fix "readonly" and "wce" defaults in cloud-hypervisor.yaml to match their respective defaults in config.rs:DiskConfig. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-02-20 18:06:54 +01:00
Samuel Ortiz	c49e31a6d9	vmm: api: Return a resize error when resize fails And not a VmCreate one. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-20 12:26:12 +01:00
Samuel Ortiz	ebc6391bea	vmm: api: Fix resize command typos Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-20 12:26:12 +01:00
Samuel Ortiz	9de755334d	vmm: openapi: Update DiskConfig It's missing a few knobs (readonly, vhost, wce) that should be exposed through the rest API. Fixes: #790 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-20 12:17:50 +01:00
Rob Bradford	ed1e7817cc	vmm: Workaround double reboot triggered by the kernel The kernel does not adhere to the ACPI specification (probably to work around broken hardware) and rather than busy looping after requesting an ACPI reset it will attempt to reset by other mechanisms (such as i8042 reset.) In order to trigger a reset the devices write to an EventFd (called reset_evt.) This is used by the VMM to identify if a reset is requested and make the VM reboot. As the reset_evt is part of the VMM and reused for both the old and new VM it is possible for the newly booted VM to immediately get reset as there is an old event sitting in the EventFd. The simplest solution is to "drain" the reset_evt EventFd on reboot to make sure that there is no spurious events in the EventFd. Fixes: #783 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-19 18:51:14 +01:00
Sebastien Boeuf	793d4e7b8d	vmm: Move codebase to GuestMemoryAtomic from vm-memory Relying on the latest vm-memory version, including the freshly introduced structure GuestMemoryAtomic, this patch replaces every occurrence of Arc<ArcSwap<GuestMemoryMmap> with GuestMemoryAtomic<GuestMemoryMmap>. The point is to rely on the common RCU-like implementation from vm-memory so that we don't have to do it from Cloud-Hypervisor. Fixes #735 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-19 13:48:19 +00:00
Rob Bradford	1f6cbad01a	vmm: Add support for spawning vhost-user-block backend If no socket is supplied when enabling "vhost_user=true" on "--disk" follow the "exe" path in the /proc entry for this process and launch the network backend (via the vmm_path field.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-18 08:43:47 +00:00
Sebastien Boeuf	3edc2bd6ab	vmm: Prevent memory overcommitment through virtio-fs shared regions When a virtio-fs device is created with a dedicated shared region, by default the region should be mapped as PROT_NONE so that no pages can be faulted in. It's only when the guest performs the mount of the virtiofs filesystem that we can expect the VMM, on behalf of the backend, to perform some new mappings in the reserved shared window, using PROT_READ and/or PROT_WRITE. Fixes #763 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-17 15:03:47 +01:00
Rob Bradford	bc75c1b4e1	vmm: Add support for spawning vhost-user-net backend If no socket is supplied when enabling "vhost_user=true" on "--net" follow the "exe" path in the /proc entry for this process and launch the network backend (via the vmm_path field.) Currently this only supports creating a new tap interface as the network backend also only supports that. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-14 17:32:49 +00:00
Rob Bradford	b04eb4770b	vmm: Follow the "exe" symlink from the PID directory in /proc It is necessary to do this at the start of the VMM execution rather than later as it must be done in the main thread in order to satisfy the checks required by PTRACE_MODE_READ_FSCREDS (see proc(5) and ptrace(2)) The alternative is to run as CAP_SYS_PTRACE but that has its disadvantages. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-14 17:32:49 +00:00
Rob Bradford	7c9e8b103f	vmm: device_manager: Shutdown all virtio devices When the DeviceManager is dropped explicitly shutdown() all virtio devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-14 17:32:49 +00:00
Sebastien Boeuf	3447e226d9	dependencies: bump vm-memory from `4237db3` to `f3d1c27` This commit updates Cloud-Hypervisor to rely on the latest version of the vm-memory crate. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-06 11:40:45 +01:00
Sebastien Boeuf	62ccccc303	vmm: Make sure to retry creating the VM on EINTR If the ioctl syscall KVM_CREATE_VM gets interrupted while creating the VM, it is expected that we should retry since EINTR should not be considered a standard error. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-02-05 12:06:21 +01:00
Samuel Ortiz	da2b3c92d3	vm-device: interrupt: Remove InterruptType dependencies and definitions Having the InterruptManager trait depend on an InterruptType forces implementations into supporting potentially very different kind of interrupts from the same code base. What we're defining through the current, interrupt type based create_group() method is a need for having different interrupt managers for different kind of interrupts. By associating the InterruptManager trait to an interrupt group configuration type, we create a cleaner design to support that need as we're basically saying that one interrupt manager should have the single responsibility of supporting one kind of interrupt (defined through its configuration). Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-04 19:32:45 +01:00
Samuel Ortiz	84fc807bc6	interrupt: Interrupt manager split We create 2 different interrupt managers for separately handling creation of legacy and MSI interrupt groups. Doing so allows us to have a cleaner interrupt manager and IOAPIC initialization path. It also prepares for an InterruptManager trait design improvement where we remove the interrupt source type dependency by associating an interrupt configuration type to the trait. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-02-04 19:32:45 +01:00
Rob Bradford	880a57c920	vmm: Remove VmInfo struct After refactoring the VmInfo struct is no longer needed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	07bc292fa5	vmm: device_manager: Get VmFd from AddressManager A reference to the VmFd is stored on the AddressManager so it is not necessary to pass in the VmInfo into all methods that need it as it can be obtained from the AddressManager. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	6411c3ae42	vmm: device_manager: Use MemoryManager to get guest memory The DeviceManager has a reference to the MemoryManager so use that to get the GuestMemoryMmap rather than the version stored in the VmInfo struct. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	066fc6c0d1	vmm: device_manager: Get VM config from the struct member Remove the use of vm_info in methods to get the config and instead use the config stored on the DeviceManager itself. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	77ae3de4f3	vmm: device_manager: Make legacy device addition a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	599275b610	vmm: device_manager: Make ACPI device creation a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	b8c1b2e174	vmm: device_manager: Make console creation a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	b5440e2d0a	vmm: device_manager: Make virtio device creation functions methods Remove some in/out parameters and instead rely on them as members of the &mut self parameter. This prepares the way to more easily store state on the DeviceManager. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	e90c6f3c44	vmm: device_manager: Make make_virtio_devices a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. A follow-up commit will change the callee functions that create the devices themselves. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	dbc09ad0ef	vmm: device_manager: Make add_vfio_devices a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	d9e1c2cd22	vmm: device_manager: Make add_virtio_pci_device a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	aaa5e2e9ea	vmm: device_manager: Make add_virtio_mmio_device a method Remove some in/out parameters and instead rely on them as members of the &mut self parameter. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	2987476e0a	vmm: device_manager: Make add_pci_devices and add_mmio_devices methods Modify these functions to take an &mut self and become methods on DeviceManager. This allows the removal of some in/out parameters and leads the way to further refactoring and simplification. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	3dbae423bb	vmm: device_manager: Only add MemoryManager to I/O bus on ACPI builds The MemoryManager should only be included on the I/O bus when doing ACPI builds as that is the only time it will be interrogated. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Rob Bradford	68fa97eb0e	vmm: device_manager: Always embed MemoryManager in the struct Currently the MemoryManager is only used on the ACPI code paths after the DeviceManager has been created. This will change in a future commit as part of the refactoring so for now always include it but name it with underscore prefix to indicate it might not always be used. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-02-03 12:28:30 +00:00
Sebastien Boeuf	ac01ceddbb	vmm: Cleanup list of PCI IDs related to virtual IOMMU Now that devices attached to the virtual IOMMU are described through virtio configuration, there is no need for the DeviceManager to store the list of IDs for all these devices. Instead, things are handled locally when PCI devices are being added. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-30 10:37:40 +01:00
Sebastien Boeuf	097cff2d85	vmm: Use virtio topology for virtio-iommu Instead of relying on the ACPI tables to describe the devices attached to the virtual IOMMU, let's use the virtio topology, as the ACPI support is getting deprecated. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-30 10:37:40 +01:00
Rob Bradford	75e6762897	vmm: Give deprecation warning for "--vhost-user-blk" syntax This will be removed in a future release. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-29 08:06:37 +00:00
Rob Bradford	969b5ee4e8	vmm: config: Add warning about specifying "wce" without "vhost-user" Currently configuring WCE is only supported when using vhost-user. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-29 08:06:37 +00:00
Rob Bradford	aeeae661fc	vmm: Support vhost-user-block via "--disks" Add a socket and vhost_user parameter to this option so that the same configuration option can be used for both virtio-block and vhost-user-block. For now it is necessary to specify both vhost_user and socket parameters as auto activation is not yet implemented. The wce parameter for supporting "Write Cache Enabling" is also added to the disk configuration. The original command line parameter is still supported for now and will be removed in a future release. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-29 08:06:37 +00:00
Rob Bradford	2c6f528c23	vmm: Give deprecation warning for "--vhost-user-net" syntax This will be removed in a future release. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-28 12:39:26 +00:00
Rob Bradford	a831aa214c	vmm: Support vhost-user-net via "--net" Add a socket and vhost_user parameter to this option so that the same configuration option can be used for both virtio-net and vhost-user-net. For now it is necessary to specify both vhost_user and socket parameters as auto activation is not yet implemented. The original command line parameter is still supported for now. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-28 12:39:26 +00:00
Sebastien Boeuf	f5b53ae4be	vm-virtio: Implement multiqueue/multithread support for virtio-blk This commit improves the existing virtio-blk implementation, allowing for better I/O performance. The cost for the end user is to accept allocating more vCPUs to the virtual machine, so that multiple I/O threads can run in parallel. One thing to notice, the amount of vCPUs must be egal or superior to the amount of queues dedicated to the virtio-blk device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-28 09:26:53 +01:00
Sebastien Boeuf	08e47ebd4b	vmm: Add num_queues and queue_size parameters to virtio-blk The number of queues and the size of each queue were not configurable. In anticipation for adding multiqueue support, this commit introduces some new parameters to let the user decide about the number of queues and the queue size. Note that the default values for each of these parameters are identical to the default values used for vhost-user-blk, that is 1 for the number of queues and 128 for the queue size. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-28 09:26:53 +01:00
Sebastien Boeuf	0fa1e2c241	vmm: Handle mapping from devices regions through vm-memory Devices like virtio-pmem and virtio-fs require some dedicated memory region to be mapped. The memory mapping from the DeviceManager is being replaced by the usage of MmapRegion from the vm-memory crate. The unmap will happen automatically when the MmapRegion will be dropped, which should happen when the DeviceManager gets dropped. Fixes #240 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-24 17:56:49 +01:00
Sebastien Boeuf	148a9ed5ce	vmm: Fix map_err losing the inner error Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-24 12:42:09 +01:00
Sebastien Boeuf	06396593c9	net_util: Fix map_err losing the inner error Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-24 12:42:09 +01:00
Rob Bradford	a34893a402	Revert "vmm: Move MemoryManager from I/O ports to MMIO region" This reverts commit `03108fb88b`.	2020-01-24 12:08:31 +01:00
Rob Bradford	57ed006992	Revert "devices, vmm: Move GED device to MMIO region" This reverts commit `5e3c62dc6a`.	2020-01-24 12:08:31 +01:00
Rob Bradford	6120d0fb1b	Revert "vmm: Move CpuManager device to MMIO region" This reverts commit `980e03fa0a`.	2020-01-24 12:08:31 +01:00
Rob Bradford	980e03fa0a	vmm: Move CpuManager device to MMIO region Move the CpuManager device from the I/O bus to living in an MMIO region. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-23 16:04:58 +00:00
Rob Bradford	5e3c62dc6a	devices, vmm: Move GED device to MMIO region Move GED device reporting of required device type to scan into an MMIO region rather than an I/O port. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-23 16:04:58 +00:00
Rob Bradford	03108fb88b	vmm: Move MemoryManager from I/O ports to MMIO region Rather than have the MemoryManager device sit on the I/O bus allocate space for MMIO and add it to the MMIO bus. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-23 16:04:58 +00:00
Sebastien Boeuf	0042f1de75	ioapic: Rely fully on the InterruptSourceGroup to manage interrupts This commit relies on the interrupt manager and the resulting interrupt source group to abstract the knowledge about KVM and how interrupts are updated and delivered. This allows the entire "devices" crate to be freed from kvm_ioctls and kvm_bindings dependencies. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-23 11:20:08 +00:00
Sebastien Boeuf	2dca959084	ioapic: Create the InterruptSourceGroup from InterruptManager The interrupt manager is passed to the IOAPIC creation, and the IOAPIC now creates an InterruptSourceGroup for MSI interrupts based on it. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-23 11:20:08 +00:00
Sebastien Boeuf	52800a871a	vmm: Create an InterruptManager dedicated to IOAPIC By introducing a new InterruptManager dedicated to the IOAPIC, we don't have to solve the chicken and eggs problem about which of the InterruptManager or the Ioapic should be created first. It's also totally fine to have two interrupt manager instances as they both share the same list of GSI routes and the same allocator. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-23 11:20:08 +00:00
Qiu Wenbo	2034fc2d84	vmm: Fix LENGTH_OFFSET_HIGH of MemoryManager Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2020-01-22 12:33:38 +00:00
Sergio Lopez	925c862f98	vmm: device_manager: Add 'direct' support for virtio-blk vhost_user_blk already has it, so it's only fair to give it to virtio-blk too. Extend DiskConfig with a 'direct' property, honor it while opening the file backing the disk image, and pass it to vm_virtio::RawFile. Fixes #631 Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-01-21 13:39:45 +00:00
Sergio Lopez	fb79e75afc	vmm: device_manager: Add read-only support for virtio-blk vhost_user_blk already has it, so it's only fair to give it to virtio-blk too. Extend DiskConfig with a 'readonly' properly, and pass it to vm_virtio::Block. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-01-21 13:39:45 +00:00
Sebastien Boeuf	9ac06bf613	ci: Run clippy for each specific feature The build is run against "--all-features", "pci,acpi", "pci" and "mmio" separately. The clippy validation must be run against the same set of features in order to validate the code is correct. Because of these new checks, this commit includes multiple fixes related to the errors generated when manually running the checks. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 11:44:40 +01:00
Sebastien Boeuf	99f39291fd	pci: Simplify PciDevice trait There's no need for assign_irq() or assign_msix() functions from the PciDevice trait, as we can see it's never used anywhere in the codebase. That's why it's better to remove these methods from the trait, and slightly adapt the existing code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	a20b383be8	vmm: Always use a reference for InterruptManager Since the InterruptManager is never stored into any structure, it should be passed as a reference instead of being cloned. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	bb8cd9eb24	vmm: Use LegacyUserspaceInterruptGroup for acpi device This commit replaces the way legacy interrupts were handled with the brand new implementation of the legacy InterruptSourceGroup for KVM. Additionally, since it removes the last bit relying on the Interrupt trait, the trait and its implementation can be removed from the codebase. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	75e22ff34e	vmm: Use LegacyUserspaceInterruptGroup for serial device This commit replaces the way legacy interrupts were handled with the brand new implementation of the legacy InterruptSourceGroup for KVM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	8d7c4ea334	vmm: Use LegacyUserspaceInterruptGroup for mmio devices This commit replaces the way legacy interrupts were handled with the brand new implementation of the legacy InterruptSourceGroup for KVM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	12657ef59f	vmm: Fully implement LegacyUserspaceInterruptGroup Relying on the previous commits, the legacy interrupt implementation can be completed. The IOAPIC handler is used to deliver the interrupt that will be triggered through the trigger() method. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	f70c9937fb	vmm: Add ioapic to KvmInterruptManager By having a reference to the IOAPIC, the KvmInterruptManager is going to be able to initialize properly the legacy interrupt source group. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	c9ea235a0e	vmm: Add LegacyUserspaceInterruptGroup skeleton for legacy interrupts In order to be able to use the InterruptManager abstraction with virtio-mmio devices, this commit introduces InterruptSourceGroup's skeleton for legacy interrupts. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	2aabf58bf5	vmm: Move irq_routes creation to specific MSI use case When KvmInterruptManager initializes a new InterruptSourceGroup, it's only for PCI_MSI_IRQ case that it needs to allocate the GSI and create a new InterruptRoute. That's why this commit moves the general code into the specific use case. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	d34f31fe7b	vmm: Fix KvmInterruptManager when base is different from 0 When the base InterruptIndex is different from 0, the loop allocating GSI and HashMap entries won't work as expected. The for loop needs to start from base, but the limit must be base+count so that we allocate a number of "count" entries. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Sebastien Boeuf	e73cb1ff80	vmm: Initialize InterruptManager sooner In order to let the InterruptManager be shared across both PCI and MMIO devices, this commit moves the initialization earlier in the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-21 10:44:48 +01:00
Rob Bradford	3901a1dd7d	vmm: Log an error if VM resize fails As well as returing an error to the API caller. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-17 23:44:21 +01:00
Rob Bradford	76d9bf2792	vmm: Start memory slots at zero After refactoring a common function is used to setup these slots and that function takes care of allocating a new slot so it is not necessary to reserve the initial region slots. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-17 23:44:21 +01:00
Rob Bradford	0ab22fea2c	vmm: Only generate GED event when new DIMM added Avoid the ACPI scan in the guest OS when no new DIMM is hotplugged. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-17 23:44:21 +01:00
Rob Bradford	211786ab42	vmm: Only generate GED interrupt when the number of vCPUs has changed Avoid activity in the the guest OS if the number of vCPUs has not changed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-17 23:44:21 +01:00
Sebastien Boeuf	4bb12a2d8d	interrupt: Reorganize all interrupt management with InterruptManager Based on all the previous changes, we can at this point replace the entire interrupt management with the implementation of InterruptManager and InterruptSourceGroup traits. By using KvmInterruptManager from the DeviceManager, we can provide both VirtioPciDevice and VfioPciDevice a way to pick the kind of InterruptSourceGroup they want to create. Because they choose the type of interrupt to be MSI/MSI-X, they will be given a MsiInterruptGroup. Both MsixConfig and MsiConfig are responsible for the update of the GSI routes, which is why, by passing the MsiInterruptGroup to them, they can still perform the GSI route management without knowing implementation details. That's where the InterruptSourceGroup is powerful, as it provides a generic way to manage interrupt, no matter the type of interrupt and no matter which hypervisor might be in use. Once the full replacement has been achieved, both SystemAllocator and KVM specific dependencies can be removed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	92082ad439	vmm: Fully implement interrupt traits After the skeleton of InterruptManager and InterruptSourceGroup traits have been implemented, this new commit takes care of fully implementing the content of KvmInterruptManager (InterruptManager trait) and MsiInterruptGroup (InterruptSourceGroup). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	0f727127d5	vmm: Implement InterruptSourceGroup and InterruptManager skeleton This commit introduces an empty implementation of both InterruptManager and InterruptSourceGroup traits, as a proper basis for further implementation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	c396baca46	vm-virtio: Modify VirtioInterrupt callback into a trait Callbacks are not the most Rust idiomatic way of programming. The right way is to use a Trait to provide multiple implementation of the same interface. Additionally, a Trait will allow for multiple functions to be defined while using callbacks means that a new callback must be introduced for each new function we want to add. For these two reasons, the current commit modifies the existing VirtioInterrupt callback into a Trait of the same name. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	2381f32ae0	msix: Add gsi_msi_routes to MsixConfig Because MsixConfig will be responsible for updating KVM GSI routes at some point, it is necessary that it can access the list of routes contained by gsi_msi_routes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	9b60fcdc39	msix: Add VmFd to MsixConfig Because MsixConfig will be responsible for updating the KVM GSI routes at some point, it must have access to the VmFd to invoke the KVM ioctl KVM_SET_GSI_ROUTING. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	86c760a0d9	msix: Add SystemAllocator to MsixConfig The point here is to let MsixConfig take care of the GSI allocation, which means the SystemAllocator must be passed from the vmm crate all the way down to the pci crate. Once this is done, the GSI allocation and irq_fd creation is performed by MsixConfig directly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sebastien Boeuf	f5704d32b3	vmm: Move gsi_msi_routes creation to be shared across all PCI devices Because we will need to share the same list of GSI routes across multiple PCI devices (virtio-pci, VFIO), this commit moves the creation of such list to a higher level location in the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-17 23:43:45 +01:00
Sergio Lopez	a14aee9213	qcow: Use RawFile as backend instead of File Use RawFile as backend instead of File. This allows us to abstract the access to the actual image with a specialized layer, so we have a place where we can deal with the low-level peculiarities. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-01-17 17:28:44 +00:00
Sergio Lopez	c5a656c9dc	vm-virtio: block: Add support for alignment restrictions Doing I/O on an image opened with O_DIRECT requires to adhere to certain restrictions, requiring the following elements to be aligned: - Address of the source/destination memory buffer. - File offset. - Length of the data to be read/written. The actual alignment value depends on various elements, and according to open(2) "(...) there is currently no filesystem-independent interface for an application to discover these restrictions (...)". To discover such value, we iterate through a list of alignments (currently, 512 and 4096) calling pread() with each one and checking if the operation succeeded. We also extend RawFile so it can be used as a backend for QcowFile, so the later can be easily adapted to support O_DIRECT too. Signed-off-by: Sergio Lopez <slp@redhat.com>	2020-01-17 17:28:44 +00:00
Cathy Zhang	652e7b9b8a	vm-virtio: Implement multiple queue support for net devices Update the common part in net_util.rs under vm-virtio to add mq support, meanwhile enable mq for virtio-net device, vhost-user-net device and vhost-user-net backend. Multiple threads will be created, one thread will be responsible to handle one queue pair separately. To gain the better performance, it requires to have the same amount of vcpus as queue pair numbers defined for the net device, due to the cpu affinity. Multiple thread support is not added for vhost-user-net backend currently, it will be added in future. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2020-01-17 12:06:19 +01:00
Cathy Zhang	404316eea1	vmm: Add multiple queue option and update config for virtio-net device Add num_queues and queue_size for virtio-net device to make them configurable, while add the associated options in command line. Update cloud-hypervisor.yaml with the new options for NetConfig. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2020-01-17 12:06:19 +01:00
Cathy Zhang	4ab88a8173	net_util: Add multiple queue support for tap Add support to allow VMMs to open the same tap device many times, it will create multiple file descriptors meanwhile. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2020-01-17 12:06:19 +01:00
Cathy Zhang	1ae7deb393	vm-virtio: Implement refactor for net devices and backend Since the common parts are put into net_util.rs under vm-virtio, refactoring code for virtio-net device, vhost-user-net device and backend to shrink the code size and improve readability meanwhile. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2020-01-17 12:06:19 +01:00
Rob Bradford	8b500d7873	deps: Bump vm-memory and linux-loader version The function GuestMemory::end_addr() has been renamed to last_addr() Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	7310ab6fa7	devices, vmm: Use a bit field for ACPI GED interrupt type Use independent bits for storing whether there is a CPU or memory device changed when reporting changes via ACPI GED interrupt. This prevents a later notification squashing an earlier one and ensure that hotplugging both CPU and memory at the same time succeeds. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	28c6652e57	vmm: Upon VmResize attempt to hotplug the memory If a new amount of RAM is requested in the VmResize command try and hotplug if it an increase (MemoryManager::Resize() silently ignores decreases.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	4e414f0d84	vmm: device_manager: Scan memory devices upon GED interrupt If there is a GED interrupt and the field indicates that the memory device has changed triggers a scan of the memory devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	284d5e011a	vmm: Add memory hotplug ACPI entries to DSDT Generate and expose the DSDT table entries required to support memory hotplug. The AML methods call into the MemoryManager via I/O ports exposed as fields. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	8ecf736982	vmm: device_manager: Add the MemoryManager to the I/O bus Now that the MemoryManager has I/O port functionality it needs to be exposed on the I/O bus. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	1218765df2	vmm: memory_manager: Expose the slots details via an I/O port Expose the details of hotplug RAM slots via an I/O port. This will be consumed by the ACPI DSDT tables to report the hotplug memory details to the guest. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	9880a2aba9	vmm: memory_manger: Add support for adding new memory to the VM Add a "resize()" method on MemoryManager which will create a new memory allocation based on the difference between the desired RAM amount and the amount already in use. After allocating the added RAM using the same backing method as the boot RAM store the details in a vector and update the KVM map and create a new GuestMemoryMmap and replace all the users. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	82fce5a4e2	vmm: Add support for resizing the memory used by the VM For now the new memory size is only used after a reboot but support for hotplugging memory will be added in a later commit. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	78dcb1862c	vmm: device_manager: Store the type of notification in a local value When the value is read from the I/O port via the ACPI AML functions to determine what has been triggered the notifiction value is reset preventing a second read from exposing the value. If we need support multiple types of GED notification (such as memory hotplug) then we should avoid reading the value multiple times. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	f5137e84bb	vmm, main: Add optional "hotplug_size" to --mem This specifies how much address space should be reserved for hotplugging of RAM. This space is reserved by adding move the start of the device area by the desired amount. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	f1b6657833	vmm: Make desired vCPUs optional in resize command In order to be able to support resizing either vCPUs or memory or both make the fields in the resize command optional. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	72b9e920a3	vmm: memory_manager: Further refactor memory region allocation This allows the memory regions to be allocated later which is necessary for hotplug memory. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Rob Bradford	1af11a7c92	vmm: memory_manager: Refactor GuestMemoryMmap construction Make the GuestMemoryMmap from a Vec<Arc<GuestRegionMmap>> by using this method we can persist a set of regions in the MemoryManager and then extend this set with a newly created region. Ultimately that will allow the hotplug of memory. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-15 20:21:22 +01:00
Samuel Ortiz	5788d36583	vmm: Do not create virtio devices when missing a transport If neither PCI or MMIO are built in, we should not bother creating any virtio devices at all. When building a minimal VMM made of a kernel with an initramfs and a serial console, the RNG virtio device is still created even though there is no way it can ever get probed. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-01-14 07:42:09 +01:00
Sebastien Boeuf	ae6f27277b	acpi: Introduce VIOT to support latest virtio-iommu implementation Because virtio-iommu is still evolving (as it's only partly upstream), some pieces like the ACPI declaration of the different nodes and devices attached to the virtual IOMMU are changing. This patch introduces a new ACPI table called VIOT, standing as the high level table overseeing the IORT table and associated subtables. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-01-08 09:27:07 +01:00
Rob Bradford	b2589d4f3f	vm-virtio, vmm, vfio: Store GuestMemoryMmap in an Arc<ArcSwap<T>> This allows us to change the memory map that is being used by the devices via an atomic swap (by replacing the map with another one). The ArcSwap provides the mechanism for atomically swapping from to another whilst still giving good read performace. It is inside an Arc so that we can use a single ArcSwap for all users. Not covered by this change is replacing the GuestMemoryMmap itself. This change also removes some vertical whitespace from use blocks in the files that this commit also changed. Vertical whitespace was being used inconsistently and broke rustfmt's behaviour of ordering the imports as it would only do it within the block. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-01-02 13:20:11 +00:00
Rob Bradford	a551398135	vmm: device_manager: Use MemoryManager to create KVM mapping Use the newly exported funtionality to reduce the amount of duplicated code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	962dec2913	vmm: memory_manager: Refactor KVM userspace mapping creation This function will be useful for other parts of the VMM that also estabilish their own mappings. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	7df88793a0	vmm: device_manager: Get device range from MemoryManager This removes the duplication of these values. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	61cfe3e72d	vmm: Obtain sequential KVM memory slot numbers from MemoryManager This removes the need to handle a mutable integer and also centralises the allocation of these slot numbers. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	260cebb8cf	vmm: Introduce MemoryManager The memory manager is responsible for setting up the guest memory and in the long term will also handle addition of guest memory. In this commit move code for creating the backing memory and populating the allocator into the new implementation trying to make as minimal changes to other code as possible. Follow on commits will further reduce some of the duplicated code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	d5682cd306	vmm: device_manager: Rewrite if chain using match To reflect updated clippy rules: error: `if` chain can be rewritten with `match` --> vmm/src/device_manager.rs:1508:25 \| 1508 \| / if ret > 0 { 1509 \| \| debug!("MSI message successfully delivered"); 1510 \| \| } else if ret == 0 { 1511 \| \| warn!("failed to deliver MSI message, blocked by guest"); 1512 \| \| } \| \|_________________________^ \| = note: `-D clippy::comparison-chain` implied by `-D warnings` = help: Consider rewriting the `if` chain to use `cmp` and `match`. = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#comparison_chain Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	21b88c3ea0	vmm: cpu: Rewrite if chain using match Address updated clippy error: error: `if` chain can be rewritten with `match` --> vmm/src/cpu.rs:668:9 \| 668 \| / if desired_vcpus > self.present_vcpus() { 669 \| \| self.activate_vcpus(desired_vcpus, None)?; 670 \| \| } else if desired_vcpus < self.present_vcpus() { 671 \| \| self.mark_vcpus_for_removal(desired_vcpus)?; 672 \| \| } \| \|_________^ \| = note: `-D clippy::comparison-chain` implied by `-D warnings` = help: Consider rewriting the `if` chain to use `cmp` and `match`. = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#comparison_chain Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	e25a47b32c	vmm: device_manager: Remove redundant clones Address updated clippy errors: error: redundant clone --> vmm/src/device_manager.rs:699:32 \| 699 \| .insert(acpi_device.clone(), 0x3c0, 0x4) \| ^^^^^^^^ help: remove this \| = note: `-D clippy::redundant-clone` implied by `-D warnings` note: this value is dropped without further use --> vmm/src/device_manager.rs:699:21 \| 699 \| .insert(acpi_device.clone(), 0x3c0, 0x4) \| ^^^^^^^^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone error: redundant clone --> vmm/src/device_manager.rs:737:26 \| 737 \| .insert(i8042.clone(), 0x61, 0x4) \| ^^^^^^^^ help: remove this \| note: this value is dropped without further use --> vmm/src/device_manager.rs:737:21 \| 737 \| .insert(i8042.clone(), 0x61, 0x4) \| ^^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone error: redundant clone --> vmm/src/device_manager.rs:754:29 \| 754 \| .insert(cmos.clone(), 0x70, 0x2) \| ^^^^^^^^ help: remove this \| note: this value is dropped without further use --> vmm/src/device_manager.rs:754:25 \| 754 \| .insert(cmos.clone(), 0x70, 0x2) \| ^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	a6878accd5	vmm: cpu: Implement CPU removal When the running OS has been told that a CPU should be removed it will shutdown the CPU and then signal to the hypervisor via the "_EJ0" method on the device that ultimately writes into an I/O port than the vCPU should be shutdown. Upon notification the hypervisor signals to the individual thread that it should shutdown and waits for that thread to end. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	7b3fc72aea	vmm: cpu: Notify guest OS that it should offline vCPUs Allow the resizing of the number of vCPUs to less than the current active vCPUs. This does not currently remove them from the system but the kernel will take them offline. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	7e81b0ded7	vmm: cpu: Create vCPU state for all possible vCPUs This will make it more straightforward when we attempt to remove vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	156ea392a2	vmm: cpu: Only do ACPI notify on newly added vCPUs When we add a vCPU set an "inserting" boolean that is exposed as an ACPI field that will be checked for and reset when the ACPI GED notification for CPU devices happens. This change is a precursor for CPU unplug. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-16 23:57:14 +01:00
Rob Bradford	e8313e3e69	vmm: acpi: Refactor ACPI CPU notification Continue to notify on all vCPUs but instead separate the notification functionality into two methods, CSCN that walks through all the CPUs and CTFY which notifies based on the numerical CPU id. This is an interim step towards only notifying on changed CPUs and ultimately CPU removal. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-16 23:57:14 +01:00
Sebastien Boeuf	d1390906c8	vmm: config: Derive Debug and PartialEq for configuration structures In anticipation for the writing of unit tests comparing two VmConfig structures, this commit derives the PartialEq trait for VmConfig and all embedded structures. This patch also derives the Debug trait for the same set of structures so that we can print them to facilitate debugging. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Sebastien Boeuf	93f5f6ed45	vmm: config: Provide a default empty command line through OpenAPI The OpenAPI should not have to provide a command line since the CLI considers the command line as an empty string if nothing is provided. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Sebastien Boeuf	43bd0e53c4	main: Move VmParams creation into a dedicated function This brings more modularity to the code, which will be helpful when we will later test the CLI and OpenAPI generate the same VmConfig output. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Samuel Ortiz	f0b7412495	vmm: device_manager: Add all virtio devices to the migratable list We want to track all migratable devices through the DeviceManager. Fixes: #341 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	37557c8b35	vmm: vm: Implement the Pausable trait Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	9756fc2dd0	vmm: cpu_manager: Implement the Pausable trait Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	35dd1523c9	vmm: device_manager: Implement the Pausable trait Since the Snapshotable placeholder and Migratable traits are provided as well, the DeviceManager object and all its objects are now Migratable. All Migratable devices are tracked as Arc<Mutex<dyn Migratable>> references. Keeping track of all migratable devices allows for implementing the Migratable trait for the DeviceManager structure, making the whole device model potentially migratable. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	35d7721683	vmm: Convert virtio devices to Arc<Mutex<T>> Migratable devices can be virtio or legacy devices. In any case, they can potentially be tracked through one of the IO bus as an Arc<Mutex<dyn BusDevice>>. In order for the DeviceManager to also keep track of such devices as Migratable trait objects, they must be shared as mutable atomic references, i.e. Arc<Mutex<T>>. That forces all Migratable objects to be tracked as Arc<Mutex<dyn Migratable>>. Virtio devices are typically migratable, and thus for them to be referenced by the DeviceManager, they now should be built as Arc<Mutex<VirtioDevice>>. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Sebastien Boeuf	64c5e3d8cb	vmm: api: Adjust FsConfig for OpenAPI The FsConfig structure has been recently adjusted so that the default value matches between OpenAPI and CLI. Unfortunately, with the current description, there is no way from the OpenAPI to describe a cache_size value "None", so that DAX does not get enabled. Usually, using a Rust "Option" works because the default value is None. But in this case, the default value is Some(8G), which means we cannot describe a None. This commit tackles the problem, introducing an explicit parameter "dax", and leaving "cache_size" as a simple u64 integer. This way, the default value is dax=true and cache_size=8G, but it lets the opportunity to disable DAX entirely with dax=false, which will simply ignore the cache_size value. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	4bfd51cc42	vmm: api: Match VhostUserBlkConfig defaults between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the VhostUserBlkConfig structure, this patch defines some default values for num_queues, queue_size and wce. num_queues is 1, queue_size is 128 and wce is true. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	1c2587f8cb	vmm: api: Match VhostUserNetConfig defaults between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the VhostUserNetConfig structure, this patch defines some default values for num_queues, queue_size and mac. num_queues is 2 since that's a pair of TX/RX queues, queue_size is 256 and mac is a randomly generated value. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	5e0bbf9c3b	vmm: Don't factorize vhost-user configurations We want to set different default configurations for vhost-user-net and vhost-user-blk, which is the reason why the common part corresponding to the number of queues and the queue size cannot be embedded. This prepares for the following commit, matching API and CLI behaviors. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	793327cff8	vmm: api: Make ConsoleConfig default match between CLI and HTTP API A simple patch making sure the field "file" is provisioned with the same default value through CLI and OpenAPI. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	cc08c44cb9	vmm: api: Make MemoryConfig default match between CLI and HTTP API Just making sure we have a serde default for the field "file" since it is not a required field in the OpenAPI definition. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	5a72225856	vmm: api: Update CpuConfig name to match the internal name All structures match between the OpenAPI definition and the internal configuration code, that's why CpuConfig is being renamed into CpusConfig. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Rob Bradford	c61104df47	vmm: Port to latest vmm-sys-util The signal handling for vCPU signals has changed in the latest release so switch to the new API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-11 14:11:11 +00:00
Sebastien Boeuf	ee528ae808	vmm: api: Make FsConfig defaults match between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the FsConfig structure, this patch defines some default values for num_queues, queue_size and the cache_size. num_queues is set to 1, queue_size is set to 1024, and cache_size is set to Some(8G) which means that DAX is enabled by default with a shared region of 8GiB. Fixes #508 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-09 23:42:23 -08:00
Sebastien Boeuf	befd342da4	vmm: api: Make NetConfig defaults match between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the NetConfig structure, this patch defines some default values for tap, ip, mask, mac and iommu. tap is None, ip is 192.168.249.1, mask is 255.255.255.0, mac is a randomly generated value, and iommu is false. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-09 23:19:24 -08:00
Jose Carlos Venegas Munoz	99e608c240	openapi: Fix schema Fix openapi schema to be a valid yaml. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-12-09 14:30:15 -08:00
Rob Bradford	f994665610	vmm: Reduce the minimum IRQ constant Now that the GED device does not use a hardcoded IRQ number the starting IRQ number can be restored (needed for the hardcoded serial port IRQ.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-09 16:58:00 +00:00
Rob Bradford	ba59c62044	vmm, devices: Remove hardcoded IRQ number for GED device Remove the previously hardcoded IRQ number used for the GED device. Instead allocate the IRQ using the allocator and use that value in the definition in the ACPI device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-09 16:58:00 +00:00
Sebastien Boeuf	aa94e9b8f3	Revert "vmm: api: Modify FsConfig to be OpenAPI friendly" This reverts commit `defc5dcd9c`.	2019-12-06 18:08:10 +00:00
Rob Bradford	9b1ba14f2d	vmm: Delegate device related ACPI DSDT table work to DeviceManager Move the code for handling the creation of the DSDT entries for devices into the DeviceManager. This will make it easier to handle device hotplug and also in the future remove some hardcoded ACPI constants. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 17:44:00 +00:00
Rob Bradford	60e6609011	vmm: Delegate CPU related ACPI tables to CpuManager Move the code for generating the MADT (APIC) table and the DSDT generation for CPU related functionality into the CpuManager. There is no functional change just code rearrangement. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 17:44:00 +00:00
Sebastien Boeuf	defc5dcd9c	vmm: api: Modify FsConfig to be OpenAPI friendly When consumer of the HTTP API try to interact with cloud-hypervisor, they have to provide the equivalent of the config structure related to each component they need. Problem is, the Rust enum type "Option" cannot be obtained from the OpenAPI YAML definition. This patch intends to fix this inconsistency between what is possible through the CLI and what's possible through the HTTP API by using simple types bool and int64 instead of Option<u64>. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-06 06:38:48 -08:00
Rob Bradford	59d01712ad	vmm: Remove kernel based IOAPIC handling from the device manager Previously the device setup code assumed that if no IOAPIC was passed in then the device should be added to the kernel irqchip. As an earlier change meant that there was always a userspace IOAPIC this kernel based code can be removed. The accessor still returns an Option type to leave scope for implementing a situation without an IOAPIC (no serial or GED device). This change does not add support no-IOAPIC mode as the original code did not either. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00
Rob Bradford	afea6a10a2	vmm: Stop initialising kernel based IOAPIC/PIC Now that we require the modern capabilities we can stop creating a kernel base irqchip. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00
Rob Bradford	9b1cb9621f	vmm: Remove pin based interrupt setup for virtio devices With MSI now required remove pin based interrupt support from all the virtio PCI device setup. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00
Rob Bradford	72fb687e3f	vmm: Check for required capabilities We now require CAP_SIGNAL_MSI, CAP_TSC_DEADLINE_TIMER and CAP_SPLIT_IRQCHIP. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00
Rob Bradford	f98b16f308	vmm: Update the configuration to preserve hot-plug CPUs after reboot Update the configuration after a resize to ensure that after a reboot the added vCPUs are preserved. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-05 16:39:19 +00:00
Rob Bradford	1722708612	vmm: Switch to storing VmConfig inside an Arc<Mutex<>> This permits the runtime reconfiguration of the VM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-05 16:39:19 +00:00
Rob Bradford	c063bb8d30	vmm: acpi: Make GED interrupt edge triggered This was causing issues when the kernel was trying to reset the interrupt and making the reboot fail. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-05 16:39:19 +00:00
Qiu Wenbo	e1af17d93a	vmm: Restore tty to canonical mode when SIGTERM or SIGINT received The tty mode remains raw mode when cloud-hypervisor is terminted by SIGTERM or SIGINT. The terminal is unusable due to echoing is disabled which is really annoying. Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2019-12-05 01:29:26 -08:00
Qiu Wenbo	5208ff86c8	vmm: Detect and handle AMD SME (Secure Memory Encryption) Some physical address bits may become reserved in page table when SME is enabled on AMD platform. Guest will trigger a reserved bit violation page fault in this case due to write these reserved bits to 1 in page table. We need reduce the reserved bits to get the right physical address range. Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2019-12-04 14:46:44 +00:00
Sebastien Boeuf	08258d5dad	vfio: pci: Allow multiple devices to be passed through The KVM_SET_GSI_ROUTING ioctl is very simple, it overrides the previous routes configuration with the new ones being applied. This means the caller, in this case cloud-hypervisor, needs to maintain the list of all interrupts which needs to be active at all times. This allows to correctly support multiple devices to be passed through the VM and being functional at the same time. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-04 08:48:17 +01:00
Rob Bradford	17badfbff5	vmm: cpu: Call vcpu configure() on the vCPU thread The function that programs the vCPUs is expected to be run from within each vCPU thread. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-03 03:22:15 -08:00
Rob Bradford	13503061e6	api: Fix OpenAPI specification entries Some renames from "cpu_count" were missing. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-03 03:28:06 +01:00
Rob Bradford	66a31c19e8	vmm: acpi: Upon GED interrupt notify on all vCPUs Call the "CTFY" method that will itself call Notify() on the CPU objects in the ACPI namespace. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	48bf141364	vmm: Trigger a hotplug device notification when resizing When adjusting the number of vCPUs generate a hotplug notification. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	b629727901	vmm: acpi: Add a CTFY method to notify on all CPU objects This method calls Notify() on all the vCPU objects in the ACPI namespace. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	ae9359c859	vmm: acpi: Create the CPU entries in the DSDT for all vCPUs CPU entries need to be created for every potential vCPU in the system. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	791ca3388f	vmm: device_manager: Add ability to notify via GED device Add ability to notify via the GED device that there is some new hotplug activity. This will be used by the CpuManager (and later DeviceManager itself) to notify of new hotplug activity. Currently it has a hardcoded IRQ of 5 as the ACPI tables also need to refer to this IRQ and the IRQ allocation does not permit the allocation of specific IRQs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	7ad68d499a	vmm: device_manager: Allocate I/O port for ACPI shutdown device The refactoring in `ce1765c8af` dropped the code to allocate the I/O port. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	86339b4cb4	vmm: Add HTTP API to resize the VM Currently only increasing the number of vCPUs is supported but in the future it will be extended. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	e7d4eae527	vmm: cpu: Add support for starting more vCPU threads Add support for starting vCPU threads after the initial boot ones. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	0ef999978c	vmm: cpu: Support only partially configuring the vCPU When configuring a processor after boot as a hotplug CPU we only configure a subset of the CPU state. In particular we should not configure the FPU, segment registers (or reconfigure the paging which is a side-effect of that) nor the main registers. Achieve this by making the function take an Option type for the start address. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	c8b3041e62	vmm: openapi: Update OpenAPI for CpuConfig struct This struct has changed in order to support differentiating between boot and max vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	b6801e355e	vmm: cpu: Refactor vCPU thread starting Refactor the vCPU thread starting so that there is the possibility to bring on extra vCPU threads. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	66d5163ee7	vmm: cpu: Encapsulate vCPU state into its own struct Currently this just holds the thread handle but will be enlarged to encompass details such as whether the vCPU is currently being inserted or ejected. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	1bbe48b24c	vmm: acpi: Mark non-boot vCPUs as disabled in the MADT table The MADT table contains the details of all the potential vCPUs and whether they are present at boot (as indicated by the flags field.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	82bc07cce4	vmm: Add boot and max vCPU handling to command line parser Also retain support (with a warning for the old behaviour.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	7543e00a07	vmm: Use new CpuManager accessor to get boot vCPUs When initialising the ACPI tables and configuring the VM use the new accessor on the CpuManager to get the number of boot vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	df0907845a	vmm: cpu: Introduce concept of maximum vs boot vCPUs in CpuManager For now the max vCPUs is the same as the boot vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Samuel Ortiz	0f21781fbe	cargo: Bump the kvm and vmm-sys-util crates Since the kvm crates now depend on vmm-sys-util, the bump must be atomic. The kvm-bindings and ioctls 0.2.0 and 0.4.0 crates come with a few API changes, one of them being the use of a kvm_ioctls specific error type. Porting our code to that type makes for a fairly large diff stat. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-29 17:48:02 +00:00
Jose Carlos Venegas Munoz	ab16af2941	openapi: make context ID vsock int64 context ID on vsock man defines a 32-bits value, openapi default integer is a signed 32-bits value. This could lead to miss one bit during castings for typed client implmentations. Lets increase the range of valid values by requesting an int64. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-11-26 08:38:59 +01:00
Sebastien Boeuf	f979380620	vmm: Mark guest persistent memory pages as mergeable In case the VM is started with the flag "--pmem mergeable=on", it means the user expects the guest persistent memory pages to be marked as mergeable. This commit relies on the madvise(MADV_MERGEABLE) system call to inform the host kernel about these pages. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-11-22 15:28:10 +00:00
Sebastien Boeuf	0f9afc3017	vmm: Add mergeable=on\|off option to --pmem flag In order to let the user indicate if the persistent memory pages should be marked as mergeable or not, a new option is being introduced. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-11-22 15:28:10 +00:00
Sebastien Boeuf	e4e8062dda	vmm: Mark guest RAM pages as mergeable In case the VM is started with the flag "--memory mergeable=on", it means the user expects the guest RAM pages to be marked as mergeable. This commit relies on the madvise(MADV_MERGEABLE) system call to inform the host kernel about these pages. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-11-22 15:28:10 +00:00
Sebastien Boeuf	880f62bab8	vmm: Add mergeable=on\|off option to --memory flag In order to let the user indicate if the guest RAM pages should be marked as mergeable or not, a new option is being introduced. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-11-22 15:28:10 +00:00
Jose Carlos Venegas Munoz	1d852e9ce5	vmm: Provide vmm version to start_vmm_thread When vmm.ping give a response, we expect get the version from the VMM not the vmm create Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-11-21 15:04:11 -08:00
Jose Carlos Venegas Munoz	a518651402	http: api: implement vmm.ping vmm.ping will help to check if http API server is up and running. This also removes the vmm.info endpoint. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-11-21 15:04:11 -08:00
Rob Bradford	348a1bc30e	vmm: cpu: Allocate I/O port for the CPU manager The CPU manager uses an I/O port and to prevent potential clashes with assignment for PCI devices ensure that it is allocated by the allocator. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	07cdb37dda	vmm: cpu & acpi: Query CPU manager for CPU status Rather than hardcode the CPU status for all the CPUs instead query from the CPU manager via the I/O port that is is on via the ACPI tables. Each CPU device has a _STA method that calls into the CSTA method which reads and writes the I/O ports via the PRST field which exposes the I/O port through and OpRegion. As we only support boot CPUS report that all the CPUs are enabled for now. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	5faf8b756c	vmm: acpi: Add an _MAT for the CPU devices containing a LAPIC The Linux kernel expects all CPUs, whether they be enabled or disabled to have an _MAT entry containing the LAPIC details for this CPU with the enabled bit set to 1 (in the flags.) In the MADT table the same bit is used to determine if the CPU is present at boot vs available later. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	1da0ff395d	vmm: cpu: Add the CpuManager onto the IO bus This allows the kernel (via ACPI based controls) to query and control the CPU state. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	50c8335d3d	vmm: device_manager: Expose the SystemAllocator This allows other code to allocate I/O ports for use on the (already) exposed IO bus. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	1ac1231292	vmm: Encase CpuManager within an Arc<Mutex<>> This is necessary to be able to add the CpuManager onto the IO bus. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Samuel Ortiz	f0e618431d	vmm: device_manager: Use consistent naming when adding devices When adding devices to the guest, and populating the device model, we should prefix the routines with add_. When we're just creating the device objects but not yet adding them we use make_. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	a2ee681665	vmm: device_manager: Add an MMIO devices creation routine In order to reduce the DeviceManager's new() complexity, we can move the MMIO devices creation code into its own routine. Fixes: #441 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	79b8f8e477	vmm: device_manager: Add a PCI devices creation routine In order to reduce the DeviceManager's new() complexity, we can move the PCI devices creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	5087f633f6	vmm: device_manager: Add an IOAPIC creation routine In order to reduce the DeviceManager's new() complexity, we can move the ACPI device creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	ce1765c8af	vmm: device_manager: Add an ACPI device creation routine In order to reduce the DeviceManager's new() complexity, we can move the ACPI device creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	cfca2759fc	vmm: device_manager: Add a legacy devices creation routine In order to reduce the DeviceManager's new() complexity, we can move the legacy devices creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	4b469b98cf	vmm: device_manager: Add a console creation routine In order to reduce the DeviceManager's new() complexity, we can move the console creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	b930b3fb41	vmm: api: Specify which integers are 64 bit wide By default, client will assume 32-bits for OpenAPI interger types. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-12 08:39:05 -08:00
Samuel Ortiz	6af2f57644	vmm: api: Fix the vm.info response payload We are returning a state and a config. Fixes: #431 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-12 08:39:05 -08:00
Rob Bradford	6958ec4922	vmm: Move CPU management code to its own module Move CpuManager, Vcpu and related functionality to its own module (and file) inside the VMM crate Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-11 15:46:24 +00:00
Samuel Ortiz	3dde848c8f	vmm: api: Update our OpenAPI document In most cases we return a 204 (No Content) and not a 201. In those cases, we do not send any HTTP body back at all. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-10 14:51:55 +01:00
Samuel Ortiz	96aa2441ad	vmm: http: Convert to micro_http HttpServer The new micro_http package provides a built-in HttpServer wrapper for running a more robust HTTP server based on the package HTTP API. Switching to this implementation allows us to, among other things, handle HTTP requests that are larger than 1024 bytes. Fixes: #423 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-10 14:51:55 +01:00
Samuel Ortiz	f34ace7673	vmm: http_endpoint: Do not sent 200 status code when our body is empty Otherwise HTTP client will not close the connection and wait for a pending body. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-10 14:51:55 +01:00
Jose Carlos Venegas Munoz	ede262684d	API: HTTP: change response content type to JSON The HTTP API responses are encoded in json Suggested-by: Samuel Ortiz <sameo@linux.intel.com> Tested-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com> Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-11-08 22:49:08 +01:00
Rob Bradford	3c715daa9d	vmm: Fix rustfmt failure by removing extra ";" Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-08 20:43:52 +00:00
Rob Bradford	a1a5fe0c93	vmm: Split CPU management into it's own struct Pull details of vCPU management (booting, pausing, resuming, shutdown) into it's own structure. This will ultimately enable this to be moved to its own file and encapsulate all the vCPU handling for the VMM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-08 11:59:21 +01:00
Rob Bradford	0319a4a09a	arch: vmm: Move ACPI tables creation to vmm crate Remove ACPI table creation from arch crate to the vmm crate simplifying arch::configure_system() GuestAddress(0) is used to mean no RSDP table rather than adding complexity with a conditional argument or an Option type as it will evaluate to a zero value which would be the default anyway. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-07 14:02:27 +00:00
Cathy Zhang	5cd4f5daeb	vmm: Release the old vm before build a new one In vm_reboot, while build the new vm, the old one pointed by self.vm is not released, that is, the tap opened by self.vm is not closed either. As a result, the associated dev name slot in host kernel is still in use state, which prevents the new build from picking it up as the new opened tap's name, but to use the name in next slot finally. Call self.vm_shutdown instead here since it has call take() on vm reference, which could ensure the old vm is destructed before the new vm build. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2019-11-05 14:40:43 +01:00
Rob Bradford	b3388c343d	vmm: device_manager: Ensure I/O ports are allocated Ensure that we tell the allocator about all the I/O ports that we are using for I/O bus attached devices (serial, i8042, ACPI device.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-05 10:13:01 +00:00
Sebastien Boeuf	5694ac2b1e	vm-virtio: Create new VirtioTransport trait to abstract ioeventfds In order to group together some functions that can be shared across virtio transport layers, this commit introduces a new trait called VirtioTransport. The first function of this trait being ioeventfds() as it is needed from both virtio-mmio and virtio-pci devices, represented by MmioDevice and VirtioPciDevice structures respectively. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	3fa5df4161	vmm: Unregister old ioeventfds when reprogramming PCI BAR Now that kvm-ioctls has been updated, the function unregister_ioevent() can be used to remove eventfd previously associated with some specific PIO or MMIO guest address. Particularly, it is useful for the PCI BAR reprogramming case, as we want to ensure the eventfd will only get triggered by the new BAR address, and not the old one. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	587a420429	cargo: Update to the latest kvm-ioctls version We need to rely on the latest kvm-ioctls version to benefit from the recent addition of unregister_ioevent(), allowing us to detach a previously registered eventfd to a PIO or MMIO guest address. Because of this update, we had to modify the current constraint we had on the vmm-sys-util crate, using ">= 0.1.1" instead of being strictly tied to "0.2.0". Once the dependency conflict resolved, this commit took care of fixing build issues caused by recent modification of kvm-ioctls relying on EventFd reference instead of RawFd. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	c7cabc88b4	vmm: Conditionally update ioeventfds for virtio PCI device The specific part of PCI BAR reprogramming that happens for a virtio PCI device is the update of the ioeventfds addresses KVM should listen to. This should not be triggered for every BAR reprogramming associated with the virtio device since a virtio PCI device might have multiple BARs. The update of the ioeventfds addresses should only happen when the BAR related to those addresses is being moved. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	de21c9ba4f	pci: Remove ioeventfds() from PciDevice trait The PciDevice trait is supposed to describe only functions related to PCI. The specific method ioeventfds() has nothing to do with PCI, but instead would be more specific to virtio transport devices. This commit removes the ioeventfds() method from the PciDevice trait, adding some convenient helper as_any() to retrieve the Any trait from the structure behing the PciDevice trait. This is the only way to keep calling into ioeventfds() function from VirtioPciDevice, so that we can still properly reprogram the PCI BAR. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	d6c68e4738	pci: Add error propagation to PCI BAR reprogramming Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	3e819ac797	pci: Use a weak reference to the AddressManager Storing a strong reference to the AddressManager behind the DeviceRelocation trait results in a cyclic reference count. Use a weak reference to break that dependency. Signed-off-by: Rob Bradford <robert.bradford@intel.com> Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	149b61b213	pci: Detect BAR reprogramming Based on the value being written to the BAR, the implementation can now detect if the BAR is being moved to another address. If that is the case, it invokes move_bar() function from the DeviceRelocation trait. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	04a449d3f3	pci: Pass DeviceRelocation to PciBus In order to trigger the PCI BAR reprogramming from PciConfigIo and PciConfigMmmio, we need the PciBus to have a hold onto the trait implementation of DeviceRelocation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	e93467a96c	vmm: Implement DeviceRelocation trait By implementing the DeviceRelocation trait for the AddressManager structure, we now have a way to let the PCI BAR reprogramming happen. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	8746c16593	vmm: Create AddressManager to own SystemAllocator In order to reuse the SystemAllocator later at runtime, it is moved into the new structure AddressManager. The goal is to have a hold onto the SystemAllocator and both IO and MMIO buses so that we can use them later. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	1870eb4295	devices: Lock the BtreeMap inside to avoid deadlocks Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Jose Carlos Venegas Munoz	78e2f7a99a	api: http: handle cpu according to openapi openapi definition defines an object for cpus not an integer Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-10-17 07:39:56 +02:00
Jose Carlos Venegas Munoz	205b8c1cd5	api: http: make consistent api and implementation vsocks: vsocks is implemented as an array Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-10-17 07:39:56 +02:00
Sebastien Boeuf	3acf9dfcf3	vfio: Don't map guest memory for VFIO devices attached to vIOMMU In case a VFIO devices is being attached behind a virtual IOMMU, we should not automatically map the entire guest memory for the specific device. A VFIO device attached to the virtual IOMMU will be driven with IOVAs, hence we should simply wait for the requests coming from the virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	63c30a6e79	vmm: Build and set the list of external mappings for VFIO When VFIO devices are created and if the device is attached to the virtual IOMMU, the ExternalDmaMapping trait implementation is created and associated with the device. The idea is to build a hash map of device IDs with their associated trait implementation. This hash map is provided to the virtual IOMMU device so that it knows how to properly trigger external mappings associated with VFIO devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	837bcbc6ba	vfio: Create VFIO implementation of ExternalDmaMapping With this implementation of the trait ExternalDmaMapping, we now have the tool to provide to the virtual IOMMU to trigger the map/unmap on behalf of the guest. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	3598e603d5	vfio: Add a public function to retrive VFIO container The VFIO container is the object needed to update the VFIO mapping associated with a VFIO device. This patch allows the device manager to have access to the VFIO container. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	9085a39c7d	vmm: Attach VFIO devices to IORT table This patch attaches VFIO devices to the virtual IOMMU if they are identified as they should be, based on the option "iommu=on". This simply takes care of adding the PCI device ID to the ACPI IORT table. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	5fc3f37c9b	vmm: Add iommu=on\|off option for --device Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a VFIO device should be attached to the virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Jose Carlos Venegas Munoz	786e33931f	api: http: Fix openpi schema. Fix invalid type for version: - VmInfo.version.type string Change Null value from enum as it has problems to build clients with openapi tools. - ConsoleConfig.mode.enum Null -> Nil Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-10-15 07:16:24 +02:00
Samuel Ortiz	2a0ba7aef8	vmm: vm: Add state validation unit test Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-14 06:35:36 +02:00
Samuel Ortiz	097b30669f	vmm: vm: Verify that state transitions are valid We should return an explicit error when the transition from on VM state to another is invalid. The valid_transition() routine for the VmState enum essentially describes the VM state machine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-14 06:35:36 +02:00
Samuel Ortiz	d2d3abb13c	vmm: Rename Booted vm state to Running Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-10 17:13:44 -07:00
Samuel Ortiz	dbbd04a4cf	vmm: Implement VM resume To resume a VM, we unpark all its vCPU threads. Fixes: #333 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-10 17:13:44 -07:00
Samuel Ortiz	4ac0cb9cff	vmm: Implement VM pause In order to pause a VM, we signal all the vCPU threads to get them out of vmx non-root. Once out, the vCPU thread will check for a an atomic pause boolean. If it's set to true, then the thread will park until being resumed. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-10 17:13:44 -07:00
Samuel Ortiz	1298b508bf	vmm: Manage the exit and reset behaviours from the control loop So that we don't need to forward an ExitBehaviour up to the VMM thread. This simplifies the control loop and the VMM thread even further. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-08 18:03:27 -07:00
Samuel Ortiz	a95fa1c4e8	vmm: api: Add a VMM shutdown command This shuts the current VM down, if any, and then exits the VMM process. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-08 18:03:27 -07:00
Samuel Ortiz	228adebc32	vmm: Unreference the VM when shutting down This way, we are forced to re-create the VM object when moving from shutdown to boot. Fixes: #321 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-08 17:24:05 +02:00
Sebastien Boeuf	b918220b49	vmm: Support virtio-pci devices attached to a virtual IOMMU This commit is the glue between the virtio-pci devices attached to the vIOMMU, and the IORT ACPI table exposing them to the guest as sitting behind this vIOMMU. An important thing is the trait implementation provided to the virtio vrings for each device attached to the vIOMMU, as they need to perform proper address translation before they can access the buffers. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	278ab05cbc	vmm: Add iommu=on\|off option for --vsock Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-vsock device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	32d07e40cc	vmm: Add iommu=on\|off option for --console Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-console device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	63869bde75	vmm: Add iommu=on\|off option for --pmem Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-pmem device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	fb4769388b	vmm: Add iommu=on\|off option for --rng Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-rng device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	20c4ed829a	vmm: Add iommu=on\|off option for --net Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-net device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	4b8d7e718d	vmm: Add iommu=on\|off option for --disk Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-blk device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". One side effect of this new option is that we had to introduce a new option for the disk path, simply called "path=". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	6e0aa56f06	vmm: Add iommu field to the VmConfig Adding a simple iommu boolean field to the VmConfig structure so that we can later use it to create a virtio-iommu device for the current VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	03352f45f9	arch: Create ACPI IORT table The virtual IOMMU exposed through virtio-iommu device has a dependency on ACPI. It needs to expose the device ID of the virtio-iommu device, and all the other devices attached to this virtual IOMMU. The IDs are expressed from a PCI bus perspective, based on segment, bus, device and function. The guest relies on the topology description provided by the IORT table to attach devices to the virtio-iommu device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	0acb1e329d	vm-virtio: Translate addresses for devices attached to IOMMU In case some virtio devices are attached to the virtual IOMMU, their vring addresses need to be translated from IOVA into GPA. Otherwise it makes no sense to try to access them, and they would cause out of range errors. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	6566c739e1	vm-virtio: Add IOMMU support to virtio-vsock Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	9ab00dcb75	vm-virtio: Add IOMMU support to virtio-rng Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	ee1899c6f6	vm-virtio: Add IOMMU support to virtio-pmem Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	392f1ec155	vm-virtio: Add IOMMU support to virtio-console Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	9fad680db1	vm-virtio: Add IOMMU support to virtio-net Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	9ebb1a55bc	vm-virtio: Add IOMMU support to virtio-blk Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	8225d4cd6e	vm-virtio: Implement reset() for virtio-console The virtio specification defines a device can be reset, which was not supported by this virtio-console implementation. The reason it is needed is to support unbinding this device from the guest driver, and rebind it to vfio-pci driver. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Samuel Ortiz	2a466132a0	vmm: api: Set the HTTP response header Server field To "Cloud Hypervisor API" and not "Firecracker API". Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	7abbad0a62	vmm: Be more idiomatic when calling into the VMM API Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	7328ecdb3b	vmm: Implement the /api/v1/vm.delete endpoint Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	f9daf2e247	vmm: Factorize the vm boot and shutdown code So that the API handling state machine is cleaner and easier to read. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	43b3642955	vmm: Clean Error handling up We used to have errors definitions spread across vmm, vm, api, and http. We now have a cleaner separation: All API routines only return an ApiResult. All VM operations, including the VMM wrappers, return a VmResult. This makes it easier to carry errors up to the HTTP caller. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	42758244a0	vmm: Implement the /api/v1/vm.info endpoint This, for now, returns the VM config and its state. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	27af983ec9	vmm: Track the VM state We will expose it through the api/v1/vm.info endpoint. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	b70344158b	vmm: Handle the missing VM error When trying to boot or shut a VM down, return an error if the VM was not previously created. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	7e0cb078ed	vmm: Only build a new VM when booting it In order to support further use cases where a VM configuration could be modified through the HTTP API, we only store the passed VM config when being asked to create a VM. The actual creation will happen when booting a new config for the first time. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	c505cfae2b	vmm: Implement the VM HTTP endpoint handlers Implement the vm.create, vm.boot, vm.shutdown and vm.reboot HTTP endpoint handlers. Fixes: #244 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	8a5e47f989	vmm: Implement the shutdown and reboot API We factorize some of the code for both the API helpers and the VMM thread. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	46cde1a38e	vmm: Rename the VM start and stop operations to boot and shutdown To match the OpenAPI description. And also to map the real life terminology. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	ce0b475ef7	vmm: Move the VM creation and startup helpers to the api module They're API wrappers, not VMM ones. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	f674019ea1	vmm: {De}serialize VmConfig We use the serde crate to serialize and deserialize the VmVConfig structure. This structure will be passed from the HTTP API caller as a JSON payload and we need to deserialize it into a VmConfig. For a convenient use of the HTTP API, we also provide Default traits implementations for some of the VmConfig fields (vCPUs, memory, etc...). Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	f2de4d0315	vmm: config: Make the cmdline config serializable The linux_loader crate Cmdline struct is not serializable. Instead of forcing the upstream create to carry a serde dependency, we simply use a String for the passed command line and build the actual CmdLine when we need it (in vm::new()). Also, the cmdline offset is not a configuration knob, so we remove it. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	6a722e5c0b	vmm: config: Make VhostUser configs serializable They point to a vm_virtio structure (VhostUserConfig) and in order to make the whole config serializable (through the serde crate for example), we'd have to add a serde dependency to the vm_virtio crate. Instead we use a local, serializable structure and convert it to VhostUserConfig from the DeviceManager code. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	aa31748781	vmm: Start the HTTP server thread Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	b14fd37db9	vmm: Make --kernel optional The kernel path was the only mandatory command line option. With the addition of the --api-socket option, we can run without a kernel path and get it later through the API. Since we can end up with VM configurations that are no longer valid by default, we need to provide a validation check for it. For now, if the kernel path is not defined, the VM configuration is invalid. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	2371325f9c	vmm: api: Add HTTP server The Cloud Hyper HTTP server runs a synchronous, multi-threaded loop that receives HTTP requests and tries to call the corresponding endpoint handlers for the requests URIs. An endpoint handler will parse the HTTP request and potentially translate it into and IPC request. The handler holds an notifier and an mspc Sender for respectively notifying and sending the IPC payload to the VMM API server. The handler then waits for an API server response and translate it back into an HTTP response. The HTTP server is responsible for sending the reponse back to the caller. The HTTP server uses a static routes hash table that maps URIs to endpoint handlers. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	8916dad2da	vmm: api: Add cloud-hypervisor OpenAPI documentation The cloud-hypervisor API uses HTTP as a transport and is accessible through a local UNIX socket. The API root path is /api/v1 and is a collection of RPC-style methods. All methods are static, unlike typical REST APIs. Variable (e.g. device IDs) are passed through the request body. Fixes: #244 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Rob Bradford	8ea4145f98	devices, vmm: Add legacy CMOS device Based off of crosvm revision b5237bbcf074eb30cf368a138c0835081e747d71 add a CMOS device. This environments that can't use KVM clock to get the current time (e.g. Windows and EFI.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-10-03 14:57:49 +01:00
Rob Bradford	833a3d456c	pci, vmm: Expose the PCI bus for configuration via MMIO Refactor the PCI datastructures to move the device ownership to a PciBus struct. This PciBus struct can then be used by both a PciConfigIo and PciConfigMmio in order to expose the configuration space via both IO port and also via MMIO for PCI MMCONFIG. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-30 18:00:31 +01:00
Rob Bradford	b5ee9212c1	vmm, devices: Use APIC address constant In order to avoid introducing a dependency on arch in the devices crate pass the constant in to the IOAPIC device creation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-27 11:48:30 -07:00
Rob Bradford	162791b571	vmm, arch: Use IOAPIC constants from layout in DeviceManager Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-27 11:48:30 -07:00
Rob Bradford	a0455167d0	vmm: Use layout constant for kernel command line Remove the unnecessary field on CmdlineConfig and switch to using the common offset. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-27 11:48:30 -07:00
Rob Bradford	0e7a1fc923	arch, vmm: Start documenting major regions of RAM and reserved memory Using the existing layout module start documenting the major regions of RAM and those areas that are reserved. Some of the constants have also been renamed to be more consistent and some functions that returned constant variables have been replaced. Future commits will move more constants into this file to make it the canonical source of information about the memory layout. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-27 08:55:47 -07:00
Samuel Ortiz	8188074300	main: Start the VMM thread We now start the main VMM thread, which will be listening for VM and IPC related events. In order to start the configured VM, we no longer directly call the VM API but we use the IPC instead, to first create and then start a VM. Fixes: #303 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	e235c6de4f	vmm: Add VM creation and startup helpers Based on the newly defined Cloud Hypervisor IPC, those helpers send VmCreate and VmStart requests respectively. This will be used by the main thread to create and start a VM based on the CLI parameters. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	151f96e454	vmm: Add a VMM thread startup routine This starts the main, single VMM thread, which: 1. Creates the VMM instance 2. Starts the VMM control loop 3. Manages the VMM control loop exits for handling resets and shutdowns. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	2f1ff23066	vmm: (Re-)Introduce a VMM structure Unlike the Vmm structure we removed with commit `bdfd1a3f`, this new one is really meant to represent the VM monitoring/management object. For that, we implement a control loop that will replace the one that's currently embedded within the Vm structure itself. This will allow us to decouple the VM lifecycle management from the VM object itself, by having a constantly running VMM control loop. Besides the VM specific events (exit, reset, stdin for now), the VMM control loop also handles all the Cloud Hypervisor IPC requests. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	4671a5831f	vmm: Move the EpollContext implementation to lib The VMM thread and control loop will be the sole consumer of the EpollContext and EpollDispatch API, so let's move it to lib.rs. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	03ab6839c1	vmm: Introduce Cloud Hypervisor IPC Cloud Hypervisor IPC is a simple, mpsc based protocol for threads to send command to the furture VMM thread. This patch adds the API definition for that IPC, which will be used by both the main thread to e.g. start a new VM based on the CLI arguments and the future HTTP server to relay external requests received from a local Unix domain socket. We are moving it to its own "api" module because this is where the external API (HTTP based) will also be implemented. The VMM thread will be listening for IPC requests from an mpsc receiver, process them and send a response back through another mpsc channel. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	6710a39b5a	vmm: Pass the exit and reset fds to the vm creation method As we're going to move the control loop to the VMM thread, the exit and reset EventFds are no longer going to be owned by the VM. We pass a copy of them when creating the Vm instead. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	feb1c33084	vmm: Add a VM config getter We will need it from the VMM thread, when trying to reboot a VM. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	47167a658e	vmm: Add a VM console handling method In order to handle the VM STDIN stream from a separate VMM thread without having to export the DeviceManager, we simply add a console handling method to the Vm structure. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	ea7abc6c80	vmm: Add a VM stop method In order to transfer the control loop to a separate VMM thread, we want to shrink the VM control loop to a bare minimum. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	e6ef9ece2c	vmm: Move the tty setting to the VM start routine We want to shrink the control loop to a bare minimal. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	2e9d815701	vmm: Use a reference counted VmConfig when creating a new VM Once passed to the VM creation routine, a VmConfig structure is immutable. We can simply carry a Arc of it instead of a reference. This also allows us to remove any lifetime bound from our VM. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	bdfd1a3f38	vmm: Remove the Vmm structure The Vmm structure is just a placeholder for the KVM instance. We can create it directly from the VM creation routine instead. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 10:12:04 +02:00
Samuel Ortiz	9c5135da7a	vmm: Simplify the VM start flow We can integrate the kernel loading into the VM start method. The VM start flow is then: Vm::new() -> vm.start(), which feels more natural. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 10:12:04 +02:00
Samuel Ortiz	b79c1f7722	vmm: Derive the clone trait for VmConfig Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	acc60b0ad5	vmm: Make VsockConfig owned Convert Path to PathBuf and remove the associated lifetime. Now we can remove the VmConfig associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	3dc7aff00e	vmm: Make vhost-user configuration owned Convert Path to PathBuf, &str to String and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	5f8a62f3d0	vmm: Make DeviceConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	36137232f0	vmm: Make ConsoleConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	79a02f9171	vmm: Make PmemConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	00674cd850	vmm: Make FsConfig owned Convert Path to PathBuf, &str to String and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	5323da031c	vmm: Make RngConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	0688bec298	vmm: Make NetConfig owned Convert str to String and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	675e46355c	vmm: Make DiskConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	036890e5be	vmm: Make KernelConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	9c5bfb8e13	vmm: Make MemoryConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Yang Zhong	4164853ec6	vmm: add vhost-user-blk support Update vm configuration and device initial process to add vhost-user-blk support. Signed-off-by: Yang Zhong <yang.zhong@intel.com>	2019-09-20 15:56:51 +02:00
Yang Zhong	c7559bb7a4	config: make error definition common Since vhost-user-blk use same error definition with vhost-user-net, those errors need define to common usage. Signed-off-by: Yang Zhong <yang.zhong@intel.com>	2019-09-20 15:56:51 +02:00
Rob Bradford	5b3ca78dac	vmm: Use the full host physical address range Probe for the size of the host physical address range and use that to establish the address range for the VM. This removes the limitation on the size of the VM RAM and gives more space for the devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-19 10:43:55 +01:00
Rob Bradford	f0360c92d9	arch: acpi: Set the upper device range based on RAM levels After the 32-bit gap the memory is shared between the devices and the RAM. Ensure that the ACPI tables correctly indicate where the RAM ends and the device area starts by patching the precompiled tables. We get the following valid output now from the PCI bus probing (8GiB guest) [ 0.317757] pci_bus 0000:00: resource 4 [io 0x0000-0x0cf7 window] [ 0.319035] pci_bus 0000:00: resource 5 [io 0x0d00-0xffff window] [ 0.320215] pci_bus 0000:00: resource 6 [mem 0x000a0000-0x000bffff window] [ 0.321431] pci_bus 0000:00: resource 7 [mem 0xc0000000-0xfebfffff window] [ 0.322613] pci_bus 0000:00: resource 8 [mem 0x240000000-0xfffffffff window] Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-19 10:43:55 +01:00
Rob Bradford	3bc11a4a2e	vmm: Make the "mmio" only build generate no errors Rerrange "use" statements and make rename variables and fields to indicate they might be unused. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-16 08:55:35 -07:00

... 9 10 11 12 13 ...

1173 Commits