cloud-hypervisor

mirror of https://github.com/cloud-hypervisor/cloud-hypervisor.git synced 2024-12-29 00:55:18 +00:00

Author	SHA1	Message	Date
Wei Liu	b00171e17d	vmm: use MemoryRegion where applicable That removes one more KVM-ism in VMM crate. Note that there are more KVM specific code in those files to be split out, but we're not at that stage yet. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-25 10:25:13 +02:00
Rob Bradford	d983c0a680	vmm: Expose counters from virtio devices to API Collate the virtio device counters in DeviceManager for each device that exposes any and expose it through the recently added HTTP API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-25 07:02:44 +02:00
Rob Bradford	bca8a19244	vmm: Implement HTTP API for obtaining counters The counters are a hash of device name to hash of counter name to u64 value. Currently the API is only implemented with a stub that returns an empty set of counters. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-25 07:02:44 +02:00
Rob Bradford	fd4aba8eae	vmm: api: Implement support for GET handlers EndpointHandler This can be used for simple API requests which return data but do not require any input. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-25 07:02:44 +02:00
Rob Bradford	80be393b16	vmm: api: Order HTTP entry points in alphabetical order Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-25 07:02:44 +02:00
Wei Liu	4cc37d7b9a	vmm: interrupt: drop a few pub keywords Those items are not used elsewhere. Restrict their scope. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-24 12:39:42 +02:00
Wei Liu	1661adbbaf	vmm: interrupt: add "Kvm" prefix to MsiInterruptGroup The structure is tightly coupled with KVM. It uses KVM specific structures and calls. Add Kvm prefix to it. Microsoft hypervisor will implement its own interrupt group(s) later. No functional change intended. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-24 12:39:42 +02:00
Sebastien Boeuf	9f4714c32a	vmm: Extend seccomp filters with KVM_KVMCLOCK_CTRL Now that the VMM uses KVM_KVMCLOCK_CTRL from the KVM API, it must be added to the seccomp filters list. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-24 12:38:56 +02:00
Sebastien Boeuf	4a81d65f79	vmm: Notify the guest about vCPUs being paused Through the newly added API notify_guest_clock_paused(), this patch improves the vCPU pause operation by letting the guest know that each vCPU is being paused. This is important to avoid soft lockups detection from the guest that could happen because the VM has been paused for more than 20 seconds. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-24 12:38:56 +02:00
Sebastien Boeuf	9fa8438063	vmm: Fill CpuManager's vCPU list on restore path It's important that on restore path, the CpuManager's vCPU gets filled with each new vCPU that is being created. In order to cover both boot and restore paths, the list is being filled from the common function create_vcpu(). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-24 12:38:56 +02:00
Sebastien Boeuf	f5150aa261	vmm: Extend seccomp filters with KVM_GET_CLOCK and KVM_SET_CLOCK Now that the VMM uses both KVM_GET_CLOCK and KVM_SET_CLOCK from the KVM API, they must be added to the seccomp filters list. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 14:36:01 +01:00
Sebastien Boeuf	8038161861	vmm: Get and set clock during pause and resume operations In order to maintain correct time when doing pause/resume and snapshot/restore operations, this patch stores the clock value on pause, and restore it on resume. Because snapshot/restore expects a VM to be paused before the snapshot and paused after the restore, this covers the migration use case too. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 14:36:01 +01:00
Rob Bradford	4b64f2a027	vmm: cpu: Reuse already allocated vCPUs if available When a request is made to increase the number of vCPUs in the VM attempt to reuse any previously removed (and hence inactive) vCPUs before creating new ones. This ensures that the APIC ID is not reused for a different KVM vCPU (which is not allowed) and that the APIC IDs are also sequential. The two key changes to support this are: * Clearing the "kill" bit on the old vCPU state so that it does not immediately exit upon thread recreation. * Using the length of the vcpus vector (the number of allocated vcpus) rather than the number of active vCPUs (.present_vcpus()) to determine how many should be created. This change also introduced some new info!() debugging on the vCPU creation/removal path to aid further development in the future. TEST=Expanded test_cpu_hotplug test. Fixes: #1338 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-23 14:11:14 +01:00
Rob Bradford	9dcd0c37f3	vmm: cpu: Clear the "kill" flag on vCPU to support reuse After the vCPU has been ejected and the thread shutdown it is useful to clear the "kill" flag so that if the vCPU is reused it does not immediately exit upon thread recreation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-23 14:11:14 +01:00
Rob Bradford	b107bfcf2c	vmm: cpu: Add info!() level debugging to vCPU handling These messages are intended to be useful to support debugging related to vCPU hotplug/unplug issues. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-23 14:11:14 +01:00
Sebastien Boeuf	e382dc6657	vmm, vm-virtio: Restore DeviceManager's devices in a paused state The same way the VM and the vCPUs are restored in a paused state, all devices associated with the device manager must be restored in the same paused state. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 10:15:03 +02:00
Sebastien Boeuf	8a165b5314	vmm: Restore the VM in "paused" state Because we need to pause the VM before it is snapshot, it should be restored in a paused state to keep the sequence symmetrical. That's the reason why the state machine regarding the valid VM's state transition needed to be updated accordingly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 10:15:03 +02:00
Sebastien Boeuf	a16414dc87	vmm: Restore vCPUs in "paused" state To follow a symmetrical model, and avoid potential race conditions, it's important to restore a previously snapshot VM in a "paused" state. The snapshot operation being valid only if the VM has been previously paused. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-23 10:15:03 +02:00
Wei Liu	7552f4db61	vmm: device_manager: restore error handling When the hypervisor crate was introduced, a few places that handled errors were commented out in favor of unwrap, but that's bad practice. Restore proper error handling in those places in this patch. We cannot use from_raw_os_error anymore because it is wrapped deep under hypervisor crate. Create new custom errors instead. Fixes: `e4dee57e81` ("arch, pci, vmm: Initial switch to the hypervisor crate") Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-22 22:02:21 +01:00
Muminul Islam	cca59bc52f	hypervisor, arch: Fix warnings introduced in hypervisor crate This commit fixes some warnings introduced in the previous hyperviosr crate PR.Removed some unused variables from arch/aarch64 module. Signed-off-by: Muminul Islam <muislam@microsoft.com>	2020-06-22 21:58:45 +01:00
Rob Bradford	d714efe6d4	vmm: cpu: Import CpuTopology conditionally on x86_64 only The aarch64 build has no use for this structure at the moment. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-22 15:00:27 +01:00
Sebastien Boeuf	a998e89375	build(deps): bump signal-hook from 0.1.15 to 0.1.16 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.1.15 to 0.1.16. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](vorner/signal-hook@v0.1.15...v0.1.16) Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-22 14:09:11 +01:00
Muminul Islam	e4dee57e81	arch, pci, vmm: Initial switch to the hypervisor crate Start moving the vmm, arch and pci crates to being hypervisor agnostic by using the hypervisor trait and abstractions. This is not a complete switch and there are still some remaining KVM dependencies. Signed-off-by: Muminul Islam <muislam@microsoft.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-06-22 15:03:15 +02:00
Rob Bradford	a74c6fc14f	vmm, arch: x86_64: Fill the CPUID leaves with the topology There are two CPUID leaves for handling CPU topology, 0xb and 0x1f. The difference between the two is that the 0x1f leaf (Extended Topology Leaf) supports exposing multiple die packages. Fixes: #1284 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-17 12:18:09 +02:00
Rob Bradford	e19079782d	vmm, arch: x86_64: Set the APIC ID on the 0x1f CPUID leaf The extended topology leaf (0x1f) also needs to have the APIC ID (which is the KVM cpu ID) set. This mirrors the APIC ID set on the 0xb topology leaf Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-17 12:18:09 +02:00
Rob Bradford	b81bc77390	vmm: cpu: Save CpusConfig into CpuManager Rather than saving the individual parts into the CpuManager save the full struct as it now also contains the topology data. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-17 12:18:09 +02:00
Rob Bradford	4a0439a993	vmm: config: Extend CpusConfig to add the topology This allows the user to optionally specify the desired CPU topology. All parts of the topology must be specified and the product of all parts must match the maximum vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-17 12:18:09 +02:00
Wei Liu	103cd61bd2	vmm: device_tree: make available remove function unconditionally Its test case calls remove unconditionally. Instead of making the test code call remove conditionally, removing the pci_support dependency simplifies things -- that function is just a wrapper around HashMap's remove function anyway. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-15 11:41:34 +02:00
Wei Liu	fb461c820f	vmm: vm: enable test_vm test case Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-12 14:46:58 +01:00
Wei Liu	b99b5777bb	vmm: vm: move some imports into test_vm They are only needed there. Not moving them causes rustc to complain about unused imports. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-12 14:46:58 +01:00
Sebastien Boeuf	b62d5d22ff	vmm: openapi: Update the OpenAPI definition Now that PCI device hotplug returns a response, the OpenAPI definition must reflect it, describing what is expected to be received. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	4fe7347fb9	vmm: Manually implement Serialize for PciDeviceInfo In order to provide a more comprehensive b/d/f to the user, the serialization of PciDeviceInfo is implemented manually to control the formatting. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	83cd9969df	vmm: Enable HTTP response for PCI device hotplug This patch completes the series by connecting the dots between the HTTP frontend and the device manager backend. Any request to hotplug a VFIO, disk, fs, pmem, net, or vsock device will now return a response including the device name and the place of the device in the PCI topology. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	3316348d4c	vmm: vm: Carry information from hotplugged PCI device Pass from the device manager to the calling code the information about the PCI device that has just been hotplugged. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	f08e9b6a73	vmm: device_manager: Return PciDeviceInfo from a hotplugged device In order to provide the device name and PCI b/d/f associated with a freshly hotplugged device, the hotplugging functions from the device manager return a new structure called PciDeviceInfo. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	0bc2b08d3a	vmm: api: Return an optional response from vm_action() Any action that relies on vm_action() can now return a response body. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Sebastien Boeuf	038180269e	vmm: api: Allow HTTP PUT request to return a response Adding the codepath to return a response from a PUT request. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-12 13:37:18 +01:00
Wei Liu	5ebd02a572	vmm: vm: fix test_vm test case We should break out from the loop after getting the HLT exit, otherwise the VM hangs forever. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2020-06-12 08:38:07 +02:00
Michael Zhao	97a1e5e1d2	vmm: Exit VMM event loop after guest shutdown for AArch64 X86 and AArch64 work in different ways to shutdown a VM. X86 exit VMM event loop through ACPI device; AArch64 need to exit from CPU loop of a SystemEvent. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	5cd1730bc4	vmm: Configure VM on AArch64 Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	917219fa92	vmm: Enable VCPU for AArch64 Added MPIDR which is needed in system configuration. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	b5f1c912d6	vmm: Enable memory manager for AArch64 Screened IO space as it is not available on AArch64. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	eeeb45bbb9	vmm: Enable device manager for AArch64 Screened IO bus because it is not for AArch64. Enabled Serial, RTC and Virtio devices with MMIO transport option. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Michael Zhao	e9488846f1	vm-allocator: Enable vm-allocator for AArch64 Implemented GSI allocator and system allocator for AArch64. Renamed some layout definitions to align more code between architectures. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-11 15:00:17 +01:00
Anatol Belski	abd6204d27	source: Fix file permissions Rust sources and some data files should not be executable. The perms are set to 644. Signed-off-by: Anatol Belski <ab@php.net>	2020-06-10 18:47:27 +01:00
Sebastien Boeuf	653087d7a3	vmm: Reduce MMIO address space by 4KiB In order to workaround a Linux bug that happens when we place devices at the end of the physical address space on recent hardware (52 bits limit) we reduce the MMIO address space by one 4k page. This way, nothing gets allocated in the last 4k of the address space, which is negligible given the amount of space in the address space. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-06-09 18:08:09 +01:00
Bo Chen	625bab69bd	vmm: api: Allow to delete non-booted VMs The action of "vm.delete" should not report errors on non-booted VMs. This patch also revised the "docs/api.md" to reflect the right 'Prerequisites' of different API actions, e.g. on "vm.delete" and "vm.boot". Fixes: #1110 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-06-09 05:58:32 +01:00
Rob Bradford	9b71ba20ac	vmm, vm-virtio: Stop always autogenerating a host MAC address This removes the need to use CAP_NET_ADMIN privileges and instead the host MAC addres is either provided by the user or alternatively it is retrieved from the kernel. TEST=Run cloud-hypervisor without CAP_NET_ADMIN permission and a preconfigured tap device: sudo ip tuntap add name tap0 mode tap sudo ifconfig tap0 192.168.249.1 netmask 255.255.255.0 up cargo clean cargo build target/debug/cloud-hypervisor --serial tty --console off --kernel ~/src/rust-hypervisor-firmware/target/target/release/hypervisor-fw --disk path=~/workloads/clear-33190-kvm.img --net tap=tap0 VM was also rebooted to check that works correctly. Fixes: #1274 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-08 17:56:10 +02:00
Rob Bradford	929d70bc7f	net_util: Only try and enable the TAP device if it not already enabled This allows an existing TAP interface to be used without needing CAP_NET_ADMIN permissions on the Cloud Hypervisor binary as the ioctl to bring up the interface is avoided. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-06-08 17:56:10 +02:00
Bo Chen	a8cdf2f070	tests,vm-virtio,vmm: Use 'socket' for all CLI/API parameters This patch unifies the inconsistent uses of 'socket' and 'sock' from our CLI/API parameters. Fixes: #1091 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-06-08 17:41:12 +02:00
Samuel Ortiz	3336e80192	vfio: Switch to the vfio-ioctls crate ch branch Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-06-04 08:48:55 +02:00
Samuel Ortiz	d24aa72d3e	vfio: Rename to vfio-ioctls Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-06-04 08:48:55 +02:00
Samuel Ortiz	53ce529875	vfio: Move the PCI implementation to the PCI crate There is a much stronger PCI dependency from vfio_pci.rs than a VFIO one from pci/src/vfio.rs. It seems more natural to have the PCI specific VFIO implementation in the PCI crate rather than the other way around. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-06-04 08:48:55 +02:00
Michael Zhao	8f7dc73562	vmm: Move Vcpu::configure() to arch crate Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-03 11:27:29 +02:00
Michael Zhao	969e5e0b51	vmm: Split configure_system() from load_kernel() for x86_64 Now the flow of both architectures are aligned to: 1. load kernel 2. create VCPU's 3. configure system 4. start VCPU's Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-03 11:27:29 +02:00
Michael Zhao	20cf21cd9d	vmm: Change booting process to cover AArch64 requirements Between X86 and AArch64, there is some difference in booting a VM: - X86_64 can setup IOAPIC before creating any VCPU. - AArch64 have to create VCPU's before creating GIC. The old process is: 1. load_kernel() load kernel binary configure system 2. activate_vcpus() create & start VCPU's So we need to separate "activate_vcpus" into "create_vcpus" and "activate_vcpus" (to start vcpus only). Setup GIC and create FDT between the 2 steps. The new procedure is: 1. load_kernel() load kernel binary (X86_64) configure system 2. create VCPU's 3. (AArch64) setup GIC 4. (AArch64) configure system 5. start VCPU's Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-06-03 11:27:29 +02:00
Rob Bradford	c31ad72ee9	build: Address issues found by 1.43.0 clippy These are mostly due to use of "bare use" statements and unnecessary vector creation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-27 19:32:12 +02:00
Bo Chen	fbd1a6c5f1	vmm: api: Return complete error responses in handle_http_request() Instead of responding only headers with error code, we now return complete error responses to HTTP requests with errors (e.g. undefined endpoints and InternalSeverError). Fixes: #472 Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-05-27 18:29:52 +01:00
Rob Bradford	0728bece0c	vmm: seccomp: Ensure that umask() can be reprogrammed When doing self spawning the child will attempt to set the umask() again. Let it through the seccomp rules so long as it the safe mask again. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-27 16:46:51 +01:00
Michael Zhao	8f1f9d9e6b	devices: Implement InterruptController on AArch64 This commit only implements the InterruptController crate on AArch64. The device specific part for GIC is to be added. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-05-26 11:09:19 +02:00
Michael Zhao	b32d3025f3	devices: Refactor IOAPIC to cover other architectures IOAPIC, a X86 specific interrupt controller, is referenced by device manager and CPU manager. To work with more architectures, a common type for all architectures is needed. This commit introduces trait InterruptController to provide architecture agnostic functions. Device manager and CPU manager can use it without caring what the underlying device is. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-05-26 11:09:19 +02:00
Michael Zhao	1befae872d	build: Fixed build errors and warnings on AArch64 This is a preparing commit to build and test CH on AArch64. All building issues were fixed, but no functionality was introduced. For X86, the logic of code was not changed at all. For ARM, the architecture specific part is still empty. And we applied some tricks to workaround lint warnings. But such code will be replaced later by other commits with real functionality. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2020-05-21 11:56:26 +01:00
Rob Bradford	af8292b623	vmm, config, vhost_user_blk: remove "wce" parameter This config option provided very little value and instead we now enable this feature (which then lets the guest control the cache mode) unconditionally. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-21 08:40:43 +02:00
Bo Chen	7c3e19c65a	vhost_user_backend, vmm: Close leaked file descriptors Explicit call to 'close()' is required on file descriptors allocated from 'epoll::create()', which is missing for the 'EpollContext' and 'VringWorker'. This patch enforces to close the file descriptors by reusing the Drop trait of the 'File' struct. Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-05-19 09:22:09 +02:00
Rob Bradford	1b8b5ac179	vhost-user_net, vm-virtio, vmm: Permit host MAC address setting Add a new "host_mac" parameter to "--net" and "--net-backend" and use this to set the MAC address on the tap interface. If no address is given one is randomly assigned and is stored in the config. Support for vhost-user-net self spawning was also included. Fixes: #1177 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-15 11:45:09 +01:00
Rob Bradford	11049401ce	vmm: seccomp: Add ioctl() commands interface hardware address This is necessary to support setting the MAC address on the tap interface on the host. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-15 11:45:09 +01:00
Sebastien Boeuf	68fc432978	vmm: Update seccomp filters with clock_nanosleep The clock_nanosleep system call needs to be whitelisted since the commit `12e00c0f45` introduced the use of a sleep() function. Without this patch, we can see an error when the VM is paused or killed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-15 12:34:53 +02:00
Rob Bradford	6aa29bdb24	vmm: api: Use a common handler for data actions too Like the actions that don't take data such as "pause" or "resume" use a common handler implementation to remove duplicated code for handling simple endpoints like the hotplug ones. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	0fe223f00e	vmm: api: Extend VmAction to reduce code duplication Many of the API requests take a similar form with a single data item (i.e. config for a device hotplug) expand the VmAction enum to handle those actions and a single function to dispatch those API events. For now port the existing helper functions to use this new API. In the future the HTTP layer can create the VmAction directly avoiding the extra layer of indirection. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	6ec605a7fb	vmm: api: Refactor generic action handler Rather than save the save a function pointer and use that instead the underlying action. This is useful for two reasons: 1. We can ensure that we generate HttpErrors in the same way as the other endpoints where API error variant should be determined by the request being made not the underlying error. 2. It can be extended to handle other generic actions where the function prototype differs slightly. As result of this refactoring it was found that the "vm.delete" endpoint was not connected so address that issue. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	c652625beb	vmm: api: Add a default implementation for simple PUT requests Extend the EndpointHandler trait to include automatic support for handling PUT requests. This will allow the removal of lots of duplicated code in the following commit from the API handling code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	a3e8bea03c	vmm: api: Move HttpError enum to http module Minor rearrangement of code to make it easier to implement refactoring. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-14 16:55:51 +01:00
Rob Bradford	9ccc7daa83	build, vmm: Update to latest kvm-ioctls The ch branch has been rebased to incorporate the latest upstream code requiring a small change to the unit tests. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-13 17:14:49 +02:00
Rob Bradford	88ec93d075	vmm: config: Add missing "id" from FsConfig parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-13 09:11:50 +01:00
Sebastien Boeuf	c37da600e8	vmm: Update DeviceTree upon PCI BAR reprogramming By passing a reference of the DeviceTree to the AddressManager, we can now update the DeviceTree whenever a PCI BAR is reprogrammed. This is mandatory to maintain the correct resources information related to each virtio-pci device, which will ensure correct information will be stored upon VM snapshot. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Sebastien Boeuf	d0ae9d7ce6	vmm: Share the DeviceTree across threads We want to be able to share the same DeviceTree across multiple threads, particularly to handle the use case where PCI BAR reprogramming might need to update the tree while from another thread a new device is being added to the tree. That's why this patch moves the DeviceTree instance into an Arc<Mutex<>> so that we can later share a reference of the same mutable tree with the AddressManager responsible for handling PCI BAR reprogramming. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Sebastien Boeuf	5e9d254564	vmm: Store and restore virtio-pci BAR resources By using the vector of resources provided by the DeviceNode, the device manager can store the information related to PCI BARs from a virtio-pci device. Based on this, and upon VM restoration, the device manager can restore the BARs in the expected location in the guest address space. One thing to note is that we only need to provide the VirtioPciDevice with the configuration BAR (BAR 0) since the SHaredMemory BAR info comes from the virtio device directly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Sebastien Boeuf	8a826ae24c	vmm: Store and restore virtio-pci device on right PCI slot Based on the new field "pci_bdf", a virtio-pci device can be restored at the same place on the PCI bus it was located before the VM snapshot. This ensures consistent placement on the PCI bus, based on the stored information related to each device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Sebastien Boeuf	98dac352b8	vmm: Add optional PCI b/d/f to each DeviceNode We need a way to store the information about where a PCI device was placed on the PCI bus before the VM was snapshotted. The way to do this is by adding an extra field to the DeviceNode structure. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-12 17:37:31 +01:00
Rob Bradford	5016fcf8d5	vhost_user_block: Use config::OptionParser to simplify block backend parsing Switch to using the recently added OptionParser in the code that parses the block backend. Fixes: #1092 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-11 09:40:40 +02:00
Rob Bradford	592de97fbd	vhost_user_net: Use config::OptionParser to simplify net backend parsing Switch to using the recently added OptionParser in the code that parses the network backend. Whilst doing this also update the net-backend syntax to use "sock" rather than socket. Fixes: #1092 Partially fixes: #1091 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-11 09:40:40 +02:00
Rob Bradford	12e00c0f45	vmm: cpu: Retry sending signals if necessary To avoid a race condition where the signal might "miss" the KVM_RUN ioctl() instead reapeatedly try sending a signal until the vCPU run is interrupted (as indicated by setting a new per vCPU atomic.) It important to also clear this atomic when coming out of a paused state. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Rob Bradford	31bde4f5da	vmm: Unpark the DeviceManager threads in shutdown To ensure that the DeviceManager threads (such as those used for virtio devices) are cleaned up it is necessary to unpark them so that they get cleanly terminated as part of the shutdown. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Rob Bradford	801e72ac6d	vmm: cpu: Unpause vCPU threads After setting the kill signal flag for the vCPU thread release the pause flag and unpark the threads. This ensures that that the vCPU thread will wake up and check the kill signal flag if the VM is paused. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Rob Bradford	91a4a2581e	vmm: cpu: When coming out of the pause event check for a kill signal Rather than immediately entering the vCPU run() code check if the kill signal is set. This allows paused VMs to be shutdown. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Rob Bradford	cd60de8f7f	Revert "vmm: vm: Unpark the threads before shutdown when the current state is paused" This reverts commit `e1a07ce3c4`. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-07 09:00:14 +02:00
Sebastien Boeuf	f6a71bec36	vmm: Add unit tests for DeviceTree Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	64e01684f9	vmm: Create new module device_tree This module will be dedicated to DeviceNode and DeviceTree definitions along with some dedicated unit tests. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	3b77be903d	vmm: Add device_node!() macro to improve code readability Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	83ec716ec4	vmm: Create breadth-first search iterator for the DeviceTree This iterator will let the VMM enumerate the resources associated with the DeviceManager, allowing for introspection. Moreover, by implementing a double ended iterator, we can get the hierarchy from the leaves to the root of the tree, which is very helpful in the context of restoring the devices in the right order. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	b91ab1e3a5	vmm: Remove the list of migratable devices Now that the device tree fully replaced the need for a dedicated list of migratable devices, this commit cleans up the codebase by removing it from the DeviceManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	1be7037229	vmm: Don't use migratable_devices for restore This commit switches from migratable_devices to device_tree in order to restore devices exclusively based on the device tree. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	bc6084390f	vmm: Add migratable field to the DeviceNode This commit adds an extra field to the DeviceNode so that the structure can hold a Migratable device. The long term plan is to be able to remove the dedicated table of migratable devices, but instead rely only on the device tree. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	7fec020f53	vmm: Create a dedicated DeviceTree structure In order to hide the complexity chosen for the device tree stored in the DeviceManager, we introduce a new DeviceTree structure. For now, this structure is a simple passthrough of a HashMap, but it can be extended to handle some DeviceTree specific operations. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	14b379dec5	vmm: Add an identifier field to DeviceNode structure Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	0805d458c4	vmm: Add support for multiple children per DeviceNode Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	daaeba5142	vmm: Change Node into DeviceNode Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	5c7df03efe	vmm: Store and restore virtio-pmem resources This device has a dedicated memory region in the guest address space, which means in case of snapshot/restore, it must be restored in the exact same location it was during the snapshot. That's through the resources that we can describe the location of this extra memory region, allowing the device for correct restoring. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	2e6895d911	vmm: Store and restore virtio-fs resources This device has a dedicated memory region in the guest address space, which means in case of snapshot/restore, it must be restored in the exact same location it was during the snapshot. That's through the resources that we can describe the location of this extra memory region, allowing the device for correct restoring. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	987f82152e	vmm: Store and restore virtio-mmio resources Based on the device tree, retrieve the resources associated with a virtio-mmio device to restore it at the right location in guest address space. Also, the IRQ number is correctly restored. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	9cb1e1cc6b	vmm: Perform MMIO allocation from virtio-mmio device creation Instead of splitting the MMIO allocation and the device creation into separate functions for virtio-mmio devices, it's is easier to move everything into the same function as we'll be able to gather resources in the same place for the same device. These resources will be stored in the device tree in a follow up patch. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	adf297066d	vmm: Create devices in different path if restoring the VM In case the VM is created from scratch, the devices should be created after the DeviceManager has been created. But this should not affect the restore codepath, as in this case the devices should be created as part of the restore() function. It's necessary to perform this differentiation as the restore must go through the following steps: - Create the DeviceManager - Restore the DeviceManager with the right state - Create the devices based on the restored DeviceManager's device tree - Restore each device based on the restored DeviceManager's device tree That's why this patch leverages the recent split of the DeviceManager's creation to achieve what's needed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	d39f91de02	vmm: Reorganize DeviceManager creation This commit performs the split of the DeviceManager's creation into two separate functions by moving anything related to device's creation after the DeviceManager structure has been initialized. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	89c2a5868c	vmm: Restore devices following the device tree Based on the device tree, we now ensure the restore can be done in the right order, as it will respect the dependencies between nodes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	52c80cfcf5	vmm: Snapshot and restore DeviceManager state The DeviceManager itself must be snapshotted in order to store the information regarding the devices associated with it, which effectively means we need to store the device tree. The mechanics to snapshot and restore the DeviceManagerState are added to the existing snapshot() and restore() implementations. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Sebastien Boeuf	5b408eec66	vmm: Create a device tree The DeviceManager now creates a tree of devices in order to store the resources associated with each device, but also to track dependencies between devices. This is a key part for proper introspection, but also to support snapshot and restore correctly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-05-05 16:08:42 +02:00
Rob Bradford	fec97e0586	vm-virtio, vmm: Delete unix socket on shutdown It's not possible to call UnixListener::Bind() on an existing file so unlink the created socket when shutting down the Vsock device. This will allow the VM to be rebooted with a vsock device. Fixes: #1083 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-05 13:01:38 +02:00
Rob Bradford	5109f914eb	vmm: config: Reject attempts to use VFIO or IOMMU without PCI Generate an error during validation if an attempt it made to place a device behind an IOMMU or using a VFIO device when not using PCI. Fixes: #751 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-05-05 11:20:52 +01:00
Rob Bradford	5115ad6e56	vmm: config: Support on/off/true/false for all booleans Migrate missing boolean controls over to the Toggle to handle all values. Fixes: #936 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-30 15:21:09 +02:00
Rob Bradford	d5bfa2dfc8	vmm, vhost_user_block: Make parameter names match --disk Make the --block-backend parameters match the --disk parameters. Fixes: #898 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-30 15:20:55 +02:00
Sebastien Boeuf	2f0bc06bec	vmm: Update default devices names as "internal" Let's put an underscore "_" in front of each device name to identify when it has been set internally. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	aaba6e777f	vmm: Add virtio-console to the list of Migratable devices The virtio-console was not added to the list of Migratable devices, which is fixed from this patch. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9ab4bb1ae2	devices: serial: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	06487131f9	vm-virtio: pci: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. It is based off the name from the virtio device attached to this transport layer. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	eeb7e10d1f	vm-virtio: mmio: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. It is based off the name from the virtio device attached to this transport layer. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9d84ef5073	vmm: Make the virtio identifier mandatory Because we know we will need every virtio device to be identified with a unique id, we can simplify the code by making the identifier mandatory. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	14350f5de4	devices: ioapic: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	556871570e	vm-virtio: iommu: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	052eff1ca7	vm-virtio: console: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	354c2a4b3d	vm-virtio: vhost-user-net: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	46e0b3ff75	vm-virtio: vhost-user-blk: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	bb7fa71fcb	vm-virtio: vhost-user-fs: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	ec5ff395cf	vm-virtio: vsock: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9b53044aae	vm-virtio: mem: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	1592a9292f	vm-virtio: pmem: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	2e91b73881	vm-virtio: rng: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	9eb7413fab	vm-virtio: net: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	be946caf4b	vm-virtio: blk: Expect an identifier upon device creation This identifier is chosen from the DeviceManager so that it will manage all identifiers across the VM, which will ensure uniqueness. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	ff9c8b847f	vmm: Always generate the next device name Even in the context of "mmio" feature, we need the next device name to be generated as we need to identify virtio-mmio devices to support snapshot and restore functionalities. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	8183141399	vmm: Add an identifier to the ioapic device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	e4386c8bb7	vmm: Add an identifier to the virtio-iommu device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	75ddd2a244	vmm: Add an identifier to the --console device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	eac350c454	vmm: Add an identifier to the virtio-mem device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	6802ef5406	vmm: Add an identifier to the --rng device This will be later used to identify each device used by the VM in order to perform introspection and snapshot/restore properly. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	d71d52e9b0	vmm: Fix virtio-console creation with virtual IOMMU If the virtio-console device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	b08fde5928	vmm: Fix virtio-rng creation with virtual IOMMU If the virtio-rng device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Sebastien Boeuf	8031ac33c3	vmm: Fix virtio-vsock creation with virtual IOMMU If the virtio-vsock device is supposed to be placed behind the virtual IOMMU, this must be explicitly propagated through the code. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 19:34:31 +01:00
Rob Bradford	8cef35745b	vmm: seccomp: Add fork, gettid and pipe2 syscalls to permitted list This is needed for self spawning with the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Rob Bradford	ce7678f29f	vmm: seccomp: Add tkill syscall to permitted list This is needed for rebooting on the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Rob Bradford	12758d7fad	vmm: seccomp: Add epoll_pwait syscall to permitted list This is needed for basic operation on the musl target. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 17:57:01 +01:00
Samuel Ortiz	86fcd19b8a	build: Initial musl support Fix all build failures and add musl to the gihub workflows. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-29 17:57:01 +01:00
Sebastien Boeuf	a5de49558e	vmm: Only allow removal of specific types of virtio device Now that all virtio devices are assigned with identifiers, they could all be removed from the VM. This is not something that we want to allow because it does not make sense for some devices. That's why based on the device type, we remove the device or we return an error to the user. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 13:33:19 +01:00
Sebastien Boeuf	9ed880d74e	vmm: Add an identifier to the --fs device By giving the devices ids this effectively enables the removal of the device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 13:33:19 +01:00
Sebastien Boeuf	7e0ab6b56d	vmm: Fix pmem device creation The parameters regarding the attachment to the virtio-iommu device was not propagated correclty, and any modification to the configuration was not stored back into it. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-29 13:33:19 +01:00
Rob Bradford	8de7448d44	vmm: api: Add "add-vsock" API entry point This allows the hotplugging of vsock devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	bf09a1e695	openapi: Add "id" field to VsockConfig Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	a76cf0865f	vmm: vm: Remove vsock device from config When doing device unplug remove the vsock device from the configuration if present. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	99422324a7	vmm: vm: Add "add_vsock()" Add the vsock device to the device manager and patch the config to add the new vsock device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	1d61c476a1	vmm: device_manager: Add support for hotplugging virtio-vsock devices Create a new VirtioVsock device and add it to the PCI bus upon hotplug. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	f8501a3bd3	vmm: config: Move --vsock syntax to VsockConfig This means it can be reused with ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Sebastien Boeuf	6e049e0da1	vmm: Add an identifier to the --vsock device It's possible to have multiple vsock devices so in preparation for hotplug/unplug it is important to be able to have a unique identifier for each device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-29 12:44:49 +01:00
Rob Bradford	10348f73e4	vmm, main: Support only zero or one vsock devices The Linux kernel does not support multiple virtio-vsock devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-28 20:07:18 +02:00
Rob Bradford	9d1f95a3cc	openapi: Add missing "id" field NetConfig/DiskConfig/PmemConfig/FsConfig were all missing the id field in the API yaml file. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-28 18:27:45 +02:00
Muminul Islam	e1a07ce3c4	vmm: vm: Unpark the threads before shutdown when the current state is paused If the current state is paused that means most of the handles got killed by pthread_kill We need to unpark those threads to make the shutdown worked. Otherwise The shutdown API hangs and the API is not responding afterwards. So before the shutdown call we need to resume the VM make it succeed. Fixes: #817 Signed-off-by: Muminul Islam <muislam@microsoft.com>	2020-04-27 09:09:12 +02:00
Rob Bradford	1df38daf74	vmm, tests: Make specifying a size optional for virtio-pmem If a size is specified use it (in particular this is required if the destination is a directory) otherwise seek in the file to get the size of the file. Add a new check that the size is a multiple of 2MiB otherwise the kernel will reject it. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-24 18:30:05 +01:00
Rob Bradford	7481e4d959	vmm: config: Validate that shared memory is enabled if using vhost-user Check that if any device using vhost-user (net & disk with vhost_user=true) or virtio-fs is enabled then check shared memory is also enabled. Fixes: #848 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-24 16:01:49 +01:00
Bo Chen	2ac6971a8b	vmm: MemoryManager: Cleanup the usage of std::ffi/io/result Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-04-23 21:39:51 +02:00
Bo Chen	3f42f86d81	vmm: Add the 'shared' and 'hugepages' controls to MemoryConfig The new 'shared' and 'hugepages' controls aim to replace the 'file' option in MemoryConfig. This patch also updated all related integration tests to use the new controls (instead of providing explicit paths to "/dev/shm" or "/dev/hugepages"). Fixes: #1011 Signed-off-by: Rob Bradford <robert.bradford@intel.com> Signed-off-by: Bo Chen <chen.bo@intel.com>	2020-04-23 21:39:51 +02:00
Martin Xu	5a380a6918	vmm: memory_manager: Support non-power-of-2 block sizes Replace alignment calculation of start address with functionally equivalent version that does not assume that the block size is a power of two. Signed-off-by: Martin Xu <martin.xu@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-22 09:11:51 +02:00
Sebastien Boeuf	c22fd39170	vmm: Remove virtio device's userspace mapping on hot-unplug When a virtio device is dynamically removed from the VM through the hot-unplug mechanism, every mapping associated with it must be properly removed. Based on the previous patches letting a VirtioDevice expose the list of userspace mappings associated with it, this patch can now remove all the KVM userspace memory regions through the MemoryManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	0a97c25464	vmm: Extend MemoryManager to remove userspace mappings The same way we added a helper for creating userspace memory mappings from the MemoryManager, this patch adds a new helper to remove some previously added mappings. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	fbcf3a7a7a	vm-virtio: Implement userspace_mappings() for virtio-pmem When hot-unplugging the virtio-pmem from the VM, we don't remove the associated userspace mapping. This patch will let us fix this in a following patch. For now, it simply adapts the code so that the Pmem device knows about the mapping associated with it. By knowing about it, it can expose it to the caller through the new userspace_mappings() function. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-21 10:02:21 +01:00
Sebastien Boeuf	18f7789a81	vmm: Add hotplugged virtio devices to the DeviceManager list The hotplugged virtio devices were not added to the list of virtio devices from the DeviceManager. This patch fixes it, as it was causing hotplugged virtio-fs devices from not supporting memory hotplug, since they were never getting the update as they were not part of the list of virtio devices held by the DeviceManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-20 20:36:26 +02:00
Dean Sheather	c2abadc293	vmm: Add ability to add virtio-fs device post-boot Adds DeviceManager method `make_virtio_fs_device` which creates a single device, and modifies `make_virtio_fs_devices` to use this method. Implements the new `vm.add-fs route`. Signed-off-by: Dean Sheather <dean@coder.com>	2020-04-20 20:36:26 +02:00
Dean Sheather	bb2139a408	vmm/api: Add vm.add-fs route Currently unimplemented. Once implemented, this API will allow for creating virtio-fs devices in the VM after it has booted. Signed-off-by: Dean Sheather <dean@coder.com>	2020-04-20 20:36:26 +02:00
Sebastien Boeuf	d35e775ed9	vmm: Update KVM userspace mapping when PCI BAR remapping In the context of the shared memory region used by virtio-fs in order to support DAX feature, the shared region is exposed as a dedicated PCI BAR, and it is backed by a KVM userspace mapping. Upon BAR remapping, the BAR is moved to a different location in the guest address space, and the KVM mapping must be updated accordingly. Additionally, we need the VirtioDevice to report the updated guest address through the shared memory region returned by get_shm_regions(). That's why a new setter is added to the VirtioDevice trait, so that after the mapping has been updated for KVM, we can tell the VirtioDevice the new guest address the shared region is located at. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-20 16:01:25 +02:00
Sebastien Boeuf	ac7178ef2a	vmm: Keep migratable devices list as a Vec The order the elements are pushed into the list is important to restore them in the right order. This is particularly important for MmioDevice (or VirtioPciDevice) and their VirtioDevice counterpart. A device must be fully ready before its associated transport layer management can trigger its restoration, which will end up activating the device in most cases. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-17 19:29:41 +02:00
Rob Bradford	e7e0e8ac38	vmm, devices: Add firmware debug port device OVMF and other standard firmwares use I/O port 0x402 as a simple debug port by writing ASCII characters to it. This is gated under a feature that is not enabled by default. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-17 12:54:00 +02:00
Rob Bradford	f9a0445c3d	vmm: vm: Remove device from configuration after unplug This ensures that a device that is removed will not reappear after a reboot. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	444e5c2a04	vmm: device_manager: Generalise NoAvailableVfioDeviceName We now support assigning device ids for VFIO and virtio-pci devices so this error can be generalised. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	5bab9c3894	vmm: device_manager: Assign ids to pmem/net/disk devices if absent If the id has not been provided by the user generate an incrementing id. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	514491a051	vmm: device_manager: Support unplugging virtio-pci devices Extend the eject_device() method on DeviceManager to also support virtio-pci devices being unplugged. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	476e4ce24f	vmm: device_manager: Add virtio-pci devices into id to BDF map In order to support hotplugging there is a map of human readable device id to PCI BDF map. As the device id is part of the specific device configuration (e.g. NetConfig) it is necessary to return the id through from the helper functions that create the devices through to the functions that add those devices to the bus. This necessitates changing a great deal of function prototypes but otherwise has little impact. Currently only if an id is supplied by the user as part of the device configuration is it populated into this map. A later commit will populate with an autogenerated name where none is supplied by the user. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	b38470df4b	vmm: config: Add "id" parameter to {Net, Disk, Pmem}Config This id will be used to unplug the device if the user has chosen an id. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	1beb62ed2d	vmm: vm: Don't panic on kernel load error Rather than panic()ing when we get a kernel loading error populate the error upwards. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-16 17:03:25 +02:00
Rob Bradford	72fdfff15d	vmm: device_manager: Remove unused "_mmap_regions" member Now that ownership of the memory regions used for the virtio-pmem and vhost-user-fs devices have been moved into those devices it is no longer necessary to track them inside DeviceManager. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Rob Bradford	70ecd6bab4	vmm, virtio: fs: Move freeing of mappped region into device Move the release of the managed memory region from the DeviceManager to the vhost-user-fs device. This ensures that the memory will be freed when the device is unplugged which will lead to it being Drop()ed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Rob Bradford	0c6706a510	vmm, virtio: pmem: Move freeing of mappped region into device Move the release of the managed memory region from the DeviceManager to the virtio-pmem device. This ensures that the memory will be freed when the device is unplugged which will lead to it being Drop()ed. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-14 17:46:11 +01:00
Sebastien Boeuf	b1554642e4	vmm: seccomp: Add missing mremap() syscall While testing self spawned vhost-user backends, it appeared that the backend was aborting due to a missing system call in the seccomp filters. mremap() was the culprit and this patch simply adds it to the whitelist. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-14 14:11:41 +02:00
Rob Bradford	28abfa9de5	vmm: openapi: Mark "initramfs" field nullable This should make it a pointer in the Go generated code so that it will be ommitted and thus not populated with an unhelpful default value. Fixes: #1015 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-09 23:25:18 +02:00
Rob Bradford	c260640fd5	vmm: config: Use Default::default() value for initramfs field This ensures that the field is filled with None when it is not specified as part of the deserialisation step. Fixes: #1015 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-09 17:28:45 +02:00
Alejandro Jimenez	7134f3129f	vmm: Allow PVH boot with initramfs We can now allow guests that specify an initramfs to boot using the PVH boot protocol. Signed-off-by: Alejandro Jimenez <alejandro.j.jimenez@oracle.com>	2020-04-09 17:28:03 +02:00
Rob Bradford	2d3f518c72	vmm: config: Error if both socket and path are specified for a disk This allows the validation of this requirement for both command line booted VMs and those booted via the API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	eeb7e2529d	vmm: config: Move max vCPUs > boot vCPUs check to validate() This allows the validation of this requirement for both command line booted VMs and those booted via the API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	12edb24678	vmm: config: Validate that serial/console file mode has a path Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	aaf382eee2	vmm: Move kernel check to VmConfig::validate() method Replace the existing VmConfig::valid() check with a call into .validate() as part of earlier config setup or boot API checks. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	3b0da2d895	vmm: vm: Validate configuration on API boot When performing an API boot validate the configuration. For now only some very basic validation is performed but in subsequent commits the validation will be extended. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Rob Bradford	99b2ada4d0	vmm: Start splitting configuration parsing and validation The configuration comes from a variety of places (commandline, REST API and restore) however some validation was only happening on the command line parsing path. Therefore introduce a new ability to validate the configuration before proceeding so that this can be used for commandline and API boots. For now move just the console and serial output mode validation under the new validation API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-08 12:06:09 +01:00
Sebastien Boeuf	0ea706faf5	vmm: openapi: Update OpenAPI definition with RestoreConfig Making sure the OpenAPI definition is up to date with newly added structure and parameters to support VM restoration. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	8d9d22436a	vmm: Add "prefault" option when restoring Now that the restore path uses RestoreConfig structure, we add a new parameter called "prefault" to it. This will give the user the ability to populate the pages corresponding to the mapped regions backed by the snapshotted memory files. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	a517ca23a0	vmm: Move restore parameters into common RestoreConfig structure The goal here is to move the restore parameters into a dedicated structure that can be reused from the entire codebase, making the addition or removal of a parameter easier. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	6712958f23	vmm: memory: Add prefault option when creating region When CoW can be used, the VM restoration time is reduced, but the pages are not populated. This can lead to some slowness from the guest when accessing these pages. Depending on the use case, we might prefer a slower boot time for better performances from guest runtime. The way to achieve this is to prefault the pages in this case, using the MAP_POPULATE flag along with CoW. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	b2cdee80b6	vmm: memory: Restore with Copy-on-Write when possible This patch extends the previous behavior on the restore codepath. Instead of copying the memory regions content from the snapshot files into the new memory regions, the VMM will use the snapshot region files as the backing files behind each mapped region. This is done in order to reduce the time for the VM to be restored. When the source VM has been initially started with a backing file, this means it has been mapped with the MAP_SHARED flag. For this case, we cannot use the CoW trick to speed up the VM restore path and we simply fallback onto the copy of the memory regions content. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	d771223b2f	vmm: memory: Extend new() to support external backing files Whenever a MemoryManager is restored from a snapshot, the memory regions associated with it might need to directly back the mapped memory for increased performances. If that's the case, a list of external regions is provided and the MemoryManager should simply ignore what's coming from the MemoryConfig. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	ee5a041a0f	vmm: memory: Add Copy-on-Write parameter when creating region Now that we can choose specific mmap flags for the guest RAM, we create a new parameter "copy_on_write" meaning that the memory mappings backed by a file should be performed with MAP_PRIVATE instead of MAP_SHARED. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	be4e1e8712	vmm: memory: Use fine grained mmap wrapper In order to anticipate the need for special mmap flags when memory mapping the guest RAM, we need to switch from from_file() wrapper to build() wrapper. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-08 10:56:14 +02:00
Sebastien Boeuf	b9f9f01fcc	vmm: Extend seccomp filters to allow snapshot/restore A few KVM ioctls were missing in order to perform both snapshot and restore while keeping seccomp enabled. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	6eb721301c	vmm: Enable restore feature This connects the dots together, making the request from the user reach the actual implementation for restoring the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	53613319cc	vmm: Enable snapshot feature This connects the dots together, making the request from the user reach the actual implementation for snapshotting the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	2cd0bc0a2c	vmm: Create initial VM from its snapshot The MemoryManager is somehow a special case, as its restore() function was not implemented as part of the Snapshottable trait. Instead, and because restoring memory regions rely both on vm.json and every memory region snapshot file, the memory manager is restored at creation time. This makes the restore path slightly different from CpuManager, Vcpu, DeviceManager and Vm, but achieve the correct restoration of the MemoryManager along with its memory regions filled with the correct content. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	b55b83c6e8	vmm: vm: Implement the Transportable trait This is only implementing the send() function in order to store all Vm states into a file. This needs to be extended for live migration, by adding more transport methods, and also the recv() function must be implemented. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	1ed357cf34	vmm: vm: Implement the Snapshottable trait By aggregating snapshots from the CpuManager, the MemoryManager and the DeviceManager, Vm implements the snapshot() function from the Snapshottable trait. And by restoring snapshots from the CpuManager, the MemoryManager and the DeviceManager, Vm implements the restore() function from the Snapshottable trait. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	20ba271b6c	vmm: memory_manager: Implement the Transportable trait This implements the send() function of the Transportable trait, so that the guest memory regions can be saved into one file per region. This will need to be extended for live migration, as it will require other transport methods and the recv() function will need to be implemented too. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	e606112cef	vmm: memory_manager: Implement the Snapshottable trait In order to snapshot the content of the guest RAM, the MemoryManager must implement the Snapshottable trait. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	50b3f008d1	vmm: cpu: Implement the Snapshottable trait Implement the Snapshottable trait for Vcpu, and then implements it for CpuManager. Note that CpuManager goes through the Snapshottable implementation of Vcpu for every vCPU in order to implement the Snapshottable trait for itself. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Sebastien Boeuf	f787c409c4	vmm: cpu: Factorize vcpu starting code Anticipating the need for a slightly different function for restoring vCPUs, this patch factorizes most of the vCPU creation, so that it can be reused for migration purposes. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-07 12:26:10 +02:00
Cathy Zhang	722f9b6628	vmm: cpu: Get and set KVM vCPU state These two new helpers will be useful to capture a vCPU state and being able to restore it at a later time. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Cathy Zhang	13756490b5	vmm: cpu: Track all Vcpus through CpuManager In anticipation for the CpuManager to aggregate all Vcpu snapshots together, this change makes sure the CpuManager has a handle onto every vCPU. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com> Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Samuel Ortiz	a0d5dbce6c	vmm: device_manager: Implement the Snapshottable trait Based on the list of Migratable devices stored by the DeviceManager, the DeviceManager can implement the Snapshottable trait by aggregating all devices snapshots together. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-07 12:26:10 +02:00
Yi Sun	93d3abfd6e	vmm: device_manager: Make serial and ioapic devices migratable Serial and Ioapic both implement the Migratable trait, hence the DeviceManager can store them in the list of Migratable devices. Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-07 12:26:10 +02:00
Rob Bradford	c7dfbd8a84	vmm: config: Implement fmt::Display for error Fixes: #367 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	d8119fda13	vmm: config: Remove unused error entries These entries are not currently used. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	1a10f16ad0	vmm: config: Consolidate size parsing code The parse_size helper function can now be consolidated into the ByteSized FromStr implementation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	f449486b9b	vmm: config: Make toggle parsing more tolerant Support "true" and "false" as well as well as capitalised forms. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	a4e0ce58c7	vmm: config: Consolidate on/off parsing Now all parsing code makes use of the Toggle and it's FromStr support move the helper function into the from_str() implementation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	c731a943d4	vmm: config: Port vsock to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	37264cf21b	vmm: config: Add unit testing for vsock Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	8665898ff3	vmm: config: Port device parsing to OptionParser Also make the "path" option required and generate an error if it is not provided. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	a85e2fa735	vmm: config: Add unit test for VFIO device parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	bed282b801	vmm: config: Add "valueless" options to OptionParser Valueless options are those like "off" or "tty" as used by the console options. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	2ae3392d32	vmm: config: Port console parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	143d63c88e	vmm: config: Add unit test for console parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	5ab58e743a	vmm: config: Port pmem option to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	233ad78b3a	vmm: config: Add parsing test for pmem Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	13dc637350	vmm: config: Port filesystem parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	7a071c28db	vmm: config: Implement unit testing for virtio-fs parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	e4cd3072d4	vmm: config: Port RNG options to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	708dbb973a	vmm: config: Add RNG parsing unit test Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	057e71d266	vmm: config: Accept empty value strings The integration tests and documentation make use of empty value strings like "--net tap=" accept them but return None so that the default value will be used as expected. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	218c780f67	vmm: config: Port network parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	8754720e2d	vmm: config: Add unit test for net parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	224e3ddef4	vmm: config: Switch disk parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e10244716	vmm: config: Add unit test for disk parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	e40ae6274b	vmm: config: Port memory option parsing to OptionParser This simplifies the parsing of the option by using OptionParser along with its automatic conversion behaviour. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	be32065aa4	vmm: config: Add "ByteSized" type for simplifying parsing of byte sizes Byte sizes are quantities ending in "K", "M", "G" and by implementing this type with a FromStr implementation the values can be converted using .parse(). Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	f01bd7d56d	vmm: config: Implement FromStr for HotplugMethod This allows the use of .parse() to automatically convert the string to the enum. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	746138039d	vmm: config: Add a Toggle type for "on/off" strings Some of the config parameters take an "on" or "off". Add a way to neatly parse that. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	929142bc2e	vmm: config: Add memory parsing unit test Before porting over to OptionParser add a unit test to validate the current memory parsing code. This showed up a bug where the "size=" was always required. Temporarily resolve this by assigning the string a default value which will later be replaced when the code is refactored. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	68203ea414	vmm: config: Port CPU parsing to OptionParser Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e6a2825ba	vmm: config: Add unit test for CPU parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Rob Bradford	9e7231cd69	vmm: config: Introduce basic OptionParser This will be used to simplify and consolidate much of the parsing code used for command line parameters. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-04-06 10:31:24 +01:00
Samuel Ortiz	447af8e702	vmm: vm: Factorize the device and cpu managers creation routine Into a new_from_memory_manager() routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	c73c9b112c	vmm: vm: Open kernel and initramfs once all managers are created Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	0646a90626	vmm: cpu: Pass CpusConfig to simplify the new() prototype Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	b584ec3fb3	vmm: memory_manager: Own the system allocator Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	ef2b11ee6c	vmm: memory_manager: Pass MemoryConfig to simplify the new() prototype Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	622f3f8fb6	vmm: vm: Avoid ioapic variable creation For a more readable VM creation routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	164e810069	vmm: cpu: Move CPUID patching to CpuManager Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	1a2c1f9751	vmm: vm: Factorize the KVM setup code Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	7a50646c02	vmm: device_manager: Convert migratable_devices to a map We must be able to map a migratable component id to its device. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-03 18:05:18 +01:00
Samuel Ortiz	8f300bed83	vmm: api: Add a /api/v1/vm.restore endpoint Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	92c73c3b78	vmm: Add a VmRestore command Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	39d4f817f0	vmm: http: Add a /api/v1/vm.snapshot endpoint Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	cf8f8ce93a	vmm: api: Add a Snapshot command Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-02 13:24:25 +01:00
Sebastien Boeuf	452475c280	vmm: Add migration helpers Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-02 13:24:25 +01:00
Samuel Ortiz	1b1a2175ca	vm-migration: Define the Snapshottable and Transportable traits A Snapshottable component can snapshot itself and provide a MigrationSnapshot payload as a result. A MigrationSnapshot payload is a map of component IDs to a list of migration sections (MigrationSection). As component can be made of several Migratable sub-components (e.g. the DeviceManager and its device objects), a migration snapshot can be made of multiple snapshot itself. A snapshot is a list of migration sections, each section being a component state snapshot. Having multiple sections allows for easier and backward compatible migration payload extensions. Once created, a migratable component snapshot may be transported and this is what the Transportable trait defines, through 2 methods: send and recv. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com> Signed-off-by: Yi Sun <yi.y.sun@linux.intel.com>	2020-04-02 13:24:25 +01:00
Sebastien Boeuf	2d17f4384a	vmm: seccomp: Add missing open() syscall On some systems, the open() system call is used by Cloud-Hypervisor, that's why it should be part of the seccomp filters whitelist. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-04-02 09:56:48 +02:00
Sebastien Boeuf	e4ea8b0bef	vmm: Add missing syscalls to the seccomp filters Both clock_gettime and gettimeofday syscalls where missing when running Cloud-Hypervisor on a Linux host without vDSO enabled. On a system with vDSO enabled, the syscalls performed by vDSO were not filtered, that's why we didn't have to whitelist them. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 16:50:52 +00:00
Sebastien Boeuf	9e18177654	vmm: Add memory hotplug support to VFIO PCI devices Extend the update_memory() method from DeviceManager so that VFIO PCI devices can update their DMA mappings to the physical IOMMU, after a memory hotplug has been performed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 09:35:39 +01:00
Sebastien Boeuf	cc67131ecc	vmm: Retrieve new memory region when memory is extended Whenever the memory is resized, it's important to retrieve the new region to pass it down to the device manager, this way it can decide what to do with it. Also, there's no need to use a boolean as we can instead use an Option to carry the information about the region. In case of virtio-mem, there will be no region since the whole memory has been reserved up front by the VMM at boot. This means only the ACPI hotplug will return a region and is the only method that requires the memory to be updated from the device manager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-27 09:35:39 +01:00
Samuel Ortiz	8fc7bf2953	vmm: Move to the latest linux-loader Commit 2adddce2 reorganized the crate for a cleaner multi architecture (x86_64 and aarch64) support. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-27 08:48:20 +01:00
Sebastien Boeuf	785812d976	vmm: Fallback to legacy boot if PVH is enabled along with initramfs For now, the codebase does not support booting from initramfs with PVH boot protocol, therefore we need to fallback to the legacy boot. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	6cce7b9560	arch: load initramfs and populate zero page * load the initramfs File into the guest memory, aligned to page size * finally setup the initramfs address and its size into the boot params (in configure_64bit_boot) Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	1f9bc68c54	openapi: Add initramfs support added InitramfsConfig property to the REST API spec Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Damjan Georgievski	4db252b418	main, vmm: add --initramfs cli option currently unused, the initramfs argument is added to the cli, and stored in vmm::config:VmConfig as an Option(InitramfsConfig(PathBuf)) Signed-off-by: Damjan Georgievski <gdamjan@gmail.com>	2020-03-26 11:59:03 +01:00
Rob Bradford	6244beb9d5	openapi: Add "vm.add-net" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	57c3fa4b1e	vmm: Add "add-net" to the API Add the HTTP and internal API entry points for adding a network device at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	f664cddec9	vmm: Add support for adding network devices to the VM The persistent memory will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	8f323e61d8	vmm: Add support to DeviceManager for hotplugging network devices Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	42a9896fe4	vmm: device_manager: Refactor make_virtio_net_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Rob Bradford	9df601a1df	bin, vmm: Centralise the net syntax This will allow the syntax to be reused with cloud-hypervsor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 17:58:06 +01:00
Samuel Ortiz	41d7b3a387	vmm: memory_manager: Only send the GED notification for the ACPI method Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	15d9ec0149	openapit: Add hotplug_method to MemoryConfig Add hotplug_method to MemoryConfig in cloud-hypervisor.yaml. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	e63f98182a	vmm: device: Add make_virtio_mem_devices Add make_virtio_mem_devices to add virtio-mem to vmm. Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Hui Zhu	e6b934a56a	vmm: Add support for virtio-mem This commit adds new option hotplug_method to memory config. It can set the hotplug method to "acpi" or "virtio-mem". Signed-off-by: Hui Zhu <teawater@antfin.com>	2020-03-25 15:54:16 +01:00
Rob Bradford	75878dd90a	openapi: Add "vm.add-pmem" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	f6f4c68fb4	vmm: Add "add-pmem" to the API Add the HTTP and internal API entry points for adding persistent memory at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	15de30f141	vmm: Add support for adding pmem devices to the VM The persistent memory will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	f7def621dd	vmm: Add support to DeviceManager for hotplugging pmem devices Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	8c3ea8cd76	vmm: device_manager: Refactor make_virtio_pmem_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	a7296bbb52	bin, vmm: Centralise the pmem syntax This will allow the syntax to be reused with cloud-hypervisor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 13:18:17 +01:00
Rob Bradford	4c9d15d44c	vmm: Fix copy and paste error message vm_remove_device was copied from vm_add_device but the error message wasn't correctly updated. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	82cad99c0b	openapi: Add "vm.add-disk" entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	f2151b2734	vmm: Add "add-disk" to the API Add the HTTP and internal API entry points for adding disks at runtime. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	164ec2b8e6	vmm: Add support for adding disks to the VM The disk will be hotplugged via DeviceManager and saved in the config for later use. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	b3082c1984	vmm: Add support to DeviceManager for hotplugging disks Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	2be703ca92	vmm: device_manager: Refactor make_virtio_block_devices Split it into a method that creates a single device which is called by the multiple device version so this can be used when dynamically adding a device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Rob Bradford	66da29d8dd	bin, vmm: Centralise the disk syntax This will allow the syntax to be reused with cloud-hypervsor binary and ch-remote. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-25 09:35:53 +00:00
Sebastien Boeuf	e54f8ec8a5	vmm: Update memory through DeviceManager Whenever the VM memory is resized, DeviceManager needs to be notified so that it can subsequently notify each virtio devices about it. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 19:01:15 +00:00
Sebastien Boeuf	feb8d7ae90	vmm: Separate seccomp filters between VMM and API threads This separates the filters used between the VMM and API threads, so that we can apply different rules for each thread. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	f1a23d712f	vmm: api: Add seccomp to the HTTP API thread Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	db62cb3f4d	vmm: Add seccomp filter to the VMM thread This commit introduces the application of the seccomp filter to the VMM thread. The filter is empty for now (SeccompLevel::None). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Sebastien Boeuf	cb98d90097	vmm: Create new seccomp_filter module Based on the seccomp crate, we create a new vmm module responsible for creating a seccomp filter that will be applied to the VMM main thread. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2020-03-24 14:59:57 +01:00
Rob Bradford	f7197e8415	vmm: Add a "discard_writes=" to --pmem This opens the backing file read-only, makes the pages in the mmap() read-only and also makes the KVM mapping read-only. The file is also mapped with MAP_PRIVATE to make the changes local to this process only. This is functional alternative to having support for making a virtio-pmem device readonly. Unfortunately there is no concept of readonly virtio-pmem (or any type of NVDIMM/PMEM) in the Linux kernel so to be able to have a block device that is appears readonly in the guest requires significant specification and kernel changes. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Rob Bradford	d11a67b0fe	vmm: Use more generic MmapRegion constructor Switch to MmapRegion::build() and fill in the fields appropriately. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Rob Bradford	7257e890ef	vmm: Add "readonly" parameter MemoryManager::create_userspace_mapping Use this boolean to turn on the KVM_MEM_READONLY flag to indicate that this memory mapping should not be writable by the VM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-20 14:46:34 +01:00
Qiu Wenbo	c503118d16	vmm: fix a corrupted stack caused by get_win_size According to `asm-generic/termios.h`, the `struct winsize` should be: struct winsize { unsigned short ws_row; unsigned short ws_col; unsigned short ws_xpixel; unsigned short ws_ypixel; }; The ioctl of TIOCGWINSZ will trigger a segfault on aarch64. Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2020-03-20 07:30:06 +01:00
Rob Bradford	0788600702	build: Remove "pvh_boot" feature flag This feature is stable and there is no need for this to be behind a flag. This will also reduce the time needed to run the integration test as we will not be running them all again under the flag. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-19 13:05:44 +00:00
Rob Bradford	477bc17f18	bin: Share VFIO device syntax between cloud-hypervisor and ch-remote Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2020-03-18 23:38:55 +00:00
Jose Carlos Venegas Munoz	a31ffef085	openapi: Add hotplug_size for memory hotplug Add hotplug_size, needed to be defined when hotplug is used. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2020-03-18 19:06:07 +00:00

... 4 5 6 7 8 ...

1034 Commits