cloud-hypervisor

mirror of https://github.com/cloud-hypervisor/cloud-hypervisor.git synced 2024-11-05 11:31:14 +00:00

Author	SHA1	Message	Date
Rob Bradford	1a2d0e6dd8	build: bump linux-loader from 0.3.0 to 0.4.0 Requires manual change to command line loading. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-24 09:11:57 +00:00
Michael Zhao	d72af85c42	vmm: Add "_CCA" field to ACPI DSDT table "_CCA" is required by DMA configuration on AArch64. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-09-24 07:57:57 +01:00
Rob Bradford	43365ade2e	vmm, pci: Implement virtio-mem support for vfio-user Implement the infrastructure that lets a virtio-mem device map the guest memory into the device. This is necessary since with virtio-mem zones memory can be added or removed and the vfio-user device must be informed. Fixes: #3025 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-21 15:42:49 +01:00
Rob Bradford	e9d67dc405	vmm: pci: Move creation of vfio_user::Client to DeviceManager By moving this from the VfioUserPciDevice to DeviceManager the client can be reused for handling DMA mapping behind an IOMMU. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-21 15:42:49 +01:00
Rob Bradford	fd4f32fa69	virtio-mem: Support multiple mappings For vfio-user the mapping handler is per device and needs to be removed when the device in unplugged. For VFIO the mapping handler is for the default VFIO container (used when no vIOMMU is used - using a vIOMMU does not require mappings with virtio-mem) To represent these two use cases use an enum for the handlers that are stored. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-21 15:42:49 +01:00
dependabot[bot]	d826b4fbdc	build: bump arc-swap from 1.3.2 to 1.4.0 Bumps [arc-swap](https://github.com/vorner/arc-swap) from 1.3.2 to 1.4.0. - [Release notes](https://github.com/vorner/arc-swap/releases) - [Changelog](https://github.com/vorner/arc-swap/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/arc-swap/compare/v1.3.2...v1.4.0) --- updated-dependencies: - dependency-name: arc-swap dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2021-09-19 17:12:50 +00:00
Rob Bradford	0faa7afac2	vmm: Add fast path for PCI config IO port Looking up devices on the port I/O bus is time consuming during the boot at there is an O(lg n) tree lookup and the overhead from taking a lock on the bus contents. Avoid this by adding a fast path uses the hardcoded port address and size and directs PCI config requests directly to the device. Command line: target/release/cloud-hypervisor --kernel ~/src/linux/vmlinux --cmdline "root=/dev/vda1 console=ttyS0" --serial tty --console off --disk path=~/workloads/focal-server-cloudimg-amd64-custom-20210609-0.raw --api-socket /tmp/api PIO exit: 17913 PCI fast path: 17871 Percentage on fast path: 99.8% perf before: marvin:~/src/cloud-hypervisor (main )$ perf report -g \| grep resolve 6.20% 6.20% vcpu0 cloud-hypervisor [.] vm_device:🚌:Bus::resolve perf after: marvin:~/src/cloud-hypervisor (2021-09-17-ioapic-fast-path )$ perf report -g \| grep resolve 0.08% 0.08% vcpu0 cloud-hypervisor [.] vm_device:🚌:Bus::resolve The compromise required to implement this fast path is bringing the creation of the PciConfigIo device into the DeviceManager::new() so that it can be used in the VmmOps struct which is created before DeviceManager::create_devices() is called. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-17 17:09:45 +01:00
Michael Zhao	b3fa56544c	virtio-devices: iommu: Support AArch64 The MSI IOVA address on X86 and AArch64 is different. This commit refactored the code to receive the MSI IOVA address and size from device_manager, which provides the actual IOVA space data for both architectures. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-09-17 12:19:46 +02:00
Michael Zhao	253c06d3ba	arch/aarch64: Add virtio-iommu device in FDT Add a virtio-iommu node into FDT if iommu option is turned on. Now we support only one virtio-iommu device. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-09-17 12:19:46 +02:00
William Douglas	46f6d9597d	vmm: Switch to using the serial_manager for serial input This change switches from handling serial input in the VMM thread to its own thread controlled by the SerialManager. The motivation for this change is to avoid the VMM thread being unable to process events while serial input is happening and vice versa. The change also makes future work flushing the serial buffer on PTY connections easier. Signed-off-by: William Douglas <william.douglas@intel.com>	2021-09-17 11:15:35 +01:00
William Douglas	7b4f56e372	vmm: Add new serial_manager for serial input handling This change adds a SerialManager with its own epoll handling that should be created and run by the DeviceManager when creating an appropriately configured console (serial tty or pty). Both stdin and pty input are handled by the SerialManager. The stdin and pty specific methods used by the VMM should be removed in a future commit. Signed-off-by: William Douglas <william.douglas@intel.com>	2021-09-17 11:15:35 +01:00
William Douglas	d6a2f48b32	vmm: device_manager: Make PtyPair implement Clone The clone method for PtyPair should have been an impl of the Clone trait but the method ended up not being used. Future work will make use of the trait however so correct the missing trait implementation. Signed-off-by: William Douglas <william.douglas@intel.com>	2021-09-17 11:15:35 +01:00
dependabot[bot]	f67b3f79ea	build: bump vmm-sys-util from 0.8.0 to 0.9.0 Bumps [vmm-sys-util](https://github.com/rust-vmm/vmm-sys-util) from 0.8.0 to 0.9.0. - [Release notes](https://github.com/rust-vmm/vmm-sys-util/releases) - [Changelog](https://github.com/rust-vmm/vmm-sys-util/blob/main/CHANGELOG.md) - [Commits](https://github.com/rust-vmm/vmm-sys-util/compare/v0.8.0...v0.9.0) --- updated-dependencies: - dependency-name: vmm-sys-util dependency-type: direct:production update-type: version-update:semver-minor ... This needed a bunch of manual updates as well, including vfio-ioctls and vhost crates. The vhost crate is being patched with the latest version from rust-vmm because the version 0.1.0 on crates.io doesn't include the patches we need yet. Signed-off-by: dependabot[bot] <support@github.com> Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-09-16 14:01:19 +01:00
dependabot[bot]	c1e896dddb	build: bump libc from 0.2.101 to 0.2.102 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.101 to 0.2.102. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.101...0.2.102) --- updated-dependencies: - dependency-name: libc dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-09-15 17:23:46 +00:00
Sebastien Boeuf	a6040d7a30	vmm: Create a single VFIO container For most use cases, there is no need to create multiple VFIO containers as it causes unwanted behaviors. Especially when passing multiple devices from the same IOMMU group, we need to use the same container so that it can properly list the groups that have been already opened. The correct logic was already there in vfio-ioctls, but it was incorrectly used from our VMM implementation. For the special case where we put a VFIO device behind a vIOMMU, we must create one container per device, as we need to control the DMA mappings per device, which is performed at the container level. Because we must keep one container per device, the vIOMMU use case prevents multiple devices attached to the same IOMMU group to be passed through the VM. But this is a limitation that we are fine with, especially since the vIOMMU doesn't let us group multiple devices in the same group from a guest perspective. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-09-15 09:08:13 -07:00
dependabot[bot]	8836715c2d	build: bump serde_json from 1.0.67 to 1.0.68 Bumps [serde_json](https://github.com/serde-rs/json) from 1.0.67 to 1.0.68. - [Release notes](https://github.com/serde-rs/json/releases) - [Commits](https://github.com/serde-rs/json/compare/v1.0.67...v1.0.68) --- updated-dependencies: - dependency-name: serde_json dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-09-15 00:06:23 +00:00
Alyssa Ross	330b5ea3be	vmm: notify virtio-console of pty resizes When a pty is resized (using the TIOCSWINSZ ioctl -- see ioctl_tty(2)), the kernel will send a SIGWINCH signal to the pty's foreground process group to notify it of the resize. This is the only way to be notified by the kernel of a pty resize. We can't just make the cloud-hypervisor process's process group the foreground process group though, because a process can only set the foreground process group of its controlling terminal, and cloud-hypervisor's controlling terminal will often be the terminal the user is running it in. To work around this, we fork a subprocess in a new process group, and set its process group to be the foreground process group of the pty. The subprocess additionally must be running in a new session so that it can have a different controlling terminal. This subprocess writes a byte to a pipe every time the pty is resized, and the virtio-console device can listen for this in its epoll loop. Alternatives I considered were to have the subprocess just send SIGWINCH to its parent, and to use an eventfd instead of a pipe. I decided against the signal approach because re-purposing a signal that has a very specific meaning (even if this use was only slightly different to its normal meaning) felt unclean, and because it would have required using pidfds to avoid race conditions if cloud-hypervisor had terminated, which added complexity. I decided against using an eventfd because using a pipe instead allows the child to be notified (via poll(2)) when nothing is reading from the pipe any more, meaning it can be reliably notified of parent death and terminate itself immediately. I used clone3(2) instead of fork(2) because without CLONE_CLEAR_SIGHAND the subprocess would inherit signal-hook's signal handlers, and there's no other straightforward way to restore all signal handlers to their defaults in the child process. The only way to do it would be to iterate through all possible signals, or maintain a global list of monitored signals ourselves (vmm:vm::HANDLED_SIGNALS is insufficient because it doesn't take into account e.g. the SIGSYS signal handler that catches seccomp violations). Signed-off-by: Alyssa Ross <hi@alyssa.is>	2021-09-14 15:43:25 +01:00
Alyssa Ross	28382a1491	virtio-devices: determine tty size in console This prepares us to be able to handle console resizes in the console device's epoll loop, which we'll have to do if the output is a pty, since we won't get SIGWINCH from it. Signed-off-by: Alyssa Ross <hi@alyssa.is>	2021-09-14 15:43:25 +01:00
dependabot[bot]	f3778a7fc7	build: bump anyhow from 1.0.43 to 1.0.44 Bumps [anyhow](https://github.com/dtolnay/anyhow) from 1.0.43 to 1.0.44. - [Release notes](https://github.com/dtolnay/anyhow/releases) - [Commits](https://github.com/dtolnay/anyhow/compare/1.0.43...1.0.44) --- updated-dependencies: - dependency-name: anyhow dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-09-14 00:22:00 +00:00
Alyssa Ross	8abe8c679b	seccomp: allow mmap everywhere brk is allowed Musl often uses mmap to allocate memory where Glibc would use brk. This has caused seccomp violations for me on the API and signal handling threads. Signed-off-by: Alyssa Ross <hi@alyssa.is>	2021-09-10 12:01:31 -07:00
Rob Bradford	b6b686c71c	vmm: Shutdown VMM if API thread panics See: #3031 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-10 10:52:08 -07:00
Rob Bradford	171d12943d	vmm: memory_manager: Increase robustness of MemoryManager control device See: #1289 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-10 10:23:19 -07:00
Rob Bradford	bdc44cd8bc	vmm: cpu: Increase robustness of CpuManager control device See: #1289 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-10 10:22:05 -07:00
Bo Chen	4f37a273d9	vmm: Fix clippy issue error: all if blocks contain the same code at the end --> vmm/src/memory_manager.rs:884:9 \| 884 \| / Ok(mm) 885 \| \| } \| \|_________^ Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-09-08 13:31:19 -07:00
Rob Bradford	d64a77a5c6	vmm: Shutdown VMM if signal thread panics See: #3031 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-08 11:26:48 -07:00
Rob Bradford	e0d05683ab	vmm: Split up functions for creating signal handler and tty setup These are quite separate and should be in their own functions. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-08 11:26:48 -07:00
Rob Bradford	387753ae1d	vmm: Remove concept of "input_enabled" This concept ends up being broken with multiple types on input connected e.g. console on TTY and serial on PTY. Already the code for checking for injecting into the serial device checks that the serial is configured. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-08 11:26:48 -07:00
Rob Bradford	951ad3495e	vmm: Only resize virtio-console when attached to TTY Fixes: #3092 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-08 11:26:48 -07:00
Rob Bradford	0dbb2683e3	vmm: Consolidate duplicated code for setting up signal handler Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-08 11:26:48 -07:00
Rob Bradford	687d646c60	virtio-devices, vmm: Shutdown VMM on virtio thread panic Shutdown the VMM in the virtio (or VMM side of vhost-user) thread panics. See: #3031 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-08 09:40:36 +01:00
Rob Bradford	54e523c302	virtio-devices: Use a common method for spawning virtio threads Introduce a common solution for spawning the virtio threads which will make it easier to add the panic handling. During this effort I discovered that there were no seccomp filters registered for the vhost-user-net thread nor the vhost-user-block thread. This change also incorporates basic seccomp filters for those as part of the refactoring. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-08 09:40:36 +01:00
Wei Liu	9c5b404415	vmm: MSHV now supports VFIO-based device passthrough Drop a few feature gates and adjust code a bit. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-09-07 15:17:08 +01:00
Wei Liu	10b954e954	build: use vfio-ioctls that supports MSHV Disable default features and propagate hypervisor selection where necessary. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-09-07 15:17:08 +01:00
dependabot[bot]	a20041ba68	build: bump thiserror from 1.0.28 to 1.0.29 Bumps [thiserror](https://github.com/dtolnay/thiserror) from 1.0.28 to 1.0.29. - [Release notes](https://github.com/dtolnay/thiserror/releases) - [Commits](https://github.com/dtolnay/thiserror/compare/1.0.28...1.0.29) --- updated-dependencies: - dependency-name: thiserror dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-09-07 08:35:50 +00:00
Henry Wang	c50051a686	device_manager: Enable power button for ACPI on AArch64 Current AArch64 power button is only for device tree using a PL061 GPIO controller device. Since AArch64 now supports ACPI, this commit extend the power button on AArch64 to: - Using GED for ACPI+UEFI boot. - Using PL061 for device tree boot. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-09-03 10:27:52 -07:00
Rob Bradford	e475b12cf7	virtio-devices, vmm: Upgrade restore related messages to info!() These happen only sporadically so can be included at the info!() level. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-03 09:30:55 -07:00
Rob Bradford	968902dfec	devices, vmm: Upgrade exit reasons to info!() level debugging These statements are useful for understanding the cause of reset or shutdown of the VM and are not spammy so should be included at info!() level. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-03 09:30:55 -07:00
Alyssa Ross	7549149bb5	vmm: ensure signal handlers run on the right thread Despite setting up a dedicated thread for signal handling, we weren't making sure that the signals we were listening for there were actually dispatched to the right thread. While the signal-hook provides an iterator API, so we can know that we're only processing the signals coming out of the iterator on our signal handling thread, the actual signal handling code from signal-hook, which pushes the signals onto the iterator, can run on any thread. This can lead to seccomp violations when the signal-hook signal handler does something that isn't allowed on that thread by our seccomp policy. To reproduce, resize a terminal running cloud-hypervisor continuously for a few minutes. Eventually, the kernel will deliver a SIGWINCH to a thread with a restrictive seccomp policy, and a seccomp violation will trigger. As part of this change, it's also necessary to allow rt_sigreturn(2) on the signal handling thread, so signal handlers are actually allowed to run on it. The fact that this didn't seem to be needed before makes me think that signal handlers were almost _never_ actually running on the signal handling thread. Signed-off-by: Alyssa Ross <hi@alyssa.is>	2021-09-02 21:33:31 +01:00
Rob Bradford	c2144b5690	vmm, virtio-console: Move input reading into virtio-console thread Move the processing of the input from stdin, PTY or file from the VMM thread to the existing virtio-console thread. The handling of the resize of a virtio-console has not changed but the name of the struct used to support that has been renamed to reflect its usage. Fixes: #3060 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-02 21:17:33 +01:00
Henry Wang	0d01eac1d4	vmm: Do the downcast of GicDevice in a safer way for AArch64 Downcasting of GicDevice trait might fail. Therefore we try to downcast the trait first and only if the downcasting succeeded we can then use the object to call methods. Otherwise, do nothing and log the failure. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-09-02 15:18:41 +01:00
Henry Wang	46c60183cd	arch, vmm: Implement GIC Pausable trait This commit implements the GIC (including both GICv3 and GICv3ITS) Pausable trait. The pause of device manager will trigger a "pause" of GIC, where we flush GIC pending tables and ITS tables to the guest RAM. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-09-02 15:18:41 +01:00
Rob Bradford	66f0b5b2b6	vmm: Open the serial PTY in non-blocking mode This prevents the boot of the guest kernel from being blocked by blocking I/O on the serial output since the data will be buffered into the SerialBuffer. Fixes: #3004 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-02 13:52:18 +01:00
Rob Bradford	d92707afc5	vmm: Introduce a SerialBuffer for buffering serial output Introduce a dynamic buffer for storing output from the serial port. The SerialBuffer implements std::io::Write and can be used in place of the direct output for the serial device. The internals of the buffer is a vector that grows dynamically based on demand up to a fixed size at which point old data will be overwritten. Currently the buffer is only flushed upon writes. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-09-02 13:52:18 +01:00
Alyssa Ross	9a634f07cb	build: update Cargo for rust-vmm branch renames The rust-vmm crates we're pulling from git have renamed their main branches. We need to update the branch names we're giving to Cargo, or people who don't have these dependencies cached will get errors like this when trying to build: error: failed to get `vm-fdt` as a dependency of package `arch v0.1.0 (/home/src/cloud-hypervisor/arch)` Caused by: failed to load source for dependency `vm-fdt` Caused by: Unable to update https://github.com/rust-vmm/vm-fdt?branch=master#031572a6 Caused by: object not found - no match for id (031572a6edc2f566a7278f1e17088fc5308d27ab); class=Odb (9); code=NotFound (-3) Signed-off-by: Alyssa Ross <hi@alyssa.is>	2021-09-02 10:38:25 +01:00
dependabot[bot]	b30a95f69a	build: bump signal-hook from 0.3.9 to 0.3.10 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.3.9 to 0.3.10. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/compare/v0.3.9...v0.3.10) --- updated-dependencies: - dependency-name: signal-hook dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-09-01 00:56:16 +00:00
Rob Bradford	63637eba31	vmm: Simplify epoll handling for VMM main loop Remove the indirection of a dispatch table and simply use the enum as the event data for the events. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-31 21:30:11 +01:00
dependabot[bot]	8841e63e2d	build: bump thiserror from 1.0.26 to 1.0.28 Bumps [thiserror](https://github.com/dtolnay/thiserror) from 1.0.26 to 1.0.28. - [Release notes](https://github.com/dtolnay/thiserror/releases) - [Commits](https://github.com/dtolnay/thiserror/compare/1.0.26...1.0.28) --- updated-dependencies: - dependency-name: thiserror dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-30 06:02:02 +00:00
dependabot[bot]	e877718b29	build: bump serde from 1.0.129 to 1.0.130 Bumps [serde](https://github.com/serde-rs/serde) from 1.0.129 to 1.0.130. - [Release notes](https://github.com/serde-rs/serde/releases) - [Commits](https://github.com/serde-rs/serde/compare/v1.0.129...v1.0.130) --- updated-dependencies: - dependency-name: serde dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-30 05:23:16 +00:00
dependabot[bot]	b0d8b50b36	build: bump serde_derive from 1.0.129 to 1.0.130 Bumps [serde_derive](https://github.com/serde-rs/serde) from 1.0.129 to 1.0.130. - [Release notes](https://github.com/serde-rs/serde/releases) - [Commits](https://github.com/serde-rs/serde/compare/v1.0.129...v1.0.130) --- updated-dependencies: - dependency-name: serde_derive dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-30 00:23:14 +00:00
dependabot[bot]	2349b7e753	build: bump serde_json from 1.0.66 to 1.0.67 Bumps [serde_json](https://github.com/serde-rs/json) from 1.0.66 to 1.0.67. - [Release notes](https://github.com/serde-rs/json/releases) - [Commits](https://github.com/serde-rs/json/compare/v1.0.66...v1.0.67) --- updated-dependencies: - dependency-name: serde_json dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-29 23:28:14 +00:00
dependabot[bot]	8fae21c10c	build: bump arc-swap from 1.3.1 to 1.3.2 Bumps [arc-swap](https://github.com/vorner/arc-swap) from 1.3.1 to 1.3.2. - [Release notes](https://github.com/vorner/arc-swap/releases) - [Changelog](https://github.com/vorner/arc-swap/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/arc-swap/compare/v1.3.1...v1.3.2) --- updated-dependencies: - dependency-name: arc-swap dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-29 22:30:31 +00:00
Bo Chen	b82bb55927	vmm: openapi: use the right default values This patch fixes couple of typos for the default values from the openapi yaml file. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-08-27 15:58:23 +01:00
dependabot[bot]	f840335922	build: bump libc from 0.2.100 to 0.2.101 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.100 to 0.2.101. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.100...0.2.101) --- updated-dependencies: - dependency-name: libc dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-26 05:11:06 +00:00
Rob Bradford	4d2a4e2805	vmm: Handle epoll events for PTYs separately Use two separate events for the console and serial PTY and then drive the handling of the inputs on the PTY separately. This results in the correct behaviour when both console and serial are attached to the PTY as they are triggered separately on the epoll so events are not lost. Fixes: #3012 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-25 13:33:32 +01:00
Rob Bradford	6233f6f68e	vmm: Send tty input to correct destination Check the config to find out which device is attached to the tty and then send the input from the user into that device (serial or virtio-console.) Fixes: #3005 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-25 10:08:25 +01:00
dependabot[bot]	8f6a5f979d	build: bump serde from 1.0.127 to 1.0.129 Bumps [serde](https://github.com/serde-rs/serde) from 1.0.127 to 1.0.129. - [Release notes](https://github.com/serde-rs/serde/releases) - [Commits](https://github.com/serde-rs/serde/compare/v1.0.127...v1.0.129) --- updated-dependencies: - dependency-name: serde dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-24 07:16:49 +00:00
dependabot[bot]	b8b16c6eec	build: bump serde_derive from 1.0.127 to 1.0.129 Bumps [serde_derive](https://github.com/serde-rs/serde) from 1.0.127 to 1.0.129. - [Release notes](https://github.com/serde-rs/serde/releases) - [Commits](https://github.com/serde-rs/serde/compare/v1.0.127...v1.0.129) --- updated-dependencies: - dependency-name: serde_derive dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-24 06:50:04 +00:00
dependabot[bot]	a969d5016c	build: bump libc from 0.2.99 to 0.2.100 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.99 to 0.2.100. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.99...0.2.100) --- updated-dependencies: - dependency-name: libc dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-24 05:53:20 +00:00
dependabot[bot]	722523f925	build: bump arc-swap from 1.3.0 to 1.3.1 Bumps [arc-swap](https://github.com/vorner/arc-swap) from 1.3.0 to 1.3.1. - [Release notes](https://github.com/vorner/arc-swap/releases) - [Changelog](https://github.com/vorner/arc-swap/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/arc-swap/compare/v1.3.0...v1.3.1) --- updated-dependencies: - dependency-name: arc-swap dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-24 04:19:39 +00:00
Fazla Mehrab	5db4dede28	block_util, vhdx: vhdx crate integration with the cloud hypervisor vhdx_sync.rs in block_util implements traits to represent the vhdx crate as a supported block device in the cloud hypervisor. The vhdx is added to the block device list in device_manager.rs at the vmm crate so that it can automatically detect a vhdx disk and invoke the corresponding crate. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Fazla Mehrab <akm.fazla.mehrab@intel.com>	2021-08-19 11:43:19 +02:00
Bo Chen	9aba1fdee6	virtio-devices, vmm: Use syscall definitions from the libc crate Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-08-18 10:42:19 +02:00
Bo Chen	864a5e4fe0	virtio-devices, vmm: Simplify 'get_seccomp_rules' Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-08-18 10:42:19 +02:00
Bo Chen	7d38a1848b	virtio-devices, vmm: Fix the '--seccomp false' option We are relying on applying empty 'seccomp' filters to support the '--seccomp false' option, which will be treated as an error with the updated 'seccompiler' crate. This patch fixes this issue by explicitly checking whether the 'seccomp' filter is empty before applying the filter. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-08-18 10:42:19 +02:00
Bo Chen	08ac3405f5	virtio-devices, vmm: Move to the seccompiler crate Fixes: #2929 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-08-18 10:42:19 +02:00
dependabot[bot]	86b2c17135	build: bump anyhow from 1.0.42 to 1.0.43 Bumps [anyhow](https://github.com/dtolnay/anyhow) from 1.0.42 to 1.0.43. - [Release notes](https://github.com/dtolnay/anyhow/releases) - [Commits](https://github.com/dtolnay/anyhow/compare/1.0.42...1.0.43) --- updated-dependencies: - dependency-name: anyhow dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-17 00:32:44 +00:00
dependabot[bot]	754ce37031	build: bump bitflags from 1.3.1 to 1.3.2 Bumps [bitflags](https://github.com/bitflags/bitflags) from 1.3.1 to 1.3.2. - [Release notes](https://github.com/bitflags/bitflags/releases) - [Changelog](https://github.com/bitflags/bitflags/blob/main/CHANGELOG.md) - [Commits](https://github.com/bitflags/bitflags/compare/1.3.1...1.3.2) --- updated-dependencies: - dependency-name: bitflags dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-16 23:54:21 +00:00
Rob Bradford	9d35a10fd4	vmm: cpu: Shutdown VMM on vCPU thread panic If the vCPU thread panics then catch it and trigger the shutdown of the VMM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-13 09:19:54 +02:00
Rob Bradford	53b2e19934	vmm: Add support for hotplugging user devices Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-12 13:19:04 +01:00
dependabot[bot]	f99462add4	build: bump bitflags from 1.2.1 to 1.3.1 Bumps [bitflags](https://github.com/bitflags/bitflags) from 1.2.1 to 1.3.1. - [Release notes](https://github.com/bitflags/bitflags/releases) - [Changelog](https://github.com/bitflags/bitflags/blob/main/CHANGELOG.md) - [Commits](https://github.com/bitflags/bitflags/compare/1.2.1...1.3.1) --- updated-dependencies: - dependency-name: bitflags dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2021-08-12 09:09:39 +00:00
Henry Wang	bcae6c41e3	vmm, doc: Forbid same memory zone in multiple NUMA nodes It is forbidden that the same memory zone belongs to more than one NUMA node. This commit adds related validation to the `--numa` parameter to prevent the user from specifying such configuration. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-08-12 10:49:02 +02:00
Henry Wang	5a0a4bc505	arch: Add optional `distance-map` node to FDT The optional device tree node distance-map describes the relative distance (memory latency) between all NUMA nodes. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-08-12 10:49:02 +02:00
Henry Wang	165364e08b	vmm: Move NUMA node data structures to `arch` This is to make sure the NUMA node data structures can be accessed both from the `vmm` crate and `arch` crate. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-08-12 10:49:02 +02:00
Henry Wang	20aa811de7	vmm: Extend NUMA setup to more than ACPI The AArch64 platform provides a NUMA binding for the device tree, which means on AArch64 platform, the NUMA setup can be extended to more than the ACPI feature. Based on above, this commit extends the NUMA setup and data structures to following scenarios: - All AArch64 platform - x86_64 platform with ACPI feature enabled Signed-off-by: Henry Wang <Henry.Wang@arm.com> Signed-off-by: Michael Zhao <Michael.Zhao@arm.com>	2021-08-12 10:49:02 +02:00
Sebastien Boeuf	4918c1ca7f	block_util, vmm: Propagate error on QcowDiskSync creation Instead of panicking with an expect() function, the QcowDiskSync::new function now propagates the error properly. This ensures the VMM will not panic, which might be the source of weird errors if only one thread exits while the VMM continues to run. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-11 16:44:28 -07:00
Sebastien Boeuf	4735cb8563	vmm, virtio-devices: Restore vhost-user devices in a dedicated way We cannot let vhost-user devices connect to the backend when the Block, Fs or Net object is being created during a restore/migration. The reason is we can't have two VMs (source and destination) connected to the same backend at the same time. That's why we must delay the connection with the vhost-user backend until the restoration is performed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-10 12:36:58 -07:00
Sebastien Boeuf	71c7dff32b	vmm: Fix the error handling logic when migration fails The code wasn't doing what it was expected to. The '?' was simply returning the error to the top level function, meaning the Err() case in the match was never hit. Moving the whole logic to a dedicated function allows to identify when something got wrong without propagating to the calling function, so that we can still stop the dirty logging and unpause the VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-10 12:36:58 -07:00
Sebastien Boeuf	db444715fd	vmm: Shutdown VM after migration succeeded In case the migration succeeds, the destination VM will be correctly running, with potential vhost-user backends attached to it. We can't let the source VM trying to reconnect to the same backends, which is why it's safer to shutdown the source VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-10 12:36:58 -07:00
Sebastien Boeuf	5a83ebce64	vmm: Notify Migratable objects about migration being complete Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-10 12:36:58 -07:00
Sebastien Boeuf	06729bb3ba	vmm: Provide a restoring state to the DeviceManager In anticipation for creating vhost-user devices in a different way when being restored compared to a fresh start, this commit introduces a new boolean created by the Vm depending on the use case, and passed down to the DeviceManager. In the future, the DeviceManager will use this flag to assess how vhost-user devices should be created. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-10 12:36:58 -07:00
Rob Bradford	5e74848ab4	vmm: seccomp: Permit syscalls used for vfio-user on vCPU thread Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-10 16:01:00 +01:00
Rob Bradford	3efccd0fef	vmm: config: Ensure shared memory is enabled if using user-devices Correct operation of user devices (vfio-user) requires shared memory so flag this to prevent it from failing in strange ways. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-10 16:01:00 +01:00
Rob Bradford	b28063a7b4	vmm: Create user devices from config Create the vfio-user / user devices from the config. Currently hotplug of the devices is not supported nor can they be placed behind the (virt-)iommu. Removal of the coldplugged device is however supported. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-10 16:01:00 +01:00
Rob Bradford	7fbec7113e	main, config: Add support for `--user-device` This allows the user to specify devices that are running in a different userspace process and communicated with vfio-user. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-10 16:01:00 +01:00
Rob Bradford	77e147f333	build: Bump dependencies This has the side effect of also removing the vm-memory 0.5.0 dependency. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-10 15:24:28 +01:00
Markus Theil	5b0d4bb398	virtio-devices: seccomp: allow unix socket connect in vsock thread Allow vsocks to connect to Unix sockets on the host running cloud-hypervisor with enabled seccomp. Reported-by: Philippe Schaaf <philippe.schaaf@secunet.com> Tested-by: Franz Girlich <franz.girlich@tu-ilmenau.de> Signed-off-by: Markus Theil <markus.theil@tu-ilmenau.de>	2021-08-06 08:44:47 -07:00
Rob Bradford	f7f2f25a57	build: Use fixed versions in Cargo.toml files This doesn't really affect the build as we ship a Cargo.lock with fixed versions in. However for clarity it makes sense to use fixed versions throughout and let dependabot update them. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-06 12:11:39 +02:00
Rob Bradford	b4f887ea80	build: Move from patched vm-memory version to released version Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-08-06 10:08:58 +01:00
Henry Wang	27a285257e	vmm: cpu: Add PPTT table for AArch64 The optional Processor Properties Topology Table (PPTT) table is used to describe the topological structure of processors controlled by the OSPM, and their shared resources, such as caches. The table can also describe additional information such as which nodes in the processor topology constitute a physical package. The ACPI PPTT table supports topology descriptions for ACPI guests. Therefore, this commit adds the PPTT table for AArch64 to enable CPU topology feature for ACPI. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-08-05 21:19:16 +08:00
Henry Wang	7fb980f17b	arch, vmm: Pass cpu topology configuation to FDT In an Arm system, the hierarchy of CPUs is defined through three entities that are used to describe the layout of physical CPUs in the system: - cluster - core - thread All these three entities have their own FDT node field. Therefore, This commit adds an AArch64-specific helper to pass the config from the Cloud Hypervisor command line to the `configure_system`, where eventually the `create_fdt` is called. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-08-05 21:19:16 +08:00
Sebastien Boeuf	5c6139bbff	vmm: Finalize migration support for all devices Make sure the DeviceManager is triggered for all migration operations. The dirty pages are merged from MemoryManager and DeviceManager before to be sent up to the Vmm in lib.rs. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-05 06:07:00 -07:00
Sebastien Boeuf	0411064271	vmm: Refactor migration through Migratable trait Now that Migratable provides the methods for starting, stopping and retrieving the dirty pages, we move the existing code to these new functions. No functional change intended. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-05 06:07:00 -07:00
Sebastien Boeuf	e9637d3733	vmm: device_manager: Fully implement Migratable trait This patch connects the dots between the vm.rs code and each Migratable device, in order to make sure Migratable methods are correctly invoked when migration happens. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-05 06:07:00 -07:00
Sebastien Boeuf	79425b6aa8	vm-migration, vmm: Extend methods for MemoryRangeTable In anticipation for supporting the merge of multiple dirty pages coming from multiple devices, this patch factorizes the creation of a MemoryRangeTable from a bitmap, as well as providing a simple method for merging the dirty pages regions under a single MemoryRangeTable. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-08-05 06:07:00 -07:00
Muminul Islam	83c44a2411	vmm, virtio-devices: Add missing seccomp rules for MSHV This patch adds all the seccomp rules missing for MSHV. With this patch MSFT internal CI runs with seccomp enabled. Signed-off-by: Muminul Islam <muislam@microsoft.com>	2021-08-03 11:09:07 -07:00
Bo Chen	902fe20d41	vmm: Add fallback handling for sending live migration This patch adds a fallback path for sending live migration, where it ensures the following behavior of source VM post live-migration: 1. The source VM will be paused only when the migration is completed successfully, or otherwise it will keep running; 2. The source VM will always stop dirty pages logging. Fixes: #2895 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-08-03 09:26:12 +01:00
Muminul Islam	3baa0c3721	vmm: Add MSHV_VP_TRANSLATE_GVA to seccomp rule This rule is needed to boot windows guest. This bug was introduced while we tried to boot windows guest on MSHV. Signed-off-by: Muminul Islam <muislam@microsoft.com>	2021-07-29 16:29:53 +01:00
Muminul Islam	81895b9b40	hypervisor: Implement start/stop_dirty_log for MSHV This patch modify the existing live migration code to support MSHV. Adds couple of new functions to enable and disable dirty page tracking. Add missing IOCTL to the seccomp rules for live migration. Adds necessary flags for MSHV. This changes don't affect KVM functionality at all. In order to get better performance it is good to enable dirty page tracking when we start live migration and disable it when the migration is done. Signed-off-by: Muminul Islam <muislam@microsoft.com>	2021-07-29 16:29:53 +01:00
Muminul Islam	fdecba6958	hypervisor: MSHV needs gpa to retrieve dirty logs Right now, get_dirty_log API has two parameters, slot and memory_size. MSHV needs gpa to retrieve the page states. GPA is needed as MSHV returns the state base on PFN. Signed-off-by: Muminul Islam <muislam@microsoft.com>	2021-07-29 16:29:53 +01:00
Sebastien Boeuf	12db6e5068	vmm: Allow restoring virtio-fs with no cache region It's totally acceptable to snapshot and restore a virtio-fs device that has no cache region, since this is a valid mode of functioning for virtio-fs itself. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-29 06:35:03 -07:00
Sebastien Boeuf	dcc646f5b1	clippy: Fix redundant allocations With the new beta version, clippy complains about redundant allocation when using Arc<Box<dyn T>>, and suggests replacing it simply with Arc<dyn T>. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-29 13:28:57 +02:00
Bo Chen	b00a6a8519	vmm: Create guest memory regions with explicit dirty-pages-log flags As we are now using an global control to start/stop dirty pages log from the `hypervisor` crate, we need to explicitly tell the hypervisor (KVM) whether a region needs dirty page tracking when it is created. This reverts commit `f063346de3`. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-07-28 09:08:32 -07:00
Bo Chen	e7c9954dc1	hypervisor, vmm: Abstract the interfaces to start/stop dirty log Following KVM interfaces, the `hypervisor` crate now provides interfaces to start/stop the dirty pages logging on a per region basis, and asks its users (e.g. the `vmm` crate) to iterate over the regions that needs dirty pages log. MSHV only has a global control to start/stop dirty pages log on all regions at once. This patch refactors related APIs from the `hypervisor` crate to provide a global control to start/stop dirty pages log (following MSHV's behaviors), and keeps tracking the regions need dirty pages log for KVM. It avoids leaking hypervisor-specific behaviors out of the `hypervisor` crate. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-07-28 09:08:32 -07:00
Bo Chen	ca09638491	vmm: Add CPUID compatibility check for snapshot/restore Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-07-28 09:26:02 +02:00
Bo Chen	0835198ddd	vmm: Factorize CPUID check for live-migration and snapshot/restore This patch adds a common function "Vmm::vm_check_cpuid_compatibility()" to be shared by both live-migration and snapshot/restore. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-07-28 09:26:02 +02:00
Bo Chen	6d9c1eb638	arch, vmm: Add CPUID check to the 'Config' step of live migration We now send not only the 'VmConfig' at the 'Command::Config' step of live migration, but also send the 'common CPUID'. In this way, we can check the compatibility of CPUID features between the source and destination VMs, and abort live migration early if needed. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-07-28 09:26:02 +02:00
Bo Chen	f063346de3	vmm: Create guest memory regions without dirty-pages-log by default With the support of dynamically turning on/off dirty-pages-log during live-migration (only for guest RAM regions), we now can create guest memory regions without dirty-pages-log by default both for guest RAM regions and other regions backed by file/device. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-07-26 09:19:35 -07:00
Bo Chen	5e0d498582	hypervisor, vmm: Add dynamic control of logging dirty pages This patch extends slightly the current live-migration code path with the ability to dynamically start and stop logging dirty-pages, which relies on two new methods added to the `hypervisor::vm::Vm` Trait. This patch also contains a complete implementation of the two new methods based on `kvm` and placeholders for `mshv` in the `hypervisor` crate. Fixes: #2858 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-07-26 09:19:35 -07:00
dependabot[bot]	49c72beda5	build: bump seccomp from v0.24.4 to v0.24.5 Bumps [seccomp](https://github.com/firecracker-microvm/firecracker) from v0.24.4 to v0.24.5. - [Release notes](https://github.com/firecracker-microvm/firecracker/releases) - [Changelog](`cd36c699f3/CHANGELOG.md`) - [Commits](`8f44986a0e...cd36c699f3`) --- updated-dependencies: - dependency-name: seccomp dependency-type: direct:production ... Signed-off-by: dependabot[bot] <support@github.com>	2021-07-26 11:18:42 +02:00
Sebastien Boeuf	0ac4545c5b	vmm: Extend seccomp filters with fcntl() for HTTP thread Whenever a file descriptor is sent through the control message, it requires fcntl() syscall to handle it, meaning we must allow it through the list of syscalls authorized for the HTTP thread. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-21 15:34:22 +02:00
Sebastien Boeuf	3e482c9c74	vmm: Limit physical address space for TDX When running TDX guest, the Guest Physical Address space is limited by a shared bit that is located on bit 47 for 4 level paging, and on bit 51 for 5 level paging (when GPAW bit is 1). In order to keep things simple, and since a 47 bits address space is 128TiB large, we ensure to limit the physical addressable space to 47 bits when runnning TDX. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-20 15:00:04 +02:00
Sebastien Boeuf	05f7651cf5	vmm: Force VIRTIO_F_IOMMU_PLATFORM when running TDX When running a TDX guest, we need the virtio drivers to use the DMA API to share specific memory pages with the VMM on the host. The point is to let the VMM get access to the pages related to the buffers pointed by the virtqueues. The way to force the virtio drivers to use the DMA API is by exposing the virtio devices with the feature VIRTIO_F_IOMMU_PLATFORM. This is a feature indicating the device will require some address translation, as it will not deal directly with physical addresses. Cloud Hypervisor takes care of this requirement by adding a generic parameter called "force_iommu". This parameter value is decided based on the "tdx" feature gate, and then passed to the DeviceManager. It's up to the DeviceManager to use this parameter on every virtio device creation, which will imply setting the VIRTIO_F_IOMMU_PLATFORM feature. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-20 14:47:01 +02:00
Bo Chen	569be6e706	arch, vmm: Move "generate_common_cpuid" from "CpuManager" to "arch" This refactoring ensures all CPUID related operations are centralized in `arch::x86_64` module, and exposes only two related public functions to the vmm crate, e.g. `generate_common_cpuid` and `configure_vcpu`. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-07-19 09:59:34 -07:00
Sebastien Boeuf	d4316d0228	vmm: http: Allow file descriptor to be sent with add-net In order to let a separate process open a TAP device and pass the file descriptor through the control message mechanism, this patch adds the support for sending a file descriptor over to the Cloud Hypervisor process along with the add-net HTTP API command. The implementation uses the NetConfig structure mutably to update the list of fds with the one passed through control message. The list should always be empty prior to this, as it makes no sense to provide a list of fds once the Cloud Hypervisor process has already been started. It is important to note that reboot is supported since the file descriptor is duplicated upon receival, letting the VM only use the duplicated one. The original file descriptor is kept open in order to support a potential reboot. Fixes #2525 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-19 09:51:32 +02:00
Muminul Islam	3937e03c02	vmm, virtio-devices: Extend mshv feature There are some seccomp rules needed for MSHV in virtio-devices but not for KVM. We only want to add those rules based on MSHV feature guard. Signed-off-by: Muminul Islam <muislam@microsoft.com>	2021-07-15 11:05:11 -07:00
Sebastien Boeuf	d68c388cac	vmm: Update seccomp filters for HTTP thread The micro-http crate now uses recvmsg() syscall in order to receive file descriptors through control messages. This means the syscall must be part of the authorized list in the seccomp filters. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-15 08:13:48 +00:00
Wei Liu	39bc444db4	vmm, vm-device: make use of the kvm feature gate in vfio-ioctls The vfio-ioctls crate now contains a KVM feature gate. Make use of it in Cloud Hypervisor. That crate has two users. For the vmm crate is it straight-forward. For the vm-device crate, we introduce a KVM feature gate as well so that the vmm crate can pass on the configuration. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-07-15 09:35:51 +02:00
Sebastien Boeuf	6b710209b1	numa: Add optional `sgx_epc_sections` field to NumaConfig This new option allows the user to define a list of SGX EPC sections attached to a specific NUMA node. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-09 14:45:30 +02:00
Sebastien Boeuf	9aedabe11e	sgx: Add mandatory `id` field to SgxEpcConfig In order to uniquely identify each SGX EPC section, we introduce a mandatory option `id` to the `--sgx-epc` parameter. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-09 14:45:30 +02:00
dependabot[bot]	5effa20a5b	build: bump libc from 0.2.97 to 0.2.98 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.97 to 0.2.98. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.97...0.2.98) --- updated-dependencies: - dependency-name: libc dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-07-08 04:15:16 +00:00
Sebastien Boeuf	17c99ae00a	vmm: Enable provisioning for SGX guest The guest can see that SGX supports provisioning as it is exposed through the CPUID. This patch enables the proper backing of this feature by having the host open the provisioning device and enable this capability through the hypervisor. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-07 14:56:38 +02:00
Sebastien Boeuf	5b6d424a77	arch, vmm: Fix TDVF section handling This patch fixes a few things to support TDVF correctly. The HOB memory resources must contain EFI_RESOURCE_ATTRIBUTE_ENCRYPTED attribute. Any section with a base address within the already allocated guest RAM must not be allocated. The list of TD_HOB memory resources should contain both TempMem and TdHob sections as well. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-07-06 11:47:43 +02:00
Henry Wang	4da3bdcd6e	vmm: Split restore device_manager and devices Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-07-05 22:51:56 +02:00
Henry Wang	95ca4fb15e	vmm: vm: Enable snapshot/restore of GICv3ITS This commit enables the snapshot/restore of GICv3ITS in the process of VM snapshot/restore. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-07-05 22:51:56 +02:00
Wei Liu	1f2915bff0	vmm: hypervisor: split set_user_memory_region to two functions Previously the same function was used to both create and remove regions. This worked on KVM because it uses size 0 to indicate removal. MSHV has two calls -- one for creation and one for removal. It also requires having the size field available because it is not slot based. Split set_user_memory_region to {create/remove}_user_memory_region. For KVM they still use set_user_memory_region underneath, but for MSHV they map to different functions. This fixes user memory region removal on MSHV. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-07-05 09:45:45 +02:00
Wei Liu	71bbaf556f	vmm: seccomp: add seccomp rules for MSHV Add a minimum set of rules that allow Cloud Hypervisor to run Linux on top of Microsoft Hypervisor. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-07-05 09:44:02 +02:00
Wei Liu	8819bb0f21	vmm: seccomp: make use of KVM feature The to-be-introduced MSHV rules don't need to contain KVM rules and vice versa. Put KVM constants into to a module. This avoids the warnings about dead code in the future. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-07-05 09:44:02 +02:00
Henry Wang	054c036e81	vmm: acpi: Add AArch64 vCPUs to SRAT table This commit introduces the `ProcessorGiccAffinity` struct for the AArch64 platform. This struct will be created and included into the SRAT table to enable AArch64 NUMA setup. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-06-25 10:22:40 +01:00
Michael Zhao	239e39ddbc	vmm: Fix clippy warnings on AArch64 Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-06-24 08:59:53 -07:00
Bo Chen	5768dcc320	vmm: Refactor slightly `vm_boot` and 'control_loop' It ensures all handlers for `ApiRequest` in `control_loop` are consistent and minimum and should read better. No functional changes. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-06-24 16:01:39 +02:00
Bo Chen	1075209e2a	vmm: Handle ApiRequest::VmCreate in a separate function It simplifies a bit the `Vmm::control_loop` and reads better to be consistent with other `ApiRequest` handlers. Also, it removes the repetitive `ApiError::VmAlreadyCreated` and makes `ApiError::VmCreate` useful. No functional changes. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-06-24 16:01:39 +02:00
Michael Zhao	3613b4c096	aarch64: Enable default build option We have been building Cloud Hypervisor with command like: `cargo build --no-default-features --features ...`. After implementing ACPI, we donot have to use specify all features explicitly. Default build command `cargo build` can work. This commit fixed some build warnings with default build option and changed github workflow correspondingly. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-06-24 13:13:27 +01:00
Bo Chen	c4be0f4235	clippy: Address the issue 'needless-collect' error: avoid using `collect()` when not needed --> vmm/src/vm.rs:630:86 \| 630 \| let node_id_list: Vec<u32> = configs.iter().map(\|cfg\| cfg.guest_numa_id).collect(); \| ^^^^^^^ ... 664 \| if !node_id_list.contains(&dest) { \| ---------------------------- the iterator could be used here instead \| = note: `-D clippy::needless-collect` implied by `-D warnings` = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#needless_collect Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-06-24 08:55:43 +02:00
Bo Chen	5825ab2dd4	clippy: Address the issue 'needless-borrow' Issue from beta verion of clippy: Error: --> vm-virtio/src/queue.rs:700:59 \| 700 \| if let Some(used_event) = self.get_used_event(&mem) { \| ^^^^ help: change this to: `mem` \| = note: `-D clippy::needless-borrow` implied by `-D warnings` = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#needless_borrow Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-06-24 08:55:43 +02:00
Bo Chen	585269ecb9	clippy: Address the issue 'field is never read' Issue from beta verion of clippy: error: field is never read: `type` --> vmm/src/cpu.rs:235:5 \| 235 \| pub r#type: u8, \| ^^^^^^^^^^^^^^ \| = note: `-D dead-code` implied by `-D warnings` Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-06-24 08:55:43 +02:00
Rob Bradford	4d25eaa24a	vmm: Add I/O port range to PCI bus resources The Linux kernel expects that any PCI devices that advertise I/O bars have use an address that is within the range advertised by the bus itself. Unfortunately we were not advertising any I/O ports associated with the PCI bus in the ACPI tables. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-06-23 16:48:52 +01:00
Rob Bradford	b56e1217b6	vmm: tdx: Add KVM_FEATURE_STEAL_TIME_BIT to filtered bits Filter out the KVM_FEATURE_STEAL_TIME_BIT when running with TDX. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-06-18 15:54:10 +01:00
Sebastien Boeuf	a36ac96444	vmm: cpu_manager: Add _PXM ACPI method to each vCPU In order to allow a hotplugged vCPU to be assigned to the correct NUMA node in the guest, the DSDT table must expose the _PXM method for each vCPU. This method defines the proximity domain to which each vCPU should be attached to. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-06-17 16:08:46 +02:00
Sebastien Boeuf	09c3ddd47d	vmm: memory_manager: Remove _PXM from ACPI memory slot The _PXM method always return 0, which is wrong since the SRAT might tell differently. The point of the _PXM method is to be evaluated by the guest OS when some new memory slot is being plugged, but this will never happen for Cloud Hypervisor since using NUMA nodes along with memory hotplug only works for virtio-mem. Memory hotplug through ACPI will only happen when there's only one NUMA node exposed to the guest, which means the _PXM method won't be needed at all. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-06-17 16:08:46 +02:00
Sebastien Boeuf	07f3075773	vmm: device_manager: Tie PCI bus to NUMA node 0 Make sure the unique PCI bus is tied to the default NUMA node 0, and update the documentation to let the users know about this special case. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-06-17 16:08:46 +02:00
Fei Li	aa27f0e743	virtio-balloon: add deflate_on_oom support Sometimes we need balloon deflate automatically to give memory back to guest, especially for some low priority guest processes under memory pressure. Enable deflate_on_oom to support this. Usage: --balloon "size=0,deflate_on_oom=on" \ Signed-off-by: Fei Li <lifei.shirley@bytedance.com>	2021-06-16 09:55:22 +02:00
Sebastien Boeuf	a6fe4aa7e9	virtio-devices, vmm: Update virtio-iommu to rely on VIOT Since using the VIRTIO configuration to expose the virtual IOMMU topology has been deprecated, the virtio-iommu implementation must be updated. In order to follow the latest patchset that is about to be merged in the upstream Linux kernel, it must rely on ACPI, and in particular the newly introduced VIOT table to expose the information about the list of PCI devices attached to the virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-06-15 17:05:59 +02:00
dependabot[bot]	428c637506	build: bump libc from 0.2.96 to 0.2.97 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.96 to 0.2.97. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.96...0.2.97) --- updated-dependencies: - dependency-name: libc dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-06-14 09:50:38 +00:00
Michael Zhao	1a7a76511f	vmm: Enable pty console on AArch64 Allowed syscall "SYS_readlinkat" on AArch64 which is required by pty. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-06-10 15:05:25 +02:00
Jianyong Wu	b8b5dccfd8	aarch64: Enable UEFI image loading Implemented an architecture specific function for loading UEFI binary. Changed the logic of loading kernel image: 1. First try to load the image as kernel in PE format; 2. If failed, try again to load it as formatless UEFI binary. Signed-off-by: Jianyong Wu <jianyong.wu@arm.com>	2021-06-09 18:36:59 +08:00
Jianyong Wu	6880692a78	vmm, acpi: Add DSM method to ACPI _DSM (Device Specific Method) is a control method that enables devices to provide device specific control functions. Linux kernel will evaluate this device then initialize preserve_config in acpi pci initialization. Signed-off-by: Jianyong Wu <jianyong.wu@arm.com>	2021-06-09 18:36:59 +08:00
dependabot[bot]	c866ccb3a1	build: bump libc from 0.2.95 to 0.2.96 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.95 to 0.2.96. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.95...0.2.96) --- updated-dependencies: - dependency-name: libc dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-06-09 07:27:37 +00:00
Rob Bradford	3dc15a9259	vmm: tdx: Don't access same locked structure twice Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-06-03 17:29:05 +02:00
Bo Chen	7839e121f6	vmm: Add dirty pages tracked by vm_memory::bitmap to live migration Live migration currently handles guest memory writes from the guest through the KVM dirty page tracking and sends those dirty pages to the destination. This patch augments the live migration support with dirty page tracking of writes from the VMM to the guest memory(e.g. virtio devices). Fixes: #2458 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-06-03 08:34:45 +01:00
Bo Chen	2c4fa258a6	virtio-devices, vmm: Deprecate "GuestMemory::with_regions(_mut)" Function "GuestMemory::with_regions(_mut)" were mainly temporary methods to access the regions in `GuestMemory` as the lack of iterator-based access, and hence they are deprecated in the upstream vm-memory crate [1]. [1] https://github.com/rust-vmm/vm-memory/issues/133 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-06-03 08:34:45 +01:00
Bo Chen	b5bcdbaf48	misc: Upgrade to use the vm-memory crate w/ dirty-page-tracking As the first step to complete live-migration with tracking dirty-pages written by the VMM, this commit patches the dependent vm-memory crate to the upstream version with the dirty-page-tracking capability. Most changes are due to the updated `GuestMemoryMmap`, `GuestRegionMmap`, and `MmapRegion` structs which are taking an additional generic type parameter to specify what 'bitmap backend' is used. The above changes should be transparent to the rest of the code base, e.g. all unit/integration tests should pass without additional changes. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-06-03 08:34:45 +01:00
dependabot[bot]	f3d3d8daed	build: bump signal-hook from 0.3.8 to 0.3.9 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.3.8 to 0.3.9. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/compare/v0.3.8...v0.3.9) --- updated-dependencies: - dependency-name: signal-hook dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2021-06-02 07:48:03 +00:00
Rob Bradford	c357adae44	vmm: tdx: Clear unsupported KVM PV features This matches with the features that QEMU clears as they are not supported with TDX. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-06-01 23:00:54 +02:00
Michael Zhao	7f3fa39d81	vmm: Remove enable_interrupt_controller() After adding "get_interrupt_controller()" function in DeviceManager, "enable_interrupt_controller()" became redundant, because the latter one is the a simple wrapper on the interrupt controller. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-06-01 16:56:43 +01:00
Michael Zhao	9a5f3fc2a7	vmm: Remove "gicr" handling from DeviceManager The function used to calculate "gicr-typer" value has nothing with DeviceManager. Now it is moved to AArch64 specific files. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-06-01 16:56:43 +01:00
Michael Zhao	7932cd22ca	vmm: Remove GIC entity set/get from DeviceManager Moved the set/get functions from vmm::DeviceManager to devices::Gic. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-06-01 16:56:43 +01:00
Michael Zhao	195eba188a	vmm: Split create_gic() from configure_system() Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-06-01 16:56:43 +01:00
Sebastien Boeuf	e9cc23ea94	virtio-devices: vhost_user: net: Move control queue back We thought we could move the control queue to the backend as it was making some good sense. Unfortunately, doing so was a wrong design decision as it broke the compatibility with OVS-DPDK backend. This is why this commit moves the control queue back to the VMM side, meaning an additional thread is being run for handling the communication with the guest. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-05-26 16:09:32 +01:00
dependabot[bot]	6c245f6cf1	build: Manual seccomp bump Seccomp needs to be bumped in the main tree and fuzz at the same time. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-26 14:39:43 +02:00
dependabot[bot]	dd92715ed2	build: Bump libc from 0.2.94 to 0.2.95 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.94 to 0.2.95. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.94...0.2.95) Signed-off-by: dependabot[bot] <support@github.com>	2021-05-26 07:18:40 +00:00
Michael Zhao	ff46fb69d0	aarch64: Fix IRQ number setting for ACPI On FDT, VMM can allocate IRQ from 0 for devices. But on ACPI, the lowest range below 32 has to be avoided. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-05-25 10:20:37 +02:00
Henry Wang	efc583c13e	acpi: AArch64: Enable PSCI in FADT This commit enables the PSCI (Power State Coordination Interface) for the AArch64 platform, which allows the VMM to manage the power status of the guest. Also, multiple vCPUs can be brought up using PSCI. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-05-25 10:20:37 +02:00
Henry Wang	d882f8c928	acpi: Implement IORT for AArch64 This commit implements the IO Remapping Table (IORT) for AArch64. The IORT is one of the required ACPI table for AArch64, since it describes the GICv3ITS node. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-05-25 10:20:37 +02:00
Henry Wang	213da7d862	acpi: Implement SPCR on AArch64 This commit implements an AArch64-required ACPI table: Serial Port Console Redirection Table (SPCR). The table provides information about the configuration and use of the serial port or non-legacy UART interface. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-05-25 10:20:37 +02:00
Henry Wang	9c5528490e	acpi: Implement GTDT on AArch64 This commit implements an AArch64-specific ACPI table: Generic Timer Description Table (GTDT). The GTDT provides OSPM with information about a system’s Generic Timers configuration. The Generic Timer (GT) is a standard timer interface implemented on ARM processor-based systems. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-05-25 10:20:37 +02:00
Michael Zhao	7bfb51489b	acpi: Implement MCFG for AArch64 Added the final PCI bus number in MCFG table. This field is mandatory on AArch64. On X86 it is optional. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-05-25 10:20:37 +02:00
Michael Zhao	e1ef141112	acpi: Enable DSDT for CpuManager on AArch64 Simplified definition block of CPU's on AArch64. It is not complete yet. Guest boots. But more is to do in future: - Fix the error in ACPI definition blocks (seen in boot messages) - Implement CPU hot-plug controller Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-05-25 10:20:37 +02:00
Michael Zhao	5f27d649a6	acpi: Enable DSDT for DeviceManager on AArch64 Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-05-25 10:20:37 +02:00
Michael Zhao	cab103d172	acpi: Implement MADT on AArch64 Added following structures in MADT table: - GICC: GIC CPU interface structure - GICD: GIC Distributor structure - GICR: GIC Redistributor Structure - GICITS: GIC Interrupt Translation Service Structure Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-05-25 10:20:37 +02:00
Rob Bradford	f840327ffb	vmm: Version MemoryManager state Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-21 15:29:52 +02:00
Sebastien Boeuf	9a7199a116	virtio-devices: vhost_user: blk: Cleanup device creation Prepare the device creation so that it can be factorized in a follow up commit. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-05-21 12:03:54 +02:00
renlei4	65a39f43cb	vmm: support restore KVM clock in migration In migration, vm object is created by new_from_migration with NULL kvm clock. so vm.set_clock will not be called during vm resume. If the guest using kvm-clock, the ticks will be stopped after migration. As clock was already saved to snapshot, add a method to restore it before vm resume in migration. after that, guest's kvm-clock works well. Signed-off-by: Ren Lei <ren.lei4@zte.com.cn>	2021-05-20 14:32:49 +01:00
ren lei	324c924ca4	vmm: fix KVM clock lost during restore form snapshot Connecting a restored KVM clock vm will take long time, as clock is NOT restored immediately after vm resume from snapshot. this is because `9ce6c3b` incorrectly remove vm_snapshot.clock, and always pass None to new_from_memory_manager, which will result to kvm_set_clock() never be called during restore from snapshot. Fixes: `9ce6c3b` Signed-off-by: Ren Lei <ren.lei4@zte.com.cn>	2021-05-20 11:31:01 +02:00
Sebastien Boeuf	5d2df70a79	virtio-devices: vhost_user: net: Remove control queue Now that the control queue is correctly handled by the backend, there's no need to handle it as well from the VMM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-05-19 18:21:47 +02:00
Rob Bradford	7ae631e0c1	vmm: Hypervisor::check_required_extensions() is not x86_64 only This function is implemented for aarch64 so should be called on it. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-19 17:11:30 +02:00
Rob Bradford	0cf9218d3f	hypervisor, vmm: Add default Hypervisor::check_required_extensions() This allows the removal of KVM specific compile time checks on this function. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-19 17:11:30 +02:00
Rob Bradford	2439625785	hypervisor: Cleanup unused Hypervisor trait members Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-19 17:11:30 +02:00
Michael Zhao	01acea3102	acpi: Refactor create_acpi_tables() Split the code of create_acpi_tables() into multiple functions for easy extension in future. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-05-19 09:59:59 +02:00
Michael Zhao	e4bb6409ae	aarch64: Change memory layout for UEFI & ACPI Before this change, the FDT was loaded at the end of RAM. The address of FDT was not fixed. While UEFI (edk2 now) requires fixed address to find FDT and RSDP. Now the FDT is moved to the beginning of RAM, which is a fixed address. RSDP is wrote to 2 MiB after FDT, also a fixed address. Kernel comes 2 MiB after RSDP. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-05-18 23:24:09 +02:00
Rob Bradford	b282ff44d4	vmm: Enhance boot with info!() level messages These messages are predominantly during the boot process but will also occur during events such as hotplug. These cover all the significant steps of the boot and can be helpful for diagnosing performance and functionality issues during the boot. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-18 20:45:38 +02:00
Dayu Liu	8160c2884b	docs: Fix some typos in docs and comments Fix some typos or misspellings without functional change. Signed-off-by: Dayu Liu <liu.dayu@zte.com.cn>	2021-05-18 17:19:12 +01:00
Rob Bradford	5296bd10c1	vmm: tdx: Reorder TDX ioctls Separate the population of the memory and the HOB from the TDX initialisation of the memory so that the latter can happen after the CPU is initialised. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-17 07:39:52 +02:00
Rob Bradford	496ceed1d0	misc: Remove unnecessary "extern crate" Now all crates use edition = "2018" then the majority of the "extern crate" statements can be removed. Only those for importing macros need to remain. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-12 17:26:11 +02:00
Rob Bradford	6e6f66de2e	build: Bulk upgrade dependencies including fuzz Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-12 13:57:16 +01:00
Rob Bradford	b8f5911c4e	misc: Remove unused errors from public interface Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-11 13:37:19 +02:00
Rob Bradford	30b74e74cd	vmm: tdx: Reject attempt to use --kernel with --tdx Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-10 15:12:41 +01:00
Sebastien Boeuf	b5c6b04b36	virtio-devices, vmm: vhost: net: Add client mode support Adding the support for an OVS vhost-user backend to connect as the vhost-user client. This means we introduce with this patch a new option to our `--net` parameter. This option is called 'server' in order to ask the VMM to run as the server for the vhost-user socket. Fixes #1745 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-05-05 16:05:51 +02:00
Rob Bradford	7e0ccce225	vmm: config: Validate that vCPUs is sufficient for MQ queue count Fixes: #2563 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-04 19:49:34 +02:00
Rob Bradford	2ad615cd32	vmm: Validate config upon hotplug Create a temporary copy of the config, add the new device and validate that. This needs to be done separately to adding it to the config to avoid race conditions that might be result in config changes being overwritten. Fixes: #2564 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-04 19:49:20 +02:00
Rob Bradford	151668b260	vmm: Simplify adding devices to config To handle that devices are stored in an Option<Vec<T>> and reduce duplicated code use generic function to add the devices to the the struct. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-04 19:49:20 +02:00
Rob Bradford	f237789753	vmm: Output config used for booting VM at info!() level This helps with debugging VM behaviour when the VM is created via an API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-05-04 14:27:04 +01:00
Rob Bradford	da8136e49d	arch, vmm: Remove support for LinuxBoot By supporting just PVH boot on x86-64 we simplify our boot path substatially. Fixes: #2231 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-04-30 16:16:48 +02:00
Mikko Ylinen	3b18caf229	sgx: update virt EPC device path and docs The latest kvm-sgx code has renamed sgx_virt_epc device node to sgx_vepc. Update cloud-hypervisor code and documentation to follow this. Signed-off-by: Mikko Ylinen <mikko.ylinen@intel.com>	2021-04-30 16:16:01 +02:00
William Douglas	a2cfe71c0a	vmm: Update the http thread's seccomp filter Because the http thread no longer needs to create the api socket, remove the socket, bind and listen syscalls from the seccomp filter. Signed-off-by: William Douglas <william.douglas@intel.com>	2021-04-29 09:44:40 +01:00
William Douglas	b8779ddc9e	vmm: Create the api socket fd to pass to the http server Instead of using the http server's method to have it create the fd (causing the http thread to need to support the socket, bind and listen syscalls). Create the socket fd in the vmm thread and use the http server's new method supporting passing in this fd for the api socket. Signed-off-by: William Douglas <william.douglas@intel.com>	2021-04-29 09:44:40 +01:00
dependabot-preview[bot]	1bad026377	build(deps): bump libc from 0.2.93 to 0.2.94 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.93 to 0.2.94. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.93...0.2.94) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-04-28 07:02:14 +00:00
William Douglas	767b4f0e59	main: Enable the api-socket to be passed as an fd To avoid race issues where the api-socket may not be created by the time a cloud-hypervisor caller is ready to look for it, enable the caller to pass the api-socket fd directly. Avoid breaking current callers by allowing the --api-socket path to be passed as it is now in addition to through the path argument. Signed-off-by: William Douglas <william.r.douglas@gmail.com>	2021-04-26 14:40:49 -07:00
Rob Bradford	01f0c1e313	vmm: Simplify memory state to support Versionize In order to support using Versionize for state structures it is necessary to use simpler, primitive, data types in the state definitions used for snapshot restore. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-04-23 14:24:16 +01:00
Sebastien Boeuf	0b00442022	vmm: acpi: Allow reading from B0EJ field Windows guests read this field upon PCI device ejection. Let's make sure we don't return an error as this is valid. We simply return an empty u32 since the ejection is done right away upon write access, which means there's no pending ejection that might be reported to the guest. Here is the error that was shown during PCI device removal: ERROR:vmm/src/device_manager.rs:3960 -- Accessing unknown location at base 0x7ffffee000, offset 0x8 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-04-21 16:11:54 +01:00
Rob Bradford	a7c4483b8b	vmm: Directly (de)serialise CpuManager, DeviceManager and MemoryManager state Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-04-20 18:58:37 +02:00
Bo Chen	78796f96b7	vmm: Refine the granularity of dirty memory tracking Instead of tracking on a block level of 64 pages, we are now collecting dirty pages one by one. It improves the efficiency of dirty memory tracking while live migration. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-04-19 17:17:14 +02:00
Michael Zhao	4c299c6c00	vmm: Update 'micro_http' crate branch to 'main' The master branch of 'micro_http' crate was renamed to 'main'. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-04-19 15:47:54 +01:00
Sebastien Boeuf	b92fe648e9	vmm: cpu: Disable KVM_FEATURE_ASYNC_PF_INT in CPUID By disabling this KVM feature, we prevent the guest from using APF (Asynchronous Page Fault) mechanism. The kernel has recently switched to using interrupts to notify about a page being ready, but for some reasons, this is causing unexpected behavior with Cloud Hypervisor, as it will make the vcpu thread spin at 100%. While investigating the issue, it's better to disable the KVM feature to prevent 100% CPU usage in some cases. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-04-15 10:08:45 +01:00
Wei Liu	810ed7e887	vmm: interrupt: drop unnecessary type from impl The original code had a generic type E. It was later replaced by a concrete type. The code should have been simplified when the replacement happened. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-04-13 14:18:14 +01:00
Rob Bradford	86e4067437	vmm: config: Reject reserved fd from network config Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-04-13 14:29:18 +02:00
Rob Bradford	e0c0d0e142	vmm: config: Validate network configuration Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-04-13 14:29:18 +02:00
Rob Bradford	17072e9a6f	vmm: seccomp: Add missing SYS_newfstatat This is used when running on a new libc like Fedora34. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-04-12 18:02:29 +02:00
Anatol Belski	e1cc702327	memory_manager: Fix address range calculation in MemorySlot The MCRS method returns a 64-bit memory range descriptor. The calculation is supposed to be done as follows: max = min + len - 1 However, every operand is represented not as a QWORD but as combination of two DWORDs for high and low part. Till now, the calculation was done this way, please see also inline comments: max.lo = min.lo + len.lo //this may overflow, need to carry over to high max.hi = min.hi + len.hi max.hi = max.hi - 1 // subtraction needs to happen on the low part This calculation has been corrected the following way: max.lo = min.lo + len.lo max.hi = min.hi + len.hi + (max.lo < min.lo) // check for overflow max.lo = max.lo - 1 // subtract from low part The relevant part from the generated ASL for the MCRS method: ``` Method (MCRS, 1, Serialized) { Acquire (MLCK, 0xFFFF) \_SB.MHPC.MSEL = Arg0 Name (MR64, ResourceTemplate () { QWordMemory (ResourceProducer, PosDecode, MinFixed, MaxFixed, Cacheable, ReadWrite, 0x0000000000000000, // Granularity 0x0000000000000000, // Range Minimum 0xFFFFFFFFFFFFFFFE, // Range Maximum 0x0000000000000000, // Translation Offset 0xFFFFFFFFFFFFFFFF, // Length ,, _Y00, AddressRangeMemory, TypeStatic) }) CreateQWordField (MR64, \_SB.MHPC.MCRS._Y00._MIN, MINL) // _MIN: Minimum Base Address CreateDWordField (MR64, 0x12, MINH) CreateQWordField (MR64, \_SB.MHPC.MCRS._Y00._MAX, MAXL) // _MAX: Maximum Base Address CreateDWordField (MR64, 0x1A, MAXH) CreateQWordField (MR64, \_SB.MHPC.MCRS._Y00._LEN, LENL) // _LEN: Length CreateDWordField (MR64, 0x2A, LENH) MINL = \_SB.MHPC.MHBL MINH = \_SB.MHPC.MHBH LENL = \_SB.MHPC.MHLL LENH = \_SB.MHPC.MHLH MAXL = (MINL + LENL) /* \_SB_.MHPC.MCRS.LENL / MAXH = (MINH + LENH) / \_SB_.MHPC.MCRS.LENH / If ((MAXL < MINL)) { MAXH += One / \_SB_.MHPC.MCRS.MAXH / } MAXL -= One Release (MLCK) Return (MR64) / \_SB_.MHPC.MCRS.MR64 */ } ``` Fixes #1800. Signed-off-by: Anatol Belski <anbelski@linux.microsoft.com>	2021-04-12 16:20:19 +02:00
dependabot-preview[bot]	23411d45ba	build(deps): bump libc from 0.2.92 to 0.2.93 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.92 to 0.2.93. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.92...0.2.93) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-04-06 17:03:25 +00:00
dependabot-preview[bot]	d052b9ec12	build(deps): bump seccomp from v0.22.0 to v0.24.2 Bumps [seccomp](https://github.com/firecracker-microvm/firecracker) from v0.22.0 to v0.24.2. - [Release notes](https://github.com/firecracker-microvm/firecracker/releases) - [Changelog](`5ba819d7b7/CHANGELOG.md`) - [Commits](`cc5387637c...5ba819d7b7`) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-04-05 21:34:34 +01:00
dependabot-preview[bot]	78f9d2cc85	build(deps): bump signal-hook from 0.3.7 to 0.3.8 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.3.7 to 0.3.8. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/commits/v0.3.8) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-04-04 08:15:46 +00:00
Sebastien Boeuf	46f96f27a4	vmm: Add missing syscall for vCPU unplug clock_nanosleep() is triggered when hot-unplugging a vCPU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-31 10:51:52 +01:00
Bo Chen	5d8de62362	vmm: openapi: Add `rate_limiter_config` to the `NetConfig` Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-03-30 19:47:43 +02:00
Bo Chen	b176ddfe2a	virtio-devices, vmm: Add rate limiter for the TX queue of virtio-net Partially fixes: #1286 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-03-30 19:47:43 +02:00
dependabot-preview[bot]	b8311cac38	build(deps): bump libc from 0.2.91 to 0.2.92 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.91 to 0.2.92. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.91...0.2.92) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-03-30 12:09:56 +00:00
Sebastien Boeuf	73e8fd4d72	clippy: Fix codebase to compile with beta toolchain Fixes the current codebase so that every cargo clippy can be run with the beta toolchain without any error. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-29 15:56:23 +01:00
Anatol Belski	9e9aba7c0b	CpuManager: Fix MMIO read handling There are two parts: - Unconditionally zero the output area. The length of the incoming vector has been seen from 1 to 4 bytes, even though just the first byte might need to be handled. But also, this ensures any possibly unhandled offset will return zeroed result to the caller. The former implementation used an I/O port which seems to behave differently from MMIO and wouldn't require explicit output zeroing. - An access with zero offset still takes place and needs to be handled. Fixes #2437. Signed-off-by: Anatol Belski <anbelski@linux.microsoft.com>	2021-03-29 13:51:31 +01:00
Rob Bradford	431c16dc44	vmm: Use definition of MmioDeviceInfo from arch Remove duplicated copies from vmm. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-29 12:06:07 +02:00
Gaelan Steele	21506a4a76	vmm: reorder constructor fields to match struct Satisfies nightly clippy. Signed-off-by: Gaelan Steele <gbs@canishe.com>	2021-03-29 09:55:29 +02:00
Gaelan Steele	b161a570ec	vmm: use Option::map and Option::cloned It's more concise, more idiomatic Rust, and satisfies nightly clippy. Signed-off-by: Gaelan Steele <gbs@canishe.com>	2021-03-29 09:55:29 +02:00
Anatol Belski	5b168f54a6	hyperv: Fix CPU hotadd The following is from the Hyper-V specification v6.0b. Cpuid leaf 0x40000003 EDX: Bit 3: Support for physical CPU dynamic partitioning events is available. When Windows determines to be running under a hypervisor, it will require this cpuid bit to be set to support dynamic CPU operations. Cpuid leaf 0x40000004 EAX: Bit 5: Recommend using relaxed timing for this partition. If used, the VM should disable any watchdog timeouts that rely on the timely delivery of external interrupts. This bit has been figured out as required after seeing guest BSOD when CPU hotplug bit is enabled. Race conditions seem to arise after a hotplug operation, when a system watchdog has expired. Closes #1799. Signed-off-by: Anatol Belski <anbelski@linux.microsoft.com>	2021-03-26 14:06:51 +01:00
Rob Bradford	40da6210f4	aarch64: Address Rust 1.51.0 clippy issue (upper_case_acroynms) error: name `GPIOInterruptDisabled` contains a capitalized acronym Error: --> devices/src/legacy/gpio_pl061.rs:46:5 \| 46 \| GPIOInterruptDisabled, \| ^^^^^^^^^^^^^^^^^^^^^ help: consider making the acronym lowercase, except the initial letter: `GpioInterruptDisabled` \| = note: `-D clippy::upper-case-acronyms` implied by `-D warnings` = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#upper_case_acronyms Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Rob Bradford	3c6dfd7709	tdx: Address Rust 1.51.0 clippy issue (upper_case_acroynms) error: name `FinalizeTDX` contains a capitalized acronym --> vmm/src/vm.rs:274:5 \| 274 \| FinalizeTDX(hypervisor::HypervisorVmError), \| ^^^^^^^^^^^ help: consider making the acronym lowercase, except the initial letter: `FinalizeTdx` \| = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#upper_case_acronyms Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Rob Bradford	c5d15fd938	devices: Address Rust 1.51.0 clippy issue (upper_case_acroynms) warning: name `AcpiPMTimerDevice` contains a capitalized acronym --> devices/src/acpi.rs:175:12 \| 175 \| pub struct AcpiPMTimerDevice { \| ^^^^^^^^^^^^^^^^^ help: consider making the acronym lowercase, except the initial letter: `AcpiPmTimerDevice` \| = note: `#[warn(clippy::upper_case_acronyms)]` on by default = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#upper_case_acronyms Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Rob Bradford	827229d8e4	pci: Address Rust 1.51.0 clippy issue (upper_case_acroynms) warning: name `IORegion` contains a capitalized acronym --> pci/src/configuration.rs:320:5 \| 320 \| IORegion = 0x01, \| ^^^^^^^^ help: consider making the acronym lowercase, except the initial letter (notice the capitalization): `IoRegion` \| = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#upper_case_acronyms Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Rob Bradford	3b8d1f1411	vmm: Address Rust 1.51.0 clippy issue (vec_init_then_push) warning: calls to `push` immediately after creation --> vmm/src/cpu.rs:630:9 \| 630 \| / let mut cpuid_patches = Vec::new(); 631 \| \| 632 \| \| // Patch tsc deadline timer bit 633 \| \| cpuid_patches.push(CpuidPatch { ... \| 662 \| \| edx_bit: Some(MTRR_EDX_BIT), 663 \| \| }); \| \|___________^ help: consider using the `vec![]` macro: `let mut cpuid_patches = vec![..];` \| = note: `#[warn(clippy::vec_init_then_push)]` on by default = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#vec_init_then_push Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Rob Bradford	9762c8bc28	vmm: Address Rust 1.51.0 clippy issue (upper_case_acroynms) warning: name `LocalAPIC` contains a capitalized acronym --> vmm/src/cpu.rs:197:8 \| 197 \| struct LocalAPIC { \| ^^^^^^^^^ help: consider making the acronym lowercase, except the initial letter: `LocalApic` \| = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#upper_case_acronyms Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Rob Bradford	7c302373ed	vmm: Address Rust 1.51.0 clippy issue (needless_question_mark) warning: Question mark operator is useless here --> vmm/src/seccomp_filters.rs:493:5 \| 493 \| / Ok(SeccompFilter::new( 494 \| \| rules.into_iter().collect(), 495 \| \| SeccompAction::Trap, 496 \| \| )?) \| \|_______^ \| = note: `#[warn(clippy::needless_question_mark)]` on by default = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#needless_question_mark help: try \| 493 \| SeccompFilter::new( 494 \| rules.into_iter().collect(), 495 \| SeccompAction::Trap, 496 \| ) \| warning: Question mark operator is useless here --> vmm/src/seccomp_filters.rs:507:5 \| 507 \| / Ok(SeccompFilter::new( 508 \| \| rules.into_iter().collect(), 509 \| \| SeccompAction::Log, 510 \| \| )?) \| \|_______^ \| = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#needless_question_mark help: try \| 507 \| SeccompFilter::new( 508 \| rules.into_iter().collect(), 509 \| SeccompAction::Log, 510 \| ) \| warning: Question mark operator is useless here --> vmm/src/vm.rs:887:9 \| 887 \| Ok(CString::new(cmdline).map_err(Error::CmdLineCString)?) \| ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^ help: try: `CString::new(cmdline).map_err(Error::CmdLineCString)` \| = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#needless_question_mark Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Rob Bradford	aa34d545f6	vm-virtio, virtio-devices: Address Rust 1.51.0 clippy issue (upper_case_acronyms) error: name `TYPE_UNKNOWN` contains a capitalized acronym --> vm-virtio/src/lib.rs:48:5 \| 48 \| TYPE_UNKNOWN = 0xFF, \| ^^^^^^^^^^^^ help: consider making the acronym lowercase, except the initial letter: `Type_Unknown` \| = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#upper_case_acronyms Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Rob Bradford	db6516931d	acpi_tables: Address Rust 1.51.0 clippy issue (upper_case_acronyms) error: name `SDT` contains a capitalized acronym --> acpi_tables/src/sdt.rs:27:12 \| 27 \| pub struct SDT { \| ^^^ help: consider making the acronym lowercase, except the initial letter: `Sdt` \| = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#upper_case_acronyms Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-26 11:32:09 +00:00
Wei Liu	27ba8133a4	vmm: interrupt: drop RoutingEntryExt trait We can now directly use associated functions. This simplifies code and causes no functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-03-25 10:13:43 +01:00
Wei Liu	f5c550affb	vmm: interrupt: simplify interrupt handling traits Drop the generic type E and use IrqRoutngEntry directly. This allows dropping a bunch of trait bounds from code. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-03-25 10:13:43 +01:00
Wei Liu	1d9f27c9fb	vmm: interrupt: extract common code from MSHV and KVM Their make_entry functions look the same now. Extract the logic to a common function. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-03-25 10:13:43 +01:00
Vineeth Pillai	68401e6e4a	hypervisor:mshv: Support the move of MSI routing to kernel Signed-off-by: Vineeth Pillai <viremana@linux.microsoft.com>	2021-03-23 11:06:13 +01:00
dependabot-preview[bot]	e9793020c2	build(deps): bump libc from 0.2.90 to 0.2.91 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.90 to 0.2.91. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.90...0.2.91) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-03-23 06:56:04 +00:00
Sebastien Boeuf	1e11e6789a	vmm: device_manager: Factorize passthrough_device creation There's no need to have the code creating the passthrough_device being duplicated since we can factorize it in a function used in both cases (both cold plugged and hot plugged devices VFIO devices). Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-22 10:47:33 +01:00
dependabot-preview[bot]	e39924d45a	build(deps): bump libc from 0.2.89 to 0.2.90 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.89 to 0.2.90. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.89...0.2.90) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-03-18 21:29:25 +00:00
Sebastien Boeuf	b06fd80fa9	vmm: device_manager: Use DeviceTree to store PCI devices Extend and use the existing DeviceTree to retrieve useful information related to PCI devices. This removes the duplication with pci_devices field which was internal to the DeviceManager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-18 15:26:25 +01:00
Sebastien Boeuf	c8c3cad8cb	vmm: device_manager: Update structure holding PCI IRQs Make the code a bit clearer by changing the naming of the structure holding the list of IRQs reserved for PCI devices. It is also modified into an array of 32 entries since we know this is the amount of PCI slots that is supported. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-18 15:26:25 +01:00
Sebastien Boeuf	e311fd66cd	vmm: device_manager: Remove the need for Any We define a new enum in order to classify PCI device under virtio or VFIO. This is a cleaner approach than using the Any trait, and downcasting it to find the object back. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-18 15:26:25 +01:00
Sebastien Boeuf	fbd624d816	vmm: device_manager: Remove pci_id_list Introduces a tuple holding both information needed by pci_id_list and pci_devices. Changes pci_devices to be a BTreeMap of this new tuple. Now that pci_devices holds the information needed from pci_id_list, pci_id_list is no longer needed. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-18 15:26:25 +01:00
Sebastien Boeuf	305e095d15	vmm: device_manager: Invert pci_id_list HashMap In anticipation for further factorization, the pci_id_list is now a hashmap of PCI b/d/f leading to each device name. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-18 15:26:25 +01:00
Sebastien Boeuf	62aaccee28	vmm: Device name verification based on DeviceTree Instead of relying on a PCI specific device list, we use the DeviceTree as a reference to determine if a device name is already in use or not. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-18 15:26:25 +01:00
Rob Bradford	9440304183	vmm: http: Error out earlier if we can't create API server This removes a panic inside the API thread. Fixes: #2395 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-17 11:30:26 +00:00
Rob Bradford	9b0996a71f	vmm, main: Optionalise creation of API server Only if we have a valid API server path then create the API server. For now this has no functional change there is a default API server path in the clap handling but rather prepares to do so optionally. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-17 11:30:26 +00:00
Henry Wang	2bb153de2b	vmm: Implement the power button method for AArch64 This commit implements the power button method for AArch64 using the PL061 GPIO controller. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-03-16 20:27:15 +08:00
Henry Wang	a59ff42a95	aarch64: Add PL061 for device tree implementation This commit adds a new legacy device PL011 for the AArch64 device tree implementation. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-03-16 20:27:15 +08:00
Henry Wang	a8cde12b14	vmm: AArch64: Use PL011 for AArch64 device tree This commit switches the default serial device from 16550 to the Arm dedicated UART controller PL011. The `ttyAMA0` can be enabled. Signed-off-by: Henry Wang <Henry.Wang@arm.com>	2021-03-16 11:53:51 +08:00
dependabot-preview[bot]	c85fba0c43	build(deps): bump libc from 0.2.88 to 0.2.89 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.88 to 0.2.89. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.88...0.2.89) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-03-15 23:40:01 +00:00
Michael Zhao	ee7fcdb3cf	aarch64: Correct wrong settings for serial device Corrected: - The device name in FDT - MMIO mapping size Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-03-15 20:59:50 +08:00
Michael Zhao	afc83582be	aarch64: Enable IRQ routing for legacy devices On AArch64, interrupt controller (GIC) is emulated by KVM. VMM need to set IRQ routing for devices, including legacy ones. Before this commit, IRQ routing was only set for MSI. Legacy routing entries of type KVM_IRQ_ROUTING_IRQCHIP were missing. That is way legacy devices (like serial device ttyS0) does not work. The setting of X86 IRQ routing entries are not impacted. Signed-off-by: Michael Zhao <michael.zhao@arm.com>	2021-03-15 20:59:50 +08:00
Rob Bradford	5bc311184e	build: Remove url crate dependency This removes multiple transitive dependencies and speeds up our build. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-12 16:52:55 +01:00
Rob Bradford	7f96eb2b67	vmm: migration: Simplify url socket handling in migration code Extract URL handling to a common function and simplify to remove url crate dependency. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-12 16:52:55 +01:00
Rob Bradford	ab4b30edd3	vmm: Switch MemoryManager::send() to url_to_path() This continues the work in `cc78a597cd` Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-12 16:52:55 +01:00
Rob Bradford	cc78a597cd	vmm: Simplify snapshot/restore path handling Extend the existing url_to_path() to take the URL string and then use that to simplify the snapshot/restore code paths. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-12 13:03:01 +01:00
Bo Chen	ebcbab739e	vmm: openapi: Add `rate_limiter_config` to the `DiskConfig Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-03-12 09:35:03 +01:00
Bo Chen	af8def364d	virtio-devices, vmm: add I/O rate limiter on block device This patch is based on the 'rate_limiter' module from firecracker[1]. To simplify dependencies, we reply on 'vmm-sys-util::TimerFd' instead of the `timerfd` crate. [1]https://github.com/firecracker-microvm/firecracker/tree/master/src/rate_limiter Fixes: #1285 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-03-12 09:35:03 +01:00
Sebastien Boeuf	00873e5f84	vmm: device_manager: Update virtio devices memory with a single region Relies on the preliminary work allowing virtio devices to be updated with a single memory at a time instead of updating the entire memory at once. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-11 19:04:21 +01:00
dependabot-preview[bot]	7b3e79cdfa	build(deps): bump signal-hook from 0.3.6 to 0.3.7 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.3.6 to 0.3.7. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/commits) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-03-11 08:45:27 +00:00
Rob Bradford	66ffaceffc	vmm: tdx: Correctly populate the 64-bit MMIO region The MMIO structure contains the length rather than the maximum address so it is necessary to subtract the starting address from the end address to calculate the length. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-10 14:45:51 +00:00
Rob Bradford	a0c07474a3	vmm: seccomp: Add KVM_MEMORY_ENCRYPT_OP ioctl to seccomp filter This is the basis for TDX based operations on the various KVM file descriptors. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-09 16:26:06 +01:00
Rob Bradford	d24aa887b6	vmm: Reject VM snapshot request if TDX in use It is not possible to snapshot the contents of a TDX VM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	835a31e283	vmm: config: Require max and boot vCPUs to be equal for TDX CPU hotplug is not possible with TDX Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	57c8c250fd	tdx: Permit starting Cloud Hypervisor without --kernel This is not required if TDX is present. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	62abc117ab	tdx: Configure TDX state for the VM Load the sections backed from the file into their required addresses in memory and populate the HOB with details of the memory. Using the HOB address initialize the TDX state in the vCPUs and finalize the TDX configuration. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	57ce0986f7	vmm: cpu: Add functionality for enabling TDX for all vCPUs Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	c8cad394b5	vmm: cpu: Expose the common/shared CPUID data for all vCPUs This allows the CPUID data to be passed into the VM level ioctl used for initalizing TDX. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	b02aff5761	vmm: memory_manager: Disable dirty page logging when running on TDX It is not permitted to have this enabled in memory that is part of a TD. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	f282cc001a	tdx: Add abstraction to call TDX ioctls to hypervisor Add API to the hypervisor interface and implement for KVM to allow the special TDX KVM ioctls on the VM and vCPU FDs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	c48e82915d	vmm: Make kernel optional in VM internals When booting with TDX no kernel is supplied as the TDFV is responsible for loading the OS. The requirement to have the kernel is still currently enforced at the validation entry point; this change merely changes function prototypes and stored state to use Option<> to support. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	66a3bed086	vmm: config: Add "--tdx" option parsing Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
Rob Bradford	e61ee6bcac	tdx: Add "tdx" feature with an empty module inside arch to implement Add the skeleton of the "tdx" feature with a module ready inside the arch crate to store implementation details. TEST=cargo build --features="tdx" Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-03-08 18:30:00 +00:00
dependabot-preview[bot]	ccfa34d066	build(deps): bump libc from 0.2.87 to 0.2.88 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.87 to 0.2.88. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.87...0.2.88) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-03-05 18:39:37 +00:00
William Douglas	56028fb214	Try to restore pty configuration on reboot When a vm is created with a pty device, on reboot the pty fd (sub only) will only be associated with the vmm through the epoll event loop. The fd being polled will have been closed due to the vm itself dropping the pty files (and potentially reopening the fd index to a different item making things quite confusing) and new pty fds will be opened but not polled on for input. This change creates a structure to encapsulate the information about the pty fd (main File, sub File and the path to the sub File). On reboot, a copy of the console and serial pty structs is then passed down to the new Vm instance which will be used instead of creating a new pty device. This resolves the underlying issue from #2316. Signed-off-by: William Douglas <william.r.douglas@gmail.com>	2021-03-05 18:34:52 +01:00
Sebastien Boeuf	933d41cf2f	vmm: Provide DMA mapping handlers to virtio-mem devices Now that virtio-mem devices can update VFIO mappings through dedicated handlers, let's provide them from the DeviceManager. Important to note these handlers should either be provided to virtio-mem devices or to the unique virtio-iommu device. This must be mutually exclusive. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-05 10:38:42 +01:00
Sebastien Boeuf	080ea31813	pci, vmm: Manage VFIO DMA mapping from DeviceManager Instead of letting the VfioPciDevice take the decision on how/when to perform the DMA mapping/unmapping, we move this to the DeviceManager instead. The point is to let the DeviceManager choose which guest memory regions should be mapped or not. In particular, we don't want the virtio-mem region to be mapped/unmapped as it will be virtio-mem device responsibility to do so. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-05 10:38:42 +01:00
Sebastien Boeuf	d6db2fdf96	vmm: memory_manager: Add ACPI hotplug region to default memory zone When memory is resized through ACPI, a new region is added to the guest memory. This region must also be added to the corresponding memory zone in order to keep everything in sync. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-03-05 10:38:42 +01:00
dependabot-preview[bot]	d433ae1656	build(deps): bump libc from 0.2.86 to 0.2.87 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.86 to 0.2.87. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.86...0.2.87) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-03-02 11:14:57 +00:00
Rob Bradford	f8875acec2	misc: Bulk upgrade dependencies In particular update for the vmm-sys-util upgrade and all the other dependent packages. This requires an updated forked version of kvm-bindings (due to updated vfio-ioctls) but allowed the removal of our forked version of kvm-ioctls. The changes to the API from kvm-ioctls and vmm-sys-util required some other minor changes to the code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-26 11:31:08 +00:00
Sebastien Boeuf	a0a89b1346	pci, vmm: Move to upstream vfio-ioctls crate This commit moves both pci and vmm code from the internal vfio-ioctls crate to the upstream one from the rust-vmm project. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-24 08:02:37 +01:00
Sebastien Boeuf	aee1155870	virtio-devices, vmm: Move to ExternalDmaMapping from vm-device Now that ExternalDmaMapping is defined in vm-device, let's use it from there. This commit also defines the function get_host_address_range() to move away from the vfio-ioctls dependency. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-24 08:02:37 +01:00
Rob Bradford	deedfcdc35	vmm: Improve restore error message about URL conversion Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-23 11:07:48 +00:00
Bo Chen	d361fc1a36	vmm: config: Fix and complete the help info for the '--disk' option The help information displayed for our `--disk` option is incorrect and incomplete, e.g. missing the `direct` and `poll_queue` field. Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-02-23 08:55:33 +01:00
Rob Bradford	05a2b3fac2	vmm: Remove "tempfile" dependency from vmm This was completely unused. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-22 14:29:53 +01:00
Sebastien Boeuf	4ed0e1a3c8	net_util: Simplify TX/RX queue handling The main idea behind this commit is to remove all the complexity associated with TX/RX handling for virtio-net. By using writev() and readv() syscalls, we could get rid of intermediate buffers for both queues. The complexity regarding the TAP registration has been simplified as well. The RX queue is only processed when some data are ready to be read from TAP. The event related to the RX queue getting more descriptors only serves the purpose to register the TAP file if it's not already. With all these simplifications, the code is more readable but more performant as well. We can see an improvement of 10% for a single queue device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-22 10:39:23 +00:00
dependabot-preview[bot]	ae04fe432c	build(deps): bump signal-hook from 0.3.4 to 0.3.6 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.3.4 to 0.3.6. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/compare/v0.3.4...v0.3.6) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-02-21 09:36:30 +00:00
dependabot-preview[bot]	bfb12b7777	build(deps): bump url from 2.2.0 to 2.2.1 Bumps [url](https://github.com/servo/rust-url) from 2.2.0 to 2.2.1. - [Release notes](https://github.com/servo/rust-url/releases) - [Commits](https://github.com/servo/rust-url/compare/v2.2.0...v2.2.1) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-02-18 21:56:54 +00:00
Rob Bradford	9260c4c10e	vmm: Use event!() for some key VM actions Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-18 16:15:13 +00:00
Rob Bradford	c1d9edbfc0	vmm: seccomp: Add getrandom to vCPU thread filter This can be triggered upon device reset. Fixes: #2278 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-18 16:15:13 +00:00
Rob Bradford	38c41a5074	vmm: memory_manager: Extract code for allocating new memory This function can then be used by the TDX code to allocate the memory at specific locations required for the TDVF to run from. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-16 18:38:57 +01:00
Rob Bradford	707bb0ba72	vmm: Simplify return path of vm_boot Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-16 18:38:57 +01:00
Rob Bradford	bc84a1c79b	build: Remove nested matches Update for clippy in Rust 1.50.0: error: Unnecessary nested match --> vmm/src/vm.rs:419:17 \| 419 \| / if let vm_device::BusError::MissingAddressRange = e { 420 \| \| warn!("Guest MMIO write to unregistered address 0x{:x}", gpa); 421 \| \| } \| \|_________________^ \| = note: `-D clippy::collapsible-match` implied by `-D warnings` help: The outer pattern can be modified to include the inner pattern. --> vmm/src/vm.rs:418:17 \| 418 \| Err(e) => { \| ^ Replace this binding 419 \| if let vm_device::BusError::MissingAddressRange = e { \| ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^ with this pattern = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#collapsible_match Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-11 18:18:44 +00:00
Rob Bradford	9c5be6f660	build: Remove unnecessary Result<> returns If the function can never return an error this is now a clippy failure: error: this function's return value is unnecessarily wrapped by `Result` --> virtio-devices/src/watchdog.rs:215:5 \| 215 \| / fn set_state(&mut self, state: &WatchdogState) -> io::Result<()> { 216 \| \| self.common.avail_features = state.avail_features; 217 \| \| self.common.acked_features = state.acked_features; 218 \| \| // When restoring enable the watchdog if it was previously enabled. We reset the timer ... \| 223 \| \| Ok(()) 224 \| \| } \| \|_____^ \| = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#unnecessary_wraps Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-11 18:18:44 +00:00
Sebastien Boeuf	9353856426	vmm: Fix seccomp filters for vCPUs Depending on the host OS the code for looking up the time for the CMOS make require extra syscalls to be permitted for the vCPU thread. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-11 11:24:57 +00:00
Sebastien Boeuf	19167e7647	pci: vfio: Implement INTx support With all the preliminary work done in the previous commits, we can update the VFIO implementation to support INTx along with MSI and MSI-X. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-10 17:34:56 +00:00
Sebastien Boeuf	2cdc3e3546	vmm: device_manager: Add PCI routing table to ACPI Here we are adding the PCI routing table, commonly called _PRT, to the ACPI DSDT. For simplification reasons, we chose not to implement PCI links as this involves dynamic decision from the guest OS, which result in lots of complexity both from an AML perspective and from a device manager perspective. That's why the _PRT creates a static list of 32 entries, each assigned with the IRQ number previously reserved by the device manager. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-10 17:34:56 +00:00
Sebastien Boeuf	de9471fc72	vmm: device_manager: Allocate IRQs for PCI devices In order to support INTx for PCI devices, each PCI device must be assigned an IRQ. This is preliminary work to reserve 8 IRQs which will be shared across the 32 PCI devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-10 17:34:56 +00:00
Sebastien Boeuf	b5450ca72c	vmm: device_manager: Store legacy interrupt manager In anticipation for accessing the legacy interrupt manager from the function creating a VFIO PCI device, we store it as part of the DeviceManager, to make it available for all methods. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-10 17:34:56 +00:00
Sebastien Boeuf	8008aad545	vmm: device_manager: Don't pass the MSI interrupt manager around The DeviceManager already has a hold onto the MSI interrupt manager, therefore there's no need to pass it through every function. Instead, let's simplify the code by using the attribute from DeviceManager's instance. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-10 17:34:56 +00:00
Sebastien Boeuf	3bd47ffdc1	interrupt: Add a notifier method to the InterruptController Both GIC and IOAPIC must implement a new method notifier() in order to provide the caller with an EventFd corresponding to the IRQ it refers to. This is needed in anticipation for supporting INTx with VFIO PCI devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-10 17:34:56 +00:00
Sebastien Boeuf	acfbee5b7a	interrupt: Make notifier function return Option<EventFd> In anticipation for supporting the notifier function for the legacy interrupt source group, we need this function to return an EventFd instead of a reference to this same EventFd. The reason is we can't return a reference when there's an Arc<Mutex<>> involved in the call chain. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-10 17:34:56 +00:00
Wei Liu	5fc12862e6	hypervisor, vmm: minor changes to VmmOps Swap the last two parameters of guest_mem_{read,write} to be consistent with other read / write functions. Use more descriptive parameter names. No functional change. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-02-10 11:31:03 +00:00
dependabot-preview[bot]	6d63018d9f	build(deps): bump vm-memory from 0.4.0 to 0.5.0 Bumps [vm-memory](https://github.com/rust-vmm/vm-memory) from 0.4.0 to 0.5.0. - [Release notes](https://github.com/rust-vmm/vm-memory/releases) - [Changelog](https://github.com/rust-vmm/vm-memory/blob/master/CHANGELOG.md) - [Commits](https://github.com/rust-vmm/vm-memory/compare/v0.4.0...v0.5.0) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-02-10 11:30:05 +00:00
Rob Bradford	50a995b63d	vmm: Rename patch_cpuid() to generate_common_cpuid() This reflects that it generates CPUID state used across all vCPUs. Further ensure that errors from this function get correctly propagated. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-09 16:02:25 +00:00
Rob Bradford	ccdea0274c	vmm, arch: Move KVM HyperV emulation handling to shared CPUID code Move the code for populating the CPUID with KVM HyperV emulation details from the per-vCPU CPUID handling code to the shared CPUID handling code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-09 16:02:25 +00:00
Rob Bradford	688ead51c6	vmm, arch: Move CPU identification handling to shared CPUID code Move the code for populating the CPUID with details of the CPU identification from the per-vCPU CPUID handling code to the shared CPUID handling code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-09 16:02:25 +00:00
Rob Bradford	9792c9aafa	vmm, arch: Move max_phys_bits handling to shared CPUID code Move the code for populating the CPUID with details of the maximum address space from the per-vCPU CPUID handling code to the shared CPUID handling code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-09 16:02:25 +00:00
William Douglas	48963e322a	Enable pty console Add the ability for cloud-hypervisor to create, manage and monitor a pty for serial and/or console I/O from a user. The reasoning for having cloud-hypervisor create the ptys is so that clients, libvirt for example, could exit and later re-open the pty without causing I/O issues. If the clients were responsible for creating the pty, when they exit the main pty fd would close and cause cloud-hypervisor to get I/O errors on writes. Ideally the main and subordinate pty fds would be kept in the main vmm's Vm structure. However, because the device manager owns parsing the configuration for the serial and console devices, the information is instead stored in new fields under the DeviceManager structure directly. From there hooking up the main fd is intended to look as close to handling stdin and stdout on the tty as possible (there is some future work ahead for perhaps moving support for the pty into the vmm_sys_utils crate). The main fd is used for reading user input and writing to output of the Vm device. The subordinate fd is used to setup raw mode and it is kept open in order to avoid I/O errors when clients open and close the pty device. The ability to handle multiple inputs as part of this change is intentional. The current code allows serial and console ptys to be created and both be used as input. There was an implementation gap though with the queue_input_bytes needing to be modified so the pty handlers for serial and console could access the methods on the serial and console structures directly. Without this change only a single input source could be processed as the console would switch based on its input type (this is still valid for tty and isn't otherwise modified). Signed-off-by: William Douglas <william.r.douglas@gmail.com>	2021-02-09 10:03:28 +00:00
dependabot-preview[bot]	aa3d5cfbfe	build(deps): bump libc from 0.2.85 to 0.2.86 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.85 to 0.2.86. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.85...0.2.86) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-02-08 16:19:06 +00:00
Wei Liu	1fc8c9165a	vmm: drop two unused errors Their last users were gone in `af3c6c34c3`. Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-02-08 16:15:31 +00:00
Rob Bradford	b91b829a28	openapi: Update API spec to include hugepage_size parameter Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-05 09:24:02 +00:00
Rob Bradford	7928a697dc	vmm: Support configurable huge pages in MemoryManager Use the newly added hugepages_size option if provided by the user to pick a huge page size when creating the memfd region. If none is specified use the system default. Sadly different huge pages cannot be tested by an integration test as creating a pool of the non-default size cannot be done at runtime (requires kernel to be booted with certain parameters.) TETS=Manually tested with a kernel booted with both 1GiB and 2MiB huge pages (hugepagesz=1G hugepages=1 hugepagesz=2M hugepages=512) Fixes: #2230 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-05 09:24:02 +00:00
Rob Bradford	29607f38ad	vmm: config: Add a hugepage_size option This allows the user to use an alternative huge page size otherwise the default size will be used. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-05 09:24:02 +00:00
Rob Bradford	9c3a73706f	vmm: config: Remove ineffectual code from test Remove a change the invalid configuration that is not ever tested again. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-02-05 09:24:02 +00:00
Sebastien Boeuf	c397c9c95e	vmm, virtio-devices: mem: Don't use MADV_DONTNEED on hugepages This commit introduces a new information to the VirtioMemZone structure in order to know if the memory zone is backed by hugepages. Based on this new information, the virtio-mem device is now able to determine if madvise(MADV_DONTNEED) should be performed or not. The madvise documentation specifies that MADV_DONTNEED advice will fail if the memory range has been allocated with some hugepages. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com> Signed-off-by: Hui Zhu <teawater@antfin.com>	2021-02-04 17:52:30 +00:00
Sebastien Boeuf	f24094392e	virtio-devices: mem: Improve semantic around Resize object By introducing a ResizeSender object, we avoid having a Resize clone with a different content than the original Resize object. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-04 17:52:30 +00:00
dependabot-preview[bot]	89008a49cf	build(deps): bump libc from 0.2.84 to 0.2.85 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.84 to 0.2.85. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.84...0.2.85) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-02-02 07:47:41 +00:00
Sebastien Boeuf	24c8cce012	block_util: Add synchronous support for fixed VHD disk files Relying on the simplified version of the synchronous support for RAW disk files, the new fixed_vhd_sync module in the block_util crate introduces the synchronous support for fixed VHD disk files. With this patch, the fixed VHD support is complete as it is implemented in both synchronous and asynchronous versions. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-01 13:45:08 +00:00
Sebastien Boeuf	c6854c5a97	block_util: Simplify RAW synchronous implementation Using directly preadv and pwritev, we can simply use a RawFd instead of a file, and we don't need to use the more complex implementation from the qcow crate. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-01 13:45:08 +00:00
Sebastien Boeuf	b2e5dbaecb	block_util, vmm: Add fixed VHD asynchronous implementation This commit adds the asynchronous support for fixed VHD disk files. It introduces FixedVhd as a new ImageType, moving the image type detection to the block_util crate (instead of qcow crate). It creates a new vhd module in the block_util crate in order to handle VHD footer, following the VHD specification. It creates a new fixed_vhd_async module in the block_util crate to implement the asynchronous version of fixed VHD disk file. It relies on io_uring. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-02-01 13:45:08 +00:00
Muminul Islam	5bbf2dca80	vmm: Remove unneeded `return` statement This unneeded return statement giving clippy warnings Signed-off-by: Muminul Islam <muislam@microsoft.com>	2021-01-30 08:58:09 +00:00
dependabot-preview[bot]	1df952726a	build(deps): bump libc from 0.2.83 to 0.2.84 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.83 to 0.2.84. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/commits) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-01-29 06:21:38 +00:00
Rob Bradford	9ea087597f	vmm: acpi: Ensure field size address matches ResourceTag dsdt.dsl 960: CreateDWordField (MR64, \_SB.MHPC.MCRS._Y00._MIN, MINL) // _MIN: Minimum Base Address Warning 3128 - ResourceTag larger than Field ^ (Size mismatch, Tag: 64 bits, Field: 32 bits) dsdt.dsl 962: CreateDWordField (MR64, \_SB.MHPC.MCRS._Y00._MAX, MAXL) // _MAX: Maximum Base Address Warning 3128 - ResourceTag larger than Field ^ (Size mismatch, Tag: 64 bits, Field: 32 bits) dsdt.dsl 964: CreateDWordField (MR64, \_SB.MHPC.MCRS._Y00._LEN, LENL) // _LEN: Length Warning 3128 - ResourceTag larger than Field ^ (Size mismatch, Tag: 64 bits, Field: 32 bits) Fixes: #2216 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-28 14:30:34 +01:00
Rob Bradford	952f9bd3fe	vmm: acpi: Remove incorrect return statement _EJx built in should not return. dsdt.dsl 813: Return (CEJ0 (0x00)) Warning 3104 - ^ Reserved method should not return a value (_EJ0) dsdt.dsl 813: Return (CEJ0 (0x00)) Error 6080 - ^ Called method returns no value Fixes: #2216 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-28 14:30:34 +01:00
Rob Bradford	c29caf2a85	vmm: acpi: Fix incorrect mutex timeout value The mutex timeout should be 0xffff rather than 0xfff to disable the timeout feature. dsdt.dsl 745: Acquire (\_SB.PRES.CPLK, 0x0FFF) Warning 3130 - ^ Result is not used, possible operator timeout will be missed dsdt.dsl 767: Acquire (\_SB.PRES.CPLK, 0x0FFF) Warning 3130 - ^ Result is not used, possible operator timeout will be missed dsdt.dsl 775: Acquire (\_SB.PRES.CPLK, 0x0FFF) Warning 3130 - ^ Result is not used, possible operator timeout will be missed Fixes: #2216 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-28 14:30:34 +01:00
Bo Chen	ff54909d2c	vmm: openapi: Update the `fd` field of the `NetConfig` definition Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-01-28 09:11:39 +00:00
Bo Chen	6664e5a6e7	net_util, virtio-devices, vmm: Accept multiple TAP fds This patch enables multi-queue support for creating virtio-net devices by accepting multiple TAP fds, e.g. '--net fds=3:7'. Fixes: #2164 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-01-28 09:11:39 +00:00
Muminul Islam	a194dad98c	arch, vmm: Run KVM specific unit tests with kvm feature guard Signed-off-by: Muminul Islam <muislam@microsoft.com>	2021-01-28 09:11:02 +00:00
dependabot-preview[bot]	36360c7630	build(deps): bump libc from 0.2.82 to 0.2.83 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.82 to 0.2.83. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.82...0.2.83) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-01-28 06:30:03 +00:00
Wei Liu	b959849149	vmm: drop unnecessary semicolon Building with 1.51 nightly produces the following warning: warning: unnecessary trailing semicolon --> vmm/src/device_manager.rs:396:6 \| 396 \| }; \| ^ help: remove this semicolon \| = note: `#[warn(redundant_semicolons)]` on by default warning: 1 warning emitted Signed-off-by: Wei Liu <liuwe@microsoft.com>	2021-01-27 14:43:20 +00:00
dependabot-preview[bot]	192d69d601	build(deps): bump log from 0.4.13 to 0.4.14 Bumps [log](https://github.com/rust-lang/log) from 0.4.13 to 0.4.14. - [Release notes](https://github.com/rust-lang/log/releases) - [Changelog](https://github.com/rust-lang/log/blob/master/CHANGELOG.md) - [Commits](https://github.com/rust-lang/log/compare/0.4.13...0.4.14) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-01-27 12:48:52 +00:00
Rob Bradford	76e15a4240	vmm: acpi: Support compiling ACPI code on aarch64 This skeleton commit brings in the support for compiling aarch64 with the "acpi" feature ready to the ACPI enabling. It builds on the work to move the ACPI hotplug devices from I/O ports to MMIO and conditionalises any code that is x86_64 only (i.e. because it uses an I/O port.) Filling in the aarch64 specific details in tables such as the MADT it out of the scope. See: #2178 Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-26 15:19:02 +08:00
Sebastien Boeuf	f6c8e4b045	vmm: device_manager: Add info!() message about disk file backend It might be useful debugging information for the user to know what kind of disk file implementation is in use. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-01-22 16:10:34 +00:00
Sebastien Boeuf	2824642e80	virtio-devices: Rename BlockIoUring to Block Now that BlockIoUring is the only implementation of virtio-block, handling both synchronous and asynchronous backends based on the AsyncIo trait, we can rename it to Block. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-01-22 16:10:34 +00:00
Sebastien Boeuf	12e20effd7	block_util: Port synchronous QCOW file to AsyncIo trait Based on the synchronous QCOW file implementation present in the qcow crate, we created a new qcow_sync module in block_util that ports this synchronous implementation to the AsyncIo trait. The point is to reuse virtio-blk asynchronous implementation for both synchronous and asynchronous backends. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-01-22 16:10:34 +00:00
Sebastien Boeuf	9fc86a91e2	block_util: Port synchronous RAW file to AsyncIo trait Based on the synchronous RAW file implementation present in the qcow crate, we created a new raw_sync module in block_util that ports this synchronous implementation to the AsyncIo trait. The point is to reuse virtio-blk asynchronous implementation for both synchronous and asynchronous backends. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-01-22 16:10:34 +00:00
Sebastien Boeuf	da8ce25abf	virtio-devices: Use asynchronous traits for virtio-blk io_uring Based on the new DiskFile and AsyncIo traits, the implementation of asynchronous block support does not have to be tied to io_uring anymore. Instead, the only thing the virtio-blk implementation knows is that it is using an asynchronous implementation of the underlying disk file. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2021-01-22 16:10:34 +00:00
Rob Bradford	d578c408b7	vmm: acpi: Move DeviceManager ACPI device to an MMIO address Migrate the DeviceManager from a fixed I/O port address to an allocated MMIO address. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-22 16:08:41 +01:00
Rob Bradford	55a3a38e14	vmm: acpi: Move CpuManager ACPI device to an MMIO address Migrate the CpuManager from a fixed I/O port address to an allocated MMIO address. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-22 16:08:41 +01:00
Rob Bradford	6006068951	vmm: acpi: Move MemoryManager ACPI device to an MMIO address Migrate the MemoryManager from a fixed I/O port address to an allocated MMIO address. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-22 16:08:41 +01:00
Rob Bradford	28ab6cea0e	vmm: acpi: Move ACPI GED device to MMIO bus Currently the GED control is in a fixed I/O port address but instead use an MMIO address that has been chosen by the allocator. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-22 16:08:41 +01:00
Rob Bradford	4ebeeb1310	vmm: device_manager: Rename shutdown ACPI device Rename the variable for the shutdown device to clarify it's purpose in the function. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-22 16:08:41 +01:00
Bo Chen	8d418e1fca	vmm: Refine the seccomp filter list for the vCPU thread This patch refines the sccomp filter list for the vCPU thread, as we are no longer spawning virtio-device threads from the vCPU thread. Fixes: #2170 Signed-off-by: Bo Chen <chen.bo@intel.com>	2021-01-21 09:41:40 +00:00
dependabot-preview[bot]	c99d3f70ac	build(deps): bump signal-hook from 0.3.3 to 0.3.4 Bumps [signal-hook](https://github.com/vorner/signal-hook) from 0.3.3 to 0.3.4. - [Release notes](https://github.com/vorner/signal-hook/releases) - [Changelog](https://github.com/vorner/signal-hook/blob/master/CHANGELOG.md) - [Commits](https://github.com/vorner/signal-hook/commits/v0.3.4) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-01-17 20:27:52 +00:00
Rob Bradford	68d90cbe65	opeenapi: Add power button API entry point Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-13 17:00:39 +00:00
Rob Bradford	981bb72a09	vmm: api: Add "power-button" API entry point This will lead to the triggering of an ACPI button inside the guest in order to cleanly shutdown the guest. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-13 17:00:39 +00:00
Rob Bradford	843a010163	vmm: vm, device_manager: Trigger a power button notification via ACPI Use the ACPI GED device to trigger a notitifcation of type POWER_BUTTON_CHANGED which will ultimately lead to the guest being notified. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-13 17:00:39 +00:00
Rob Bradford	ff4d378f1b	vmm: device_manager: acpi: Add a power button (PNP0C0C) Add a power button to the devices defined in the \_SB Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-13 17:00:39 +00:00
Rob Bradford	7b376fa8e0	devices: acpi: Generalise the HotPlugNotificationFlags Renamed this bitfield as it will also be used for non-hotplug purposes such as synthesising a power button. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2021-01-13 17:00:39 +00:00
dependabot-preview[bot]	d83c9a74f4	build(deps): bump tempfile from 3.1.0 to 3.2.0 Bumps [tempfile](https://github.com/Stebalien/tempfile) from 3.1.0 to 3.2.0. - [Release notes](https://github.com/Stebalien/tempfile/releases) - [Changelog](https://github.com/Stebalien/tempfile/blob/master/NEWS) - [Commits](https://github.com/Stebalien/tempfile/commits) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-01-12 08:21:47 +00:00
dependabot-preview[bot]	d26866e018	build(deps): bump libc from 0.2.81 to 0.2.82 Bumps [libc](https://github.com/rust-lang/libc) from 0.2.81 to 0.2.82. - [Release notes](https://github.com/rust-lang/libc/releases) - [Commits](https://github.com/rust-lang/libc/compare/0.2.81...0.2.82) Signed-off-by: dependabot-preview[bot] <support@dependabot.com>	2021-01-12 07:06:21 +00:00

... 5 6 7 8 9 ...

1755 Commits