cloud-hypervisor

mirror of https://github.com/cloud-hypervisor/cloud-hypervisor.git synced 2024-12-24 06:35:20 +00:00

Author	SHA1	Message	Date
Rob Bradford	a551398135	vmm: device_manager: Use MemoryManager to create KVM mapping Use the newly exported funtionality to reduce the amount of duplicated code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	962dec2913	vmm: memory_manager: Refactor KVM userspace mapping creation This function will be useful for other parts of the VMM that also estabilish their own mappings. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	7df88793a0	vmm: device_manager: Get device range from MemoryManager This removes the duplication of these values. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	61cfe3e72d	vmm: Obtain sequential KVM memory slot numbers from MemoryManager This removes the need to handle a mutable integer and also centralises the allocation of these slot numbers. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	260cebb8cf	vmm: Introduce MemoryManager The memory manager is responsible for setting up the guest memory and in the long term will also handle addition of guest memory. In this commit move code for creating the backing memory and populating the allocator into the new implementation trying to make as minimal changes to other code as possible. Follow on commits will further reduce some of the duplicated code. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-23 10:25:40 +00:00
Rob Bradford	d5682cd306	vmm: device_manager: Rewrite if chain using match To reflect updated clippy rules: error: `if` chain can be rewritten with `match` --> vmm/src/device_manager.rs:1508:25 \| 1508 \| / if ret > 0 { 1509 \| \| debug!("MSI message successfully delivered"); 1510 \| \| } else if ret == 0 { 1511 \| \| warn!("failed to deliver MSI message, blocked by guest"); 1512 \| \| } \| \|_________________________^ \| = note: `-D clippy::comparison-chain` implied by `-D warnings` = help: Consider rewriting the `if` chain to use `cmp` and `match`. = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#comparison_chain Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	21b88c3ea0	vmm: cpu: Rewrite if chain using match Address updated clippy error: error: `if` chain can be rewritten with `match` --> vmm/src/cpu.rs:668:9 \| 668 \| / if desired_vcpus > self.present_vcpus() { 669 \| \| self.activate_vcpus(desired_vcpus, None)?; 670 \| \| } else if desired_vcpus < self.present_vcpus() { 671 \| \| self.mark_vcpus_for_removal(desired_vcpus)?; 672 \| \| } \| \|_________^ \| = note: `-D clippy::comparison-chain` implied by `-D warnings` = help: Consider rewriting the `if` chain to use `cmp` and `match`. = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#comparison_chain Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	e25a47b32c	vmm: device_manager: Remove redundant clones Address updated clippy errors: error: redundant clone --> vmm/src/device_manager.rs:699:32 \| 699 \| .insert(acpi_device.clone(), 0x3c0, 0x4) \| ^^^^^^^^ help: remove this \| = note: `-D clippy::redundant-clone` implied by `-D warnings` note: this value is dropped without further use --> vmm/src/device_manager.rs:699:21 \| 699 \| .insert(acpi_device.clone(), 0x3c0, 0x4) \| ^^^^^^^^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone error: redundant clone --> vmm/src/device_manager.rs:737:26 \| 737 \| .insert(i8042.clone(), 0x61, 0x4) \| ^^^^^^^^ help: remove this \| note: this value is dropped without further use --> vmm/src/device_manager.rs:737:21 \| 737 \| .insert(i8042.clone(), 0x61, 0x4) \| ^^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone error: redundant clone --> vmm/src/device_manager.rs:754:29 \| 754 \| .insert(cmos.clone(), 0x70, 0x2) \| ^^^^^^^^ help: remove this \| note: this value is dropped without further use --> vmm/src/device_manager.rs:754:25 \| 754 \| .insert(cmos.clone(), 0x70, 0x2) \| ^^^^ = help: for further information visit https://rust-lang.github.io/rust-clippy/master/index.html#redundant_clone Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-20 00:52:03 +01:00
Rob Bradford	a6878accd5	vmm: cpu: Implement CPU removal When the running OS has been told that a CPU should be removed it will shutdown the CPU and then signal to the hypervisor via the "_EJ0" method on the device that ultimately writes into an I/O port than the vCPU should be shutdown. Upon notification the hypervisor signals to the individual thread that it should shutdown and waits for that thread to end. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	7b3fc72aea	vmm: cpu: Notify guest OS that it should offline vCPUs Allow the resizing of the number of vCPUs to less than the current active vCPUs. This does not currently remove them from the system but the kernel will take them offline. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	7e81b0ded7	vmm: cpu: Create vCPU state for all possible vCPUs This will make it more straightforward when we attempt to remove vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-18 08:23:53 +00:00
Rob Bradford	156ea392a2	vmm: cpu: Only do ACPI notify on newly added vCPUs When we add a vCPU set an "inserting" boolean that is exposed as an ACPI field that will be checked for and reset when the ACPI GED notification for CPU devices happens. This change is a precursor for CPU unplug. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-16 23:57:14 +01:00
Rob Bradford	e8313e3e69	vmm: acpi: Refactor ACPI CPU notification Continue to notify on all vCPUs but instead separate the notification functionality into two methods, CSCN that walks through all the CPUs and CTFY which notifies based on the numerical CPU id. This is an interim step towards only notifying on changed CPUs and ultimately CPU removal. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-16 23:57:14 +01:00
Sebastien Boeuf	d1390906c8	vmm: config: Derive Debug and PartialEq for configuration structures In anticipation for the writing of unit tests comparing two VmConfig structures, this commit derives the PartialEq trait for VmConfig and all embedded structures. This patch also derives the Debug trait for the same set of structures so that we can print them to facilitate debugging. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Sebastien Boeuf	93f5f6ed45	vmm: config: Provide a default empty command line through OpenAPI The OpenAPI should not have to provide a command line since the CLI considers the command line as an empty string if nothing is provided. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Sebastien Boeuf	43bd0e53c4	main: Move VmParams creation into a dedicated function This brings more modularity to the code, which will be helpful when we will later test the CLI and OpenAPI generate the same VmConfig output. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-16 16:48:59 +01:00
Samuel Ortiz	f0b7412495	vmm: device_manager: Add all virtio devices to the migratable list We want to track all migratable devices through the DeviceManager. Fixes: #341 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	37557c8b35	vmm: vm: Implement the Pausable trait Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	9756fc2dd0	vmm: cpu_manager: Implement the Pausable trait Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	35dd1523c9	vmm: device_manager: Implement the Pausable trait Since the Snapshotable placeholder and Migratable traits are provided as well, the DeviceManager object and all its objects are now Migratable. All Migratable devices are tracked as Arc<Mutex<dyn Migratable>> references. Keeping track of all migratable devices allows for implementing the Migratable trait for the DeviceManager structure, making the whole device model potentially migratable. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Samuel Ortiz	35d7721683	vmm: Convert virtio devices to Arc<Mutex<T>> Migratable devices can be virtio or legacy devices. In any case, they can potentially be tracked through one of the IO bus as an Arc<Mutex<dyn BusDevice>>. In order for the DeviceManager to also keep track of such devices as Migratable trait objects, they must be shared as mutable atomic references, i.e. Arc<Mutex<T>>. That forces all Migratable objects to be tracked as Arc<Mutex<dyn Migratable>>. Virtio devices are typically migratable, and thus for them to be referenced by the DeviceManager, they now should be built as Arc<Mutex<VirtioDevice>>. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-12-12 08:50:36 +01:00
Sebastien Boeuf	64c5e3d8cb	vmm: api: Adjust FsConfig for OpenAPI The FsConfig structure has been recently adjusted so that the default value matches between OpenAPI and CLI. Unfortunately, with the current description, there is no way from the OpenAPI to describe a cache_size value "None", so that DAX does not get enabled. Usually, using a Rust "Option" works because the default value is None. But in this case, the default value is Some(8G), which means we cannot describe a None. This commit tackles the problem, introducing an explicit parameter "dax", and leaving "cache_size" as a simple u64 integer. This way, the default value is dax=true and cache_size=8G, but it lets the opportunity to disable DAX entirely with dax=false, which will simply ignore the cache_size value. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	4bfd51cc42	vmm: api: Match VhostUserBlkConfig defaults between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the VhostUserBlkConfig structure, this patch defines some default values for num_queues, queue_size and wce. num_queues is 1, queue_size is 128 and wce is true. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	1c2587f8cb	vmm: api: Match VhostUserNetConfig defaults between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the VhostUserNetConfig structure, this patch defines some default values for num_queues, queue_size and mac. num_queues is 2 since that's a pair of TX/RX queues, queue_size is 256 and mac is a randomly generated value. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	5e0bbf9c3b	vmm: Don't factorize vhost-user configurations We want to set different default configurations for vhost-user-net and vhost-user-blk, which is the reason why the common part corresponding to the number of queues and the queue size cannot be embedded. This prepares for the following commit, matching API and CLI behaviors. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	793327cff8	vmm: api: Make ConsoleConfig default match between CLI and HTTP API A simple patch making sure the field "file" is provisioned with the same default value through CLI and OpenAPI. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	cc08c44cb9	vmm: api: Make MemoryConfig default match between CLI and HTTP API Just making sure we have a serde default for the field "file" since it is not a required field in the OpenAPI definition. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Sebastien Boeuf	5a72225856	vmm: api: Update CpuConfig name to match the internal name All structures match between the OpenAPI definition and the internal configuration code, that's why CpuConfig is being renamed into CpusConfig. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-11 15:50:24 +00:00
Rob Bradford	c61104df47	vmm: Port to latest vmm-sys-util The signal handling for vCPU signals has changed in the latest release so switch to the new API. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-11 14:11:11 +00:00
Sebastien Boeuf	ee528ae808	vmm: api: Make FsConfig defaults match between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the FsConfig structure, this patch defines some default values for num_queues, queue_size and the cache_size. num_queues is set to 1, queue_size is set to 1024, and cache_size is set to Some(8G) which means that DAX is enabled by default with a shared region of 8GiB. Fixes #508 Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-09 23:42:23 -08:00
Sebastien Boeuf	befd342da4	vmm: api: Make NetConfig defaults match between CLI and HTTP API In order to let the CLI and the HTTP API behave the same regarding the NetConfig structure, this patch defines some default values for tap, ip, mask, mac and iommu. tap is None, ip is 192.168.249.1, mask is 255.255.255.0, mac is a randomly generated value, and iommu is false. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-09 23:19:24 -08:00
Jose Carlos Venegas Munoz	99e608c240	openapi: Fix schema Fix openapi schema to be a valid yaml. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-12-09 14:30:15 -08:00
Rob Bradford	f994665610	vmm: Reduce the minimum IRQ constant Now that the GED device does not use a hardcoded IRQ number the starting IRQ number can be restored (needed for the hardcoded serial port IRQ.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-09 16:58:00 +00:00
Rob Bradford	ba59c62044	vmm, devices: Remove hardcoded IRQ number for GED device Remove the previously hardcoded IRQ number used for the GED device. Instead allocate the IRQ using the allocator and use that value in the definition in the ACPI device. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-09 16:58:00 +00:00
Sebastien Boeuf	aa94e9b8f3	Revert "vmm: api: Modify FsConfig to be OpenAPI friendly" This reverts commit `defc5dcd9c`.	2019-12-06 18:08:10 +00:00
Rob Bradford	9b1ba14f2d	vmm: Delegate device related ACPI DSDT table work to DeviceManager Move the code for handling the creation of the DSDT entries for devices into the DeviceManager. This will make it easier to handle device hotplug and also in the future remove some hardcoded ACPI constants. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 17:44:00 +00:00
Rob Bradford	60e6609011	vmm: Delegate CPU related ACPI tables to CpuManager Move the code for generating the MADT (APIC) table and the DSDT generation for CPU related functionality into the CpuManager. There is no functional change just code rearrangement. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 17:44:00 +00:00
Sebastien Boeuf	defc5dcd9c	vmm: api: Modify FsConfig to be OpenAPI friendly When consumer of the HTTP API try to interact with cloud-hypervisor, they have to provide the equivalent of the config structure related to each component they need. Problem is, the Rust enum type "Option" cannot be obtained from the OpenAPI YAML definition. This patch intends to fix this inconsistency between what is possible through the CLI and what's possible through the HTTP API by using simple types bool and int64 instead of Option<u64>. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-06 06:38:48 -08:00
Rob Bradford	59d01712ad	vmm: Remove kernel based IOAPIC handling from the device manager Previously the device setup code assumed that if no IOAPIC was passed in then the device should be added to the kernel irqchip. As an earlier change meant that there was always a userspace IOAPIC this kernel based code can be removed. The accessor still returns an Option type to leave scope for implementing a situation without an IOAPIC (no serial or GED device). This change does not add support no-IOAPIC mode as the original code did not either. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00
Rob Bradford	afea6a10a2	vmm: Stop initialising kernel based IOAPIC/PIC Now that we require the modern capabilities we can stop creating a kernel base irqchip. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00
Rob Bradford	9b1cb9621f	vmm: Remove pin based interrupt setup for virtio devices With MSI now required remove pin based interrupt support from all the virtio PCI device setup. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00
Rob Bradford	72fb687e3f	vmm: Check for required capabilities We now require CAP_SIGNAL_MSI, CAP_TSC_DEADLINE_TIMER and CAP_SPLIT_IRQCHIP. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-06 12:34:06 +01:00
Rob Bradford	f98b16f308	vmm: Update the configuration to preserve hot-plug CPUs after reboot Update the configuration after a resize to ensure that after a reboot the added vCPUs are preserved. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-05 16:39:19 +00:00
Rob Bradford	1722708612	vmm: Switch to storing VmConfig inside an Arc<Mutex<>> This permits the runtime reconfiguration of the VM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-05 16:39:19 +00:00
Rob Bradford	c063bb8d30	vmm: acpi: Make GED interrupt edge triggered This was causing issues when the kernel was trying to reset the interrupt and making the reboot fail. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-05 16:39:19 +00:00
Qiu Wenbo	e1af17d93a	vmm: Restore tty to canonical mode when SIGTERM or SIGINT received The tty mode remains raw mode when cloud-hypervisor is terminted by SIGTERM or SIGINT. The terminal is unusable due to echoing is disabled which is really annoying. Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2019-12-05 01:29:26 -08:00
Qiu Wenbo	5208ff86c8	vmm: Detect and handle AMD SME (Secure Memory Encryption) Some physical address bits may become reserved in page table when SME is enabled on AMD platform. Guest will trigger a reserved bit violation page fault in this case due to write these reserved bits to 1 in page table. We need reduce the reserved bits to get the right physical address range. Signed-off-by: Qiu Wenbo <qiuwenbo@phytium.com.cn>	2019-12-04 14:46:44 +00:00
Sebastien Boeuf	08258d5dad	vfio: pci: Allow multiple devices to be passed through The KVM_SET_GSI_ROUTING ioctl is very simple, it overrides the previous routes configuration with the new ones being applied. This means the caller, in this case cloud-hypervisor, needs to maintain the list of all interrupts which needs to be active at all times. This allows to correctly support multiple devices to be passed through the VM and being functional at the same time. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-12-04 08:48:17 +01:00
Rob Bradford	17badfbff5	vmm: cpu: Call vcpu configure() on the vCPU thread The function that programs the vCPUs is expected to be run from within each vCPU thread. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-03 03:22:15 -08:00
Rob Bradford	13503061e6	api: Fix OpenAPI specification entries Some renames from "cpu_count" were missing. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-03 03:28:06 +01:00
Rob Bradford	66a31c19e8	vmm: acpi: Upon GED interrupt notify on all vCPUs Call the "CTFY" method that will itself call Notify() on the CPU objects in the ACPI namespace. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	48bf141364	vmm: Trigger a hotplug device notification when resizing When adjusting the number of vCPUs generate a hotplug notification. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	b629727901	vmm: acpi: Add a CTFY method to notify on all CPU objects This method calls Notify() on all the vCPU objects in the ACPI namespace. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	ae9359c859	vmm: acpi: Create the CPU entries in the DSDT for all vCPUs CPU entries need to be created for every potential vCPU in the system. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	791ca3388f	vmm: device_manager: Add ability to notify via GED device Add ability to notify via the GED device that there is some new hotplug activity. This will be used by the CpuManager (and later DeviceManager itself) to notify of new hotplug activity. Currently it has a hardcoded IRQ of 5 as the ACPI tables also need to refer to this IRQ and the IRQ allocation does not permit the allocation of specific IRQs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	7ad68d499a	vmm: device_manager: Allocate I/O port for ACPI shutdown device The refactoring in `ce1765c8af` dropped the code to allocate the I/O port. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	86339b4cb4	vmm: Add HTTP API to resize the VM Currently only increasing the number of vCPUs is supported but in the future it will be extended. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	e7d4eae527	vmm: cpu: Add support for starting more vCPU threads Add support for starting vCPU threads after the initial boot ones. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	0ef999978c	vmm: cpu: Support only partially configuring the vCPU When configuring a processor after boot as a hotplug CPU we only configure a subset of the CPU state. In particular we should not configure the FPU, segment registers (or reconfigure the paging which is a side-effect of that) nor the main registers. Achieve this by making the function take an Option type for the start address. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	c8b3041e62	vmm: openapi: Update OpenAPI for CpuConfig struct This struct has changed in order to support differentiating between boot and max vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	b6801e355e	vmm: cpu: Refactor vCPU thread starting Refactor the vCPU thread starting so that there is the possibility to bring on extra vCPU threads. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	66d5163ee7	vmm: cpu: Encapsulate vCPU state into its own struct Currently this just holds the thread handle but will be enlarged to encompass details such as whether the vCPU is currently being inserted or ejected. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	1bbe48b24c	vmm: acpi: Mark non-boot vCPUs as disabled in the MADT table The MADT table contains the details of all the potential vCPUs and whether they are present at boot (as indicated by the flags field.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	82bc07cce4	vmm: Add boot and max vCPU handling to command line parser Also retain support (with a warning for the old behaviour.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	7543e00a07	vmm: Use new CpuManager accessor to get boot vCPUs When initialising the ACPI tables and configuring the VM use the new accessor on the CpuManager to get the number of boot vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Rob Bradford	df0907845a	vmm: cpu: Introduce concept of maximum vs boot vCPUs in CpuManager For now the max vCPUs is the same as the boot vCPUs. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-12-02 13:49:04 +00:00
Samuel Ortiz	0f21781fbe	cargo: Bump the kvm and vmm-sys-util crates Since the kvm crates now depend on vmm-sys-util, the bump must be atomic. The kvm-bindings and ioctls 0.2.0 and 0.4.0 crates come with a few API changes, one of them being the use of a kvm_ioctls specific error type. Porting our code to that type makes for a fairly large diff stat. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-29 17:48:02 +00:00
Jose Carlos Venegas Munoz	ab16af2941	openapi: make context ID vsock int64 context ID on vsock man defines a 32-bits value, openapi default integer is a signed 32-bits value. This could lead to miss one bit during castings for typed client implmentations. Lets increase the range of valid values by requesting an int64. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-11-26 08:38:59 +01:00
Sebastien Boeuf	f979380620	vmm: Mark guest persistent memory pages as mergeable In case the VM is started with the flag "--pmem mergeable=on", it means the user expects the guest persistent memory pages to be marked as mergeable. This commit relies on the madvise(MADV_MERGEABLE) system call to inform the host kernel about these pages. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-11-22 15:28:10 +00:00
Sebastien Boeuf	0f9afc3017	vmm: Add mergeable=on\|off option to --pmem flag In order to let the user indicate if the persistent memory pages should be marked as mergeable or not, a new option is being introduced. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-11-22 15:28:10 +00:00
Sebastien Boeuf	e4e8062dda	vmm: Mark guest RAM pages as mergeable In case the VM is started with the flag "--memory mergeable=on", it means the user expects the guest RAM pages to be marked as mergeable. This commit relies on the madvise(MADV_MERGEABLE) system call to inform the host kernel about these pages. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-11-22 15:28:10 +00:00
Sebastien Boeuf	880f62bab8	vmm: Add mergeable=on\|off option to --memory flag In order to let the user indicate if the guest RAM pages should be marked as mergeable or not, a new option is being introduced. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-11-22 15:28:10 +00:00
Jose Carlos Venegas Munoz	1d852e9ce5	vmm: Provide vmm version to start_vmm_thread When vmm.ping give a response, we expect get the version from the VMM not the vmm create Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-11-21 15:04:11 -08:00
Jose Carlos Venegas Munoz	a518651402	http: api: implement vmm.ping vmm.ping will help to check if http API server is up and running. This also removes the vmm.info endpoint. Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-11-21 15:04:11 -08:00
Rob Bradford	348a1bc30e	vmm: cpu: Allocate I/O port for the CPU manager The CPU manager uses an I/O port and to prevent potential clashes with assignment for PCI devices ensure that it is allocated by the allocator. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	07cdb37dda	vmm: cpu & acpi: Query CPU manager for CPU status Rather than hardcode the CPU status for all the CPUs instead query from the CPU manager via the I/O port that is is on via the ACPI tables. Each CPU device has a _STA method that calls into the CSTA method which reads and writes the I/O ports via the PRST field which exposes the I/O port through and OpRegion. As we only support boot CPUS report that all the CPUs are enabled for now. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	5faf8b756c	vmm: acpi: Add an _MAT for the CPU devices containing a LAPIC The Linux kernel expects all CPUs, whether they be enabled or disabled to have an _MAT entry containing the LAPIC details for this CPU with the enabled bit set to 1 (in the flags.) In the MADT table the same bit is used to determine if the CPU is present at boot vs available later. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	1da0ff395d	vmm: cpu: Add the CpuManager onto the IO bus This allows the kernel (via ACPI based controls) to query and control the CPU state. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	50c8335d3d	vmm: device_manager: Expose the SystemAllocator This allows other code to allocate I/O ports for use on the (already) exposed IO bus. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Rob Bradford	1ac1231292	vmm: Encase CpuManager within an Arc<Mutex<>> This is necessary to be able to add the CpuManager onto the IO bus. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-21 09:17:15 -08:00
Samuel Ortiz	f0e618431d	vmm: device_manager: Use consistent naming when adding devices When adding devices to the guest, and populating the device model, we should prefix the routines with add_. When we're just creating the device objects but not yet adding them we use make_. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	a2ee681665	vmm: device_manager: Add an MMIO devices creation routine In order to reduce the DeviceManager's new() complexity, we can move the MMIO devices creation code into its own routine. Fixes: #441 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	79b8f8e477	vmm: device_manager: Add a PCI devices creation routine In order to reduce the DeviceManager's new() complexity, we can move the PCI devices creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	5087f633f6	vmm: device_manager: Add an IOAPIC creation routine In order to reduce the DeviceManager's new() complexity, we can move the ACPI device creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	ce1765c8af	vmm: device_manager: Add an ACPI device creation routine In order to reduce the DeviceManager's new() complexity, we can move the ACPI device creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	cfca2759fc	vmm: device_manager: Add a legacy devices creation routine In order to reduce the DeviceManager's new() complexity, we can move the legacy devices creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	4b469b98cf	vmm: device_manager: Add a console creation routine In order to reduce the DeviceManager's new() complexity, we can move the console creation code into its own routine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-19 13:36:21 -08:00
Samuel Ortiz	b930b3fb41	vmm: api: Specify which integers are 64 bit wide By default, client will assume 32-bits for OpenAPI interger types. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-12 08:39:05 -08:00
Samuel Ortiz	6af2f57644	vmm: api: Fix the vm.info response payload We are returning a state and a config. Fixes: #431 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-12 08:39:05 -08:00
Rob Bradford	6958ec4922	vmm: Move CPU management code to its own module Move CpuManager, Vcpu and related functionality to its own module (and file) inside the VMM crate Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-11 15:46:24 +00:00
Samuel Ortiz	3dde848c8f	vmm: api: Update our OpenAPI document In most cases we return a 204 (No Content) and not a 201. In those cases, we do not send any HTTP body back at all. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-10 14:51:55 +01:00
Samuel Ortiz	96aa2441ad	vmm: http: Convert to micro_http HttpServer The new micro_http package provides a built-in HttpServer wrapper for running a more robust HTTP server based on the package HTTP API. Switching to this implementation allows us to, among other things, handle HTTP requests that are larger than 1024 bytes. Fixes: #423 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-10 14:51:55 +01:00
Samuel Ortiz	f34ace7673	vmm: http_endpoint: Do not sent 200 status code when our body is empty Otherwise HTTP client will not close the connection and wait for a pending body. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-11-10 14:51:55 +01:00
Jose Carlos Venegas Munoz	ede262684d	API: HTTP: change response content type to JSON The HTTP API responses are encoded in json Suggested-by: Samuel Ortiz <sameo@linux.intel.com> Tested-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com> Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-11-08 22:49:08 +01:00
Rob Bradford	3c715daa9d	vmm: Fix rustfmt failure by removing extra ";" Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-08 20:43:52 +00:00
Rob Bradford	a1a5fe0c93	vmm: Split CPU management into it's own struct Pull details of vCPU management (booting, pausing, resuming, shutdown) into it's own structure. This will ultimately enable this to be moved to its own file and encapsulate all the vCPU handling for the VMM. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-08 11:59:21 +01:00
Rob Bradford	0319a4a09a	arch: vmm: Move ACPI tables creation to vmm crate Remove ACPI table creation from arch crate to the vmm crate simplifying arch::configure_system() GuestAddress(0) is used to mean no RSDP table rather than adding complexity with a conditional argument or an Option type as it will evaluate to a zero value which would be the default anyway. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-07 14:02:27 +00:00
Cathy Zhang	5cd4f5daeb	vmm: Release the old vm before build a new one In vm_reboot, while build the new vm, the old one pointed by self.vm is not released, that is, the tap opened by self.vm is not closed either. As a result, the associated dev name slot in host kernel is still in use state, which prevents the new build from picking it up as the new opened tap's name, but to use the name in next slot finally. Call self.vm_shutdown instead here since it has call take() on vm reference, which could ensure the old vm is destructed before the new vm build. Signed-off-by: Cathy Zhang <cathy.zhang@intel.com>	2019-11-05 14:40:43 +01:00
Rob Bradford	b3388c343d	vmm: device_manager: Ensure I/O ports are allocated Ensure that we tell the allocator about all the I/O ports that we are using for I/O bus attached devices (serial, i8042, ACPI device.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-11-05 10:13:01 +00:00
Sebastien Boeuf	5694ac2b1e	vm-virtio: Create new VirtioTransport trait to abstract ioeventfds In order to group together some functions that can be shared across virtio transport layers, this commit introduces a new trait called VirtioTransport. The first function of this trait being ioeventfds() as it is needed from both virtio-mmio and virtio-pci devices, represented by MmioDevice and VirtioPciDevice structures respectively. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	3fa5df4161	vmm: Unregister old ioeventfds when reprogramming PCI BAR Now that kvm-ioctls has been updated, the function unregister_ioevent() can be used to remove eventfd previously associated with some specific PIO or MMIO guest address. Particularly, it is useful for the PCI BAR reprogramming case, as we want to ensure the eventfd will only get triggered by the new BAR address, and not the old one. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	587a420429	cargo: Update to the latest kvm-ioctls version We need to rely on the latest kvm-ioctls version to benefit from the recent addition of unregister_ioevent(), allowing us to detach a previously registered eventfd to a PIO or MMIO guest address. Because of this update, we had to modify the current constraint we had on the vmm-sys-util crate, using ">= 0.1.1" instead of being strictly tied to "0.2.0". Once the dependency conflict resolved, this commit took care of fixing build issues caused by recent modification of kvm-ioctls relying on EventFd reference instead of RawFd. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	c7cabc88b4	vmm: Conditionally update ioeventfds for virtio PCI device The specific part of PCI BAR reprogramming that happens for a virtio PCI device is the update of the ioeventfds addresses KVM should listen to. This should not be triggered for every BAR reprogramming associated with the virtio device since a virtio PCI device might have multiple BARs. The update of the ioeventfds addresses should only happen when the BAR related to those addresses is being moved. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	de21c9ba4f	pci: Remove ioeventfds() from PciDevice trait The PciDevice trait is supposed to describe only functions related to PCI. The specific method ioeventfds() has nothing to do with PCI, but instead would be more specific to virtio transport devices. This commit removes the ioeventfds() method from the PciDevice trait, adding some convenient helper as_any() to retrieve the Any trait from the structure behing the PciDevice trait. This is the only way to keep calling into ioeventfds() function from VirtioPciDevice, so that we can still properly reprogram the PCI BAR. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-31 09:30:59 +01:00
Sebastien Boeuf	d6c68e4738	pci: Add error propagation to PCI BAR reprogramming Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	3e819ac797	pci: Use a weak reference to the AddressManager Storing a strong reference to the AddressManager behind the DeviceRelocation trait results in a cyclic reference count. Use a weak reference to break that dependency. Signed-off-by: Rob Bradford <robert.bradford@intel.com> Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	149b61b213	pci: Detect BAR reprogramming Based on the value being written to the BAR, the implementation can now detect if the BAR is being moved to another address. If that is the case, it invokes move_bar() function from the DeviceRelocation trait. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	04a449d3f3	pci: Pass DeviceRelocation to PciBus In order to trigger the PCI BAR reprogramming from PciConfigIo and PciConfigMmmio, we need the PciBus to have a hold onto the trait implementation of DeviceRelocation. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	e93467a96c	vmm: Implement DeviceRelocation trait By implementing the DeviceRelocation trait for the AddressManager structure, we now have a way to let the PCI BAR reprogramming happen. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	8746c16593	vmm: Create AddressManager to own SystemAllocator In order to reuse the SystemAllocator later at runtime, it is moved into the new structure AddressManager. The goal is to have a hold onto the SystemAllocator and both IO and MMIO buses so that we can use them later. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Sebastien Boeuf	1870eb4295	devices: Lock the BtreeMap inside to avoid deadlocks Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-29 16:48:02 +01:00
Jose Carlos Venegas Munoz	78e2f7a99a	api: http: handle cpu according to openapi openapi definition defines an object for cpus not an integer Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-10-17 07:39:56 +02:00
Jose Carlos Venegas Munoz	205b8c1cd5	api: http: make consistent api and implementation vsocks: vsocks is implemented as an array Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-10-17 07:39:56 +02:00
Sebastien Boeuf	3acf9dfcf3	vfio: Don't map guest memory for VFIO devices attached to vIOMMU In case a VFIO devices is being attached behind a virtual IOMMU, we should not automatically map the entire guest memory for the specific device. A VFIO device attached to the virtual IOMMU will be driven with IOVAs, hence we should simply wait for the requests coming from the virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	63c30a6e79	vmm: Build and set the list of external mappings for VFIO When VFIO devices are created and if the device is attached to the virtual IOMMU, the ExternalDmaMapping trait implementation is created and associated with the device. The idea is to build a hash map of device IDs with their associated trait implementation. This hash map is provided to the virtual IOMMU device so that it knows how to properly trigger external mappings associated with VFIO devices. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	837bcbc6ba	vfio: Create VFIO implementation of ExternalDmaMapping With this implementation of the trait ExternalDmaMapping, we now have the tool to provide to the virtual IOMMU to trigger the map/unmap on behalf of the guest. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	3598e603d5	vfio: Add a public function to retrive VFIO container The VFIO container is the object needed to update the VFIO mapping associated with a VFIO device. This patch allows the device manager to have access to the VFIO container. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	9085a39c7d	vmm: Attach VFIO devices to IORT table This patch attaches VFIO devices to the virtual IOMMU if they are identified as they should be, based on the option "iommu=on". This simply takes care of adding the PCI device ID to the ACPI IORT table. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Sebastien Boeuf	5fc3f37c9b	vmm: Add iommu=on\|off option for --device Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a VFIO device should be attached to the virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-16 07:27:06 +02:00
Jose Carlos Venegas Munoz	786e33931f	api: http: Fix openpi schema. Fix invalid type for version: - VmInfo.version.type string Change Null value from enum as it has problems to build clients with openapi tools. - ConsoleConfig.mode.enum Null -> Nil Signed-off-by: Jose Carlos Venegas Munoz <jose.carlos.venegas.munoz@intel.com>	2019-10-15 07:16:24 +02:00
Samuel Ortiz	2a0ba7aef8	vmm: vm: Add state validation unit test Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-14 06:35:36 +02:00
Samuel Ortiz	097b30669f	vmm: vm: Verify that state transitions are valid We should return an explicit error when the transition from on VM state to another is invalid. The valid_transition() routine for the VmState enum essentially describes the VM state machine. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-14 06:35:36 +02:00
Samuel Ortiz	d2d3abb13c	vmm: Rename Booted vm state to Running Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-10 17:13:44 -07:00
Samuel Ortiz	dbbd04a4cf	vmm: Implement VM resume To resume a VM, we unpark all its vCPU threads. Fixes: #333 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-10 17:13:44 -07:00
Samuel Ortiz	4ac0cb9cff	vmm: Implement VM pause In order to pause a VM, we signal all the vCPU threads to get them out of vmx non-root. Once out, the vCPU thread will check for a an atomic pause boolean. If it's set to true, then the thread will park until being resumed. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-10 17:13:44 -07:00
Samuel Ortiz	1298b508bf	vmm: Manage the exit and reset behaviours from the control loop So that we don't need to forward an ExitBehaviour up to the VMM thread. This simplifies the control loop and the VMM thread even further. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-08 18:03:27 -07:00
Samuel Ortiz	a95fa1c4e8	vmm: api: Add a VMM shutdown command This shuts the current VM down, if any, and then exits the VMM process. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-08 18:03:27 -07:00
Samuel Ortiz	228adebc32	vmm: Unreference the VM when shutting down This way, we are forced to re-create the VM object when moving from shutdown to boot. Fixes: #321 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-08 17:24:05 +02:00
Sebastien Boeuf	b918220b49	vmm: Support virtio-pci devices attached to a virtual IOMMU This commit is the glue between the virtio-pci devices attached to the vIOMMU, and the IORT ACPI table exposing them to the guest as sitting behind this vIOMMU. An important thing is the trait implementation provided to the virtio vrings for each device attached to the vIOMMU, as they need to perform proper address translation before they can access the buffers. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	278ab05cbc	vmm: Add iommu=on\|off option for --vsock Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-vsock device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	32d07e40cc	vmm: Add iommu=on\|off option for --console Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-console device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	63869bde75	vmm: Add iommu=on\|off option for --pmem Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-pmem device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	fb4769388b	vmm: Add iommu=on\|off option for --rng Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-rng device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	20c4ed829a	vmm: Add iommu=on\|off option for --net Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-net device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	4b8d7e718d	vmm: Add iommu=on\|off option for --disk Having the virtual IOMMU created with --iommu is one thing, but we also need a way to decide if a virtio-blk device should be attached to this virtual IOMMU or not. That's why we introduce an extra option "iommu" with the value "on" or "off". By default, the device is not attached, which means "iommu=off". One side effect of this new option is that we had to introduce a new option for the disk path, simply called "path=". Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	6e0aa56f06	vmm: Add iommu field to the VmConfig Adding a simple iommu boolean field to the VmConfig structure so that we can later use it to create a virtio-iommu device for the current VM. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	03352f45f9	arch: Create ACPI IORT table The virtual IOMMU exposed through virtio-iommu device has a dependency on ACPI. It needs to expose the device ID of the virtio-iommu device, and all the other devices attached to this virtual IOMMU. The IDs are expressed from a PCI bus perspective, based on segment, bus, device and function. The guest relies on the topology description provided by the IORT table to attach devices to the virtio-iommu device. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	0acb1e329d	vm-virtio: Translate addresses for devices attached to IOMMU In case some virtio devices are attached to the virtual IOMMU, their vring addresses need to be translated from IOVA into GPA. Otherwise it makes no sense to try to access them, and they would cause out of range errors. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	6566c739e1	vm-virtio: Add IOMMU support to virtio-vsock Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	9ab00dcb75	vm-virtio: Add IOMMU support to virtio-rng Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	ee1899c6f6	vm-virtio: Add IOMMU support to virtio-pmem Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	392f1ec155	vm-virtio: Add IOMMU support to virtio-console Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	9fad680db1	vm-virtio: Add IOMMU support to virtio-net Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	9ebb1a55bc	vm-virtio: Add IOMMU support to virtio-blk Adding virtio feature VIRTIO_F_IOMMU_PLATFORM when explicitly asked by the user. The need for this feature is to be able to attach the virtio device to a virtual IOMMU. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Sebastien Boeuf	8225d4cd6e	vm-virtio: Implement reset() for virtio-console The virtio specification defines a device can be reset, which was not supported by this virtio-console implementation. The reason it is needed is to support unbinding this device from the guest driver, and rebind it to vfio-pci driver. Signed-off-by: Sebastien Boeuf <sebastien.boeuf@intel.com>	2019-10-07 10:12:07 +02:00
Samuel Ortiz	2a466132a0	vmm: api: Set the HTTP response header Server field To "Cloud Hypervisor API" and not "Firecracker API". Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	7abbad0a62	vmm: Be more idiomatic when calling into the VMM API Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	7328ecdb3b	vmm: Implement the /api/v1/vm.delete endpoint Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	f9daf2e247	vmm: Factorize the vm boot and shutdown code So that the API handling state machine is cleaner and easier to read. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	43b3642955	vmm: Clean Error handling up We used to have errors definitions spread across vmm, vm, api, and http. We now have a cleaner separation: All API routines only return an ApiResult. All VM operations, including the VMM wrappers, return a VmResult. This makes it easier to carry errors up to the HTTP caller. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	42758244a0	vmm: Implement the /api/v1/vm.info endpoint This, for now, returns the VM config and its state. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	27af983ec9	vmm: Track the VM state We will expose it through the api/v1/vm.info endpoint. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	b70344158b	vmm: Handle the missing VM error When trying to boot or shut a VM down, return an error if the VM was not previously created. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	7e0cb078ed	vmm: Only build a new VM when booting it In order to support further use cases where a VM configuration could be modified through the HTTP API, we only store the passed VM config when being asked to create a VM. The actual creation will happen when booting a new config for the first time. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	c505cfae2b	vmm: Implement the VM HTTP endpoint handlers Implement the vm.create, vm.boot, vm.shutdown and vm.reboot HTTP endpoint handlers. Fixes: #244 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	8a5e47f989	vmm: Implement the shutdown and reboot API We factorize some of the code for both the API helpers and the VMM thread. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	46cde1a38e	vmm: Rename the VM start and stop operations to boot and shutdown To match the OpenAPI description. And also to map the real life terminology. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	ce0b475ef7	vmm: Move the VM creation and startup helpers to the api module They're API wrappers, not VMM ones. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	f674019ea1	vmm: {De}serialize VmConfig We use the serde crate to serialize and deserialize the VmVConfig structure. This structure will be passed from the HTTP API caller as a JSON payload and we need to deserialize it into a VmConfig. For a convenient use of the HTTP API, we also provide Default traits implementations for some of the VmConfig fields (vCPUs, memory, etc...). Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	f2de4d0315	vmm: config: Make the cmdline config serializable The linux_loader crate Cmdline struct is not serializable. Instead of forcing the upstream create to carry a serde dependency, we simply use a String for the passed command line and build the actual CmdLine when we need it (in vm::new()). Also, the cmdline offset is not a configuration knob, so we remove it. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	6a722e5c0b	vmm: config: Make VhostUser configs serializable They point to a vm_virtio structure (VhostUserConfig) and in order to make the whole config serializable (through the serde crate for example), we'd have to add a serde dependency to the vm_virtio crate. Instead we use a local, serializable structure and convert it to VhostUserConfig from the DeviceManager code. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	aa31748781	vmm: Start the HTTP server thread Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	b14fd37db9	vmm: Make --kernel optional The kernel path was the only mandatory command line option. With the addition of the --api-socket option, we can run without a kernel path and get it later through the API. Since we can end up with VM configurations that are no longer valid by default, we need to provide a validation check for it. For now, if the kernel path is not defined, the VM configuration is invalid. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	2371325f9c	vmm: api: Add HTTP server The Cloud Hyper HTTP server runs a synchronous, multi-threaded loop that receives HTTP requests and tries to call the corresponding endpoint handlers for the requests URIs. An endpoint handler will parse the HTTP request and potentially translate it into and IPC request. The handler holds an notifier and an mspc Sender for respectively notifying and sending the IPC payload to the VMM API server. The handler then waits for an API server response and translate it back into an HTTP response. The HTTP server is responsible for sending the reponse back to the caller. The HTTP server uses a static routes hash table that maps URIs to endpoint handlers. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Samuel Ortiz	8916dad2da	vmm: api: Add cloud-hypervisor OpenAPI documentation The cloud-hypervisor API uses HTTP as a transport and is accessible through a local UNIX socket. The API root path is /api/v1 and is a collection of RPC-style methods. All methods are static, unlike typical REST APIs. Variable (e.g. device IDs) are passed through the request body. Fixes: #244 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-10-04 09:36:33 +02:00
Rob Bradford	8ea4145f98	devices, vmm: Add legacy CMOS device Based off of crosvm revision b5237bbcf074eb30cf368a138c0835081e747d71 add a CMOS device. This environments that can't use KVM clock to get the current time (e.g. Windows and EFI.) Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-10-03 14:57:49 +01:00
Rob Bradford	833a3d456c	pci, vmm: Expose the PCI bus for configuration via MMIO Refactor the PCI datastructures to move the device ownership to a PciBus struct. This PciBus struct can then be used by both a PciConfigIo and PciConfigMmio in order to expose the configuration space via both IO port and also via MMIO for PCI MMCONFIG. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-30 18:00:31 +01:00
Rob Bradford	b5ee9212c1	vmm, devices: Use APIC address constant In order to avoid introducing a dependency on arch in the devices crate pass the constant in to the IOAPIC device creation. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-27 11:48:30 -07:00
Rob Bradford	162791b571	vmm, arch: Use IOAPIC constants from layout in DeviceManager Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-27 11:48:30 -07:00
Rob Bradford	a0455167d0	vmm: Use layout constant for kernel command line Remove the unnecessary field on CmdlineConfig and switch to using the common offset. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-27 11:48:30 -07:00
Rob Bradford	0e7a1fc923	arch, vmm: Start documenting major regions of RAM and reserved memory Using the existing layout module start documenting the major regions of RAM and those areas that are reserved. Some of the constants have also been renamed to be more consistent and some functions that returned constant variables have been replaced. Future commits will move more constants into this file to make it the canonical source of information about the memory layout. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-27 08:55:47 -07:00
Samuel Ortiz	8188074300	main: Start the VMM thread We now start the main VMM thread, which will be listening for VM and IPC related events. In order to start the configured VM, we no longer directly call the VM API but we use the IPC instead, to first create and then start a VM. Fixes: #303 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	e235c6de4f	vmm: Add VM creation and startup helpers Based on the newly defined Cloud Hypervisor IPC, those helpers send VmCreate and VmStart requests respectively. This will be used by the main thread to create and start a VM based on the CLI parameters. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	151f96e454	vmm: Add a VMM thread startup routine This starts the main, single VMM thread, which: 1. Creates the VMM instance 2. Starts the VMM control loop 3. Manages the VMM control loop exits for handling resets and shutdowns. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	2f1ff23066	vmm: (Re-)Introduce a VMM structure Unlike the Vmm structure we removed with commit `bdfd1a3f`, this new one is really meant to represent the VM monitoring/management object. For that, we implement a control loop that will replace the one that's currently embedded within the Vm structure itself. This will allow us to decouple the VM lifecycle management from the VM object itself, by having a constantly running VMM control loop. Besides the VM specific events (exit, reset, stdin for now), the VMM control loop also handles all the Cloud Hypervisor IPC requests. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	4671a5831f	vmm: Move the EpollContext implementation to lib The VMM thread and control loop will be the sole consumer of the EpollContext and EpollDispatch API, so let's move it to lib.rs. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	03ab6839c1	vmm: Introduce Cloud Hypervisor IPC Cloud Hypervisor IPC is a simple, mpsc based protocol for threads to send command to the furture VMM thread. This patch adds the API definition for that IPC, which will be used by both the main thread to e.g. start a new VM based on the CLI arguments and the future HTTP server to relay external requests received from a local Unix domain socket. We are moving it to its own "api" module because this is where the external API (HTTP based) will also be implemented. The VMM thread will be listening for IPC requests from an mpsc receiver, process them and send a response back through another mpsc channel. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	6710a39b5a	vmm: Pass the exit and reset fds to the vm creation method As we're going to move the control loop to the VMM thread, the exit and reset EventFds are no longer going to be owned by the VM. We pass a copy of them when creating the Vm instead. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	feb1c33084	vmm: Add a VM config getter We will need it from the VMM thread, when trying to reboot a VM. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	47167a658e	vmm: Add a VM console handling method In order to handle the VM STDIN stream from a separate VMM thread without having to export the DeviceManager, we simply add a console handling method to the Vm structure. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	ea7abc6c80	vmm: Add a VM stop method In order to transfer the control loop to a separate VMM thread, we want to shrink the VM control loop to a bare minimum. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	e6ef9ece2c	vmm: Move the tty setting to the VM start routine We want to shrink the control loop to a bare minimal. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	2e9d815701	vmm: Use a reference counted VmConfig when creating a new VM Once passed to the VM creation routine, a VmConfig structure is immutable. We can simply carry a Arc of it instead of a reference. This also allows us to remove any lifetime bound from our VM. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-26 16:21:14 +02:00
Samuel Ortiz	bdfd1a3f38	vmm: Remove the Vmm structure The Vmm structure is just a placeholder for the KVM instance. We can create it directly from the VM creation routine instead. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 10:12:04 +02:00
Samuel Ortiz	9c5135da7a	vmm: Simplify the VM start flow We can integrate the kernel loading into the VM start method. The VM start flow is then: Vm::new() -> vm.start(), which feels more natural. Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 10:12:04 +02:00
Samuel Ortiz	b79c1f7722	vmm: Derive the clone trait for VmConfig Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	acc60b0ad5	vmm: Make VsockConfig owned Convert Path to PathBuf and remove the associated lifetime. Now we can remove the VmConfig associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	3dc7aff00e	vmm: Make vhost-user configuration owned Convert Path to PathBuf, &str to String and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	5f8a62f3d0	vmm: Make DeviceConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	36137232f0	vmm: Make ConsoleConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	79a02f9171	vmm: Make PmemConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	00674cd850	vmm: Make FsConfig owned Convert Path to PathBuf, &str to String and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	5323da031c	vmm: Make RngConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	0688bec298	vmm: Make NetConfig owned Convert str to String and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	675e46355c	vmm: Make DiskConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	036890e5be	vmm: Make KernelConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Samuel Ortiz	9c5bfb8e13	vmm: Make MemoryConfig owned Convert Path to PathBuf and remove the associated lifetime. Fixes #298 Signed-off-by: Samuel Ortiz <sameo@linux.intel.com>	2019-09-24 08:39:39 +01:00
Yang Zhong	4164853ec6	vmm: add vhost-user-blk support Update vm configuration and device initial process to add vhost-user-blk support. Signed-off-by: Yang Zhong <yang.zhong@intel.com>	2019-09-20 15:56:51 +02:00
Yang Zhong	c7559bb7a4	config: make error definition common Since vhost-user-blk use same error definition with vhost-user-net, those errors need define to common usage. Signed-off-by: Yang Zhong <yang.zhong@intel.com>	2019-09-20 15:56:51 +02:00
Rob Bradford	5b3ca78dac	vmm: Use the full host physical address range Probe for the size of the host physical address range and use that to establish the address range for the VM. This removes the limitation on the size of the VM RAM and gives more space for the devices. Signed-off-by: Rob Bradford <robert.bradford@intel.com>	2019-09-19 10:43:55 +01:00

... 2 3 4 5 6 ...

475 Commits