qemu-server

mirror of https://git.proxmox.com/git/qemu-server synced 2025-10-27 22:32:20 +00:00

Author	SHA1	Message	Date
Fabian Ebner	c2c96d7378	fix checks for transfering replication state/switching job target In some cases $self->{replicated_volumes} will be auto-vivified to {} by checks like next if $self->{replicated_volumes}->{$volid} and then {} would evaluate to true in a boolean context. Now the replication information is retrieved once in prepare, and used to decide whether to make the calls or not. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-11-09 10:08:22 +01:00
Fabian Ebner	68980d6626	Repeat check for replication target in locked section No need to warn twice, so the warning from the outside check was removed. Suggested-by: Fabian Grünbichler <f.gruenbichler@proxmox.com> Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-11-09 10:08:22 +01:00
Thomas Lamprecht	e5d611c382	fix various conditionally declared vars Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-10-16 16:52:11 +02:00
Fabian Ebner	1264d6c511	Use correct option for storage_migrate Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-08-04 13:57:09 +02:00
Stefan Reiter	b53ba8d0f1	fixup: use parse_property_string instead of parse_cpu_conf_basic The latter was removed and replaced with a validator. Signed-off-by: Stefan Reiter <s.reiter@proxmox.com>	2020-07-09 14:45:21 +02:00
Fabian Ebner	9b29cbd0ed	update_disksize: make interface leaner Pass new size directly, so the function doesn't need to know about how some hash is organized. And return a message directly, instead of both size-strings. Also dropped the wantarray, because both existing callers use the message anyways. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-07-01 09:18:13 +02:00
Fabian Ebner	1c2174833b	sync_disks: fix check Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-07-01 09:13:06 +02:00
Fabian Grünbichler	ae194a5c5e	migrate: cleanup forwarding code fixing the following two issues: - the legacy code path was never converted to the new fork_tunnel signature (which probably means that nothing triggers it in practice anymore?) - the NBD Unix socket got forwarded multiple times if more than one disk was migrated via NBD (this is harmless, but wrong) for the second issue I opted to keep the code compatible with the possibility that Qemu starts supporting multiple NBD servers in the future (and the target node could thus return multiple UNIX socket paths). currently we can only start one NBD server on one socket, and each drive-mirror simply starts a new connection over that single socket. I took the liberty of renaming the variables/keys since I found 'tunnel_addr' and 'sock_addr' rather confusing. Reviewed-By: Mira Limbeck <m.limbeck@proxmox.com> Tested-By: Mira Limbeck <m.limbeck@proxmox.com> Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-05-06 16:16:50 +02:00
Dominik Csapak	cd37203880	migrate: skip rescan for efidisk and shared volumes we really only want to rescan the disk size of the disks we actually need, and that are only the local disks (for which we have to allocate the correct size on the target) also we want to always skip the efidisk, since we get the wanted size after the loop, and this produced a confusing log line (for details why we do not want the 'real' size, see commit `818ce80ec1`) Signed-off-by: Dominik Csapak <d.csapak@proxmox.com>	2020-05-04 17:35:12 +02:00
Fabian Grünbichler	6f4b11e9db	migrate: don't accidentally take NBD code paths by avoiding auto-vivification of $self->{online_local_volumes} via iteration. most code paths don't care whether it's undef or a reference to an empty list, but this caused the (already) fixed bug of calling nbd_stop without having started an NBD server in the first place. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-05-04 17:34:58 +02:00
Thomas Lamprecht	3e802221e1	migrate: only stop NBD if we got a NBD url from the target Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-04-29 16:22:33 +02:00
Fabian Ebner	ae180b8f08	Include vmstate and unused volumes in foreach_volid and refactor the test_volid closure. Like this get_replicatable_volumes doesn't need a separate loop for unused volumes anymore. For get_vm_volumes, which is used for activation/deactivation of volumes at migration and deactivation in vm_stop_cleanup, includes those volumes now. For migration it's an improvement, because those volumes might need to be migrated and for vm_stop_cleanup it shouldn't hurt. The last user of foreach_volid is check_vm_disks_local used by migrate_vm_precondition, where information about the additional volumes doesn't hurt either. Note that replicate is (still) set by default, so the behavior for get_replicatable_volumes for unused volumes should not change. Hibernation vmstate files are now also included and recognized as 'is_vmstate'. The 'size' attribute will not be overwritten by subsequent iterations for the same volid anymore (a volid may appear both in the config and in snapshots), so the size from the current config is now preferred. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-29 12:14:40 +02:00
Fabian Ebner	b24f07d406	Fix test_volid call for vmstate and fix check for snapshots on migration by excluding vmstate. It is referenced by snapshots, but is not a volume containing a snapshot. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-29 12:14:40 +02:00
Fabian Grünbichler	90ff65b63a	migrate: simplify replicated_volume loop (no change compared to previous iteration except for readability) Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-04-20 11:24:23 +02:00
Fabian Ebner	cee620e671	Fix live migration with replicated unused volumes by counting only local volumes that will be live-migrated via qemu_drive_mirror, i.e. those listed in $self->{online_local_volumes}. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-20 11:12:56 +02:00
Thomas Lamprecht	38311a1d17	migrate: workaround issues with format switch on storage live migration Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-04-17 15:27:38 +02:00
Fabian Ebner	ea5b400812	sync_disks: log output of storage_migrate Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-08 22:11:54 +02:00
Fabian Ebner	49a5a0d84b	sync_disks: be more verbose if storage_migrate fails If storage_migrate dies, the error message might not include the volume ID or the target storage ID, but those might be good to know. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-08 22:11:54 +02:00
Fabian Ebner	cc1a3820db	sync_disks: use allow_rename to avoid collisions on the target storage This makes it possible to migrate a VM with volumes store1:vm-123-disk-0 store2:vm-123-disk-0 to some targetstorage. Also prevents migration failure when there is an orphaned disk with the same volid on the target. To avoid confusion, the name should not change for 'vmstate'-volumes. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-08 22:11:54 +02:00
Fabian Ebner	97ece9ddce	Update volume IDs in one go Use 'update_volume_ids' for the live-migrated disks as well. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-08 22:11:54 +02:00
Fabian Ebner	37666e4caa	Take note of changes to the volume IDs when migrating and update the config Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-08 22:11:54 +02:00
Fabian Ebner	1f726e0a85	Use new storage_migrate interface Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-08 22:11:54 +02:00
Fabian Ebner	912792e245	Switch to using foreach_volume instead of foreach_drive It was necessary to move foreach_volid back to QemuServer.pm In VZDump/QemuServer.pm and QemuMigrate.pm the dependency on QemuConfig.pm was already there, just the explicit "use" was missing. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-04-08 22:11:54 +02:00
Stefan Reiter	58c64ad5d9	Include "-cpu" parameter with live-migration This is required to support custom CPU models, since the "cpu-models.conf" file is not versioned, and can be changed while a VM using a custom model is running. Changing the file in such a state can lead to a different "-cpu" argument on the receiving side. This patch fixes this by passing the entire "-cpu" option (extracted from /proc/.../cmdline) as a "qm start" parameter. Note that this is only done if the VM to migrate is using a custom model (which we can check just fine, since the <vmid>.conf is versioned with pending changes), thus not breaking any live-migration directionality. Signed-off-by: Stefan Reiter <s.reiter@proxmox.com>	2020-04-07 17:27:58 +02:00
Fabian Grünbichler	bf8fc5a307	migrate: allow arbitrary source->target storage maps the syntax is backwards compatible, providing a single storage ID or '1' works like before. the new helper ensures consistent behaviour at all call sites. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-04-02 17:47:14 +02:00
Stefan Reiter	c05f1b33ea	migration: fix downtime limit auto-increase `485449e37` ("qmp: use migrate-set-parameters in favor of deprecated options") changed the initial "migrate_set_downtime" QMP call to the more recent "migrate-set-parameters", but forgot to do so for the auto-increase code further below. Since the units of the two calls don't match, this would have caused the auto-increase to increase the limit to absurd levels as soon as it kicked in (ms treated as s). Update the second call to the new version as well, and while at it remove the unnecessary "defined()" check for $migrate_downtime, which is always initialized from the defaults anyway. Signed-off-by: Stefan Reiter <s.reiter@proxmox.com>	2020-04-02 16:48:51 +02:00
Fabian Grünbichler	6a039d06e9	migrate: improve cleanup_remotedisks to also handle cases where disk allocation failed in the remote vm_start, and we only have a bitmap but no target drive information. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-04-01 17:41:07 +02:00
Dominik Csapak	818ce80ec1	fix efidisks on storages with minimum sizes bigger than OVMF_VARS.fd on storages where the minimum size of images is bigger than the real OVMF_VARS.fd file, they get padded to their minimum size when using such an image, qemu maps it fully to the vm, but the efi does not find the vars region and creates a file on the first efi partition it finds this breaks some settings in the ovmf, such as resolution to fix this, we have to specify the size for the pflash, so that qemu only maps the first n bytes in the vm (this only works for raw files, not for qcow2) we also have to use the correct size when converting between storages in 'clone_disk' (used for move disk and cloning vms) and when live migrating to different storages when we now expect that the source image is always correctly used/created (e.g. raw with size=x in pflash argument) then we always create the target correctly when encountering users which have a non-valid image (e.g. a efidisk moved from zfs to qcow2 before this patch), we have to tell them to recreate the efidisk and the settings on it we have to version_guard it to 4.1+pve2 (since we haven't bumped yet since the change to pve2) also add 2 tests, one for the old version and one for the new Signed-off-by: Dominik Csapak <d.csapak@proxmox.com> Tested-by: Stefan Reiter <s.reiter@proxmox.com> Reviewed-by: Stefan Reiter <s.reiter@proxmox.com> [ Thomas: rebased to master ] Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-03-30 09:41:55 +02:00
Fabian Ebner	5c50a84f23	migration with targetstorage: check if target storage supports images This makes sure that live migration also respects content types. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-03-27 14:32:42 +01:00
Thomas Lamprecht	2cd808d331	migrate sync disks: split long line Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-03-27 10:17:54 +01:00
Thomas Lamprecht	b10afa311d	migrate sync_disks: use own variable for often referenced storage config also fix two places where we used $self->{vmid} even if $vmid was in scope (and the same). Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-03-27 10:13:10 +01:00
Fabian Grünbichler	9b3f5a5c99	migrate: cleanup disk/bitmaps if 'qm start' failed since bitmaps are set early on, and 'qm start' potentially has allocated the disks but still failed. we can only clean up what we know about anyway, so the disk part is still only best effort. also use replicated_volumes instead of bitmap existence to check for replicated volumes, since 'qm start' on an old node that does not understand replicated volumes might have allocated a new volume that we DO want to clean up, and not skip. also cleanup disks after stopping target VM, otherwise we might end up in a situation where the target VM is still running and using the disks, thus blocking the disk cleanup. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-03-27 07:54:44 +01:00
Fabian Grünbichler	7f5fb49a7c	migrate: fix auto-vivification in cleanup_bitmaps this does not currently trigger since nothing uses $self->{target_drive} afterwards. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-03-27 07:54:44 +01:00
Fabian Grünbichler	88126be3f7	migrate: fix replication false-positives by only checking for replicatable volumes when a replication job is defined, and passing only actually replicated volumes to the target node via STDIN, and back via STDOUT. otherwise this can pick up theoretically replicatable, but not actually replicated volumes and treat them wrong. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-03-27 07:54:44 +01:00
Fabian Ebner	47250f03ef	Fix calls to get_replicateable_volumes There is a need to set $noerr, because otherwise migration for a VM with a non-replicatable volume fails with: missing replicate feature on volume 'myfs:107/vm-107-disk-2.raw' Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-03-25 14:53:17 +01:00
Thomas Lamprecht	6d7450cbec	qemu migrate: sort and split module usage Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-03-25 10:05:58 +01:00
Thomas Lamprecht	28e6e180bc	add basic version check for live-migration with replicated disks as we need at least pve-qemu in 4.2 for this to work, the target side is implicitly checked with "to old version" check for migrate or the mirror will fail anyway. Just use the simple "qemu binary version check", as we could stil live migrate an older snapshot with older machine versions if both sides have a recent enough qemu. Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-03-25 10:02:36 +01:00
Fabian Grünbichler	9b6efe436d	migrate: add live-migration of replicated disks with incremental drive-mirror and dirty-bitmap tracking. 1.) get replicated disks that are currently referenced by running VM 2.) add a block-dirty-bitmap to each of them 3.) replicate ALL replicated disks 4.) pass bitmaps from 2) to drive-mirror for disks from 1) 5.) skip replicated disks when cleaning up volumes on either source or target added error handling is just removing the bitmaps if an error occurs at any point after 2, except when the handover to the target node has already happened, since the bitmaps are cleaned up together with the source VM in that case. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com> Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com> Tested-by: Stefan Reiter <s.reiter@proxmox.com>	2020-03-24 12:22:32 +01:00
Fabian Grünbichler	b9f44d2773	migrate: add replication info to disk overview to make migration logs a bit easier to grasp with a quick glance. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com> Tested-by: Stefan Reiter <s.reiter@proxmox.com>	2020-03-24 11:54:32 +01:00
Thomas Lamprecht	1e0074c437	migrate phase3: add to comment why a blockjob cancel is OK here Clarify why a cancel is actually not really canceling here, because we're already finished with storage migration and the block jobs are all in ready state and we (source) are going to stop soon to hand over to target. > Note that if you issue 'block-job-cancel' after 'drive-mirror' has > indicated (via the event BLOCK_JOB_READY) that the source and > destination are synchronized, then the event triggered by this > command changes to BLOCK_JOB_COMPLETED, to indicate that the > mirroring has ended and the destination now has a point-in-time > copy tied to the time of the cancellation -- qapi/block-core.json (QEMU 4.2) Signed-off-by: Thomas Lamprecht <t.lamprecht@proxmox.com>	2020-03-20 11:08:23 +01:00
Mira Limbeck	ff09c795ed	revert spice_ticket prefix change in `7827de4` The change to the prefixed version broke migration from new to old qemu-server version. This reverts the change and adds a TODO comment for 7.0 to change it to the prefixed version then. Signed-off-by: Mira Limbeck <m.limbeck@proxmox.com>	2020-03-20 10:37:33 +01:00
Fabian Grünbichler	db1f8b39e1	drive_mirror: rename variables and values and add some more details to comments. Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-03-18 08:21:29 +01:00
Mira Limbeck	7827de41a2	add unix socket support for NBD storage migration The reuse of the tunnel, which we're opening to communicate with the target node and to forward the unix socket for the state migration, for the NBD unix socket requires adding support for an array of sockets to forward, not just a single one. We also have to change the $sock_addr variable to an array for the cleanup of the socket file as SSH does not remove the file. To communicate to the target node the support of unix sockets for NBD storage migration, we're specifying an nbd_protocol_version which is set to 1. This version is then passed to the target node via STDIN. Because we don't want to be dependent on the order of arguments being passed via STDIN, we also prefix the spice ticket with 'spice_ticket: '. The target side handles both the spice ticket and the nbd protocol version with a fallback for old source nodes passing the spice ticket without a prefix. All arguments are line based and require a newline in between. When the NBD server on the target node is started with a unix socket, we get a different line containing all the information required to start the drive-mirror. This contains the unix socket path used on the target node which we require for forwarding and cleanup. Signed-off-by: Mira Limbeck <m.limbeck@proxmox.com>	2020-03-18 08:03:44 +01:00
Mira Limbeck	e02fb12620	add qemu_drive_mirror_monitor completion modes With Qemu 4.2 we encountered a problem with unix sockets and SSH socket forwarding for drive-mirror. It seems the socket gets reopened again and again after it closes for some reason. This can be worked around by specifying 'block-job-cancel' instead of 'block-job-complete' when we're not interested in swapping the disks again from NBD to their original protocol. This is always the case when we use drive-mirror for live migrating a VM. qemu_drive_mirror is used for migration and for clone_disk. All in all we have 3 cases to handle. Either the 'skip' case which skips the completion of the job. The 'wait' case which was the default before and still is when $completion is undefined. And the new 'wait_noswap' case which is used for the live migration. If 'wait_noswap' is specified, we issue a 'block-job-cancel' once the block job is in 'ready' state. This completes the block job without swapping the disks. clone_disk always uses 'block-job-cancel' via the qemu_blockjobs_cancel sub. Signed-off-by: Mira Limbeck <m.limbeck@proxmox.com>	2020-03-18 08:03:44 +01:00
Fabian Ebner	0ad295f9fb	Consistently use format determined in 'PVE::Storage::foreach_volid' Signed-off-by: Fabian Ebner <f.ebner@proxmox.com> LGTM-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-03-09 19:36:58 +01:00
Fabian Ebner	5eca0c3643	sync_disks: Always set 'snapshots' for qcow2 and vmdk volumes This fixes an issue when migrating a VM with an unused volume with format qcow2 or vmdk. Since 'snapshots' wasn't set, storage_migrate wanted to export/import with format raw+size instead. Therefore it used (instead of just 'dd') 'qemu-img convert', which fails when its output leaves through a pipe. Upon importing, a second error is present, because the format from the volume ID doesn't match the format of the stream and there is no conversion yet. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com> LGTM-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-03-09 19:36:45 +01:00
Fabian Ebner	e0fd2b2f84	Create Drive.pm and move drive-related code there The initialization for the drive keys in $confdesc is changed to be a single for-loop iterating over the keys of $drivedesc_hash and the initialization of the unusedN keys is move to directly below it. To avoid the need to change all the call sites, functions with more than a few callers are exported from the submodule and imported into QemuServer.pm. For callers of the now imported functions within QemuServer.pm, the prefix PVE::QemuServer is dropped, because it is unnecessary and now even confusing. Signed-off-by: Fabian Ebner <f.ebner@proxmox.com>	2020-03-07 18:23:57 +01:00
Stefan Reiter	29eb909ee0	fix #2611 : use correct operation in get_bandwidth_limit Signed-off-by: Stefan Reiter <s.reiter@proxmox.com>	2020-03-03 11:47:13 +01:00
Stefan Reiter	485449e37b	qmp: use migrate-set-parameters in favor of deprecated options migrate_set_downtime, migrate_set_speed and migrate-set-cachesize have all been deprecated since 2.8 or 2.11 [0]. They still work, but no reason not to use the correct version. Note that the downtime-limit parameter switched from seconds to milliseconds, so convert to that. Slightly improve log output with units while at it. [0] https://qemu.weilnetz.de/doc/qemu-doc.html#Deprecated-features Signed-off-by: Stefan Reiter <s.reiter@proxmox.com>	2020-02-06 13:50:33 +01:00
Fabian Grünbichler	683ab65491	migrate: re-order lines to improve readability Signed-off-by: Fabian Grünbichler <f.gruenbichler@proxmox.com>	2020-02-05 09:43:09 +01:00

1 2 3 4 5

204 Commits