惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

H
Help Net Security
F
Fortinet All Blogs
Engineering at Meta
Engineering at Meta
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
T
The Exploit Database - CXSecurity.com
H
Hackread – Cybersecurity News, Data Breaches, AI and More
I
Intezer
P
Privacy & Cybersecurity Law Blog
M
MIT News - Artificial intelligence
MyScale Blog
MyScale Blog
P
Privacy International News Feed
MongoDB | Blog
MongoDB | Blog
Project Zero
Project Zero
C
Cyber Attacks, Cyber Crime and Cyber Security
T
Tenable Blog
Security Latest
Security Latest
Stack Overflow Blog
Stack Overflow Blog
L
Lohrmann on Cybersecurity
V
Vulnerabilities – Threatpost
Microsoft Azure Blog
Microsoft Azure Blog
NISL@THU
NISL@THU
T
Threat Research - Cisco Blogs
L
LangChain Blog
Simon Willison's Weblog
Simon Willison's Weblog
WordPress大学
WordPress大学
SecWiki News
SecWiki News
博客园 - 三生石上(FineUI控件)
Forbes - Security
Forbes - Security
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
G
GRAHAM CLULEY
K
Kaspersky official blog
W
WeLiveSecurity
A
Arctic Wolf
TaoSecurity Blog
TaoSecurity Blog
Recorded Future
Recorded Future
AI
AI
T
The Blog of Author Tim Ferriss
宝玉的分享
宝玉的分享
K
KPMG report finds enterprise disconnect between AI and its ROI | CIO
Last Week in AI
Last Week in AI
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
雷峰网
雷峰网
GbyAI
GbyAI
S
SegmentFault 最新的问题
N
News and Events Feed by Topic
C
CXSECURITY Database RSS Feed - CXSecurity.com
Google Online Security Blog
Google Online Security Blog
博客园 - Franky
罗磊的独立博客

Proxmox Support Forum

[SOLVED] - Github Auth for Mirrors-Kernel Repo? [Automation] Mass migration tool for MS Win11/Server Proxmox GUI hang - not response is it possible to reject or quarantine spam based on conditions I set ? The PVENode task list in PVE9 is partially obscured due to the terminal font being too large. About 100% error reporting due to pveproxy.service hooks Kubernetes overlay networking breaks when upgrading from PVE 9.1 to PVE 9.2.3 Zentraler Speicher No space left on device Combine datastore and direct file archival to tape Kernel panic VFS: Unable to mount root fs on unknown-block (0,0) sobald ein 7.x Kernel verwendet wird. How to migrate disk of a VM from one ZFS to another Windows Server 2025 fails to boot after PVE 9.2 / Linux 7.0 Kernel upgrade Cannot Install Proxmox on T610 Poweredge with H700 PERC card sdn Config. gateway not reachable How to safely change domain/FQDN? Welche Filterquote erreicht ihr? NFS Share status unknown on 2 of 5 nodes Can't connect to PVE9 consoles [solved] Can't connect to PVE9 consoles [solved] [SOLVED] - Use secondary network for PVE commands Created cluster, one node storage gone BUG: proxmox mail gateway FROM = null bypass spam filtering Moving existing PBS from VMWare workstation to PVE cluster Does eBGP SDN fabric support external peering? Bug: PDM 1.1 not recognizing valid license status Proxmox GUI hang - not response PVE crashes unexpectedly Proxmox Backup Server 4.2 released! Advice ceph-osd crashes with kernel 6.17.2-1-pve on Dell system [META] Links on Proxmox Forum Website Hardwarer oder Software RAID Joining a cluster with already created guests VM PDM missing backup jobs from PVE / Log retention Remove VM.Monitor from all users/roles, PVE 9.2 Proxmox Freezing (new instalation) 9.2.2 - Intel 12700T No Web gui and random connection reset by peer [SOLVED] - i40e module for X710 Intel NIC Dutch Proxmox Day 2026 How pools use the space Corosync initiiert Reboot trotz Verfügbarkeit der Systeme Opt-in Linux 7.0 Kernel for Proxmox VE 9 available After PVE 8to9 upgrade, unable to check guest fs freeze status Problem with MegaRAID SAS3508 controller proxmox-kernel-7.0.2-6-pve failing network service Auto sync guest time after rollback of VM snapshot with RAM/state Broadcom BCM57504 (100G) bnxt_en TX timeout and NIC reset on Proxmox 8.1.5 — while BCM57414 (25G) works fine on same host QEMU 11.0 available on pve-test and pve-no-subscription as of now 350 MPM Solventless Lamination Machine for High-Speed Flexible Packaging Making sense of NVMe zfs and SMART errors [SOLVED] - PVE loses network connection after kernel upgrade to proxmox-kernel-7.0.0-3-pve [SOLVED] - Remove or reset cluster configuration. Proxmox 8.4.1 Fresh Install BCM57416 10G Ethernet Adapter Not Recognized PDM 1.1.1 unable to add AD realm with anonymous search [TUTORIAL] - Developer Workstation (Proxmox-VE 9) with cinnamon (LMDE7) SDN zone shows "pending" on peer nodes after node reboot (9.2.x) Cluster not quorate - extending auth key lifetime! Proxmox not rebooting properly (SOLVED) Proxmox 9 Stuck on loading initial ramdisk With new HA-Disarm Feature is there a Documentation for NUT Setup on Clusters? Proxmox 8.3 Installation Issue on ProLiant DL380 Gen9 Cluster networking setup LXC System images unavailable [SOLVED] - Fix: NVIDIA Drivers Failing after upgrade to Proxmox 9.2.2 (Kernel 7.0.2-6-pve) / NovaCore Conflict Install NUT directly on Proxmox VE and control guests from here driver usb for windows 7 System startup error and no network: Failed to start ifupdown2-pre.service - Helper to synchronize boot up for ifupdown. PBS backup space grow up constantly Proxmox Datacenter Manager 1.1 released! IPv4 not available in newly created VM Recommended Setup for Offsite Proxmox Backups? Hetzner Storage Box & Remote PBS Challenges duplicate, please delete this passthrought an USB device "by ID" to CT PDM Installer Freezes at 66% Tried PDM for the first time (version 1.1) - had issues PDM 1.1 automated install Suche Server-Provider für Proxmox connecting sdn to edge firewall SDN, IPAM & DHCP Migrating from read-only file system Ubuntu 26.04 installation fails for unknown reason Status Unbekannt nach Cluster Join Installing Proxmox Backup Server on Mac Mini (Late 2012) kernel 7.0 performance issue with zfs pools PVE becomes unreachable via ethernet but OS is running [SOLVED] - New 9.2 install - can't find 7.0.2-6-pve , not all the time [SOLVED] - Backup and dedupe a VM with LUKS Gibt es mit PVE 2.x ggf. Änderungen bei der RAM-Nutzung, bzw. deren Anzeige bei VMs? I need help for setting up backup solution Way more NAGware, very little functionality, bugs galore Root squashing virtiofsd with --uid-map Intel ixgbe Driver Update Fail Passkey Login (not 2FA) Roblox VM detection - can be overcome? [TUTORIAL] - ZFS-Autosnaptshot inkl. Rollback und Daten direkt recovern (Windows/Linux) How to stop PVE Kernel upgrade [SOLVED] - very long waiting to log in to lxc debian 11 ssh [TUTORIAL] - Configuring Fusion-Io (SanDisk) ioDrive, ioDrive2, ioScale and ioScale2 cards with Proxmox Increase maximum USB devices in vm.conf
Virtual machine freezes with IO error
invalid@exam · 2026-04-16 · via Proxmox Support Forum

Hello everyone,

I'm having problems with only one virtual machine. After several days, it displays the error: Status: io-error. I've checked the disk and there's no problem there. It only happens with this machine. After shutting it down and turning it back on, it continues to function normally.

1.png

I have no idea what to check.

Thanks!

Are there any errors in `journalctl` or `dmesg`?
You can gather the VM configuration/status and the storage status with the following commands:

Code:

qm status $VMID --verbose
qm config $VMID
qm showcmd $VMID --pretty
pvesm status

Hello d.oshi,

No errors were observed with journalctl or dmesg.
I also don't see any errors with the commands you sent me, unless I don't know how to interpret them; I've attached a file with the output of each command.

Thanks!

  • info.txt

    8.4 KB · Views: 17

cwt

Renowned Member

Your VM is on the local-lvm which is not really suitable for VM storage.

If the error re-occurs look for storage related messages:

Code:

dmesg -T | egrep -i "error|fail|reset|nvme|sd"

The log indicates slow writes:

Code:

wr_operations: 802447
wr_total_time_ns: 1731275382833

That‘s around 2ms per write. And this:

Code:

account_failed: 1
account_invalid: 1

comes directly from the QEMU block-layer and means that there was an IO error.

How does did you setup your storage for local-lvm? NVME? SSD?

There does not seem to be a local-lvm. Also why would local-lvm not be suitable? For local I agree.

Hello cwt,

Attached is the result of dmesg grep, at this moment it froze again, I turned it off and on again.

Thanks!

  • info.txt

    5.1 KB · Views: 7

cwt

Renowned Member

Jepp, typo. Storage is local, not local-lvm.

@uzisuicida: your VM has

Code:

cache.direct=true
no-flush=false

But dmesg shows:

Code:

sd 0:0:0:0: [sda] Write cache: enabled, read cache: enabled, doesn't support DPO or FUA
sd 1:0:0:0: [sdb] Write cache: enabled, read cache: enabled, doesn't support DPO or FUA

FUA is Force Unit Access = write blocks directly on the storage.

qcow2 (used by your VM) relies heavily on metadata updates and requires reliable flush operations to keep the filesystem consistent.
If the underlying storage does not support FUA, flush requests may be acknowledged before data is actually written to disk.
This creates a mismatch where the VM assumes data is safely stored, while it may still reside in volatile cache.
Under load or failure conditions, this can lead to data corruption or I/O errors, potentially crashing the VM.

Hello cwt,
Do I need to disable the disk cache? Or what should I do in this case? Thank you so much for your help.

This creates a mismatch where the VM assumes data is safely stored, while it may still reside in volatile cache. Under load or failure conditions, this can lead to data corruption or I/O errors, potentially crashing the VM.

How would it lead to data corruption or I/O errors under load? When data is requested to be read again, if it's in the cache (and hasn't YET been written to disk), then the cached data would be given back to the requestor.

Where you run into issues is if there's a crash and cached data can't be flushed to disk

Jepp, typo. Storage is local, not local-lvm.

@uzisuicida: your VM has

Code:

cache.direct=true
no-flush=false

But dmesg shows:

Code:

sd 0:0:0:0: [sda] Write cache: enabled, read cache: enabled, doesn't support DPO or FUA
sd 1:0:0:0: [sdb] Write cache: enabled, read cache: enabled, doesn't support DPO or FUA

FUA is Force Unit Access = write blocks directly on the storage.

qcow2 (used by your VM) relies heavily on metadata updates and requires reliable flush operations to keep the filesystem consistent.
If the underlying storage does not support FUA, flush requests may be acknowledged before data is actually written to disk.
This creates a mismatch where the VM assumes data is safely stored, while it may still reside in volatile cache.
Under load or failure conditions, this can lead to data corruption or I/O errors, potentially crashing the VM.

Hello,

I have other servers with Proxmox, and I ran the same command, it shows the same thing, but I don't have this problem of a virtual machine freezing and having to turn it off and on.

Thanks.

fiona

Proxmox Staff Member

Yes, there was an IO error, but

Code:

account_failed: 1
account_invalid: 1

comes directly from the QEMU block-layer and means that there was an IO error.

this is not what these two values mean:

  • account_invalid (<span>boolean</span>) – Whether invalid operations are included in thelast access statistics (Since 2.5)
  • account_failed (<span>boolean</span>) – Whether failed operations are included in thelatency and last access statistics (Since 2.5)

https://qemu.readthedocs.io/en/mast...f.html#object-QMP-block-core.BlockDeviceStats

cwt

Renowned Member

Agree.

The IO error is not indicated by account_failed / account_invalid (those are just accounting flags), but by failed_*_operations together with qmpstatus: io-error

Hello,

I have other servers with Proxmox, and I ran the same command, it shows the same thing, but I don't have this problem of a virtual machine freezing and having to turn it off and on.

Thanks.

Hi uzisuicida,

I am having the same problems with "io-errors" with two of my VMs - no logs from the kernel or anything else.

DId you ever find a solution for your issue?

fiona

Proxmox Staff Member

Hi @nils-the-oldone1980,
is there enough free space? Can you share the excerpt from the system logs/journal from around the time the issue happened as well as the output of pveversion -v and the VM configuration qm config ID with the numerical ID.

Hi @nils-the-oldone1980,
is there enough free space? Can you share the excerpt from the system logs/journal from around the time the issue happened as well as the output of pveversion -v and the VM configuration qm config ID with the numerical ID.

Hi fiona,

please see attached output. I've already tried all available "Async IO" mechanisms, disabling "IO thread" at all and/or using "VirtIO SCSI" instead of "VirtIO SCSI single". Even tried without KSM. No success - still locks up with "io error" eventually - cannot force it; it just happens.

edit: VM 999 locked up at around 05:08 today - cannot say for sure because I am using Nagios for out-of-work-time monitoring. Nagios has a delay of two minutes before reporting not-reachable guest agents.

Question: whenever one of the VMs should lock up again, which commands/outputs can I execute/provide to investigate further?

  • output.txt

    38 KB · Views: 2

Last edited:

fiona

Proxmox Staff Member

Please share the full journal, not just warnings. Are you using CommVault or Naviko as a backup solution? In the log it can be seen that something creates NBD block devices and maybe the issue is related to that.

The output of

Code:

echo '{"execute": "qmp_capabilities"}{"execute": "query-block"}' | socat - /run/qemu-server/123.qmp | jq

would also be interesting. Replace 123 with your VM ID. You might need to install socat and jq first.

Please share the full journal, not just warnings. Are you using CommVault or Naviko as a backup solution? In the log it can be seen that something creates NBD block devices and maybe the issue is related to that. [...]

Yes, we are using NAKIVO. But your mentioning of looking at the full journal and not just warnings gave me the solution, i think.

Yesterday, the VM 999 locked up again at around 09:49 CEST. Here are the relevant journal entries:

Code:

May 27 09:49:52 proxkon01 zed[2586197]: eid=43733 class=dio_verify_wr pool='datastore' size=131072 offset=2737142026240 priority=1 err=5 flags=0x200080 bookmark=54:6784:0:984520
May 27 09:49:59 proxkon01 pvedaemon[2574504]: VM 999 qga command failed - VM 999 qga command 'guest-ping' failed - got timeout

Looking up the failure "dio_verify_wr" in combination with "err=5" led me to another thread:

Hi everyone,

We’ve been deploying several new Proxmox 9 nodes using ZFS as the primary storage, and we’re encountering issues where virtual machines become I/O locked.

When it happens, the VMs are paused with an I/O error. We’re aware this can occur when a host runs out of disk space, but in our case there is plenty of free storage available.
We’ve seen this behavior across multiple hosts, different clusters, and different hardware platforms.

Furthermore, we’ve been running ZFS on Proxmox 8 without any issues, but since these problems started with our Proxmox 9 installations, we’re...

I have disabled Direct I/O on our ZFS storages now using "zfs set direct=disabled datastore" and am very sure that this was the cause. Will let you know the outcome in a day or two.

Hello again,

since deactivation of Direct I/O there were no more io errors of the VMs, and all of them are running absolutely fine.