惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Google DeepMind News
Google DeepMind News
人人都是产品经理
人人都是产品经理
S
Securelist
P
Proofpoint News Feed
H
Help Net Security
S
Schneier on Security
T
Tenable Blog
C
Cisco Blogs
S
Security @ Cisco Blogs
博客园 - 司徒正美
博客园 - 叶小钗
Cisco Talos Blog
Cisco Talos Blog
Google DeepMind News
Google DeepMind News
C
Cybersecurity and Infrastructure Security Agency CISA
Google Online Security Blog
Google Online Security Blog
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
Hacker News: Ask HN
Hacker News: Ask HN
NISL@THU
NISL@THU
云风的 BLOG
云风的 BLOG
V
Vulnerabilities – Threatpost
T
The Blog of Author Tim Ferriss
aimingoo的专栏
aimingoo的专栏
W
WeLiveSecurity
www.infosecurity-magazine.com
www.infosecurity-magazine.com
Jina AI
Jina AI
腾讯CDC
WordPress大学
WordPress大学
Simon Willison's Weblog
Simon Willison's Weblog
Vercel News
Vercel News
小众软件
小众软件
N
Netflix TechBlog - Medium
有赞技术团队
有赞技术团队
AWS News Blog
AWS News Blog
雷峰网
雷峰网
Forbes - Security
Forbes - Security
The Hacker News
The Hacker News
博客园 - 聂微东
F
Full Disclosure
量子位
Scott Helme
Scott Helme
宝玉的分享
宝玉的分享
A
About on SuperTechFans
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Schneier on Security
Schneier on Security
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
K
Kaspersky official blog
AI
AI
SecWiki News
SecWiki News
Webroot Blog
Webroot Blog
Martin Fowler
Martin Fowler

Proxmox Support Forum

[SOLVED] - Github Auth for Mirrors-Kernel Repo? [Automation] Mass migration tool for MS Win11/Server Proxmox GUI hang - not response is it possible to reject or quarantine spam based on conditions I set ? The PVENode task list in PVE9 is partially obscured due to the terminal font being too large. About 100% error reporting due to pveproxy.service hooks Kubernetes overlay networking breaks when upgrading from PVE 9.1 to PVE 9.2.3 Zentraler Speicher No space left on device Combine datastore and direct file archival to tape Kernel panic VFS: Unable to mount root fs on unknown-block (0,0) sobald ein 7.x Kernel verwendet wird. How to migrate disk of a VM from one ZFS to another Windows Server 2025 fails to boot after PVE 9.2 / Linux 7.0 Kernel upgrade Cannot Install Proxmox on T610 Poweredge with H700 PERC card sdn Config. gateway not reachable How to safely change domain/FQDN? Welche Filterquote erreicht ihr? NFS Share status unknown on 2 of 5 nodes Can't connect to PVE9 consoles [solved] Can't connect to PVE9 consoles [solved] [SOLVED] - Use secondary network for PVE commands Created cluster, one node storage gone BUG: proxmox mail gateway FROM = null bypass spam filtering Moving existing PBS from VMWare workstation to PVE cluster Does eBGP SDN fabric support external peering? Bug: PDM 1.1 not recognizing valid license status Proxmox GUI hang - not response PVE crashes unexpectedly Proxmox Backup Server 4.2 released! Advice ceph-osd crashes with kernel 6.17.2-1-pve on Dell system [META] Links on Proxmox Forum Website Hardwarer oder Software RAID Joining a cluster with already created guests VM PDM missing backup jobs from PVE / Log retention Remove VM.Monitor from all users/roles, PVE 9.2 Proxmox Freezing (new instalation) 9.2.2 - Intel 12700T No Web gui and random connection reset by peer [SOLVED] - i40e module for X710 Intel NIC Dutch Proxmox Day 2026 How pools use the space Corosync initiiert Reboot trotz Verfügbarkeit der Systeme Opt-in Linux 7.0 Kernel for Proxmox VE 9 available After PVE 8to9 upgrade, unable to check guest fs freeze status Problem with MegaRAID SAS3508 controller proxmox-kernel-7.0.2-6-pve failing network service Auto sync guest time after rollback of VM snapshot with RAM/state Broadcom BCM57504 (100G) bnxt_en TX timeout and NIC reset on Proxmox 8.1.5 — while BCM57414 (25G) works fine on same host QEMU 11.0 available on pve-test and pve-no-subscription as of now 350 MPM Solventless Lamination Machine for High-Speed Flexible Packaging [SOLVED] - PVE loses network connection after kernel upgrade to proxmox-kernel-7.0.0-3-pve [SOLVED] - Remove or reset cluster configuration. Proxmox 8.4.1 Fresh Install BCM57416 10G Ethernet Adapter Not Recognized PDM 1.1.1 unable to add AD realm with anonymous search [TUTORIAL] - Developer Workstation (Proxmox-VE 9) with cinnamon (LMDE7) SDN zone shows "pending" on peer nodes after node reboot (9.2.x) Cluster not quorate - extending auth key lifetime! Proxmox not rebooting properly (SOLVED) Proxmox 9 Stuck on loading initial ramdisk With new HA-Disarm Feature is there a Documentation for NUT Setup on Clusters? Proxmox 8.3 Installation Issue on ProLiant DL380 Gen9 Cluster networking setup LXC System images unavailable [SOLVED] - Fix: NVIDIA Drivers Failing after upgrade to Proxmox 9.2.2 (Kernel 7.0.2-6-pve) / NovaCore Conflict Install NUT directly on Proxmox VE and control guests from here driver usb for windows 7 System startup error and no network: Failed to start ifupdown2-pre.service - Helper to synchronize boot up for ifupdown. PBS backup space grow up constantly Proxmox Datacenter Manager 1.1 released! IPv4 not available in newly created VM Recommended Setup for Offsite Proxmox Backups? Hetzner Storage Box & Remote PBS Challenges duplicate, please delete this passthrought an USB device "by ID" to CT PDM Installer Freezes at 66% Tried PDM for the first time (version 1.1) - had issues PDM 1.1 automated install Suche Server-Provider für Proxmox connecting sdn to edge firewall SDN, IPAM & DHCP Migrating from read-only file system Ubuntu 26.04 installation fails for unknown reason Status Unbekannt nach Cluster Join Installing Proxmox Backup Server on Mac Mini (Late 2012) kernel 7.0 performance issue with zfs pools PVE becomes unreachable via ethernet but OS is running [SOLVED] - New 9.2 install - can't find 7.0.2-6-pve , not all the time [SOLVED] - Backup and dedupe a VM with LUKS Gibt es mit PVE 2.x ggf. Änderungen bei der RAM-Nutzung, bzw. deren Anzeige bei VMs? I need help for setting up backup solution Way more NAGware, very little functionality, bugs galore Root squashing virtiofsd with --uid-map Intel ixgbe Driver Update Fail Help to fix Proxmox access issues after power cut Passkey Login (not 2FA) Roblox VM detection - can be overcome? [TUTORIAL] - ZFS-Autosnaptshot inkl. Rollback und Daten direkt recovern (Windows/Linux) How to stop PVE Kernel upgrade [SOLVED] - very long waiting to log in to lxc debian 11 ssh [TUTORIAL] - Configuring Fusion-Io (SanDisk) ioDrive, ioDrive2, ioScale and ioScale2 cards with Proxmox Increase maximum USB devices in vm.conf
Making sense of NVMe zfs and SMART errors
invalid@exam · 2026-05-29 · via Proxmox Support Forum

Hope you’re all well.

I have a question that’s been wrecking my head for months now.
I have a 3 node cluster (Dell PowerEdge R7525) with the following drive configuration:
Node 1: 10x 2TB NVMe KINGSTON SKC2500M82000G in raidz2
Node 2: 10x 2TB NVMe KINGSTON SKC2500M82000G in raidz2
Node 3: 8x 4TB Crucial MX500 in raidz2

A little while ago I got an email about a zfs scrub_finish for the pool on Node 2. The issue was with two drives.
The error mentioned: One or more devices are faulted in response to persistent errors.

The array became degraded but still accessible. I moved the VMs off of it immediately. SMART didn’t show any issues for the two problematic drives.

I then left it as I was going to remove the two drives, and rebuild the pool when I upgrade the node to PVE 9. I didn’t have any spare drives, so I planned to create a new smaller pool.

Today I did just that, and as soon as I upgraded to PVE 9, the degraded state went away, and I got an email saying that a zfs resilver finished with 0 errors. The pool changed to online and the drives were fine. I then ran: zpool upgrade as suggested.

But then I got an email saying the two drives (the ones that were marked as faulty prior to upgrading to PVE 9), had the following error:
Media and Data Integrity Errors changed from 0 to 12.

The zpool was still online, and I did a scrub that reported 0 errors. I then ran short SMART tests, and they did not report any errors.

I then upgraded Node 1 to PVE 9. After the upgrade I received an email saying one of the drives reported an error: Media and Data Integrity Errors changed from 0 to 16

The pool is still online, and the drive in question has passed a short SMART test.

Would anyone have any idea what could be happening? If the drives are bad, that’s not a problem, I can just remove them and create a new smaller pool. But I’m a bit hesitant to do that if these are just false positives. Has anyone encountered this? Could this just be some strange behaviour related to the Dell servers? I know the drives should be enterprise class, but we could not afford them at the time, and especially now.

Thank you so much!

I suspect that the upgrade changed something about the smart monitoring to reset the tracked stats, or maybe to start tracking stats that weren't tracked before. This is probably why it jumped from 0 to 12 or 0 to 16, and it's likely that you had 12 or whatever errors for a while.

The resilver probably didn't take long to fix the raid because you had moved all the data off to other nodes. Unlike hardware raid, ZFS knows when parts of the raid are free space, and can skip rebuilding those empty parts. And for that same reason the scrub doesn't tell you anything about the health of empty parts of the disks. Scrub reads the data from the disks, and verifies that it matches the hashes recorded when that data was written, but empty space isn't checked.

The short SMART test passing is a good sign, and suggests you might be able to keep using these disks for a while. I would first do the full long test on all of the disks in that first system, and if that passes that's an even better sign, since that should test even the empty parts of the drives.

SMART is a weird tech. Different manufacturers (or indeed, different models from the same manufacturer) seem to implement it differently. I have a drive in use with ~20 (I don't remember the exact number) of those "Media and Data Integrity Errors" that it has had for more than a year. So far there's been no data loss, and that number hasn't gone up any. So those 20-ish problems either the drive recovered from, or ZFS did. I keep an eye on it, but since the number isn't going up, I'm not too worried about it. I don't know exactly what event caused that. Was it a bad sector that was swapped out for a reserve one, or just read that failed but worked when retried? I don't know, but the drive passed a long self-test, and the number hasn't gone up since I started watching it, so I've stopped worrying about it until the number does go up.

It's hard to suggest you treat it as I did. I choose that because I don't have strict SLAs, and do have strict budgets. If that drive did fail and I had to restore a few VMs from backups, it would be fine as long as I got that started within an hour. So I can't say that I'd do the same if I were in your place, but my experience suggests that as long as you keep watch out for notifications of that number starting to climb again, its probably fine to keep using those drives.