@j.theisen. I have just read that LVM-Thin is NOT supported/recommended for shared iSCSI in a cluster (especially with HA). Metadata is not cluster-aware and this kind of corruption is a known risk. I am running LVM-Thin in this configuration...
Hi @southwalesowl
thanks for posting in the forum!
Can you please share your /etc/pve/storage.cfg and the journal for the time period of the simulated outage for both hosts?
Yours sincerely
Jonas
I have 2 Proxmox hosts configured and an additional QDevice so that I can configure HA. Each host has 4 physical NICS. I have 2 of them configured with static I.Ps, on different subnets, for iSCSI and I have multipathing configured. I have HA...
Hi @etfz ,
As far as I know there is no network or storage monitor integrated into default PVE cluster distribution. You will need to create your own.
Cheers
Blockbridge : Ultra low latency all-NVME shared storage for Proxmox -...
Hi,
I have a cluster with separate networks for VM upstream traffic, Ceph and corosync. When a host loses corosync connectivity, fencing kicks in, as expected, but can I do the same for the VM and Ceph networks? A host without upstream or Ceph...
AFAIK that doesn't have any m.2 native controller/connection. You must be using some PCI-riser card as a controller - or some other shenanigans through something else. (A waste of that m.2 Gen 4 drive!). But my point being - I'm not sure you are...
hi,
i added 2x micron 7400 (m.2) as zfs special device in kernel 6.8.x
and had random one of the drives drop/offline after 1-2 days.
i think this fixed it for me:
/etc/kernel/cmdline
pcie_aspm=off pcie_port_pm=off
now with 6 days of uptime on...
We have many points in common:
The hardware worked well in a previous software configuration
Ceph is used for SSDs, is it a CephRBD volume?
Several SSDs and several nodes are affected
The failure is random but the more the system is loaded, the...
Happy to give some background and thx for offering to give your thoughts…
Nodes: 3
Drive details per node: 1&2) 1x Samsung PM1733a 15.36TB, 1x Optane 905P 380GB as db; 3) 2x Intel D4502 7.68TB (dual link running on single link), 2x Optane P1600X...
I also want to add that I’ve been having problems with 6.8.12-9 and even 6.11.11-2. When I was using 6.8.12-4 it was somewhat ok. For reference my disks are Intel D4502 7.68TB U.2 NVMe. Other NVMe drives are fine however. Maybe I’ll revert to...
Hi,
installed a new node today, but when trying to update, I am encountering ssl issues.
Could it be that proxmox has certificate problems on their repo server?
root@pve:~# curl -v https://download.proxmox.com/debian/pve/dists/trixie/InRelease...
Hi Team,
Before connecting the SSDs with an M.2 adapter in a PCIe slot, I noticed that the disconnection of the SSDs occurs only under very specific conditions.
My configuration is based on 4 nodes and 1 qdevice with the latest PVE community...