Falk hat absolut recht - Netzwerkprobleme können zu einer OSD-degradierung (OSD = out) führen, die nur manuell aufgelöst werden kann. Siehe https://docs.ceph.com/en/latest/rados/configuration/mon-osd-interaction
Ursachen hierfür kann sein, dass...
Ich würde auch das Netzwerk nicht vernachlässigen. Hängende OSDs kenne ich nur von defekten Switches die das LACP nicht mehr sauber bedienen und zu viele Pakete verloren gehen oder fehlerhafte Konfigurationen die zu ähnlichen Phänomenen führen...
Hallo @Winet.maier ,
auf den ersten Blick sehe ich hier zwei Probleme:
1. Die OSD muss neu gestartet werden (ist mir in der Praxis noch nie vorgekommen) und
2. der Cluster bleibt "stehen" sobald eine OSD vom Netz geht.
Um letzteres wirklich...
Wouldn't it be more precise to say that PVE on DHCP isn't supported without other tweaks even on single node setups?
Some services require to resolve the host name to an IP that is currently assigned to the host to work properly.
Hi,
first off - using DHCP with Proxmox VE (and physical servers in general, IMO) isn't really recommended.
In any case - this is totally expected, as nic0 is attached to the vmbr0 bridge. In such network scenarios, the bridge should have the...
Sind die OSDs mit Version 19 erstellt worden?
https://docs.clyso.com/docs/kb/known-bugs/squid/
Wobei in der Fehlerbeschreibung kein Hängen sondern ein Crash erwähnt wird.
If you don't want to use proxmox offline mirror then I'd suggest using a HTTP-Proxy. This way you can whitelist the URLs. Allowing CDN IPs is (depending on the CDN) not really the most secure solution anyways since IPs are reused by multiple...
Beim Interface warnt Checkmk, weil sich die Geschwindigkeit seit dem ersten Service Discovery geändert hat.
Ist das tatsächlich der Fall (von 10G auf 1G)?
Oder ist das jetzt einfach anderes Interface an dem Index? Dann sollte besser der...
Du hast doch im CheckMK Forum gesagt bekommen, das es ein Kindprozess (PVE Sheduler) handelt kann und der Zähler zu klein im CheckMK ist.
Siehe: https://forum.checkmk.com/t/proxmox-ntp-time-pveschedule-interface-speed/59729/4
Es wurde dort...
cluster network is for osd->osd replication.
osd->monitor && osd->client is done on public network.
They are differents hearbeats, osd hearbeat between osd (on clusternetwork)
hearbeat from monitor-> osd (on public network)...
It is preferable to have two rings as a misconfigured switching equipment on a bond might make the link down or flap for no reason.
Imagine a failing switch with one of two members of a bonded link dropping say, 50% of the traffic on a bonded...
Any advice on powering down two clusters ?
I have a 6 node closure on 8.3.1 (Enterprise repo) with 50-65 VMs pre node. Storage is a separate 5 node Proxmox cluster with Ceph same 8.3.1 version - this cluster is not running nay VMs.
We have to...
OSDs need both networks (Ceph public_network and cluster_network) to function on all nodes to be able to properly work .
If one of these two networks do not work the OSDs flap.
That’s what I was trying to say…AFAIK the cluster network is only for replication. The VMs talk to it on the public network so that needs to be working.
Why did you open a new thread and why didn't you link the old one with it's critism?
I still think that handing the keys to a vibecoded ai agent is a bad idea...
Thanks for sharing. Please note that Proxmox is a registered trademark, so please don't use it as part of your own product/app/tool name.
You can indicate that your work is for Proxmox VE by, e.g., using a wording that avoids suggesting that this...