This isn’t surprising: XML’s core purpose was to simplify SGML for a wider breadth of applications on the web.
HTML also descended from SGML, and it’s hard to imagine a more deeply grooved structure in these models, given their training data.
So if you want to annotate text with semantics in a way models will understand…