Why this is a reconstruction rather than a conversion
A PDF page contains text-showing operators: set this font at this size, move to this coordinate, draw this string of glyphs. That is the whole model. There is no paragraph object, no sentence, no reading order, no notion that two lines belong to the same block of prose. The information was discarded when the document was written, because a PDF only needs to know where to put ink.
A converter therefore reverses an inference. It groups glyph runs into lines by their vertical position, groups lines into paragraphs by their spacing and left edges, and guesses at headings from font size. Word then needs all of that, because a .docx describes paragraphs that reflow.
Where the guesses are right the output is genuinely editable. Where they are wrong you get the familiar symptoms: a paragraph broken at every visual line, a two-column page interleaved line by line, a header repeated as body text, a table rendered as rows of tab-separated text. None of these are defects in the reading of the PDF — the reading was accurate — they are the cost of having to invent structure.