Why LlamaIndex defaults to Markdown output for document parsing
Original titleMarkdown is all you need.
AISummary
LlamaIndex says Markdown is its default output for document parsing because it preserves headings, lists, and tables, which helps models read content correctly. The post notes that parsers can extract every word yet lose which column a number belongs to, forcing models to guess. For tables with merged headers, LlamaIndex switches to HTML.
Source: LlamaIndex · x.comPublished · added here