DataToolsLab
XML Tools

XML Formatter

Pretty-print messy or minified XML with a configurable indent and one attribute per line for long tags.

Loading tool…

What is XML Formatter?

An XML formatter that re-indents a document for reading and diffing. It parses first and prints second, so a document that cannot be parsed is reported instead of being reformatted into something that only looks right.

How it works

The parser builds a tree with attributes, comments, CDATA sections and processing instructions preserved. Printing indents element-only content — a child element whose parent has no text of its own — and leaves mixed content on one line, because moving whitespace around inside `<p>Hello <b>world</b></p>` changes the text a consumer sees. Start tags longer than a set width wrap one attribute per line. Empty elements can be written as `<a/>` or `<a></a>`, and unused namespace declarations can be dropped.

  1. Paste the document. Drop in XML, a schema, a feed or an XPath expression. Everything is parsed in the page: no upload, no server round trip, and the tools keep working with the network switched off.
  2. Set the options that match your document. Indent unit, whether to keep comments, which dialect, which output method. The defaults are the safe ones — nothing that changes meaning is enabled for you.
  3. Read the result, then copy or download it. Results appear as you type, copy straight to the clipboard, and download with a sensible filename (.xml, .xsd, .csv). Reset returns every field to its default.

Examples

A minified document from an API response

Paste one line of XML and the Formatter indents it, wraps long start tags one attribute per line and normalises empty elements — without touching text inside mixed content.

An XPath that works in Chrome and not in your servlet

The XPath Tester evaluates against a real tree and reports the axis or function it cannot support, instead of returning zero nodes and letting you guess why.

A document that will not parse

The Validator gives line and column for every mismatched tag, stray ampersand and duplicate attribute, with the reason in words: which open element was expected, which was found.

A schema that only half matches your data

The XSD Validator marks each violation with the path of the offending node, and the Schema Generator goes the other way — inferring a starting-point XSD from real documents.

Common mistakes

Assuming a pretty-printer cannot change meaning

Whitespace between elements is often insignificant, but inside mixed content it is not: `<p>Hello <b>world</b></p>` must stay on one line. Formatting here re-indents element-only content and leaves mixed content alone.

Treating a validation pass as proof of interoperability

The DTD and XSD validators implement honest subsets. A document reported valid here can still be rejected by a full validator that understands identity constraints, substitution groups or XSD 1.1 assertions — each page lists what it does not cover.

Reading the whole document because the parser did

A DOCTYPE is not inert. If your parser resolves external entities, pasting an untrusted document is enough to read a file or call an internal URL. The XXE Risk Checker shows what a document asks for; the fix is in the parser configuration.

Forgetting that attributes and elements are not interchangeable

`<id>1</id>` and `id="1"` mean different things to every schema and every XPath. When converting to CSV or YAML the distinction is preserved (`@id`), because flattening it away is the change that costs an afternoon later.

Frequently asked questions

Why did part of my document not get re-indented?

Because re-indenting it would change the document. Text between elements is content when it sits alongside text — in `<p>Hello <b>world</b></p>` the space before `<b>` is part of the paragraph. The formatter detects mixed content and leaves it alone, and it also respects xml:space="preserve".

Does formatting change the file's meaning?

Not for element-only content: whitespace between tags there is insignificant by the XML specification, so adding indentation is safe. Where it is significant the formatter refuses to touch it. Attribute order is never changed, because `diff` on two differently-ordered documents would otherwise become noise.

Can it fix XML that does not parse?

No, and anything claiming to is guessing. Use the Validator to find the mismatched tag, fix it, and format afterwards — the formatter's error messages point at the same line and column.

Online