Legacy Content · XML & JSON Conversion · AI-Enabled · Accessibility & Metadata
We convert books, journals, educational materials, and legacy archives into the format your workflow requires — JATS, BITS, NLM DTD, DocBook, DITA, QTI, JSON, and more — using AI-assisted processing, with accessibility tagging and metadata enrichment built in.
What each format enables
Each format we convert to opens up a specific set of platforms, workflows, and opportunities. Here's what becomes possible once your content is properly structured.
Proof, not promises
A global academic content platform serving 1.6M+ readers across 195 countries, working with 130+ publishers, needed older articles brought up to JATS so they’d meet indexing requirements. We mapped a transformation roadmap, built custom XSLT transforms, converted 40,000+ articles to JATS-compliant XML, and converted select content to NIMAS for accessibility — then republished it directly into their production pipeline.
Read the full case study →How it works
No long procurement process before you see a result — the sample batch comes before any commitment.
Tell us your source format, schema, and indexing requirements.
We confirm the target schema — JATS, BITS, NLM DTD, DocBook, DITA, QTI, JSON, or any custom format — and any publisher-specific customisations.
5 articles or 2 book titles, converted at no cost, in 5 working days.
Your backlist converted at an agreed turnaround and price per page.
Validated output delivered into your platform or repository.
How AI fits into the workflow
We use AI-assisted analysis to identify document structure, tables, equations, references, and metadata — then route content through the right processing path based on complexity. Validation and enrichment are automated where repeatable; human reviewers handle exceptions and edge cases.
Why publishers choose our approach
Automation handles repeatable work. Experts focus on exceptions, edge cases, and output quality.
JATS, BITS, NLM DTD, DocBook, DITA, QTI, JSON, and custom schemas — we work to the format your workflow requires.
Structured content preserves and enriches the information needed for accessible digital outputs as standard.
What we handle
We work to the schema your indexing service or repository requires — JATS, BITS, NLM DTD, DocBook, and publisher-specific DTD customisations.
Word, InDesign, born-digital and scanned PDF, LaTeX, and legacy XML in custom schemas all come into the same pipeline.
Equations in MathML, chemical structures, multilingual text, complex tables, figures with captions, footnotes, and cross-references are handled as standard, not as exceptions.
Schema and Schematron rule validation, custom XSLT checks, and a manual editorial QA pass — before anything is delivered.
oXygen XML Editor, XSLT 2.0/3.0 transforms, Python-based automation, and custom CMS integrations for repeat batches.
Our team converts up to 35,000 pages a month, across journals, books, educational content, and legacy archives.
Send us 5 articles or 2 book titles. We'll convert them to your schema in 5 working days, at no cost, with no commitment.