#XML-formatted
Fade In now supports import and/or export of: FDX, .docx, Fountain, PDF, RTF, .txt (formatted or unformatted), HTML, XML, Markdown, and .epub, plus Highland, Scrivener, Celtx, StudioBinder, Avid, and Adobe Story.
September 20, 2026 at 5:34 PM
Last post for today – an update on the PM Transcripts section of the GLAM Workbench: https://updates.timsherratt.org/2026/09/15/pm-transcripts-in-the-glam.html
The Department of Prime Minister and Cabinet has compiled a large collection of speeches, interviews, press releases etc by Australian prime ministers and made them available through the PM Transcripts site. You can search through the transcripts, and download the underlying XML files using a simple API. Back in 2016 (and again in 2019), I downloaded all the XML-formatted transcript files and shared them through a GitHub repository to make it easier for researchers to use the collection as a whole. Many more transcripts have been added in the years since, so an update was long overdue! Last week I reharvested the complete collection, which now comprises 28,560 transcripts. The only change to the harvesting process was that the file that lists all of the transcript file names is now generated by a batch process and has to be downloaded manually. Once that was saved, I could just plug it in to the old code. As previously, I created a CSV index to the transcripts, and did a little analysis and visualisation. Here, for example, is the number of transcripts by year and prime minister – you can see a lot of Scomo and Albo has been added since my original harvest. I also aggregated the transcript files by prime minister – creating a single text file for each PM, as well as a zip file containing all the XML transcripts. That should make it easier to load up all the words of a particular PM for analysis. I completely overhauled the PM Transcripts section of the GLAM Workbench in the process, updating it to the latest format. There you’ll find all the notebooks I used to process the data, as well as details of the complete dataset. The transcripts are made available by the Commonwealth under a CC-BY licence.
updates.timsherratt.org
September 15, 2026 at 6:52 AM
We can't really plan for what the software eco-system will look like in 20-30 years. But we can preserve the data. Current plan is to produce a formatted dump of the contents in a nice, human-readable PDF (or 12), plus a dump into an open XML format file.
August 6, 2026 at 8:29 PM
PDFs are formatted legal codes. This uses whatever *that* is sourced from, generally XML. It’s designed for laws—I have zero knowledge of its applicability to other domains, so if it worked, that would be coincidental.
July 14, 2026 at 12:34 PM
I vastly prefer XML to HTML because you can mess around with it in a way thats not possible in HTML.

As long as your tags are correctly formatted, you can otherwise just make shit up.

Like: <invader>Zim</invader>
June 8, 2026 at 9:28 AM
Then I will definitely look into it! Thank you for bringing it to my attention.

From a cursory glance, it seems like the core is just websites publishing a specially formatted XML, that RSS readers poll for updates.

I already use XML for the Series page backend, so that shouldn't be too difficult.
May 30, 2026 at 11:58 PM
if the underlying machine format is XML, they should be able to publish formatted diffs of the affected statutes too
May 18, 2026 at 4:09 PM
can we talk about what this implies about how legislation is being drafted and formatted by the united states congress
May 18, 2026 at 3:57 PM
Parts 1-2/2 — EFTA01226838.jpg
#epsteinweb #efta01226838
https://epsteinweb.org
Available in the iOS app store now!
https://apps.apple.com/us/app/epstein-web/id6758880661
May 14, 2026 at 12:38 AM
Neat new M&A database with thousands of XML-formatted and tagged agreements: pandects.org

Accompanying paper: papers.ssrn.com/sol3/papers....
Pandects - Open-Source M&A Agreement Search & Data
Search and download structured M&A agreements from SEC EDGAR. Tag clauses, extract terms, and export CSVs.
pandects.org
April 27, 2026 at 4:07 PM
April 19, 2026 at 3:18 AM
April 19, 2026 at 3:05 AM
April 19, 2026 at 3:05 AM
April 19, 2026 at 2:52 AM
April 17, 2026 at 9:55 PM
April 12, 2026 at 12:01 AM
April 11, 2026 at 7:55 PM
April 11, 2026 at 7:54 PM
April 11, 2026 at 7:54 PM
April 11, 2026 at 7:52 PM
April 10, 2026 at 10:30 PM
Rewind ⏪ Symfony Messenger #46
When a message is handled async by a Messenger worker, it's dispatched *back* through the message bus! And *stamps* are a key part of the process. Let's take a closer look at this & add a missing BusNameStamp in our custom serializer symfonycasts.com/screencast/m...
Custom Transport Serializer
If an external system sends messages to a queue that we're going to read, those messages will probably be sent as JSON or XML. We added a message formatted as JSON
symfonycasts.com
April 10, 2026 at 8:01 AM
Rewind ⏪ Symfony Messenger #45
One of the most *versatile* parts of a Messenger transport is its *serializer*. Let's create a custom serializer & use it to decode JSON messages from the queue into the exact message objects we need. symfonycasts.com/screencast/m...
Custom Transport Serializer
If an external system sends messages to a queue that we're going to read, those messages will probably be sent as JSON or XML. We added a message formatted as JSON
symfonycasts.com
April 9, 2026 at 8:01 AM