I’ve scripted a fair bit of scraping in the past just by looking for those kinds of entries. But only focussed on a few dozen websites so I can’t say how common it is. Still, if the fetching function can validate the data it should be able to decide to continue without a date in those cases.
A lot of them use this in the HTML:
<meta property="article:post_date" content="2026-09-24T15:42:36+0100"> <meta property="article:post_modified" content="2026-09-24T15:54:02+0100"> <meta property="article:published_time" content="2026-09-24T15:42:36Z"> <meta property="article:modified_time" content="2026-09-24T15:54:02Z">I’ve scripted a fair bit of scraping in the past just by looking for those kinds of entries. But only focussed on a few dozen websites so I can’t say how common it is. Still, if the fetching function can validate the data it should be able to decide to continue without a date in those cases.
IT would be neat to auto-add yeah.