Yeah, dates in submission links would be great. Wonder if it can be automated so when I submit a URL, lemmy fetches the creation date while it’s fetching the headline.
I’ve scripted a fair bit of scraping in the past just by looking for those kinds of entries. But only focussed on a few dozen websites so I can’t say how common it is.
If the fetching function can validate the data it should be able to continue without a date if needed.
Yeah, dates in submission links would be great. Wonder if it can be automated so when I submit a URL, lemmy fetches the creation date while it’s fetching the headline.
If the site metadata is made to contain fhe date then yeah, but otherwise I don’t think theres a way to get the date of the html file from its url…?
If not, then there isn’t any standardized publication date format which allows us to automate retrieval.
A lot of them use this in the HTML:
<meta property="article:post_date" content="2026-09-24T15:42:36+0100"> <meta property="article:post_modified" content="2026-09-24T15:54:02+0100"> <meta property="article:published_time" content="2026-09-24T15:42:36Z"> <meta property="article:modified_time" content="2026-09-24T15:54:02Z">I’ve scripted a fair bit of scraping in the past just by looking for those kinds of entries. But only focussed on a few dozen websites so I can’t say how common it is.
If the fetching function can validate the data it should be able to continue without a date if needed.
IT would be neat to auto-add yeah.