The platform says hackers targeted a firm that helped to verify the ages of its users.

From 2025.

you are viewing a single comment's thread
view the rest of the comments
[–] 4 points 5 days ago* (2 children)

If the site metadata is made to contain fhe date then yeah, but otherwise I don't think theres a way to get the date of the html file from its url...?

If not, then there isn't any standardized publication date format which allows us to automate retrieval.

  • source
  • parent
  • hideshow 2 child comments
  • [–] 4 points 4 days ago* (last edited 4 days ago) (1 child)

    A lot of them use this in the HTML:

    <meta property="article:post_date" content="2026-09-24T15:42:36+0100">
    <meta property="article:post_modified" content="2026-09-24T15:54:02+0100">
    <meta property="article:published_time" content="2026-09-24T15:42:36Z">
    <meta property="article:modified_time" content="2026-09-24T15:54:02Z">
    

    I’ve scripted a fair bit of scraping in the past just by looking for those kinds of entries. But only focussed on a few dozen websites so I can’t say how common it is.

    If the fetching function can validate the data it should be able to continue without a date if needed.

  • source
  • parent
  • hideshow 1 child comment