The problem is the downloaded pages of the papers themselves, which often completely lack proper bibliographic data (or even header/footer info other than a page number). Compare this with a page from a corporate tech report, where each page might have the title and document number somewhere.
Depending on the source, it might not be published yet, or the PDF you grabbed might be a pre-print while the 'officially' published article is behind a journal paywall.
Generally, I take the publication dates of the cited works as representative of the age of the paper.