Denmark-based data scientist Tommy Carstensen launched and maintains an interactive public archive of the U.S. Department of Justice's unclassified files related to the prosecution of Jeffrey Epstein, after the department missed a legally mandated December 19, 2025 deadline to release the materials. Lawmakers in December 2025 accused the department of failing to comply with a law that required declassification and release of the files by that date.

Carstensen, a bioinformatician, has published interactive graphics of Epstein's properties and financial transactions, an analysis of more than 1 million documents released by the department that groups records into subject areas, and court records along with transcripts of audio and video files from the releases. He has also published a facial recognition tool that allows users to upload an image of a face to check whether it appears in any images in the files.

Carstensen spends as much as 50 hours per week maintaining the archive in addition to his full-time job. He was motivated to build the archive after lawmakers accused the Justice Department of failing to comply with the law on the Epstein files. Prior to this project, he participated in online efforts to identify participants in the January 6, 2021 Capitol insurrection.

To guard against republishing sensitive material, Carstensen wrote code to monitor the department's website for changes to the released files. He maintains a list of victim names and names of victims' family members that are automatically redacted from his archive, and images of known survivors' faces and of minors are also redacted. He has been responsive to takedown requests from survivors. Journalists and researchers have praised his archive efforts.

The Justice Department previously removed or retroactively redacted documents after erroneously failing to redact identifying information on victims. The department has conceded it made some errors during the release of the files but maintains it has ultimately complied with the Epstein Files Transparency Act. A department watchdog is investigating the redaction and release issues.

A number of journalists, researchers, and activists have applied technical analyses to the Epstein files to extract information not readily available in the department's raw data dumps. Earlier this month, the non-profit Decoherence Media published a searchable database of faces of individuals who appear in original images in the files. Faces were identified in part using Amazon Web Services Rekognition facial recognition technology, with identifications double-checked using multiple recognition models and manual review. The database includes images of more than 100 individuals not mentioned in Epstein's email files, and images of nearly 200 individuals who have not been reported on, including a Hollywood agent and the head of a large fitness chain. Appearing in the Epstein records does not indicate wrongdoing.

Distributed Denial of Secrets, which publishes leaked and hacked datasets, obtained an archive of more than 20,000 unredacted emails from Epstein's Yahoo account and has made the complete archive available only to vetted researchers and journalists. Those cleared to access it may republish only emails relevant to their reporting and must assume people in the emails are potential victims unless determined otherwise. The Yahoo emails, with redactions to protect victims and minors, are being released by Jmail, a browser-based archive of Epstein's emails and other files developed by a group of volunteer tech workers and engineers.