We currently only calculate hashes of pdfs and htmls
We currently only calculate hashes of pdfs and htmls