aboutcode-org / aboutcode-org/typecode

PDF file detected as non-binary

Aperta
#41 1 commento 0 reazioni 0 assegnatari Vedi su GitHub
Lingua principale
Python
Stelle
10
Fork
15
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

I am feeding a PDF file to `typecode.contenttype.is_binary`. As PDF files are usually considered as binary files, I would have expected the file to be detected as binary, but apparently the first bytes used for detection are looking like plain-text, leading to a wrong classification.

Example file: [antartica-3427135_640_1_libtiff.pdf](https://github.com/user-attachments/files/16411602/antartica-3427135_640_1_libtiff.pdf)

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.