Specify how filenames from the OS map to File's `name` property
Nadie ha tomado este issue todavía.
- Lenguaje dominante
- HTML
- Estrellas
- 118
- Forks
- 52
- Merge medio
- 9 d 16 h
- PR fusionados (30 d)
- 1
Descripción
I've been testing how files coming from the OS get exposed as File objects, and in particular how filenames that aren't in the OS's default encoding get mapped to the name property.
In Windows systems, filenames are sequences of UTF-16 code units (not UTF-16 encoded text, as is sometimes claimed, because the system APIs don't check for lone surrogates), and as expected, they directly map to a DOMString. An initial BOM doesn't get removed. There doesn't seem to be any browser differences here.
In Unix systems (tested on Fedora Linux; my understanding is all other modern Unix variants/distros work the same), filenames are byte sequences, which are usually taken to be UTF-8. Here's how the various browsers behave on them:
- Firefox does the equivalent of UTF-8 decode without BOM, decoding bytes which aren't valid UTF-8 as a replacement character.
- WebKit does the equivalent of UTF-8 decode without BOM or fail, and handles failures by returning a
Fileobject with the empty string as filename, empty contents, and MIME typeapplication/octet-streaminstead. The language aroundnamein the spec might allow for an empty string to substitute a filename that cannot be decoded, but it doesn't allow the content to be dropped. Note that the resultingFileobject is identical to theFileobject that HTML's "construct the entry list" creates when a file input has no selected files. - Chrome also does the equivalent of UTF-8 decode without BOM or fail, except that for file inputs, any file whose filename isn't UTF-8 gets dropped from the selection. For drag and drop, Chrome behaves the same as WebKit.
Since it doesn't seem good to drop files or replace them by an empty file, even when their filenames don't match the OS's conventions, it seems like it would be best to agree on Firefox's behavior.
Guía de contribución
Primeros pasos
- Lee el issue completo y luego la guía de contribución del proyecto.
- Comenta en el issue que vas a ocuparte — evita que dos personas hagan lo mismo.
- Haz un fork del repositorio y trabaja en una rama.
- Abre un pull request que haga referencia al número del issue.
Línea de trabajo
Lee el texto de File API que describe cómo el nombre de un archivo del sistema operativo se convierte en el nombre de un File, y luego compara los algoritmos vinculados de Encoding y HTML. El issue estará terminado cuando la especificación defina normativamente el manejo de nombres de archivo no válidos y el comportamiento resultante de File, sin dejar sin resolver diferencias entre navegadores.
Escrito por el modelo de indexación a partir del texto del issue.
Evaluación
- Stack tecnológico
- html
- Área
- api, web-dev
- Tipo de issue
- Nueva funcionalidad
- Dificultad
- 4/5
- Tiempo estimado
- 3-5 días
- Estado de actividad
- Estancado
- Claridad
- Bastante claro
- Aptitud para principiantes
- 25/100