Convert Document Encoding to UTF-16
Description
Section titled “Description”Rewrites a text file in UTF-16. Useful when a downstream system requires UTF-16, or when a source mixes encodings that other tools reject.
Inputs
Section titled “Inputs”- Input File or Directory — the text file, or a folder of text files, to convert.
- Destination — where the converted file is written.
- Compression — No Compression, Zip, GZip or BZip2 for the output.
- Endianness (Byte Order) — Little Endian or Big Endian.
- Include BOM — tick to start the file with a byte-order mark (
FF FEfor little-endian,FE FFfor big-endian). Leave it clear for no mark.
How It Works
Section titled “How It Works”- The source’s encoding is detected. A UTF-8 file is read as UTF-8, and Windows characters pasted into one — curly quotes from Word or Excel, say — come through intact, however many. A file in another encoding is read in the encoding detected from its first part that isn’t valid UTF-8.
- The text is written unchanged — accented letters, symbols and other characters come through exactly as they were.
- A file whose encoding cannot be detected fails the step, naming the file.
- If the input path does not exist or has no files in it, the step finishes with a warning that names the path and processes nothing, and the workflow continues.
- For a folder, each file is converted and written into the destination under its own name. Files from subfolders are written into the destination itself, not into matching subfolders.
Examples
Section titled “Examples”Select the input file and browse for the file within that location. Select the desired output location, choose the byte order, and tick Include BOM if the receiving system expects one. Save and run.