hexdump confusion

Question

Welcome To Ask or Share your Answers For Others

hexdump confusion

posted Oct 7, 2021 in Technique[技术] by 深蓝 (71.8m points)

hexdump confusion

I am playing with the Unix hexdump utility. My input file is UTF-8 encoded, containing a single character ?, which is C3 B1 in hexadecimal UTF-8.

hexdump test.txt
0000000 b1c3
0000002

Huh? This shows B1 C3 - the inverse of what I expected! Can someone explain?

For getting the expected output I do:

hexdump -C test.txt
00000000  c3 b1                                             |..|
00000002

I was thinking I understood encoding systems.

question from:https://stackoverflow.com/questions/2847360/hexdump-confusion

与恶龙缠斗过久,自身亦成为恶龙；凝视深渊过久,深渊将回以凝视…

1 Reply

深蓝 · Answer 1 · 2021-10-06T17:42:12+0000

This is because hexdump defaults to using 16-bit words and you are running on a little-endian architecture. The byte sequence b1 c3 is thus interpreted as the hex word c3b1. The -C option forces hexdump to work with bytes instead of words.

Categories

hexdump confusion

hexdump confusion

Please log in or register to add a comment.

Please log in or register to reply this article.

1 Reply

Please log in or register to add a comment.

Just Browsing Browsing

Most popular tags