invalid byte 2 of 2-byte UTF-8 sequence

3 Answers

Most commonly it's due to feeding ISO-8859-x (Latin-x, like Latin-1) but parser thinking it is getting UTF-8. Certain sequences of Latin-1 characters (two consecutive characters with accents or umlauts) form something that is invalid as UTF-8, and specifically such that based on first byte, second byte has unexpected high-order bits.

This can easily occur when some process dumps out XML using Latin-1, but either forgets to output XML declaration (in which case XML parser must default to UTF-8, as per XML specs), or claims it's UTF-8 even when it isn't.

179

answered Oct 22 '22 17:10

Related questions
                            
                                Android M onRequestPermissionsResult in non-Activity
                            
                                How to set DPI information in an image?
                            
                                Scala collection standard practice
                            
                                Google guava vs Scala collection framework comparison
                            
                                readValue and readTree in Jackson: when to use which?
                            
                                How to prevent Gson from converting a long number (a json string ) to scientific notation format?
                            
                                Java: A "prime" number or a "power of two" as HashMap size?
                            
                                Form submit in Spring MVC 3 - explanation
                            
                                How to check if a String matches a pattern in Groovy
                            
                                Is there a way to use Java 8 functional interfaces on Android API below 24?
                            
                                queryPurchases() vs queryPurchaseHistoryAsync() in order to 'restore' functionality?
                            
                                Java JDBC Lazy-Loaded ResultSet
                            
                                Does closing the inputstream of a socket also close the socket connection?
                            
                                Advantages of Java's enum over the old "Typesafe Enum" pattern?
                            
                                How Iterator's remove method actually remove an object
                            
                                What does the @code java annotation do
                            
                                ImageIO reading slightly different RGB values than other methods
                            
                                Exception in thread "main" java.lang.NoSuchMethodError: java.nio.ByteBuffer.flip()Ljava/nio/ByteBuffer
                            
                                Modular web apps
                            
                                Spring MVC Form tags: Is there a standard way to add "No selection" item?

Donate For Us

If you love us? You can donate to us via Paypal or buy me a coffee so we can maintain and grow! Thank you!

Donate Us With

invalid byte 2 of 2-byte UTF-8 sequence

Tags:

java

xml

encoding

flyingfromchina

People also ask

3 Answers

StaxMan

Ignacio Vazquez-Abrams

atott

Recent Activity

Donate For Us