html - PHP Parsing Problem -   and Â

Question

Welcome To Ask or Share your Answers For Others

html - PHP Parsing Problem -   and Â

asked Oct 17, 2021 in Technique[技术] by 深蓝 (71.8m points)

When I try to parse some html that has   sprinkled through it and then echo it, the   "turns into" this character: ?. Also, html_entity_decode() and str_replace() doesn't change it.

Why is this happening? How can I remove the ?'s?

See Question&Answers more detail:os

与恶龙缠斗过久,自身亦成为恶龙；凝视深渊过久,深渊将回以凝视…

498 views

1 Answer

深蓝 · Answer 1 · 2021-10-17T01:06:54+0000

The non-breaking space exist in UTF-8 of two bytes: 0xC2 and 0xA0.

When those bytes are represented in ISO-8859-1 (a single-byte encoding) instead of UTF-8 (a multi-byte encoding) then those bytes becomes respectively the characters ? and another non-breaking space .

Apparently you're parsing the HTML using UTF-8 and echoing the results using ISO-8859-1. To fix this problem, you need to either parse HTML using ISO-8859-1 or echo the results using UTF-8. I'd recommend to use UTF-8 all the way. Go through the PHP UTF-8 cheatsheet to align it all out.

Categories

html - PHP Parsing Problem -   and Â

Please log in or register to add a comment.

Please log in or register to answer this question.

1 Answer

Please log in or register to add a comment.

Just Browsing Browsing

Most popular tags

Categories

html - PHP Parsing Problem - &nbsp; and &#194;

Please log in or register to add a comment.

Please log in or register to answer this question.

1 Answer

Please log in or register to add a comment.

Just Browsing Browsing

Most popular tags

html - PHP Parsing Problem - and Â