When I try to parse some html that has   sprinkled through it and

Question

0

Asked: May 18, 20262026-05-18T21:03:31+00:00 2026-05-18T21:03:31+00:00

When I try to parse some html that has   sprinkled through it and

0

When I try to parse some html that has   sprinkled through it and then echo it, the   “turns into” this character: Â. Also, html_entity_decode() and str_replace() doesn’t change it.

Why is this happening? How can I remove the Â’s?

Report

Leave an answer
Cancel reply

You must login to add an answer.

Need An Account,

1 Answer

Editorial Team · Answer 1 · 2026-05-18T21:03:32+00:00

The non-breaking space exist in UTF-8 of two bytes: 0xC2 and 0xA0.

When those bytes are represented in ISO-8859-1 (a single-byte encoding) instead of UTF-8 (a multi-byte encoding) then those bytes becomes respectively the characters Â and another non-breaking space .

Apparently you’re parsing the HTML using UTF-8 and echoing the results using ISO-8859-1. To fix this problem, you need to either parse HTML using ISO-8859-1 or echo the results using UTF-8. I’d recommend to use UTF-8 all the way. Go through the PHP UTF-8 cheatsheet to align it all out.

Sign Up

Sign In

Forgot Password

The Archive Base Latest Questions

When I try to parse some html that has &nbsp; sprinkled through it and

Leave an answerCancel reply

1 Answer

See also:

When I try to parse some html that has sprinkled through it and

Leave an answer
Cancel reply