Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

> How do you identify the end of a Unicode text stream

"Why not just a zero byte?"



You'd need 4 of them to avoid ambiguity. The C string delineation model has a lot of issues we probably don't want to keep. :)


In some encodings? One zero is sufficient for UTF-8.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: