# UTF8 space/difficulty tradeoffs

**URL:** <https://discourse.julialang.org/t/utf8-space-difficulty-tradeoffs/11205>\
**Category:** Internals & Design\
**Created:** [May 28, 2018, 6:14am UTC](https://discourse.julialang.org/t/utf8-space-difficulty-tradeoffs/11205 "2018-05-28T06:14:03Z")\
**Posts on this page:** 1\
**Showing post:** 1

<div class="post-metadata">

**Author:** ![ScottPJones](https://sea2.discourse-cdn.com/julialang/user_avatar/discourse.julialang.org/scottpjones/32/146_2.png) [@ScottPJones](https://discourse.julialang.org/u/ScottPJones)\
**Post date:** [May 28, 2018, 6:14am UTC](https://discourse.julialang.org/t/utf8-space-difficulty-tradeoffs/11205/1 "2018-05-28T06:14:03Z")

</div>

> [@Strftime & strptime bug #27239 is present on all platforms, not just Windows](https://discourse.julialang.org/t/strftime-strptime-bug-27239-is-present-on-all-platforms-not-just-windows/11191/2):
>
> For example, `한` is `0xc7d0` in `EUC-KR` , but `U+d55c` in Unicode and will be encoded as `0xed9f9c` in `UTF-8` .

This is something that has always concerned me about the “UTF-8 is always best” mentality, text from the languages of the vast majority of the world’s population generally takes up 50% to 200% more space encoded with `UTF-8` compared to their older national character sets (compared to `UTF-16`, which is the same size or only 100% larger, not 200%).

---

_[View the full topic](https://discourse.julialang.org/t/utf8-space-difficulty-tradeoffs/11205)._
