# LLaMA in Julia?

**URL:** <https://discourse.julialang.org/t/llama-in-julia/95979>\
**Category:** Offtopic\
**Created:** [March 13, 2023, 1:02am UTC](https://discourse.julialang.org/t/llama-in-julia/95979 "2023-03-13T01:02:28Z")\
**Posts on this page:** 1\
**Showing post:** 12

<div class="post-metadata">

**Author:** ![ImreSamu](https://sea2.discourse-cdn.com/julialang/user_avatar/discourse.julialang.org/imresamu/32/20677_2.png) [@ImreSamu](https://discourse.julialang.org/u/ImreSamu)\
**Post date:** [August 7, 2023, 7:32am UTC](https://discourse.julialang.org/t/llama-in-julia/95979/12 "2023-08-07T07:32:45Z")

</div>

> [@Tomas\_Pevny](#):
>
> Why there is about half speed comparing to the C version? What is their secret sauce they use?

`-Ofast -march=native .... `

- [GitHub - karpathy/llama2.c: Inference Llama 2 in one file of pure C](https://github.com/karpathy/llama2.c#performance)

---

_[View the full topic](https://discourse.julialang.org/t/llama-in-julia/95979)._
