# Combining CPU and GPU

**URL:** <https://discourse.julialang.org/t/combining-cpu-and-gpu/78097>\
**Category:** New to Julia\
**Tags:** gpu, performance\
**Created:** [March 18, 2022, 5:07pm UTC](https://discourse.julialang.org/t/combining-cpu-and-gpu/78097 "2022-03-18T17:07:47Z")\
**Posts on this page:** 1\
**Showing post:** 7

<div class="post-metadata">

**Author:** ![goerch](https://sea2.discourse-cdn.com/julialang/user_avatar/discourse.julialang.org/goerch/32/29122_2.png) [@goerch](https://discourse.julialang.org/u/goerch)\
**Post date:** [March 18, 2022, 8:55pm UTC](https://discourse.julialang.org/t/combining-cpu-and-gpu/78097/7 "2022-03-18T20:55:16Z")

</div>

> [@Ribeiro](#):
>
> I do think there’s a lot I can do to improve memory access and so on.

Stupid question: did you measure with `@btime` and check with a profiler? I’m only halfway qualified to talk about the CPU part of the question, but would suspect `@batch` or `@tturbo` should do better for the CPU, see [this thread](https://discourse.julialang.org/t/another-slowdown-when-using-threads-threads/78009/3) for example.

And is an exemplary MWE really out of reach?

As for your original question: I found [this paper](http://www.ziti.uni-heidelberg.de/ziti/uploads/ce_group/seminar/2015-Steffen_Lammel.pdf), but I haven’t seen any mention of such technology in this group recently. The first hit when searching Discourse is [this thread](https://discourse.julialang.org/t/several-questions-about-kernelabstractions/74740).

---

_[View the full topic](https://discourse.julialang.org/t/combining-cpu-and-gpu/78097)._
