# Most iportant from GTC Cuda on x86 hello emulation mode

**URL:** <https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897>\
**Category:** CUDA Programming and Performance\
**Created:** [September 22, 2010, 9:17pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897 "2010-09-22T21:17:06Z")\
**Posts on this page:** 20\
**Page:** 1

<div class="post-metadata">

**Author:** ![Lev](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@Lev](https://forums.developer.nvidia.com/u/Lev)\
**Post date:** [September 22, 2010, 9:17pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/1 "2010-09-22T21:17:06Z")

</div>

[url=“[Breaking News: Jen-Hsun Announces CUDA for x86 Architecture - BSN\*](http://www.brightsideofnews.com/news/2010/9/21/breaking-news-jen-hsun-announces-cuda-for-x86-architecture.aspx)”][http://www.brightsideofnews.com/news/2010/...chitecture.aspx[/url]](http://www.brightsideofnews.com/news/2010/...chitecture.aspx%5B/url%5D)

---

<div class="post-metadata">

**Author:** ![cbuchner1](https://sea2.discourse-cdn.com/nvidia/user_avatar/forums.developer.nvidia.com/cbuchner1/32/14168_2.png) [@cbuchner1](https://forums.developer.nvidia.com/u/cbuchner1)\
**Post date:** [September 22, 2010, 11:47pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/2 "2010-09-22T23:47:56Z")

</div>

> [@](#):
>
> http://www.brightsideofnews.com/news/2010/…chitecture.aspx

According to Heise News, there will be a commercial compiler for CUDA code by Portland Group (PGI) available (or officially intruduced) starting November 13th. Do you really think x86 CUDA will become part of the toolkit? I doubt that.

http://www.heise.de/newsticker/meldung/GTC…ig-1083447.html

---

<div class="post-metadata">

**Author:** ![cbuchner1](https://sea2.discourse-cdn.com/nvidia/user_avatar/forums.developer.nvidia.com/cbuchner1/32/14168_2.png) [@cbuchner1](https://forums.developer.nvidia.com/u/cbuchner1)\
**Post date:** [September 22, 2010, 11:47pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/3 "2010-09-22T23:47:56Z")

</div>

> [@](#):
>
> http://www.brightsideofnews.com/news/2010/…chitecture.aspx

According to Heise News, there will be a commercial compiler for CUDA code by Portland Group (PGI) available (or officially intruduced) starting November 13th. Do you really think x86 CUDA will become part of the toolkit? I doubt that.

http://www.heise.de/newsticker/meldung/GTC…ig-1083447.html

---

<div class="post-metadata">

**Author:** ![E.D\_Riedijk](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@E.D\_Riedijk](https://forums.developer.nvidia.com/u/E.D_Riedijk)\
**Post date:** [September 23, 2010, 7:35am UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/4 "2010-09-23T07:35:31Z")

</div>

This is also no emulation mode as it was before. It is a conversion from CUDA C to x86 machine code.

---

<div class="post-metadata">

**Author:** ![E.D\_Riedijk](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@E.D\_Riedijk](https://forums.developer.nvidia.com/u/E.D_Riedijk)\
**Post date:** [September 23, 2010, 7:35am UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/5 "2010-09-23T07:35:31Z")

</div>

This is also no emulation mode as it was before. It is a conversion from CUDA C to x86 machine code.

---

<div class="post-metadata">

**Author:** ![eyalhir74](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@eyalhir74](https://forums.developer.nvidia.com/u/eyalhir74)\
**Post date:** [September 23, 2010, 8:54am UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/6 "2010-09-23T08:54:15Z")

</div>

> [@](#):
>
> This is also no emulation mode as it was before. It is a conversion from CUDA C to x86 machine code.

I’ve recently moved to 3.2… I just dont have enough words to say how much I miss emulation mode…

it was just perfect :)

eyal

---

<div class="post-metadata">

**Author:** ![eyalhir74](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@eyalhir74](https://forums.developer.nvidia.com/u/eyalhir74)\
**Post date:** [September 23, 2010, 8:54am UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/7 "2010-09-23T08:54:15Z")

</div>

> [@](#):
>
> This is also no emulation mode as it was before. It is a conversion from CUDA C to x86 machine code.

I’ve recently moved to 3.2… I just dont have enough words to say how much I miss emulation mode…

it was just perfect :)

eyal

---

<div class="post-metadata">

**Author:** ![Lev](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@Lev](https://forums.developer.nvidia.com/u/Lev)\
**Post date:** [September 23, 2010, 11:56am UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/8 "2010-09-23T11:56:39Z")

</div>

> [@](#):
>
> This is also no emulation mode as it was before. It is a conversion from CUDA C to x86 machine code.

So your cuda program is run same way on x86, seems you can debug it etc on x86. You compile your code to x86, sounds like it is better than emulation mode cause it is faster and more precise.

---

<div class="post-metadata">

**Author:** ![Lev](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@Lev](https://forums.developer.nvidia.com/u/Lev)\
**Post date:** [September 23, 2010, 11:56am UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/9 "2010-09-23T11:56:39Z")

</div>

> [@](#):
>
> This is also no emulation mode as it was before. It is a conversion from CUDA C to x86 machine code.

So your cuda program is run same way on x86, seems you can debug it etc on x86. You compile your code to x86, sounds like it is better than emulation mode cause it is faster and more precise.

---

<div class="post-metadata">

**Author:** ![Lev](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@Lev](https://forums.developer.nvidia.com/u/Lev)\
**Post date:** [September 23, 2010, 12:00pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/10 "2010-09-23T12:00:09Z")

</div>

> [@](#):
>
> According to Heise News, there will be a commercial compiler for CUDA code by Portland Group (PGI) available (or officially intruduced) starting November 13th. Do you really think x86 CUDA will become part of the toolkit? I doubt that.
> 
> http://www.heise.de/newsticker/meldung/GTC…ig-1083447.html

Maybe it will have free licence for debug purpose. I.e. if you do not distribute your program with it, just use it for debug. I think it is good solution. Those who want thier cuda program run on x86 may buy licence.

---

<div class="post-metadata">

**Author:** ![Lev](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@Lev](https://forums.developer.nvidia.com/u/Lev)\
**Post date:** [September 23, 2010, 12:00pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/11 "2010-09-23T12:00:09Z")

</div>

> [@](#):
>
> According to Heise News, there will be a commercial compiler for CUDA code by Portland Group (PGI) available (or officially intruduced) starting November 13th. Do you really think x86 CUDA will become part of the toolkit? I doubt that.
> 
> http://www.heise.de/newsticker/meldung/GTC…ig-1083447.html

Maybe it will have free licence for debug purpose. I.e. if you do not distribute your program with it, just use it for debug. I think it is good solution. Those who want thier cuda program run on x86 may buy licence.

---

<div class="post-metadata">

**Author:** ![Ken\_Domino](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@Ken\_Domino](https://forums.developer.nvidia.com/u/Ken_Domino)\
**Post date:** [September 23, 2010, 12:42pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/12 "2010-09-23T12:42:48Z")

</div>

> [@](#):
>
> http://www.brightsideofnews.com/news/2010/…chitecture.aspx

Are there any details known yet? For example, are they going to “loosely” integrate an x86 processor to the “GPU” (not sure what it will be called then), having direct access to the Interconnection Network along with the TPC’s? Or, are they going to tightly integrate x86 by somehow replacing the SP’s in the TPC’s with x86-like stream processors, replacing or adding to PTX with x86?

---

<div class="post-metadata">

**Author:** ![Ken\_Domino](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@Ken\_Domino](https://forums.developer.nvidia.com/u/Ken_Domino)\
**Post date:** [September 23, 2010, 12:42pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/13 "2010-09-23T12:42:48Z")

</div>

> [@](#):
>
> http://www.brightsideofnews.com/news/2010/…chitecture.aspx

Are there any details known yet? For example, are they going to “loosely” integrate an x86 processor to the “GPU” (not sure what it will be called then), having direct access to the Interconnection Network along with the TPC’s? Or, are they going to tightly integrate x86 by somehow replacing the SP’s in the TPC’s with x86-like stream processors, replacing or adding to PTX with x86?

---

<div class="post-metadata">

**Author:** ![E.D\_Riedijk](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@E.D\_Riedijk](https://forums.developer.nvidia.com/u/E.D_Riedijk)\
**Post date:** [September 23, 2010, 12:49pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/14 "2010-09-23T12:49:37Z")

</div>

> [@](#):
>
> Are there any details known yet? For example, are they going to “loosely” integrate an x86 processor to the “GPU” (not sure what it will be called then), having direct access to the Interconnection Network along with the TPC’s? Or, are they going to tightly integrate x86 by somehow replacing the SP’s in the TPC’s with x86-like stream processors, replacing or adding to PTX with x86?

No, they will just compile CUDA code to x86 binaries. It will not run on a gpu, but on the cpu.

---

<div class="post-metadata">

**Author:** ![E.D\_Riedijk](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@E.D\_Riedijk](https://forums.developer.nvidia.com/u/E.D_Riedijk)\
**Post date:** [September 23, 2010, 12:49pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/15 "2010-09-23T12:49:37Z")

</div>

> [@](#):
>
> Are there any details known yet? For example, are they going to “loosely” integrate an x86 processor to the “GPU” (not sure what it will be called then), having direct access to the Interconnection Network along with the TPC’s? Or, are they going to tightly integrate x86 by somehow replacing the SP’s in the TPC’s with x86-like stream processors, replacing or adding to PTX with x86?

No, they will just compile CUDA code to x86 binaries. It will not run on a gpu, but on the cpu.

---

<div class="post-metadata">

**Author:** ![seibert](https://sea2.discourse-cdn.com/nvidia/user_avatar/forums.developer.nvidia.com/seibert/32/520688_2.png) [@seibert](https://forums.developer.nvidia.com/u/seibert)\
**Post date:** [September 23, 2010, 1:50pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/16 "2010-09-23T13:50:57Z")

</div>

> [@](#):
>
> Are there any details known yet? For example, are they going to “loosely” integrate an x86 processor to the “GPU” (not sure what it will be called then), having direct access to the Interconnection Network along with the TPC’s? Or, are they going to tightly integrate x86 by somehow replacing the SP’s in the TPC’s with x86-like stream processors, replacing or adding to PTX with x86?

I agree with E.D. Riedijk. This is almost certainly a commercially supported compiler that does what many other academic projects have been dabbling in for years: Take CUDA source code and generate multithreaded SSE x86 code. If done well, I bet a lot of people would find that CUDA on x86 is faster than even their normal CPU implementations. (Because most compilers are terrible at generating SSE instructions from all but the simplest C code and most of us are terrible at writing SSE by hand.)

---

<div class="post-metadata">

**Author:** ![seibert](https://sea2.discourse-cdn.com/nvidia/user_avatar/forums.developer.nvidia.com/seibert/32/520688_2.png) [@seibert](https://forums.developer.nvidia.com/u/seibert)\
**Post date:** [September 23, 2010, 1:50pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/17 "2010-09-23T13:50:57Z")

</div>

> [@](#):
>
> Are there any details known yet? For example, are they going to “loosely” integrate an x86 processor to the “GPU” (not sure what it will be called then), having direct access to the Interconnection Network along with the TPC’s? Or, are they going to tightly integrate x86 by somehow replacing the SP’s in the TPC’s with x86-like stream processors, replacing or adding to PTX with x86?

I agree with E.D. Riedijk. This is almost certainly a commercially supported compiler that does what many other academic projects have been dabbling in for years: Take CUDA source code and generate multithreaded SSE x86 code. If done well, I bet a lot of people would find that CUDA on x86 is faster than even their normal CPU implementations. (Because most compilers are terrible at generating SSE instructions from all but the simplest C code and most of us are terrible at writing SSE by hand.)

---

<div class="post-metadata">

**Author:** ![E.D\_Riedijk](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@E.D\_Riedijk](https://forums.developer.nvidia.com/u/E.D_Riedijk)\
**Post date:** [September 23, 2010, 3:11pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/18 "2010-09-23T15:11:28Z")

</div>

> [@](#):
>
> I agree with E.D. Riedijk. This is almost certainly a commercially supported compiler that does what many other academic projects have been dabbling in for years: Take CUDA source code and generate multithreaded SSE x86 code. If done well, I bet a lot of people would find that CUDA on x86 is faster than even their normal CPU implementations. (Because most compilers are terrible at generating SSE instructions from all but the simplest C code and most of us are terrible at writing SSE by hand.)

And it might make OpenCL a lot less attractive from a hybrid computing perspective. It will be very interesting to see the performance difference between OpenCL on multicore processors and CUDA code compiled with this compiler, and also the amount of tweaking required for both.

---

<div class="post-metadata">

**Author:** ![E.D\_Riedijk](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@E.D\_Riedijk](https://forums.developer.nvidia.com/u/E.D_Riedijk)\
**Post date:** [September 23, 2010, 3:11pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/19 "2010-09-23T15:11:28Z")

</div>

> [@](#):
>
> I agree with E.D. Riedijk. This is almost certainly a commercially supported compiler that does what many other academic projects have been dabbling in for years: Take CUDA source code and generate multithreaded SSE x86 code. If done well, I bet a lot of people would find that CUDA on x86 is faster than even their normal CPU implementations. (Because most compilers are terrible at generating SSE instructions from all but the simplest C code and most of us are terrible at writing SSE by hand.)

And it might make OpenCL a lot less attractive from a hybrid computing perspective. It will be very interesting to see the performance difference between OpenCL on multicore processors and CUDA code compiled with this compiler, and also the amount of tweaking required for both.

---

<div class="post-metadata">

**Author:** ![SPWorley](https://developer.download.nvidia.com/images/forums/profile-default-devtalk-84.png) [@SPWorley](https://forums.developer.nvidia.com/u/SPWorley)\
**Post date:** [September 23, 2010, 3:54pm UTC](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897/20 "2010-09-23T15:54:42Z")

</div>

I wonder if it uses a warp size of 4 or of 32.

I also hope it’s very CPU locality aware, trying to keep threads from the same block running on the same physical CPU to improve cache coherence. That gets tricky when you’re creating and destroying new blocks all the time.

[Next page](https://forums.developer.nvidia.com/t/most-iportant-from-gtc-cuda-on-x86-hello-emulation-mode/18897.md?page=2)
