<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Problems running scan SDK examples from NVidia in OpenCL* for CPU</title>
    <link>https://community.intel.com/t5/OpenCL-for-CPU/Problems-running-scan-SDK-examples-from-NVidia/m-p/765629#M66</link>
    <description>Have just tried to inline the code manually which made it work!</description>
    <pubDate>Mon, 11 Apr 2011 09:05:15 GMT</pubDate>
    <dc:creator>jrimestad</dc:creator>
    <dc:date>2011-04-11T09:05:15Z</dc:date>
    <item>
      <title>Problems running scan SDK examples from NVidia</title>
      <link>https://community.intel.com/t5/OpenCL-for-CPU/Problems-running-scan-SDK-examples-from-NVidia/m-p/765628#M65</link>
      <description>Hi&lt;DIV&gt;&lt;/DIV&gt;&lt;DIV&gt;I have tried running NVidia's OpenCL example of scan (prefixsum) on an Intel CPU but without luck.&lt;/DIV&gt;&lt;DIV&gt;&lt;/DIV&gt;&lt;DIV&gt;The program fails with a memory error.&lt;/DIV&gt;&lt;DIV&gt;&lt;/DIV&gt;&lt;DIV&gt;If I remove the barrier synchronization lines in the following part of the code it runs without errors, but the results are offcause wrong:&lt;/DIV&gt;&lt;DIV&gt;&lt;PRE&gt;[cpp]inline uint scan1Inclusive(uint idata, __local uint *l_Data, uint size){
	uint pos = 2 * get_local_id(0) - (get_local_id(0) &amp;amp; (size - 1));
    l_Data[pos] = 0;
    pos += size;
    l_Data[pos] = idata;

	for(uint offset = 1; offset &amp;lt; size; offset &amp;lt;&amp;lt;= 1){
		barrier(CLK_LOCAL_MEM_FENCE); //Fails with Intel openCL
        uint t = l_Data[pos] + l_Data[pos - offset];
		barrier(CLK_LOCAL_MEM_FENCE); //Fails with Intel openCL
		l_Data[pos] = t;
	}
	return l_Data[pos];
}[/cpp]&lt;/PRE&gt; &lt;/DIV&gt;&lt;DIV&gt;I have tested other code using barriers that worked fine. The main difference is that this example is using a three layer nesting of inlining. The inserted code is the innermost function.&lt;/DIV&gt;&lt;DIV&gt;&lt;/DIV&gt;&lt;DIV&gt;The code runs fine on a NVidia graphics card and on CPU compiled with the AMD openCL compiler.&lt;/DIV&gt;&lt;DIV&gt;&lt;/DIV&gt;&lt;DIV&gt;-Jens&lt;/DIV&gt;&lt;DIV&gt;&lt;/DIV&gt;</description>
      <pubDate>Mon, 11 Apr 2011 07:26:45 GMT</pubDate>
      <guid>https://community.intel.com/t5/OpenCL-for-CPU/Problems-running-scan-SDK-examples-from-NVidia/m-p/765628#M65</guid>
      <dc:creator>jrimestad</dc:creator>
      <dc:date>2011-04-11T07:26:45Z</dc:date>
    </item>
    <item>
      <title>Problems running scan SDK examples from NVidia</title>
      <link>https://community.intel.com/t5/OpenCL-for-CPU/Problems-running-scan-SDK-examples-from-NVidia/m-p/765629#M66</link>
      <description>Have just tried to inline the code manually which made it work!</description>
      <pubDate>Mon, 11 Apr 2011 09:05:15 GMT</pubDate>
      <guid>https://community.intel.com/t5/OpenCL-for-CPU/Problems-running-scan-SDK-examples-from-NVidia/m-p/765629#M66</guid>
      <dc:creator>jrimestad</dc:creator>
      <dc:date>2011-04-11T09:05:15Z</dc:date>
    </item>
    <item>
      <title>Problems running scan SDK examples from NVidia</title>
      <link>https://community.intel.com/t5/OpenCL-for-CPU/Problems-running-scan-SDK-examples-from-NVidia/m-p/765630#M67</link>
      <description>Hello jrimestad,

Thanks for the report. We're working on reproducing and fixing the problem.</description>
      <pubDate>Mon, 11 Apr 2011 10:00:40 GMT</pubDate>
      <guid>https://community.intel.com/t5/OpenCL-for-CPU/Problems-running-scan-SDK-examples-from-NVidia/m-p/765630#M67</guid>
      <dc:creator>Eli_Bendersky__Intel</dc:creator>
      <dc:date>2011-04-11T10:00:40Z</dc:date>
    </item>
  </channel>
</rss>

