<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Can an MPI task be moved from a CPU to a GPU in oneAPI Fortran in Intel® Moderncode for Parallel Architectures</title>
    <link>https://community.intel.com/t5/Intel-Moderncode-for-Parallel/Can-an-MPI-task-be-moved-from-a-CPU-to-a-GPU-in-oneAPI-Fortran/m-p/1345219#M8179</link>
    <description>&lt;P&gt;Hi,&lt;/P&gt;
&lt;P&gt;I am presently running Fortran with MPI under WIN10 to calculate typically 5000 matrix elements using 60 threads during each iteration step. I can't use OpenMP! Will it be possible, in the new version of Fortran, to move this calculation into a GPU with, lets say, 5000 cores. This should mean that each core could calculate one matrix element. As I understand it this should require that each core in the GPU has a reserved memory area or am I wrong?&amp;nbsp; The amount of data allocated to each core for the calculations is rather small. Will the memory data transfer to the GPU and/or the MPI overhead to collect the calculated matrix elements kill this use of the GPU?&lt;/P&gt;
&lt;P&gt;Best regards&lt;/P&gt;
&lt;P&gt;Anders S&lt;/P&gt;</description>
    <pubDate>Sun, 19 Dec 2021 17:07:45 GMT</pubDate>
    <dc:creator>Anders_S_1</dc:creator>
    <dc:date>2021-12-19T17:07:45Z</dc:date>
    <item>
      <title>Can an MPI task be moved from a CPU to a GPU in oneAPI Fortran</title>
      <link>https://community.intel.com/t5/Intel-Moderncode-for-Parallel/Can-an-MPI-task-be-moved-from-a-CPU-to-a-GPU-in-oneAPI-Fortran/m-p/1345219#M8179</link>
      <description>&lt;P&gt;Hi,&lt;/P&gt;
&lt;P&gt;I am presently running Fortran with MPI under WIN10 to calculate typically 5000 matrix elements using 60 threads during each iteration step. I can't use OpenMP! Will it be possible, in the new version of Fortran, to move this calculation into a GPU with, lets say, 5000 cores. This should mean that each core could calculate one matrix element. As I understand it this should require that each core in the GPU has a reserved memory area or am I wrong?&amp;nbsp; The amount of data allocated to each core for the calculations is rather small. Will the memory data transfer to the GPU and/or the MPI overhead to collect the calculated matrix elements kill this use of the GPU?&lt;/P&gt;
&lt;P&gt;Best regards&lt;/P&gt;
&lt;P&gt;Anders S&lt;/P&gt;</description>
      <pubDate>Sun, 19 Dec 2021 17:07:45 GMT</pubDate>
      <guid>https://community.intel.com/t5/Intel-Moderncode-for-Parallel/Can-an-MPI-task-be-moved-from-a-CPU-to-a-GPU-in-oneAPI-Fortran/m-p/1345219#M8179</guid>
      <dc:creator>Anders_S_1</dc:creator>
      <dc:date>2021-12-19T17:07:45Z</dc:date>
    </item>
  </channel>
</rss>

