<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Artem, in Intel® MPI Library</title>
    <link>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009992#M3887</link>
    <description>&lt;P&gt;Artem,&lt;/P&gt;

&lt;P&gt;Absolutely, thank you for the response.&lt;/P&gt;

&lt;P&gt;I ran the tests with the following IMPI variables&lt;SPAN style="font-size: 1em; line-height: 1.5;"&gt;:&lt;/SPAN&gt;&lt;/P&gt;

&lt;P&gt;I_MPI_FABRICS=shm:ofa&lt;SPAN style="font-size: 1em; line-height: 1.5;"&gt;&amp;nbsp;&lt;/SPAN&gt;&lt;/P&gt;

&lt;P&gt;It was submitted through SLURM scheduling w&lt;SPAN style="font-size: 1em; line-height: 1.5;"&gt;ith the following batch script:&lt;/SPAN&gt;&lt;/P&gt;

&lt;P&gt;#!/bin/bash&lt;BR /&gt;
	#SBATCH --job-name=HPCGeval&lt;BR /&gt;
	#SBATCH --output=slurm.out&lt;BR /&gt;
	#SBATCH --error=slurm.err&lt;BR /&gt;
	#SBATCH --partition=batch&lt;BR /&gt;
	#SBATCH --time=02:00:00&lt;BR /&gt;
	#SBATCH --account=support&lt;BR /&gt;
	#SBATCH --nodes=64&lt;BR /&gt;
	#SBATCH --ntasks-per-node=16&lt;BR /&gt;
	#SBATCH --cpus-per-task=1&lt;BR /&gt;
	#SBATCH --exclusive&lt;BR /&gt;
	#SBATCH --constraint=hpcf2013&lt;/P&gt;

&lt;P&gt;export KMP_AFFINITY=compact&lt;BR /&gt;
	export OMP_NUM_THREADS=1&lt;BR /&gt;
	srun &amp;nbsp;../../xhpcg &amp;gt; /dev/null&lt;/P&gt;

&lt;P&gt;OS: Red Hat release 6.6 (Santiago), OFED:&amp;nbsp;OFED-1.5.4.1&lt;/P&gt;

&lt;P&gt;All tests were run on 64 total nodes, with two Intel E5-2650v2 CPUs (16 total cores) per node, linked with QLogic Corp. IBA7322 Infiniband HCA (rev 02) cards connected to a QLogic 12800-180 switch.&lt;/P&gt;

&lt;P&gt;We rely on another company to handle our licenses and updates with Intel-MPI, although I believe that we will be upgrading to Intel MPI Library v5.x soon.&lt;/P&gt;

&lt;P&gt;Best,&lt;/P&gt;

&lt;P&gt;Jack&lt;/P&gt;

&lt;P&gt;&amp;nbsp;&lt;/P&gt;</description>
    <pubDate>Fri, 22 May 2015 22:06:52 GMT</pubDate>
    <dc:creator>Jack_S_2</dc:creator>
    <dc:date>2015-05-22T22:06:52Z</dc:date>
    <item>
      <title>Problem with Intel MPI on &gt;1023 processes</title>
      <link>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009990#M3885</link>
      <description>&lt;P&gt;I have been testing code using Intel MPI (version 4.1.3 &amp;nbsp;build 20140226) and the Intel compiler (version 15.0.1 build 20141023) with 1024 or more total processes. When we attempt to run on 1024 or more processes we receive the following error:&amp;nbsp;&lt;/P&gt;

&lt;P&gt;MPI startup(): ofa fabric is not available and fallback fabric is not enabled&amp;nbsp;&lt;/P&gt;

&lt;P&gt;Anything less than 1024 processes does not produce this error, and I also do not receive this error with 1024 processes using OpenMPI and GCC.&lt;/P&gt;

&lt;P&gt;I am using the High Performance Conjugate Gradient benchmark as my test code, although we have received the same errors with other test codes.&amp;nbsp;&lt;/P&gt;</description>
      <pubDate>Wed, 20 May 2015 17:53:48 GMT</pubDate>
      <guid>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009990#M3885</guid>
      <dc:creator>Jack_S_2</dc:creator>
      <dc:date>2015-05-20T17:53:48Z</dc:date>
    </item>
    <item>
      <title>Hi Jack,</title>
      <link>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009991#M3886</link>
      <description>&lt;P&gt;Hi Jack,&lt;BR /&gt;
	&lt;BR /&gt;
	Could you please provide more details about your MPI runs (IMPI environment variables, command line options, OS/OFED versions, processor type, InfiniBand adapter name, number of involved hosts and so on)?&lt;BR /&gt;
	Are you able to run&amp;nbsp;with newer Intel MPI Library (5.x)?&lt;BR /&gt;
	&amp;nbsp;&lt;/P&gt;</description>
      <pubDate>Thu, 21 May 2015 06:44:10 GMT</pubDate>
      <guid>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009991#M3886</guid>
      <dc:creator>Artem_R_Intel1</dc:creator>
      <dc:date>2015-05-21T06:44:10Z</dc:date>
    </item>
    <item>
      <title>Artem,</title>
      <link>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009992#M3887</link>
      <description>&lt;P&gt;Artem,&lt;/P&gt;

&lt;P&gt;Absolutely, thank you for the response.&lt;/P&gt;

&lt;P&gt;I ran the tests with the following IMPI variables&lt;SPAN style="font-size: 1em; line-height: 1.5;"&gt;:&lt;/SPAN&gt;&lt;/P&gt;

&lt;P&gt;I_MPI_FABRICS=shm:ofa&lt;SPAN style="font-size: 1em; line-height: 1.5;"&gt;&amp;nbsp;&lt;/SPAN&gt;&lt;/P&gt;

&lt;P&gt;It was submitted through SLURM scheduling w&lt;SPAN style="font-size: 1em; line-height: 1.5;"&gt;ith the following batch script:&lt;/SPAN&gt;&lt;/P&gt;

&lt;P&gt;#!/bin/bash&lt;BR /&gt;
	#SBATCH --job-name=HPCGeval&lt;BR /&gt;
	#SBATCH --output=slurm.out&lt;BR /&gt;
	#SBATCH --error=slurm.err&lt;BR /&gt;
	#SBATCH --partition=batch&lt;BR /&gt;
	#SBATCH --time=02:00:00&lt;BR /&gt;
	#SBATCH --account=support&lt;BR /&gt;
	#SBATCH --nodes=64&lt;BR /&gt;
	#SBATCH --ntasks-per-node=16&lt;BR /&gt;
	#SBATCH --cpus-per-task=1&lt;BR /&gt;
	#SBATCH --exclusive&lt;BR /&gt;
	#SBATCH --constraint=hpcf2013&lt;/P&gt;

&lt;P&gt;export KMP_AFFINITY=compact&lt;BR /&gt;
	export OMP_NUM_THREADS=1&lt;BR /&gt;
	srun &amp;nbsp;../../xhpcg &amp;gt; /dev/null&lt;/P&gt;

&lt;P&gt;OS: Red Hat release 6.6 (Santiago), OFED:&amp;nbsp;OFED-1.5.4.1&lt;/P&gt;

&lt;P&gt;All tests were run on 64 total nodes, with two Intel E5-2650v2 CPUs (16 total cores) per node, linked with QLogic Corp. IBA7322 Infiniband HCA (rev 02) cards connected to a QLogic 12800-180 switch.&lt;/P&gt;

&lt;P&gt;We rely on another company to handle our licenses and updates with Intel-MPI, although I believe that we will be upgrading to Intel MPI Library v5.x soon.&lt;/P&gt;

&lt;P&gt;Best,&lt;/P&gt;

&lt;P&gt;Jack&lt;/P&gt;

&lt;P&gt;&amp;nbsp;&lt;/P&gt;</description>
      <pubDate>Fri, 22 May 2015 22:06:52 GMT</pubDate>
      <guid>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009992#M3887</guid>
      <dc:creator>Jack_S_2</dc:creator>
      <dc:date>2015-05-22T22:06:52Z</dc:date>
    </item>
    <item>
      <title>Hi Jack,</title>
      <link>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009993#M3888</link>
      <description>&lt;P&gt;Hi Jack,&lt;BR /&gt;
	&lt;BR /&gt;
	Thanks for the clarification.&lt;BR /&gt;
	As far as I see you use Intel True Scale (aka QLogic) IBAs, 'shm:ofa' may work nonoptimal on such IBAs.&lt;BR /&gt;
	You can use 'tmi/shm:tmi' fabric which is designed for such cases.&lt;BR /&gt;
	&lt;BR /&gt;
	Usage:&lt;BR /&gt;
	export I_MPI_FABRICS=shm:tmi&lt;BR /&gt;
	export TMI_CONFIG=&amp;lt;path_to_impi&amp;gt;/intel64/etc/tmi.conf&lt;BR /&gt;
	&amp;nbsp;&lt;/P&gt;</description>
      <pubDate>Mon, 25 May 2015 06:56:40 GMT</pubDate>
      <guid>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009993#M3888</guid>
      <dc:creator>Artem_R_Intel1</dc:creator>
      <dc:date>2015-05-25T06:56:40Z</dc:date>
    </item>
    <item>
      <title>Artem,</title>
      <link>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009994#M3889</link>
      <description>&lt;P&gt;Artem,&lt;/P&gt;

&lt;P&gt;Thank you so much for your help, this solved the issue we were having, as well as another issue that we were having!&lt;/P&gt;

&lt;P&gt;I'm just curious, do you have any idea why this problem only seemed to surface after going over 1023 processes?&lt;/P&gt;

&lt;P&gt;Best,&lt;/P&gt;

&lt;P&gt;Jack&lt;/P&gt;</description>
      <pubDate>Mon, 25 May 2015 15:39:50 GMT</pubDate>
      <guid>https://community.intel.com/t5/Intel-MPI-Library/Problem-with-Intel-MPI-on-gt-1023-processes/m-p/1009994#M3889</guid>
      <dc:creator>Jack_S_2</dc:creator>
      <dc:date>2015-05-25T15:39:50Z</dc:date>
    </item>
  </channel>
</rss>

