openmpi

Автор	SHA1	Сообщение	Дата
Gilles Gouaillardet	35e7d86eb1	configury: make build Reproducible If defined, use SOURCE_DATE_EPOCH environment variable; make the build Reproducible by forcing timestamps. See https://reproducible-builds.org/docs/source-date-epoch/ for more information. Thanks Bernhard M. Wiedemann for bringing this to our attention. Fixes open-mpi/ompi#3759 NOTE: This was cherry-picked from master, and slightly modified / amended for the v4.1.x branch. Signed-off-by: Gilles Gouaillardet <gilles@rist.or.jp> Signed-off-by: Bernhard M. Wiedemann <bwiedemann@suse.de> Signed-off-by: Jeff Squyres <jsquyres@cisco.com> (cherry picked from commit `7b4e8ba4aa`)	2020-10-29 08:17:26 -07:00
Jeff Squyres	96ba8b7279	Merge pull request #8142 from AboorvaDevarajan/fix_zero_byte_v4.1.x [v4.1.x] pml/ucx: fix zero sized datatype transfers	2020-10-29 11:05:43 -04:00
George Bosilca	9d4e3b1649	Fix HAN issues reported by Coverity. Signed-off-by: George Bosilca <bosilca@icl.utk.edu>	2020-10-28 18:43:54 -04:00
Aboorva Devarajan	3518ebf94f	pml/ucx: fix zero sized datatype transfers Signed-off-by: Aboorva Devarajan <aburvadevarajan@gmail.com> (cherry picked from commit `202b81d95c`)	2020-10-27 10:38:07 -04:00
George Bosilca	6d735ba052	A complete overhaul of the HAN code. Among many other things: - Fix an imbalance bug in MPI_allgather - Accept more human readable configuration files. We can now specify the collective by name instead of a magic number, and the component we want to use also by name. - Add the capability to have optional arguments in the collective communication configuration file. Right now the capability exists for segment lengths, but is yet to be connected with the algorithms. - Redo the initialization of all HAN collectives. Cleanup the fallback collective support. - In case the module is unable to deliver the expected result, it will fallback executing the collective operation on another collective component. This change make the support for this fallback simpler to use. - Implement a fallback allowing a HAN module to remove itself as potential active collective module, and instead fallback to the next module in line. - Completely disable the HAN modules on error. From the moment an error is encountered they remove themselves from the communicator, and in case some other modules calls them simply behave as a pass-through. Communicator: provide ompi_comm_split_with_info to split and provide info at the same time Add ompi_comm_coll_preference info key to control collective component selection COLL HAN: use info keys instead of component-level variable to communicate topology level between abstraction layers - The info value is a comma-separated list of entries, which are chosen with decreasing priorities. This overrides the priority of the component, unless the component has disqualified itself. An entry prefixed with ^ starts the ignore-list. Any entry following this character will be ingnored during the collective component selection for the communicator. Example: "sm,libnbc,^han,adapt" gives sm the highest preference, followed by libnbc. The components han and adapt are ignored in the selection process. - Allocate a temporary buffer for all lower-level leaders (length 2 segments) - Fix the handling of MPI_IN_PLACE for gather and scatter. COLL HAN: Fix topology handling - HAN should not rely on node names to determine the ordering of ranks. Instead, use the node leaders as identifiers and short-cut if the node-leaders agree that ranks are consecutive. Also, error out if the rank distribution is imbalanced for now. Signed-off-by: Xi Luo <xluo12@vols.utk.edu> Signed-off-by: Joseph Schuchart <schuchart@icl.utk.edu> Signed-off-by: George Bosilca <bosilca@icl.utk.edu> Conflicts: ompi/mca/coll/adapt/coll_adapt_ibcast.c	2020-10-26 21:38:00 -04:00
bsergentm	94c817ceff	Coll/han Bull * first import of Bull specific modifications to HAN * Cleaning, renaming and compilation fixing Changed all future into han. * Import BULL specific modifications in coll/tuned and coll/base * Fixed compilation issues in Han * Changed han_output to directly point to coll framework output. * The verbosity MCA parameter was removed as a duplicated of coll verbosity * Add fallback in han reduce when op cannot commute and ppn are imbalanced * Added fallback wfor han bcast when nodes do not have the same number of process * Add fallback in han scatter when ppn are imbalanced + fixed missing scatter_fn pointer in the module interface Signed-off-by: Brelle Emmanuel <emmanuel.brelle@atos.net> Co-authored-by: a700850 <pierre.lemarinier@atos.net> Co-authored-by: germainf <florent.germain@atos.net>	2020-10-26 21:35:12 -04:00
Xi Luo	feb5a7113c	Initial import of the HAN collective module a hierarchical, architecture-aware collective communication module. Add Reduce and remove up_seg_size and low_seg_size in Bcast Increase HAN's priority Signed-off-by: Xi Luo <xluo12@vols.utk.edu> Signed-off-by: George Bosilca <bosilca@icl.utk.edu>	2020-10-26 21:35:12 -04:00
Raghu Raja	c55d3d3469	mtl/ofi: Fix erroneous FI_PEEK/FI_CLAIM usage The current iprobe/improbe implementations merely checks the return code on the posted receive operation to tell if there is a match or not. This commit moves the check to the probe's error callback instead. Per the semantics defined in libfabric, the peek operation is asynchronous and the results are to be fetched from the completion queue. If no message is found matching the tags specified in the peek request, then a completion queue error entry with err field set to FI_ENOMSG will be available. Signed-off-by: Raghu Raja <craghun@amazon.com> (cherry picked from commit `39f8a86b65`)	2020-10-26 17:48:21 +00:00
Raghu Raja	b41680783f	mtl/ofi: Do not fail if error CQ is empty In multi-threaded scenarios, any thread that attempts to read a CQ when there's a pending error CQ entry gets an -FI_EAVAIL. Without any serialization here (which is okay, since libfabric will protect access to critical CQ objects), all threads proceed to read from the error CQ, but only one thread fetches the entry while others get -FI_EAGAIN indicating an empty queue, which is not erroneous. Signed-off-by: Raghu Raja <craghun@amazon.com> (cherry picked from commit `415dddb9af`)	2020-10-22 18:57:12 +00:00
Mark Allen	994186fa00	symbol pollution Made a couple vars static if they didn't look like they were used more than one place, and added prefixes to a few. Signed-off-by: Mark Allen <markalle@us.ibm.com> (cherry picked from commit `34b42b1b09`)	2020-10-07 14:07:55 -04:00
Jeff Squyres	edf03e52f3	Merge pull request #7944 from bosilca/4.1/adapt Import the ADAPT collective into the 4.1	2020-09-23 16:42:49 -04:00
Xi Luo	e65fa4ff5c	Bring ADAPT collective to 4.1 This is a meta commit, that encapsulate all the ADAPT commits in the master into a single PR for 4.1. The master commits included here are: `fe73586`, `a4be3bb`, `d712645`, `c2970a3`, `e59bde9`, `ee592f3` and `c98e387`. Here is a detailed list of added capabilities: * coll/adapt: Fix naming conventions and C11 atomic use * coll/adapt: Remove unused component field in module * Consistent handling of zero counts in the MPI API. * Correctly handle non-blocking collectives tags * As it is possible to have multiple outstanding non-blocking collectives provided by different collective modules, we need a consistent mechanism to allow them to select unique tags for each instance of a collective. * Add support for fallback to previous coll module on non-commutative operations (#30) * Replace mutexes by atomic operations. * Use the correct nbc request type (for both ibcast and ireduce) * coll/base: document type casts in ompi_coll_base_retain_* * add module-wide topology cache * use standard instead of synchronous send and add mca parameter to control mode of initial send in ireduce/ibcast * reduce number of memory allocations * call the default request completion. * Remove the requests from the Fortran lookup conversion tables before completing and free it. * piggybacking Bull functionalities Signed-off-by: Xi Luo <xluo12@vols.utk.edu> Signed-off-by: George Bosilca <bosilca@icl.utk.edu> Signed-off-by: Marc Sergent <marc.sergent@atos.net> Co-authored-by: Joseph Schuchart <schuchart@hlrs.de> Co-authored-by: Lemarinier, Pierre <pierre.lemarinier@atos.net> Co-authored-by: pierrele <31764860+pierrele@users.noreply.github.com>	2020-09-23 11:45:45 -04:00
William Zhang	7922430c66	coll/tuned: Fix dynamic message size for gather and scatter The gather and scatter operations did not use the correct message size (Only did datatype size * com size). This did not correctly reflect the total message size and prevents fine tuning within a com size. This patch multiplies the value by the number of elements sent. Signed-off-by: William Zhang <wilzhang@amazon.com> (cherry picked from commit `50823fe9a9`)	2020-09-15 08:56:56 -07:00
Howard Pritchard	fae374ce54	OFI: patch OFI MTL for GNI provider Uncovered a problem using the GNI provider with the OFI MTL. See https://github.com/ofiwg/libfabric/issues/6194. Related to #8001 Signed-off-by: Howard Pritchard <hppritcha@gmail.com> (cherry picked from commit `d6ac41cbbd`)	2020-08-28 14:30:31 -06:00
William Zhang	dceea5ad87	coll/tuned: Revert RSB and RS default algorithms Reduce scatter block and reduce scatter algorithms were hitting correctness issues for non commutative strided tests. We will revert to the original default algorithms for those two collectives (basic linear and non overlapping respectively) in the non commutative op case. See #8010 Signed-off-by: William Zhang <wilzhang@amazon.com> (cherry picked from commit `57b95bcb45`)	2020-08-25 15:48:45 -07:00
Howard Pritchard	833f8b2b41	ofi mtl: fix problem with mrecv the ofi mtl mrecv was not properly setting the message in/out arg to MPI_MRECV to MPI_MESSAGE_NULL. Signed-off-by: Howard Pritchard <hppritcha@gmail.com> (cherry picked from commit `e6f81ed6d6`)	2020-08-19 10:22:37 -06:00
George Bosilca	eb9ced786a	Use the unaligned SSE memory access primitive. Alter the test to validate misaligned data. Fixes #7954. Signed-off-by: George Bosilca <bosilca@icl.utk.edu> (cherry picked from commit `b6d71aa893`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-08-12 02:32:01 +00:00
William Zhang	abe3aaa4a7	btl/ofi: Use common provider include/exclude list The btl/ofi does not currently utilize the common ofi include/exclude list. Added verification code similar to the mtl/ofi that will check if the info object is in the include or exclude list. If it isn't in the include list or is in the exclude list, validate_info will return OPAL_ERROR. The btl/ofi will no longer pass a provider name as a hint when calling getinfo, instead filtering the provider during validate_info. This patch also moves the is_in_list MTL function into common code and adds additional debugging output to the BTL to match the MTL standard. Signed-off-by: William Zhang <wilzhang@amazon.com> (cherry picked from commit `9b8f463a76`)	2020-08-10 14:28:59 -07:00
William Zhang	5a13c5352f	coll/tuned: Change the default collective algorithm selection The default algorithm selections were out of date and not performing well. After gathering data from OMPI developers, new default algorithm decisions were selected for: allgather allgatherv allreduce alltoall alltoallv barrier bcast gather reduce reduce_scatter_block reduce_scatter scatter These results were gathered using the ompi-collectives-tuning package and then averaged amongst the results gathered from multiple OMPI developers on their clusters. You can access the graphs and averaged data here: https://drive.google.com/drive/folders/1MV5E9gN-5tootoWoh62aoXmN0jiWiqh3 Signed-off-by: William Zhang <wilzhang@amazon.com> (cherry picked from commit `ce40cfbaa5`)	2020-07-28 15:48:20 -07:00
Jeff Squyres	36bcc48dc1	Merge pull request #7902 from vspetrov/v4.1.x_hcoll_reduce_scatter V4.1.x hcoll reduce scatter	2020-07-27 15:50:09 -04:00
Joseph Schuchart	3d08d790e9	osc/rdma: fail query_btls if no endpoint for non-local peer is found Signed-off-by: Joseph Schuchart <schuchart@hlrs.de> (cherry picked from commit `eebc451ec8`)	2020-07-17 08:38:37 +02:00
dongzhong	b4e04bbd8a	Add supports for MPI_OP using AVX512, AVX2 and MMX Add logic to handle different architectural capabilities Detect the compiler flags necessary to build specialized versions of the MPI_OP. Once the different flavors (AVX512, AVX2, AVX) are built, detect at runtime which is the best match with the current processor capabilities. Add validation checks for loadu 256 and 512 bits. Add validation tests for MPI_Op. Signed-off-by: Jeff Squyres <jsquyres@cisco.com> Signed-off-by: Gilles Gouaillardet <gilles@rist.or.jp> Signed-off-by: dongzhong <zhongdong0321@hotmail.com> Signed-off-by: George Bosilca <bosilca@icl.utk.edu> (cherry picked from commit `14b3c70628`)	2020-07-13 13:49:00 -07:00
Gilles Gouaillardet	79a737ca94	mpi/c: fix param checks in [I]Neighbor_alltoall{v,w} do not check some input parameters when an {in,out}degree is zero Thanks Junchao Zhang for analyzing and reporting this issue. Signed-off-by: Gilles Gouaillardet <gilles@rist.or.jp> (cherry picked from commit `5655d64bd3`)	2020-07-10 16:58:52 -06:00
Jeff Squyres	d17685a6c9	Merge pull request #7890 from tkordenbrock/topic/v4.1.x/portals4.call-pml-add_procs v4.1.x: mtl-portals4: use the active PML to call add_procs()	2020-07-06 07:34:41 -04:00
Jeff Squyres	a878569386	Merge pull request #7895 from tkordenbrock/topic/v4.1.x/portals4.fix-inappropriate-use-of-abort v4.1.x: portals4: fix inappropriate use of abort() in mtl-portals4 and coll-portals4 components	2020-07-06 07:32:32 -04:00
Valentin Petrov	6f401186f7	coll/hcoll: compile warning fix Signed-off-by: Valentin Petrov <valentinp@mellanox.com>	2020-07-02 08:44:25 +03:00
Valentin Petrov	2441fb2baf	coll/hcoll: reduce_scatter(block) interface Signed-off-by: Valentin Petrov <valentinp@mellanox.com>	2020-07-02 08:44:22 +03:00
Jeff Squyres	bc6587d3fa	Merge pull request #7873 from devreal/osc-ucx-rget-rput-fetch-alignment-v4.1.x OSC UCX: make sure no-op fetch in rget/rput is properly aligned (v4.1.x)	2020-06-29 15:20:12 -04:00
Jeff Squyres	249c57a4bc	Merge pull request #7889 from devreal/osc-rdma-noncontig-requests-v4.1.x osc rdma: check for outstanding fragments before completing a request (II) (v4.1.x)	2020-06-29 15:19:49 -04:00
Todd Kordenbrock	20f9ed98f2	mtl-portals4: replace abort() with ompi_rte_abort() coll-portals4: replace abort() with ompi_rte_abort() Signed-off-by: Todd Kordenbrock <thkgcode@gmail.com> (cherry picked from commit `04b94637dd`)	2020-06-29 10:06:12 -05:00
Todd Kordenbrock	540b14fc32	Use the active PML to call add_procs() ompi_mtl_portals4_get_endpoint() was incorrectly making a direct call to ompi_mtl_portals4_add_procs(). Instead use the actve PML to call add_procs(). If add_procs() fails, call ompi_rte_abort() to terminate the job. Signed-off-by: Todd Kordenbrock <thkgcode@gmail.com> (cherry picked from commit `0a637967fa`)	2020-06-29 09:55:23 -05:00
Joseph Schuchart	2d3f862f1d	osc rdma: check for outstanding fragments before completing a request in ompi_osc_rdma_put_complete_flush as well Signed-off-by: Joseph Schuchart <schuchart@hlrs.de> (cherry picked from commit `caed3b2eed`)	2020-06-29 15:44:44 +02:00
Brian Barrett	a6d97e2d6d	Merge pull request #7815 from raafatfeki/topic/ime Topic/ime: Bring over IME component from master to v4.1	2020-06-26 12:40:35 -07:00
raafatfeki	6e145188d9	fs/ime & fbtl/ime: Support of IME file system Signed-off-by: raafatfeki <fekiraafat@gmail.com>	2020-06-26 12:26:51 -04:00
William Zhang	db6ed187b2	coll/tuned: Add NULL check to prevent segfault Signed-off-by: William Zhang <wilzhang@amazon.com> cr https://code.amazon.com/reviews/CR-23837553 (cherry picked from commit `771f9c011d`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
William Zhang	03758b1ef7	coll/tuned: Fix typos Signed-off-by: William Zhang <wilzhang@amazon.com> (cherry picked from commit `50640402ab`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Brinskii	7eb94164a0	COLL/TUNED: Add linear scatter using isend for mlnx platform Signed-off-by: Mikhail Brinskii <mikhailb@mellanox.com> (cherry picked from commit `f2cbd4806e`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Gilles Gouaillardet	221fad6862	coll/cuda: remove unnecessary references to ORTE Signed-off-by: Gilles Gouaillardet <gilles@rist.or.jp> (cherry picked from commit `531171ca50`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Tomislav Janjusic	f51bd8ca0c	Coll/hcoll: adding scatterv interface Signed-off-by: Valentin Petrov valentinp@mellanox.com (cherry picked from commit `6ea920e225`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Alex Anenkov	2891a23329	coll/libnbc: add recursive doubling algorithm for MPI_Iallreduce Signed-off-by: Alex Anenkov <anenkov.ru@gmail.com> (cherry picked from commit `77d466edf3`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Kurnosov	ba11f31fc8	coll/libnbc: remove debug output 1. Remove debug output in iallgather (I have forgotten to remove it). 2. Remove an incorrect comment in description of ibcast Signed-off-by: Mikhail Kurnosov <mkurnosov@gmail.com> (cherry picked from commit `64abd0f405`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Kurnosov	bf1c8bb394	coll/libnbc/ireduce: silence Coverity warning CID 1440360 Signed-off-by: Mikhail Kurnosov <mkurnosov@gmail.com> (cherry picked from commit `8b511c7889`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Kurnosov	5ee1fb62b9	coll/libnbc: add Rabenseifner's algorithm for MPI_Iallreduce An implementation of R. Rabenseifner's algorithm for MPI_Iallreduce. This algorithm is a combination of a reduce-scatter implemented with recursive vector halving and recursive distance doubling, followed either by an allgather. Limitations: -- count >= 2^{\floor{\log_2 p}} -- commutative operations only -- intra-communicators only Signed-off-by: Mikhail Kurnosov <mkurnosov@gmail.com> (cherry picked from commit `73e048b62a`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
George Bosilca	fd29cce114	Remove few warnings in libnbc identified by clang-1000.11.45.2 Signed-off-by: George Bosilca <bosilca@icl.utk.edu> (cherry picked from commit `66182a294d`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Kurnosov	91a4b4c799	coll/libnbc: add recursive doubling algorithm for MPI_Iallgather Implements recursive doubling algorithm for MPI_Iallgather. The algorithm can be used only for power-of-two number of processes. Signed-off-by: Mikhail Kurnosov <mkurnosov@gmail.com> (cherry picked from commit `a7386c1e09`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Kurnosov	6971dab943	coll/libnbc: add knomial tree algorithm for MPI_Ibcast Signed-off-by: Mikhail Kurnosov <mkurnosov@gmail.com> (cherry picked from commit `b0429d25df`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Kurnosov	a318f117f6	coll/libnbc: add Rabenseifner's algorithm for MPI_Ireduce An implementation of R. Rabenseifner's algorithm for MPI_Ireduce. This algorithm is a combination of a reduce-scatter implemented with recursive vector halving and recursive distance doubling, followed either by a gather. Limitations: -- count >= 2^{\floor{\log_2 p}} -- commutative operations only -- intra-communicators only Signed-off-by: Mikhail Kurnosov <mkurnosov@gmail.com> (cherry picked from commit `7bd63e79c8`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Brian Barrett	6f6d8180a3	coll libnbc: Remove dead code Remove dead code that was causing warnings about unused static functions. Signed-off-by: Brian Barrett <bbarrett@amazon.com> (cherry picked from commit `2e24e6ec08`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Kurnosov	de5e435dee	coll/libnbc: add recursive doubling algorithm for MPI_Iexscan Implements recursive doubling algorithm for MPI_Iexscan. The algorithm preserves order of operations so it can be used both by commutative and non-commutative operations. The MCA parameter 'coll_libnbc_iexscan_algorithm' was added for dynamic algorithm selection. Signed-off-by: Mikhail Kurnosov <mkurnosov@gmail.com> (cherry picked from commit `dfe203e167`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00
Mikhail Kurnosov	65990af3ad	coll/libnbc: add recursive doubling algorithm for MPI_Iscan Implements recursive doubling algorithm for MPI_Iscan. The algorithm preserves order of operations so it can be used both by commutative and non-commutative operations. The MCA parameter coll_libnbc_iscan_algorithm was added for dynamic algorithm selection. Signed-off-by: Mikhail Kurnosov <mkurnosov@gmail.com> (cherry picked from commit `3d43ff0f32`) Signed-off-by: Brian Barrett <bbarrett@amazon.com>	2020-06-25 23:06:51 +00:00

1 2 3 4 5 ...

10467 Коммитов