1
1

30592 Коммитов

Автор SHA1 Сообщение Дата
Jeff Squyres
6b5f67acb4
Merge pull request #7598 from alex-ross-dev/pr/readme-fix
README: Updated Tested Systems
2020-04-04 13:54:02 -04:00
Alex Ross
afb6cef9f2 README: Updated Tested Systems
Added clang version to linux system requirements.
Changed macOS version from 10.12 to 10.14-10.15.
Fixed a minor mistake where the architecture was labeled as x85_64
instead of x86_64

Signed-off-by: Alex Ross <rossaj16@wfu.edu>
2020-04-04 14:09:21 +00:00
Ralph Castain
d533cfd506
Merge pull request #7597 from rhc54/topic/fd
Update PRRTE
2020-04-03 19:07:50 -07:00
Ralph Castain
c3f2ac9eb6
Update PRRTE
Restore daemon debugging flags and prior default ranking policy

Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-04-03 17:55:07 -07:00
Jeff Squyres
9cd8d7d430
Merge pull request #7594 from rhc54/topic/slm
Cover the case of using mpirun from inside Slurm allocation
2020-04-03 16:46:02 -04:00
Ralph Castain
3fbfeabff2
Update PRRTE schizo framework
Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-04-03 11:37:19 -07:00
Josh Hursey
5580566d3d
Merge pull request #7595 from yzluka/README_edit
Updated the start date for the "abbrviated list of release note" section in the README file
2020-04-03 11:22:55 -05:00
Yixin Zhang
c497f79812 Updated the latest review date for the abbreviated list of release note in the README file
Signed-off-by: Yixin Zhang <zhany217@wfu.edu>
2020-04-03 10:58:23 -04:00
Yixin Zhang
3bcab0f3ba updated the start date for the abbrviated list of release note
Signed-off-by: Yixin Zhang <zhany217@wfu.edu>
2020-04-03 00:08:43 -04:00
Ralph Castain
7ed3f11fd5
Cover the case of using mpirun from inside Slurm allocation
Update PRRTE schizo components

Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-04-02 19:26:13 -07:00
Jeff Squyres
fc0f0b38fd
Merge pull request #7590 from jsquyres/pr/update-to-https
Update text references to HTTPS
2020-04-02 20:46:58 -04:00
Jeff Squyres
4652ba5518
Merge pull request #7593 from jsquyres/pr/fix-fortran-preprocessor-configure-check
Fortran: fix the F90 compiler preprocessor check
2020-04-02 19:50:19 -04:00
Jeff Squyres
a7e4ca4dc0 Fortran: fix the F90 compiler preprocessor check
Only check the if the Fortran compiler needs additional CLI flags for
preprocessing .F90 files if we actually have an F90 compiler.

Also fix a the AC_MSG_* usage.

Signed-off-by: Jeff Squyres <jsquyres@cisco.com>
2020-04-02 16:12:09 -07:00
Ralph Castain
7d31a99ef8
Merge pull request #7589 from rhc54/topic/env
Update PMIx and PRRTE, plus PRRTE config integration
2020-04-02 15:55:21 -07:00
Ralph Castain
44c97e3842
Update again and hope to fix the integration points
Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-04-02 11:58:58 -07:00
Josh Hursey
7907f5e3dc
Merge pull request #7591 from paulkefer/fix-readme
README: Remove SunCC bullet
2020-04-02 13:33:00 -05:00
Paul Kefer
48cad13e07 README: Remove SunCC bullet
This bullet can be removed since SunCC is no longer supported as of
version 5.0.0.

Signed-off-by: Paul Kefer <24892969+paulkefer@users.noreply.github.com>
2020-04-02 13:25:32 -04:00
Jeff Squyres
9687d5e867 Upgrade all www.open-mpi.org URLs to https
Found a handful of other URLs that weren't https-ized, so I updated
them, too (after verifying that they support https, of course).

Signed-off-by: Jeff Squyres <jsquyres@cisco.com>
2020-04-02 10:43:50 -04:00
Jeff Squyres
81fec3ec84 README: Fix URL for knem
Signed-off-by: Jeff Squyres <jsquyres@cisco.com>
2020-04-02 10:41:18 -04:00
Ralph Castain
8615761470
Incremental step forward on configury for PRRTE
Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-04-02 07:40:46 -07:00
Ralph Castain
50d05e7b64
Revert "Add extra libs to PRRTE binaries for external deps"
This reverts commit 1aabbe456d5de8fc3c3d6f91eabb8851db7d63eb.

Update PMIx and PRRTE, plus PRRTE config integration

Cleanup how we pass the extra libs and LDFLAGS for linking against
external libevent, hwloc, and pmix installs.

Catch the flag indicating that PMIx provided the user-level default MCA
params so we don't go looking for them ourselves.

Cleanup misc config warnings

Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-04-01 20:17:32 -07:00
Jeff Squyres
92b58987d6
Merge pull request #7584 from ggouaillardet/topic/configury_fpp
configury: try if -fpp flag is needed to preprocess .F90 files
2020-04-01 12:53:57 -04:00
Jeff Squyres
4c89999266
Merge pull request #7587 from jsquyres/pr/resolve-prrte-and-iquote-conflict
configure: fix PRRTE conflict with -iquote
2020-04-01 12:38:07 -04:00
Jeff Squyres
a0e34785e3 configure: fix PRRTE conflict with -iquote
PRRTE needs hwloc and libevent, so it needs to be setup "late" in
configure.ac.  However, we don't want to do it at the absolute bottom
of configure.ac, because right near the bottom, we setup CPPFLAGS (and
others) with values that are expected to be used only in
Makefile[.am]'s -- i.e., "$(foo)" values.  Such values are not able to
be used here in configure.

Hence, move the PRRTE setup up above where we do these "final"/
only-relevant-to-Makefile[.am] CPPFLAGS (etc.) updates occur.

Signed-off-by: Jeff Squyres <jsquyres@cisco.com>
2020-04-01 08:58:06 -07:00
Josh Hursey
288f6ebb61
Merge pull request #7585 from rlangefe/pr/README-Fix
Grammatical Fix in README
2020-04-01 10:13:57 -05:00
rlangefe
24a45f90ae Grammatical fix in README
Signed-off-by: rlangefe <langrc18@wfu.edu>

Fixed misplaced closing parenthesis on line 184 of the README file
2020-04-01 10:18:59 -04:00
Gilles Gouaillardet
a2c711b54b configury: try if -fpp flag is needed to preprocess .F90 files
.F90 files are preprocessed by gfortran and other compilers.
NAG compilers only preprocess .{ff,ff90,ff95} files, and the -fpp
flag is required to process .F90 files.

Fixes open-mpi/ompi#7583

Signed-off-by: Gilles Gouaillardet <gilles@rist.or.jp>
2020-04-01 16:26:47 +09:00
Ralph Castain
538d2de860
Merge pull request #7566 from rhc54/topic/up2
Update PMIx and PRRTE
2020-03-31 10:29:19 -07:00
Ralph Castain
556b3fcc00
PRRTE: Return non-zero status on timeout
Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-03-31 07:03:40 -07:00
Ralph Castain
1aabbe456d
Add extra libs to PRRTE binaries for external deps
libevent, hwloc, and pmix can be external and may require that their
libs be explicitly linked into the PRRTE binaries

Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-03-30 16:06:40 -07:00
Yossi Itigin
5dcd1f4e6c
Merge pull request #7575 from yosefe/topic/pml-ucx-fix-usage-of-mca-pml
pml/ucx: Fix usage of mca_pml_base_pml_check_selected()
2020-03-30 20:06:12 +03:00
Nathan Hjelm
160ff188b8
Merge pull request #7169 from hjelmn/fix_what_wg21_calls_our_problem_not_theirs_seriously__in_some_ways_they_are_correct_but_wtf
configure: use -iquote for non-system include paths
2020-03-30 09:22:54 -07:00
Ralph Castain
f88f271054
Cleanup few errors associated with tool support
Properly mark/detect that a daemon sourced the event broadcast to avoid
reinjecting it into the PMIx server library. Correct the source field
for the event notify call on launcher ready.

Update event notification for tool support
Deal with a variety of race conditions related to tool reconnection to a
different server.

Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-03-29 11:58:43 -07:00
Yossi Itigin
124f0c0d1f pml/ucx: Fix usage of mca_pml_base_pml_check_selected()
Pass the correct ompi_proc_t and array length to
mca_pml_base_pml_check_selected() during dynamic modex.

Signed-off-by: Yossi Itigin <yosefe@mellanox.com>
2020-03-29 17:46:45 +03:00
Ralph Castain
96759e8d9f
Merge pull request #7573 from rhc54/topic/usfoo
Fix typo in usnic btl
2020-03-28 14:42:50 -07:00
Ralph Castain
95db66d0c8
Fix typo in usnic btl
Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-03-27 20:27:45 -07:00
Howard Pritchard
f136a20cae
Merge pull request #6578 from hppritcha/topic/thread_framework2
Implement a MCA framework for threads
2020-03-27 15:55:48 -06:00
Austen Lauria
8a624ab613
Merge pull request #7523 from mkurnosov/fix-bcast-scatterallgather
Fix Bcast scatter_allgather
2020-03-27 14:17:53 -04:00
Shintaro Iwasaki
d7fba60de8 mca/threads: remove libevent hack
Argobots/Qthreads-aware libevent should be used instead.

Signed-off-by: Shintaro Iwasaki <siwasaki@anl.gov>
2020-03-27 10:16:04 -06:00
Shintaro Iwasaki
a7ea0d9bd7 ompi/request: move REQUEST constants from mca/threads to ompi/request
Signed-off-by: Shintaro Iwasaki <siwasaki@anl.gov>
2020-03-27 10:16:04 -06:00
Shintaro Iwasaki
69e8af536a mca/threads: fix tsd management
To suppress Valgrind warnings, opal_tsd_keys_destruct() needs to explicitly
release TSD values of the main thread.  However, they were not freed if keys are
created by non-main threads.  This patch fixes it.

This patch also optimizes allocation of opal_tsd_key_values by doubling its size
when count >= length instead of increasing the size by one.

Signed-off-by: Shintaro Iwasaki <siwasaki@anl.gov>
2020-03-27 10:16:03 -06:00
Shintaro Iwasaki
8cab081770 test/class: fix opal_fifo and opal_lifo
Signed-off-by: Shintaro Iwasaki <siwasaki@anl.gov>
2020-03-27 10:16:03 -06:00
Noah Evans
ee3517427e Add threads framework
Add a framework to support different types of threading models including
user space thread packages such as Qthreads and argobot:

https://github.com/pmodels/argobots

https://github.com/Qthreads/qthreads

The default threading model is pthreads.  Alternate thread models are
specificed at configure time using the --with-threads=X option.

The framework is static.  The theading model to use is selected at
Open MPI configure/build time.

mca/threads: implement Argobots threading layer

config: fix thread configury

- Add double quotations
- Change Argobot to Argobots
config: implement Argobots check

If the poll time is too long, MPI hangs.

This quick fix just sets it to 0, but it is not good for the
Pthreads version. Need to find a good way to abstract it.

Note that even 1 (= 1 millisecond) causes disastrous performance
degradation.

rework threads MCA framework configury

It now works more like the ompi/mca/rte configury,
modulo some edge items that are special for threading package
linking, etc.

qthreads module
some argobots cleanup

Signed-off-by: Noah Evans <noah.evans@gmail.com>
Signed-off-by: Shintaro Iwasaki <siwasaki@anl.gov>
Signed-off-by: Howard Pritchard <howardp@lanl.gov>
2020-03-27 10:15:45 -06:00
Brian Barrett
36e5863a9a
Merge pull request #7571 from bwbarrett/bugfix/mtl-add-procs
ofi: Call add_procs through PML
2020-03-27 08:24:35 -07:00
Ralph Castain
43e4f6c84b
Merge pull request #7572 from rhc54/topic/auth
Add .mailmap entry for Harumi Kuno
2020-03-27 08:19:06 -07:00
Jeff Squyres
5964f25748 .mailmap: Add entry for Harumi Kuno
Signed-off-by: Jeff Squyres <jsquyres@cisco.com>
2020-03-27 10:41:08 -04:00
Austen Lauria
8128c309c8
Merge pull request #6080 from ggouaillardet/topic/configury_lustre
configure: fix lustre detection
2020-03-27 09:24:08 -04:00
Brian Barrett
64d70b3076 ofi: Call add_procs through PML
Change ompi_mtl_ofi_get_endpoint() to call the active PML's
add_procs() rather than the OFI MTL add_procs() directly when
discovering a new process during operation.

Functionally, this has no impact in correct operation.  However,
the current behavior means that the heterogenous and active PML
checks are not being executed in the dynamic discovery case.

Signed-off-by: Brian Barrett <bbarrett@amazon.com>
2020-03-27 06:06:42 -07:00
Ralph Castain
1cf972dcaf
Update PMIx and PRRTE
Deprecate --am and --amca options

Avoid default param files on backend nodes
Any parameters in the PRRTE default or user param files will have been
picked up by prte and included in the environment sent to the prted, so
don't open those files on the backend.

Avoid picking up MCA param file info on backend
Avoid the scaling problem at PRRTE startup by only reading the system
and user param files on the frontend.

Complete revisions to cmd line parser for OMPI
Per specification, enforce following precedence order:

1. system-level default parameter file
1. user-level default parameter file
1. Anything found in the environment
1. "--tune" files. Note that "--amca" goes away and becomes equivalent to "--tune". Okay if it is provided more than once on a cmd line (we will aggregate the list of files, retaining order), but an error if a parameter is referenced in more than one file with a different value
1. "--mca" options. Again, error if the same option appears more than once with a different value. Allowed to override a parameter referenced in a "tune" file
1. "-x" options. Allowed to overwrite options given in a "tune" file, but cannot conflict with an explicit "--mca" option
1. all other options

Fix special handling of "-np"

Get agreement on jobid across the layers
Need all three pieces (PRRTE, PMIx, and OPAL) to agree on the nspace
conversion to jobid method

Ensure prte show_help messages get output
Print abnormal termination messages
Cleanup error reporting in persistent operations

Signed-off-by: Ralph Castain <rhc@pmix.org>

dd

Signed-off-by: Ralph Castain <rhc@pmix.org>
2020-03-26 16:01:11 -07:00
Ralph Castain
c704ed4cc5
Merge pull request #7554 from rhc54/topic/proc1
ompi_proc_t size reduction: part 1
2020-03-26 13:23:06 -07:00