openmpi

Автор	SHA1	Сообщение	Дата
Brian Barrett	e4698f5cd4	Shell of the Portals 4 collectives componetn This commit was SVN r28703.	2013-07-02 15:23:55 +00:00
Joshua Ladd	5d2d5e958c	Deleting garbage I accidentally committed. Thanks, Nathan\! This commit was SVN r28698.	2013-07-01 22:50:54 +00:00
Joshua Ladd	d7a50343bf	Per the details and schedule outlined in the attached RFC, Mellanox Technologies would like to CMR the new 'coll/hcoll' component. This component enables Mellanox Technologies' latest HPC middleware offering - 'Hcoll'. 'Hcoll' is a high-performance, standalone collectives library with support for truly asynchronous, non-blocking, hierarchical collectives via hardware offload on supporting Mellanox HCAs (ConnectX-3 and above.) To build the component, libhcoll must first be installed on your system, then you must configure OMPI with the configure flag: '--with-hcoll=/path/to/libhcoll'. Subsequent to installing, you may select the 'coll/hcoll' component at runtime as you would any other coll component, e.g. '-mca coll hcoll,tuned,libnbc'. This has been reviewed by Josh Ladd and should be added to cmr:v1.7:reviewer=jladd This commit was SVN r28694.	2013-07-01 22:39:43 +00:00
Jeff Squyres	c8258c06e2	In coll_sm, we alloc a huge chunk of shared memory, divvy it into lots of individual regions (each region is a multiple of page size in length), and each process claims its own regions by binding it to its local memory. Each process would end up membining something like 16 individual regions in the overall shmem segment. There were two errors in this code relating to the memory affinity pinning. Some combination of these two errors would lead to kernel panics (!) on my RHEL 6.2 x86_64 machines when used with mmap'ed shared memory (not posix or sysv shared memory, curiously enough): 1. The shared memory segment is initially divided into two regions: control and data. The control starts at the beginning of the shmem segment, the data starts after that. The data portion, unfortunately, was ''not'' aligned to a page. So all the multiple-of-page-size regions that we divvy up were also not alined on page boundaries. And therefore all the regions we tried to membind were not on page boundaries. The solution was to ensure that the data portion started on a page boundary. Then all of the individual regions were on page boundaries, too. That being said, in my tests, Linux mbind() fails gracefully when the address is not on a page boundary. So I'm not sure how this worked at all / led to a kernel panic... 2. There was some bad pointer math that resulted in membinding regions larger than they should have been, resulting in region overlaps. There were definitely overlaps between regions in the same process; it's likely that there were overlaps between regions of multiple processes, too -- I'm not sure (and don't care to figure out :-) ). The solution was to fix the pointer math so that each region membinds exactly only itself and no neighboring/overlapping regions. cmr:v1.7.2:reviewer=samuel This commit was SVN r28442.	2013-05-03 12:49:35 +00:00
Nathan Hjelm	9d4a26f47d	Update OMPI frameworks to use the MCA framework system. Notes: - This commit also eliminates the need for an available components list in use in several frameworks. None of the code in question was making use of the priority field of the priority component list item so these extra lists were removed. - Cleaned up selection code in several frameworks to sort lists using opal_list_sort. - Cleans up the ompi/orte-info functions. Expose the functions that construct the list of params so they can be used elsewhere. patches for mtl/portals4 from brian missed a few output variables in openib This commit was SVN r28241.	2013-03-27 21:17:31 +00:00
Nathan Hjelm	cf377db823	MCA/base: Add new MCA variable system Features: - Support for an override parameter file (openmpi-mca-param-override.conf). Variable values in this file can not be overridden by any file or environment value. - Support for boolean, unsigned, and unsigned long long variables. - Support for true/false values. - Support for enumerations on integer variables. - Support for MPIT scope, verbosity, and binding. - Support for command line source. - Support for setting variable source via the environment using OMPI_MCA_SOURCE_<var name>=source (either command or file:filename) - Cleaner API. - Support for variable groups (equivalent to MPIT categories). Notes: - Variables must be created with a backing store (char *, int , or bool *) that must live at least as long as the variable. - Creating a variable with the MCA_BASE_VAR_FLAG_SETTABLE enables the use of mca_base_var_set_value() to change the value. - String values are duplicated when the variable is registered. It is up to the caller to free the original value if necessary. The new value will be freed by the mca_base_var system and must not be freed by the user. - Variables with constant scope may not be settable. - Variable groups (and all associated variables) are deregistered when the component is closed or the component repository item is freed. This prevents a segmentation fault from accessing a variable after its component is unloaded. - After some discussion we decided we should remove the automatic registration of component priority variables. Few component actually made use of this feature. - The enumerator interface was updated to be general enough to handle future uses of the interface. - The code to generate ompi_info output has been moved into the MCA variable system. See mca_base_var_dump(). opal: update core and components to mca_base_var system orte: update core and components to mca_base_var system ompi: update core and components to mca_base_var system This commit also modifies the rmaps framework. The following variables were moved from ppr and lama: rmaps_base_pernode, rmaps_base_n_pernode, rmaps_base_n_persocket. Both lama and ppr create synonyms for these variables. This commit was SVN r28236.	2013-03-27 21:09:41 +00:00
Ralph Castain	a4b6fb241f	Remove all remaining vestiges of the Windows integration This commit was SVN r28137.	2013-02-28 17:31:47 +00:00
Ralph Castain	8d2fa3693b	First cut at removing the native Windows support. Remove all the Windows-specific components, and the .windows files sprinkled around. Remove the Windows platform files and MTT scripts. Update the NEWS to point Windows users to the cygwin package. This commit was SVN r28116.	2013-02-26 20:44:56 +00:00
Brian Barrett	312f37706e	In talking about this with Jeff and Ralph, we don't actually need ompi_show_help, because opal_show_help is replaced with an aggregating version when using ORTE, so there's no reason to directly call orte_show_help. This commit was SVN r28051.	2013-02-12 21:10:11 +00:00
Pavel Shamis	a31bc57849	Moving mca/common/netpatterns and commpaterns to ompi/patterns. This commit was SVN r28035.	2013-02-05 21:52:55 +00:00
Igor Usarov	8d80af6c10	Support FCA v3.0 This commit was SVN r27988.	2013-01-31 11:14:27 +00:00
Brian Barrett	b8442ba505	Revamp the handling of wrapper compiler flags. The user flags, main configure flags, and mca flags are kept seperate until the very end. The main configure wrapper flags should now be modified by using the OPAL_WRAPPER_FLAGS_ADD macro. MCA components should either let <framework>_<component>_{LIBS,LDFLAGS} be copied over OR set <framework>_<component>_WRAPPER_EXTRA_{LIBS,LDFLAGS}. The situations in which WRAPPER CPPFLAGS can be set by MCA components was made very small to match the one use case where it makes sense. This commit was SVN r27950.	2013-01-29 00:00:43 +00:00
Pavel Shamis	1f1e1efb7b	Removing leftovers of old infrastructure. cmr:v1.7 This commit was SVN r27942.	2013-01-28 19:11:42 +00:00
Brian Barrett	f42783ae1a	Move the RTE framework change into the trunk. With this change, all non-CR runtime code goes through one of the rte, dpm, or pubsub frameworks. This commit was SVN r27934.	2013-01-27 23:25:10 +00:00
Brian Barrett	14f4aa1198	Fix memory leak in nbc init This commit was SVN r27884.	2013-01-21 22:45:59 +00:00
Joshua Ladd	77df51c516	Fixes the definition of the first fragment and does not assume that first frag has offset_into_user_buff equal to zero. This fix should be added to cmr:v1.7.1:reviewer=pasha This commit was SVN r27775.	2013-01-08 20:24:58 +00:00
Mike Dubman	889d46e966	support for FCA v3.0 and up This commit was SVN r27731.	2012-12-31 05:49:22 +00:00
Nathan Hjelm	3e1b13b13a	Re-add support for old flex (2.5.4a and earlier) while still cleaning up properly in new flex. This commit was SVN r27657.	2012-12-07 00:12:43 +00:00
Nathan Hjelm	87e5f97400	add missing #include of opal/util/output.h This commit was SVN r27599.	2012-11-13 07:14:41 +00:00
Pavel Shamis	cbe6d6548a	Cleaning warnings in ml, sbgp, bcol. cmr:v1.7 This commit was SVN r27598.	2012-11-12 22:30:32 +00:00
Nathan Hjelm	e0f5137e46	add prototypes for lex destroy functions This commit was SVN r27580.	2012-11-09 22:00:27 +00:00
Mike Dubman	9392f1894e	Fixes MPI_Allgather when sendbuf is MPI_IN_PLACE. fixes 3342 This commit was SVN r27579.	2012-11-09 18:05:10 +00:00
Nathan Hjelm	8658bbc902	instead of relying on yyterminate to clean up the lex context call the destroy functions directly (after closing the file) This commit was SVN r27577.	2012-11-09 16:10:55 +00:00
Nathan Hjelm	7fb5caea92	Remove the finish_parsing function from various .l files. The function is incomplete (doesn't clean up the lex state) and should be replaced by *_yylex_destroy which correctly cleans up the state. Checked with the flex 2.5.35. Verified with valgrind that this fixes several "still reachable" leaks. cmr:v1.7 This commit was SVN r27571.	2012-11-06 19:26:14 +00:00
Nathan Hjelm	bdedd8b0d3	Per RFC modify the behavior of mca_base_components_close to NOT close the output. Modify frameworks to always close their output and set to -1. Reasoning: The old behavior was a little confusing. mca_base_components_open does not open an output stream so it is a little unexpected that mca_base_components_close does. To add to this several frameworks (that don't use mca_base_components_close) failed to close their output in the framework close function and others closed their output a second time. This change is an improvement to the symantics of mca_base_components_open/close as they are now symetric in their functionality. This commit was SVN r27570.	2012-11-06 19:09:26 +00:00
Nathan Hjelm	f3ce12e71a	Per RFC fix several leaks in opal and ompi. Details below. pml/v: - If vprotocol is not being used vprotocol_include_list is leaked. Assume vprotocol never takes ownership (see below) and always free the string. coll/ml: - (patch verified) calling mca_base_param_lookup_string after mca_base_param_reg_string is unnecessary. The call to mca_base_param_lookup_string causes the value returned by mca_base_param_reg_string to be leaked. - Need to free mca_coll_ml_component.config_file_name on component close. btl/openib: - calling mca_base_param_lookup_string after mca_base_param_reg_string is unnecessary. The call to mca_base_param_lookup_string causes the value returned by mca_base_param_reg_string to be leaked. vprotocol/base: - There was no way for pml/v to determine if vprotocol took ownership of vprotocol_include_list. Fix by always never ownership (use strdup). mca/base: - param_lookup will result in storage->stringval to be a newly allocated string if the mca parameter has a string value. ensure this string is always freed. cmr:v1.7 This commit was SVN r27569.	2012-11-06 18:57:46 +00:00
Mike Dubman	ca308974e0	Switched FCA collectives component from dlopen to compile-time linking to libfca This commit was SVN r27557.	2012-11-02 17:30:00 +00:00
Nathan Hjelm	2acd0f83de	Revert "Revert r27451 and r27456 - the cmd line parser is incorrectly marking the application as an MCA parameter". It appears the problem was not with the command line parser but the rsh plm. I don't know why this problem was not occuring before the command line parser changes but it appears to be resolved now. This commit was SVN r27527. The following SVN revision numbers were found above: r27451 --> open-mpi/ompi@d59034e6ef r27456 --> open-mpi/ompi@ecdbf34937	2012-10-30 19:45:18 +00:00
Brian Barrett	c0f1775620	Fix warnings in nbc This commit was SVN r27514.	2012-10-29 19:52:43 +00:00
Brian Barrett	8b40c0de9b	* Lock around tag management, so that it's thread safe * Only register the progress function on first call to a non-blocking collective operation, to try to reduce overall performance impact * Fix tag management in roll-over case This commit was SVN r27498.	2012-10-26 15:36:09 +00:00
Ralph Castain	e6014bf2e1	Revert r27451 and r27456 - the cmd line parser is incorrectly marking the application as an MCA parameter This commit was SVN r27477. The following SVN revision numbers were found above: r27451 --> open-mpi/ompi@d59034e6ef r27456 --> open-mpi/ompi@ecdbf34937	2012-10-24 18:38:44 +00:00
Nathan Hjelm	d59034e6ef	MCA: remove deprecated mca_base_param functions (mca_base_param_register_int, mca_base_param_register_string, mca_base_param_environ_variable). Remove all uses of deprecated functions. cmr:v1.7 This commit was SVN r27451.	2012-10-17 20:17:37 +00:00
George Bosilca	9984a7143f	Reorder the loop index. This commit was SVN r27423.	2012-10-08 21:34:26 +00:00
George Bosilca	b46167fc4a	Fix some issues with the MPI_IN_PLACE support. This commit was SVN r27422.	2012-10-08 21:34:04 +00:00
Ralph Castain	54db4c35eb	Get the trunk to build again when --without-hwloc is specified. Move a couple of key type definitions and utilities out from under the HAVE_HWLOC test so they are always available as they don't really depend on hwloc's presence. Tell two compnents not to build if hwloc is disabled: ompi/mca/sbgp/basesmsocket orte/mca/rmaps/lama Remove stale configure.params files from the sbgp framework as the OMPI build system no longer looks at those files. This commit was SVN r27377.	2012-09-26 23:24:27 +00:00
Pavel Shamis	1e7b958c2a	Cleaning warning in collectives code This commit was SVN r27331.	2012-09-12 19:47:23 +00:00
Pavel Shamis	8cf3c95494	Fixing ML COLL compilation issues on some SUN platforms. For more detail see following mail thread: http://www.open-mpi.org/community/lists/devel/2012/08/11448.php A lot of thanks to Paul Hargrove for the issue analysis and patch testing. Refs trac:3243 This commit was SVN r27178. The following Trac tickets were found above: Ticket 3243 --> https://svn.open-mpi.org/trac/ompi/ticket/3243	2012-08-29 14:10:42 +00:00
Yevgeny Kliteynik	8b5d634231	Enable support for FCA v2.5 This commit was SVN r27145.	2012-08-26 15:20:46 +00:00
Shiqing Fan	d141d94bd7	Include the new .windows files into the tarball. This commit was SVN r27121.	2012-08-23 12:50:51 +00:00
Shiqing Fan	95b9552546	include several components for Windows build. This commit was SVN r27108.	2012-08-22 14:46:49 +00:00
Pavel Shamis	5cedbb843c	Fixing compilation problems in ML collective component on SUN's systems. Thank you to Eugene Loh (Oracle) for discovering the problem and pin-pointing the solution. Refs trac:3243. This commit was SVN r27100. The following Trac tickets were found above: Ticket 3243 --> https://svn.open-mpi.org/trac/ompi/ticket/3243	2012-08-21 17:43:24 +00:00
Pavel Shamis	6fac989588	Cleaning warnings in collectives code. Refs trac:3243. This commit was SVN r27089. The following Trac tickets were found above: Ticket 3243 --> https://svn.open-mpi.org/trac/ompi/ticket/3243	2012-08-17 15:36:13 +00:00
Jeff Squyres	fc3ecd5d5a	Remove generated file. This commit was SVN r27080.	2012-08-16 22:08:04 +00:00
Ralph Castain	eda4cd5aa7	Cleanup warnings for improper use of C++ comment style, set ignores This commit was SVN r27079.	2012-08-16 21:52:14 +00:00
Pavel Shamis	b89f8fabc9	Adding Hierarchical Collectives project to the Open MPI trunk. The project includes following components and frameworks: - ML Collective component - NETPATTERNS and COMMPATTERNS common components - BCOL framework - SBGP framework Note: By default the ML collective component is disabled. In order to enable new collectives user should bump up the priority of ml component (coll_ml_priority) ============================================= Primary Contributors (in alphabetical order): Ishai Rabinovich (Mellanox) Joshua S. Ladd (ORNL / Mellanox) Manjunath Gorentla Venkata (ORNL) Mike Dubman (Mellanox) Noam Bloch (Mellanox) Pavel (Pasha) Shamis (ORNL / Mellanox) Richard Graham (ORNL / Mellanox) Vasily Filipov (Mellanox) This commit was SVN r27078.	2012-08-16 19:11:35 +00:00
Jeff Squyres	a4e97fb4c0	Ensure we assign "err" properly when invoking MCA_PML_CALLs. Although technically this is a necessary thing to do, it wasn't a tragedy that we didn't have it because err was initialize to 0 in the beginning of the functions where this problem occurred. Also, OMPI will likely abort if one of the MCA_PML_CALLs actually incurs an error (or, even if it doesn't, MPI doesn't define the behavior anyway ;-) ). But looking forward to an FT-aware world, fixing this issue is a Good Thing. Many thanks to Hristo Iliev for pointing out the issue. This commit was SVN r27070.	2012-08-16 17:49:48 +00:00
Shiqing Fan	2f442799f8	fix several typecasts This commit was SVN r26957.	2012-08-07 10:41:53 +00:00
Eugene Loh	10e3dc396b	Add a missing return value. This commit was SVN r26815.	2012-07-20 01:32:06 +00:00
Brian Barrett	2518014037	Fix a number of issues with IN_PLACE This commit was SVN r26814.	2012-07-19 21:29:43 +00:00
Eugene Loh	a3e02fdaff	With non-blocking collectives, a "round schedule" could fall on any address alignment, which typically causes problems on SPARC. Further, the pointer manipulation to access elements in a round schedule was clumsy. This change introduces macros to facilitate addressing and make it more portable. This commit was SVN r26802.	2012-07-18 17:08:24 +00:00
Brian Barrett	58413fa1e4	* properly setup communication infrastructure for libnbc. * Prevent infinite recursion in progress loop. Should fix improper barrier eugene was seeing. This commit was SVN r26758.	2012-07-06 13:59:03 +00:00
Brian Barrett	e0ceabd486	Need to set MPI_ERROR in the status before calling ompi_request_complete. This commit was SVN r26757.	2012-07-06 01:14:35 +00:00
Brian Barrett	27d45ad550	Implement reduce_scatter_block and ireduce_scatter_block, although possibly not nearly as optimal as they should be. This commit was SVN r26756.	2012-07-05 22:11:48 +00:00
George Bosilca	63278df92d	Prevent the coll SM from looking for information about remote procs during the init phase. This information is only available at a later stage. This commit was SVN r26746.	2012-07-04 21:15:40 +00:00
Brian Barrett	d56de80b5d	* Properly initialize handle variable as a request (since the coll_libnbc_request contains everything an NBC_Handle used to contain). Not sure how this slipped through... This commit was SVN r26710.	2012-07-02 16:39:42 +00:00
Brian Barrett	7e67bfa175	Use OMPI's ops instead of the libnbc ops. This commit was SVN r26708.	2012-07-02 15:47:22 +00:00
Brian Barrett	0b887ab5a1	* Remove unneeded prototype that was causing compile issues anyway * Use proper tag space (the negatives below the blocking communicators) instead of the point-to-point space * Use the PML interface instead of the MPI interface, since the MPI interface 1) shouldn't be used by components and 2) doesn't like negative tags This commit was SVN r26693.	2012-06-28 16:52:03 +00:00
Ralph Castain	a1344bc5c0	Add missing header to tarball This commit was SVN r26689.	2012-06-28 13:07:18 +00:00
Brian Barrett	32e70b691a	Re-enable non-blocking collectives in libnbc after finding issue with the definition of NBC_CACHE_SCHEDULE not being propogated to all uses. This commit was SVN r26686.	2012-06-27 22:08:19 +00:00
Brian Barrett	d85fdd2605	temporarily back out r26682 and r26683 until I can figure out why they cause crashes during shutdown This commit was SVN r26684. The following SVN revision numbers were found above: r26682 --> open-mpi/ompi@15a30af11f r26683 --> open-mpi/ompi@f6ea4b7234	2012-06-27 19:32:53 +00:00
Brian Barrett	f6ea4b7234	Remove now unneeded header file This commit was SVN r26683.	2012-06-27 18:43:40 +00:00
Brian Barrett	15a30af11f	Turn on all the non-blocking collectives provided by libnbc... This commit was SVN r26682.	2012-06-27 18:32:57 +00:00
Brian Barrett	3933d0a8f0	Ibarrier works! :) This commit was SVN r26680.	2012-06-27 15:58:17 +00:00
Josh Hursey	28681deffa	Backout the ORCA commit. :( There is a linking issue on Mac OSX that needs to be addressed before this is able to come back into the trunk. This commit was SVN r26676.	2012-06-27 01:28:28 +00:00
Josh Hursey	542330e3a7	Commit of ORCA: Open MPI Runtime Collaborative Abstraction This is a runtime interposition project that sits between the OMPI and ORTE layers in Open MPI. The project is described on the wiki: https://svn.open-mpi.org/trac/ompi/wiki/Runtime_Interposition And on this email thread: http://www.open-mpi.org/community/lists/devel/2012/06/11109.php This commit was SVN r26670.	2012-06-26 21:42:16 +00:00
Brian Barrett	7bdeafb772	Start bringing in libnbc. .ompi_ignored, as there's still a long way to go This commit was SVN r26658.	2012-06-25 22:38:06 +00:00
Brian Barrett	b9e8e4aeb9	* Initial merge of the non-blocking collectives interface. No implementation of the back-end yet, coming real soon now, need to solve some tag issues first. This commit was SVN r26641.	2012-06-22 20:54:12 +00:00
Jeff Squyres	5451ee46bd	Per r26575, the sync coll module is no longer necessary! (the crowd goes wild) This commit was SVN r26583. The following SVN revision numbers were found above: r26575 --> open-mpi/ompi@59e529cf1d	2012-06-08 19:19:19 +00:00
Jeff Squyres	2ba10c37fe	Per RFC, bring in the following changes: * Remove paffinity, maffinity, and carto frameworks -- they've been wholly replaced by hwloc. * Move ompi_mpi_init() affinity-setting/checking code down to ORTE. * Update sm, smcuda, wv, and openib components to no longer use carto. Instead, use hwloc data. There are still optimizations possible in the sm/smcuda BTLs (i.e., making multiple mpools). Also, the old carto-based code found out how many NUMA nodes were ''available'' -- not how many were used ''in this job''. The new hwloc-using code computes the same value -- it was not updated to calculate how many NUMA nodes are used ''by this job.'' * Note that I cannot compile the smcuda and wv BTLs -- I ''think'' they're right, but they need to be verified by their owners. * The openib component now does a bunch of stuff to figure out where "near" OpenFabrics devices are. '''THIS IS A CHANGE IN DEFAULT BEHAVIOR!!''' and still needs to be verified by OpenFabrics vendors (I do not have a NUMA machine with an OpenFabrics device that is a non-uniform distance from multiple different NUMA nodes). * Completely rewrite the OMPI_Affinity_str() routine from the "affinity" mpiext extension. This extension now understands hyperthreads; the output format of it has changed a bit to reflect this new information. * Bunches of minor changes around the code base to update names/types from maffinity/paffinity-based names to hwloc-based names. * Add some helper functions into the hwloc base, mainly having to do with the fact that we have the hwloc data reporting ''all'' topology information, but sometimes you really only want the (online \| available) data. This commit was SVN r26391.	2012-05-07 14:52:54 +00:00
George Bosilca	f09e3ce5a4	Spring cleanup. Nothing important. This commit was SVN r26247.	2012-04-06 15:48:07 +00:00
George Bosilca	654c75ff24	As suggested on the mailing list a while back, switch the default alltoallv algorithm to pairwise exchange instead of the default one. This might improve the scheduling and relax the pressure on the network. This commit was SVN r26246.	2012-04-06 15:47:29 +00:00
Josh Hursey	d1571b027a	Fix a few error return paths This commit was SVN r26233.	2012-04-04 15:11:03 +00:00
Mike Dubman	a45898ea9c	fix support for fca 2.2, warning fixes on rhel 6.x This commit was SVN r26166.	2012-03-20 10:00:52 +00:00
George Bosilca	a78a7bd8e8	The tuned collectives can now deal with more than 2Gb of data. This commit was SVN r26103.	2012-03-05 22:23:44 +00:00
George Bosilca	762b3e13a9	Use the correct name for the datatype destruction function. This commit was SVN r26100.	2012-03-05 15:54:53 +00:00
George Bosilca	7d523a8852	Avoid calling the bcast with counts larger than INT_MAX. This commit was SVN r26098.	2012-03-05 14:30:30 +00:00
George Bosilca	e8c358c188	Allow Open MPI to deal with size_t internally. This commit was SVN r26097.	2012-03-05 14:10:26 +00:00
George Bosilca	f83670211e	Allow the user to define dynamic rules for messages larger than 2GB. This commit was SVN r26084.	2012-03-02 21:16:23 +00:00
George Bosilca	8791ade293	Help he selection of the right algorithm for large data (> 2Gb). Thanks to Fujitsu for the patch. This commit was SVN r26080.	2012-03-02 19:12:22 +00:00
George Bosilca	72f731f25f	The SM2 collective component has not been updated in a long time. Rich, the original developer, agrees with this removal. This commit was SVN r25368.	2011-10-25 22:07:09 +00:00
Rainer Keller	4e6a6fc146	- Check, whether the compiler supports __builtin_clz (count leading zeroes); if so, use it for bit-operations like opal_cube_dim and opal_hibit. Implement two versions of power-of-two. In case of opal_next_poweroftwo, this reduces the average execution time from 83 cycles to 4 cycles (Intel Nehalem, icc, -O2, inlining, measured rdtsc, with loop over 2^27 values). Numbers for other functions are similar (but of course heavily depend on the usage, e.g. opal_hibit() with a start of 4 does not save much). The bsr instruction on AMD Opteron is also not as fast. - Replace various places where the next power-of-two is computed. Tested on Intel Nehalem Cluster with openib, compilers GNU-4.6.1 and Intel-12.0.4 using mpi_testsuite -t "Collective" with 128 processes. This commit was SVN r25270.	2011-10-11 22:49:01 +00:00
George Bosilca	2fefd3a928	Don't forget to move the pointer back by the true_lb. This commit was SVN r25262.	2011-10-11 20:15:49 +00:00
George Bosilca	ce7935c8fa	Obviously these were not needed. This commit was SVN r25231.	2011-10-04 14:56:34 +00:00
George Bosilca	80c02647c8	Each level (OPAL/ORTE/OMPI) should only return it's own constants, instead of the current mismatch. This commit was SVN r25230.	2011-10-04 14:50:31 +00:00
Rolf vandeVaart	0749a220e8	Add support for MPI_IN_PLACE to MPI_Exscan. Required for MPI 2.2 compliance. Reviewed by Jeff Squyres. This fixes trac:2221. This commit was SVN r25165. The following Trac tickets were found above: Ticket 2221 --> https://svn.open-mpi.org/trac/ompi/ticket/2221	2011-09-20 14:54:41 +00:00
George Bosilca	9687e7f38e	This commit fixes trac:2679 and should be added to cmr:v1.4:reviewer=jsquyres and cmr:v1.5:reviewer=jsquyres This commit was SVN r25155. The following Trac tickets were found above: Ticket 2679 --> https://svn.open-mpi.org/trac/ompi/ticket/2679	2011-09-18 00:58:26 +00:00
Wesley Bland	4e7ff0bd5e	By popular demand the epoch code is now disabled by default. To enable the epochs and the resilient orte code, use the configure flag: --enable-resilient-orte This will define both: ORTE_ENABLE_EPOCH ORTE_RESIL_ORTE This commit was SVN r25093.	2011-08-26 22:16:14 +00:00
Mike Dubman	96ef2fc0e4	fix handling datatypes which have a gap in the beginning This commit was SVN r24936.	2011-07-25 06:30:09 +00:00
Jeff Squyres	b2b781e537	Fix a few miscelaneous memory leaks. This commit was SVN r24865.	2011-07-08 16:39:58 +00:00
Wesley Bland	e1ba09ad51	Add a resilience to ORTE. Allows the runtime to continue after a process (or ORTED) failure. Note that more work will be necessary to allow the MPI layer to take advantage of this. Per RFC: http://www.open-mpi.org/community/lists/devel/2011/06/9299.php This commit was SVN r24815.	2011-06-23 20:38:02 +00:00
Samuel Gutierrez	81f38b258a	commit of new shared memory backing facility framework (shmem) and its components. This commit was SVN r24795.	2011-06-21 15:41:57 +00:00
George Bosilca	65661a3cb4	Dont use a temporary string. This commit was SVN r24786.	2011-06-20 09:29:19 +00:00
Mike Dubman	36db9c6233	* updated copyrights * added support for non-contig data layout in FCA This commit was SVN r24702.	2011-05-16 14:43:11 +00:00
Jeff Squyres	ec90a3ba6d	Fix a few memory leaks, and ensure that coll sm is also registering the common SM MCA params. This commit was SVN r24497.	2011-03-08 17:36:59 +00:00
Mike Dubman	70392ac1dc	fca: broadcast comm_new return status to from rank0 to all ranks prior to exiting with an error This commit was SVN r24481.	2011-03-02 22:18:43 +00:00
George Bosilca	87f3109df4	Cleanups. This commit was SVN r24458.	2011-02-25 00:28:32 +00:00
Mike Dubman	89ba89e812	- added support for upcomming FCA v2.1 version This commit was SVN r24418.	2011-02-21 14:08:24 +00:00
Mike Dubman	81222e1fe7	* fix PGI compiler support which does not have __BASE_FILE__ macro This commit was SVN r24369.	2011-02-10 06:42:37 +00:00
Mike Dubman	4a2e29eb32	updated Makefile with a new file This commit was SVN r24199.	2011-01-01 14:11:49 +00:00
Mike Dubman	c56e3141cb	fca: fix segmentation fault when no underlying collective implementation is found This commit was SVN r24198.	2010-12-31 12:03:49 +00:00
Mike Dubman	3d517c0285	ABI cleanups This commit was SVN r24193.	2010-12-28 07:11:46 +00:00
Mike Dubman	b339a7a07b	Add FCA 1.2/2.0 backward compatibility, depending on OMPI_FCA_VERSION_xx macro definition. This commit was SVN r24192.	2010-12-27 21:32:34 +00:00
Shiqing Fan	f43862420c	Convert the bad dos line endings to unix style for all windows related files. This commit was SVN r24137.	2010-12-02 12:08:08 +00:00
Mike Dubman	956e030f28	support for dynamic rules to control offload This commit was SVN r24094.	2010-11-29 04:11:57 +00:00
Mike Dubman	5a7d76bb9c	resolve many warnings, comply to c99 This commit was SVN r24040.	2010-11-11 12:14:31 +00:00
Mike Dubman	f9bebe53f9	- fix fca support for MPI_IN_PLACE in allgather and allgatherv collectives This commit was SVN r23841.	2010-10-06 19:09:02 +00:00
Mike Dubman	f525245498	- support for MPI_IN_PLACE during gather ops - fix ABI check and message This commit was SVN r23840.	2010-10-06 16:27:45 +00:00
Jeff Squyres	73bcc4a36b	Fix mistake that came in via the ompi-agen tree in r23764. The mistake wasn't part of the core autogen upgrade; it was an additional 'bonus' cleanup. Oops. The mistake will always create a set of directories under installdir, even if you do not --with-devel-headers. The set of directories will be empty, but still -- they should not be there at all. This commit fixes that -- the directories are not created at all if you do not --with-devel-headers This commit was SVN r23801. The following SVN revision numbers were found above: r23764 --> open-mpi/ompi@40a2bfa238	2010-09-24 22:53:28 +00:00
Mike Dubman	58aa7fd161	enabling gather This commit was SVN r23773.	2010-09-20 06:29:54 +00:00
Mike Dubman	f754bde8eb	fixing r23764 leftovers, adopting Jeff's note This commit was SVN r23772. The following SVN revision numbers were found above: r23764 --> open-mpi/ompi@40a2bfa238	2010-09-20 06:27:43 +00:00
Mike Dubman	bd9a1f28a3	revert r23764 in ompi/mca/coll/fca This commit was SVN r23771. The following SVN revision numbers were found above: r23764 --> open-mpi/ompi@40a2bfa238	2010-09-20 06:06:45 +00:00
Ralph Castain	40a2bfa238	WARNING: Work on the temp branch being merged here encountered problems with bugs in subversion. Considerable effort has gone into validating the branch. However, not all conditions can be checked, so users are cautioned that it may be advisable to not update from the trunk for a few days to allow MTT to identify platform-specific issues. This merges the branch containing the revamped build system based around converting autogen from a bash script to a Perl program. Jeff has provided emails explaining the features contained in the change. Please note that configure requirements on components HAVE CHANGED. For example. a configure.params file is no longer required in each component directory. See Jeff's emails for an explanation. This commit was SVN r23764.	2010-09-17 23:04:06 +00:00
Mike Dubman	104d57f69a	* Support allgatherv, convert displs and rcounts arrays to bytes. * change comm_init API - no need to pass local rank groups, fca calculates that on its own. * remove local rank list from module - libfca maintains that now. * in fca_bcast and fca_reduce - pass root rank index and let libfca figure out the local rank index. This commit was SVN r23716.	2010-09-05 09:49:59 +00:00
Mike Dubman	48274c1c77	better control for enable/disable specific coll APIs This commit was SVN r23708.	2010-09-02 09:22:24 +00:00
Mike Dubman	8ef56bf258	* drop support for FCA v1.2 * add support for FCA ABI * add support for allgather This commit was SVN r23705.	2010-09-01 11:29:10 +00:00
Mike Dubman	fca50c4a09	comply to code-style: no c++ style commends This commit was SVN r23645.	2010-08-24 13:42:21 +00:00
Mike Dubman	9cb2e0490b	removed #if 0 This commit was SVN r23643.	2010-08-24 13:32:28 +00:00
Mike Dubman	a036c24253	revert fix to comply with #2534 - use op->o_name directly - cosmetic prints This commit was SVN r23614.	2010-08-15 11:04:34 +00:00
Mike Dubman	16d7169680	refactoring: * split fca_open() into fca_register() and fca_open() This commit was SVN r23598.	2010-08-12 12:05:23 +00:00
Mike Dubman	ba5bc9b674	fixes: * fixup lookup of supported ops by name: in ompi 1.5.x the op string representation were changed from MPI_XXX to MPI_OP_XXX (relative to OMPI 1.4.x) * keep compat between diff versions of FCA * better error handling (return error if symbol not found) * register to opal_progress and call fca_progress API This commit was SVN r23597.	2010-08-12 08:15:55 +00:00
Mike Dubman	7d1a8a154d	fca: - keep compat to fca v1.2 and fca 2.0 - fix segv - keep compat to ompi 1.4.x This commit was SVN r23569.	2010-08-08 13:28:41 +00:00
Mike Dubman	2914d11793	fix datatype API This commit was SVN r23552.	2010-08-04 14:01:54 +00:00
Shiqing Fan	a096cc9082	it's not a component for Windows, so get rid of the Windows support files. This commit was SVN r23543.	2010-08-02 17:12:40 +00:00
Mike Dubman	7cbe9b43c2	initial release of Voltaire FCA (fabric collective accelerator) collective component - compatible with FCA v1.2 This commit was SVN r23539.	2010-08-02 11:25:53 +00:00
Jeff Squyres	87e17a41da	Ensure that the com_rules[] array entries are initialized to NULL in case individual entries aren't used, but dynamic rules are enabled (i.e., at least one or more of them are not NULL, meaning that they'll all be assumed to be either NULL or a valid value). This commit was SVN r23361.	2010-07-07 14:04:18 +00:00
Jeff Squyres	c8bb7537e7	Remove include/opal/sys/cache.h -- its only purpose in life was to #define CACHE_LINE_SIZE to 128. This name has a conflict on NetBSD, and it seems kinda odd to have a header file that ''only'' defines a single value. Also, we'll soon be raising hwloc to be a first-class item, so having this file around seemed kinda weird. Therefore, I replaced CACHE_LINE_SIZE with opal_cache_line_size, an int (in opal/runtime/opal_init.c and opal/runtime/opal.h) on the rationale that we can fill this in at runtime with hwloc info (trunk and v1.5/beyond, only). The only place we ''needed'' a compile-time CACHE_LINE_SIZE was in the BTL SM (for struct padding), so I made a new BTL_SM_ preprocessor macro with the old CACHE_LINE_SIZE value (128). That use isn't suitable for run-time hwloc information, anyway. This commit was SVN r23349.	2010-07-06 14:33:36 +00:00
Samuel Gutierrez	2fb7c344fc	Added a new System V (sysv) shared memory component for Open MPI. Configure Option: --enable-sysv MCA Parameter: mpi_common_sm mpi_common_sm accepts a comma delimited list of: [sysv],mmap (order dependent). The first component that is successfully selected is used. For example, -mca mpi_common_sm sysv,mmap will first try sysv. If sysv is not successfully selected, then mmap will be used. mmap will be used if mpi_common_sm is not provided. Notes: Please make certain that your system's shmmax limit, or equivalent, is larger than mpool_sm_min_size. Otherwise, shmget may fail. This commit was SVN r23260.	2010-06-09 16:58:52 +00:00
Edgar Gabriel	f6598138ba	fix some instances, where we might have allocated 0 bytes. Also, for allgather make sure that we do not call coll_gather and coll_bcast in the very same instances, since some collective (intra) modules do not seem to like the fact if they are called for scount or rcount being zero (for regular intra-communicator operations, this is handled on the MPI API layer). Fixes trac:2405 This commit was SVN r23188. The following Trac tickets were found above: Ticket 2405 --> https://svn.open-mpi.org/trac/ompi/ticket/2405	2010-05-20 22:23:44 +00:00
George Bosilca	c51932c250	Don't forget to initialize "line" in all cases. This commit was SVN r23178.	2010-05-19 21:19:45 +00:00
Abhishek Kulkarni	afbe3e99c6	* Wrap all the direct error-code checks of the form (OMPI_ERR_* == ret) with (OMPI_ERR_* = OPAL_SOS_GET_ERR_CODE(ret)), since the return value could be a SOS-encoded error. The OPAL_SOS_GET_ERR_CODE() takes in a SOS error and returns back the native error code. * Since OPAL_SUCCESS is preserved by SOS, also change all calls of the form (OPAL_ERROR == ret) to (OPAL_SUCCESS != ret). We thus avoid having to decode 'ret' to get the native error code. This commit was SVN r23162.	2010-05-17 23:08:56 +00:00
Jeff Squyres	181331d65e	Very minor nits/updates. This commit was SVN r22977.	2010-04-15 14:44:55 +00:00
Rainer Keller	f6e4694d67	- Print the name correctly when a certain sync module is disabled This should be cmr'd to v1.5 and v1.4.2 (but the svn post hook won't let me at the moment). This commit was SVN r22827.	2010-03-13 21:07:34 +00:00
Jeff Squyres	5ec2d8764b	Amendment to r22671: change the name of the new communicator flag from INTERNAL to EXTRA_RETAIN, because not all "internal" communicators have this flag set (only internal communicators with CIDs less than their parent). Hence, what this flag ''really'' means is that there was an extra RETAIN performed on it. So name the flag just that -- EXTRA_RETAIN -- indicating that an extra RETAIN has occurred. This commit was SVN r22690. The following SVN revision numbers were found above: r22671 --> open-mpi/ompi@61dee816db	2010-02-23 21:24:07 +00:00
Edgar Gabriel	61dee816db	This commit fixes a bug on how to deal with the potential if a 'dependent' communicator that we created has a lower CID than the parent comm. This can happen when using the hierarch collective communication module or for inter-communicators (since we make a duplicate of the original communicator). This is not a problem as long as the user calls MPI_Comm_free on the parent communicator. However, if the communicators are not freed by the user but released by Open MPI in MPI_Finalize, we walk through the list of still available communicators and free them one by one. Thus, local_comm is freed before the actual inter-communicator. However, the local_comm pointer in the inter communicator will still contain the 'previous' address of the local_comm and thus this will lead to a segmentation violation. In order to prevent that from happening, we increase the reference counter local_comm by one if its CID is lower than the parent. We cannot increase however its reference counter if the CID of local_comm is larger than the CID of the inter communicators, since a regular MPI_Comm_free would leave in that the case the local_comm hanging around and thus we would not recycle CID's properly, which was the reason and the cause for this trouble. This commit fixes tickets 2094 and 2166. Note however, that I want to close them manually, since a slightly different patch is required for the 1.4 series. This commit will have to be applied for the 1.5 series. And I will need a volunteer to review it. This commit was SVN r22671.	2010-02-19 23:45:30 +00:00
George Bosilca	3bceb20b1c	Only get the receive datatype extent on the root process, as every other process should ignore this value. Thanks to Michael Hofmann for investigating this issue. This commit closes trac:2268. This commit was SVN r22639. The following Trac tickets were found above: Ticket 2268 --> https://svn.open-mpi.org/trac/ompi/ticket/2268	2010-02-17 16:01:50 +00:00
George Bosilca	144143a3ff	Remove an unused local variable. This commit was SVN r22566.	2010-02-05 22:27:24 +00:00
George Bosilca	bc7ceb3587	We enable the dynamic decision if the user force it via an MCA argument or set it in the decision file. In addition do a fine grain activation, i.e. per collective function. This commit was SVN r22510.	2010-01-29 09:03:59 +00:00
Shiqing Fan	ad763c327d	Restore several linked libraries that were deleted by mistake in r22405. This commit was SVN r22415. The following SVN revision numbers were found above: r22405 --> open-mpi/ompi@872a4047ba	2010-01-14 21:50:42 +00:00
Shiqing Fan	872a4047ba	Fix the bug that caused by ADD_DEPENDENCIES() from different version of CMake. In CMake 2.6 and earlier, this function add dependencies for targets and also link the target libraries automatically, but in CMake 2.8,this behavior has been changed, i.e. it will only add the dependencies but no link, which will cause linking errors at compilation time. This commit was SVN r22405.	2010-01-14 18:10:20 +00:00
George Bosilca	f0303a8b25	Indentation. This commit was SVN r22254.	2009-12-02 22:03:52 +00:00
George Bosilca	3a2f071018	If the user asked for dynamic rules but "forget" to provide them, nicely complain and switch back to the default behavior (fixed rules). This commit was SVN r22109.	2009-10-19 17:58:47 +00:00
Jeff Squyres	0f8ac9223f	Refs trac:2023, #2027 . This commit does a bunch of things: * Address all remaining code review items from CMR #2023: * Defer mmap setup to be lazy; only set it up the first time we invoke a collective. In this way, we don't penalize apps that make lots of communicators but don't invoke collectives on them (per #2027). * Remove the extra assignments of mca_coll_sm_one (fixing a convertor count setup that was the real problem). * Remove another extra/unnecessary assignment. * Increase libevent polling frequency when using the RML to bootstrap mmap'ed memory. * Fix a minor procs-related memory leak in btl_sm. * Commit a datatype fix that George and I discovered along the way to fixing the coll sm. * Improve error messages when mmap fails, potentially trying to de-alloc any allocated memory when that happens. * Fix a previously-unnoticed confusion between extent and true_extent in coll sm reduce. This commit was SVN r22049. The following Trac tickets were found above: Ticket 2023 --> https://svn.open-mpi.org/trac/ompi/ticket/2023	2009-10-02 17:13:56 +00:00
Shiqing Fan	96e9ffa016	Fix a type cast. This commit was SVN r22034.	2009-09-30 14:02:47 +00:00
Jeff Squyres	152bc14079	Rename the help file to be consistent with others; add it to the Makefile.am. This commit was SVN r22005.	2009-09-23 20:28:49 +00:00
Jeff Squyres	1ef988c3d9	A slight optimization: no longer call sched_yield() when polling for shmem progress (or the Windows equiv). Instead, poll hard on the condition, but periocially call opal_progress(). This allows badly-formed apps (e.g., the ibm test communicator/bsend_free) to actually complete. To be clear, there are far too many apps out there that assume that MPI collectives will actually progress the rest of MPI. I don't like putting in a feature to enable broken apps, but I have a dim recollection of this issue coming up before (apps "hanging" when testing the sm coll because they assumed that calling collectives would trigger other MPI progress). Rather than have people claim that OMPI is broken, I prefer to put in this "workaround". :-( Indeed, the bsend_free test ''may'' be coded that way for exactly that reason...? I don't remember offhand... This commit was SVN r21984.	2009-09-21 22:20:44 +00:00
Jeff Squyres	533633b8cb	Fixes trac:1988. The little bug that turned out to be huge. Yoinks. * Various cosmetic/style updates in the btl sm * Clean up concept of mpool module (I think that code was written way back when the concept of "modules" was fuzzy) * Bring over some old fixes from the /tmp/timattox-sm-coll/ tree to fix potential segv's when mmap'ed regions were at different addresses in different processes (thanks Tim!). * Change sm coll to no longer use mpool as its main source of shmem; rather, just mmap its own segment (because it's fixed size -- there was nothing to be gained by using mpool; shedding the use of mpool saved a lot of complexity in the sm coll setup). This effectively made Tim's fixes moot (because now everything is an offset into the mmap that is computed locally; there are no global pointers). :-) * Slightly updated common/sm to allow making mmap's for a specific set of procs (vs. ''all'' procs in the process). This potentially allows for same-host-inter-proc mmaps -- yay! * Fixed many, many things in the coll sm (particularly in reduce): * Fixed handling of MPI_IN_PLACE in reduce and allreduce * Fixed handling of non-contiguous datatypes in reduce * Changed the order of reductions to go from process (n-1)'s data to process 0's data, because that's how all other OMPI coll components work * Fixed lots of usage of ddt functions * When using a non-contiguous datatype, if the root process is not (n-1), now we used a 2nd convertor to copy from shmem to the rbuf (saves a memory copy vs. what was done before) * Lots and lots of little cleanups, clarifications, and minor optimizations (although still more could be done -- e.g., I think the use of write memory barriers is fairly sub-optimal; they could be ganged together at the root, for example) I'm marking this as "fixes trac:1988" and closing the ticket; if something is still broken, we can re-open the ticket. This commit was SVN r21967. The following Trac tickets were found above: Ticket 1988 --> https://svn.open-mpi.org/trac/ompi/ticket/1988	2009-09-15 00:25:21 +00:00
Rainer Keller	8e1b23779f	- Replace combinations of #if defined (c_plusplus) defined (__cplusplus) followed by extern "C" { and the closing counterpart by BEGIN_C_DECLS and END_C_DECLS. Notable exceptions are: - opal/include/opal_config_bottom.h: This is our generated code, that itself defines BEGIN_C_DECL and END_C_DECL - ompi/mpi/cxx/mpicxx.h: Here we do not include opal_config_bottom.h: - Belongs to external code: opal/mca/backtrace/darwin/MoreBacktrace/MoreDebugging/MoreBacktrace.c opal/mca/backtrace/darwin/MoreBacktrace/MoreDebugging/MoreBacktrace.h - opal/include/opal/prefetch.h: Has C++ specific macros that are protected: - Had #if ... } #endif _and_ END_C_DECLS (aka end up with 2x END_C_DECLS) ompi/mca/btl/openib/btl_openib.h - opal/event/event.h has #ifdef __cplusplus as BEGIN_C_DECLS... - opal/win32/ompi_process.h: had extern "C"\n {... opal/win32/ompi_process.h: dito - ompi/mca/btl/pcie/btl_pcie_lex.l: needed to add *_C_DECLS ompi/mpi/f90/test/align_c.c: dito - ompi/debuggers/msgq_interface.h: used #ifdef __cplusplus - ompi/mpi/f90/xml/common-C.xsl: Amend Tested on linux using --with-openib and --with-mx The following do not contain either opal_config.h, orte_config.h or ompi_config.h (but possibly other header files, that include one of the above): ompi/mca/bml/r2/bml_r2_ft.h ompi/mca/btl/gm/btl_gm_endpoint.h ompi/mca/btl/gm/btl_gm_proc.h ompi/mca/btl/mx/btl_mx_endpoint.h ompi/mca/btl/ofud/btl_ofud_endpoint.h ompi/mca/btl/ofud/btl_ofud_frag.h ompi/mca/btl/ofud/btl_ofud_proc.h ompi/mca/btl/openib/btl_openib_mca.h ompi/mca/btl/portals/btl_portals_endpoint.h ompi/mca/btl/portals/btl_portals_frag.h ompi/mca/btl/sctp/btl_sctp_endpoint.h ompi/mca/btl/sctp/btl_sctp_proc.h ompi/mca/btl/tcp/btl_tcp_endpoint.h ompi/mca/btl/tcp/btl_tcp_ft.h ompi/mca/btl/tcp/btl_tcp_proc.h ompi/mca/btl/template/btl_template_endpoint.h ompi/mca/btl/template/btl_template_proc.h ompi/mca/btl/udapl/btl_udapl_eager_rdma.h ompi/mca/btl/udapl/btl_udapl_endpoint.h ompi/mca/btl/udapl/btl_udapl_mca.h ompi/mca/btl/udapl/btl_udapl_proc.h ompi/mca/mtl/mx/mtl_mx_endpoint.h ompi/mca/mtl/mx/mtl_mx.h ompi/mca/mtl/psm/mtl_psm_endpoint.h ompi/mca/mtl/psm/mtl_psm.h ompi/mca/pml/cm/pml_cm_component.h ompi/mca/pml/csum/pml_csum_comm.h ompi/mca/pml/dr/pml_dr_comm.h ompi/mca/pml/dr/pml_dr_component.h ompi/mca/pml/dr/pml_dr_endpoint.h ompi/mca/pml/dr/pml_dr_recvfrag.h ompi/mca/pml/example/pml_example.h ompi/mca/pml/ob1/pml_ob1_comm.h ompi/mca/pml/ob1/pml_ob1_component.h ompi/mca/pml/ob1/pml_ob1_endpoint.h ompi/mca/pml/ob1/pml_ob1_rdmafrag.h ompi/mca/pml/ob1/pml_ob1_recvfrag.h ompi/mca/pml/v/pml_v_output.h opal/include/opal/prefetch.h opal/mca/timer/aix/timer_aix.h opal/util/qsort.h test/support/components.h This commit was SVN r21855. The following SVN revision numbers were found above: r2 --> open-mpi/ompi@58fdc18855	2009-08-20 11:42:18 +00:00
George Bosilca	23e8ce91ba	Rework the selection logic for the tuned collectives. All supported collectives now are able to use the dynamic rules. Moreover, these rules are loaded only once, and stored at the component level. All communicators are able to use these rules (not only MPI_COMM_WORLD as until now). A lot of minor corrections, memory management issues and reduction in the amount of memory used by the tuned collectives. This commit was SVN r21825.	2009-08-14 21:06:23 +00:00
Shiqing Fan	bce2f44154	Update related .windows files with proper compiling properties, in order to have a successful DSO build. This commit was SVN r21805.	2009-08-12 08:55:58 +00:00
Rainer Keller	6c5532072a	- Split the datatype engine into two parts: an MPI specific part in OMPI and a language agnostic part in OPAL. The convertor is completely moved into OPAL. This offers several benefits as described in RFC http://www.open-mpi.org/community/lists/devel/2009/07/6387.php namely: - Fewer basic types (int* and float* types, boolean and wchar - Fixing naming scheme to ompi-nomenclature. - Usability outside of the ompi-layer. - Due to the fixed nature of simple opal types, their information is completely known at compile time and therefore constified - With fewer datatypes (22), the actual sizes of bit-field types may be reduced from 64 to 32 bits, allowing reorganizing the opal_datatype structure, eliminating holes and keeping data required in convertor (upon send/recv) in one cacheline... This has implications to the convertor-datastructure and other parts of the code. - Several performance tests have been run, the netpipe latency does not change with this patch on Linux/x86-64 on the smoky cluster. - Extensive tests have been done to verify correctness (no new regressions) using: 1. mpi_test_suite on linux/x86-64 using clean ompi-trunk and ompi-ddt: a. running both trunk and ompi-ddt resulted in no differences (except for MPI_SHORT_INT and MPI_TYPE_MIX_LB_UB do now run correctly). b. with --enable-memchecker and running under valgrind (one buglet when run with static found in test-suite, commited) 2. ibm testsuite on linux/x86-64 using clean ompi-trunk and ompi-ddt: all passed (except for the dynamic/ tests failed!! as trunk/MTT) 3. compilation and usage of HDF5 tests on Jaguar using PGI and PathScale compilers. 4. compilation and usage on Scicortex. - Please note, that for the heterogeneous case, (-m32 compiled binaries/ompi), neither ompi-trunk, nor ompi-ddt branch would successfully launch. This commit was SVN r21641.	2009-07-13 04:56:31 +00:00
Jeff Squyres	92e40cb20a	Enable the coll sync component to barrier before each 1000th collective. This commit was SVN r21594.	2009-07-02 20:16:45 +00:00
Jeff Squyres	cad12fda5f	* Remove an extra blank line from the help file * Add the help file to the Makefile.am so that it gets installed This commit was SVN r21567.	2009-06-30 18:58:09 +00:00
Rainer Keller	fbb2834977	- Missed string.h to get rid of warnings... This commit was SVN r21265.	2009-05-22 23:47:49 +00:00
Rainer Keller	225a1d6d8e	- For memcpy and memset need string.h This commit was SVN r21259.	2009-05-21 22:36:06 +00:00
Greg Koenig	60485ff95f	This is a very large change to rename several #define values from OMPI_* to OPAL_*. This allows opal layer to be used more independent from the whole of ompi. NOTE: 9 "svn mv" operations immediately follow this commit. This commit was SVN r21180.	2009-05-06 20:11:28 +00:00
Shiqing Fan	cd565923d3	Completely remove ltdl support for Windows build. This commit was SVN r21170.	2009-05-05 18:59:13 +00:00
George Bosilca	039fed1973	Fix Coverity CID #264 . This commit was SVN r21162.	2009-05-05 13:54:55 +00:00
George Bosilca	db096d7d3a	Fix Coverity CID #304 . This commit was SVN r21159.	2009-05-05 13:47:47 +00:00
Rainer Keller	9736af1191	- Fix Coverity CID 182: Well, well, just do not "call" ompi_comm_rank twice but rather reuse variable... - Fix Coverity CID 1262: Using uninitialized value "(statuses[err_index]).MPI_ERROR" Sure, these statuses are only initialized after ompi_request_wait_all, so introduce a short-circuit label to jump to... This commit was SVN r21153.	2009-05-05 12:28:51 +00:00
Rainer Keller	221fb9dbca	... Delayed due to notifier commits earlier this day ... - Delete unnecessary header files using contrib/check_unnecessary_headers.sh after applying patches, that include headers, being "lost" due to inclusion in one of the now deleted headers... In total 817 files are touched. In ompi/mpi/c/ header files are moved up into the actual c-file, where necessary (these are the only additional #include), otherwise it is only deletions of #include (apart from the above additions required due to notifier...) - To get different MCAs (OpenIB, TM, ALPS), an earlier version was successfully compiled (yesterday) on: Linux locally using intel-11, gcc-4.3.2 and gcc-SVN + warnings enabled Smoky cluster (x86-64 running Linux) using PGI-8.0.2 + warnings enabled Lens cluster (x86-64 running Linux) using Pathscale-3.2 + warnings enabled This commit was SVN r21096.	2009-04-29 01:32:14 +00:00
Shiqing Fan	3d4e0472d6	Add windows support files into the tarball, including .windows, CMakeLists.txt files, and CMake modules. Thanks to Jeff for testing it on Linux. This commit was SVN r21069.	2009-04-24 16:39:33 +00:00
George Bosilca	c5b1bdd57c	Correctly deal with the error case. The problem is tricky: the MPI standard doesn't allow MPI_ERR_IN_STATUS to be returned from any functions that return only one completed request (few exception here: wait_some and wait_all and the test versions). As we use an wait_all in these send_receive functions we should convert the MPI_ERR_IN_STATUS to the real error, i.e. the one comming from the MPI_ERROR field in the status corresponding to the failed request. This commit was SVN r20907.	2009-03-31 23:44:59 +00:00
Rainer Keller	d8cf4c0fec	- Get pgcc on XT to complain less: In case we use memcmp, strlen, strup and friends include <string.h> Also several constants.h are not included directly - Let's have mca_topo_base_cart_create return ompi-errors in ompi/mca/topo/base/topo_base_cart_create.c This commit was SVN r20773.	2009-03-13 02:10:32 +00:00
Jeff Squyres	14ee1b7ba2	Refs trac:1826: remove barriers before all non-rooted collective ops. This commit was SVN r20763. The following Trac tickets were found above: Ticket 1826 --> https://svn.open-mpi.org/trac/ompi/ticket/1826	2009-03-12 02:23:08 +00:00
Rainer Keller	ec0ed48718	- Revert r20739 This commit was SVN r20742. The following SVN revision numbers were found above: r20739 --> open-mpi/ompi@781caee0b6	2009-03-05 21:56:03 +00:00
Rainer Keller	781caee0b6	- First of two or three patches, in orte/util/proc_info.h: Adapt orte_process_info to orte_proc_info, and change orte_proc_info() to orte_proc_info_init(). - Compiled on linux-x86-64 - Discussed with Ralph This commit was SVN r20739.	2009-03-05 20:36:44 +00:00
Shiqing Fan	99b415a7e0	On windows, the mca_common_* libraries should be installed in bin, otherwise the libraries that are dependent on them, e.g. shared build of mca_btl_sm, couldn't be loaded at runtime. This commit fixes the problem. This commit was SVN r20735.	2009-03-05 14:57:35 +00:00
Rainer Keller	9dea63d63a	- Last of intrusive commits (promised)... err for now. Anyway, this is blocking the move: do not include pml.h if not really needed, aka none of the following used: mca_pml MCA_PML_CALL OMPI_ANY_TAG OMPI_ANY_SOURCE OMPI_PROC_NULL - Notable exceptions (deleting in one header->adding): - ompi/mca/mtl/psm/ - ompi/mca/osc/rdma/ - ompi/mca/btl/openib/btl_openib_endpoint.c depended on pml_base_sendreq.h - Tested on Linux/x86-64, this time including make check (thanks Jeff and Ralph) This commit was SVN r20725.	2009-03-04 17:06:51 +00:00
Rainer Keller	811f2bd9b4	- As discussed on RFC, move the ompi_bitmap to the opal layer. Add a check against a maximum (actually get rid of ifs internally to opal_bitmap.c) -- the functionality to set the current maximum size opal_bitmap_set_max_size() is currently only used in attribute.c to set the maximum OMPI_FORTRAN_HANDLE_MAX... Tested on linux/x86-64 with intel-tests with all_tests_no_perf_f run with 6 procs. Let's look into MTT as well... This commit was SVN r20708.	2009-03-03 22:25:13 +00:00
Rich Graham	7ef1550267	add an index to indicate which socket group I belong to. This commit was SVN r20672.	2009-03-02 14:39:54 +00:00
Rich Graham	daf7673aff	gather socket information - not debugged.` This commit was SVN r20670.	2009-03-02 10:58:12 +00:00
Rainer Keller	96e1b9b747	- Header orte/mca/rml/rml.h is not needed if no occurence of orte_rml or ORTE_RML. As the others compiles fine with -Wimplicit-function-declaration This commit was SVN r20639.	2009-02-26 03:52:31 +00:00
Terry Dontje	0178b6c45f	Added padding to predefined handle structures to maintain library version to version compatibility. This commit was SVN r20627.	2009-02-24 17:17:33 +00:00
Shiqing Fan	2148220ce4	Update the share libs dependency for windows build. This commit was SVN r20625.	2009-02-23 17:49:46 +00:00
Jeff Squyres	3742c3550c	Add "sync" collective component. This component is totally deactivated by default. It is activated by setting either of the following two MCA parameters to values greater than 0: * coll_sync_barrier_before * coll_sync_barrier_after If !_before is >0, then the sync coll collective will insert itself before the underlying collective operations and invoke a barrier before every Nth barrier (N == coll_sync_barrier_before). Similar for !_after. Note that N is a _per communicator_ value; not global to the MPI process. If both are 0 (which is the default), this component returns NULL for the comm query, meaning that it is not insertted into the coll module stack. The intent of this component is to provide a a workaround for applications with large numbers of collectives of short messages that can cause unbounded unexpected messages. Specifically, it is possible for some iterative collective communication patterns to cause unbounded unexpected messages. Forcing a barrier before or after every Nth collective operation would prevent that behavior by forcing applications to synchronize (and thereby consume any outstanding unexpected messages caused by collectives on the same communicator). Open MPI still needs to bound unexpected messages resource consumption at the receiver, but this is a viable workaround for at least some symptoms of the problem. Additionally, there has been anecdotal evidence of some applications that "perfom better" when they put barriers after other collective operations. This could be due to many factors -- including shortening the unexpected message queue. Putting this component in Open MPI allows people to try this with their own applications and give real world feedback on this kind of behavior. This commit was SVN r20584.	2009-02-18 23:32:44 +00:00
Rainer Keller	d81443cc5a	- On the way to get the BTLs split out and lessen dependency on orte: Often, orte/util/show_help.h is included, although no functionality is required -- instead, most often opal_output.h, or orte/mca/rml/rml_types.h Please see orte_show_help_replacement.sh commited next. - Local compilation (Linux/x86_64) w/ -Wimplicit-function-declaration actually showed two missing #include "orte/util/show_help.h" in orte/mca/odls/base/odls_base_default_fns.c and in orte/tools/orte-top/orte-top.c Manually added these. Let's have MTT the last word. This commit was SVN r20557.	2009-02-14 02:26:12 +00:00
Jeff Squyres	8b29e27ead	Some minor valgrind-inspired cleanups: fix some memory leaks This commit was SVN r20543.	2009-02-13 03:45:32 +00:00
Tim Mattox	9b83df22ec	Fix some "is proc on local node?" logic that got accidentally flipped by r20496 for the sm BTL, openib BTL on iWarp, and the sm & sm2 coll modules. This commit was SVN r20515. The following SVN revision numbers were found above: r20496 --> open-mpi/ompi@4cdf91a8d4	2009-02-11 15:02:38 +00:00
Ralph Castain	4cdf91a8d4	Per the RFC, extend the current use of the ompi_proc_t flags field (without changing the field itself). The prior ompi_proc_t structure had a uint8_t flag field in it, where only one bit was used to flag that a proc was "local". In that context, "local" was constrained to mean "local to this node". This commit provides a greater degree of granularity on the term "local", to include tests to see if the proc is on the same socket, PC board, node, switch, CU (computing unit), and cluster. Add #define's to designate which bits stand for which local condition. This was added to the OPAL layer to avoid conflicting with the proposed movement of the BTLs. To make it easier to use, a set of macros have been defined - e.g., OPAL_PROC_ON_LOCAL_SOCKET - that test the specific bit. These can be used in the code base to clearly indicate which sense of locality is being considered. All locations in the code base that looked at the current proc_t field have been changed to use the new macros. Also modify the orte_ess modules so that each returns a uint8_t (to match the ompi_proc_t field) that contains a complete description of the locality of this proc. Obviously, not all environments will be capable of providing such detailed info. Thus, getting a "false" from a test for "on_local_socket" may simply indicate a lack of knowledge. This commit was SVN r20496.	2009-02-10 02:20:16 +00:00
Jeff Squyres	4d8a187450	Two major things in this commit: * New "op" MPI layer framework * Addition of the MPI_REDUCE_LOCAL proposed function (for MPI-2.2) = Op framework = Add new "op" framework in the ompi layer. This framework replaces the hard-coded MPI_Op back-end functions for (MPI_Op, MPI_Datatype) tuples for pre-defined MPI_Ops, allowing components and modules to provide the back-end functions. The intent is that components can be written to take advantage of hardware acceleration (GPU, FPGA, specialized CPU instructions, etc.). Similar to other frameworks, components are intended to be able to discover at run-time if they can be used, and if so, elect themselves to be selected (or disqualify themselves from selection if they cannot run). If specialized hardware is not available, there is a default set of functions that will automatically be used. This framework is ''not'' used for user-defined MPI_Ops. The new op framework is similar to the existing coll framework, in that the final set of function pointers that are used on any given intrinsic MPI_Op can be a mixed bag of function pointers, potentially coming from multiple different op modules. This allows for hardware that only supports some of the operations, not all of them (e.g., a GPU that only supports single-precision operations). All the hard-coded back-end MPI_Op functions for (MPI_Op, MPI_Datatype) tuples still exist, but unlike coll, they're in the framework base (vs. being in a separate "basic" component) and are automatically used if no component is found at runtime that provides a module with the necessary function pointers. There is an "example" op component that will hopefully be useful to those writing meaningful op components. It is currently .ompi_ignore'd so that it doesn't impinge on other developers (it's somewhat chatty in terms of opal_output() so that you can tell when its functions have been invoked). See the README file in the example op component directory. Developers of new op components are encouraged to look at the following wiki pages: https://svn.open-mpi.org/trac/ompi/wiki/devel/Autogen https://svn.open-mpi.org/trac/ompi/wiki/devel/CreateComponent https://svn.open-mpi.org/trac/ompi/wiki/devel/CreateFramework = MPI_REDUCE_LOCAL = Part of the MPI-2.2 proposal listed here: https://svn.mpi-forum.org/trac/mpi-forum-web/ticket/24 is to add a new function named MPI_REDUCE_LOCAL. It is very easy to implement, so I added it (also because it makes testing the op framework pretty easy -- you can do it in serial rather than via parallel reductions). There's even a man page! This commit was SVN r20280.	2009-01-14 23:44:31 +00:00
Edgar Gabriel	1072812bcf	not every element in the pointer array list contains a valid entry. Thus, do not try to free elements if the list returns NULL. This commit was SVN r20275.	2009-01-14 19:11:30 +00:00
George Bosilca	01adc999c5	Correctly forward the right module if we call another collective function. Kudos to Edgar for figuring out this tricky bug. This commit was SVN r20267.	2009-01-14 03:22:54 +00:00
Jeff Squyres	11b375f8b5	CIDs 1080-1090: assert() checks were not sufficient to check for NEGATIVE_RETURNS from _reg_int() because those are not always checked. So replace them with real if() checks. This commit was SVN r20195.	2009-01-03 15:56:25 +00:00
Jeff Squyres	f13ea32830	Remove the code checkig the MCA "coll" parameter for a list of coll components to use. This code was rendered obsolete (albiet harmless) by the MCA base improvements that only open the components that were specified by each framework's MCA parameter. This commit was SVN r20176.	2008-12-31 13:40:51 +00:00
Jeff Squyres	759a295cc9	Gaah -- missed one s/m/component/g This commit was SVN r20175.	2008-12-31 13:35:37 +00:00
Jeff Squyres	955d1e132d	Rename a variable to be "component" (not "m"), to emphasize that it is the component struct, not a module. This commit was SVN r20174.	2008-12-31 13:32:46 +00:00
Jeff Squyres	865900dd27	Nothing of substance; just indenting changes (''finally'' update this framework base to 4 space tabs!). This commit was SVN r20173.	2008-12-31 12:17:08 +00:00
Jeff Squyres	ce313fa391	Minor fixes to a few comments This commit was SVN r20172.	2008-12-31 11:34:27 +00:00
Jeff Squyres	d533215dac	Fix a comment to reflect the right version number This commit was SVN r20169.	2008-12-30 12:39:32 +00:00
Nysal Jan	ee8ec6f6b5	Remove dead/redundant code. Minimize number of calloc invocations This commit was SVN r20121.	2008-12-12 10:55:50 +00:00
Shiqing Fan	a5281f0434	- 1/4 commit for Windows Visual Studio and CCP support: CMakeLists and .windows files. In contribs preconfigured and precompiled parts. This commit was SVN r20108.	2008-12-10 20:59:20 +00:00
Rolf vandeVaart	137729d2f9	Fix warnings (thanks Jeff) from previous fix. This is extra fix for ticket #1554. This commit was SVN r19728.	2008-10-10 14:35:52 +00:00
Tim Mattox	de623ea161	Remove a redundant if & goto. This commit was SVN r19724.	2008-10-09 15:07:56 +00:00
Rolf vandeVaart	aad4427caa	Fix the implementation of MPI_Reduce_scatter on intercommunicators. We still do an interreduce but it is now followed by an intrascatterv. This fixes trac:1554. This commit was SVN r19723. The following Trac tickets were found above: Ticket 1554 --> https://svn.open-mpi.org/trac/ompi/ticket/1554	2008-10-09 14:35:20 +00:00
Rolf vandeVaart	13e8975f83	In the case where we detect a value of 0 in the recvcount array, fall back to the simpler algorithms. This is not the optimal solution, but it works. This commit was SVN r19702.	2008-10-07 19:44:51 +00:00
Rolf vandeVaart	0a0ddfc934	Handle MPI_IN_PLACE correctly in the ompi_coll_tuned_reduce_scatter_intra_ring function. We were not adjusting the sendbuf in this case so we were reducing garbage. This fixes ticket #1506. This commit was SVN r19673.	2008-10-02 20:01:27 +00:00
George Bosilca	325d006577	Mostly cleanups, and eventually a little bit more scalable add_procs. There was an argument that was barely used, and on return at the PML level it contained nothing usable. It has been removed, so now we're using less memory ... This commit was SVN r19657.	2008-09-30 15:47:43 +00:00
George Bosilca	6a9514ee08	Make the code match the comment. I checked with Jelena, and based on the papers we published this is the expected algorithm for the specified message and communicator size. This commit closes ticket #1330. This commit was SVN r19563.	2008-09-15 23:28:40 +00:00
Edgar Gabriel	ef2bb46e45	no need to create and free the groups. We just want to translate the ranks and we can use the internal group structures right away for that operation. Fixes an issue with groups that have not been freed previously, due to the fact that ompi_group_free was not visible here (I know, this could have been solved also by setting OMPI_DECLSPEC on ompi_group_free, but this solution should be faster.) This commit was SVN r19362.	2008-08-19 13:59:58 +00:00
Edgar Gabriel	149ecb8d7d	1. debug the four new algorithms 2. fix a bug in the initial communicator creation of llcomm 3. fix a bug which showed up as the result of fixing issue number 2: we have to check now whether llcomm has really be created before freeing the according llcomm in hierarch_destruct. This commit was SVN r19361.	2008-08-18 21:54:35 +00:00
Edgar Gabriel	7cbc4a4077	adding four different algorithms for a hierarchical bcast which try to generate an overlap between the different layers. Why four versions? Because there is right now always the trade-off between using non-blocking operations on a layer with a trivial, linear algorithm and using the more sophisticaed algorithms in a blocking manner. - bcast_intra_seg used the bcast of lcomm and llcomm, similarly to original algorithm in hierarch. However, it can segment the message, such that we might get an overlap between the two layers. This overlap is based on the assumption, that a process might be done early with a bcast and can start the next one. - bcast_intra_seg1: replaces the llcomm->bcast by isend/irecvs to increase the overlap, keeps the lcomm->bcast however - bcast_intra_seg2: replaced lcomm->bcast by isend/irecvs to increase the overlap, keeps however llcomm->bcast - bcast_intra_seg3: replaced both lcomm->bcast and llcomm->bcast by isend/irecvs The code is lightly tested, more testing to follow right now. This commit was SVN r19358.	2008-08-18 16:05:44 +00:00
George Bosilca	a6e3a47102	Fix typo. This commit was SVN r19312.	2008-08-17 20:08:38 +00:00
Rich Graham	e64f028d62	add missing header file for errno. This commit was SVN r19246.	2008-08-12 01:34:13 +00:00
Jeff Squyres	54ab811426	Fix CID 1036: minor resource leak on error This commit was SVN r19236.	2008-08-11 20:37:36 +00:00
Rainer Keller	ee1fe9015a	- Make sure, that the *param_index are > 0 (here, we don't pass errors up...). Coverity CID 1080 - 1090 - Really make sure, the user does not specify stupid negative values. This commit was SVN r19233.	2008-08-11 11:21:04 +00:00
George Bosilca	3c8d43deed	Remove unused variable (Coverty fix 178). This commit was SVN r19195.	2008-08-06 14:09:43 +00:00
George Bosilca	567c691354	Remove unused variable (Coverty fix 177). This commit was SVN r19194.	2008-08-06 14:08:34 +00:00
George Bosilca	f6ebdf8896	Remove unused variable (Coverty fix 176). This commit was SVN r19193.	2008-08-06 14:07:20 +00:00
George Bosilca	c021427002	Remove unused variable (Coverty fix 175). This commit was SVN r19192.	2008-08-06 14:06:08 +00:00
George Bosilca	6c8017e9b7	Remove unused variable (Coverty fix 174). This commit was SVN r19191.	2008-08-06 14:04:54 +00:00
George Bosilca	afc79d1651	Remove unused variable (Coverty fix 173). This commit was SVN r19190.	2008-08-06 14:03:33 +00:00
George Bosilca	5e3a5b7c13	Remove unused variable (Coverty fix 172). This commit was SVN r19188.	2008-08-06 14:01:33 +00:00
George Bosilca	d897710e4f	Remove unused variable (Coverty fix 171). This commit was SVN r19187.	2008-08-06 14:00:22 +00:00
George Bosilca	417b727006	Remove unused variable (Coverty fix 170). This commit was SVN r19186.	2008-08-06 13:59:03 +00:00
George Bosilca	4f91b7806c	Remove unused variable (Coverty fix 169). This commit was SVN r19185.	2008-08-06 13:57:43 +00:00
Rainer Keller	23c2292478	- Fix variable set but not used Coverity CID1058 This commit was SVN r19184.	2008-08-06 13:57:38 +00:00
Jeff Squyres	0af7ac53f2	Fixes trac:1392, #1400 * add "register" function to mca_base_component_t * converted coll:basic and paffinity:linux and paffinity:solaris to use this function * we'll convert the rest over time (I'll file a ticket once all this is committed) * add 32 bytes of "reserved" space to the end of mca_base_component_t and mca_base_component_data_2_0_0_t to make future upgrades [slightly] easier * new mca_base_component_t size: 196 bytes * new mca_base_component_data_2_0_0_t size: 36 bytes * MCA base version bumped to v2.0 * '''We now refuse to load components that are not MCA v2.0.x''' * all MCA frameworks versions bumped to v2.0 * be a little more explicit about version numbers in the MCA base * add big comment in mca.h about versioning philosophy This commit was SVN r19073. The following Trac tickets were found above: Ticket 1392 --> https://svn.open-mpi.org/trac/ompi/ticket/1392	2008-07-28 22:40:57 +00:00
Jeff Squyres	d37a25a2d0	Remove per http://www.open-mpi.org/community/lists/devel/2008/07/4386.php This commit was SVN r18972.	2008-07-22 00:57:23 +00:00
Edgar Gabriel	798f47b430	Fixes ticket #1334 hierarch disables itself now if the pml module used is not ob1. The reason is, that the multi-level hierarchy detection algorithm checks the names of the btl modules used. In case there are no btl's, we would segfault. Furthermore, three minor changes: - the 2-level hierarchy detection is now the default (sm vs. everything else in the world). - add udapl to the list of protocols checked for by the multi-level hierarch detection - some of the verbose statements of hierarch were inaccurate. Fixed those comments/messages. This commit was SVN r18817.	2008-07-07 18:44:48 +00:00
Ralph Castain	9613b3176c	Effectively revert the orte_output system and return to direct use of opal_output at all levels. Retain the orte_show_help subsystem to allow aggregation of show_help messages at the HNP. After much work by Jeff and myself, and quite a lot of discussion, it has become clear that we simply cannot resolve the infinite loops caused by RML-involved subsystems calling orte_output. The original rationale for the change to orte_output has also been reduced by shifting the output of XML-formatted vs human readable messages to an alternative approach. I have globally replaced the orte_output/ORTE_OUTPUT calls in the code base, as well as the corresponding .h file name. I have test compiled and run this on the various environments within my reach, so hopefully this will prove minimally disruptive. This commit was SVN r18619.	2008-06-09 14:53:58 +00:00
Ralph Castain	c992e99035	Remove the tags from orte_output_open and the filtering operation from orte_output - this will be handled differently to improve the XML output interface This commit was SVN r18557.	2008-06-03 14:24:01 +00:00
Rolf vandeVaart	18879285c7	Fix the selection logic to prevent memory leaks. More work may be done in the priority logic but for now we just fix the leaks and preserve current behavior. This commit fixes trac:1307. This commit was SVN r18504. The following Trac tickets were found above: Ticket 1307 --> https://svn.open-mpi.org/trac/ompi/ticket/1307	2008-05-27 14:16:39 +00:00
Rolf vandeVaart	5baa733ad5	Fix another warning (using a variable before it was initialized.) Thanks Jeff for pointing this out. This commit was SVN r18489.	2008-05-23 13:57:55 +00:00
Rich Graham	b08839f9f5	change reduce-scatter/gather for non-power of 2. Spreading out the load for the non-power of 2 phase of the reduction. This commit was SVN r18486.	2008-05-22 21:42:42 +00:00
Rich Graham	f2a4b67809	automate the allreduce selection logic. This commit was SVN r18484.	2008-05-22 20:53:35 +00:00
Rich Graham	5900415a25	for non-powers of 2, distribute the work on the first step among all the procs doing the work. This commit was SVN r18480.	2008-05-22 18:50:53 +00:00
George Bosilca	c31cc5b270	Remove a warning about line being unused. This commit was SVN r18472.	2008-05-21 20:46:22 +00:00
Edgar Gabriel	0500420bec	fixing a bug in the inter-communicator scatter operation, where we used accidentally rcount instead of scounts. This commit was SVN r18466.	2008-05-20 21:17:19 +00:00
Rolf vandeVaart	74d0259480	Add new implentation of barrier. This shows better performance on some clusters. However, no decision logic is changed by this commit so default behavior has not changed. This is only selectable by runtime parameters. This commit was SVN r18464.	2008-05-20 17:37:41 +00:00
Rolf vandeVaart	71091a19c3	Fix bug in spacing of code per https://svn.open-mpi.org/trac/ompi/wiki/CodingStyle . This commit was SVN r18463.	2008-05-20 14:11:10 +00:00
Rolf vandeVaart	763f5259a8	Fix memory leak of 88 bytes that occurred on each call to MPI_Comm_dup. Need to release the items and the item list after selecting the collective modules that are being used. Reviewed by Jeff Squyres. This commit was SVN r18457.	2008-05-19 21:34:01 +00:00
Jeff Squyres	7154776465	Removed unused variable / compiler warning. This commit was SVN r18454.	2008-05-19 13:41:45 +00:00
Rolf vandeVaart	375406e1fa	Remove the ignore files as decided at Tuesday's developers conference call. Now, hierarchical collectives will be compiled in but the priority is still at 0 requiring a user to set mca parameters to enable them. This commit was SVN r18440.	2008-05-15 01:26:52 +00:00
Jeff Squyres	671f0c379d	Remove a whole pile of orte/util/show_help.h's that I missed. :-( This commit was SVN r18437.	2008-05-14 11:32:33 +00:00
Jeff Squyres	e7ecd56bd2	This commit represents a bunch of work on a Mercurial side branch. As such, the commit message back to the master SVN repository is fairly long. = ORTE Job-Level Output Messages = Add two new interfaces that should be used for all new code throughout the ORTE and OMPI layers (we already make the search-and-replace on the existing ORTE / OMPI layers): * orte_output(): (and corresponding friends ORTE_OUTPUT, orte_output_verbose, etc.) This function sends the output directly to the HNP for processing as part of a job-specific output channel. It supports all the same outputs as opal_output() (syslog, file, stdout, stderr), but for stdout/stderr, the output is sent to the HNP for processing and output. More on this below. * orte_show_help(): This function is a drop-in-replacement for opal_show_help(), with two differences in functionality: 1. the rendered text help message output is sent to the HNP for display (rather than outputting directly into the process' stderr stream) 1. the HNP detects duplicate help messages and does not display them (so that you don't see the same error message N times, once from each of your N MPI processes); instead, it counts "new" instances of the help message and displays a message every ~5 seconds when there are new ones ("I got X new copies of the help message...") opal_show_help and opal_output still exist, but they only output in the current process. The intent for the new orte_* functions is that they can apply job-level intelligence to the output. As such, we recommend that all new ORTE and OMPI code use the new orte_* functions, not thei opal_* functions. === New code === For ORTE and OMPI programmers, here's what you need to do differently in new code: * Do not include opal/util/show_help.h or opal/util/output.h. Instead, include orte/util/output.h (this one header file has declarations for both the orte_output() series of functions and orte_show_help()). * Effectively s/opal_output/orte_output/gi throughout your code. Note that orte_output_open() takes a slightly different argument list (as a way to pass data to the filtering stream -- see below), so you if explicitly call opal_output_open(), you'll need to slightly adapt to the new signature of orte_output_open(). * Literally s/opal_show_help/orte_show_help/. The function signature is identical. === Notes === * orte_output'ing to stream 0 will do similar to what opal_output'ing did, so leaving a hard-coded "0" as the first argument is safe. * For systems that do not use ORTE's RML or the HNP, the effect of orte_output_* and orte_show_help will be identical to their opal counterparts (the additional information passed to orte_output_open() will be lost!). Indeed, the orte_* functions simply become trivial wrappers to their opal_* counterparts. Note that we have not tested this; the code is simple but it is quite possible that we mucked something up. = Filter Framework = Messages sent view the new orte_* functions described above and messages output via the IOF on the HNP will now optionally be passed through a new "filter" framework before being output to stdout/stderr. The "filter" OPAL MCA framework is intended to allow preprocessing to messages before they are sent to their final destinations. The first component that was written in the filter framework was to create an XML stream, segregating all the messages into different XML tags, etc. This will allow 3rd party tools to read the stdout/stderr from the HNP and be able to know exactly what each text message is (e.g., a help message, another OMPI infrastructure message, stdout from the user process, stderr from the user process, etc.). Filtering is not active by default. Filter components must be specifically requested, such as: {{{ $ mpirun --mca filter xml ... }}} There can only be one filter component active. = New MCA Parameters = The new functionality described above introduces two new MCA parameters: * '''orte_base_help_aggregate''': Defaults to 1 (true), meaning that help messages will be aggregated, as described above. If set to 0, all help messages will be displayed, even if they are duplicates (i.e., the original behavior). * '''orte_base_show_output_recursions''': An MCA parameter to help debug one of the known issues, described below. It is likely that this MCA parameter will disappear before v1.3 final. = Known Issues = * The XML filter component is not complete. The current output from this component is preliminary and not real XML. A bit more work needs to be done to configure.m4 search for an appropriate XML library/link it in/use it at run time. * There are possible recursion loops in the orte_output() and orte_show_help() functions -- e.g., if RML send calls orte_output() or orte_show_help(). We have some ideas how to fix these, but figured that it was ok to commit before feature freeze with known issues. The code currently contains sub-optimal workarounds so that this will not be a problem, but it would be good to actually solve the problem rather than have hackish workarounds before v1.3 final. This commit was SVN r18434.	2008-05-13 20:00:55 +00:00
Rolf vandeVaart	0e32dd1022	Add MPI_Alltoallv to tuned collectives and add a pairwise implementation of MPI_Alltoallv. However, do not change the default behavior for now. The only way to use new pairwise implementation is via mca parameters. This commit was SVN r18394.	2008-05-07 02:31:24 +00:00
Rich Graham	4d1ae7b05f	accidentally made a change in the wrong place. This commit was SVN r18262.	2008-04-23 17:32:05 +00:00
Rich Graham	293dd6ad4e	add myself to list of people building this module. This commit was SVN r18261.	2008-04-23 17:25:36 +00:00
Rich Graham	7658cc79e4	Pass in the correct module to the reduction call. This commit was SVN r18260.	2008-04-23 17:23:30 +00:00
Tim Mattox	0215474cb8	Fix two bugs in coll_sm_module.c from bit-rot: Fixed a selection bug, and removed a bogus "free(proc)" call which ultimately caused MPI_Finalize to crash. This commit was SVN r18235.	2008-04-22 18:41:21 +00:00
Rich Graham	df35223603	add selection logic for barrier and reduce. This commit was SVN r18215.	2008-04-19 22:40:04 +00:00
Rich Graham	bee8b42f29	remove debug code that would not let people run. Add infrastructure for blocking-barrier. This commit was SVN r18214.	2008-04-19 01:34:04 +00:00
Rich Graham	6c77fa4921	add a blocking shared memory algorithm. This commit was SVN r18185.	2008-04-16 22:10:23 +00:00
Rich Graham	249445d61f	added reduce-scatter followed by gather to root. This commit was SVN r18133.	2008-04-11 13:49:08 +00:00
Rich Graham	a6bdbfab97	implement allreduce as reduce-scatter, followed by an allgather. This commit was SVN r18132.	2008-04-11 04:06:29 +00:00
Rich Graham	70f3aab5f2	remove some code that is not needed. This commit was SVN r18128.	2008-04-10 17:32:04 +00:00
Rich Graham	5c7db1e315	remove 2 race conditions in the buffer recycling logic. This commit was SVN r18127.	2008-04-10 17:20:52 +00:00
Edgar Gabriel	4964434205	reverting commit 18122, since the commit was executed accidentally in the wring directory. The UH copyrights do belong into this file (i.e. because of the fix which is in the 1.2 branch, the UH copyright notes are in the header there alreary), but I want to have the proper log for that. This commit was SVN r18124.	2008-04-10 15:09:31 +00:00
Edgar Gabriel	f87830767a	the verification of recvcount==0 and rank = root was braking inter-communicator scatter, since the root (root==MPI_ROOT) might very well have recvcount=0. The same fix has been applied to gather.c just the other way round. Fixes the bug reported on the mainling list by Martin Audet. If there is a 1.2.7 this fix might be worthwhile porting it over. Please note, that while the test works now for basic and for inter, we get a 0byte malloc warning from the inter module, which we still have to fix in a separate patch. This commit was SVN r18122.	2008-04-10 14:58:51 +00:00
Rich Graham	c6783549ef	getting old This commit was SVN r18110.	2008-04-09 16:55:16 +00:00
Rich Graham	1a20c3ce51	more debug. This commit was SVN r18109.	2008-04-09 16:19:52 +00:00
Rich Graham	e7e18303f6	more debug. This commit was SVN r18108.	2008-04-09 15:10:58 +00:00
Rich Graham	b14c6b17d5	adding debug output. This commit was SVN r18107.	2008-04-09 13:32:01 +00:00
Rich Graham	10434fb2f1	add barrier synchorinzation at the end of the module init, to avoid initializing shared memory variables in use. This commit was SVN r18105.	2008-04-09 03:44:40 +00:00
Rich Graham	19bb1a2e86	fix initialization bug. This commit was SVN r18104.	2008-04-08 23:34:06 +00:00
Rich Graham	a69a8d9626	initialize the flags. This commit was SVN r18102.	2008-04-08 22:16:39 +00:00
Rich Graham	8765a2bbdd	more debug code. This commit was SVN r18101.	2008-04-08 20:38:20 +00:00
Rich Graham	08becf33b5	add more debugging. This commit was SVN r18100.	2008-04-08 18:44:50 +00:00
Rich Graham	aa1b7dd406	more debug This commit was SVN r18099.	2008-04-08 03:56:47 +00:00
Rich Graham	0c18bdeff7	more debug code. This commit was SVN r18098.	2008-04-08 03:04:20 +00:00
Rich Graham	9d5a7238df	Add some debugging code. This commit was SVN r18097.	2008-04-07 23:20:15 +00:00
Rich Graham	fa696734d5	add some debug code. This commit was SVN r18096.	2008-04-07 21:03:23 +00:00
Rich Graham	1b54e8b76e	fix buffer management for nb-barrier. This commit was SVN r18081.	2008-04-05 21:59:04 +00:00
Rich Graham	94f8fd365c	a few reduction optimizations. Add bcast. This commit was SVN r18075.	2008-04-02 19:02:33 +00:00
George Bosilca	a00ca20446	More cleanups. This commit was SVN r18069.	2008-04-02 06:38:33 +00:00
Rich Graham	eb5d6096f1	add reduction routine - fix buffer recycling logic which was totally broken. This commit was SVN r18065.	2008-04-01 22:56:18 +00:00
Rich Graham	90e53ca9ee	debug the pipeline algorithm. This commit was SVN r18008.	2008-03-28 15:10:07 +00:00
Rich Graham	e2ad9c4be2	adjust to change in orte_process_info. This commit was SVN r17986.	2008-03-27 01:25:28 +00:00
Rich Graham	441fb9fb9e	checkpoint. This commit was SVN r17985.	2008-03-27 01:16:32 +00:00
Ralph Castain	cca449e379	Move an OMPI RML tag to the OMPI layer This commit was SVN r17950.	2008-03-25 13:30:48 +00:00
Ralph Castain	dc7f45dafd	Remove the obsolete and largely unused orte_system_info structure. The only fields that were used in that struct were nodeid and nodename - these have been transferred to the orte_process_info structure. Only one place used the user name field - session_dir, when formulating the name of the top-level directory. Accordingly, the code for getting the user's id has been moved to the session_dir code. This commit was SVN r17926.	2008-03-23 23:10:15 +00:00
Rich Graham	a7c836a2b0	fix location of the restrict key word. Make the tag in the fan-in/fan-out algorithm be fragment based. This commit was SVN r17903.	2008-03-21 01:40:36 +00:00
Rich Graham	2c66d396b7	take care of some bit-rot with the fanin-fanout method. This commit was SVN r17902.	2008-03-21 01:08:49 +00:00
Rich Graham	b9520e61dc	get the sm optimized allreduce working for all but user defined operations. Added to the reduction operations a set of reduction functions that take 2 input buffers and one output buffer to avoid some extra memory copies. These can't be used with user defined operations. The intel c collective suite passes both original, and new (new, not the user defined operations). This commit was SVN r17901.	2008-03-20 23:51:16 +00:00
Edgar Gabriel	570bbea5e0	fixing the allgather problem reported on the mailing list. The problem was that at one locatin we had the local-size instead of the remote size as a receive argument. This commit was SVN r17849.	2008-03-17 19:42:18 +00:00
Rich Graham	27182afb67	get the timers in correctly. This commit was SVN r17832.	2008-03-16 03:25:16 +00:00
Rich Graham	afcd1016fd	move temp buffer allocation out of the iteration loop - i.e. always use the same temp loop. The algorithm is rather synchronous already... This commit was SVN r17831.	2008-03-16 03:20:46 +00:00
Rich Graham	a1766b29f6	fix some barrier addressing errors. This commit was SVN r17830.	2008-03-15 22:46:19 +00:00
Rich Graham	0453e7d2f4	bug in management memory allocation - too much memory allocated. This commit was SVN r17829.	2008-03-15 18:12:20 +00:00
Rich Graham	3c2f1eb8bf	reduce the number of temp buffers used. This commit was SVN r17828.	2008-03-15 17:23:04 +00:00
Rich Graham	0f9d642d51	temp buffer pointers are computed when they are set up. A bit more efficient, but more important, it is much easier to play around with memory layout now. This commit was SVN r17827.	2008-03-15 16:36:35 +00:00
Rich Graham	e3e336b5ab	check point This commit was SVN r17826.	2008-03-15 13:31:21 +00:00
Rich Graham	ebcf928c24	add some diagnostics. This commit was SVN r17789.	2008-03-07 22:27:41 +00:00
Rich Graham	9131461511	move some test code to another machine. This commit was SVN r17785.	2008-03-07 19:18:02 +00:00
Rich Graham	c230b65543	fix a couple of bugs. Recursive doubling seems to be working. This commit was SVN r17777.	2008-03-07 02:51:38 +00:00
Rich Graham	70157166f9	checkpoint - compiles, now neeed to debug. This commit was SVN r17775.	2008-03-07 00:39:59 +00:00
Rich Graham	4eace9d020	starting to implement recursive doubling algorithm. This commit was SVN r17765.	2008-03-06 18:38:58 +00:00
Rich Graham	67ad9b6d6b	increase max data segments size. This commit was SVN r17677.	2008-03-02 19:11:09 +00:00
Rich Graham	53126fa7bd	add calls to opal_progress() This commit was SVN r17673.	2008-02-29 23:25:09 +00:00
Rich Graham	d37db14901	get the shared memory collectives working again with the new version of orte. This commit was SVN r17672.	2008-02-29 22:28:57 +00:00
Rich Graham	c253a7bda1	simplify the code abit. This commit was SVN r17664.	2008-02-29 03:55:12 +00:00
Rich Graham	1632d8b299	revert to an older (not previosly checked in) version to get around a regression. This commit was SVN r17663.	2008-02-29 03:12:12 +00:00
Rich Graham	827e8d877e	fix bug in node type, and some memory copy optimizations. This commit was SVN r17661.	2008-02-29 01:20:11 +00:00
Rich Graham	940d6732c9	remove compiler warnings. This commit was SVN r17656.	2008-02-28 22:01:19 +00:00
Rich Graham	2b5fab9d51	avoid 0 byte malloc. This commit was SVN r17653.	2008-02-28 21:11:42 +00:00
Rich Graham	4b26adef00	remove some debug output. This commit was SVN r17650.	2008-02-28 20:54:35 +00:00
Rich Graham	5df6c6d043	fix several race conditions. This commit was SVN r17645.	2008-02-28 19:40:19 +00:00
Ralph Castain	d70e2e8c2b	Merge the ORTE devel branch into the main trunk. Details of what this means will be circulated separately. Remains to be tested to ensure everything came over cleanly, so please continue to withhold commits a little longer This commit was SVN r17632.	2008-02-28 01:57:57 +00:00
Rich Graham	68aa691171	checkpoint work. This commit was SVN r17620.	2008-02-27 14:56:36 +00:00
Rich Graham	b4bbb70bb7	got it all, but for the mem copies. Also, need to make sure volatile declarations are all inplace, as well as memory barriers. This commit was SVN r17572.	2008-02-25 00:16:21 +00:00
Rich Graham	2d8c2420e8	checkpoint. This commit was SVN r17571.	2008-02-24 20:54:16 +00:00
Rich Graham	771584bff5	generate reduction tree. This commit was SVN r17569.	2008-02-24 03:25:40 +00:00
Rich Graham	b9bb78484d	a bit of omptimization. This commit was SVN r17528.	2008-02-20 16:19:49 +00:00
Rich Graham	09afc36f5f	correct addressing. This commit was SVN r17519.	2008-02-20 01:12:43 +00:00
Rich Graham	b87b15580c	fix memory allocation error. Initialize pointer. This commit was SVN r17514.	2008-02-19 20:01:42 +00:00
Rich Graham	1cd8a2e578	checkpoint - works for 2 procs, but not more. This commit was SVN r17477.	2008-02-17 05:21:58 +00:00
Rich Graham	8006927ae8	free buffer, rather than ask for another one, when done with the memory. This commit was SVN r17468.	2008-02-15 04:21:58 +00:00
Rich Graham	2277b47ab9	register mca_coll_sm2_allreduce_intra - function still does not do any reduction operations. This commit was SVN r17467.	2008-02-15 04:13:00 +00:00
Rich Graham	9b0687e6df	add buffer allocation and deallocation calls to the allreduce routine, so I can start debugging the memory management code. The allreduce fucntion does nothing at this stage. This commit was SVN r17466.	2008-02-15 03:59:14 +00:00
Rich Graham	41943dbd76	adding missing files. This commit was SVN r17462.	2008-02-15 00:59:28 +00:00
Rich Graham	41f4b06b39	buffer allocate/release code is fully written, and compiles. Now need to debug. This commit was SVN r17461.	2008-02-15 00:57:44 +00:00
Rich Graham	7cc58768cd	checkpoint something that compiles This commit was SVN r17460.	2008-02-15 00:33:14 +00:00
Rich Graham	292d930eea	check point. This commit was SVN r17457.	2008-02-14 20:00:26 +00:00
Edgar Gabriel	77057a50a3	- adding the two-level hierarchy detection algorithm - minor fix in the temporary collectives - removing the symmetric parameter, since it didn't really make sense. This commit was SVN r17359.	2008-02-01 17:11:36 +00:00
Rich Graham	fda485ff9c	backing file is allocated and deallocated. This commit was SVN r17358.	2008-02-01 15:26:20 +00:00
Rich Graham	165fc3f8cc	memory allocation implemented and debugged. Still need to finish file allocation/dealocation and control information initialization. This commit was SVN r17291.	2008-01-29 03:09:12 +00:00
Rich Graham	e24c2ebbc0	have a working skeleton for the SM-V2 component. It does nothing at this stage. This commit was SVN r17241.	2008-01-25 21:16:36 +00:00
Rich Graham	1d0334f4f2	skeleton for new shared memory collective component. This commit was SVN r17235.	2008-01-25 19:35:26 +00:00
Rich Graham	432ba0cecd	add comments about the life-cycle of a collective module. This commit was SVN r17223.	2008-01-25 03:46:31 +00:00
George Bosilca	31390c0074	We should take in account the extent of the datatype when we compute the initial displacement in bytes. Thanks to Daniel G. Hyams for the fix. This commit was SVN r17165.	2008-01-19 05:34:53 +00:00
George Bosilca	3fca3973d3	The PTLs are now long gone !!! This commit was SVN r17104.	2008-01-10 00:18:45 +00:00
George Bosilca	906e8bf1d1	Replace the ompi_pointer_array with opal_pointer_array. The next step (sometimes after the merge with the ORTE branch), the opal_pointer_array will became the only pointer_array implementation (the orte_pointer_array will be removed). This commit was SVN r17007.	2007-12-21 06:02:00 +00:00
Jeff Squyres	213b5d5c6e	Per long threads on the mailing list and much confusion discussion about linkers, have all OPAL, ORTE, and OMPI components '''not'' link against the OPAL, ORTE, or OMPI libraries. See ttp://www.open-mpi.org/community/lists/users/2007/10/4220.php for details (or https://svn.open-mpi.org/trac/ompi/wiki/Linkers for a better-formatted version of the same info). This commit was SVN r16968.	2007-12-15 13:32:02 +00:00
Andrew Friedley	c15047b264	Add LLNL copyright to the file i modified yesterday This commit was SVN r16404.	2007-10-09 15:18:23 +00:00
Andrew Friedley	fd51d9cf28	The call to opal_list_insert() had an off by one error (I think), causing selected components to get lost with certain load orderings. I went ahead and rewrote the code to use opal_list_insert_pos() instead, which gives a cleaner flow and more speed. This commit was SVN r16392.	2007-10-08 23:01:36 +00:00
Jeff Squyres	f92d9097d8	Some more changes to update to coll v1.1.0 that were missed yesterday. This actually exposed a very, very long-standing bug where part of the coll base was incorrectly checking the coll API version against the MCA API version. When coll went to v1.1 (yesterday) and was no longer the same as the MCA v1.0, the test started failing. This commit fixes to check for v1.1 everywhere in the coll base, and to ensure to check coll framework/API version numbers against coll framework/API version numbers (vs. against the MCA API version number). This commit was SVN r16373.	2007-10-07 12:20:22 +00:00
Jeff Squyres	3d34bff596	No technical/functional changes: simply change the name of the "data" parameter to "module" everywhere, just to be a little more clear what the purpose of that parameter is. This commit was SVN r16372.	2007-10-07 08:36:45 +00:00
Jeff Squyres	fc2b4376e9	Update forgotten macro. This commit was SVN r16368.	2007-10-06 14:11:35 +00:00
Jelena Pjesivac-Grbovic	ada43fef9e	This fixes bug #1157 in coll/self module. All vector functions had incorrect handling of the offset. This commit was SVN r16360.	2007-10-05 17:40:16 +00:00
Andrew Friedley	2e66590993	Fix mistakes in the basic component.. can't call collectives on the communicator and always pass the basic module.. have to give them the module off the communicator. This commit was SVN r16329.	2007-10-04 16:29:24 +00:00
George Bosilca	1e7a791349	Remove some of the problems identified by Coverty. This commit was SVN r16112.	2007-09-12 20:13:26 +00:00
George Bosilca	c755938eb0	Coverty: release the temporary buffer on error. This commit was SVN r16104.	2007-09-12 17:45:12 +00:00
Shiqing Fan	a0660f4deb	- Just some type casts. This commit was SVN r16100.	2007-09-12 15:29:58 +00:00
Jeff Squyres	c4a38f47f6	Resolve Coverity CID 467: remove unused variable / dead code. This commit was SVN r15997.	2007-08-29 01:23:18 +00:00
Edgar Gabriel	a2f5cada1a	convert the hiearch component to the new structure. More testing required before we remove the .ompi_ignore flag again. This commit was SVN r15954.	2007-08-23 20:41:29 +00:00
Shiqing Fan	a497a3fcad	- Fix some small bugs, copy-paste mistakes. This commit was SVN r15941.	2007-08-21 19:57:28 +00:00
Sven Stork	3985a35c35	- export required symbol This commit was SVN r15939.	2007-08-21 18:46:11 +00:00
Brian Barrett	af4e86c25f	Update collectives selection logic to allow for multiple components to be used at nce (up to one unique collective module per collective function). Matches r15795:15921 of the tmp/bwb-coll-select branch This commit was SVN r15924. The following SVN revisions from the original message are invalid or inconsistent and therefore were not cross-referenced: r15795 r15921	2007-08-19 03:37:49 +00:00
Jelena Pjesivac-Grbovic	9bd9c92dbd	Making sure that the decision function for scatter and gather correctly computes everything for MPI_IN_PLACE case. This commit was SVN r15841.	2007-08-13 17:35:50 +00:00
Jelena Pjesivac-Grbovic	b558e820cb	removing compiler wraning This commit was SVN r15803.	2007-08-08 15:22:01 +00:00
Jelena Pjesivac-Grbovic	daa10b277e	modifying scatter decision function to use binomial algorithm for small message sizes. This commit was SVN r15798.	2007-08-07 22:16:13 +00:00
Mohamad Chaarawi	59a7bf8a9f	Merging in the Sparse Groups.. This commit includes config changes.. This commit was SVN r15764.	2007-08-04 00:41:26 +00:00
Sven Stork	855434de59	- fixes several coverty issues - add missing initialisation for variables - use strncpy instead of strcpy This commit was SVN r15683.	2007-07-30 14:44:37 +00:00
Jelena Pjesivac-Grbovic	1b66a52c50	Modifying type of binomial tree used for binomial reduce: switching: 0 0 / \ \ / \ \ 1 \ \ --> 4 \ \ / \ \ / \ \ 3 2 \ 3 2 \ 4 1 (duh). The first form is the bmtree suitable for bcast, but the latter is better for reduce. Updating default decision function accordingly. This commit was SVN r15422.	2007-07-13 21:07:51 +00:00
Jelena Pjesivac-Grbovic	d677db9b5f	cleaning up alltoall implementation: - removing MPI_* calls from bruck implementation - simplifying 2 process case - identation, etc. This commit was SVN r15301.	2007-07-07 01:06:19 +00:00
Jelena Pjesivac-Grbovic	483222085e	Fixing compiler warnings. In gather, the ptmp += incr is irrelevant, since ptmp is set within the loop. This commit was SVN r15293.	2007-07-05 20:40:50 +00:00
Jelena Pjesivac-Grbovic	3b0a52a104	adding tuned allgatherv implementation using bruck, ring, and neighbor-exchange algorithms. The implementations passed intel and imb tests up to 40 processes. This commit was SVN r15280.	2007-07-03 23:33:12 +00:00
Jelena Pjesivac-Grbovic	d55b415bb0	fixing typo This commit was SVN r15240.	2007-06-28 20:56:55 +00:00
Jelena Pjesivac-Grbovic	8fc8b44d11	Modifying reduce decision function for large, single element reduces (again). Binary algorithm without segmentation tends to outperform binomial algorithm in this case. This commit was SVN r15226.	2007-06-27 22:01:56 +00:00
Jelena Pjesivac-Grbovic	0ecef1750d	Modifying the default reduce decision function to use binomial algorithm for single-element reduce (segmented algorithms make no sense in this case and can cause performance degradation). This commit was SVN r15209.	2007-06-26 20:14:03 +00:00

... 5 6 7 8 9 ...

915 Коммитов