openmpi

Автор	SHA1	Сообщение	Дата
Tim Woodall	4fc5b2105a	this is currently an int - we shouldn't restrict it unless required This commit was SVN r7895.	2005-10-27 17:06:58 +00:00
Tim Woodall	13409ec53b	correction for hang, check for additional fragments before callback, which may queue a new fragment This commit was SVN r7889.	2005-10-27 01:39:39 +00:00
Graham Fagg	5bb0d4a053	enable allreduce to be selected This commit was SVN r7888.	2005-10-26 23:55:37 +00:00
Graham Fagg	2587d7ade9	added some more linear functions. minor corrections on naming and debug info This commit was SVN r7887.	2005-10-26 23:51:56 +00:00
Graham Fagg	c3e1dc410d	Started to add basic linear functions Also started to add the allreduce algorithms as I test them (i.e. if it goes in its after testing from now on) This commit was SVN r7886.	2005-10-26 23:11:32 +00:00
Jeff Squyres	d47ce065e9	Minor Makefile.am fix for static builds. This commit was SVN r7882.	2005-10-26 15:57:58 +00:00
Edgar Gabriel	ba3bf6592f	fixing some warnings. No idea yet why the static builds fail... This commit was SVN r7879.	2005-10-26 12:56:56 +00:00
Jeff Squyres	23ab9e0277	A better solution to the previous commit -- RETAIN/RELEASE the MPI_Op at the top-level MPI API function. This allows two kinds of scenarios: 1. MPI_Ireduce(..., op, ...); MPI_Op_free(op); MPI_Wait(...); For the non-blocking collectives that we're someday planning -- to make them analogous to non-blocking point-to-point stuff. 2. Thread 1: MPI_Reduce(..., op, ...); Thread 2: MPI_Op_free(op); Granted, for #2 to occur would tread a fine line between a correct and erroneous MPI program, but it is possible (as long as the Op_free was after MPI_reduce() had started to execute). It's more realistic with case #1, where the Op_free() could be executed in the same thread or a different thread. This commit was SVN r7870.	2005-10-25 19:20:42 +00:00
Edgar Gabriel	d009d8de57	opening the hierarchical collective component to the public. I am at this stage fairly confident that - it works in most scenarious (with symmetric hierarchies, with asymmetric hierarchies, wihout hierarchies - it just removes itself) - it does not create too many problems (I am not aware of any at least) - it does not slow down startup anymore dramatically (thanks to the fixes of Brian, Jeff, Tim and a significant reduction in the number of collective operations in the comm_query) Any feedback is highly welcome. This commit was SVN r7868.	2005-10-25 18:38:43 +00:00
Edgar Gabriel	00c04ab56a	moving the hierarch collective component to the new parameter registration interface. This commit was SVN r7867.	2005-10-25 18:34:47 +00:00
Edgar Gabriel	3633605010	moving op_init further up in ompi_mpi_init, since it is required when quering some of the collective components. Up to now, it just worked somehow, but now with correct reference counting for ops in place, it refused :-) This commit was SVN r7866.	2005-10-25 18:33:48 +00:00
Jeff Squyres	ef09e768e0	Ensure to OBJ_DESTRUCT to free memory during finalize (caught by Brian). This commit was SVN r7864.	2005-10-25 17:27:58 +00:00
Jeff Squyres	f8fd10715c	- Minor style fix - Be sure to properly OBJ_CONSTRUCT the intrinsic MPI_Op's - RETAIN/RELEASE the op's when used in the invoke function This commit was SVN r7863.	2005-10-25 16:24:00 +00:00
Jeff Squyres	a9f04c7573	Only do the extra va_* stuff if we're compiling with the compiler that cares about it (PGI). This commit was SVN r7860.	2005-10-25 13:08:52 +00:00
Graham Fagg	382f05c7ad	Infastructure changes. started to add static (fixed if) statement based decision rules based on gigE numbers added mca params so that a user can force a certain algorithm/segment/topo on a per collective basis (this is not in the fixed call path but only in the dynamic (at com create) call path). (these params can be used by test suites such as OCC to choice which algorithm they are using). This commit was SVN r7854.	2005-10-25 03:55:58 +00:00
Graham Fagg	d8e32464cb	ops. setting/reading mca option from the right varible would help. This commit was SVN r7850.	2005-10-24 21:33:48 +00:00
Brian Barrett	1e2f7d6a3d	* make sure to expose ompi_op_t as an object This commit was SVN r7848.	2005-10-24 20:31:14 +00:00
Rainer Keller	d6120d32d6	- Only minor white-space changes, to clean up This commit was SVN r7843.	2005-10-24 10:36:16 +00:00
George Bosilca	b45651988b	Protect against elements with ZERO length. Remove all the useless code. This commit was SVN r7827.	2005-10-21 06:48:51 +00:00
George Bosilca	1fb8ec646a	Add the homogeneous flag back in the convertor. Correct/improve one of the comments. Descrease the amount of memory required for the stack. This commit was SVN r7826.	2005-10-21 06:47:57 +00:00
Galen Shipman	cb84a57c57	add endpoint and srq flow-control.. Note, we are failing the ring tests in the intel p2p test suite, but we seem to fail the same tests under the current trunk.. will look into this further. This commit was SVN r7823.	2005-10-21 02:21:45 +00:00
Galen Shipman	c4889ac759	update openib mpool to properly deregister and release (carry over from mvapi). Still need to add endpoint and srq flow control as in mvapi This commit was SVN r7816.	2005-10-20 03:57:17 +00:00
Galen Shipman	0d1d231169	convert to new mca params, adding description strings. changed mca param rr_buf_min/max to rd_min/max Add bandwidth param to openib This commit was SVN r7815.	2005-10-20 02:55:21 +00:00
George Bosilca	75bc3dd43c	Dont mess around with the OBJ_DESTRUCT on the communicator. It's quicker (and safer) to call directly the communicator cleanup function (ompi_convertor_cleanup). This commit was SVN r7814.	2005-10-19 21:28:52 +00:00
George Bosilca	1d75b7972f	Solve thee problem with the reference count on the datatype (RT bug 1492). The problem is that the convertor (when prepared) increase the reference count on the used datatype. This reference count will be released only when the OBJ_DESTRUCT is called on a convertor. However, having to call OBJ_CONSTRUCT and OBJ_DESTRUCT on each request every time we want to use it (even when it come from the cache) is an expensive operation. This can be avoided is the OBJ_DESTRUCT will leave the convertor in exactly the same state as OBJ_CONSTRUCT. With this approach we just have to call OBJ_CONSTRUCT for each convertor once when we initially create the request. This commit was SVN r7813.	2005-10-19 20:57:39 +00:00
George Bosilca	63c5013fe6	After a OBJ_DESTRUCT a convertor has to be in a usable state. Read the comment for more informations. This commit was SVN r7812.	2005-10-19 20:51:52 +00:00
George Bosilca	8987bcabe2	Remove the memcpy we can do it as we parse the datatypes in order to increase their references. This commit was SVN r7811.	2005-10-19 20:51:11 +00:00
George Bosilca	6c6f17628f	Remove a double OMPI_DECLSPEC from the definition of one of the predefined data-types. This commit was SVN r7810.	2005-10-19 20:50:25 +00:00
Brian Barrett	de5e501519	Rather than hard spinning waiting for something to happen when doing shared memory initialization, call opal_progress() to push any pending events around and possibly yield the processor if nothing entertaining is happening. This should probably go to the 1.0 branch. This commit was SVN r7808.	2005-10-19 00:56:14 +00:00
George Bosilca	d2f831cd18	Construct the convertor attached to the receive request. This should happens only on the first allocation of a request object. This commit was SVN r7807.	2005-10-18 21:53:05 +00:00
Brian Barrett	bcebd1b6b7	Fix a couple of places where headers didn't get installed correctly when --with-devel-headers is given to configure: * allocator, rcache, and mpool were putting things in the wrong place * timer wasn't installing the inline implementations at all This commit was SVN r7805.	2005-10-18 20:12:55 +00:00
Edgar Gabriel	3a7efaf4d9	fix for reduce and allreduce for an unsymmetric case This commit was SVN r7802.	2005-10-18 19:20:48 +00:00
Edgar Gabriel	818b4af554	- reverting the logic in the hierarchy detection stuff. This can reduce the number of collective operations and simplifies the logic significantly. - introducing a special case if size of comm == 1, avoiding thus collective operations as well ( i.e. no need for hierarchies) - fix for an unsymmetric case. Still to be tested. This commit was SVN r7799.	2005-10-18 18:17:50 +00:00
Tim Woodall	b570c8cad4	need to specify a size, base address will match This commit was SVN r7798.	2005-10-18 17:01:36 +00:00
Galen Shipman	4d2d39b0a6	intial checking of SRQ flow control support for mvapi This commit was SVN r7796.	2005-10-18 14:55:11 +00:00
Jeff Squyres	f9974f72e0	construct/destruct convertor when requests are constructed and allocated to free lists This commit was SVN r7791.	2005-10-18 12:19:43 +00:00
Jeff Squyres	a459659a33	Print the string name of the return code This commit was SVN r7789.	2005-10-17 20:47:44 +00:00
Galen Shipman	3efecaaeda	convert openib btl to use new mca_param registration.. Also, change rr_buf_min and rr_buf_max to rd_min and rd_max This commit was SVN r7786.	2005-10-17 20:00:34 +00:00
Tim Woodall	c944988b9e	merge in changes from release branch - acquire/release send token for put/get This commit was SVN r7784.	2005-10-17 18:59:28 +00:00
Jeff Squyres	89931ac05f	- Correct typo in comment - Add DIST_SUBDIRS to ompi/tools/Makefile.am This commit was SVN r7780.	2005-10-17 11:55:55 +00:00
Brian Barrett	1302cb4072	The next in a long line of crazed build system changes from Brian. This was originally suggested by Ralf Wildenhues, to try to speed autogen, configure, and make (and possibly even make install). Use automake's include directive to drastically reduce the number of Makefile files (although the number of Makefile.am files is the same - most are just included in a top-level Makefile.am). Also use an Automake SUBDIRs feature to eliminate the dynamic-mca tree, which was no longer really needed. This makes adding a framework easier (since you don't have to remember the dynamic-mca tree) and makes building faster (as make doesn't have to recurse through the dynamic-mca tree) This commit was SVN r7777.	2005-10-17 00:21:10 +00:00
George Bosilca	6e3c23ec3b	Do not allow the use of the optimized path for predefined non contiguous datatypes (like MPI_SHORT_INT on most of the architectures). This commit was SVN r7776.	2005-10-16 19:41:40 +00:00
Edgar Gabriel	7e45f64065	reduce has now been tested quite extensively for all (predefined) operations and for all root nodes and passed all tests. First cut on barrier (which from my perspective does not make sense from the performance point of view) and on allreduce (which might make sense), This commit was SVN r7774.	2005-10-15 22:24:44 +00:00
Edgar Gabriel	3fab9c628c	switching the root and creating (if necessary) the new local leader sub-communicators seems to work as well. Thoroughly tested with bcast, not yet that exhaustivly tested for the reduction. This commit was SVN r7773.	2005-10-15 21:13:44 +00:00
Edgar Gabriel	7d34770456	further bugfixes. The hierarchy detection works now as far as I can see (even in unsymmetric sitations). Bcast and reduce work as well. Still to test: the code which generates new local leader communicators, in case the root of the operation is not yet part of the lleader comm. This commit was SVN r7772.	2005-10-15 19:36:54 +00:00
Edgar Gabriel	63554d245f	further bugfixes This commit was SVN r7771.	2005-10-15 18:44:57 +00:00
Edgar Gabriel	92c7b77cbc	minor bug fixes This commit was SVN r7770.	2005-10-15 18:32:40 +00:00
Edgar Gabriel	ba163c611c	checkpoint before moving to a real cluster. Most of the recoding should be done. This version also doesn't break ompi (at least if its not chosen :-) ). New features compared to the version from last Thursday (where bcast and reduce seemed to work in most scenarios): - clearer internal infrastructure - ability to handle all root processes with a (hopefully) minimal number of local leader communicators. This commit was SVN r7769.	2005-10-15 17:04:01 +00:00
Jeff Squyres	e097ee635a	Silence compiler warnings. This commit was SVN r7768.	2005-10-14 22:06:25 +00:00
Jeff Squyres	237bd4c6cd	Fix ompi_info -- cxx:bindings was somehow hard-coded to "yes" instead of reflecting whether the C++ bindings were supported or not. This commit was SVN r7766.	2005-10-14 20:07:05 +00:00

1 2 3 4 5 ...

701 Коммитов