TAU¶
Description¶
TAU Performance System® is a portable profiling and tracing toolkit for performance analysis of parallel programs written in Fortran, C, C++, Java, Python. TAU (Tuning and Analysis Utilities) is capable of gathering performance information through instrumentation of functions, methods, basic blocks, and statements. All C++ language features are supported including templates and namespaces. The API also provides selection of profiling groups for organizing and controlling instrumentation. The instrumentation can be inserted in the source code using an automatic instrumentor tool based on the Program Database Toolkit (PDT), dynamically using DyninstAPI, at runtime in the Java Virtual Machine, or manually using the instrumentation API. This module defines the following defaults: TAU_TRACE=1, TAU_CALLPATH=1, TAU_PROFILE=1, TAU_MAKEFILE=Makefile.tau-callpath-icpc-papi-pdt-profile-trace, TAU_METRICS=TIME:PAPI_FP_OPS:PAPI_L2_DCM. For normal usage use the defaults: just load up tau, recompile, and run. See the available for commands (e.g. paraprof, tauf90, etc.), and the application program interface. To compile your code with TAU, use one of the TAU compiler wrappers: tau_f90.sh tau_cc.sh tau_cxx.sh for constructing an instrumented code (instead of icc, ifort, etc.). E.g. tau_f90.sh hello.f90, tau_cc.sh hello.c, etc. These may also be used in make files, using macro definitions: E.g. F90=tau_f90.sh, CC=tau_cc.sh, Cxx=tau_cxx.sh.
Environment Modules¶
Run module spider tau
to find out what environment modules are available for this application.
Environment Variables¶
pdf Quickstart and User Guide in the $HPC_TAU_DOC directory. Man pages are * HPC_TAU_DIR - installation directory * HPC_TAU_BIN - executable directory * HPC_TAU_LIB - library directory * HPC_TAU_DOC - examples directory * HPC_PDT_DIR - installation directory * HPC_PDT_BIN - executable directory * HPC_PDT_INC - includes directory
Citation¶
If you publish research that uses {{{app}}} you have to cite it as follows:
- TAU: The TAU Parallel Performance System, by S. Shende and A. D. Malony. International Journal of High Performance Computing Applications, Volume 20 Number 2 Summer 2006. Pages 287-311.
- PDT: A Tool Framework for Static and Dynamic Analysis of Object-Oriented Software with Templates, by K. A. Lindlan, J. Cuny, A. D. Malony, S. Shende, B. Mohr, R. Rivenburgh, C. Rasmussen. Proceedings of SC2000: High Performance Networking and Computing Conference, Dallas, November 2000.
- CCA: Performance Technology for Parallel and Distributed Component Software, by A. Malony, S. Shende, N. Trebon, J. Ray, R. Armstrong, C. Rasmussen, and M. Sottile. Concurrency and Computation: Practice and Experience, Vol. 17, Issue 2-4, pp. 117-141, John Wiley & Sons, Ltd., Feb - Apr, 2005.
Categories¶
library, performance_analysis