Sigmoid is used for binary classification at the output layer.
ReLU is mostly used at the hidden layers. (It does not make sense to be used at the output layer.)
Same goes for tanh
Linear is an appropriate activation function for the output layer in a regression problem. Typically, this is just the identity function; that is, $g(z)=z$.