Robust Environmental Sound Recognition With Sparse Key-Point Encoding and Efficient Multispike Learning