In one embodiment, a method determines a first local binary pattern for a first image in a video and a second local binary pattern for a second image in the video. The roller tappets comprising a substantially cylindrical tappet body are retained in receptions of the mounting aid that is made of plastic. When the text data is output from each of the speech recognizing units, the text data is associated with the information indicating each speaker and the text data is displayed. The transceiver additionally includes a plurality of subtraction circuits.