I have run an experiment where in a series of trials, each of 10 animals was presented with each of 6 stimuli, and each stimulus was offered once to the left and once to the right nostril. Overall, each animal participated in 12 trials (6 stimuli x 2 nostrils). My point is to calculate the reliability of the data by measuring the correlation between the observations from the left vs right nostril. Previously, I used Pearson's r, but a reviewer pointed out that it is unsuitable for repeated measurements and non-normally distributed data. Could you advise me on a test I could use instead of Pearson? I would be very grateful - I tried to check by myself but with no success.

Regarding details: I measured the frequency (in numbers) and duration (in seconds) of chosen behaviors (e.g., stomping - it was about horses). So, my goal was to simply correlate the frequency of behaviors in a given category when presented on the left with the frequency of the same behaviors when presented on the right (e.g., frequency of stomping when the stimulus was presented on the left with the frequency of the same behavior when presented on the right). I wanted to do it for all behavioral categories, separately for frequency and duration.