Thank you TBP, Dave & Alex for you suggestions,
I have duly increased my LRUS to 128, and CLEANERS & NUMAIOVPS after
looking at onstat -g iov and checking the io/wup column.
We've found the culprit which affects all hosts connected to the same
Shark (and there are quite a few) - it was a SQLServer box doing a large
delete. How come the Shark misbehaves so badly due to one particular
operation (ie. SQLServer delete, it's not volume related, the SAN is
quiet and the reports on the Shark don't show any high throughput) is
beyond me - and apparently beyond local support as well since we've
struggled with this since November last year and we've blamed everything
from the usual firmware thing to backups and what not (even found some
new entries for my BOFH DIY excuse board).
Sooo, you do something quite innocent on one box and you bring all
critical systems in your entire enterprise to a grinding halt, and
worse, you're not even able to tell. Now isn't that what mass storage is
all about :-...
Kind regards,
-------------------------------------------
Willem Roos - (+27) 21 980 4941
Per sercas vi malkovri
Disclaimer
http://www.shoprite.co.za/disclaimer.html
sending to informix-list