Has anyone run into any isi_job_d issues? I'm seeing issues with the jobs hung on our cluster and I cannot cancel them. I've actually disabled and re-enabled the job service to no degree. EMC mentioned a threadcount issue increase to 240 which I did, but since I cannot restart the isi_job_d service, this has no affect. The only option I can see is to perform a rolling reboot (which I hate to do during production hours).
Nah, I checked. The message is saying the job stopped but every time I check it's already been restarted. I do what to check the PAPI option to see if that is the problem.
EMC said the thread count for FSAnalyze needed to be bumped up to 240 but as I check this, it appears the problem went away.
I HATE issues that fix themselves and I have no clue WHY!
Yeah, that's where EMC told me to stop the isi_job_d service, change the thread count to 240 then restart the service. I did that, still got the error until between the time I did that, opened this question and then responded to you -- it fixed itself.
I'm dumbfounded. EMC has the logs though and that is where they told me that FSAnalyzer will need the threadcount increased on OneFS 8.0.0.4. That should be a part of the post upgrade config, no?
Eric_W1
33 Posts
5131
1
Posted June 5th, 2017 06:00
Connection refused? Could this be a PAPI issue at this point?