Help!
Posted in 1999
Topics: High Availability & Replication, Error Codes & Troubleshooting, Logging & Checkpoints
Dear all.
Can anyone help me, what happen with our Informix Server!
Thanks in advance!
21:55:17 Checkpoint Completed: duration was 0 seconds.
21:58:39 CDR connection to server lost, id 3, name <PI>Reason: Connect failed reason: -908
21:58:39 CDR: Re-connected to server, id 3, name <PI>
21:58:41 Assert Failed: No Exception Handler
21:58:41 Informix Dynamic Server Version 7.31.TC2
21:58:41 Who: Session(143629, informix@in4mix.panpi.com.cn, 0, 0)
Thread(199152, CDRD_4, 0, 1)
Exception at Addr: 0x70ff77, TOS: 0x2697904c, FP:0x26979050, Exc: c0000005
21:58:41 Results: Exception Caught. Type: MT_EX_OS, Context: mem
21:58:41 Action: Please notify Informix Technical Support.
21:59:03 See Also: af.dd81f11, shmem.dd81f11.0
22:00:13 \
As it says, call Informix Tech Support.
paul wrote in message <849m2b$i4k@netnews.hinet.net>...
>Dear all.
>Can anyone help me, what happen with our Informix Server!
>Thanks in advance!
>
>21:55:17 Checkpoint Completed: duration was 0 seconds.
>21:58:39 CDR connection to server lost, id 3, name <PI>>Reason: Connect failed reason: -908
>21:58:39 CDR: Re-connected to server, id 3, name <PI>
>21:58:41 Assert Failed: No Exception Handler
>21:58:41 Informix Dynamic Server Version 7.31.TC2
>21:58:41 Who: Session(143629, informix@in4mix.panpi.com.cn, 0, 0)
> Thread(199152, CDRD_4, 0, 1)
> Exception at Addr: 0x70ff77, TOS: 0x2697904c, FP:0x26979050, Exc:c0000005
>21:58:41 Results: Exception Caught. Type: MT_EX_OS, Context: mem
>21:58:41 Action: Please notify Informix Technical Support.
>21:59:03 See Also: af.dd81f11, shmem.dd81f11.0
>22:00:13 \
You lost the connection to one of this machines' replication partners due to
that machine going offline or the network being unavailable. The engine
paniced which should not have happened. I'd say call tech support and most
likely look to upgrade.
Art S. Kagel
paul wrote:
>
> Dear all.
> Can anyone help me, what happen with our Informix Server!
> Thanks in advance!
>
> 21:55:17 Checkpoint Completed: duration was 0 seconds.
> 21:58:39 CDR connection to server lost, id 3, name <PI>> Reason: Connect failed reason: -908
> 21:58:39 CDR: Re-connected to server, id 3, name <PI>
> 21:58:41 Assert Failed: No Exception Handler
> 21:58:41 Informix Dynamic Server Version 7.31.TC2
> 21:58:41 Who: Session(143629, informix@in4mix.panpi.com.cn, 0, 0)
> Thread(199152, CDRD_4, 0, 1)
> Exception at Addr: 0x70ff77, TOS: 0x2697904c, FP:0x26979050, Exc: c0000005
> 21:58:41 Results: Exception Caught. Type: MT_EX_OS, Context: mem
> 21:58:41 Action: Please notify Informix Technical Support.
> 21:59:03 See Also: af.dd81f11, shmem.dd81f11.0
> 22:00:13 \
In article <3868CFE5.ADFFE25E@bloomberg.net>, Art S. Kagel
<kagel@bloomberg.net> writes
>You lost the connection to one of this machines' replication partners due to
>that machine going offline or the network being unavailable. The engine
>paniced which should not have happened. I'd say call tech support and most
>likely look to upgrade.
>
Depends what the af file af.dd81f11 says.
>Art S. Kagel
>
>paul wrote:
>>
>> Dear all.
>> Can anyone help me, what happen with our Informix Server!
>> Thanks in advance!
>>
>> 21:55:17 Checkpoint Completed: duration was 0 seconds.
>> 21:58:39 CDR connection to server lost, id 3, name <PI>>> Reason: Connect failed reason: -908
>> 21:58:39 CDR: Re-connected to server, id 3, name <PI>
>> 21:58:41 Assert Failed: No Exception Handler
>> 21:58:41 Informix Dynamic Server Version 7.31.TC2
>> 21:58:41 Who: Session(143629, informix@in4mix.panpi.com.cn, 0, 0)
>> Thread(199152, CDRD_4, 0, 1)
>> Exception at Addr: 0x70ff77, TOS: 0x2697904c, FP:0x26979050, Exc: c0000005
>> 21:58:41 Results: Exception Caught. Type: MT_EX_OS, Context: mem
>> 21:58:41 Action: Please notify Informix Technical Support.
>> 21:59:03 See Also: af.dd81f11, shmem.dd81f11.0
>> 22:00:13 0 >> 22:00:15 PANIC: Attempting to bring system down
--
David Williams
If this were pre 7.31 I'd say that the failure occured because of a race
condition that occured when you lost/reconnected your network connection to
serverID 3. But this is 7.31.
In all honesty I rarely see ER failures with one of the DataSync threads, but
there are a couple of possibilities. One would be if an inplace alter had been
done and had somehow made it to a target system. However if that had occured,
you would not have been able to get the system up without first manually
removing the offending transaction. Another possibility would be Bug 116187
(fixed in UC5) that had to do with replication of blobs and threads whose
threadID is larger than 99999. I'm going to guess that this is what you
encountered, but without a stack trace it is impossible to say for sure. Since
your threadID is larger than 99999, I'm going to guess that B116187 is what
caused your problem. I'm also going to guess that the problem did not cause the
engine to subsequently abort.
Please open a case with tech support and give them this email as well. It should
help with the problem resolution.
paul wrote:
> Dear all.
> Can anyone help me, what happen with our Informix Server!
> Thanks in advance!
>
> 21:55:17 Checkpoint Completed: duration was 0 seconds.
> 21:58:39 CDR connection to server lost, id 3, name <PI>> Reason: Connect failed reason: -908
> 21:58:39 CDR: Re-connected to server, id 3, name <PI>
> 21:58:41 Assert Failed: No Exception Handler
> 21:58:41 Informix Dynamic Server Version 7.31.TC2
> 21:58:41 Who: Session(143629, informix@in4mix.panpi.com.cn, 0, 0)
> Thread(199152, CDRD_4, 0, 1)
> Exception at Addr: 0x70ff77, TOS: 0x2697904c, FP:0x26979050, Exc: c0000005
> 21:58:41 Results: Exception Caught. Type: MT_EX_OS, Context: mem
> 21:58:41 Action: Please notify Informix Technical Support.
> 21:59:03 See Also: af.dd81f11, shmem.dd81f11.0
> 22:00:13 \