Problem with latin Characters
Posted in 2016
Topics: Migration, Import/Export & Data Conversion, Internationalization & Character Sets
Hi to all
I need to correct in an IDS database a lot of latin characters that appear
incorrectly.
Those characters were initially inserted using characterset UTF-8 but after
that they were exposed (imported and exported) to en_US_819. Now the database
is using UTF-8 and a lot of information were introduced before the error were
discovered.
After export with dbexport I tried to look inside the unl files using ASCII in
Notepad++ and I can see some correspondence like the following:
ã = ã
ç = ç
ã = ã
õ = õ
á = á
é = é
à = í
ó = ó
ê = ê
â = Â
I'm thinking to replace it in the unl files and sql file of the export using
Notepadd++ and after do a dbimport using UTF-8 coding.
Do anyone have another better idea?
Thanks in advance
Best Regards
Hello Pedro:
Did you try 'iconv'? iconv is a linux command tool for this purpoose.
We use this tool before load unl files in order to transform from utf-8 yo
iso....., or whatever.
http://www.gnu.org/savannah-checkouts/gnu/libiconv/documentation/libiconv-1.13/i
conv.1.html
Regards.
2016-04-06 9:29 GMT+01:00 PEDRO NEVES <pedroivomz@gmail.com>:
> Hi to all
>
> I need to correct in an IDS database a lot of latin characters that appear
> incorrectly.
>
> Those characters were initially inserted using characterset UTF-8 but after
> that they were exposed (imported and exported) to en_US_819. Now the
> database
> is using UTF-8 and a lot of information were introduced before the error
> were
> discovered.
>
> After export with dbexport I tried to look inside the unl files using
> ASCII in
> Notepad++ and I can see some correspondence like the following:
>
> ã = ã
> ç = ç
> ã = ã
> õ = õ
> á = á
> é = é
> Ã = í
> ó = ó
> ê = ê
> â = Â
>
> I'm thinking to replace it in the unl files and sql file of the export
> using
> Notepadd++ and after do a dbimport using UTF-8 coding.
>
> Do anyone have another better idea?
>
> Thanks in advance
>
> Best Regards
>
>
>
>
*******************************************************************************
> Forum Note: Use "Reply" to post a response in the discussion forum.
>
>
--089e01536f8c864849052fcced35
Pedro: I think that your suggestion of export-replace-import is probably
the best option. You could do the replacements in place in the database but
it would be significantly slower than using awk or Perl etc. to patch the
unload files.
Art
Art S. Kagel, President and Principal Consultant
ASK Database Management
www.askdbmgt.com
Blog: http://informix-myview.blogspot.com/
Disclaimer: Please keep in mind that my own opinions are my own opinions
and do not reflect on the IIUG, nor any other organization with which I am
associated either explicitly, implicitly, or by inference. Neither do
those opinions reflect those of other individuals affiliated with any
entity with which I am affiliated nor those of the entities themselves.
On Wed, Apr 6, 2016 at 4:29 AM, PEDRO NEVES <pedroivomz@gmail.com> wrote:
> Hi to all
>
> I need to correct in an IDS database a lot of latin characters that appear
> incorrectly.
>
> Those characters were initially inserted using characterset UTF-8 but after
> that they were exposed (imported and exported) to en_US_819. Now the
> database
> is using UTF-8 and a lot of information were introduced before the error
> were
> discovered.
>
> After export with dbexport I tried to look inside the unl files using
> ASCII in
> Notepad++ and I can see some correspondence like the following:
>
> ã = ã
> ç = ç
> ã = ã
> õ = õ
> á = á
> é = é
> Ã = í
> ó = ó
> ê = ê
> â = Â
>
> I'm thinking to replace it in the unl files and sql file of the export
> using
> Notepadd++ and after do a dbimport using UTF-8 coding.
>
> Do anyone have another better idea?
>
> Thanks in advance
>
> Best Regards
>
>
>
>
*******************************************************************************
> Forum Note: Use "Reply" to post a response in the discussion forum.
>
>
--001a113eb9b20db1d9052fce90c1