Beefy Boxes and Bandwidth Generously Provided by pair Networks
Perl Monk, Perl Meditation
 
PerlMonks  

Re: Re: Invalid UTF8 data: namespace suggestion needed or wheel reinvented?

by liz (Monsignor)
on Mar 02, 2004 at 11:23 UTC ( #333220=note: print w/ replies, xml ) Need Help??


in reply to Re: Invalid UTF8 data: namespace suggestion needed or wheel reinvented?
in thread Invalid UTF8 data: namespace suggestion needed or wheel reinvented?

How is the utf-8 usually being corrupted?

A good question. It's a system that has it roots in Perl 4 and which consists of hundreds of pages (CGI scripts) that are being used by many different people 24x7. And which is currently being migrated to a more manageable situation, so there is little point in trying to find out which users are doing what on which page causing the problem.

What exactly are you doing (are you just removing illegal characters? what else)?

Indeed. They're mostly free text entries of different sources in various languages.

Rules (is it configurable, or will it be, how)?

The only configuration that I see, would be the ability to list the tables to correct (defaulting to all). And how fields fixed are reported (either STDERR or calling a subroutine).

...it sounds like a script to me

Well, it is is a script now. But I dislike scripts as such, because you are never flexible enough. I'd rather have a simple script that calls a module, than that people start copying a script and start adapting it to their needs.

Liz


Comment on Re: Re: Invalid UTF8 data: namespace suggestion needed or wheel reinvented?

Log In?
Username:
Password:

What's my password?
Create A New User
Node Status?
node history
Node Type: note [id://333220]
help
Chatterbox?
and the web crawler heard nothing...

How do I use this? | Other CB clients
Other Users?
Others taking refuge in the Monastery: (10)
As of 2014-07-23 05:02 GMT
Sections?
Information?
Find Nodes?
Leftovers?
    Voting Booth?

    My favorite superfluous repetitious redundant duplicative phrase is:









    Results (133 votes), past polls