Beefy Boxes and Bandwidth Generously Provided by pair Networks Joe
"be consistent"
 
PerlMonks  

4Re: How do I extract text from an HTML page?

by jeffa (Chancellor)
on Aug 03, 2003 at 19:02 UTC ( #280493=note: print w/ replies, xml ) Need Help??


in reply to Re: Re: Re: How do I extract text from an HTML page?
in thread How do I extract text from an HTML page?

Two suggestions: test before you post (and then test some more) and don't use regexes to parse HTML. Granted, this is trivial HTML to parse, but the more you use parsers, the better you get at it. Also, the more you use templating modules, the better you get at them. I hate to just outright solve the problem, but i did. Here is link that you will have to click to see my solution - so any readers have been warned.

<blink>
WARNING *SPOILERS* CLICK AT OWN RISK!
</blink>

It's a lot more code than you posted, but it does a lot more as well. ;)

jeffa

L-LL-L--L-LL-L--L-LL-L--
-R--R-RR-R--R-RR-R--R-RR
B--B--B--B--B--B--B--B--
H---H---H---H---H---H---
(the triplet paradiddle with high-hat)


Comment on 4Re: How do I extract text from an HTML page?
Download Code

Log In?
Username:
Password:

What's my password?
Create A New User
Node Status?
node history
Node Type: note [id://280493]
help
Chatterbox?
and the web crawler heard nothing...

How do I use this? | Other CB clients
Other Users?
Others perusing the Monastery: (10)
As of 2014-04-19 06:27 GMT
Sections?
Information?
Find Nodes?
Leftovers?
    Voting Booth?

    April first is:







    Results (478 votes), past polls