Beefy Boxes and Bandwidth Generously Provided by pair Networks
Syntactic Confectionery Delight
 
PerlMonks  

Re^3: Finding dates from web pages

by marto (Archbishop)
on Feb 25, 2020 at 20:53 UTC ( #11113416=note: print w/replies, xml ) Need Help??


in reply to Re^2: Finding dates from web pages
in thread Finding dates from web pages

"I think the Python library applies some heuristics in such cases"

You could look at the Python code and implement the same thing in perl. Let me know if you get stuck.

Replies are listed 'Best First'.
Re^4: Finding dates from web pages
by cormanaz (Chaplain) on Feb 26, 2020 at 15:22 UTC
    Looking at the code, that would appear to be quite an undertaking. It turns out it can be installed and run from the command line, so I will probably just go that route, though it will slow things down to load it every time it needs to be run.

      Well, it looks fairly well documented, should you encounter performance problems using the Python based command like utility.

Log In?
Username:
Password:

What's my password?
Create A New User
Node Status?
node history
Node Type: note [id://11113416]
help
Chatterbox?
and the web crawler heard nothing...

How do I use this? | Other CB clients
Other Users?
Others examining the Monastery: (6)
As of 2020-04-07 15:58 GMT
Sections?
Information?
Find Nodes?
Leftovers?
    Voting Booth?
    The most amusing oxymoron is:
















    Results (43 votes). Check out past polls.

    Notices?