This software is designed to extract formatted content from HTML pages. Many webmasters have pages that need updating of repetitive code, but don\'t want to manually change each page. Thus, a content extractor such as this is born. Outputs in XML.

Project Activity

See All Activity >

Categories

Site Management

License

GNU General Public License version 2.0 (GPLv2)

Follow The HTML Content Extractor

The HTML Content Extractor Web Site

Other Useful Business Software
Stop Cyber Threats with VM-Series Next-Gen Firewall on Azure Icon
Stop Cyber Threats with VM-Series Next-Gen Firewall on Azure

Native application identity and user-based security for your Azure cloud

Gain integrated visibility across all traffic in a single pass. Deploy Palo Alto Networks VM-Series to determine application identity and content while automating security policy updates via rich APIs.
Get a free trial
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of The HTML Content Extractor!

Additional Project Details

Languages

English

Intended Audience

Developers

Programming Language

Perl

Related Categories

Perl Site Management Software

Registered

2003-10-08