OpenBSD Handbook

OpenBSD 7.9 release · amd64 · Generated 2026-09-09

p5-HTML-Parser-3.83

modules to parse and extract information from HTML

Back to search · Project homepage

Description

This is a collection of modules that parse and extract information from HTML documents. Bug reports and discussions about these modules can be sent to the <libwww@perl.org> mailing list. Remember to also look at the HTML-Tree package that creates and extracts information from HTML syntax trees. The modules present in this collection are: HTML::Parser - The parser base class. It receives arbitrary sized chunks of the HTML text, recognizes markup elements, and separates them from the plain text. As different kinds of markup and text are recognized, the corresponding event handlers are invoked. HTML::Entities - Provides functions to encode and decode text with embedded HTML &gt;entities&gt;. HTML::HeadParser - A lightweight HTML::Parser subclass that extracts information from the <HEAD> section of an HTML document. HTML::LinkExtor - An HTML::Parser subclass that extracts links from an HTML document. HTML::TokeParser - An alternative interface to the basic parser that does not require event driven programming. Most simple parsing needs are probably best attacked with this module.

Package information

Ports path
www/p5-HTML-Parser
Package architecture
amd64
Maintainer
The OpenBSD ports mailing-list <ports@openbsd.org>
Categories
www, perl5
Available flavors
None listed

These are ports metadata. Binary availability depends on the release, architecture and mirror. Build and test dependencies are not an installation checklist.

Direct dependencies

Runtime

Test

Used by (139 dependency relationships)

Includes library, runtime, build and test relationships. Results load 100 at a time.

Installing and updating packages · Package details as JSON