in reply to The unicode / utf8 struggle, part 2: regexes
HTML::Encoding helps to determine the encoding of HTML and XML/XHTML documents...
use HTML::Encoding 'encoding_from_http_message'; use LWP::UserAgent; use Encode; my $resp = LWP::UserAgent->new->get('http://www.example.org'); my $enco = encoding_from_http_message($resp); my $utf8 = decode($enco => $resp->content);
|
|---|