Get all href urls from an HTML string
Install get-hrefs
using npm:
npm install --save get-hrefs
var getHrefs = require('get-hrefs');
getHrefs(`
<body>
<a href="http://example.com">Example</a>
</body>
`);
// ["http://example.com"]
getHrefs(`
<head>
<base href="http://example.com/path1/">
</head>
<body>
<a href="path2/index.html">Example</a>
</body>
`);
// ["http://example.com/path1/path2/index.html"]
$> get-hrefs --help
Get all href urls from an HTML string
Usage:
get-hrefs <html file>
cat <html file> | get-hrefs
Options:
-b, --base-url Set baseUrl
Examples:
curl -s example.com | get-hrefs
Name | Type | Description |
---|---|---|
html | String |
The HTML string to extract hrefs from |
options | Object |
Optional options |
Returns: Array<String>
, all unique and normalized hrefs resolved from any provided baseUrl
and <base href="...">
in the HTML document.
Type: String
Default: ""
The baseUrl to use for relative hrefs. The module also takes <base ...>
tags into account.
- get-urls - Get all urls in a string
MIT © Joakim Carlstein