Structured Web ExtractionService
structured-web-extraction
Fetches a web page and returns structured data from it. POST JSON {"url": "https://..."} to /v1/url. Without a schema, fields holds the page standard fields (title, description, canonical, language, h1, OpenGraph og_title, og_description, og_image, og_type, site_name, author, published) and page holds headings (h1-h3), JSON-LD blocks, links_count and a text excerpt. With "schema": ["field", ...] or a JSON Schema, each field is taken from the standard fields or from a "Label: value" line in the page text; with {"fields": {"name": "<regex>"}} each regex is applied to the page HTML. JavaScript is not executed. Pages are cached 10 minutes, so results can change over time.
Active Suppliers
1
of 1 ever configured
Service Info
Service ID
structured-web-extraction
Compute Units / Relay
50,000
Owner
Active Applications
1
Relay Mining EMA
1 relays
| Supplier | Stake | Endpoint |
|---|---|---|
| pokt1v…7aay | 60,000.00 POKT | https://services.supnodes.com |
| Application | Stake |
|---|---|
| pokt1c…k55p | 2,449.92 POKT |
Raw Servicesource: GraphQL indexer
{6 items
"id":"structured-web-extraction"
"name":"Structured Web Extraction"
"computeUnitsPerRelay":"50000"
"ownerId":"pokt10a8l4nc03uncmecjy0rnv27sfl266nrjmq87pn"
"owner":{1 item
"id":"pokt10a8l4nc03uncmecjy0rnv27sfl266nrjmq87pn"
}
"latestDiff":{1 item
"nodes":[1 item
0:{3 items
"newNumRelaysEma":"1"
"newTargetHashHexEncoded":"ffffffffffffffffffffffffffffffffffffffffffffffffffffffffffffffff"
"blockId":"937493"
}
]
}
}