Component

parse-web

Extract source-attributed main content from a fetched web page.

Inputhtml · sourceUrlOutputdocument

Inputs

FieldTypeContract
html *stringFetched HTML bytes.
sourceUrl *stringOriginal source URL.

Outputs

FieldTypeContract
documentobjectTitle, main text, canonical URL and extraction diagnostics.

Behavior and limits

Extract source-attributed main content from a fetched web page.

Try it in context

Use the linked loops to inspect actual configuration and how output is passed to the next step. Keep credentials in your local environment.

Validation and run evidence

Evidence not supplied. Release acceptance: pending.

No new validation or execution is performed by viewing this page.

No output artifact is available for this material.

Make it your own

One prompt to your result

Your agent handles setup, configuration and checking the result in your project.