Puts Claude in your terminal to understand your codebase, edit files and handle git workflows.
H2m Parser
gustavovalverde · App
H2m Parser turns raw HTML into Markdown that is ready to feed to a language model. It can extract the main article with Readability, resolve URLs, strip tracking parameters and convert tables, figures and footnotes, then add YAML front matter, a content hash and optional chunks. Per-tag translators, ignore lists and regex replacements let you adjust the output, and NDJSON transform helpers are included for pipelines.
Written by ooruby from the README of h2m-parser 1.0.0.
Not claimed by its maker. Is this yours? Prove it and answer the findings
Why it's useful
Web pages arrive as Markdown ready for a language model, with the main article isolated when you turn Readability on.
What you could do with it
- Pipe a saved web page through the command line with Readability enabled to get Markdown of its main article.
- Add YAML front matter and a content hash to each converted page so documents can be tracked in a pipeline.
- Split long converted pages into overlapping chunks by setting a target size and an overlap in tokens.
Written from its documentation: what it says it can do, not what our checks found.
Maintenance
Not measuredNo release date is on record for this package yet. The nightly publisher check records one as it reaches each npm package, about once every ten nights.
Release activity only: it says nothing about quality or safety, and a finished small package can be fine without releases. How it is measured
What this does not cover
- Everything. Treat it as you would software from anywhere else.
Receipt history: every signed receipt for every version of this listing, and what changed between them.
In its maker's words
LLM-ready HTML to Markdown pipeline with Readability, htmlparser2, and post-processing utilities.
Compare with
Compare all 4How far this has been checked
Static scanned
- What it establishes
- Its published package, or for a hosted server its source repository at a recorded commit, was read file by file, without running it, by every check in our current scan.
- What it does not
- How it behaves when it runs, or anything the scan does not read yet: several rubric checks, the repository's history, and compiled code in folders named dist or build. For a hosted server, that the endpoint runs the code that was read. The ooruby Index sets out exactly what was read.
- Exactly what was read
- Version 1.0.0, the file with digest sha512-6ODKW0rkefuAAXMyZIRtRuCRIWiDgao7Cd6ijm7vdKkiM1dSpg3D5yt6AIu6TO1XH1/4CVy8pdyTeK/rg+I09A==, as the registry publishes it. A published version cannot be replaced, so this names the same bytes for anyone who checks.
- Authentication
- Runs locally. No hosted endpoint is listed for this package. You run it on your own machine, so there is no endpoint of ours to authenticate against.
Adding it
Run it with:
npx -y h2m-parser@1.0.0Pinned to 1.0.0, which is the version the findings above were found in. Drop the version to take whatever is newest, and the report on this page stops describing what you installed.
Where this came from
- Registry
- npm
- Package
- h2m-parser
- Version read
- 1.0.0
- Resolved
- 2026-10-10
Built from source, attested
npm publishes a signed statement that this exact package was built from the repository below, at this commit. It is verifiable without taking our word for it.
- Repository
- github.com/gustavovalverde/h2m-parser
- Commit
- 1b6cf0188e42
- Built by
- github.com/actions/runner/github-hosted
- Registry publish attestation
- signed by npm · 2026-10-10
- Checked
- 2026-10-10
The badge, if this is your listing
It renders the current rung (static scanned) and links back here, where what that does and does not establish is one click away. It updates itself as the evidence deepens.
[](https://ooruby.com/market/h2m-parser)About this listing
- Kind
- App
- Category
- Developer tools
- Pricing
- Free
- Sandbox
- No
- Hosted
- No
- Updated
- 2026-10-06
Similar apps
All appsGives your app one API for streaming from many AI model providers, with typed and validated tools.
Estimates how many tokens a text will use without loading a full tokeniser, from code or the shell.
Verification records what our published tests found on a specific version at a specific date. It is not a warranty, and it does not certify that software is free of defects.