npm package discovery and stats viewer.

Discover Tips

  • General search

    [free text search, go nuts!]

  • Package details

    pkg:[package-name]

  • User packages

    @[username]

Sponsor

Optimize Toolset

I’ve always been into building performant and accessible sites, but lately I’ve been taking it extremely seriously. So much so that I’ve been building a tool to help me optimize and monitor the sites that I build to make sure that I’m making an attempt to offer the best experience to those who visit them. If you’re into performant, accessible and SEO friendly sites, you might like it too! You can check it out at Optimize Toolset.

About

Hi, 👋, I’m Ryan Hefner  and I built this site for me, and you! The goal of this site was to provide an easy way for me to check the stats on my npm packages, both for prioritizing issues and updates, and to give me a little kick in the pants to keep up on stuff.

As I was building it, I realized that I was actually using the tool to build the tool, and figured I might as well put this out there and hopefully others will find it to be a fast and useful way to search and browse npm packages as I have.

If you’re interested in other things I’m working on, follow me on Twitter or check out the open source projects I’ve been publishing on GitHub.

I am also working on a Twitter bot for this site to tweet the most popular, newest, random packages from npm. Please follow that account now and it will start sending out packages soon–ish.

Open Software & Tools

This site wouldn’t be possible without the immense generosity and tireless efforts from the people who make contributions to the world and share their work via open source initiatives. Thank you 🙏

© 2026 – Pkg Stats / Ryan Hefner

@dudko.dev/pdf-to-md-cli

v0.3.5

Published

Command-line PDF → Markdown converter

Readme

@dudko.dev/pdf-to-md-cli

pdf2md — PDF to Markdown on the command line. The parsing is a WebAssembly module running locally; nothing is uploaded anywhere.

npx @dudko.dev/pdf-to-md-cli report.pdf
npm install -g @dudko.dev/pdf-to-md-cli && pdf2md report.pdf

The hosted service

pdf2md.dudko.dev does the same in a browser, with the file never leaving the tab — worth a look before installing anything. For a pipeline that is not this command, the HTTP API is at pdf2md.dudko.dev/api and an MCP endpoint for assistants at https://pdf2md.dudko.dev/mcp; both need a free account at auth.dudko.dev.

A paid account carries commercial permission for that service, not for this package: running the software yourself needs the written licence below.

Pictures rather than documents: vectorize.dudko.dev, with @dudko.dev/vectorize-cli on the command line.

Use

pdf2md report.pdf                     # Markdown to stdout
pdf2md report.pdf -o report.md
pdf2md *.pdf -o out/                  # each input named after itself
cat report.pdf | pdf2md -             # from stdin

pdf2md report.pdf -o out/ --images    # plus out/report.assets/img-1.png…, linked from the Markdown
pdf2md scan.pdf -o out/ --images --jpx keep   # JPEG 2000 kept as .jp2 rather than decoded to PNG
pdf2md report.pdf -o out/ --images --no-vectors   # pictures only; no SVG of the charts and diagrams
pdf2md report.pdf --json              # the whole result, not just the Markdown
pdf2md report.pdf --detect            # type, pages, which pages need OCR (JSON; add --json to pretty-print)
pdf2md report.pdf --text              # plain text, no Markdown structure

pdf2md report.pdf --pages 1,3,5-7 --profile compact
pdf2md secret.pdf --password hunter2
pdf2md paper.pdf --page-markers --strip-headers-footers

--help lists everything, including the --no-* flags that turn off individual detectors (headings, lists, code, bold, italic, URL rewriting, hyphenation repair) and --underline, which turns on the one that is off by default: <u> around text with a line under it.

Exit codes

| Code | Meaning | | --- | --- | | 0 | everything converted | | 1 | bad usage, or nothing could be read | | 2 | nothing could be parsed | | 3 | some files converted, some did not |

Made for pipelines: 3 is the one that means "look at stderr".

Not OCR

A scanned PDF is reported as Scanned with no Markdown, and the command exits 2. Run it through an OCR tool first, then convert the result.

Licence

Free for noncommercial use under PolyForm Noncommercial 1.0.0, with a 32-day trial for evaluation at work. Commercial use needs a licence: [email protected].

Your rights come from LICENSE and, for commercial use, from a written agreement. No web page grants them; a paid account at pdf2md.dudko.dev is permission to use that service.

The notices for the third-party code compiled into the module are in THIRD-PARTY-NOTICES.md in @dudko.dev/pdf-to-md-core, which this package installs.