last week we open sourced:
anydoc: parse 14 document formats, ~5ms per page
pdf-inspector: parses + classifies pdfs, no waiting on OCR
14k stars each, built in rust. now powering
@firecrawl /parse, which adds OCR for the hard ones
what should we open source next?