Papero is a lightweight PDF parser that extracts text, tables, formulas, and layout information without machine learning, running on CPU in browsers, Python, or as an API. It converts PDFs to Markdown, JSON, Word, Excel and other formats while preserving document structure and block positions for use in RAG, search, and LLM applications.