officeParser

officeParser

harshankur

A robust, strictly-typed Node.js and Browser library for parsing office files into a rich Abstract Syntax Tree (AST) and generating high-fidelity output in multiple formats. Parses: docx · pptx · xlsx · odt · odp · ods · pdf · rtf · csv · md · html. Generates: Markdown · HTML · CSV · RTF · PDF · Plain Text · RAG Chunks

513 Stars
50 Forks
5 Watchers
Rich Text Format Language
mit License
100 SrcLog Score
Cost to Build
$3.60M
Market Value
$16.02M

Growth over time

13 data points  ·  2026-04-07 → 2026-08-03
Stars Forks Watchers
💬

How do you feel about this project?

Ask AI about officeParser

Question copied to clipboard

What is the harshankur/officeParser GitHub project? Description: "A robust, strictly-typed Node.js and Browser library for parsing office files into a rich Abstract Syntax Tree (AST) and generating high-fidelity output in multiple formats. Parses: docx · pptx · xlsx · odt · odp · ods · pdf · rtf · csv · md · html. Generates: Markdown · HTML · CSV · RTF · PDF · Plain Text · RAG Chunks". Written in Rich Text Format. Explain what it does, its main use cases, key features, and who would benefit from using it.

Question is copied to clipboard — paste it after the AI opens.

How to clone officeParser

Clone via HTTPS

git clone https://github.com/harshankur/officeParser.git

Clone via SSH

[email protected]:harshankur/officeParser.git

Download ZIP

Download master.zip

Found an issue?

Report bugs or request features on the officeParser issue tracker:

Open GitHub Issues