Morphik
An open-source AI-native knowledge base and research agent accelerating complex multi-modal document synthesis for enterprise teams.
Published entries across all sections carrying the “Extraction” tag, newest first by publication date on this site.
4 entries
An open-source AI-native knowledge base and research agent accelerating complex multi-modal document synthesis for enterprise teams.
An ontology-first Graph RAG platform that structures disparate documentation into verifiable, reusable knowledge packs for autonomous agents.
A self-hostable, TypeScript web scraping toolkit that covers everything from static pages to JS-rendered pages across three engines, with bulk search-result collection, full-site crawling, and LLM structured extraction, all outputting LLM-friendly data.
MonkeyOCR is an open-source document-parsing tool built on a lightweight multimodal LLM that parses PDFs in three stages—layout, recognition, relation—to reconstruct formulas, tables, and reading order into Markdown, with local GPU inference.