Document extraction
Parsed PDFs, images, websites, emails and audio into more structured data.
Historical tool archive
PandaETL was an open-source, no-code tool for extracting structured data from documents. The hosted product is no longer active, and its GitHub repository was archived in February 2026.
No longer active

Archive
PandaETL was built to turn PDFs, images, emails, websites and audio into structured, exportable data. Users could define the fields they needed, inspect linked source evidence and work with the results in spreadsheet-style outputs.
The product is no longer active. Its official open-source repository was archived on February 16, 2026 and is now read-only, so this page is preserved as a historical reference rather than a current product recommendation.
Past capabilities
Parsed PDFs, images, websites, emails and audio into more structured data.
Provided a project interface for configuring extraction processes without writing the full pipeline.
The launched product emphasized traceable spreadsheet outputs connected to the original document context.
Availability
PandaETL no longer has current hosted-product pricing. Historical launch materials described a limited personal option and custom business or on-premise packages, but those offers should not be treated as available today.
Status checked . Review the historical source ↗
Current options
Business Operations
A current visual automation platform that can combine AI extraction with broader business workflows.
Explore Gumloop →Business Operations
A current workflow automation platform with self-hosting options and a large integration ecosystem.
Explore n8n →Students
A simpler current option for source-grounded exploration and synthesis of uploaded documents.
Explore Gemini Notebook (formerly NotebookLM) →Questions
No. The hosted product is no longer active, and the official GitHub repository was archived in February 2026.
The source code remains available in a read-only GitHub repository under the MIT license. Running it now means maintaining, securing and adapting the software yourself.
It extracted user-defined data from documents and other unstructured sources, then presented results in traceable, exportable tables.

Get access to all our AI courses, hundreds of real-world AI use cases, live expert-led workshops, an exclusive network of AI early adopters, and more.
Get unlimited access to all of our current & upcoming industry-specific AI courses for the duration of your subscription.
To keep up with the rapid pace of AI, our team publishes AI implementation guides daily. Our library contains 300+ practical use cases to automate real-world work.
Join weekly, live, interactive sessions with industry leaders who are at the forefront of AI for hands-on implementation guidance and exclusive insights.
Network with an exclusive community of AI-first professionals who are working smarter with AI. Learn how early adopters are using AI in their work and businesses.