The most practical innovation in spreadsheet work right now might be one that removes the most tedious step: copying data out of PDFs by hand. A developer working with aviation leasing clients has built an add-in that does exactly that, connect a folder of PDFs, ask it to extract the numbers, and watch the cells fill. Each value comes with a comment linking back to the source page. That is not a minor convenience. For anyone who has spent hours cross-referencing documents and double-checking which row came from which page, this is the kind of fix that changes how you approach an entire engagement.
The design is straightforward precisely because it respects the user's existing workflow. The add-in lives in the taskpane, not in some separate interface you have to learn. You link a folder, you ask it to extract, and the results land in your spreadsheet with provenance attached. The developer started from an empty sheet and ended with a populated grid where every cell points back to its source. That traceability matters. In aviation leasing, where a single number can determine a payment schedule or a maintenance obligation, knowing exactly which page of which PDF produced that figure is not a nice-to-have, it is the difference between a defensible audit trail and a guess.
What makes this worth watching is how it reframes the relationship between spreadsheets and documents. For years, the standard approach has been to extract data into a spreadsheet and then lose the connection to the source. You paste the number, you note the page in a separate column, and you hope you never have to re-verify. This add-in inverts that pattern. The spreadsheet becomes a live index into the PDFs, not a graveyard of copied figures. The comment on each cell is not metadata; it is the thread back to the original context. That is a genuinely better way to work.
The developer acknowledges it is early, and the question of what other workflows could benefit is the right one to ask. But the pattern is clear: any industry that lives on PDFs, legal, insurance, real estate, compliance, has the same pain point. The solution is not a better PDF reader or a smarter OCR tool. It is a spreadsheet that knows where its data came from and never lets you forget. That is the direction we should be moving.