
Sign up to save your podcasts
Or


Summary
When H.R. 1 put new pressure on states to cut SNAP payment error rates, Nava PBC had six weeks to help Pennsylvania stand up a way to catch bad document uploads before they reached a caseworker. That rapid-response build became Document AI, an open-source service that checks whether an uploaded document is legible and the right type, then extracts key fields for review.
Product manager Sophia Philip and senior software engineer Laurence Goolsby of Nava Labs join Ryan to talk about the project, which has since been adapted for a Maryland pilot that's set to scale statewide by December. They cover:
- Why the tool is deliberately kept out of eligibility decisions, with no policy coded into it
- How applicants can opt in, opt out, or proceed anyway when a document is flagged
- How it handles gig-work and handwritten invoices, plus Spanish-language documents
- Why a confidence score is not the same as accuracy
- How they're approaching AWS dependence and data security
- How other agencies can fork the repo and adapt it to their own needs
This episode is sponsored by Nava PBC.
Keywords
Document AI, intelligent document processing, Nava PBC, Nava Labs, open source, SNAP, H.R. 1, payment error rates, Maryland, Pennsylvania, benefits delivery, administrative burden, human-in-the-loop, consent, Amazon Bedrock, government technology, civic tech
Key Topics
- How AI-assisted prototyping ("demos over memos") is changing collaboration between product and engineering
- Document processing as a shared problem across benefits, licensing, permits, and more
- Building a six-week rapid response to H.R. 1 in Pennsylvania, then adapting it for Maryland
- Why open source: avoiding vendor lock-in while keeping PII inside each agency's environment
- The three users: applicants, caseworkers, and the people who administer the service
- Consent that is plain-language, just-in-time, and reversible
- Supporting nonstandard documents like handwritten invoices from gig and informal work
- Guarding against caseworker over-reliance: confidence scores, and no determinations made by the tool
- Under the hood: Amazon Bedrock and Bedrock Data Automation, plain-language document schemas, and Textract
- Meeting states on the infrastructure they already use, and the path to being cloud-agnostic
- Security by default: data that never leaves the agency, and extracted values that aren't returned unless requested
- How to contribute, deploy, or request demo access
Sound Bites
- "Confidence does not mean accuracy."
- "Fundamentals are the building blocks of fun."
- "Document AI isn't making any decisions about whether the document should be uploaded. That lives with the person who's uploading it."
- "It's not the person who uploads it's fault that we haven't defined it."
- "You are not in a vacuum."
Chapters
00:00 Intro
00:21 Meet Sophia Philip and Laurence Goolsby of Nava Labs
00:49 Their personal "why"
02:16 What's changed, and what hasn't, in product management and engineering
05:59 How AI prototyping changes collaboration across practices
08:33 Why document processing is a problem worth solving
12:38 Designing around burdens imposed by policy, and building reversible consent
15:15 What is Document AI?
16:26 Origin story: a six-week rapid response to H.R. 1 in Pennsylvania
20:10 Did it move payment error rates?
21:21 What "open source" means here, and why it matters
24:25 The three user personas
26:23 The applicant experience and opt-out fallback
28:41 The caseworker experience, gig-work invoices, and multilingual documents
33:26 Avoiding rubber-stamping: human-in-the-loop and no automated determinations
37:36 Shaping the product vision, from Postman demos to a management layer
41:32 Handling surprises: agile delivery and the Maryland pilot timeline
46:04 Under the hood: how Document AI uses AI
50:12 Why confidence isn't accuracy
51:13 Unsupported languages and undefined document types
54:07 AWS dependence and the path to cloud-agnostic
58:39 Security, privacy, and keeping data inside the agency
1:01:37 Governing Document AI after adoption
1:02:26 How to contribute
1:04:10 Final takeaways
Links
- Document AI open-source GitHub repository
- Nava demo day
Rebecca Woodbury and Luke Fretwell discuss their collaborative process in writing 'Proudly Serving,' a book aimed at making civic tech and government more accessible, responsive, and effective. They share insights on open source principles, the importance of plain language, and fostering a culture of transparency and public service.
The Book: https://proudlyservingbook.com
Key Topics
Chapters
AI tools have quietly moved out of isolated dev environments and into the middle of how real work gets done. That shift is genuinely exciting, and it brings a fresh set of risks worth sitting with. In this solo episode, Ryan works through what it takes to govern AI well, all of it anchored on one idea he keeps coming back to: a human has to stay accountable for the decisions that matter. He gets into why AI strains the governance habits IT already leans on, how to weigh centralized, decentralized, and hybrid approaches against your own risk tolerance, what ISO 42001 and the NIST AI RMF actually ask of you, and where the law is heading. He closes with a practical playbook for pulling shadow AI into the open while keeping the room for creativity that made folks reach for these tools in the first place.
In this episode
Key takeaways
Resources and shoutouts
We're joined by David D'Silva, Principal Product Manager at Nava, for a conversation about legacy lockpicks: the reusable tools, playbooks, and patterns that help government teams move out of legacy systems faster and with less risk. We talk about why trust is the real currency of modernization work, why so many efforts stall at the data layer, and how to choose between patterns like strangler fig and lift and shift. We also dig into where AI is genuinely useful in modernization today and where it's mostly hype.
- David’s LinkedIn: https://www.linkedin.com/in/daviddsilva/
- Prior episode on Strata: https://civictech.chat/episodes/nava-strata-open-source-government-technology
- Strata page: https://www.navapbc.com/strata
- Legacy modernization page: https://www.navapbc.com/legacy-modernization
- Our prior episode covering the strangler fig pattern: https://civictech.chat/episodes/the-specops-method
In this interview, Brian Chidester from Adobe discusses the importance of AI readiness in government, data governance, and the future of AI in public sector transformation. Learn how organizations can start small, build trust, and leverage AI to improve citizen services and drive mission outcomes.
Music Credit: Tumbleweeds by Monkey Warhol
Christian Martinez shares his journey from cognitive neuroscience to civic tech, highlighting the importance of open data, reproducible research, and building accessible tools like R packages. Discover how relevance drives rigor in education and data projects, and learn practical tips for getting started with open data and open source development.
Resources and Shoutouts
NYC Open Data Portal: https://opendata.cityofnewyork.us/
NYS Open Data Portal (MTA data also lives here): https://data.ny.gov/
NYC Open Data Student Gallery Book: https://martinezc1-nyc-open-data-student-gallery.share.connect.posit.cloud/
Music Credit: Tumbleweeds by Monkey Warhol
Ryan Koch discusses the importance of understanding supply chain attacks in cybersecurity, illustrating how they exploit trust in software components and build systems, and offering practical strategies to mitigate these risks.
Resources and Shoutouts:
- GRC Engineering Presentation by AJ Yawn: https://engage.isaca.org/events/event-description?CalendarEventKey=bd0c09c0-662f-4fe8-a0d2-019adaab6cc4&Home=%2fevents%2fcalendar
We're joined by Mark Headd author The SpecOps Method: A New Approach to Modernizing Legacy Technology Systems. We talk about good modernization practices like the strangler fig pattern as well as ways teams can use artificial intelligence to develop specifications that can be understood and validated across both technical and programmatic teams.
Music Credit: Tumbleweeds by Monkey Warhol
We’re joined by Ed Mullen, Technical Solutions Director and Lauren Ciferri Director of Product, over at Nava PBC. We’ll have a conversation about open source government technology, using Nava’s Strata project as a key example. The chat will dive into importance of open source in the space, maintaining projects like this one, and how one can get involved.
Resources and shoutouts:
Music Credit: Tumbleweeds by Monkey Warhol
We're joined by Mike Paciello (https://www.linkedin.com/in/mike-paciello-1231741/), Chief Accessibility Officer at AudioEye (https://audioeye.com) for a conversation about web accessibility standards. We'll talk about upcoming deadlines for local and state governments, foundational principles, and what goes into establishing and updating standards. Resources and shoutouts: - Civic Plus (www.civicplus.com) - US Access Board (https://www.access-board.gov/) - ADA Fact Sheet (https://www.ada.gov/resources/2024-03-08-web-rule/) ##### Music Credit: Tumbleweeds by Monkey Warhol
From the publisher's feed

17,639 Listeners

4,591 Listeners