Initiative #72
Extracting impacts data from regulatory documents
Listed in the registry Ideation Original language : English
About the initiative
Description of the initiative
Use OCR pdf readers, topic detection and language models to detect specific data elements inside written text documents, for extraction and analysis in a cumulative impacts context. Large lanugage models and topic clustering tools
AI family
- AI capability
- Document and Text Data Digitization and Interpretation
Bucket rationale
Perceives and classifies unstructured text, extracting key elements—perceptual natural language processing.
Lifecycle
-
Ideation
Stage status : In progress
May 29, 2025
-
Assessment
Stage status : Upcoming
-
Approved
Stage status : Upcoming
-
Development
Stage status : Upcoming
-
Pilot
Stage status : Upcoming
-
Production
Stage status : Upcoming
-
Archived
Stage status : Upcoming
Schedule for this initiative
- Completed
- In progress
- Planned
- Not applicable
- Today
A similar need?
Nobody has come forward yet. If your team faces the same problem, say so: it helps bring initiatives together and share a solution.
Contacts
- Requester
- Quan, Eric
- Sector contact
- Tom Bird/ Neil Fisher
- Subject matter expert
- Bird, Tom
Data and tools
- Sector
- Ecosystems and Oceans Science
- Region
- Pacific
- Submitted on
- May 29, 2025