Senior Data Ingestion Engineer (AWS / Document Extraction / AI-RAG Prep)
- Рівень:
- senior
- Джерело:
- djinni.co
Що робити
- Pipeline Design & Engineering: Design and deploy scalable data ingestion pipelines for high-volume structured/unstructured documents (PDFs, scans, emails, Word, Excel, PowerPoint).
- Document Extraction & OCR: Implement OCR and document processing workflows using AWS Textract (or equivalent) for text extraction, cleaning, normalization, and metadata tagging.
- RAG & Vector Storage Prep: Execute semantic chunking, metadata extraction, vector storage schema design, and retrieval mechanism preparation for downstream AI models.
- Integrations & Connectors: Build robust connectors with enterprise sources, including SharePoint, email servers, and public cloud repositories.
- Validation & Observability: Implement automated error monitoring, OCR extraction validation, and CI/CD automated testing using AWS Step Functions and CloudWatch.
Схожі вакансії
- Senior Data Engineer (Multimodal Data and Entity Modelling)Osavulsenior
Data Entry Specialist (Medical Records)Pharmbills, до $1200remote
Data Processing CoordinatorPharmbills, до $1200remote- 3D Full-Stack Geometry Engineer (C++ / Three.js / Angular / Node.js / AWS)orthoeye.digital
Epic Data & Analytics Consultant (Presales / Freelance)Yalantisremote
З блогу Trackr
Усі статті →Знайдено через trackr.help/jobs · Канал: @trackrhelp · Бот для персональних сповіщень: @trackrhelpBot


