11 sep
|
Kreitech
|
Argentina
11 sep
Kreitech
Argentina
Join the Kreitech Team!
We're looking for a Senior Data Engineer who understands that clean, well-governed, AI-ready data is what separates a working agent from a broken one. You'll own the data layer that feeds our AI systems — designing pipelines, structuring unstructured data, integrating external sources, and making sure what flows into the model is trustworthy and traceable.
Responsibilities:
- Design and build production data pipelines for ingestion, transformation, and delivery across structured and unstructured sources (documents, APIs, logs, database exports, enterprise feeds)
- Integrate external systems — CRMs, ERPs, document stores, data platforms, third-party APIs — into unified analytical and AI-consumption layers
- Implement data governance across the pipeline: quality controls, lineage tracking, access scoping, and audit trails that hold up under client compliance requirements
- Prepare data specifically for AI/LLM consumption: chunking strategies, embedding-ready formats, retrieval-optimized schemas, and context scoping
- Own ETL/ELT processes across cloud platforms (AWS, GCP, Azure) using modern tools such as DBT, PySpark, and Data Factory
- Collaborate directly with AI engineers to define what data the agent needs and ensure it is available, scoped, and verifiable
- Document data flows and maintain engineering standards that can survive client handoff
Requirements:
- 5+ years in data engineering with end-to-end production pipeline experience
- Strong Python and SQL; hands-on experience with DBT,
PySpark, or equivalent transformation frameworks
- Multi-cloud experience across at least two of: AWS (S3, Redshift, Glue, Lambda), GCP (BigQuery, GCS), Azure (Data Factory, Synapse)
- Proven ability to ingest and structure unstructured data — documents, text, API responses, logs — into analyzable formats
- Experience integrating external systems via REST APIs and building reusable data connectors
- Solid understanding of data governance: quality, lineage, access control, and compliance-aware design
- Understands how data is consumed by ML/AI systems — not just how to move it
Desirable:
- Hands-on experience with AI/LLM pipelines: RAG architectures, text embeddings, vector databases (pgvector, Pinecone, Qdrant, ChromaDB)
- Background in fintech, healthcare, insurance, or other regulated industries (GDPR, SOC 2, compliance-aware pipeline design)
- Experience with MLOps tooling: MLflow, SageMaker, Azure ML, or model monitoring pipelines
- Familiarity with LangChain, LangGraph, or other agent frameworks from a data-layer perspective
Why join Kreitech?
- Remote work – enjoy the flexibility to work from anywhere
- Versátil schedule – prioritize work-life balance
- Continuous learning opportunities – stay ahead in your field
- Competitive compensation in USD
Apply now for the Senior Data Engineer position at Kreitech: https://airtable.com/appkPBsD5JS0Tp5wx/pagi7lO7hibX0U4bN?prefill_Position=Senior%20Data%20Engineer%20AI
📌 Data Engineer - AI speciality (Argentina)
🏢 Kreitech
📍 Argentina