Role Overview
You will join a global consultancy's Digital Services group as an AI Data Engineer Consultant, working with North American clients to build production-grade intelligent systems. Your day-to-day focus is turning raw, often messy environmental and operational data into structured, reliable inputs for AI-driven workflows, then scaling those workflows so they genuinely inform business decisions. This role sits at the intersection of data engineering, software development, and applied AI, so you will not just experiment with models — you will own the pipelines, APIs, and integrations that make them usable.
Key Responsibilities
- Design and implement AI-powered data pipelines that extract, clean, enrich, and transform both structured and unstructured data from multiple sources.
- Orchestrate large language models to automate data manipulation, validation, cross-referencing, and quality control tasks that would otherwise require manual effort.
- Generate polished, client-ready deliverables in Excel, PDF, and Word formats by programmatically assembling data and narrative content.
- Prototype emerging AI tools and techniques, then convert successful proofs-of-concept into scalable, maintainable production features.
- Build and integrate full-stack applications that connect AI-driven back-ends to modern web interfaces, ensuring smooth end-to-end user experiences.
- Design and optimize relational and vector databases to support high-volume AI workflows, including embedding storage and semantic search.
- Contribute to engineering best practices around code quality, documentation, testing, and system design so the team can move quickly without accruing technical debt.
Requirements & Qualifications
- A university degree in an environmental or technical discipline such as Environmental Sciences, Information Technology, Computer Science, Engineering, Management Information Systems, or a related business field.
- 4–6 years of relevant experience in AI data engineering, data-intensive software development, or an environment/health/safety (EHS) related domain.
- Hands-on experience with large language models such as OpenAI, Anthropic, Mistral, or open-source equivalents, including their APIs and limitations.
- A solid grasp of embeddings, vector search, semantic similarity, and retrieval-augmented generation (RAG) architectures.
- Strong Python skills, including async patterns, data libraries like pandas and polars, and AI SDKs.
- Practical experience implementing AI workflows using LangChain, LlamaIndex, or similar orchestration frameworks.
- Solid SQL and relational data modeling skills, with attention to query performance and data integrity.
- Ability to generate and manipulate Excel, PDF, and Word documents programmatically in a production context.
- Experience building APIs with FastAPI, .NET, or Node.js and integrating them into live systems.
- Preferred: front-end development experience with Vue 3 (Composition API) and TypeScript.
- Preferred: familiarity with vector databases and AI-oriented storage patterns.
- Preferred: exposure to containerization, cloud platforms, or serverless architectures.
- Preferred: background in NLP, document intelligence, or data enrichment pipelines, plus familiarity with monorepo tooling like Turborepo or Nx.
What We Offer / Why Join
You will work within a flexible, collaborative environment where your technical judgment directly shapes how clients operationalize AI responsibly. The organization emphasizes safety, well-being, and professional development, and offers a competitive salary. If you are motivated by sustainability and want to build solutions that have real impact on business and environmental outcomes, this team provides the platform to do so.
124 open positions on Semasocial right now
· 9080 open positions in Nairobi County, Kenya
· 9 posted in the last 7 days
Contact Information