Data Engineer, AI & Distributed Systems
Location: Remote
Compensation: Salary
Reviewed: Wed, Aug 05, 2026
This job expires in: 21 days
Job Summary
Building and operating data pipelines, the full-time Data Engineer, AI & Distributed Systems will develop and maintain systems that process massive volumes of unstructured data while collaborating with cross-functional teams in a fully remote environment.
Key responsibilities
- Build and maintain batch and streaming data pipelines, owning components from implementation to production monitoring
- Support AI systems by developing data pathways for NLP, LLM, and retrieval services
- Integrate with search and storage layers to enhance semantic search and real-time data retrieval
Required qualifications
- 3+ years of experience building and operating data pipelines in production
- Strong programming skills in a JVM language such as Scala, Java, or Kotlin
- Working proficiency in Python for data and scripting tasks
- Hands-on experience with distributed processing frameworks, particularly Apache Spark
- Practical experience with AWS and familiarity with Docker and Kubernetes
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 47,198 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee