Job Description
Apple is looking for an experienced Data Integration Engineer who will create and operate big data processing systems which handle both structured and unstructured information. The position requires working with AIML datasets together with cloud databases and ETL processes while supporting analytics and data science teams who need to transform data for Apple Sales projects.
Qualification: BS/MS in Computer Science or equivalent industry experience
Experience: 5+ Years
Job ID: 200647260-1052
Role: Data Integration Engineer, Cloud & AIML Data Operations
Apply: Click Here
Primary Responsibilities
- Develop and design data integration and ingestion processes for internal and external Apple data.
- Build and maintain pipelines for structured and unstructured data to support AIML models.
- Implement data models, mapping rules, and a semantic layer to transform raw data into actionable insights.
- Collaborate with analytics, data science teams, business partners, and vendors to meet data requirements.
- Ensure data quality, validation, governance, and ethical handling for AIML datasets.
Technical & Support Responsibilities
- Expertise in SQL, Python, Java, R, Hadoop, Spark, Snowflake, Redshift and Databricks cloud database platforms.
- Continuous integration and continuous delivery pipelines through automated processes which they will manage with Apache Airflow and Git and current cloud technologies.
- Use ETL tools and API integrations and debt to transform data which includes using macros and Jinja templating.
- Monitor and troubleshoot and optimize pipelines to achieve peak performance and reliable operation and ability to handle increased workloads.
- The AI/ML model development lifecycle requires high-quality clean datasets which need proper annotation support.
Qualifications Required
- Five years of experience in creating and developing scalable data systems which I later maintained for operational use.
- Expertise in SQL and cloud databases together with programming languages and ETL systems and data pipeline technologies.
- Experience working with unstructured data formats which include JSON, Parquet, text, images, audio, and video.
- Comprehensive understanding of data modeling together with warehousing, governance systems and AIML data requirements.
- Demonstrates capacity to handle multiple tasks while working in a fast-paced environment which requires quick adaptation to changing situations.