job description
Join Oppstar Technology as a Senior Data Engineer and play a pivotal role in designing, building, and maintaining cutting-edge data infrastructure that powers next-generation AI agents. In this high-impact position, you will focus on delivering high-quality, scalable data solutions tailored for semiconductor manufacturing and internet-based applications, ensuring seamless data flow, integrity, and performance.
Based in the vibrant tech hub of Canggu, Bali, youâll collaborate with cross-functional teams to architect robust ETL pipelines, optimize data warehouses, and implement advanced data governance frameworks. Your work will directly enable AI-driven decision-making, predictive analytics, and real-time processing for industry-leading clients in semiconductor and digital ecosystems.
If youâre passionate about big data technologies, cloud platforms, and AI/ML integration, this is your chance to shape the future of data engineering in a dynamic, innovative environment.
Responsibility
- Design, develop, and maintain scalable data pipelines (batch/real-time) to support AI agents and semiconductor data processing.
- Architect and optimize data warehouses (e.g., Snowflake, BigQuery, Redshift) for high-performance analytics.
- Implement data quality frameworks to ensure accuracy, consistency, and reliability of datasets.
- Collaborate with AI/ML teams to integrate data infrastructure with machine learning models and predictive algorithms.
- Automate data workflows using Airflow, Kafka, or Spark to streamline ETL processes.
- Monitor and troubleshoot data systems, ensuring uptime, latency, and cost-efficiency.
- Develop APIs and data services to enable seamless access to structured/unstructured data.
- Stay ahead of industry trends in data engineering, semiconductor analytics, and AI-driven data solutions.
Qualifications
- Bachelorâs or Masterâs degree in Computer Science, Data Engineering, or a related field.
- 5+ years of experience in data engineering, with a focus on large-scale data systems.
- Proficiency in Python, SQL, or Scala and experience with big data tools (Hadoop, Spark, Databricks).
- Hands-on experience with cloud platforms (AWS, GCP, or Azure) and data storage solutions (e.g., S3, Delta Lake).
- Strong understanding of data modeling, ETL/ELT processes, and data governance.
- Experience with semiconductor or manufacturing data is a plus.
- Familiarity with AI/ML pipelines and tools like TensorFlow, PyTorch, or MLflow.
- Excellent problem-solving skills and ability to work in a fast-paced, collaborative environment.