Full-time · Sydney, Melbourne & India
Data Engineer
Design, Build, and optimise the pipelines and infrastructure powering the Zetaris Platform
About us
Zetaris is the world's leading data harness for AI - the open, vendor-agnostic layer that harnesses an organization's data engines and AI models so every agent, analyst, and application can build on enterprise data instantly. Zetaris lets teams query, govern, and activate their data in place, across every source and system they own, without data movement, ETL pipelines, or proprietary lock-in. By harnessing underlying engines such as Spark, Presto, and Trino, Zetaris delivers governed, real-time access to data wherever it lives - with full sovereignty over security, cost, and compliance. As enterprises face mounting pressure to deliver AI ROI on fragmented, distributed data estates, Zetaris provides the open data infrastructure that turns that ambition into measurable outcomes.
Role Overview
A Data Engineer who is passionate about designing, building, and optimising the data pipelines and infrastructure that power the Zetaris Platform.
This role is responsible for building scalable, reliable data pipelines, integrating diverse structured and unstructured data sources, and ensuring data is accurate, well-governed, and readily accessible for analytics and AI use cases. You will work closely with Software Engineering, Product, and Customer-facing teams to deliver high-quality data solutions across cloud, on-premises, and hybrid environments.
Basic Zetaris platform knowledge: Basic knowledge of the Zetaris platform is important for this role. As part of your application, please register and navigate the platform at https://www.zetaris.com/cloud. No deep product knowledge is required - we're interested in seeing how you navigate and apply the platform based on your understanding.
Responsibilities
Design, build, and maintain scalable data pipelines and integrations supporting the Zetaris Networked Data Platform, including:
-
Design, develop, and maintain ETL/ELT pipelines to ingest data from structured, semi-structured, and unstructured sources.
-
Build and optimise data models and schemas to support analytics, reporting, and AI/ML use cases.
- Develop and manage data integration workflows across cloud, on-premises, and hybrid environments.
- Implement data quality checks, validation rules, and monitoring to ensure data accuracy and reliability
- Optimise query performance and data storage across distributed and multi-source environments.
- Collaborate with Software Engineering teams to support the Zetaris query engine and data virtualisation capabilities.
- Support data governance practices, including data cataloguing, lineage, and access control.
- Automate data workflows using orchestration tools and CI/CD pipelines.
- Troubleshoot data pipeline issues and perform root cause analysis to improve reliability.
- Work effectively within Agile cross-functional teams and contribute to continuous improvement of engineering practices.
- Document data pipelines, schemas, and integration processes to support knowledge-sharing and onboarding.
- Support customer and partner data integration requirements as needed.
Attributes
- Strong understanding of data engineering principles, distributed systems, and cloud data architecture.
- Proven experience building and maintaining ETL/ELT pipelines at scale.
- Strong experience with SQL and query optimisation across large, multi-source datasets.
- Experience with data pipeline orchestration tools such as Apache Airflow, or similar.
- Experience with one or more major cloud platforms including Azure, AWS, or GCP.
- Proficiency in a programming language commonly used for data engineering, such as Python, Scala, or Java.
- Experience working with both structured and unstructured data sources, including APIs, databases, and file-based systems.
- Understanding of data modelling techniques, including dimensional modelling and data warehousing concepts.
- Experience with big data processing frameworks such as Spark or similar distributed processing tools.
- Knowledge of data governance, data quality, and data security best practices.
- Experience with version control and CI/CD tools to support automated data pipeline deployment.
- Experience working in fast-paced Agile environments.
- Strong troubleshooting and analytical problem-solving skills.
- Strong communication and collaboration skills.
- Up to date with modern data engineering tools and best practices.
Behavioural & Soft Skills
- Collaborates effectively within Agile, cross-functional teams to achieve shared goals.
- Demonstrates strong analytical and problem-solving abilities with a practical approach to challenges.
- Adapts quickly to shifting priorities and thrives in a fast-paced, evolving environment.
- Takes ownership of tasks, works independently, and consistently delivers high-quality outcomes.
- Communicates complex ideas clearly and confidently with both technical and non-technical audiences.
- Shows curiosity and a commitment to continuous learning, exploring new technologies and best practices.
- Balances innovation with pragmatic execution to deliver reliable, scalable solutions.
- Acts with integrity, professionalism, and accountability in all interactions and decisions.
- Contributes to a supportive, knowledge-sharing, and growth-oriented team culture.
- Maintains a proactive, positive, and results-focused mindset, even under pressure.
Why you will love working with Zetaris
- Be at the forefront of an exciting international expansion backed by a global co-investment partnership.
- Shape the architecture and commercial direction of a growing software company.
- Work in a collaborative, agile environment where your ideas directly influence success.
- Competitive remuneration with performance incentives and potential for equity participation.
Apply for this role
Tell us about yourself and attach your CV (PDF or Word). We read every application.
Not quite the right role?
See all open positions at Zetaris.