Verified today
Data Engineer II
About the role
Fam (formerly FamPay) is India’s first payments app for users above 11, and its data team is building a high-performance, scalable Data Lakehouse to enable sub-minute data latency and unified batch/streaming compute. As a Data Engineer II (SDE-2) based in Bengaluru on-site, you will manage complex CDC flows, optimize distributed query engines, design domain models for OLAP, drive key tech initiatives and design reviews, and leverage AI tools to accelerate development while partnering with product and stakeholders on data and analytics.
What you’ll do
- Develop a high-performing and scalable Data Lakehouse toward sub-minute data latency and unified batch/streaming compute.
- Manage complex CDC flows and optimize distributed query engines.
- Prepare TRD and actively involve in design reviews to drive key tech initiatives.
- Design Domain models for OLAP including Fact, Dimension, Cumulative, types of SCDs, and OBT pattern tables.
- Lead the team technically and bring in new ideas to contribute to the growth of the charter.
- Interact with Product and key stakeholders to add value to business workflows with data and analytics.
- Use AI tools to write, test, and document code efficiently.
What you’ll bring
- 3–5 years in Data Engineering, specifically with distributed systems and cloud-native architectures.
- Expert-level Python/PySpark and SQL.
- Hands-on experience with AWS (S3, EKS, MSK) and Infrastructure-as-Code.
- Experience with Airflow or Temporal for complex workflow management.
- Proficiency in using AI tools (Claude, Codex, Copilot) to write, test, and document code efficiently.
- Ability to explain trade-offs between different storage formats and processing frameworks.
- Hands-on in designing Domain models for OLAP such as Fact, Dimension, Cumulative, types of SCDs, and OBT pattern tables.
- Lead the team technically, bring in new ideas, and interact with Product and key stakeholders to add value to business workflows with data and analytics.
Nice to have
- Familiarity with Go/Java/Scala.
- Ownership of high-throughput ingestion from RDBMS to Lakehouse using Debezium, PeerDB.
- Designing and optimizing table formats (Iceberg, Delta, Hudi) for performance and storage efficiency.
- Developing robust ETL/ELT frameworks in PySpark and Flink for batch and streaming workloads.
- Managing data workloads on AWS (EMR, EKS, MSK, S3) and automating via Gitlab/Github Actions.
- Tuning Trino or Clickhouse to power real-time dashboards in Metabase, Superset, and PowerBI.
- Exposure to Delta/Hudi and BigQuery.
Skills
Benefits
- Relocation assistance.
- Free office meals (lunch & dinner).
- Generous leave policy including birthday leave, period leave, paternity and maternity support.
- Salary advance and loan policies.
- Comprehensive health insurance for you and your family, plus mental health support.