Specialist - Data Engineering

  • Pune, Maharashtra, India
  • Full-Time
  • On-Site

Job Description:

Location: Pune

Experience: 5–7 Years

About the Role

We are seeking a Data Engineer with strong expertise in PostgreSQL, SQL, data pipelines, and data modeling. The role focuses on building reliable data platforms and reusable data assets that support enterprise analytics, querying, APIs, feature engineering, and real-time and analytical consumption.

Job Description

The role involves designing, developing, and supporting data pipelines that securely acquire, aggregate, refine, and move data from source systems to target platforms. The candidate will develop production-grade SQL and optimize PostgreSQL performance through query tuning, indexing, partitioning, statistics management, and concurrency handling.

The position also requires designing logical and physical data models, establishing data contracts, managing schema evolution, and implementing resilient data processing with error handling, retries, idempotency, and reprocessing capabilities. The candidate will collaborate with data scientists, data wranglers, and technical development teams to deliver high-quality data solutions.

Key Responsibilities

  • Design, build, and support reliable data pipelines for enterprise data platforms.
  • Develop and maintain reusable data assets for analytics, querying, APIs, and feature engineering.
  • Write production-grade PostgreSQL SQL using CTEs, window functions, and set-based processing.
  • Optimize PostgreSQL performance through query tuning, execution plan analysis, indexing, partitioning, and statistics management.
  • Use EXPLAIN ANALYZE and BUFFERS to diagnose slow queries and identify I/O and execution bottlenecks.
  • Design effective indexes including B-tree, Hash, GiST, and GIN, aligned with filtering, joins, and sorting requirements.
  • Implement resilient pipelines with error handling, retry logic, idempotency, reprocessing, and backfill capabilities.
  • Design and maintain logical and physical data models using normalized and dimensional modeling approaches.
  • Define and maintain data contracts covering schemas, keys, constraints, naming conventions, SCD approaches, and business definitions.
  • Manage schema evolution with backward compatibility, deprecation planning, and impact analysis.
  • Implement data integration using ETL and other data tools across databases and platforms.
  • Establish data quality checks and monitoring for analytics data flows.

Required Skills

  • PostgreSQL
  • Advanced SQL
  • Query tuning and execution plan analysis
  • Indexing and partitioning
  • VACUUM/ANALYZE and table statistics
  • Locking and concurrency management
  • EXPLAIN ANALYZE / BUFFERS
  • Data pipeline development
  • Data modeling — logical and physical
  • 3NF and dimensional modeling
  • Data contracts and schema management
  • Schema evolution and backward compatibility
  • ETL and data integration
  • Data quality and monitoring
  • Version control and peer review
  • Error handling, retry, idempotency, and data reprocessing

Mandatory Skills

  • PostgreSQL

Candidate Profile

The ideal candidate will have 5–7 years of experience in Data Engineering, with strong hands-on expertise in PostgreSQL and production-grade SQL development. The candidate should have solid experience in data pipelines, performance optimization, data modeling, data integration, and data quality, along with the ability to collaborate effectively with data and technical teams.

Benefits Package

  • Competitive salary and benefits package
  • Opportunities for professional growth and development
  • A collaborative and inclusive work environment
  • The opportunity to work on exciting and challenging projects with leading clients