⚡ New
Data Engineer - Contract (Mail at nirmal.francis@
Inspire
ChennaiFull-timeMid LevelOn-site
Job Description
Job Description
Role Description:
\n- \n
- Cloud Infrastructure Architecture (AWS) \n
- Architect, deploy, and manage scalable AWS data infrastructure including Redshift, S3, Lambda, IAM, VPC, and related services \n
- Define and enforce infrastructure-as-code standards and CI/CD practices for all data platform components \n
- Monitor infrastructure health, optimize performance, and drive cost efficiency across cloud resources \n
- Ensure high availability and disaster recovery posture for all critical data systems \n
- Manage multi-account AWS environments with cross-account IAM roles, VPC peering, and environment-specific security boundaries (dev, test, production) \n
- Data Warehouse Design & Development \n
- Architect and evolve the organization's Redshift data warehouse using dimensional modeling (Kimball methodology) and star schema best practices \n
- Architect and engineer the dbt project end-to-end: data modeling, testing, documentation, and deployment via Jenkins across development, testing, and production environments \n
- Write complex, optimized SQL and Python to build and maintain production data models, snapshots, and incremental loads \n
- Establish and enforce data modeling standards, naming conventions, and development workflows \n
- Data Pipelines & Integrations \n
- Design and build robust ELT/ETL pipelines using custom Python and managed connectors via Fivetran \n
- Manage third-party data integrations including SaaS platforms, APIs, and partner data feeds \n
- Implement pipeline observability, alerting, and data quality frameworks to ensure reliable data delivery \n
- Evaluate and adopt new ingestion tools and patterns as the data ecosystem evolves \n
- Data Security & Compliance \n
- Own data platform security including IAM roles and policies, Redshift access controls, encryption at rest and in transit, and network security (VPC, security groups) in partnership with Infrastructure \n
- Ensure compliance with HIPAA, SOC 2, and applicable healthcare data privacy regulations \n
- Conduct regular access reviews, security audits, and vulnerability assessments of the data ecosystem \n
- Partner cross-functionally to implement and maintain data governance standards \n
- Design, implement and maintain PII anonymization pipelines to provision safe, production-representative data to development and test environments \n
- Identify and automate recurring compliance workflows (e.g., data deletion requests) to ensure consistent adherence to GDPR and other applicable privacy regulations at scale \n
- Platform Reliability & Availability \n
- Define and maintain SLAs for the data platform, proactively identifying and resolving availability risks \n
- Lead incident response and root cause analysis for data platform outages or degraded performance \n
- Build and maintain runbooks, architecture documentation, and operational playbooks \n
- Lead legacy system migrations with minimal disruption, including data warehouse migrations, pipeline platform transitions, and infrastructure modernization initiatives \n
- Technical Leadership & Collaboration \n
- Serve as a technical mentor and escalation point for data analytics and AI/ML engineering team members \n
- Collaborate with Data Analytics, AI/ML Engineering, Infrastructure, Product, and Software Engineering teams to align infrastructure with roadmap priorities \n
- Contribute to architecture reviews and technology evaluations \n
- Developer Experience & Tooling \n
- Build and maintain internal developer tooling to automate operational tasks (e.g., permission management, environment provisioning, data sanitization) \n
- Establish and enforce code quality standards including linting, formatting, and review workflows across data platform repositories \n
Posted Today