Resident Data Architect - SageMaker Unified Studio, Data Foundation
Office in India: Gurugram, & 5 others
Data Solution Architecture& 6 others
Looking for something else?
Find a vacancy that works for you. Send us your CV to receive a personalized offer.
Find me a jobChoose an option
We're seeking a Resident Data Architect to support Amazon SageMaker Unified Studio along with the wider data infrastructure for an extended engagement, with focus areas spanning serverless data pipelines, contemporary data lake design, and agentic AI integration built around AWS cloud-native principles.
Responsibilities
- Configure, administer, and maintain Amazon SageMaker Unified Studio, covering data catalog management, blueprint handling, and serverless compute functionality
- Facilitate self-service data exploration, discovery, and analysis capabilities for both technical and business users through SMUS
- Keep SMUS blueprints and templates up to date to enable repeatable and well-governed data workflows
- Supervise SMUS IAM domains, access control policies, and governance-related configurations
- Design the client's underlying data architecture on AWS with attention to scalability, governance, and performance
- Refine data lake structures leveraging Amazon S3, incorporating modern storage formats like Parquet, Iceberg, and Delta
- Put in place strategies for data cataloging, partitioning, and lifecycle management across large-scale data environments
- Bring Amazon OpenSearch Service into the architecture to support search, analytics, and data exploration needs
- Construct serverless data pipelines utilizing AWS Glue and AWS Lambda
- Build ETL/ELT jobs to handle data ingestion, transformation, and delivery throughout the data platform
- Facilitate the integration and rollout of Quick Suite to support analytics, dashboards, and self-service reporting capabilities
- Apply principles from the AWS Well-Architected Framework, with particular attention to security, reliability, performance, and cost efficiency
Requirements
- 8 or more years of background in data architecture, data engineering, or a related discipline within AWS cloud settings
- Strong command of Amazon SageMaker Unified Studio and data lake architectural design
- Skilled in AWS Glue and AWS Lambda for building serverless data compute workflows
- Capability working with Amazon S3, including modern storage formats such as Parquet, Iceberg, and Delta
- Background using Amazon OpenSearch Service for search, analytics, and data exploration purposes
- Solid grasp of data foundation architecture, covering cataloging, governance, and self-service capabilities
- English language proficiency at a B2 level or higher
Nice to have
- Familiarity with Quick Suite or Amazon QuickSight for building analytics and dashboard solutions
- Awareness of agentic AI services, such as Bedrock Agents and AgentCore
- Working knowledge of Amazon Lake Formation
- Capability with infrastructure as code using CDK or Terraform
- Strong command of Python
