SonicJobs Logo
Left arrow iconBack to search

Principal Data Engineer - AWS

Metis Technology Solutions Inc
Posted a day ago, valid for a month
Location

Seven Trees, CA, US

Salary

$140,000 - $200,000 per year

Contract type

Full Time

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.

Sonic Summary

info
  • Metis Technology Solutions is looking for a Principal Data Engineer – AWS with a minimum of 10 years of relevant experience to join their Software Operations team.
  • The role involves designing and maintaining scalable ETL/ELT data pipelines in AWS, along with developing production-quality data-processing software using Python and SQL.
  • Candidates must have at least 5 years of hands-on experience with AWS services and strong SQL skills, particularly with relational databases like PostgreSQL/PostGIS.
  • The position requires collaboration with various stakeholders and the implementation of AWS security practices, along with the ability to diagnose complex problems in data pipelines and cloud infrastructure.
  • The salary for this position is competitive and commensurate with experience, and applicants must be eligible for a U.S. Government Public Trust Clearance.

POSITION SUMMARY: 

Metis Technology Solutions is seeking an experienced Principal Data Engineer – AWS to join our Software Operations team. This engineer will work closely with software developers, researchers, analysts, database engineers, and system administrators to design, develop, deploy, and operate the data infrastructure and pipelines connecting research platforms, external data sources, cloud-based services, and the project's Sherlock data warehouse. 


This position:

  • Architects, designs, develops, deploys, and maintains scalable and reliable ETL/ELT data pipelines in Amazon Web Services (AWS).
  • Designs automated ingestion and processing solutions for structured, semi-structured, unstructured, batch, and streaming data.
  • Develops production-quality data-processing and pipeline software using Python, SQL, and shell scripting.
  • Designs and maintain workflow orchestration solutions using technologies such as Apache Airflow, Dagster, AWS Step Functions, or equivalent platforms.
  • Designs and implements real-time and near-real-time streaming and event-driven data pipelines using technologies such as AWS Kinesis, Amazon Data Firehose, Kafka, RabbitMQ, SQS/SNS, or equivalent technologies.
  • Develops and maintains cloud data solutions using AWS services such as Amazon S3, AWS Lambda, Amazon Redshift, Amazon RDS, Amazon DynamoDB, Amazon EC2, API Gateway, IAM, and CloudWatch, as appropriate to project requirements.
  • Administers and optimizes relational databases and cloud data warehouses, including PostgreSQL/PostGIS and Amazon Redshift or comparable technologies.
  • Develops and maintains infrastructure as code (IaC) using Terraform or equivalent technologies.
  • Develops and maintains automated deployment and CI/CD processes for data applications and supporting infrastructure.
  • Designs pipelines and supporting infrastructure for reliability, scalability, maintainability, observability, security, and efficient use of AWS resources.
  • Implements appropriate AWS security practices, including IAM roles and policies, encryption, secrets management, network controls, logging, and least-privilege access.
  • Develops monitoring, logging, metrics, and alerting that provide operational visibility into data-pipeline health and performance.
  • Documents data architectures, data flows, interfaces, infrastructure, deployment processes, operational procedures, and troubleshooting practices.
  • Collaborates with researchers, analysts, and application developers to translate research and application requirements into reliable and maintainable data-processing solutions.


MINIMUM QUALIFICATIONS:

 

Education: 

Bachelor’s degree or higher in computer science, computer engineering, information systems, or a related technical discipline.



Required Skills and knowledge:


  • Minimum 10 years of progressively responsible software engineering, data engineering, database engineering, or closely related technical experience.
  • Minimum 5 years of substantial hands-on AWS experience, including the design, implementation, deployment, and operation of production data-processing or data-pipeline solutions.
  • Demonstrated experience architecting and developing ETL/ELT pipelines involving large, heterogeneous, or rapidly changing datasets.
  • Strong hands-on experience with AWS data and compute services. Relevant technologies may include S3, Lambda, Redshift, RDS, DynamoDB, Kinesis, Data Firehose, EC2, API Gateway, IAM, and CloudWatch.
  • Advanced SQL skills and substantial experience working with relational database systems, preferably PostgreSQL/PostGIS.
  • Demonstrated experience with cloud data warehouses such as Amazon Redshift or comparable technologies.
  • Experience with NoSQL databases or document/key-value data stores such as DynamoDB, MongoDB, or comparable technologies.
  • Demonstrated experience with data-pipeline and workflow orchestration using Apache Airflow, Dagster, AWS Step Functions, or comparable technologies.
  • Experience designing or implementing streaming, message-oriented, or event-driven data-processing systems using technologies such as Kinesis, Data Firehose, Kafka, RabbitMQ, SQS/SNS, or comparable technologies.
  • Demonstrated experience diagnosing and correcting database and data-pipeline performance problems.
  • Hands-on experience with infrastructure as code, preferably Terraform or an equivalent technology.
  • Working knowledge of AWS security concepts including IAM, encryption, secrets management, network security, logging, and least-privilege access.
  • Strong understanding of modern software-engineering practices, including modular design, automated testing, version control, documentation, security, and maintainable code.
  • Strong analytical and troubleshooting skills and demonstrated ability to diagnose complex problems spanning applications, databases, data pipelines, and cloud infrastructure.
  • Proven ability to evaluate technical alternatives, make sound architectural decisions, and communicate the rationale and tradeoffs associated with those decisions.
  • Excellent written and verbal communication skills and demonstrated ability to collaborate effectively with software engineers, researchers, analysts, system administrators, and other technical stakeholders.


LOCATION SPECIFIC REQUIRMENTS: NASA Ames Research Center, Moffett Field CA


SECURITY CLEARANCE: Applicant must be eligible to obtain a U.S. Government Public Trust Clearance. Must be a U.S. Citizen or Permanent Resident.


EEOE Including Vets and Disability






Learn more about this Employer on their Career Site

Apply now in a few quick clicks

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.