Unknown Company

Data Operations Engineer

washington, dc • Posted 3 days ago
Onsite Full Time General

Data Operations EngineerWashington, District of Columbia, United StatesAbout the JobEducation / Qualifications / ExperienceB.S. in computer science or information systems fields required, or 5+ years related work experience.Strong analytical, critical thinking skills used to solve complex problemsStrong technical background with a mix of development and automation skillsOutstanding attention to detail and consistently meets deadlinesExceptional communication and interpersonal skillsAbility to work alongside a highly collaborative team, but also a self-starter, able to work independently with little guidanceExperience in troubleshooting, performance tuning, and optimizationProficient in shell scripting, Python, Scala or other programming languagesKnowledge of Spark/PySparkExcellent SQL knowledge, ability to read/write SQL queriesSkilled in Hive (HQL) and HDFSExperience working with both unstructured and structured data sets, including flat files, JSON, XML, ORC, Parquet and AVROComfortable working with big data environments and dealing with large diverse data setsProficient in Linux environmentsFamiliarity with source code management/versioning tools such as GithubUnderstanding of CI/CD principles and best practices in data processingExperience building data visualization dashboards to capture data quality metrics using tools like Tableau, Big Data StudioUnderstanding of public cloud technologies such as AWS, GCP and Azure is a plusMajor Job ResponsibilitiesContribute to the maintenance, documentation, and monitoring of supported data pipelinesContinuously analyze supported data workflows for opportunities to improve reliability and timeliness against established SLAsConceive, develop, and apply improvements to workflows and monitoring to minimize the occurrence and impact of defectsCommunicate with stakeholders when data is in error or is delayed, with clear plans and timelines for recovery, and future preventionDevelop modifications to workflows using git and githubAssist with production support tickets and inquiries from consumers of supported data pipelinesFacilitate the onboarding of new products and pipelines into our suite of supported production processesAbout YouYou are passionate about improving the integrity, accuracy and reliability of data across the organization. You are a highly motivated individual with excellent analytical, critical thinking and problem solving skills. You bring substantial value to the team with your prior experience in building and supporting production data and reporting pipelines.

Strong verbal, written and interpersonal skills provide the flexibility to work collaboratively with a team or independently with minimal supervision. You’re a detail-oriented, self-starter with the ability to multitask and thrive in a dynamic environment. Furthermore, your familiarity with techniques for automating, cleansing and standardizing data at rest, and in motion, make you a great fit for this role.What You’d Be DoingAs a Data Operations Engineer you will be responsible for monitoring and managing maintenance of multiple data ETL pipelines, which power hundreds to thousands of business-critical applications and reports, used by many teams throughout the organization, as well as external customers, every day. Due to the scale, variety, and complexity of the processes we support, standardized or automated practices and tools are needed for monitoring and maintenance of the pipelines.

It is your duty to ensure that monitoring alerts Data Operations to any issues with data pipelines, with appropriate timeliness and sensitivity. Any alerts must be dealt with in a timely and appropriate manner. This can include a variety of things, including job regeneration, communication to stakeholders, definition and development of process enhancements via code, or updates to process documentation. In addition, you will assist with production support inquiries from stakeholders about the quality or timeliness of data in reporting.

Finally, you will work with other teams to facilitate the onboarding of new data pipelines and products into our suite of supported production processes. Strong analytic skills and problem solving skills are needed throughout to model, monitor, and troubleshoot the production processes.

Back to Job Search