Job Title: Remote Data ArchitectJoin an innovative R&D group focused on enhancing the efficiency and automation of their data pipeline, including data collection, transformation, and system architecture. Work with advanced technologies such as smart displays on off-highway equipment, telematics, edge devices, and cloud computing.ResponsibilitiesDesign and architect scalable and reliable data pipelines to automate data processing workflows.Develop and test software to implement prototype code to prove out designs.Collaborate with data scientists, engineers, and stakeholders to understand data needs and requirements.Develop and implement ETL processes to transform and load data from various sources into our data warehouse.Optimize data pipelines for performance, scalability, and efficiency.Create data architectural designs and develop related best practices.Transform business requirements into conceptual, logical, and physical data models.Promote and support a data-driven culture of quality, collaboration, rapid delivery, and business impact.Collaborate with Product Managers, Delivery leads, System Engineers, Data Analysts, Data Stewards, Business Stakeholders, and Software Development teams during the solution design process.Author, refine, and collaborate on the improvement of corporate data architecture standards, procedures, and metrics.Communicate data architecture-related concepts to both business and development teams.Integrate Data Architecture with the existing software delivery deployment process.Essential Skills7+ years of experience with Data Modeling, Data Warehousing, and working with large-scale datasets.Experience working with Business Architects, System Architects, Engineering, Data Architects, and/or Data Stewards to capture business requirements in a Logical Data Model.Expertise in data architecture best practices, relational data modeling, and data warehouse/data mart concepts.Experience in integrating Data Architecture with CI/CD processes.Comprehensive understanding of data lineage, data mapping, data security/confidentiality/privacy, and related data standards.Deep knowledge of entity-relationship diagrams, normalization, abstraction, denormalization, dimensional modeling, and metadata modeling practices.Working knowledge of Relational Database Management Systems (RDBMS) and SQL.Strong collaboration skills with experience driving conversations and influencing complex data design challenges.Ability to make decisions with limited knowledge while balancing best practices, schedule, and business needs.Strong communication skills to model and convey complex concepts both in writing and verbally.Ability to work with limited supervision, break down complex data architectural requirements, prioritize work, and execute.BS/MS in Computer Engineering or Computer Science.Proficiency in programming languages such as Python, Golang, or others.Experience with ETL tools and frameworks (Apache tools, AWS Lambda, etc.).Experience with cloud platforms (e.g., AWS, Azure, Google Cloud) and their data services and deployment strategies.Knowledge of data modeling, database design, and data warehousing concepts.Experience with container orchestration tools (Docker, Kubernetes).Ability to create supporting documentation such as design documents, architecture diagrams, test procedures, and reports.Good oral and written communication skills with the ability to professionally support periodic communication to management and technical teams.Experience with Agile Scrum development methodologies and common workflow tools: GIT, Jira, etc.Ability to code (hands-on architect).Experience with ER/Studio.Additional Skills & QualificationsExperience in the off-highway heavy machinery, automotive, or industrial control industry.Experience designing software for a distributed ECU system using CAN or Ethernet communication.Understanding of machine learning and AI concepts.Experience developing software in additional programming languages.5+ years of experience leveraging metadata within repositories.Working experience in AWS or exposure to other cloud technologies.Working knowledge of or exposure to Data Engineering tools and technologies (e.g., Databricks, Data Lake concepts).Working knowledge of AWS services such as Lambda, RDS, ECS, DynamoDB, API Gateway, S3, etc.Experience with REST API development or providing data architectural support during API development.Experience implementing and/or leveraging Data Governance and Stewardship program capabilities.Passionate, creative, and eager to learn new complex technical areas.Accountable, curious, and collaborative with a focus on quality.Skilled in interpersonal communications and ability to communicate complex topics to non-technical audiences.Experience working in an agile team environment.Ability to function in a fast-paced, collaborative team environment distributed across time zones and locations.Work EnvironmentWork in a dynamic R&D environment utilizing advanced technologies such as telematics, edge devices, and cloud computing. Collaborate closely with cross-functional teams to drive innovation and efficiency in data architecture and processing workflows.
The role may require working with diverse technologies and tools, including AWS, Python, Golang, Docker, and Kubernetes.