Unknown Company

Principal Software Engineer, Core Infrastructure

reston, virginia • Posted Today
Remote Full Time General
hackajob is collaborating with Oracle to connect them with exceptional professionals for this role.
About the team
At Technical Strategy and Oversight (TSO) organization, our mission is to support customer choice, transparency, and value when it comes to cloud infrastructure. We're embarking on ambitious new initiatives such as building new innovative platforms, high performance primitives, frameworks to support OCI developers, and new container runtime that will allow us to run the full variety of OCI services, including our most demanding, high-performance, high-availability services. We're also working on providing canonical implementation of core components for data planes through a data-plane runtime framework, developing a remote persistent storage solution with the latency and performance comparable to that of a local nVME drive, and developing standards and tools to identify critical performance improvements across OCI data-planes.
Description
As Oracle Cloud Infrastructure (OCI) continues its rapid expansion, we are seeking a skilled Software Engineer to join our newly established Cloud Performance Organization. This team plays a key role in addressing service inefficiencies, reducing cloud expenses, improving customer experience, and ensuring scalability. Your work will focus on optimizing the performance of OCI's critical components, internal tools, and applications while fostering a culture of performance engineering.
This is a greenfield opportunity to design and build new cloud services from the ground up. We are growing fast, still at an early stage, and working on ambitious new initiatives. You will be part of a team of smart, motivated, diverse people, and given the autonomy as well as support to do your best work. It is a dynamic and flexible workplace where you'll belong and be encouraged.
Leads development and architecture of scalable, elastic distributed systems for high-throughput, hyperscale workloads. Designs fault-tolerant, highly available systems with robust observability, testing, replication, and resilience mechanisms to meet SLOs. Drives operational readiness, production troubleshooting, and peer mentorship while implementing security, compliance, IaC, and automation for safe patching, upgrades, and rollbacks.
Responsibilities
Key Responsibilities
System Design & Architecture
* Lead development and architecture of scalable, elastic distributed systems for high-throughput, hyperscale workloads.
* Define scalability requirements, optimize performance, and leverage distributed state management and data-plane platforms .
* Design fault-tolerant, highly available systems using redundancy, replication, failover, load shedding, throttling, and rate limiting.
* Establish SLOs, KPIs, telemetry, dashboards, and alerts to ensure reliability and performance .
* Design performance, load, fault-injection, and brownout testing, and implement replication and synchronization for correctness and availability.
Operational Excellence
* Proactively diagnose production issues, guide incident response and root cause analysis , and ensure operational readiness.
* Enable in-service maintenance and upgrades with minimal customer impact.
* Mentor engineers in troubleshooting and operational practices.
Security & Compliance
* Implement encryption, access controls, and security remediation for multi-tenant environments.
* Ensure compliance with applicable standards and maintain required documentation.
Automation & Change Management
* Develop and maintain IaC and automation for cloud infrastructure.
* Enable safe and repeatable patching, updates, and rollbacks through effective change-management practices.
Core Responsibilities
Planning & Execution
* Manage moderately complex initiatives, prioritizing work, timelines, resources, and deliverables while providing technical oversight.
Collaboration & Partnership
* Collaborate across teams and stakeholders to align objectives and deliver solutions that meet business and customer needs .
* Promote inclusive collaboration and diverse perspectives.
Problem Solving
* Analyze and resolve moderately complex issues, escalating critical concerns with clear assessments and recommended solutions.
* Document and share effective problem-solving practices.
Continuous Learning
* Stay current with industry trends and continuously develop technical skills.
* Coach and mentor junior engineers and promote knowledge sharing.
Continuous Improvement
* Identify and implement improvements to processes, workflows, and team effectiveness .
* Evaluate outcomes and incorporate stakeholder feedback.
Performance & Development
* Support talent development through candidate interviews, assessments, and hiring recommendations .
Qualifications
Disclaimer:
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $114,600 to $234,600 per annum. May be eligible for bonus, equity, and compensation deferral.
Oracle maintains broad salary ranges for its roles in order to

Principal Software Engineer, Core Infrastructure in reston at Unknown Company

This position is listed as full time and able to be worked remotely.

Back to Job Search