Jobs / Del***
Applied AI Site Reliability Engineer III - PxE Talent
Del*** · Dallas, TX, United States
Visa sponsorship details are locked. Unlock company name and apply link with .
Dallas, TX, United States102,500-210,600 USD/yearlyRemote
Remuneration
102,500-210,600 USD/yearly
Location
Dallas, TX, United States
Visa sponsorship
Sponsors visa
Job summary
Applied AI Site Reliability Engineer III Role Overview: As an Applied AI Site Reliability Engineer III , you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity of high-visibility products and platforms and the environments they run in.
Qualifications
- into service-level objectives, runbooks, and production tooling.
- Be a valuable, flexible, and dedicated team member, supportive of teammates, and focused on quality and tech debt payoff.
- Effective Communication and Influence: Exhibit exceptional communication
- A bachelor's degree in computer science, software engineering, data science, machine learning, or related discipline.
- Experience is the most relevant factor.
- NET, SQL/NoSQL, Kubernetes, Terraform, ArgoCD, as well as CI/CD and observability stacks.
Responsibilities
- Outcome-Driven Accountability: Embrace and drive a culture of accountability for reliability, performance, and cost outcomes, measured in service-level objectives and error budgets, not raw uptime.
- Engineering Craftsmanship: Maintain accountability for the operational integrity of production and pre-production environments, and for the production standards that systems are admitted against.
- Stay hands-on, self-driven, and continuously learn new approaches, languages, and frameworks-operating as an infrastructure-focused engineer, not a tool operator.
- Demonstrate collaborative
- line as a dedicated, embedded function while partnering with security and risk on the control objectives you enforce.
- Foster a collaborative environment that enhances team synergy and innovation.
- Strive to be a role model, leveraging these techniques to optimize reliability, performance, and operational delivery.
- Demonstrate strong understanding of the full lifecycle of platform and product development, focusing on continuous improvement and learning.
- Translate reliability needs, reference architectures, and operational
- controls (least-privilege/RBAC, deploy approvals, secrets management) in partnership with security and risk.
- Prior experience using methodologies &
Skills
CommunicationLeadership
Degrees
AssociateBachelorDegree
Work schedule
On-call
Travel
Travel
Industry
AutomotiveEnergyOil-gas
Company size
Smb
Security clearance
Secret