LPL Financial Logo

LPL Financial

Senior Engineer, Reliability

Posted 18 Days Ago
Be an Early Applicant
In-Office
Austin, TX, USA
102K-169K Annually
Senior level
In-Office
Austin, TX, USA
102K-169K Annually
Senior level
Ensures the reliability, availability, and performance of enterprise observability platforms and supporting applications. Responsibilities include production support, incident response, root cause analysis, platform maintenance, health checks, release execution, CI/CD maturation, monitoring improvements, change management, and operational documentation. The role partners with SRE, Platform Engineering, DevOps, Operations, and global delivery teams to maintain platform stability and drive continuous improvement.
The summary above was generated by AI

Where Ambition Meets Innovation

Build a career that matches all your initiative with an impressive dose of innovation. From cutting-edge resources and a collaborative environment to the freedom to make an impact and more, you’ll find the ingredients you need at LPL Financial to shape your success while helping clients pursue their financial goals.

Job Overview:

The Senior Engineer, Observability and Platform Stability is responsible for ensuring the reliability, availability, and performance of enterprise observability platforms and supporting applications. This role drives operational excellence through proactive monitoring, incident response, platform maintenance, automation, and continuous improvement initiatives. The position partners closely with Site Reliability Engineering (SRE), Platform Engineering, DevOps, Operations, and Incident Management teams to enhance platform stability, mature CI/CD practices, and support strategic technology initiatives. The ideal candidate brings strong cloud technology experience and a proven ability to support production environments while delivering actionable insights through observability capabilities.

Responsibilities:

Delivery Support

  • Partner with Product Owners and engineering teams to provide operational and release support for technology initiatives.

  • Ensure platform changes meet operational readiness requirements, including rollback procedures, runbook documentation, integration standards, and support handoffs.

  • Maintain production stability throughout platform upgrades, enhancements, and enterprise initiatives.

  • Support platform ownership transitions and operational readiness activities across global delivery teams.

Production Support & Incident Response

  • Troubleshoot application and platform issues to restore services and minimize business impact.

  • Serve as an escalation point for complex production incidents and operational challenges.

  • Participate in incident triage, root cause analysis, corrective action planning, and resolution activities.

  • Collaborate with engineering teams to implement long-term solutions that reduce recurring incidents.

  • Support mission-critical environments through timely incident response and service restoration.

Platform Stability & Proactive Operations

  • Conduct health checks, configuration reviews, and performance assessments to identify operational risks.

  • Support reliability initiatives focused on improving recovery times, reducing incident recurrence, and minimizing change-related defects.

  • Validate vendor releases, hotfixes, and configuration changes prior to production deployment.

  • Partner with observability and analytics teams to enhance monitoring, alerting, and issue detection capabilities.

Release Execution & CI/CD Maturation

  • Execute platform changes through established SDLC, change management, and release management processes.

  • Collaborate with Platform Engineering, DevOps, Quality Engineering, and Scrum teams to improve release and deployment practices.

  • Support the adoption of source control, environment separation, release automation, and CI/CD capabilities.

  • Ensure solutions are testable, deployable, and operationally supported before and after production implementation.

Documentation & Operational Excellence

  • Maintain runbooks, support documentation, configuration records, incident playbooks, and release procedures.

  • Document incident findings, lessons learned, and process improvement opportunities.

  • Contribute to the development of standardized, repeatable, and scalable operational practices.

What are we looking for?

We seek professionals who pursue greatness, act with integrity, are driven to help our clients succeed, win together, and create and share joy. The ideal candidate brings strong technical expertise in observability platforms, site reliability engineering, production support, release management, and operational excellence while demonstrating a commitment to platform reliability, cross-functional collaboration, and continuous improvement.

Requirements:

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field.

  • 6+ years of experience in SRE, platform engineering, production support, enterprise application operations, or related technology environments.

  • 5+ years of experience supporting enterprise platforms, including AWS, Dynatrace, ELK, ServiceNow, and SolarWinds.

  • Experience troubleshooting complex production incidents within enterprise-scale technology environments.

  • Experience executing technology changes through formal change management and release management processes.

Preferences:

  • Experience within financial services or another regulated industry.

  • Experience implementing or supporting CI/CD pipelines, release automation, and deployment processes across development, testing, and production environments.

  • Experience with observability platforms, monitoring tools, performance dashboards, or application monitoring solutions.

  • Experience collaborating with offshore, nearshore, or global delivery teams.


 

Pay Range:

$101,558.00 - $169,229.00
 
Actual base salary varies based on factors, including but not limited to, relevant skill, prior experience, education, base salary of internal peers, demonstrated performance, and geographic location. Additionally, LPL Total Rewards package is highly competitive, designed to support your success at work, at home, and at play – such as 401K matching, health benefits, employee stock options, paid time off, volunteer time off, and more. Your recruiter will be happy to discuss all that LPL has to offer!
 

Company Overview:

LPL Financial Holdings Inc. (Nasdaq: LPLA) is among the fastest growing wealth management firms in the U.S. As a leader in the financial advisor-mediated marketplace(6) , LPL supports over 32,000 financial advisors and the wealth management practices of approximately 1,100 financial institutions, servicing and custodying approximately $2.3 trillion in brokerage and advisory assets on behalf of approximately 8 million Americans. The firm provides a wide range of advisor affiliation models, investment solutions, fintech tools and practice management services, ensuring that advisors and institutions have the flexibility to choose the business model, services, and technology resources they need to run thriving businesses. For further information about LPL, please visit www.lpl.com.


At LPL, independence means that advisors and institution leaders have the freedom they deserve to choose the business model, services, and technology resources that allow them to run a thriving business. They have the flexibility to do business their way. And they have the freedom to manage their client relationships, because they know their clients best. Simply put, we take care of our advisors and institutions, so they can take care of their clients.


For further information about LPL, please visit www.lpl.com.


Join the LPL team and help us make a difference by turning life’s aspirations into financial realities. Please log in or create an account to apply to this position. Principals only. EOE.


Information on Interviews:

LPL will only communicate with a job applicant directly from an @lplfinancial.com email address and will never conduct an interview online or in a chatroom forum.  During an interview, LPL will not request any form of payment from the applicant, or information regarding an applicant’s bank or credit card.  Should you have any questions regarding the application process, please contact LPL’s Human Resources Solutions Center at (855) 575-6947.


EAC 5.19.26

LPL Financial Austin, Texas, USA Office

Austin, United States

Similar Jobs

3 Days Ago
In-Office
107K-223K Annually
Senior level
107K-223K Annually
Senior level
Other • Utilities
Strengthen reliability and resilience across T-Mobile payment systems by automating deployment and operational processes, managing cloud infrastructure, responding to incidents, conducting root cause analysis, and improving distributed database scalability, system stability, and software delivery.
Top Skills: AWSBashCassandraCi/CdInfrastructure As CodeKubernetesNoSQLPythonSQL
5 Days Ago
In-Office
110K-210K Annually
Senior level
110K-210K Annually
Senior level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Lead aircraft reliability engineering across development, integration, testing, and operations. Establish reliability requirements and budgets, perform predictions and analyses including FMEA, FTA, RBD, and Weibull analysis, manage FRACAS and reliability growth activities, analyze test and field data, assess supplier performance, and recommend design improvements. Communicate reliability risks and metrics to engineering leadership and customers while mentoring junior engineers and improving organizational processes.
Top Skills: Aerospace SystemsEprdFault Tree AnalysisFmeaFmecaFracasIsographMil-Hdbk-217Model-Based Systems EngineeringMtbcfMtbfNprdRam CommanderReliability Block DiagramsReliability EngineeringReliasoftRequirements Management ToolsRiac 217PlusWeibull AnalysisWindchill Quality Solutions
6 Days Ago
Hybrid
151K-187K Annually
Senior level
151K-187K Annually
Senior level
Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Design and implement reliable, scalable IT infrastructure; automate processes; monitor systems; resolve incidents; conduct performance and load testing; optimize cloud environments; maintain storage architecture; and lead continuous improvement. The role requires troubleshooting complex system issues, developing automation solutions, managing incidents, collaborating across teams, mentoring others, and supporting secure business operations across AWS, Google Cloud, and Microsoft Azure.
Top Skills: AWSGoogle Cloud PlatformAzure

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account