Thales Logo

Thales

Site Reliability Engineer

Posted 3 Days Ago
Be an Early Applicant
In-Office
Austin, TX, USA
Senior level
In-Office
Austin, TX, USA
Senior level
Design, build, and maintain cloud infrastructure and CI/CD for a high-availability telecommunications product. Define SLOs/SLIs, manage incident response and on-call rotations, implement observability, perform performance and capacity planning, run blameless postmortems, and collaborate with security teams to ensure compliance and access control.
The summary above was generated by AI
Location: Austin, United States of America

Thales people architect identity management and data protection solutions at the heart of digital security. Business and governments rely on us to bring trust to the billions of digital interactions they have with people. Our technologies and services help banks exchange funds, people cross borders, energy become smarter and much more. More than 30,000 organizations already rely on us to verify the identities of people and things, grant access to digital services, analyze vast quantities of information and encrypt data to make the connected world more secure.

Austin, TX - Hybrid (3 days a week)

Position Summary

We are seeking a Site Reliability Engineer to ensure the high level of service and operation excellence for the development of the innovative and ambitious Telecommunication solution  (high availability, strong performance constraints) deployed in the public cloud. This product requires the establishment of a product specific SRE team.

Essential Functions

  • Automation & Infrastructure as Code: Design, build, and maintain scalable infrastructure using tools such as Terraform, Ansible, and Kubernetes. Develop automated CI/CD pipelines via GitLab to reduce manual toil.
  • Availability & Reliability Engineering: Define and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Manage "Error Budgets" to balance the velocity of new features with the stability of the platform.
  • Incident Management & On-Call Support: Participate in 24/7 on-call rotations to provide emergency response and perform deep-dive troubleshooting for production issues.
  • Performance & Capacity Planning: Conduct system performance analysis, identify bottlenecks, and perform capacity planning to ensure the infrastructure can handle growth and peak loads.
  • Observability & Monitoring: Implement and refine symptom-based alerting and comprehensive monitoring strategies using platforms like Datadog to ensure high visibility into system health.
  • Continuous Improvement & Postmortems: Lead blameless postmortems after incidents to identify root causes and implement long-term technical fixes to prevent recurrence.
  • Security & Compliance Collaboration: Partner with Cloud Security teams to implement security best practices, manage access controls, and respond to security breaches or vulnerabilities.
  • Support customer relationship
  • Interface with other stakeholders to define solution improvement plan
  • You will have the ownership of solution service availability.

Minimum Requirements

Education:

  • Engineer or equivalent

Experience:

  • at least 5 years of experience

Skills and Abilities:

  • Java development skill is required.
  • You are familiar with Public Cloud (GCP, AWS), containers and microservices (Docker, Kubernetes, Java), CI/CD and automation (Jenkins, Gitlab, Helm), NoSQL database.

Certification

  • GCP cloud architect certification is a plus

Preferred Qualifications

  • You have already set up product monitoring and the underlying infrastructure
  • You have development experience in a distributed systems and/or high availability context
  • You are familiar with microservices development
  • You participated in the definition of architectures, data structures, algorithms with performance, security, reliability constraints, etc.
  • Public cloud architect certification
  • You are interested in aspects of Site Reliability Engineer: CI/CD, automation, monitoring and observability, and continuous improvement.
  • You are an accomplished, versatile and multi-tasking developer engineer.

Must have U.S. or Dual Citizenship and be able to obtain post-hire clearance from the Committee on Foreign Investments in the U.S. (CFIUS) and Department of Treasury

If you’re excited about working with Thales, but not meeting the requirements for this position, we encourage you to join our Talent Community! https://careers.thalesgroup.com/global/en/jointalentcommunity. You can upload your CV and our recruiters can get in touch with any new opportunities that may be of interest to you.

Why Join Us?

Say HI and learn more about working at Thales click here

#LI-MG1

#LI-Hybrid

This position will require successfully completing a post-offer background check. Qualified candidates with [a] criminal history will be considered and are not automatically disqualified, consistent with federal law, state law, and local ordinances.

Thales champions inclusion and we believe diversity strengthens the fabric of our culture. Thales is an Equal Opportunity Employer, including disability/veterans.

If you need an accommodation or assistance in order to apply for a position with Thales, please contact us at [email protected].


The reference Total Target Compensation (TTC) market range for this position, inclusive of annual base salary and the variable compensation target, is between


Total Target Cash (TTC) 109,653.00 - 182,755.00 USD Annual

This reflects how companies in a similar industry and geographic region generally pay for similar jobs. This range helps the Company make pay decisions as one data point among many. Where a position falls within this range is also dependent on other factors including – but not limited to – the employee’s career path history, competencies, skills and performance, as well as the company’s annual salary budget, the customer’s program requirements, and the company’s internal equity. Thales may offer additional benefits and other compensation, depending on circumstances not related to an applicant’s status protected by local, state, or federal law.


(For Internal candidate, if you need more information, please raise HR request through MyThales)


Thales provides an extensive benefits program for all full-time employees working 30 or more hours per week and their eligible dependents, including the following:

•Elective Health, Dental, Vision, FSA/HSA, Voluntary Life and AD&D, Whole Group Life w/LTC, Critical Illness, Hospital Indemnity, Accident Insurance, Legal Plan, Identity Theft, and Pet Insurance

•Retirement Savings Plan after 30 days of employment with a company contribution and a match, and with no vesting period

•Company paid holidays and Paid Time Off

•Company provided Life Insurance, AD&D, Disability, Employee Assistance Plan, and Well-being Program

Thales Austin, Texas, USA Office

9442 Capital of Texas Highway North Suite 400, Austin, United States, 78759

Similar Jobs

Yesterday
Hybrid
179K-225K Annually
Mid level
179K-225K Annually
Mid level
Fintech • Machine Learning • Payments • Software • Financial Services
Lead a portfolio of cloud-native full‑stack and SRE projects, mentor engineers, collaborate with product managers, and deliver resilient AWS/GCP/Azure services using Java, Python, containers, orchestration, and observability tooling.
Top Skills: Ai ToolingAWSDockerGCPHTML/CSSJavaKubernetesAzureNoSQLPythonRdbms
6 Days Ago
Hybrid
Mid level
Mid level
Financial Services
Designs, implements, monitors, and optimizes cloud-based application infrastructure and reliability. Uses IaC/NaC, observability, SLOs, CI/CD pipelines, and enterprise-authorized AI to prevent and resolve incidents, improve availability and scalability, and mentor peers on SRE best practices.
Top Skills: Ci/CdCloudContainer OrchestrationContainersContinuous DeliveryContinuous IntegrationEnterprise-Authorized AiInfrastructure As CodeJavaMonitoringNetwork As CodeNetworkingObservabilityPysparkPythonService Level Objectives (Slos)Spring BootTelemetry Collection
7 Days Ago
Hybrid
Senior level
Senior level
Financial Services
Lead SRE responsible for driving reliability culture, conducting resiliency design reviews, leading incident response, improving service levels with data-driven analytics, adopting AI-assisted SRE workflows, mentoring engineers, and documenting knowledge across the organization.
Top Skills: .NetAi-Assisted ToolsCi/CdDatadogDockerDynatraceEcsGitlabGrafanaJava Spring BootJenkinsKubernetesObservabilityPrometheusPythonSplunkTelemetryTerraform

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account