Senior Site Reliability Engineer (SRE) Job at GT, Brazil

QlM3VXpuUFJhcnYxWDRqQUlRaTlRMnFmaUE9PQ==
  • GT
  • Brazil

Job Description

About the Project

You’ll join a consumer mobile product with an engineering and product organization of around 50 people distributed across Europe and the US.

The team works in small, autonomous product squads, each responsible for a specific area of the product and critical user journeys. Because the team operates across multiple regions without a full follow-the-sun model, strong observability, monitoring and reliable incident response are essential.

From a technical perspective, the team is focused on building reliable, observable systems that allow engineers to identify issues early, understand their impact and respond quickly when incidents occur.

  • Technology stack: Node.js , TypeScript , AWS , Cloudflare , CloudWatch , Sentry . React Native is used on the mobile side.

  • Team: Cross-functional product squads of approximately 6–8 people , distributed across Europe, the US and LATAM.

About the Role

We are looking for an experienced Site Reliability Engineer with a strong backend engineering background in Node.js and TypeScript .

The ideal profile is someone who started in backend/software engineering and has moved into SRE or reliability-focused work, combining a strong understanding of application code with hands-on experience in observability, monitoring and production incident management.

You will be embedded within a product squad and take ownership of the reliability and observability of critical user journeys. An important part of the role is being able to interpret production signals, identify when something is going wrong and begin mitigating incidents independently while bringing in the wider engineering team when needed.

Responsibilities:

  • Own observability for critical product and user journeys within your squad.

  • Define, build and maintain meaningful metrics, dashboards and alerts.

  • Define and maintain SLIs/SLOs for key services and product-level metrics.

  • Improve monitoring, logging, tracing and alerting across the squad’s systems.

  • Act as the first responder for critical P0/P1 production incidents, including out-of-hours incidents.

  • Investigate production signals, identify potential root causes and begin mitigating issues independently.

  • Coordinate with other engineers when broader support or escalation is required.

  • Participate in incident triage, mitigation and postmortems.

  • Identify recurring reliability issues and drive improvements to infrastructure, tooling and incident-response processes.

  • Work closely with backend and product engineers in a distributed, autonomous squad.

Essential knowledge, skills & experience:

  • Strong previous experience as a Backend / Software Engineer , with senior-level hands-on experience in Node.js and TypeScript .

  • Hands-on experience working in an SRE, Production Engineering or similar reliability-focused role .

  • Strong production experience with AWS .

  • Experience with Cloudflare and CloudWatch .

  • Experience with monitoring and observability across metrics, logging, tracing and alerting .

  • Practical experience responding to production incidents, including triage, mitigation and postmortems .

  • Ability to interpret monitoring signals and independently investigate and begin resolving production issues.

  • Understanding of both the application and infrastructure layers rather than infrastructure-only experience.

  • Strong communication skills and the ability to work autonomously within a distributed engineering team.

  • Comfortable participating in out-of-hours incident response as part of the team’s coverage model.

Nice-to-have

  • Experience with observability tools such as Sentry .

  • Experience defining SLIs and SLOs for product-level metrics.

  • Experience with React Native or exposure to mobile application environments.

  • Previous experience with consumer mobile products or high-traffic B2C systems.

Job Tags

Full time

Similar Jobs

Crossroads Campaigns Solutions

Web and Graphic Designer Job at Crossroads Campaigns Solutions

 ...provide top-notch service and a high return on investment for our clients. We are looking for a part-time, mid-level web and graphic designer to join our growing team. Our ideal candidate thrives in a creative setting using design software such as Adobe's Creative Suite,... 

Everglades Equipment Group

Parts Support Representative Job at Everglades Equipment Group

Position Specifics Location: North Port Reports to: Jackson Sherer Supervises: None Parts Support Representative Job Summary Build customer relationships by traveling to customers locations, pass on any customer concerns to the CSR and Jackson, organize consignment cabinets...

NBBJ

Intermediate Designer - Corporate Commercial Practice Job at NBBJ

 ...NBBJ is an award-winning design firm recognized as aTIME100 Most Influential Company, aFast Company Most Innovative Architecture...  ...and Urban Environment projects. We also have several areas of service expertise including: Architecture, Environmental Graphic Design... 

Merck & Co.

Senior Director, Advanced Pharmacometrics, Quantitative Pharmacology and Pharmacometrics - Immuno-Oncology (Hybrid or Remote) Job at Merck & Co.

Job DescriptionWe are seeking an accomplished scientific leader to join the Quantitative Pharmacology and Pharmacometrics - Immuno-Oncology (QP2-IO) team as-Senior Director, Advanced Pharmacometrics. QP2-IO is part of the Global Clinical Development organization and is ...

Med Source Consultants

Family Medicine Physician - 4109 Job at Med Source Consultants

 ...Family Medicine Physician 4109 Family Medicine Physician for Vibrant Waterfront Community in Fairfield County, CT Family Medicine Physician for Vibrant Waterfront Community in Fairfield County, CT Job Type Permanent Specialty Family Medicine/General...