Senior Software Engineer (Backend/Fullstack) — Reliability Focused
CodeRoad
Latin AmericaRemote
Try resume matching
Upload your resume to see how it matches a sample of current jobs. No account needed.
About this role
Latin America | 100% Remote
A bout the Role
We're looking for a senior engineer who can design, build, and operate production systems end-to-end. You'll work across our core stack building features and services, while also owning the reliability, observability, and performance of what you ship. You won't be handing your code off to a separate ops team — you'll be on-call for it, instrumenting it, and improving it based on how it behaves in production.
What You'll Do
Design and build backend/fullstack features across Java 21, Spring/Spring boot, JPA/Hibernate, NodeJS, React/React Native
Instrument services with logging, metrics, and tracing (e.g., Splunk, Prometheus, Grafana, Datadog, OpenTelemetry) as a standard part of development, not an afterthought
Define and monitor SLIs/SLOs for the services you own; use error budgets to guide prioritization between feature work and reliability work
Participate in an on-call rotation; respond to and resolve production incidents affecting your services
Write and maintain postmortems/RCAs, and drive follow-up fixes to prevent recurrence
Contribute to infrastructure-as-code (Terraform, CloudFormation, etc.) for the systems you build
Perform capacity planning and load testing for services ahead of scale events
Collaborate with platform/infra teams on shared tooling, but take primary ownership of your service's health
Participate in code reviews, architecture discussions, and mentor junior engineers
What We're Looking For
5+ years of professional software engineering experience, with deep expertise in Java 21, Spring/Spring boot, JPA/Hibernate, NodeJS, React/React Native
Demonstrated experience owning services in production — not just writing code, but debugging, scaling, and maintaining it live
Comfort with observability tooling and reading dashboards/logs/traces to diagnose issues under pressure
Experience with incident response processes (on-call, paging, postmortems)
Working knowledge of cloud infrastructure with AWS and containerization (Docker, Kubernetes) — enough to reason about deployment and scaling, even if you're not a dedicated platform engineer
Strong communication skills — you can explain a production issue to both engineers and stakeholders
Bonus: experience with CI/CD pipelines, infrastructure-as-code, or chaos engineering practices
What This Role Is Not
This is not a dedicated SRE/DevOps/Platform Engineering role. You won't be building the observability platform itself or managing infrastructure for other teams — you'll be a strong practitioner of reliability engineering within your own feature work.
What You’ll Love
100% Remote
Holidays off
Paid Time Off
Health insurance assistance
Competitive USD compensation
Growth opportunities