At BairesDev®, we've been leading the way in technology projects for over 15 years. We deliver cutting-edge solutions to giants like Google and the most innovative startups in Silicon Valley.
Our diverse 4,000+ team, composed of the world's Top 1% of tech talent, works remotely on roles that drive significant impact worldwide.
When you apply for this position, you're taking the first step in a process that goes beyond the ordinary. We aim to align your passions and skills with our vacancies, setting you on a path to exceptional career development and success.
SRE (Observability) at BairesDev
As an SRE specialized in Observability, you will architect and maintain the infrastructure required to gain deep visibility into complex, large-scale systems. You will act as a key link between systems reliability and actionable data, ensuring that engineering teams can detect and resolve issues before they impact the user experience.
What You'll Do
Design and deploy scalable monitoring architectures that provide real-time insights across metrics, logs, and traces.
Build and optimize data pipelines for centralized log aggregation and high-cardinality telemetry.
Implement distributed tracing frameworks to map service dependencies and identify latency bottlenecks across microservices.
Develop intelligent alerting strategies and automated runbooks to reduce noise and improve incident response efficiency.
Define and track reliability governance standards, including SLOs, SLIs, and error budgets, to drive data-driven decision-making.
Collaborate with development teams to instrument applications and embed observability best practices into the software lifecycle.
What We Are Looking For
4+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure.
Proven expertise in designing and operating large-scale observability infrastructure for metrics, logs, and traces.
Proficiency with monitoring and visualization tools such as Prometheus, Grafana, and the ELK stac