Advertisement
H
Salary: Competitive and Location-Based
Posted August 21, 2026
30 views
0 apply clicks
Role
About the Role
Honeycomb is seeking a high-impact Field Reliability Engineer based in LATAM to join our Platform Engineering team. As a pioneer in the observability space, Honeycomb is redefining how developers interact with their production systems, working with industry leaders like Slack, HelloFresh, and Vanguard. In this role, you will be a key contributor to our Managed Services and Infrastructure team, ensuring the reliability, scalability, and performance of our sophisticated customer-facing deployments. You will join a fully distributed, fiercely inclusive team of engineers who value autonomy, technical excellence, and a deep commitment to our customers' success.
Key Responsibilities- Own and operate business-critical customer-facing managed infrastructure, including Refinery as a Service (RaaS) and Honeycomb Private Cloud (HnyPC) deployments across diverse AWS accounts and regions.
- Design, build, and maintain robust Terraform modules and Helm charts to standardize and automate the provisioning of cloud infrastructure.
- Develop and refine deployment automation pipelines to ensure seamless updates and high availability for managed services.
- Collaborate with cross-functional teams to troubleshoot complex distributed systems issues and optimize performance for high-throughput observability data.
- Act as a technical bridge for our LATAM-based customers, providing expert guidance on infrastructure reliability and architectural best practices.
- Contribute to the evolution of our platform by identifying opportunities for automation, cost optimization, and improved operational resilience.
- Participate in a distributed on-call rotation, ensuring the health and stability of our global managed service offerings.
- Proven experience as a Site Reliability Engineer, DevOps Engineer, or Platform Engineer managing production-grade distributed systems.
- Deep expertise in Amazon Web Services (AWS) ecosystem, including multi-region and multi-account architecture management.
- Strong proficiency with Infrastructure as Code (IaC) tools, specifically Terraform, and container orchestration using Kubernetes and Helm.
- Experience managing managed services or private cloud deployments for external customers at scale.
- Proficiency in at least one programming or scripting language, such as Go, Python, or Ruby, for automation and systems engineering tasks.
- Solid understanding of observability principles and experience with monitoring, logging, and distributed tracing tools.
- Strong communication skills in English, with the ability to work effectively across time zones in a fully remote, distributed environment.
- A proactive mindset with the ability to take full ownership of projects from inception to production operation.
- A fully remote and distributed work environment that prioritizes impact and delivery over physical presence.
- Competitive compensation package including equity in a fast-growing, Series D startup named to Forbes’ Best Startups list.
- A culture of high trust, autonomy, and accountability where your contributions are valued from Day 1.
- Investment in professional development and the opportunity to work with a world-class team of humble, talented engineers.
- Comprehensive health and wellness benefits tailored to our LATAM workforce.
- Generous time-off policies to ensure a healthy work-life balance while building the future of observability.
Advertisement
Skills
Required Skills
AWS
Terraform
Kubernetes
Helm
Go
Distributed Systems
SRE
Infrastructure as Code
Interested in this role?
Sign in to your free seeker account to apply.
Advertisement