- Location
- Rio de Janeiro
- Workplace
- Onsite
- Type
- Full-time
- Department
- Engineering
- Seniority
- Senior
- Experience
- 6+ years
- Closing date
- Today
- Source
- CareersPage
Description
About hireworks
hireworks is building a community of top talent in key international markets by unlocking unparalleled access to positions at leading U.S. based companies. As your employer, hireworks will ensure you have a seamless interview, onboarding, and employee experience - providing ongoing support and resources along the way. Established in 2023, hireworks is forging corp-to-corp relationships with leading U.S. based organizations looking to grow their teams with best-in-class talent around the world. Working with hireworks means unlocking access to a network of local peers and mentors and career opportunities through our client network.
About Our client
A leading provider of innovative software and services to K-12 schools and educational institutions worldwide. Their mission is to empower schools to transform education through technology. They believe in the power of education to shape the future, and are committed to helping schools prepare students for success and make a positive impact on the world.
Position Overview
Our client's Events platform powers event marketing, ticket sales, ticket management, and on-site ticket distribution and scanning for organizers running live events. It's a high-throughput, event-driven application on AWS, and it needs a Senior SRE to keep it reliable, well-instrumented, and continuously improving in place.
You will operate and strengthen the platform's AWS infrastructure, CI/CD, and observability, working closely with the engineers who build and evolve the application so the system holds up under real event-day load — on-sales, high-traffic scans, and everything in between.
Key Responsibilities
Operate and maintain the Events platform's AWS infrastructure — compute, data stores, CDN, and the event bus connecting bookings, tickets, users, and scanning
Build and extend observability — dashboards, tracing, and alerting that catches problems before customers do, especially around live on-sales and event days
Maintain and improve CI/CD pipelines for the booking, box office, and ticket-scanning applications, working toward deployments that are repeatable, auditable, and low-risk
Partner with the engineers building and evolving the Events platform, understanding the event-driven architecture well enough to contribute code, diagnose production issues, and suggest infrastructure improvements
Participate in on-call and incident response for the platform, with particular attention to the operational patterns unique to ticket scanning at scale
Apply IaC discipline to infrastructure changes, avoiding manual, untracked changes to production
Use AI-assisted development tools (Claude Code or comparable) as part of your workflow to navigate and document a complex codebase and infrastructure estate
Required Qualifications
~6 years of experience as an SRE, DevOps engineer, or backend/infrastructure engineer supporting production systems
Strong, hands-on AWS experience — comfortable operating services like ECS, Lambda, Aurora/RDS, CloudFront, and IAM in a real production environment
Solid understanding of event-driven architecture — message/event buses (EventBridge, SQS/SNS, Kafka, or comparable), asynchronous workflows, and the failure modes specific to distributed, event-driven systems
Proficient as a software engineer, not just an operator — comfortable reading and writing code (Node.js, TypeScript, Python, C#/.NET, or similar) to fix issues, write tooling, and contribute directly to the applications you support
Experience with CI/CD pipelines and deployment automation
Working knowledge of observability practices — monitoring, alerting, and using tracing/logging tools (CloudWatch, X-Ray, Datadog, or similar) to diagnose production issues
Familiarity with Infrastructure-as-Code (CloudFormation, CDK, Terraform, or comparable)
Comfortable using AI-assisted development tools such as Claude Code, Codex, or similar to navigate and understand large codebases
Preferred Qualifications
Experience with ticketing, event-management, booking, or e-commerce platforms, especially systems with sharp demand spikes (on-sales, high-traffic events)
Exposure to Azure DevOps
Familiarity with high-throughput, queue-driven architectures and autoscaling patterns (SQS, KEDA, or similar)
Experience with mobile or field-device systems that must operate reliably online and offline (e.g., ticket scanning, POS, field service apps)
Exposure to Angular, React Native, or .NET applications, even if your primary focus is infrastructure
Background in EdTech, SaaS, or other domains where system reliability has direct, real-time consequences for end users
Benefits:
hireworks is cultivating a growing community of top talent across LATAM. In addition to unlocking access to positions at top tier U.S. based companies, we offer a variety of benefits to enhance your experience:
Competitive Pay - compensation that reflects your experience and accomplishments
Remote Flexibility - work from anywhere within your home country (Brazil, Colombia or Argentina)
Paid Time Off - ample vacation days to rest and recharge
Public Holidays - local federal holidays are fully paid days off