Hiring.Camp

Site Reliability Engineer

Intermedia Intelligent Communications

·

Mar 12, 2026

Location
Tbilisi, Georgia
Workplace
Remote
Type
Full-time
Department
Engineering
Education
Bachelor
Source
Pinpoint

Description

Site Reliability Engineer

Department: Tech Operations

Employment Type: Full Time

Location: Tbilisi, Georgia



Description

**ALL CANDIDATES MUST BE LOCATED IN GEORGIA (Country)**

About Intermedia

Are you looking for a company where YOUR VOICE is heard? Where you can MAKE A DIFFERENCE? Do you THRIVE in a FAST-PACED work environment? Do you wake every morning EXCITED to work with GREAT PEOPLE and create SUCCESS TOGETHER? Then Intermedia is the place for you.

Intermedia has established itself as a leading provider of cloud communications and collaboration tech that allows companies to connect better. We have a strong track record of growth, profitability, and creating an environment where everyone matters. Everyone. While we are fast-paced and admittedly a bit intense, we promise that you won’t be bored. You will find Intermedia is a place where you can indulge your passion for creating and supporting great cloud technology. What’s more, we always look to promote from within and have many employees who have been with us 10, 15, and 20+ years!

Culture at Intermedia is built on teamwork and transparency. We hold each other accountable and always have each other’s back!

Are you ready to make your mark?

About the Role:
We are looking for an SRE to improve reliability and operational readiness with a strong focus on metrics, alerting, and event management. You will build and maintain monitoring using Prometheus/VictoriaMetrics, integrate alerts and events with BigPanda, and participate in on-call rotations to drive fast incident response and continuous improvement across Windows and Linux environments.



Key Responsibilities

  • Build and operate metrics/monitoring platforms: Prometheus and/or VictoriaMetrics (scrape configs, exporters, recording rules)
  • Design and maintain alerting strategy: thresholds, anomaly detection where applicable, alert routing, deduplication, and noise reduction
  • Integrate monitoring/alerting and events with BigPanda (correlation, enrichment, routing, incident workflows)
  • Create and maintain dashboards and operational visibility (Grafana or equivalent)
  • Develop and maintain runbooks, operational playbooks, and incident response procedures
  • Participate in on-call shifts: triage alerts, manage incidents, coordinate response, and lead communication during outages
  • Perform root-cause analysis, postmortems, and implement corrective/preventive actions
  • Improve service reliability via SLOs/SLIs, capacity planning, and automation to reduce toil
  • Support monitoring for core infrastructure and services on Windows and Linux, including HA components and clusters
  • Collaborate with DevOps/Engineering to instrument applications and standardize telemetry (metrics, logs, traces where applicable)



Skills, Knowledge and Expertise

  • Bachelor in Computer Science or related field 
  • Experience in SRE / Operations / DevOps with production incident ownership
  • Hands-on experience with Prometheus and/or VictoriaMetrics (exporters, alert rules, recording rules, troubleshooting)
  • Experience integrating alerting/event pipelines with BigPanda (or similar event correlation tools)
  • Strong troubleshooting skills across Linux and Windows systems (networking, OS, services)
  • Ability to build reliable alerting with minimal noise (correlation, grouping, suppression, maintenance windows)
  • Experience with Git-based workflows for monitoring-as-code and configuration management

Nice to have

  • Grafana administration and dashboard design standards
  • Log management (ELK/EFK, Loki) and/or tracing (OpenTelemetry)
  • Automation skills (Python, PowerShell, Bash) and configuration tools (Ansible)
  • Messaging/cache/proxy operations: RabbitMQ, Redis, Nginx
  • Experience with Windows clustering or HA environments
  • Experience defining SLOs/SLIs and operational KPIs
  • Experience in managing VOIP components and protocols (SIP , FreeSwitch, OpenSIP, session border controllers)
  • Experience with load balancing components ( F5 LTM, F5 GTM)
  • Experience with Virtualization platforms such as VMWare or HyperV
  • Experience with administering AWS or Azure tenants

On-call expectations
  • Participation in a rotating on-call schedule (including nights/weekends as needed)
  • Ownership of incident response: rapid triage, escalation, mitigation, and follow-up improvements
  • Commitment to improving monitoring quality to reduce alert fatigue and improve MTTR



Diversity, Inclusion, and Equal Opportunity

We hire, promote, and compensate employees based on their ability to perform their job responsibilities, without regard to race, color, creed, religion, sex, gender, marital status, national origin, ancestry, age, citizenship, physical or mental disability, sexual orientation, or any other basis protected by applicable law (collectively referred to in our Code of Conduct as “Protected Classes”). We do not tolerate employment discrimination in the workplace, and we are committed to making reasonable accommodations for identified disabilities or other limitations as required by all applicable laws. We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Skills

PythonAWSAzureAnsibleLinuxNginxVMwareRedisGitDevOpsSRE

Similar Jobs

30

Site Reliability Engineer

Experian · Cyberjaya, Selangor, Malaysia · Hybrid

Yesterday

Site Reliability Engineer

Cisco · USA-SAN FRANCISCO, United States of America +11 · Remote

Yesterday

Site Reliability Engineer

LSEG · ROU-Bucharest-Iuliu Maniu Boulevard, Romania

Yesterday

Site Reliability Engineer

Rbs · Bengaluru, India

Yesterday

Site Reliability Engineer

Huntington · Easton Ops Cols C Oh, United States of America +10 · Onsite

4 days ago

Site Reliability Engineer

Westpac Group · Sydney, NSW, Australia

5 days ago

Site Reliability Engineer

Professional Kyndryl · PELML Lima (PELML) La Molina, Peru · Remote

5 days ago

Site Reliability Engineer

"Arena Intelligence, Inc." · Bay Area · Hybrid

5 days ago

Site Reliability Engineer

Damia Group · Lisbon, Coimbra, Braga · Remote, Hybrid, Onsite

5 days ago

Site Reliability Engineer

Accenture · Taguig, Uptown Bonifacio Tower 3, Philippines

5 days ago

Site Reliability Engineer

NielsenIQ · Mexico City, MEX, Mexico

5 days ago

Site Reliability Engineer

Tinybird · Spain · Remote

6 days ago

Site Reliability Engineer

Cisco · USA-RESEARCH TRIANGLE PARK, United States of America +4 · Hybrid

6 days ago

Site Reliability Engineer

Miqdigital · Bengaluru, India

6 days ago

Site Reliability Engineer

Forwardnetworks · Santa Clara, CA +1

6 days ago

Site Reliability Engineer

Bosonai · Toronto · Remote

6 days ago

Site Reliability Engineer

Experian · Cyberjaya, Selangor, Malaysia · Hybrid

1 week ago

Site Reliability Engineer

CRC Careers · Charlotte NC - 600 S Tryon St., United States of America

1 week ago

Site Reliability Engineer

Citi Bank · 5900 HURONTARIO STREET MISSISSAUGA, Canada · Hybrid

1 week ago

Site Reliability Engineer

HostPapa · Remote

1 week ago

Site Reliability Engineer

citibank · Mississauga, ON,CA, CA · Hybrid

1 week ago

Site Reliability Engineer

CRC Careers · CRC - Charlotte, NC 600 S. Tryon St., United States of America

1 week ago

Site Reliability Engineer

LSEG · Taipei - Nan Shan Plaza, Taiwan

1 week ago

Site Reliability Engineer

Careers Home · Zagreb (Croatia) +4 · Hybrid

1 week ago

Site Reliability Engineer

Braiins · Braiins · Hybrid

1 week ago

Site Reliability Engineer

Acronis · Serbia +2 · Remote

1 week ago

Site Reliability Engineer

Earnin · Mountain View, US +1 · Hybrid, Onsite

1 week ago

Site Reliability Engineer

Quberesearchandtechnologies · Hong kong +1

1 week ago

Site Reliability Engineer

Runloop · San Francisco, CA · Onsite

1 week ago

Site Reliability Engineer

Triomics · India Office · Hybrid

1 week ago
Remote Site Reliability Engineer at Intermedia Intelligent Communications | Hiring.Camp