HirePortal

Experienced Site Reliability Engineer

Proxiad SEE is an IT consulting and software development company, part of Proxiad Group France. Operating in Bulgaria and North Macedonia, we employ over 250 professionals skilled in various sectors: healthcare, finance, e-commerce and retail, transport & logistics, and procurement. We are also the company behind two software solutions for office work – our workplace management software, Desk Buddy, and our performance management software, GROW.

We're seeking an Experienced Site Reliability Engineer to join our existing team in Sofia, working for a French email marketing platform.

You'll play an instrumental role in the day-to-day management of our client's global infrastructure. This includes monitoring and tracking key performance indicators (KPIs), collaborating with engineers to ensure our products and services are properly resourced, automating processes, and planning for future growth and scalability.

Responsibilities:

  • Partner with product engineering teams to identify systems requirements
  • Build and support our cloud-based infrastructure
  • Automate routine processes and remediation tasks through Infrastructure as Code
  • Develop, monitor and track Service Level Objectives (SLOs) for the systems under management
  • Proactively troubleshoot, resolve, and plan for issues that typically come from support staff, other engineering teams, and our automated monitoring system
  • Ensure our datastores are healthy and operate at optimal performance levels
  • Contribute to the growth and culture of our engineering team

Requirements:

  • Experienced background in infrastructure, operations or software engineering
  • Strong hands-on experience with Ansible and Terraform is mandatory — the role relies heavily on Infrastructure as Code
  • Experience with any major public cloud provider (AWS, Azure, GCP, or similar); experience with GCP is considered an advantage
  • Hands-on proficiency with modern monitoring tools like Prometheus and Grafana
  • Experience with data stores and search technologies, including PostgreSQL and Elasticsearch
  • Experience with Apache Cassandra is a big plus — the role involves planning, executing, and supporting Cassandra cluster upgrades in production environments, so background here will be highly valued
  • Experience with Python and Bash is beneficial
  • Experience operating and maintaining production systems in a Linux and public cloud environment
  • Strong technical skills across various infrastructure technologies
  • Strong communication skills

Benefits:

  • Hybrid work model with a minimum of two days per month onsite. Remote options available
  • 25 days of paid annual leave
  • Additional health insurance package (including dental and optical services)
  • Language classes and L&D opportunities
  • Internal recognition system with different levels of financial bonuses
  • Team buildings, game nights, hikes and many other social events
  • Presents for your work anniversary, birthday, etc., employee discounts

All applications will be treated with strict confidentiality.

Only short-listed applicants will be contacted.

Skills

  • Site Reliability Engineering
  • Infrastructure as Code
  • Cloud Infrastructure
  • Monitoring and Alerting
  • Service Level Objectives
  • Automation
  • Troubleshooting

Related jobs

Proxiad SEEApply for this job