A1group
Site Reliability Engineer (AIOps) (f/m/d) @ A1 Competence Delivery Center
София
Apply through hirly
Upload your resume and get a version tailored to this job, plus a cover letter, in about thirty seconds — before you create an account.
Apply with hirlyhirly's read of this role
- Role family
- Engineering
- Seniority
- Mid level
- Country
- BG
- Work mode
- On-site / unstated
- First seen by hirly
- 16 Sept 2026
Derived automatically from the posting. Sign up to see how the role scores against your own resume.
the posting
Strength. Care. Growth
A1 Competence Delivery Center is a vital component of A1’s telecommunications business. Acting as an expertise hub, CDC is dedicated to delivering a full range of high-quality IT, network, financial and other services to support A1’s operations across all OpCos, independent of location.
Using the power of being OneGroup and leveraging synergies, CDC enables transparency of resources, key skills and knowledge expansion and personal career growth opportunities’ enhancement, paired with job stability.
You will know we are the right place for you, if you are driven by:
- Opportunities to learn and build your career.
- Meaningful work in a stable and fast-paced company.
- Diversity of people, projects, and platforms.
- A supportive, fun, and inspiring place to work.
Job Overview:
We're looking for an Site Reliability Engineer to build and optimize our monitoring ecosystem during a major infrastructure transformation. You'll ensure end-to-end visibility across Azure and Exoscale, leveraging modern observability and AIOps practices to improve reliability, reduce alert fatigue, and proactively identify issues before they impact our services.
Role Insights:
- Design and optimize multi-cloud observability pipelines for metrics, logs, and traces across Azure and Exoscale.
- Implement AIOps solutions for anomaly detection, event correlation, and faster incident resolution.
- Ensure seamless monitoring and visibility during cloud migration from Azure to Exoscale.
- Integrate observability into GitHub Actions CI/CD pipelines to detect deployment and performance issues.
- Develop automation, self-healing solutions, and runbooks together with DevOps engineers.
What Makes You Unique:
- Experience in DevOps/SRE with knowledge of telemetry, observability, or data analytics.
- Strong Kubernetes and observability stack expertise (Prometheus, Grafana, OpenTelemetry, ELK/PLG).
- Hands-on experience with AIOps platforms (Datadog, Dynatrace, New Relic) or custom monitoring solutions in Python/R.
- Understanding of cloud migrations and resilient, cloud-agnostic monitoring practices.
- Fluent English and the ability to translate complex technical insights into actionable improvements.
Our gratitude for the job done will be eternal, but we’ll also offer you:
- Innovative technologies and platforms to “play” with.
- Modern working environment for your comfort.
- Friendly, ambitious, and motivated teammates to support each other.
- Thousands of online and in-person learning opportunities for you to grow.
- Challenging assignments and career development opportunities in multinational environment.
- Attractive compensation package.
- Flexible working schedule and opportunity for home office.
- Numerous additional benefits, including, but not limited to free A1 services.
If you have any questions, please do not hesitate to contact Maria Ivanova.
Similar jobs
- Site Reliability Engineer - Solution PartnersAppfire · SpainFirst seen todayremote
- Site Reliability Engineer - Solution PartnersAppfire · BulgariaFirst seen todayremote
- Site Reliability EngineerAmpeco Global · SofiaFirst seen 26d ago
- Senior Site Reliability EngineerDraftkings · Remote - BulgariaFirst seen yesterdayremote
- Senior Site Reliability EngineerLiveperson · Sofia, BulgariaFirst seen 13d agoremote
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job