Stellar Cyber logo

Stellar Cyber

Staff SRE Engineer

Work from home

Remote role where the employee must remain based in a particular country.

Hungary only

Employer listed it 5 months ago · Added today

Been open since 5 months ago. Long-running listings are sometimes left up after the role is filled.

Salary

Not stated

Location

Hungary only

Timezone

Not stated

Contract

Full-time

Experience

Lead

Category

Software

This employer didn't state pay. Jobs like this usually pay around $205k–$275k a year, a typical range taken from 599 lead-level software roles on Nomaders that do state pay. It's a guide, not an offer.

Remote flexibility

Work from home

This is a remote role, but the employee must be based in Hungary. It is work from home rather than work from anywhere.

What the employer says

  • Source listing states candidate location: "Hungary, Remote"

What Nomaders makes of it

  • Residency required in Hungary
  • Payroll and tax are likely handled in that country only

The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.

About the role

Join Stellar Cyber, a fast-growing global leader in cybersecurity trusted by some of the biggest names in the industry. Besides many enterprises and government agencies, nearly 30% of the world’s top MSSPs rely on our platform, and that number is growing every day as more companies recognize the value of next-generation security solutions. We're at the forefront of protecting organizations against sophisticated cyber threats using cutting-edge AI and automation technologies. Our culture is built on diversity, openness, and collaboration, fostering creativity and innovation that drives real impact in the market. We are seeking a highly skilled Staff Site Reliability Engineer (SRE) to join our team and drive reliability, scalability, and efficiency across our production systems. The ideal candidate will have deep expertise in cloud infrastructure, Kubernetes administration, observability, and incident management, with a proven track record of building and maintaining highly available and resilient platforms. As a senior member of the SRE team, you will not only operate complex distributed systems but also influence architecture, tooling, and best practices to ensure operational excellence. Please note, as part of our interview process, we may invite candidates for an in-person interview to meet with our team.

Responsibilities:

Administer and maintain container orchestration platforms and containerized workloads.

Monitor and troubleshoot production systems, participating in on-call rotations to ensure reliability.

Drive observability improvements by enhancing monitoring, logging, and alerting capabilities across systems and data platforms.

Administer and optimize cloud-based environments across multiple providers.

Manage and support distributed data platforms and real-time processing systems.

Develop and maintain continuous integration and delivery pipelines for efficient and reliable deployments.

Own and implement Infrastructure as Code (IaC) practices to ensure consistency and scalability.

Automate and orchestrate infrastructure using programming and scripting languages.

Perform system administration and networking tasks to support internal and external environments.

Collaborate effectively with engineers and stakeholders across different time zones.

Requirements

5+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles.

Proven success leading large-scale production systems in cloud environments (AWS, GCP, Azure, or OCI).

Demonstrated leadership in driving incident response, on-call best practices, and reliability-focused culture.

Strong experience with production on-call operations and incident management.

Advanced proficiency in Kubernetes administration and troubleshooting.

Hands-on experience with observability tools: Prometheus, Grafana, Loki, and Alertmanager.

Knowledge in chat-based operations interfaces and/or auto-remediation controllers using AI agentic framework.

Understanding of AI agents for Auto-triaging alerts, correlate signals and suggest/root-cause hypotheses

Expertise in operating data platforms (Elasticsearch, MongoDB, Spark, Kafka, Redis).

Proficiency with public cloud services (AWS, Azure, GCP, or OCI).

Strong programming and automation skills in Python and Bash.

Requirements

  • ·5+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles.
  • ·Proven success leading large-scale production systems in cloud environments (AWS, GCP, Azure, or OCI).
  • ·Demonstrated leadership in driving incident response, on-call best practices, and reliability-focused culture.
  • ·Strong experience with production on-call operations and incident management.
  • ·Advanced proficiency in Kubernetes administration and troubleshooting.

Benefits

No benefits package published with this listing. Ask about it at first interview.

How to apply

  1. 1Check the flexibility label above, work from home, matches where you plan to live and work.
  2. 2Tailor your CV to the role at Stellar Cyber, mentioning your remote working experience.
  3. 3Apply directly on the employer's careers page using the button below. Nomaders never handles your application.

Found 15h ago. Last checked today. Always confirm the details on the original posting, salary and location can change after publication.

Listing sourced from Company boards.

Similar roles

Other open software roles with comparable remote rules.

Browse all open roles

Free to apply, no account needed.

Typically $205k to $275k per year · You'll be taken to the employer's careers page.