← Back to jobs

Senior Platform Engineer

  • Remote
  • Sweden
  • English
  • Posted 17.09.26 15:25

We are looking for an experienced and pragmatic Senior Platform Engineer to join our Shared Services and Common Tools team. In this role, you will take direct technical ownership of our cloud native observability ecosystem and core developer platform capabilities that support our internal engineering teams. You will design, scale, and maintain and evolve our multi-tenant telemetry infrastructure while driving automated delivery pipelines and GitOps practices. Working alongside the Tech Lead, you will champion automated CI/CD workflows, declarative deployments, infrastructure as code, and AI assisted development practices within an Agile/Scrum environment. Key Responsibilities: Observability Platform Ownership: Design, operate, and scale our multi-tenant distributed observability stack (LGTM: Grafana, Loki, Tempo, Mimir) hosted on Kubernetes. Telemetry Pipelines & Ingestion: Configure, tune, and maintain telemetry ingestion using Grafana Alloy (metrics scraping, log stream processing, and distributed tracing export via OTLP). Declarative GitOps Delivery: Manage and reconcile application states and platform deployments declaratively using ArgoCD. CI/CD Automation: Design, build, and optimize automated pipelines using GitLab CI/CD, defining reusable CI templates, testing stages, and deployment triggers.Infrastructure as Code (IaC): Automate and manage underlying cloud infrastructure, object storage, and Kubernetes resources using Terraform and Helm. Core Identity & DevEx: Support shared platform services, including centralized IAM integration with Keycloak and developer onboarding via Backstage (Internal Developer Portal). Reliability & Capacity Planning: Monitor platform health, manage high-cardinality telemetry data, configure retention/compactor rules, and establish alerting and SLOs.Modern Engineering Workflows: Leverage AI-assisted coding tools to accelerate platform development, documentation, and troubleshooting Mandatory Requirements: 5+ years of experience in Platform Engineering, SRE, or DevOps roles.Solid Kubernetes Expertise (CKA desiderable): Extensive production experience running, troubleshooting, scaling, and maintaining workloads, StatefulSets, and DaemonSets.Hands-on LGTM Observability Stack Experience: Demonstrated proficiency operating and querying Loki (LogQL) and Mimir / Prometheus (PromQL).Working experience with distributed tracing via Tempo (OpenTelemetry standards).Direct experience with modern telemetry collectors, specifically Grafana Alloy (or OpenTelemetry Collector / Promtail).Clear understanding of high cardinality, multi-tenancy, and object storage lifecycles for metrics and logs.GitOps Mastery with ArgoCD: Proven experience deploying, configuring, and managing multi cluster applications declaratively using ArgoCD (ApplicationSets, sync policies, and rollback strategies).GitLab CI/CD Experience: Proven ability to build robust .gitlab-ci.yml pipelines, configure custom runners, and structure CI stages for containerized platform tooling.Infrastructure as Code & Packaging: Proficiency writing modular Terraform and packaging applications with Helm.AI-Assisted Development: Active, hands-on experience leveraging AI-assisted development tools (e.g., Claude Code, Cursor, or GitHub Copilot) to streamline coding, debugging, and automation tasks.Scripting: Strong automation skills in Bash, Python, or Go.Language: Professional working proficiency in English (written and verbal). Nice-to-Have Supply Chain Security & Governance: Experience collaborating on software supply chain security initiatives, including vulnerability tracking and SBOM governance using tools such as Dependency-Track. Experience with Backstage service catalogs and software templates. Familiarity with Keycloak administration, OIDC/OAuth2 concepts, and realm configurations What We Offer Direct ownership over foundational observability and delivery systems powering enterprise-scale products. A modern, cloud-native tech stack with high autonomy over technical execution. Flexible working options (Remote within Europe). A collaborative, engineering-first environment focused on continuous delivery, modern workflows, and blameless operations