Senior Storage Engineer

Verda (formerly DataCrunch) is a Helsinki-headquartered full-stack AI cloud provider offering on-demand GPU compute, self-service clusters, serverless containers, storage, and inference infrastructure.

Helsinki, Finland
About Verda

Founded in Helsinki in 2020, Verda operates European AI cloud infrastructure across physical data centers, hardware, a cloud platform, and AI research. DataCrunch renamed to Verda in November 2025 without changing its service offering.

View jobs by Verda

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will lead the design, deployment, and operation of production Ceph clusters supporting petabyte-scale customer data and GPU workloads. You will define storage architecture and operational standards, launch managed object storage capabilities, improve observability and automation, diagnose distributed-storage performance issues, and lead production incident response.

Requirements

  • Ceph deployment, operations, troubleshooting, and performance optimization
  • Experience operating multi-petabyte production Ceph environments
  • Technical leadership of storage infrastructure
  • Engineering mentorship
  • Linux systems and DevOps skills
  • Bare-metal and systems-internals knowledge
  • Networking and distributed-systems debugging
  • Mission-critical production infrastructure operations
  • Infrastructure automation development
  • Cross-team communication

Responsibilities

  • Direct the design, deployment, and operation of large-scale production Ceph clusters
  • Establish technical direction for storage architecture and operational standards
  • Launch a managed Object Storage product with access keys, bucket management, and SLOs
  • Scale storage systems to petabyte and hundreds-of-petabyte capacity
  • Manage CephFS, RBD, and RADOS Gateway interfaces
  • Mentor engineers through code reviews, design collaboration, and shared on-call responsibilities
  • Enhance observability, automation, and operational tooling
  • Diagnose complex performance issues in distributed storage environments
  • Collaborate on GPU cluster integration
  • Oversee capacity planning, infrastructure upgrades, and lifecycle management
  • Lead production operations and incident response

Benefits

  • Equity compensation
  • Local benefits