DevOps SRE
Solidus Labs provides an AI-powered compliance platform for modern financial and digital-asset markets. Its HALO platform supports trade surveillance, transaction monitoring, execution-quality supervision, token and stablecoin monitoring, and case management for regulated market participants.
Funding history
About Solidus Labs
Solidus Labs operates a multidimensional compliance and market-integrity platform built for crypto and broader financial markets. Its technology correlates trading activity with on- and off-chain signals, funding flows, and social sentiment to detect market abuse and financial-crime risks. The company serves broker-dealers, trading platforms, exchanges, institutional firms, custodians, market makers, token and stablecoin issuers, and regulators, alongside professional services for digital-asset compliance operations.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will own the reliability, availability, and performance of production environments. You will operate Kubernetes on EKS, manage AWS infrastructure, build infrastructure as code with Terraform and Helm, support GitLab CI/CD, improve observability, troubleshoot networking and security issues, lead incident response, perform root-cause analysis, and participate in on-call rotations.
Requirements
- 3+ years of hands-on DevOps or SRE experience
- Production experience with Docker and Kubernetes
- Knowledge of AWS, including EKS, EC2, Organizations, RDS, S3, CloudWatch, Lambda, and DynamoDB
- Experience with monitoring, logging, and alerting systems
- Proficiency with Terraform, Helm, and GitLab CI or similar tools
- Troubleshooting skills across infrastructure, CI/CD, and networking
- Bash and Python scripting experience
- Willingness to participate in on-call rotations
- Familiarity with pub/sub systems such as SQS or Kafka
- Experience with Redis, Airflow, Databricks, Spark, or EMR is a plus
- Experience with GitOps workflows and advanced Git usage is a plus
- Experience supporting Postgres, Snowflake, or ClickHouse is a plus
Responsibilities
- Own the reliability, availability, and performance of production environments
- Operate production Kubernetes on EKS, including cluster upgrades and Helm deployments
- Manage scaling and capacity with KEDA, Karpenter, and HPA
- Manage AWS Cloud environments, including EC2, Lambda, AWS Batch, ElastiCache, and RDS
- Evolve infrastructure as code with Terraform and Helm
- Support GitLab CI/CD pipelines and improve deployment stability
- Design observability systems with Prometheus, Grafana, and EFK
- Troubleshoot TLS, load balancing, VPC, NAT, and VPN issues
- Support compliance initiatives and respond to security incidents
- Use AI-powered tools for automation and productivity
- Lead incident response from troubleshooting through resolution
- Perform deep-dive root-cause analysis
- Participate in on-call rotations
