
Recruiter
Ekaterina Bulanova
Roles:
DevOps
Must-have skills:
KubernetesGCPTerraformGrafana
Nice-to-have skills:
Go
Considering candidates from:
Baltics, Eastern Europe, Western Balkans, Western Europe, Armenia, Cyprus, Georgia, Greece, Moldova and Portugal
Baltics, Eastern Europe, Western Balkans, Western Europe, Armenia, Cyprus, Georgia, Greece, Moldova and Portugal
Work arrangement: Remote
Industry: Travel Arrangements
Language: English
Level: Senior
Required experience: 6+ years
Size: 51 - 200 employees
Company
Hellotickets is building the largest global marketplace for travel experiences, from a Broadway musical show to a helicopter tour in Rio de Janeiro. Although they are a young startup, their platform is already presented in 15 countries — all with their local currencies and payment methods. The company is financially backed by key players in the entertainment and travel industry, like LCD SoundSystem’s James Murphy, Sony Music’s Managing Director and the Founder of TripAdvisor.
Description
At the moment, the company is looking for an SRE/DevOps Engineer to work with Go microservices on GKE, Cloud SQL PostgreSQL, Cloudflare, Terraform and Datadog. Traffic is spiky (on-sales, peak season), so reliability and autoscaling directly affect revenue. We need an engineer who owns production, not just pipelines.
Tasks:
Tasks:
- Own reliability: SLIs/SLOs, error budgets, on-call, incident response and blameless post-mortems.
- Operate and evolve our GKE platform: upgrades, autoscaling, right-sizing, Helm, secrets.
- Keep everything in Terraform and GitOps (ArgoCD or similar) - fast, safe, observable deploys.
- Observability in Datadog: alerts tied to SLOs, dashboards and tracing that actually help find root causes.
- Cloud SQL PostgreSQL operations: backups and restore drills, pooling, performance diagnostics.
- Edge & security: Cloudflare WAF/rate limiting/Bot Management, Cloud Armor, IAM, scanning in CI.
- Watch and cut GCP/Datadog spend.
- Build the paved road for ~30 engineers in 5 squads: templates, self-service, runbooks. Use AI coding agents in infra work.
Must-have:
- 4+ years in DevOps/SRE/platform roles with real production ownership.
- Strong Kubernetes in production (GKE preferred), Terraform, one major cloud (GCP preferred).
- Solid Linux and networking (TCP/HTTP/TLS/DNS, load balancing, CDN).
- Observability experience (Datadog, Prometheus/Grafana); you think in SLOs, not just alerts.
- CI/CD & GitOps in practice; working knowledge of PostgreSQL operations.
- English and Russian language
Nice to have:
- Go, Cloudflare Workers, Teamcity, Openresty, Kong/Traefik, FinOps, load/chaos testing.
Benefits:
- Full-time, remote job
- 4 days working week (Monday - Thursday, Friday day off)
- Paid vacation (20 days) and sick leaves
- Flexible working schedule
- Friendly professional staff and warm atmosphere
Interview process:
- Intro call with Toughbyte
- Call with HR
- Technical interview with CTO and System Architect
Questions
Have questions about this position? Try the company page or sign up to ask one.
