Alexander Kolesen

IT Infrastructure & Service Architect

Summary

IT Infrastructure and Service Architect with 15+ years of experience designing, implementing, and operating cloud infrastructure for high-load and business-critical systems β€” from a NYC ride-sharing service (millions of users) to an online game platform (600+ developers) and a national tax system for the Danish government. Deep hands-on expertise in AWS, Infrastructure as Code with Terraform, CI/CD, PostgreSQL and MySQL. Experienced in translating business requirements into technical solutions, driving implementations, coordinating external vendors and distributed engineering teams, and ensuring stable service operations. Working in English-speaking international environments since 2012.

Core Skills

Cloud & Infrastructure
AWSEC2ECS LambdaRDSS3 SQSAPI GatewayKinesis DynamoDBCloudWatchCloudFront ElastiCacheRoute 53GameLift BatchLinuxDocker Kubernetesnginx
IaC & DevOps
TerraformAtlantisGitHub Actions GoCDCI/CDCapacity planning Fault tolerancePrometheusGrafana Kibana
Databases
PostgreSQLMySQL/InnoDBMongoDB DynamoDBSnowflakeElasticsearch
Architecture & Service Management
Requirements gathering and solution design for business needs, service ownership and production readiness, incident and change management, vendor evaluation and coordination, stakeholder communication, team leadership.
Programming
PythonClojureGolang Java/KotlinCC++
Languages
English β€” professional working proficiency (B2+)

Work Experience

Cloud Infrastructure / Service Architect

2025 β€” present

Self-employed Β· remote

Developing a cloud orchestration product for ETL and data-processing workloads, with end-to-end responsibility for architecture and implementation.

  • Design and implement AWS infrastructure for ETL and data-processing workloads using Step Functions, ECS, Batch, and S3.
  • Manage the full infrastructure lifecycle with Terraform, ensuring reliable and observable operation of production services.

Stack: Terraform, Python, PostgreSQL, AWS (Step Functions, ECS, Batch, Lambda, S3, RDS, CloudWatch)

Principal Site Reliability Engineer

2023 β€” 2025

Online Services, People Can Fly Β· ~600 people Β· Warsaw, PL (remote)

People Can Fly is a global video game development company.

  • Built and launched the foundation for a self-hosted online game platform serving 600+ developers and supporting multiple new titles.
  • Built the content pipeline for the staging environment for content developers.
  • Partnered with internal stakeholders and external game-hosting providers (Unity Multiplay, AWS GameLift) to gather requirements, evaluate options, define the target architecture, and coordinate implementation.
  • Owned the platform's production readiness: deployment processes, observability, reliability requirements, and coordination of operational responsibilities.
  • Built the CI/CD pipeline for the game platform.

Stack: AWS, Terraform, Java, Kotlin, GitHub Actions, GameLift, Multiplay

Software Developer

2022 β€” 2023

ActVPN and wscp, Actmobile Β· ~10 people Β· San Francisco, US (part-time, remote)

Actmobile builds VPN products. Worked on ActVPN, a VPN written in C, and wscp, a maximum-throughput network copying tool.

  • Integrated LwIP, a userspace TCP/IP stack, into ActVPN.
  • Tracked down memory leaks and stabilized ActVPN.
  • Implemented a dual-channel model (UDP for data, TCP for retransmitting lost fragments) that outperformed standard TCP by 2–10Γ— in bandwidth utilization under certain network conditions.

Stack: Linux, C, LwIP, OpenSSH, rsync, TCP, UDP

Senior Backend Developer & Principal Infrastructure Engineer

2021 β€” 2022

Palta Data Platform Β· ~10 people Β· Limassol, CY (part-time, remote)

Palta is a startup incubator for health and well-being applications; its data platform replaced third-party data collection services with an internal one.

  • Designed and implemented a cost-effective, multi-tenant data pipeline (API Gateway β†’ Kinesis Firehose β†’ S3 β†’ SQS β†’ Lambda β†’ Snowflake) at one-tenth the cost of Amplitude on the same data volume.
  • Designed and implemented a reliable HTTP callbacks system.
  • Designed infrastructure for a payments system to support rapid load scaling.

Stack: Docker, PostgreSQL, Apache Kafka, Snowflake, Terraform, Atlantis, Grafana, GitHub Actions, AWS (API Gateway, Lambda, S3, SQS, Kinesis, DynamoDB, RDS, CloudWatch)

Software Developer

2019 β€” 2021

ICE, Denmark's Ministry of Taxation (Flexiana) Β· 500–1000 people Β· Copenhagen, DK (remote)

A public system covering the full property taxation process in Denmark, built in Clojure/ClojureScript on AWS.

  • Maintained and extended several components to eliminate performance bottlenecks.
  • Recognized and bridged a gap between developers and infrastructure engineers, significantly increasing the velocity of infrastructure-related delivery.
  • Sped up the CD pipeline 4–10Γ— by introducing more granular deployments.

Stack: Clojure, Docker, PostgreSQL, Jenkins, Terraform, Atlantis, GitHub Actions, AWS (Batch, API Gateway, EC2, ECS, Lambda, S3, SQS, CloudWatch, RDS)

Solo Backend Engineer

2019

Wanna Kicks Β· ~20 people Β· Minsk, BY (part-time, remote)

Virtual try-on application that lets users see how sneakers would look on their feet before buying.

  • Implemented a CRUD backend to manage the model catalog.
  • Distributed the model catalog as encrypted bundles via CDN, in a multi-versioned and multi-tenant fashion.
  • Implemented statistics collection and SQL-based analysis using AWS Kinesis Firehose and AWS Athena.

Stack: Python, Chalice, Docker, Terraform, AWS (API Gateway, Lambda, S3, SQS, DynamoDB, CloudFront)

Principal Infrastructure Engineer & Senior Software Developer

2015 β€” 2018

Juno Β· 100–200 people Β· Minsk, BY β€” New York, US (onsite in Minsk)

Juno was a ride-sharing service operating in New York, acquired by Gett in 2017.

  • Scaled from a fresh team to 40K daily trips in under one year in the highly competitive NYC market, serving millions of users.
  • Established continuous delivery of 100+ backend microservices, running several times a day and sometimes exceeding 600 deploys daily.
  • Worked on capacity planning and fault tolerance, isolating parts of the computing cluster into separate shards.
  • Led a team of 5 infrastructure engineers serving 40 core backend engineers and 15 QA engineers.
  • Provided developers with tooling to inspect production logs and metrics and to maintain their own alert sets.

Stack: Linux, Clojure, Golang, Python, MySQL/InnoDB, PostgreSQL, MongoDB, Docker, GoCD, nginx, Elasticsearch, Kibana, Prometheus, Grafana, AWS (EC2, ECS, RDS, S3, ElastiCache, Route 53)

Infrastructure Engineer & Software Developer β€” Various Projects

2007 β€” 2015

Kontur Β· Dyn (via EPAM) Β· Iron.io Β· Wargaming.net

  • Kontur, Inc. (GIS service provider) β€” built backend infrastructure for Kontur Platform Services (Kubernetes, Java, nginx).
  • Dyn Inc. (DNS, via EPAM, acquired by Oracle) β€” designed an API bridging legacy and new DNS systems; built Docker + Chef + Jenkins pipelines.
  • Iron.io (cloud platform) β€” migrated MongoDB clusters, automated monitoring, extended deployment infrastructure.
  • Wargaming.net (gaming) β€” built deployment tooling and the infrastructure team; scaled web systems to 1M+ concurrent players.

Education

M.Sc. in Artificial Intelligence / Information System Security

2004 β€” 2009

Belarusian State University of Informatics and Radioelectronics β€” Faculty of Information Systems and Management

Master's thesis: β€œIntellectual collection and analysis of Internet service statistics in real time under high load.”