Principal Site Reliability Engineer
Administration, Software Engineering
London, UK
Description
Who We Are
Veson Nautical empowers the global maritime industry to navigate complexity on all sides of the trade. Veson's platform combines AI-driven workflows, trusted data, and seamless collaboration, to deliver the insight and context needed for confident, competitive decision-making.
The Opportunity
As a Principal Site Reliability Engineer at Veson Nautical, you will design, build, monitor, and support the cloud infrastructure that underpins our rapidly growing SaaS platform. This is a multi-cloud role spanning both AWS and Google Cloud Platform, with an immediate focus on growing our GCP footprint and the systems that connect our environments across regions, accounts, and clouds.
The work is a mix of greenfield and stewardship. You'll stand up infrastructure for new applications from scratch, and you'll take on the harder problem of making our existing estate more consistent, more observable, and easier to operate. You will have significant influence over the architectural direction of the platform.
Our suite of products includes:
· Veson Platform - the core commercial maritime platform used by the world's leading shipping organizations to manage vessel communication, operations, and trade decisions
· Oceanbolt - a dynamic data intelligence platform, tracking over 23,000 vessels in real time to deliver accurate, timely market intelligence to drive decision making
· Shipfix - using proprietary AI-driven tools to infer cargo and vessel information, extracting, anonymizing, and aggregating billions of data points with near real-time processing of email exchanges in the shipping market
The Team
You'll join a global Site Reliability Engineering team with members in the United States and the United Kingdom. This is a senior individual contributor role without direct reports, but with real leadership expectations. If you are the kind of engineer who measures success partly by what your teammates can do without you, this role will suit you.
This position is based in London in a hybrid model, with an expectation of 2-3 days per week in our office. We think those days are worth showing up for: our beautiful office is stocked with snacks, there's a weekly team lunch on us, and the atmosphere is friendly, informal, and genuinely collaborative.
The team participates in an on-call rotation -more details will be provided in the interview process.
Our Stack
· Google Cloud Platform - primarily PaaS services (Bigtable, Cloud SQL, Dataflow, Datastore, GKE, GCS, KMS, Pub/Sub)
· Amazon Web Services - multi-region, multi-account, with a broad range of managed services
· Containers and orchestration - Kubernetes (GKE and EKS)
· Infrastructure-as-Code - Terraform, Terragrunt, and Atlantis
· CI/CD - GitLab Pipelines, ArgoCD, Octopus Deploy
· Data - ElasticSearch hosted with Kubernetes Operator, PostgreSQL, SQL Server, BigQuery
· Monitoring and Security - Splunk, Grafana / Grafana Tempo, OpenTelemetry, Cloud Armor Enterprise, OpsGenie, Renovate, Sentry
· AI Tools - Claude, Amazon Bedrock, Gemini, Vertex AI
Key Responsibilities
· Design, implement, and operate scalable, reliable, and secure infrastructure across Google Cloud Platform and AWS
· Lead greenfield infrastructure builds for new applications, and modernize existing infrastructure toward common patterns
· Solve cross-region, cross-account, and cross-cloud problems — networking, identity, data movement, and the operational patterns that hold across environments
· Drive automation of infrastructure provisioning and configuration management using Terraform and related IaC tooling
· Establish and maintain comprehensive monitoring, alerting, and observability practices
· Cross-train and mentor other engineers, with the explicit goal of broadening GCP and multi-cloud capability across the team
· Partner closely with development teams to ensure the reliability, performance, and scalability of our platforms
· Set technical direction through design reviews, architecture proposals, and clear written documentation
· Participate in and improve incident response, and drive the follow-through that keeps the same incident from recurring
· Improve the cost effectiveness of our cloud footprint through visibility, analysis, and sound architectural choices
Skills/Experience Needed to Be Successful in This Role
Required:
· Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience
· Previous experience working on a large-scale Software-as-a-Service (SaaS) platform supporting thousands of global users in a 24x7x365 environment
· 5+ years of hands-on experience with Google Cloud Platform services and architecture, including running production workloads at scale
· Production Kubernetes experience, particularly Google Kubernetes Engine (GKE)
· Proficiency with Infrastructure-as-Code, preferably Terraform with the GCP provider
· Experience with Google Cloud networking, including VPC, Cloud Load Balancing, and Cloud CDN
· Strong programming skills in Python, Go, or TypeScript for automation and tool development
· Demonstrated ability to raise the technical capability of a team through mentorship, documentation, and knowledge sharing
· A track record of building consensus around technical decisions that span multiple teams
Highly Desirable:
· Hands-on experience with both GCP and AWS, and a clear point of view on where multi-cloud helps and where it hurts
· Working knowledge of Google Cloud security best practices and IAM implementation
· Experience with Google Cloud data services including BigQuery and Dataflow
· Experience with cloud cost management (budgeting, anomaly detection, cost analysis and reporting)
· Experience working on a geographically distributed team across time zones
Nice to have:
· Google Cloud Professional certifications (Cloud Architect, Data Engineer, or DevOps Engineer)
· Experience with GitLab CI or Octopus Deploy
This role is scoped at the Principal level, and we may also consider well-qualified candidates at the Senior level.
We are focused on building a diverse and inclusive workforce. If you're excited about this role, but do not meet 100% of the qualifications listed above, we encourage you to apply. While we try to be thorough with our job descriptions, not everything about you as a candidate can be condensed into a list of bullet points.
About Veson Nautical
Veson Nautical is a successful, rapidly growing global software company. Our clients are the world's leading commercial maritime owners, operators, and commodity trading companies. Veson's solutions enable our clients to identify new opportunities and proactively manage their business to make more profitable decisions. With offices in Singapore, Tokyo, London, Houston and headquarters in Boston, USA, Veson Nautical is a dynamic organization with a committed team of professionals. Dedicated to ensuring the highest levels of client satisfaction, Veson Nautical brings decades of experience, technical knowledge, enthusiasm, and commitment to clients around the world.