You're seeing this page as if you were . The main menu is still yours, though. Exit from immersion
Gordon MurrayGM

Gordon Murray

Staff Cloud & AI Infrastructure Engineer

€700/day
Cork, IE
15+ years

Average response time: 1 hour

About Gordon

I’m a Staff-level AWS and Platform Engineer specialising in designing, building and operating reliable cloud infrastructure and distributed systems.

My core expertise is AWS, Terraform, cloud architecture, automation, reliability, observability and production operations. I have extensive experience taking ownership of complex production environments, simplifying infrastructure, improving resilience, reducing operational overhead and controlling cloud costs.

I work particularly well on projects that need an experienced engineer who can independently understand an existing system, identify practical improvements and deliver them safely into production.

My experience includes:

• AWS architecture and infrastructure
• Terraform and Infrastructure as Code
• Platform and DevOps engineering
• Highly available and distributed systems
• Containers and Kubernetes
• CI/CD and automation
• Observability, monitoring and production reliability
• Cloud cost optimisation
• APIs, data infrastructure and object storage
• AI infrastructure, vector search and search platforms

I also develop open-source software and have hands-on experience building modern search infrastructure combining vector and full-text search with object storage and tiered caching.

I’m particularly interested in AWS platform engineering, AI/LLM infrastructure, data and search systems, and technically challenging infrastructure projects.

Available for focused consulting and part-time engineering engagements where senior-level ownership and practical delivery are required.
  • English

    Native or bilingual

Remote only
Primarily works remotely

Experience

  • Canto
    Staff AWS Engineer
    April 2024 - Today (2 years and 4 months)
    Cork, Ireland
    Took full ownership of AWS infrastructure across multiple AWS accounts and regions previously managed by contractors, establishing standardised Terraform/IaC, reusable modules, deployment patterns, and operational visibility. Rebuilt a manual disaster-recovery process as a complete Terraform-defined recovery region reducing recovery time from weeks to hours. Resolved DynamoDB hot-partition issues by redesigning partition keys and access patterns, eliminating throttling in production workloads.
    Redesigned OpenSearch architecture from a shared multi-tenant index to tenant-isolated indices with improved shard configuration, reducing query contention and improving reliability for high-traffic workloads. Founded and now lead an internal AWS operations and observability function using CloudWatch, New Relic, and Terraform, improving alerting coverage, incident response, and system visibility across engineering teams. Coordinated infrastructure and platform work daily with US-based leadership and with developers in Germany and India, across US, European and Indian working days. Reduced monthly AWS spend by double-digit percentages through rightsizing, Savings Plans and reserved-instance planning, and ongoing cost-governance practices. Introduced tagging, guardrails, and automated policy checks to reduce configuration drift and support compliance across AWS accounts.
    Reverse-engineered an undocumented Jenkins/Bash/S3 deployment system and rebuilt it as a GitLab CI/CD workflow using ECR, Terraform and ECS, with Octopus Deploy for release management, turning fragile manual releases into repeatable, version-controlled deployments. Partnered with backend engineers to containerize internal APIs and background workers, improving deployment cadence, service isolation, and production reliability. Supported colleagues in scaling NVIDIA Triton Inference Server and OpenCLIP models for production use.
    Terraform Amazon Web Services AI infrastructure Cloud Architecture DevOps
  • Teamwork.com
    Platform Engineer
    February 2023 - April 2024 (1 year and 2 months)
    Cork, Ireland
    Designed and deployed a production Change Data Capture pipeline using Apache Flink, MariaDB, Debezium, Kafka, and KSQLDB to capture project-management activity across 600,000 tenants in US and EU regions, optimising application caches and enriching the data lake for analytical and operational needs. Handled high-volume data streams with strategies to ensure data integrity and availability, including diagnosing replication lag and consumer backpressure under production load. Instituted data governance practices preventing transmission of personally identifiable information, architecting and enforcing data privacy measures for GDPR compliance. Worked as part of a dedicated senior platform team focused on improving integration and collaboration across departments.
  • Teamwork.com
    SysOps Team Lead
    February 2016 - February 2023 (7 years)
    Cork, Ireland
    Founded the SysOps function as the first dedicated infrastructure hire and grew it to a seven-person team of two DBAs and five SysOps/platform engineers. Supported growth from three customer-facing products in one AWS region to five Go-based multi-tenant SaaS products across two AWS regions. Owned production reliability, patching, infrastructure operations, monitoring, alerting, and a 24/7 on-call rotation for customer-facing systems. Worked with remote colleagues and teams in the US, Amsterdam and Australia across a distributed engineering organisation. Led migration from EC2/ECS workloads to Kubernetes, including Kops and later EKS, introducing Helm based deployments, autoscaling, and GitOps workflows.
    Operated production database, cache, search, and traffic-management layers including MySQL, MariaDB, Redis/ElastiCache, OpenSearch/Elasticsearch, HAProxy, and relational replicas. Delivered repeatable database maintenance and zero-downtime schema-change workflows using ProxySQL, Percona tooling, and pt-archiver under sustained production pressure. Introduced HashiCorp Vault for centralised secret management and supported the operational controls needed for SOC 2 and ISO 27001 certification.
    Adopted Cloud Custodian and policy automation to detect exposed assets, enforce tagging, reduce configuration drift, and improve cost reporting, alongside regular pentests and automated scanning. Built a custom domain service supporting customer custom domains with HTTPS.
    Implemented centralised logging, alerting, runbooks, emergency procedures, access-control processes, and offboarding controls.
    Operated Datadog and Sentry for application monitoring and error tracking across the customer-facing products, and led internal work to reduce their cost as usage and ingest volumes grew.

Recommendations

Be the first to recommend Gordon

Help this freelancer shine by sharing your experience working together.

These freelancer profiles also match your criteria

AgathaA

Agatha Frydrych

Backend Java Software Engineer

4.7

(3)

2

BaptisteB

Baptiste Duhen

Fullstack developer

4.6

(4)

5

AmedA

Amed Hamou

Senior Lead Developer

4

(2)

7

AudreyA

Audrey Champion

Web developer

4.3

(3)

4

Education

  • AWS Certified Solutions Architect - Associate AWS Certified Cloud Practitioner
    AWS Certified Solutions Architect - Associate AWS Certified Cloud Practitioner
  • Certified: Terraform Associate
    HashiCorp
    Certified: Terraform Associate

Categories