Lucía Romero, ai infrastructure engineer

Lucía Romero

Senior AI Infrastructure Engineer

Barcelona, Spain · Europe/Madrid (UTC+2) · 9 years of experience

Lucía focuses on reliable, observable, and cost-efficient compute platforms for AI workloads.

Request a shortlist

Why this profile fits

Reliable AI compute

Runs Kubernetes and GPU platforms with workload isolation, SLOs, capacity controls, and recovery procedures.

Infrastructure as code

Makes environments repeatable through reviewed modules, GitOps, policy checks, and controlled change.

Cost and capacity discipline

Connects utilization, latency, availability, and unit economics to practical scaling decisions.

Relevant toolkit

KubernetesTerraformNVIDIA GPUsPrometheusArgo CD

Delivery evidence

Selected work

AI Service Observability Stack

travel technology

Designed and delivered a AI compute platform focused on observability for AI services; reduced mean time to isolate inference incidents from 74 to 19 minutes. The release included capacity controls, workload isolation, SLOs, disaster recovery, and unit-economics reporting.

KubernetesTerraform

Inference Incident Workbench

travel technology operations

Built the supporting control and measurement layer for the primary system, covering review queues, regression checks, operational visibility, and a documented handoff to the owning team.

KubernetesLinux
View more profile detail →

Relevant experience

Career timeline

Senior AI Infrastructure Engineer

2023–Present

Devlyn client assignments · Remote

  • Led observability for AI services delivery for a travel technology team and reduced mean time to isolate inference incidents from 74 to 19 minutes.

AI Infrastructure Engineer

2017–2022

travel technology product company (confidential) · Barcelona, Spain

  • Built production systems in travel technology, with increasing ownership of reliability, testing, and stakeholder delivery.

Skill depth

Show experience in context.

VerbalCommunicationDomainUnderstandingProblemSolvingCodeQualitySystemDesignDeliverySpeed
Competency shapeAssessment across the same six dimensions used for every profile.

Relevant experience by skill

Kubernetes9 years
Terraform8 years
NVIDIA GPUs7 years
Linux6 years
Prometheus9 years
Argo CD8 years
See the complete skill index

Languages

PythonTypeScriptSQLBash

Frameworks

TerraformNVIDIA GPUsLinuxHelmAnsiblePython

Tools

OpenTelemetryKarpenterVaultGitHub ActionsDocker

Platforms

AWS EKSGoogle Kubernetes EngineAzure Kubernetes ServiceNVIDIA DGX

Storage

S3CephNVMePostgreSQL

Paradigms

GPU schedulingInfrastructure as codeSRECapacity engineering

Credentials

Certifications

Certified Kubernetes Security Specialist

Cloud Native Computing Foundation · 2023 · Certified

HashiCorp Certified: Terraform Associate (004)

HashiCorp · 2026 · Certified

Communication

Languages

English

Fluent

More examples

Similar AI Infrastructure Engineer profiles.