Senior Platform Engineer – AWS / Terraform / Linux
2 days ago
Barcelona
ph3About the role /h3 pWe are looking for a bSenior /b bCloud Platform Engineer /b to join a technology team within a large international environment, working on the platform foundation that supports modern AI, data, and cloud-based systems. This role is highly focused on cloud infrastructure, platform engineering, AWS, Terraform, Linux troubleshooting, networking, automation, and production operations. The position is not primarily focused on building AI models or developing GenAI applications. Instead, you will work on the infrastructure and platform layer that enables technical teams to deploy, operate, monitor, and scale reliable systems in production. You will be expected to bring a strong bPlatform / DevOps / Cloud Engineering mindset /b, with the ability to analyse real production issues, debug infrastructure problems, design AWS-based solutions, and make sound technical decisions in complex environments. Experience with AI/ML platforms, MLOps, SageMaker, MLflow, or LLM tooling is valuable, but the core of the role is bAWS Cloud Platform Engineering /b. /p h3Responsibilities /h3 ul liDesign, build, and maintain cloud infrastructure solutions on AWS /li liWork with Terraform / Infrastructure as Code to provision, manage, and standardize infrastructure /li liAnalyse and troubleshoot production issues across cloud infrastructure, Linux systems, networking, storage, permissions, deployments, and platform services /li liInvestigate infrastructure drift, unexpected production changes, Terraform state inconsistencies, and configuration mismatches /li liDesign AWS-based solutions, selecting the right services and explaining technical trade-offs around scalability, reliability, security, cost, and maintainability /li liSupport and improve CI/CD pipelines for infrastructure, platform services, and cloud workloads /li liWork with Linux environments, including debugging issues related to disk usage, permissions, processes, logs, networking, and system performance /li liContribute to monitoring, observability, alerting, logging, and operational readiness of production systems /li liCollaborate with engineering, data, AI, and platform teams to ensure systems are reliable, automated, secure, and scalable /li liApply DevOps, SRE, and platform engineering best practices to improve reliability, automation, and operational excellence /li liSupport cloud environments that may include AI/ML workloads, MLOps tooling, training/inference environments, or AI platform components /li /ul h3Qualifications /h3 ul liSolid experience in Platform Engineering, Cloud Engineering, DevOps, Infrastructure Engineering, or SRE /li liStrong hands‑on experience with AWS in production environments /li liStrong experience with Terraform and Infrastructure as Code /li liGood understanding of cloud infrastructure design, including networking, compute, storage, IAM/security, monitoring, and scalability /li liStrong troubleshooting skills in Linux environments /li liAbility to debug real infrastructure issues using command‑line tools, logs, metrics, system resources, and cloud‑native services /li liExperience with CI/CD pipelines and automation /li liUnderstanding of networking fundamentals, including VPCs, subnets, routing, DNS, load balancers, security groups, firewalls, and connectivity troubleshooting /li liExperience with production operations, incident analysis, root cause investigation, and reliability improvement /li liAbility to design technical solutions in AWS and explain the reasoning behind the selected services and architecture /li liStrong ownership mindset and ability to work independently in complex technical environments /li liGood communication skills and ability to explain technical decisions clearly /li liExperience with MLOps / AI Platform environments, SageMaker, MLflow, feature stores, model deployment, model serving, or training/inference platforms /li liExperience with Docker and Kubernetes /li liFamiliarity with LLM tooling such as LangChain, Langfuse, LangSmith, or similar /li liExperience with observability tools, monitoring platforms, logging, tracing, and alerting systems /li liExperience with cost optimisation in AWS environments /li liExperience with data pipelines or workflow orchestration tools such as Airflow or Prefect /li liKnowledge of security, governance, compliance, and best practices for cloud platforms /li liExperience working in Agile / Scrum environments /li /ul h3Hybrid model /h3 pHybrid model: 2 days onsite per week /p h3Information Security Notice /h3 ul liThe employee will have access to confidential information related to Capitole and the assigned project. /li liCompliance with internal security and information protection policies is mandatory. /li /ul /p #J-18808-Ljbffr