← Back to list
Job · Senior

Platform Engineer - Agentic AI & Harness Engineering

Platform Engineer • Senior • Hybrid • Full-time • Spain Madrid, Spain

You will build the internal agent harness: the runtime, orchestration and guardrails that turn foundation models into AI agents doing reliable, governed work at scale. You treat the harness as a product whose users are the teams building and running agents.

Responsibilities

  • ▹Design and build the agent harness: runtime, orchestration, tool and MCP integration, context and memory, evaluation, guardrails and observability
  • ▹Build golden paths and reusable patterns so teams ship reliable, governed and measurable agents instead of one-off scripts
  • ▹Define the oversight model: policy-as-code, human-in-the-loop checkpoints, loop and failure detection, evaluation harnesses and cost control
  • ▹Integrate agents with the systems they act on (CI/CD, infrastructure, cloud services, observability and ticketing) through APIs and MCP
  • ▹Make agentic productivity measurable: define what productive means for each use case, instrument it and run the harness on that data
  • ▹Keep the harness model-agnostic and versioned so it survives model upgrades and provider changes without rewrites
  • ▹Mentor teams, run enablement and reduce friction between Dev, Ops, Security and the people adopting agents

Requirements

  • ▹5+ years in Platform Engineering, DevOps or SRE, building developer platforms or runtime and orchestration systems in production
  • ▹Hands-on experience building agent harnesses or agentic systems: runtime, tool and MCP integration, orchestration, evaluation and guardrails
  • ▹Solid software engineering, primarily in Python, with the depth to build production systems and abstractions, not only prompts
  • ▹Strong cloud-native background: Kubernetes, containers, Infrastructure as Code and one major cloud (AWS, Azure or GCP)
  • ▹A working grasp of how LLM agents fail and the engineering that makes them reliable: context design, loop and failure detection, verification and evaluation
  • ▹Experience treating a platform as a product: contracts, SLOs, adoption metrics and real users
  • ▹Professional working English and Spanish

Nice to have

  • ▹Agent frameworks such as LangGraph, CrewAI, AutoGen and the Model Context Protocol
  • ▹Serious production use of AI-assisted development tools (Claude Code, Copilot, Codex)
  • ▹LLMOps, agent evaluation or observability for agentic systems
  • ▹Prior work on harnesses that stay stable across model upgrades
  • ▹Platform engineering in regulated environments (finance, insurance, public sector)

Soft skills

Mentoring and enablementCross-team collaboration (Dev, Ops, Security)Growth mindsetCustomer focus

What we offer

  • ▹Hybrid-friendly culture
  • ▹Be Well programs supporting financial, mental, physical and social health
  • ▹Personalized development goals and continuous feedback
  • ▹Learning opportunities, certifications with Microsoft, Google and Amazon, and coaching

About the company

Kyndryl runs and reimagines the mission-critical technology systems of leading businesses.

Languages: Angol: munkaszintű, Spanyol: munkaszintű

Similar jobs