Hybrid Madrid role at Fujitsu evaluating and monitoring LLM/GenAI models in production: benchmarking providers like OpenAI and Anthropic on performance, latency and cost, building testing/evaluation pipelines and dashboards, and working with IAOps/MLOps teams on observability and AI governance. Needs 2+ years' experience, AWS skills, and fluent Spanish and English.