Senior Data Engineer – Customer AI Analytics
United Airlines · Chicago, Illinois ·
- Seniority
- Senior
- Category
- Data engineering
- Visa sponsorship
- No
- Company size
- 1000+
- Salary
- USD 117,610 – 153,146 / year
United Airlines · Chicago, Illinois ·
The Travelers Companies · CT - Hartford
ShyftLabs · Toronto, Ontario
UNITED NETWORK FOR ORGAN SHARING · Richmond, VA 23219
SITA · Madrid, Spain
Senior Data Engineer on United Airlines' Retail AI Automation team building and optimizing large‑scale PySpark pipelines in Databricks to process conversational and AI telemetry data, creating analytics data products and PowerBI dashboards that measure AI performance and business impact.
United's Digital Technology team is comprised of many talented individuals all working together with cutting-edge technology to build the best airline in the history of aviation. Our team designs, develops and maintains massively scaling technology solutions brought to life with innovative architectures, data analytics, and digital solutions.
Job overview and responsibilities
We are seeking a highly skilled Senior Data Engineer to join our Retail AI Automation team. This role is critical in transforming massive volumes of conversational data, customer interaction logs, and AI agent performance metrics into actionable business insights that drive continuous improvement of our AI automation capabilities.
Our Vision:
We are giving our people the backup they've always deserved, so that when a customer needs a human, they get the most outstanding customer experience they can imagine from AI. You will be building the data infrastructure and analytics capabilities that measure, optimize, and prove the value of AI automation at scale.You will be responsible for designing and building robust data pipelines that process millions of customer conversations, creating analytics data products that enable business stakeholders to understand AI performance, customer behavior, and operational impact across voice, chat, and future channels.
Data Product Architecture:
- Transform raw conversational data, AI agent logs, and customer interaction events into structured, reliable analytics data assets
- Design and build scalable data models that support AI performance monitoring, customer journey analytics, and business impact measurement
- Create reusable data products that enable self-service analytics for business stakeholders, data scientists, and AI engineers
Pipeline Development & Optimization:
- Build highly optimized, modular PySpark pipelines within Databricks to process large-scale conversational data and AI telemetry
- Convert ad-hoc analytical queries into production-ready data pipelines with rigorous testing and monitoring
- Implement incremental processing patterns and Delta Lake optimization techniques to minimize compute costs and improve query performance
Conversational Data Processing:
- Parse and structure complex, semi-structured conversational data including chat transcripts, voice call logs, AI agent decision traces, and customer intent classifications
- Standardize ingestion of diverse data sources including LLM prompt/response pairs, agent orchestration logs, and customer feedback signals
- Build precise conversation funnel metrics, AI containment rates, resolution accuracy, and customer satisfaction analytics
AI Performance Analytics:
- Design data models that enable comprehensive AI agent performance monitoring including response accuracy, latency, escalation patterns, and customer satisfaction
- Create analytics frameworks that measure business impact of AI automation including cost savings, wait time reduction, and operational efficiency gains
- Build attribution logic that connects AI interactions to downstream business outcomes such as bookings, customer retention, and contact center volume reduction
Data Governance & Quality:
- Enforce rigorous technical standards across all data products including schema enforcement, data quality validation, and comprehensive metadata documentation
- Implement standardized naming conventions, data lineage tracking, and documentation practices
- Ensure data products meet security, privacy, and compliance requirements for customer interaction data
Visualization & Reporting:
- Expose Databricks Delta tables to PowerBI and other visualization tools for business stakeholder consumption
- Partner with business analysts and product managers to design intuitive dashboards and reports that drive decision-making
- Create automated reporting frameworks that deliver regular insights on AI performance and business impact
Production Excellence:
- Transition experimental analytics work into robust, production-ready data assets with comprehensive monitoring and alerting
- Partner with analytics teams, AI engineers, and business stakeholders to ensure data products meet evolving business needs
- Document data products thoroughly to enable handoff to broader engineering teams for long-term maintenance
What's needed to succeed (Minimum Qualifications):
Preferred:
TEKsystems · Newark, New Jersey