Lead Data Engineer - Data Scientist
Manager
Analytics
Artificial Intelligence
Business Intelligence
Cloud
Cloud Native
Cloud Platforms
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Integration
Data Pipeline
Data Platform
Data Processing
Data Science
Data Warehouse
Database
Databases
Digital Marketing
ETL
GCP
Google Cloud
Google Cloud Bigquery
Google Cloud Platform
Graph Database
Informatica
Information Technology (IT)
Integration
Lead Data Engineering
Machine Learning
Reporting and Analytics
SQL
Job Description
Lead data engineering and AI/ML delivery in Cybersecurity within Identity Access Management, turning identity event analytics into safer access governance and anomaly detection.
Responsibilities
- Lead complex initiatives with broad impact; contribute as a key participant in large scale software planning for Identity and Access Management
- Design, develop, and run tooling to discover issues in data and applications; escalate findings to engineering and product leadership
- Apply statistical and data science methods to Identity and Access Management business problems
- Act as a subject matter expert on ML, AI, and mathematical and statistical techniques applied to large datasets
- Design, support, and operate data pipelines, data models, dashboards, and API integrations for real-time and batch analytics
- Design and conduct experiments, statistical analyses, and hypothesis testing to evaluate proposals supporting controls, policies, and operational processes
- Build AI capabilities to identify inappropriate access, recommend entitlements prior to user requests, and detect anomalous behavior across millions of identity events
- Lead IAM team members and coordinate with operations, onboarding, initiatives, engineering teams, and lines of business and lines of defense to translate analytical needs into technical solutions
- Establish and implement engineering and analytical best practices for solution development
- Develop solutions in alignment with security, privacy, model risk, and regulatory guidelines
- Develop, test, deploy, and support ML-enabled analytical solutions; guide teammates in ML, AI, and statistical techniques to drive AI/ML adoption across IAM
- Assist with monitoring model health, reliability, and drift to enable continuous improvement and required remediation
- Demonstrate proficiency with AI-assisted development and analysis tools, including GitHub Copilot and approved code-centric agents
- Use AI to accelerate system design, coding, testing, analysis, and troubleshooting
- Validate and integrate AI-assisted outputs using strong technical judgment, accounting for model limitations, security risks, and operational considerations
- Apply AI responsibly in development and production environments, ensuring alignment with security, compliance, privacy, and ethical standards
Requirements
- 5+ years of Database Engineering experience, or equivalent demonstrated through a combination of work experience, training, military experience, and education
- 5+ years of experience with Python data science libraries including Pandas, NumPy, and Scikit-Learn, plus exposure to a deep learning framework such as TensorFlow or PyTorch
Technologies
- Python, Pandas, NumPy, Scikit-Learn
- TensorFlow, PyTorch
- Vertex AI, GCP, BigQuery
- Neo4j
- GitHub
- Power BI, Tableau, Alteryx
- GitHub Copilot, Claude Code, Cursor, Devin
Benefits
- Health benefits
- 401(k) Plan
- Paid time off
- Disability benefits
- Life insurance, critical illness insurance, and accident insurance
- Parental leave
- Critical caregiving leave
- Discounts and savings
- Commuter benefits
- Tuition reimbursement
- Scholarships for dependent children
- Adoption reimbursement
Desired Qualifications
- Knowledge of Vertex AI and GCP environments, BigQuery, and deploying models into production on these platforms
- Knowledge of graph networks for anomaly detection (Neo4j)
- Learning mindset to stay current with developments in the field
- Knowledge of GitHub for code management and familiarity with Power BI, Tableau, and Alteryx
- Understanding of software engineering fundamentals such as testing, debugging, and code reviews (highly desirable)
- Knowledge of the mathematical foundations of statistics, machine learning, and modern AI techniques
- Strong understanding of model evaluation techniques for classification and clustering, including precision, accuracy, F1-scores, confusion matrices, and ROC curves
- Comfort with AI coding agents (GitHub Copilot, Claude Code, Cursor, Devin or others) and ability to critically examine AI code for fitness for use
- Familiarity with LLMs, RAG, and AI agents
- Strong SQL skills and ability to work with relational and analytical databases
- Strong book-of-work management and organizational skills
- Hands-on ability to manipulate data and tools to prototype and present solutions
- Confident, self-motivated producer of original ideas and solutions with sound judgment for when to escalate issues
- Strong written and verbal communication and excellent presentation skills, including ability to communicate complex technical concepts
- Bachelor’s or Master’s degree in Computer Science, Data Science, AI, Engineering, or related field
Location and Work Model
- Irving, TX (onsite)
Compensation
- Salary range: $119,000 - $206,000 per year
Posting Notes
- Job posting may come down early due to volume of applicants
- Required location(s) listed; relocation assistance is not available
- Salary range is determined by location of the job; may be considered for discretionary bonus, Restricted Share Rights, or other long-term incentive awards
- Not eligible for visa sponsorship
- Role is not eligible for 100% remote work
Posting end date: 22 Sep 2026