ISC2

Lakehouse Machine Learning Engineer

Job Locations US-Remote
Posted Date 4 hours ago(9/18/2026 12:51 PM)
Job ID
2026-2543
# of Openings
1
Category
Information Technology

Overview

Your Future. Secured. ISC2 is a force for good. As the world’s leading nonprofit member organization for cybersecurity professionals, our core values — Integrity, Advocacy, Commitment, Inclusion, and Excellence — drive everything we do in support of our vision of a safe and secure cyber world. Our globally recognized, award-winning portfolio of certifications provide an independent and globally recognized endorsement of cybersecurity knowledge, skills and experience for all career levels. Our charitable arm, the Center for Cyber Safety and Education, enables ISC2 and our members to serve the public by educating the most vulnerable about cyber risks and empowering access to enter and thrive in the cyber profession. Learn more at ISC2 online and connect with us on Twitter, Facebook and LinkedIn. When you join ISC2, you’ll demonstrate your commitment to an inclusive and equitable environment. Your support of the unique perspectives and experiences shared by our global cybersecurity workforce and profession will be recognized. We invite you to take an active role in helping us create a true sense of belonging across our organization — an environment of authenticity, trust, empowerment and connectedness that empowers all of our successes. Learn more.

Position Summary

ISC2 is building a modern lakehouse and using it to run machine learning across the membership business. This role covers the whole path, from the pipelines that produce the data through to results a stakeholder actually uses. The Lakehouse Machine Learning (ML) Engineer writes production Python and Spark against a governed lakehouse, spends real time exploring data before deciding what’s worth modeling, and deploys and monitors models once they’re live. As the platform matures there is room to take on applied LLM work. The lakehouse runs on Azure Databricks; deep experience on a comparable stack transfers.

 

**This position is not available to residents of California**.

Responsibilities

  • Build and maintain Python/Spark pipelines through bronze, silver, and gold layers, plus the semantic datasets and ML models that consume them.
  • Dig into the data before modeling it. Work out what it can support, build and test the features that come out of that exploration, then coordinate with stakeholders about which questions are worth answering and which aren’t.
  • Work across a range of model types: survival and time-to-event, forecasting, classification and propensity, sequence models, recommenders, causal evaluation.
  • Put models into production and keep them there - Experiment tracking, model registry, scheduled inference, and monitoring for drift and decay once they’re live.
  • Get results to the people who need them. That means landing output in governed semantic tables feeding dashboards and CDP systems, and being able to walk a business team through what the numbers mean.
  • Take on applied LLM work as the platform matures, including structured extraction from free text, and retrieval over governed data.
  • Build inside the platform’s security and governance requirements rather than around them. Access controls, data protection, auditability, and human review where model output drives a decision that affects a member.
  • Turn what works into reusable patterns, including project templates, shared feature and evaluation code, and implementation standards, so the next model doesn’t start from a blank notebook.
  • Prove things out before they get built for real - Small proofs of concept that establish whether the data supports the theory, whether the approach holds up, and whether the result can actually be operated once it’s live.
  • Perform miscellaneous duties, as required.

Behavioral Competencies

  • Highly organized with strong attention to detail and documentation rigor.
  • Collaborative, intellectually curious, and proactive in identifying analytical opportunities.

Qualifications

  • Strong Extract/Transform/Load (ETL) skills, with the ability to assemble a dataset rather than request one.
  • Fluent Python and SQL skills, comfortable working across enterprise source systems, and fluent with the standard ML stack — scikit-learn at minimum, and at least one deep learning framework such as PyTorch or TensorFlow.
  • Knowledge range in ML, including knowing when timing matters enough to warrant survival analysis over a churn classifier, and being able to tell a causal question from a predictive one.
  • Familiarity with hyperparameter tuning, cross-validation, and the techniques used to confirm a model holds up on data it has not seen.
  • Ability to perform careful model validation, and to provide clarity about uncertainty. Also, to perform model evaluation and bias mitigation, judgment about where a person needs to stay in the loop, and the ability to explain the result to executives in non-technical terms.
  • Understanding of data security, privacy, and compliance constraints, as well as how they shape what can be built with member data. Ability to work within the confines of access controls, data protection, and auditability.
  • Knowledge of Databricks, including Unity Catalog, Workflows, MLflow, a plus.
  • Ability to perform cohort-based or hierarchical forecasting at scale, a plus.
  • Working knowledge of Salesforce, a plus.
  • Relevant certifications: Databricks Data Engineer or Machine Learning Associate/Professional, Azure AI Engineer or Data Scientist Associate, or equivalent, a plus.

Education and Work Experience

  • Bachelor’s or Master's degree in an IT field preferred. Will consider candidates with a high school diploma or equivalent and 7+ years of hands-on experience in data engineering and applied machine learning, preferably in enterprise environments.
  • 3+ years of hands-on experience in data engineering and applied machine learning, preferably in enterprise environments.
  • Experience deploying and monitoring models in production. ML flow or an equivalent tracking, registry, and scheduled-inference stack.
  • Production experience building in a medallion architecture — bronze, silver, and gold, or an equivalent layered model — in a data-catalog-governed environment. Distributed processing with Spark or a comparable engine, an open table format such as Delta or Iceberg, and catalog-managed schemas, lineage, and access control. Demonstrated experience building and operating in that environment, not just querying it.
  • Practical Large Language Model (LLM) experience including embeddings and retrieval, structured extraction, evaluation, a plus.

  • Experience with subscription or membership-lifecycle data, a plus.

  • Experience running a build-versus-buy evaluation. Hands-on assessment of tools and vendors, and a recommendation that can be defended, a plus.

Physical and Mental Demands

  • Up to 5% travel may be required.
  • Work normal business hours and extended hours when necessary.
  • Remain in a stationary position, often standing or sitting, for prolonged periods.
  • Regular use of office equipment in a remote environment such as a computer/laptop and monitor computer screens.
  • Dexterity of hands and fingers to operate a computer keyboard, mouse, and other computer components.

Total Rewards

The pay range for this position is $93,000 - $118,900/Yr.

 

Final pay is based on several factors including but not limited to internal equity, market data, and the applicant’s education, work experience, certifications, etc.

 

Information regarding our comprehensive benefits package is available here.

 

This position will be posted for a minimum of 5 calendar days. This is a current vacancy, and the employer intends to fill this position within approximately 30 days.

Equal Employment Opportunity Statement

All qualified applicants will receive consideration for employment without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic as protected by applicable law. Job candidates will not be obligated to disclose sealed or expunged records of conviction or arrest as part of the hiring process.

Options

Sorry the Share function is not working properly at this moment. Please refresh the page and try again later.
Share on your newsfeed