SonicJobs Logo
Left arrow iconBack to search

Senior Developer

Intercontinental Exchange Holdings, Inc.
Posted 2 days ago, valid for 12 days
Location

Atlanta, GA, US

Salary

Competitive

Contract type

Full Time

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.

Sonic Summary

info
  • Intercontinental Exchange, Inc. (ICE) is seeking a full-time AI Platform Technical Lead to architect and manage the enterprise-wide platform for AI model training, deployment, and inference at scale.
  • The ideal candidate should have extensive experience in AI/ML training pipeline architecture and production model serving platforms, particularly with cloud infrastructure such as AWS, Azure, or GCP.
  • Candidates must possess strong leadership skills, advanced technical proficiency, and the ability to mentor MLOps and platform engineering teams effectively, with a focus on innovation and operational excellence.
  • An advanced degree in Computer Science, Machine Learning, or a related field is preferred, along with significant experience in AI infrastructure challenges and strategic thinking for platform scalability.
  • The position offers a competitive salary, though specific figures are not disclosed, and requires several years of experience in relevant fields.

Overview

Job Purpose

Intercontinental Exchange, Inc. (ICE) presents an opportunity for a full-time AI Platform Technical Lead to join and lead a team responsible for architecting and managing the enterprise-wide platform for AI model training, deployment, and inference at scale. The candidate will serve as a Technical Lead within the AI Center of Excellence team, playing a pivotal role in advancing the firm's strategic initiative to integrate Generative AI technologies responsibly and sustainably across the enterprise through robust training and inference infrastructure.

 

The ideal candidate must possess deep expertise in AI/ML training pipeline architecture, inference optimization, and production model serving platforms leveraging the latest advancements in Generative AI, distributed computing, GPU clusters, model optimization techniques, and high-performance inference systems.

 

This position demands advanced technical proficiency in training orchestration, model deployment pipelines, inference scaling, and performance optimization, innovative problem-solving capabilities, strong leadership qualities, and the ability to mentor and guide MLOps and platform engineering teams effectively. The role requires strategic vision for AI training and inference infrastructure roadmaps, including compute resource management, model lifecycle optimization, and real-time serving architectures. Exceptional professionalism, proactive collaboration, and outstanding communication skills are essential.

 

The candidate will actively engage and influence diverse stakeholders across the organization to align training and inference platform capabilities with AI model requirements and business SLAs, ensuring efficient resource utilization, optimal model performance, and cost-effective scaling. Strong written and verbal communication skills are imperative, given the candidate's responsibility to articulate training efficiency metrics, inference latency optimizations, resource allocation strategies, and platform ROI clearly and persuasively to both technical teams and executive audiences, including presenting model performance benchmarks, infrastructure cost optimization, and platform scalability roadmaps to senior leadership.

 

 

Responsibilities

  • Architecting, implementing, and managing enterprise-wide AI inference and training platform infrastructure.
  • Driving innovation, operational excellence, and scalability within AI/ML model serving and training environments.
  • Leading technical strategy for AI platform development and optimization across the organization.

 

Knowledge and Experience

  • Extensive experience and demonstrated leadership in designing and managing AI/ML training and inference platforms using cloud infrastructure (AWS, Azure, GCP).
  • Deep expertise in ML model serving frameworks (e.g., TensorFlow Serving, TorchServe, MLflow, Kubeflow).
  • Proficiency with GPU cluster management, distributed training and model optimization techniques.
  • Strong experience with AI/ML orchestration platforms, particularly Kubernetes for ML workloads and container technologies including Docker.
  • Comprehensive knowledge of MLOps pipelines, model versioning, A/B testing frameworks, and continuous integration for ML models.
  • Advanced programming skills in Python, experience with ML frameworks (TensorFlow, PyTorch, Hugging Face), and proficiency in performance optimization.
  • Experience with high-performance computing, inference optimization, and real-time model serving architectures.
  • Exceptional problem-solving skills in AI infrastructure challenges and strategic thinking for platform scalability.
  • Proven leadership abilities in guiding cross-functional AI/ML engineering teams and mentoring MLOps engineers.
  • Excellent written and verbal communication skills for technical and executive audiences.
  • Ability to effectively collaborate with data scientists, ML engineers, and business stakeholders to align AI platform capabilities with strategic objectives.

 

Preferred Knowledge and Experience

  • Advanced degree (PhD with few years' experience, or MS/BS with extensive experience) in Computer Science, Machine Learning, Data Engineering, or related field with focus on AI/ML systems.
  • Strong programming skills in Python with deep knowledge of ML libraries (scikit-learn, TensorFlow, PyTorch, Transformers).
  • Proficiency in ML model deployment frameworks, inference engines, and real-time serving APIs.
  • Working knowledge of vector databases, model registries, and feature stores (e.g., Feast, Tecton).
  • Experience with distributed computing frameworks (Spark, Ray) and GPU programming (CUDA) is highly beneficial.
  • Experience with AI model monitoring, performance tracking, and observability tools (Prometheus, Grafana, MLflow).
  • Extensive experience in cloud ML platforms (AWS SageMaker, Azure ML, Google AI Platform).
  • Deep experience with Kubernetes for ML workloads, Helm charts, and container orchestration for training pipelines.
  • Experience leading AI platform development in cross-functional teams of data scientists and ML engineers.
  • Expertise in CI/CD pipelines specifically for ML model deployment and automated retraining workflows.
  • Excellence in explaining complex AI infrastructure solutions to technical teams and business stakeholders.
  • Experience in enterprise AI/ML environments, working with governance, compliance, and responsible AI practices.

 

#LI-MA1

----------

Intercontinental Exchange, Inc. is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to legally protected characteristics.



Learn more about this Employer on their Career Site

Apply now in a few quick clicks

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.