Why This Role Stands Out
This hybrid role offers an exciting opportunity to leverage cutting-edge LLM-based methods and contribute to impactful digital intelligence solutions within a reputable financial institution. You'll thrive here if you possess a strong software engineering background and hands-on machine learning development experience, eager to turn innovative ideas into scalable products. Apply today to join a high-caliber team and advance your career in machine learning.
Quick Overview
Job Description
hackajob is collaborating with J.P. Morgan to connect them with exceptional professionals for this role.
JOB DESCRIPTION
Consumer & Community Banking division serves our Chase customers through a range of financial services, including personal banking, credit cards, mortgages, auto financing, investment advice, small business loans and payment processing. We're proud to lead the U.S. in credit card sales and deposit growth and have the most-used digital solutions - all while ranking first in customer satisfaction. In this role, you'll apply strong technical judgment to choose the right approaches (including modern LLM-based methods where appropriate), evaluate performance with rigorous metrics, and ensure solutions are reliable, secure, and scalable in real-world environments. You'll also contribute to improving data quality and feedback loops, monitoring models in production, and continuously iterating to reduce agent effort, shorten resolution times, and increase consistency and quality across operational workflows.
As a Senior Machine Learning Engineer-Digital intelligence in the Digital Intelligence team, you will be collaborating with a high-caliber team of software developers and deep learning experts, you will specialize in large language modeling, optimization, interpretability, and related algorithms
The ideal candidate brings a strong software engineering foundation combined with hands-on, zero-to-one machine learning development experience. You will possess broad expertise in post-training machine learning models - including quality and performance optimization - alongside deep knowledge of large language models and modern deep learning techniques. Above all, you will have a demonstrated ability to operate at the intersection of research and engineering, turning promising ideas into scalable, real-world products within a fast-paced, collaborative environment.
Job Responsibilities
-
Research and prototype next-generation architectures for structured and unstructured data
-
Develop novel pre-training objectives tailored to financial event sequences and heterogeneous profile data
-
Implement research ideas in production-quality code
-
Mentor engineers on ML best practices; translate research advances into deployable systems
-
Optimize training throughput for large data sources
-
Collaborate with other teams to design solutions for product use cases.
Required qualifications, capabilities, and skills:
-PhD with 2+ years OR Master's degree with 4+ in Computer Science, with training and work experience in Machine Learning, LLM/NLP or similar fields.
-Deep LLM and Transformer expertise - strong command of attention mechanisms, positional encodings such as RoPE, and the ability to handle multi-modal data inputs effectively.
-PyTorch proficiency at scale - hands-on experience with distributed training frameworks including FSDP and DeepSpeed, alongside practical memory optimization techniques.
-Foundation model training - proven experience in pre-training from scratch and designing tokens and vocabularies for complex, heterogeneous data sources including tabular, temporal, and graphical formats.
-Strong software engineering skills - ability to build robust, production-quality systems that perform reliably at scale.
-Prior experience with financial data and recommendation systems.
Preferred qualifications, capabilities, and skills:
-
Publication record at top AI/ML venues.
-
Experience optimizing serving infrastructure is a plus.
-
Experience with post-training LLMs and network optimization algorithms, as well as interpretability or steering techniques for LLMs.
-
Experience working with large-scale compute infrastructure.
-
Experience shipping a real-world product, project, or feature.
Similar jobs
- DE
Staff Inference Engineer
NewDesignworkstalent
Bellevue🇺🇸Hybrid8 hours agoEngineering - JM
Software Engineer III - Applied AI/ML Engineer (Asset & Wealth Management)
NewJ.P. Morgan
Jersey City, New Jersey🇺🇸On-site1 hour agoAWSAgilePythonTechnology - BU
Staff Machine Learning / Operations Research Engineer
NewBurq, Inc.
United States🇺🇸Remote7 hours agoMLOpsDeep LearningForecasting+2Technology - NT
Senior Machine Learning Engineer Computer Vision // HYBRID
NewNeumeric Technologies Corporation
San Francisco, CA🇺🇸HybridYesterdayAWSMachine LearningComputer Vision+2Technology - VI
Sr. ML Engineer
NewVisa
Austin, Texas🇺🇸$123.4k - $191.1k/yrHybrid5 hours agoDockerShellAWS+9Technology - CO
Director, Machine Learning Engineer
NewCapital One
Mc Lean, Virginia🇺🇸$269.1k - $307.2k/yrHybrid6 hours agoGCPScalaAWS+12Technology