Scientist II / Senior ML Scientist, Cofolding and Structure-Aware ML

Lila Sciences · Boston, MA · $228k - $358k
full-time senior Posted 3 weeks ago

Before you apply

Build my evidence-backed draft — free Apply on company site →

Paste your relevant resume section or 2–4 true bullets. See supported requirements and honest gaps. No account and no application sent.

Get weekly job alerts like this →

About this role

Your Impact at LILA Lila Sciences is seeking a Machine Learning Scientist, Cofolding and Structure-Aware ML to train next-generation cofolding models for drug discovery. This role is focused on improving models that reason over proteins, ligands, binding context, and experimental data, potentially using contrastive learning and related representation-learning approaches. This person should have direct experience training modern scientific ML models, not only using pretrained systems. You will work with ML researchers, computational chemists, computational biophysicists, data engineers, and drug discovery teams to develop models that learn from DEL and related datasets, connect molecular and protein context, and improve AI-driven discovery decisions. The models developed in this role should produce outputs that medicinal and computational chemists as well as biophysicists can interrogate, validate, and use in downstream agent-driven discovery decisions. What You'll Be Building Train and evaluate cofolding models for protein-ligand and related molecular discovery applications. Use contrastive learning, representation learning, self-supervised learning, or related methods where they help improve cofolding models trained on molecules, proteins, structures, and experimental readouts. Develop modeling approaches that make DEL data more useful for learning binding, enrichment, selectivity, and structure-activity signals. Build and evaluate models informed by Boltz, AlphaFold-style cofolding, equivariant GNNs, and related structure-aware ML methods. Design training objectives, including contrastive, self-supervised, or multimodal objectives, that connect ligands, proteins, structures, assays, simulations, and experimental data. Build rigorous evaluation frameworks that distinguish meaningful molecular learning from dataset artifacts, leakage, or spurious correlations. Collaborate with data and platform teams to define datasets, labels, negatives, controls, and metadata needed for model training. Partner with computational chemistry and biophysics teams to connect model outputs to physically and chemically meaningful hypotheses. Work with low-data learning scientists to identify which DEL, assay, simulation, or structural data would most improve model performance in focused chemical spaces. Work with research engineers to scale training, inference, and evaluation workflows. Help expose trained models and model-derived capabilities as tools for scientists and AI agents. What You'll Need to Succeed PhD or equivalent experience in machine learning, computational biology, computational chemistry, bioinformatics, computer science, or a related field. Hands-on experience training deep learning models for molecular, protein, structural biology, or scientific data applications. Experience with contrastive learning, representation learning, self-supervised learning, or multimodal learning. Familiarity with DEL or related selection, enrichment, screening, or molecular assay datasets. Experience with protein-ligand modeling, cofolding, structure prediction, geometric deep learning, or structure-aware molecular ML. Practical experience with PyTorch, JAX, or an equivalent ML framework. Ability to design careful experiments, ablations, and evaluations for scientific ML models. Strong understanding of data quality, leakage risks, negative construction, and benchmark design. Ability to collaborate across ML, data, computational science, and drug discovery functions. Bonus Points For Hands-on experience with DEL data. Drug discovery experience, especially in protein-ligand modeling or molecular optimization contexts. Experience with Boltz, AlphaFold or AlphaFold-derived methods, equivariant GNNs, diffusion models, protein language models, or molecular encoders. Experience training or extending cofolding, protein-ligand, protein-protein, structure prediction, diffusion, or geometric deep learning models. Experience with distributed model training and large-scale scientific data pipelines. Familiarity with active learning or closed-loop molecular design. Experience integrating ML models into agentic scientific workflows. Compensation We offer competitive base compensation with bonus potential and generous early-stage equity. Your final offer will reflect your background, expertise, and expected impact. U.S. Benefits. Full-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program. International Benefits. Full-time employees outside the U.S. receive a comprehensive benefits program tailored to their region. USD salary ranges apply only to U.S.-based positions; inter

Similar Jobs

Related searches:

On-site Jobs Senior Jobs On-site Senior Jobs Senior AI Agents & RAGSenior Data EngineeringSenior Generative AISenior Machine LearningSenior Computer Vision AI Jobs in Boston AI Agents & RAG in BostonData Engineering in BostonGenerative AI in BostonMachine Learning in BostonComputer Vision in Boston diffusion-modelssearchpytorchdata-pipelineagentsdeep-learning

Get jobs like this delivered weekly

Free AI jobs newsletter. No spam.