Senior ML Research Engineer/Scientist
It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. That moment inspired Fred to build a company that could do that for everyone—freeing people from busywork so they could focus on meaningful work. Today, ServiceNow is the AI control tower for business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any workflow— helping 85% of the Fortune 500® work smarter, faster, and better. We're building an AI-native culture where technology and talent are unstoppable together. And we're just getting started.Join us to put AI to work for people. What you'll doOwn LLM mid-training and post-training research, including continued pretraining, SFT, preference optimization, and RL; make data-mixture and experimental decisions and determine how training changes affect downstream agent behavior.Research and prototype novel agentic architectures and algorithms across planning, reasoning, memory, skills, tool use, retrieval, and multi-agent collaboration, advancing beyond existing approaches where appropriate.Design and build research harnesses and experimentation methodologies that enable systematic experimentation, trajectory analysis, reproducibility, and rigorous comparison across models, checkpoints, and agent architectures.Define evaluation methodologies and develop novel benchmarks for measuring agent reasoning, planning, tool use, reliability, factuality, and safety; establish rigorous approaches for LLM-as-a-Judge, trajectory-based, and human evaluation.Identify systematic model and agent failure patterns, determine their root causes, and translate those insights into new research directions or improvements in training data, architecture, context, or evaluation.Independently identify research problems, formulate novel hypotheses, and drive projects from research idea to validated prototype and measurable impact, collaborating with engineering to transition successful approaches into production and communicating results through publications, patents, or open-source work. 5+ years of experience in machine learning, deep learning, AI research, or a related field, with demonstrated applied research experience and a track record of independently driving research projects.Hands-on experience with LLM training and post-training, including one or more of continued pretraining, SFT, preference optimization, or RL; experience making training-data decisions and understanding training dynamics and failure modes at scale.Strong Python and advanced PyTorch expertise, with experience modifying models, training pipelines, or research infrastructure to support novel experimentation.Strong practical depth in agentic AI and context engineering, including planning, reasoning, memory, skills, tool use, retrieval, long-context processing, and knowledge grounding.Experience designing evaluation methodologies, not just running evaluations, including benchmark desig...