Datadog brand banner

Senior Applied Scientist - AI Platform

DatadogParis, FrancePosted 2h ago
via Greenhouse

AI Platform builds the foundations of Datadog's AI efforts. The org is 70+ people organised in three pillars: training and serving (GPU clusters, distributed training, low-level infrastructure), agents (agent harnesses, memory systems, the internal AI gateway that routes every LLM request at Datadog), and evaluation and experimentation. This role sits in the evaluation and experimentation pillar, which owns Datadog's shared annotation and evaluation infrastructure — including the evaluation scenario store and the telemetry archival systems used across the Bits org. Together they let an agent travel back in time and query what Datadog looked like at the exact moment an incident happened, so scenarios can be replayed and agent performance tracked over time.

Specifically, you'll be the first applied scientist on GenSim (Generative Simulations), the team that builds the environments Datadog's agents learn in. GenSim doesn't replay sampled telemetry — it stands up real, fully instrumented applications that talk to Datadog, drives them with representative traffic, injects controlled failures, and records what happens. Because GenSim injected the failure, it knows the ground truth. That corpus — hundreds of postmortem-derived scenarios and thousands of runnable applications — is today the primary source of post-training data for Datadog's own SRE model, and the substrate that Bits AI SRE and our other agents are trained and evaluated against.

The team has no applied science support today and is learning post-training data methodology on the fly. That's the gap this role fills, and the open questions are the interesting part. How do you tell whether a generated environment is actually representative of the messy, incomplete telemetry real customers run — rather than a suspiciously clean one where every monitor exists and every service emits complete logs? How do you make injected problems genuinely hard, and how do you even measure difficulty? How do you define and control the quality of post-training data when correctness, representativeness and difficulty pull in different directions? How do you evaluate an agent end to end when the trajectory is non-deterministic? Creating simulated agent environments for monitoring and SRE work is not well solved in open source or in published research, and Datadog is the leading company in this field. If those are the problems you want to spend your time on, come build this with us.

What You'll Do:

  • Own the applied science direction for GenSim: set the methodology and the forward-looking technical calls on how simulated environments and post-training data should be built, on a team where that decision-making does not exist yet.
  • Define, measure and raise the quality of post-training data — basic correctness, representativeness against the real distribution of customer systems and production telemetry, and difficulty — and make those measures something the team can act on release over release.
  • Close the realism gap. Simulated environments today are too clean and the injected problems are not yet hard enough; you'll drive the research and the engineering that make them look like real, imperfect production systems.
  • Build scalable, production-grade systems rather than research scripts. The output is not just a dataset — it is a system of synthetic environments that must be reliable and invokable inside a training loop.
  • Determine how this data is best applied, in LLM post-training and in evaluation, and own the agent and LLM application evaluation approaches for these environments.
  • Work cross-functionally with the engineers and applied scientists on adjacent teams — Bits AI SRE, the model training effort, and the wider evaluation and experimentation pillar — so that what you learn moves freely in both directions.

Who You Are:

  • You have a PhD, MS or equivalent research experience in a scientific field, with strong applied mathematics grounding.
  • 6+ years of relevant applied science or ML engineering experience, including setting technical direction for others.
  • You have hands-on experience with LLM and agent post-training data: how it is created, managed, and how training-data quality is controlled. This is the requirement that matters most.
  • You have real domain expertise in LLMs and agentic applications — not classical ML fine-tuning. Fine-tuning classifiers or traditional models is a different problem from the one this team is solving.
  • You have evaluated agents or LLM applications, and can define what 'good' means before you measure it.
  • You are a strong programmer and production software engineer. Python at minimum, plus the ability to ship scalable production systems and work with distributed systems.
  • You collaborate well across engineering and science teams, and you're comfortable being the domain expert who decides what comes next.
  • You thrive in ambiguity and can make sound technical calls when the path isn't yet defined.

Bonus Points:

  • Hands-on LLM fine-tuning, post-training or model training experience.
  • Background in statistics, experiment design and data analysis.
  • Experience deploying production-level ML infrastructure.
  • Observability or monitoring systems background.
  • Architecture-level understanding of LLMs.

Benefits & Growth:

  • New hire stock equity (RSUs) and employee stock purchase plan (ESPP)
  • Continuous professional development, product training, and career pathing
  • Intra-departmental mentor and buddy program for in-house networking
  • An inclusive company culture and the ability to join our Community Guilds
  • Access to Inclusion Talks, our internal panel discussions
  • Free, global Spring Health benefits for employees and dependents age 6+
  • Competitive global benefits

Benefits and Growth listed above may vary based on the country of your employment and the nature of your employment with Datadog.

#LI-Hybrid


About Datadog: 

Datadog is the leading observability and security platform for the AI era, providing businesses with unified visibility across the technology stack to manage complexity at scale. It brings applications, infrastructure, data, models, and security into one place, using AI to detect and resolve issues before they impact customers. Trusted globally by Fortune 500 companies and high-growth AI leaders, Datadog enables businesses to move faster with clarity and confidence. Learn more about #DatadogLife on Instagram, LinkedIn, and Datadog Learning Center.


Equal Opportunity at Datadog:

Datadog is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and other characteristics protected by law. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. Here are our Candidate Legal Notices for your reference. 

Datadog endeavors to make our Careers Page accessible to all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please complete this form. This form is for accommodation requests only and cannot be used to inquire about the status of applications. 

Privacy and AI Guidelines:

Any information you submit to Datadog as part of your application will be processed in accordance with Datadog’s Applicant and Candidate Privacy Notice. For information on our AI policy, please visit Interviewing at Datadog AI Guidelines.

Similar roles

  • Applied Materials logo

    College Intern

    Applied Materials·Singapore, Singapore

    **Who We Are** Applied Materials is the global leader in materials science and engineering solutions that are at the foundation of virtually every new semiconductor chip and advanced display in the world. The equipment that we create and service is essential to advancing AI and accelerating the commercialization of next-generation semiconductor chips. Join us and push the boundaries of materials…

    • Full-time
    • Tech.Passself-sponsored, 2-yr
  • Rightangled logo

    R&D / Genetic Researcher

    Rightangled·Dubai, United Arab Emirates

    SIBIE LTD HOLDING GROUP Sibie Ltd is a healthcare group. We are an innovative company rapidly expanding our services across the UK, Europe, the USA, and the Gulf region. We are a preferred partner for several leading healthcare organisations due to the high quality of our products and services. Role Overview We are looking for a highly skilled and research-driven…

    • Full-time
    • Green Visaself-sponsored, no tie
  • Upstart logo

    Applied Scientist

    Upstart·Worldwide

    About Upstart At Upstart, we’re united by a mission that matters: to radically reduce the cost and complexity of borrowing for all Americans. Every day, we bring creativity, experimentation, and advanced AI to reshape access to credit, helping millions move forward financially with clarity and confidence. As the leading AI lending marketplace, we partner with banks and credit unions to…

    • Remote
    • Full-time
  • Booyco Electronics  Ltd logo

    Incoming Quality Inspector

    Booyco Electronics Ltd·South Africa

    Main Purpose of the Job Conduct thorough inspection and testing incoming products to verify if it meets all Booyco specifications and quality requirements. This role is pivotal in ensuring our electronic products adhere to rigorous quality standards through meticulous inspection and testing procedures. Education, experience and competencies Matric or N3 minimum. N4 or higher is preferential. Qualification in quality Inspections…

    • Full-time
  • Agence Appel Médical Seine et Marne logo

    AIDE SOIGNANT (F/H)

    Agence Appel Médical Seine et Marne·Rebais, France

    Parce que la santé est un enjeu essentiel, dont l'humain est la pierre angulaire, les femmes et les hommes d'Appel Médical sont engagés depuis plus de 50 ans aux côtés des soignants.Grâce à notre expertise métiers, nos spécialistes sauront vous accompagner au quotidien pour concrétiser vos souhaits professionnels. Pour un poste dans le secteur médical, nous mettons notre expertise en…

    • Contract
  • Agence Appel Médical Clermont Ferrand logo

    INFIRMIER DE (F/H)

    Agence Appel Médical Clermont Ferrand·Riom-ès-Montagnes, France

    Rejoignez les 30 000 collaborateurs de l'Appel Médical et bénéficiez de nombreuses missions et emplois les plus adaptés à vos envies et compétences tout en profitant des nombreux services et avantages exclusifs. Les fonctions ou intitulés se déclinent au féminin comme au masculin. Comment aimeriez-vous contribuer, en tant qu'Infirmier(ère), au bien-être des patients dans notre Établissement de Soins de Suite…

    • Contract
  • Merge logo

    Risk Operations Specialist

    Merge·France

    About Merge: Merge is a regulated stablecoin payments infrastructure platform enabling enterprises to move money globally in real time. We combine stablecoin rails with local instant payment systems to deliver faster, cheaper, and more transparent cross-border payments...

    • Full-time
  • Designer Fund logo

    Algorithm Researcher

    Designer Fund·United Kingdom

    Via is on a mission to create public transportation systems that provide far greater access to jobs, healthcare, and education. Our platform serves as the technology backbone for modern transit networks, transforming antiquated and siloed public transportation systems ...

    • Full-time