
Senior Research Engineer (ML infrastructure )
We are looking for a Senior Research Engineer to build and improve the ML infrastructure that powers our deep learning research and trading.
Our DL strategies are becoming an increasingly important driver of the business, and the infrastructure behind them has a direct impact on how quickly researchers can test ideas, how efficiently we use compute, and how reliably successful research reaches production across markets and asset classes.
You will work at the boundary between deep learning research and infrastructure. Your job is not just to maintain ML systems, but to identify what is slowing research down, solve technically difficult problems across the stack, and turn research prototypes into reliable, scalable production systems.
The scope is broad: training frameworks, datasets and data pipelines, distributed training, experiment infrastructure, reproducibility, observability, and tooling around the research lifecycle. Depending on your strengths, you may go deep in one of these areas or work across several of them.
This is a highly hands-on IC role with substantial technical ownership. You will work closely with researchers, other research engineers, and infrastructure teams, and will be expected to independently drive important problems from diagnosis to production.
Responsibilities
- Turn successful research prototypes into robust implementations that can be trained, validated, and deployed across multiple markets and asset classes;
- Work directly with researchers to remove infrastructure bottlenecks from the research loop — make new ideas easier to prototype, experiments faster to run, and failures easier to understand;
- Build and improve our internal ML platform — training framework, datasets and data pipelines, orchestration, experiment tracking and reproducibility tooling, GPU/compute infrastructure, and tooling for releasing models into production;
- Own technically challenging areas of the ML stack end-to-end: identify problems, design solutions, implement them, measure their impact, and maintain them in production;
- Make training reproducible and observable: investigate regressions, debug models that no longer reproduce from master, and improve tooling around data, experiments, model quality, and training behavior;
- Take ownership of shared training, dataset, and research infrastructure code — proactively find bugs, reduce technical debt, improve abstractions, and maintain a high bar through rigorous code review;
- Work across team boundaries when a research problem spans datasets, storage, compute, training infrastructure, or production systems;
- Identify areas where researchers are repeatedly paying an infrastructure tax and build reusable solutions instead of fixing the same problem case by case;
- When useful, contribute directly to research: run experiments, investigate model behavior, prototype architectural or optimization ideas, and help push model quality forward.
Requirements
- A strong, versatile ML systems / research engineer who is comfortable working on ambiguous problems at the intersection of deep learning and infrastructure;
- Strong background in at least one of:
- large-scale ML training systems and distributed training;
- ML/data infrastructure, including datasets, storage formats, orchestration, experiment tracking, and configuration;
- research engineering around large-scale deep learning systems;
- Hands-on experience making ML research faster or more reliable: for example, by reducing training time, improving hardware utilization, accelerating data access, improving experimentation workflows, or eliminating recurring infrastructure bottlenecks;
- Ability to independently investigate complex systems: profile them, form hypotheses, run experiments, trace problems across abstraction boundaries, and arrive at practical solutions;
- Strong engineering judgment and a high bar for code quality, reproducibility, observability, and maintainability in a large shared ML codebase;
- Strong product sense toward internal research infrastructure: you care not only whether a system works, but whether researchers can use it effectively and iterate quickly;
- Ability to work closely with researchers, understand what they are trying to achieve, and translate research needs into infrastructure improvements that compound over time;
- Strong ownership and initiative: you notice important problems that do not have a clear owner and are willing to take them from an unclear state to a working solution;
- Ability to influence technical direction through expertise, design work, code review, and collaboration without requiring formal management authority;
- While prior trading experience can be useful, it is not a prerequisite. Our priority is a first-principles mindset, strong technical depth, and a willingness to rethink how quantitative ML research infrastructure should work.
What we offer
- High base salary and social benefits;
- Generous bonus structure. We are very flexible in discussing salary and conditions of employment;
- Cutting-edge hardware and software in production, as well as high technical expertise across the company, which allows us to implement bold ideas and achieve great results. Ownership over initiatives that directly solve business problems;
- Ability to trade on dozens of international exchanges;
- Flexible workflow (lack of formalism and bureaucracy, no pressure and over-management) and working schedule;
- Tuition reimbursement, conference and training sponsorship.
Similar roles
Graduate Engineer (Mechanical/HVAC)
Johnson Controls·Sydney, Australia
Johnson Controls , a global leader in thermal management, mission-critical building systems, energy efficiency, and decarbonization, helps customers use energy more productively, reduce carbon emissions, and operate with the precision and resilience required in rapidly expanding industries such as data centers, healthcare, pharmaceuticals, advanced manufacturing, and higher education. For more than 140 years, Johnson Controls has delivered performance where it…
- Full-time
- Skilled Independent 189 — PR, no sponsor
Head of Methodology and R&D
Meridia Land·Amsterdam, Netherlands
ABOUT MERIDIA Meridia is an AgTech company specialised in software solutions for field data within complex agri-commodity supply chains. Meridia Verify®, a SaaS product, verifies supply chain field data against regulatory and voluntary sustainability frameworks in minutes and provides hands-on guidance for risk mitigation. Procurement and sustainability teams eliminate commercial and reputational risk, prevent supply-chain disruptions, and shift from guesswork…
- Hybrid
- Full-time
- Highly Skilled Migrant — sponsor-tied, fast settlement
Senior Geotechnical Engineer/Engineering Geologist
Civic·London, United Kingdom
Civic is a team of system thinkers in the built environment, creating positive impact for people, places and the planet. We want our work to have a positive impact on the environment — helping people to live healthier, happier lives while respecting natural limits. We build without destroying, explore without exploiting, and seek to reuse rather than just consume. Above…
- Hybrid
- Full-time
Senior Mechanical Design Engineer
IONATE·London, United Kingdom
London / Hybrid: 3 - 5 days in office per week IONATE is a deep tech scale-up building the hardware and software backbone for smart grids. Our mission is to transform power systems - from grids and microgrids to renewables and data centers - unlocking the massive potential in this under-innovated sector that touches every aspect of our modern lives.…
- Hybrid
- Full-time
Technical Escalations Engineer 2 (Database Monitoring) - EMEA
Datadog·France
THE TEAM As Datadog’s in-house product experts, the Technical Escalation Engineering (TEE) team plays a critical role in driving our global success. We enable our customers, from the world’s most innovative startups to the largest enterprises. Through deep technical expertise, relentless problem-solving, and exceptional customer engagement, we educate, guide, and troubleshoot, delivering high-impact solutions that shape the customer experience. Whether…
- Full-time
Principal Solutions Engineer - UKI
Amplitude·London, Canada
<div class="content-intro"><p>Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data,…
- On-site
- Full-time
- Express Entry — PR day one, no employer
Human Factors Graduate / Human Factors Engineer
Team Consulting·Ickleton, United Kingdom
Cambridge, UK | Hybrid Working Are you passionate about understanding people and designing products that improve lives? We're looking for a Human Factors Graduate or Human Factors Engineer to join our established Human Factors team. Whether you're starting your career or already have industry experience, you'll help develop innovative medical devices and healthcare technologies used by patients and healthcare professionals…
- Hybrid
- Full-time
Fieldmanager Sales
CPM Benelux·The Hague, Netherlands
Ervaring met consumenten elektronica én klaar voor een volgende stap? Word Field Sales Manager bij CPM Benelux en ga aan de slag voor een groot CE merk! Als Field Sales Manager draag je bij aan de ontwikkeling van ons Field Sales Team binnen het Media Markt kanaal en zorg jij ervoor dat de sales van de nieuwste televisies en audio…
- Full-time
- Highly Skilled Migrant — sponsor-tied, fast settlement