The English-language job board for Spain

Model Serving Engineer

Remote, Europe +1·Full-time·Added 5 months ago

This job was published 5 months ago

Fundamental

12 open roles

Overview

Job details

  • Full-time hours

    Per the ad.

  • Fully remote

    Only from certain places, per the ad: “Job Location: Europe / United States”

  • A relocation package

    Help with the move, per the ad.

  • Relocation help

    Indicative — the ad doesn’t promise sponsorship.

Requirements

  • Work in English

    English is required, per the ad.

  • Have 5+ years of experience

    Mid-level role.

  • University degree

    “Bachelor's or Master's degree in Computer Science, Engineering, or a related field (or equivalent practical experience)” — per the ad.

Pay & benefits

What you'll get

  • Competitive compensation with salary and equity

  • Comprehensive health coverage for you and your dependents

  • Paid parental leave for all new parents, inclusive of adoptive and surrogate journeys

  • Relocation support for employees moving to join the team in one of our office locations

  • A mission-driven, low-ego culture that values diversity of thought, ownership, and bias toward action

Requirements

Nice to have

  • Understanding of GPU architecture, performance characteristics, and resource utilization in high-performance compute workloads

  • Experience working with tabular and structured-data ML systems

  • Understanding of neural networks and modern deep learning architectures

  • Experience with Kubernetes and cloud infrastructure

  • Familiarity with DevOps and production infrastructure tooling, including containers, Helm, observability, and CI/CD systems

The role

Our Serving team is responsible for turning NEXUS, our Large Tabular Model, into a reliable and scalable production system. We own the infrastructure and execution stack that serves the model across multiple deployment environments, each with different requirements around scale, isolation, performance, and trust.

The team sits at the intersection of research and production engineering. We work closely with researchers to bring new model architectures into production, while building the systems needed to operate them efficiently and predictably under real-world workloads. Tabular foundation models introduce serving challenges that differ meaningfully from traditional LLM inference, including irregular computational behavior and complex resource tradeoffs across CPU, GPU, memory, and networking.

As a Model Serving Engineer, you'll work across the full inference stack - from Python runtime performance and concurrency behavior to distributed orchestration, GPU serving infrastructure, and deployment architecture. You'll identify bottlenecks, improve throughput and latency, and help define how new generations of the model are translated from research artifacts into production-grade systems.

This is a deeply technical, Python-heavy role for engineers who enjoy distributed systems, performance optimization, and low-level infrastructure challenges close to modern ML systems.

What you'll do


Key responsibilities

  • Optimize Python inference code for performance under real concurrency constraints, including GIL contention, multi-threading, multiprocessing, async execution, and long-running production workloads

  • Work closely with research to understand model internals and support the continuous evolution of the architecture, especially around complex and non-obvious computational behavior under production load

  • Collaborate with research and infrastructure teams to reason about hardware utilization and serving tradeoffs across GPU, CPU, memory, networking, batching, and concurrency

  • Define and evolve the architecture behind our distributed inference and asynchronous execution stack, including orchestration, worker coordination, and end-to-end concurrency patterns

  • Own the Triton serving layer for NEXUS, including how models are packaged, configured, and executed as part of our production inference pipeline

  • Build observability and performance tooling across the serving stack, and use production metrics to drive tuning decisions around latency, throughput, and resource efficiency

  • Solve cross-cutting serving challenges that emerge from deploying the same model across environments with very different scale, isolation, and reliability constraints

  • Evaluate and integrate new inference runtimes, serving strategies, and infrastructure approaches as the model ecosystem evolves

Must have

  • Bachelor's or Master's degree in Computer Science, Engineering, or a related field (or equivalent practical experience)

  • 5+ years of experience in model serving, ML infrastructure, or a closely related backend engineering role

  • Deep expertise in Python concurrency, including GIL behavior, multi-threading, thread safety, multiprocessing

  • Experience building asynchronous and message-driven systems

  • High-performance, large-scale distributed systems

  • Ability to read and reason about ML model implementations at a computational level, including compute behavior, batching, memory usage, and inference characteristics

  • Experience profiling and optimizing performance across CPU, memory, I/O, and ideally GPU workloads, and translating findings into architectural improvements

About Fundamental

Fundamental is a London-based AI research and engineering company that develops advanced machine learning systems and developer tools. The company focuses on applied AI, bridging cutting-edge research with production-grade engineering to build products that serve developers and enterprises. Its team is composed of researchers, ML engineers, and software engineers working on model serving, MLOps, and applied AI challenges.

Fundamental is expanding its engineering and research presence in Spain, with a dedicated hub that hosts roles across ML research, MLOps, backend, and full-stack development. The company offers a research-driven culture with a strong emphasis on technical excellence, making it an attractive destination for international engineers and researchers looking to work on challenging AI problems while based in Spain.

Employees
11–50
Headquarters
London, UK

Good to know if you are moving

  • Fundamental is headquartered in London, UK, and is expanding its engineering hub in Spain, offering international professionals the chance to work on AI research and applied engineering from Spain.
  • The company's open roles in Spain span ML research, MLOps, backend, and full-stack engineering, indicating a strong focus on technical and research-oriented positions.
  • As a smaller, research-driven company, Fundamental likely offers a collaborative environment with direct impact on product decisions, appealing to engineers seeking meaningful contributions.
  • The Spain office appears to be a key part of the company's growth strategy, with multiple team lead roles (e.g., MLOps Team Lead, DevOps Team Lead) suggesting opportunities for career advancement.

In their own words

About the company & team

Fundamental is an AI company pioneering the future of enterprise decision-making. Founded by DeepMind alumni, Fundamental has developed NEXUS - the world's most powerful Large Tabular Model (LTM) - purpose-built for the structured records that actually drive enterprise decisions. Backed by world class investors and trusted by Fortune 100 companies, Fundamental unlocks trillions of dollars of value by giving businesses the Power to Predict.

At Fundamental, you'll work on unprecedented technical challenges in foundation model development and build technology that transforms how the world's largest companies make decisions. This is your opportunity to be part of a category-defining company from the ground-up. Join the team defining the future of enterprise AI.

All open roles at Fundamental
WhatsApp
BarcelonaRelocation helpNo Spanish needed
BarcelonaRelocation helpNo Spanish needed

ML Researcher

Fundamental11mo ago
BarcelonaRelocation helpNo Spanish needed

MLOps Team Lead

Fundamental3mo ago
Remote, Europe +3Relocation helpNo Spanish needed

Data Scientist - Extensions

Fundamental3mo ago
Remote, Europe +3Relocation helpNo Spanish needed

Backend Engineer - Extensions

Fundamental3mo ago
Remote, Europe +3Relocation helpNo Spanish needed

Applied AI Engineer

Fundamental4mo ago
Remote, Europe +3Relocation helpNo Spanish needed

DevOps Team Lead

Fundamental8mo ago
Remote, Europe +3Relocation helpNo Spanish needed

Full-Stack Engineer

Fundamental8mo ago
Remote, Europe +1Relocation helpNo Spanish needed

DevOps Engineer

Fundamental10mo ago
Remote, Europe +1Relocation helpNo Spanish needed

MLOps Engineer

Fundamental2mo ago
Remote, Europe +3No Spanish needed
See all 12 jobs