HC

H Company

Research Engineer, Model Inference & Serving - London

Hybrid

Part remote, part office, you need to live within commuting distance of a named location.

Hybrid · Hybrid London

Employer listed it 6 months ago · Added 5 days ago

Been open since 6 months ago. Long-running listings are sometimes left up after the role is filled.

Salary

Not stated

Location

Hybrid · Hybrid London

Timezone

Not stated

Contract

Full-time

Experience

Mid

Category

Software

This employer didn't state pay. Jobs like this usually pay around $140k–$245k a year, a typical range taken from 591 mid-level software roles on Nomaders that do state pay. It's a guide, not an offer.

Remote flexibility

Hybrid

This role is only partly remote, the employer expects time in the office around Hybrid London, Hybrid, so you need to live within commuting distance.

What the employer says

  • Source listing states candidate location: "Hybrid London, Hybrid"
  • Listing mentions "Hybrid" and 3 days a week in the office

What Nomaders makes of it

  • Not suitable if you plan to move between countries

The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.

About the role

Research Engineer, Model Inference & Serving

About H: H exists to push the boundaries of superintelligence with agentic AI. By automating complex, multi-step tasks typically performed by humans, AI agents will help unlock full human potential. H is hiring the world's best AI talent, seeking those who are dedicated as much to building safely and responsibly as to advancing disruptive agentic capabilities. We promote a mindset of openness, learning, and collaboration, where everyone has something to contribute.

About the Team: The Inference team builds and operates the systems that serve H's foundational models in production. We focus on multimodal inference and serving for Computer Use Agents, optimizing across both the inference engine layer (e.g., vLLM, SGLang) and the model serving layer (e.g., disaggregated inference, intelligent routing). Agentic inference brings constraints around context length, multimodality, and tool calls, which we address by co-designing with the Models team on training-time choices and with the agent teams on how models are deployed. We operate at the intersection of research and production, translating cutting-edge inference techniques into the systems that power H's next generation of agents. We are looking for strong engineers excited about inference to join the team and help shape the systems behind superintelligent AI.

Key Responsibilities:

Build and operate the inference stack that serves H's multimodal agentic models

Improve latency, throughput, and cost of model serving across the stack

Research and implement inference techniques tailored to agent workloads

Co-design with the Models team on training-time decisions that affect inference

Collaborate with cross-functional teams to integrate inference into agentic AI products

Evaluate inference, serving, and hardware platforms, and communicate findings to stakeholders

Stay current with advancements in inference, model serving, and accelerator technology

Requirements:

Technical skills:

Strong software engineering track record

Proficient in Python and at least one systems language (Rust, C++, or Go)

Hands-on experience with deep learning frameworks (PyTorch, JAX), preferably in an industry setting

Solid distributed systems fundamentals

Experience working in a modern cloud environment and with production ML infrastructure (Kubernetes, etc.)

Working knowledge of modern ML, including transformers and multimodal architectures

Research skills:

Research engagement: an advanced degree with research output, or publications at top-tier AI or systems venues (e.g., NeurIPS, ICML, MLSys, OSDI), research internships, or substantive open-source contributions

Soft skills:

Excellent communication and presentation skills

Strong collaboration and teamwork skills

Requirements

  • ·Technical skills:
  • ·Strong software engineering track record
  • ·Proficient in Python and at least one systems language (Rust, C++, or Go)
  • ·Hands-on experience with deep learning frameworks (PyTorch, JAX), preferably in an industry setting
  • ·Solid distributed systems fundamentals

Benefits

  • ·Join the exciting journey of shaping the future of AI
  • ·Collaborate with a fun, dynamic and multicultural team, working alongside world-class AI talent in a highly collaborative environment
  • ·Enjoy a competitive salary
  • ·Unlock opportunities for professional growth, continuous learning, and career development

How to apply

  1. 1Check the flexibility label above, hybrid, matches where you plan to live and work.
  2. 2Tailor your CV to the role at H Company, mentioning your remote working experience.
  3. 3Apply directly on the employer's careers page using the button below. Nomaders never handles your application.

Found 6d ago. Last checked today. Always confirm the details on the original posting, salary and location can change after publication.

Listing sourced from Company boards.

Similar roles

Other open software roles with comparable remote rules.

Browse all open roles

Free to apply, no account needed.

Typically $140k to $245k per year · You'll be taken to the employer's careers page.