Jobs / United States / OpenAI UK Ltd

Software Engineer, Data Infrastructure - Research

OpenAI UK Ltd · 🇺🇸 San Francisco

Sponsorship verdict

No sponsorship evidence yet

No government record and no wording either way. Not a refusal — ask the recruiter.

  • No government sponsor record hereThis employer does not appear with approvals in the government data SponsorApply holds.
  • The posting doesn’t mention sponsorshipSilence isn’t a refusal — ask the recruiter before investing much time.
  • No salary bar for this routeH-1B has no fixed salary bar: the employer must pay at least the prevailing wage for the role and area. Cap-subject employers enter a lottery weighted by wage level. Source: https://www.federalregister.gov/documents/2025/12/29/2025-23853/weighted-selection-process-for-registrants-and-petitioners-seeking-to-file-cap-subject-h-1b, rules effective 2026-02-27.
  • Confirmed live todayWhen a source last listed this job as open.

US H-1B: cap-subject employers enter a lottery weighted by wage level — Level I gets 1 entry, Level IV gets 4 (DHS projected selection odds ≈15% at Level I to ≈61% at Level IV). Universities and non-profit research employers are cap-exempt. The $100,000 fee for new petitions from abroad is currently blocked by a court order (appeal pending).

A verdict summarises public evidence; it is not legal advice and never a guarantee — the employer and the immigration authority decide. Sign in to factor in where you can already work.

Start free →

Or apply yourself on the official page →

Why not apply?

Sponsorship unknown

No register record and no sponsorship wording in the posting. Worth asking the employer before investing significant time.

SponsorApply flags time-wasters so your applications go where they can land. These come from the posting's own wording — read the original listing to confirm. See better-fit alternatives →

Sponsor Radar — OpenAI UK Ltd

No sponsorship history found

This employer does not appear with approvals in the government data SponsorApply holds.

Past sponsorship or register membership never guarantees sponsorship for this vacancy or for you. Full Sponsor Radar for OpenAI UK Ltd →

About the role

ABOUT THE TEAM The Workload team is responsible for designing and running OpenAI’s LLM training and inference infrastructure that powers frontier models at massive scale. Our systems unify how researchers train and serve models, abstracting away the complexity of performance, parallelism, and execution across vast GPU/accelerator fleets. By providing this foundation, the Workload team ensures that researchers can focus on advancing model capabilities while we handle the scale, efficiency, and reliability required to bring those models to life. ABOUT THE ROLE We are looking for an engineer to design and implement the dataset infrastructure that powers OpenAI’s next-generation training stack. You will be responsible for building standardized dataset interfaces, scaling pipelines across thousands of GPUs, and proactively testing performance bottlenecks. In this role, you will collaborate closely with the multimodal researchers, and other infra groups to ensure datasets are unified, efficient, and easy to consume. IN THIS ROLE, YOU WILL: - Design and maintain standardized dataset APIs, including for multimodal (MM) data that cannot fit in memory. - Build proactive testing and scale validation pipelines for dataset loading at GPU scale. - Collaborate with teammates to integrate datasets seamlessly into training and inference pipelines, ensuring smooth adoption and a great user experience. - Document and maintain dataset interfaces so they are discoverable, consistent, and easy for other teams to adopt. - Establish safeguards and validation systems to ensure datasets remain reproducible and unchanged once standardized. - Debug and resolve performance bottlenecks in distributed dataset loading (e.g., straggler systems slowing global training). - Provide visualization and inspection tools to surface errors, bugs, or bottlenecks in datasets. YOU MIGHT THRIVE IN THIS ROLE IF YOU: - Have strong engineering fundamentals with experience in distributed systems, data pipelines, or infrastructure. - Have experience building APIs, modular code, and scalable abstractions, while recognizing that abstractions ultimately serve the users and UX is an important part of the abstractions design. - Are comfortable debugging bottlenecks across large fleets of machines. - Take pride in building infrastructure that “just works,” and find joy in being the guardian of reliability and scale. - Are collaborative, humble, and excited to own a foundational (if not glamorous) part of the ML stack. Bonus points if you: - Have background knowledge in data math, probability, or distributed data theory. - Have worked with GPU-scale distributed systems or dataset scaling for real-time data About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.  We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement https://cdn.openai.com/policies/eeo-policy-statement.pdf. Background checks for applicants will be administered in accordance with applicable law, an

View the official posting →

Source: Ashby (employer board) First seen: 2025-09-18 Last confirmed: 2026-10-03 How our data works → Report this job

Similar opportunities