← Back to jobs

Industrial Compute

openai

Remoto US - Remote
Uncategorized

Job Score

90 pts
Remote model (+90)

About the Team

OpenAI’s Compute organization turns ambitious AI research into real-world capability by delivering the compute infrastructure behind our most advanced models. The team works across software, hardware, facilities, operations, and engineering disciplines to make enormous amounts of compute available, reliable, and efficient.

As the demand for frontier AI grows, so does the complexity of the systems required to support it. Scaling this infrastructure means solving problems that cut across distributed systems, ML infrastructure, GPU fleets, power, cooling, networking, manufacturing, supply chain, and data center delivery.

Our work is focused on expanding the compute foundation that enables OpenAI to train more capable models, including systems like GPT-5.6, and make frontier AI available to more people, products, and workflows. We’re looking for exceptional people across many disciplines to help build the next generation of AI infrastructure at a scale few organizations have attempted.

About the Role

We are hiring across a broad range of roles to help design, build, scale, and operate OpenAI’s compute infrastructure. Depending on your background, you may work on large-scale distributed systems, ML infrastructure, hardware systems, manufacturing, supply chain, data center development, or the physical engineering systems required to bring massive compute capacity online.

You’ll work with teams across research, engineering, hardware, operations, and infrastructure to solve high-impact problems at extraordinary scale. This may include improving system reliability, accelerating deployment timelines, increasing operational efficiency, designing new infrastructure, or helping bring new compute platforms and facilities from concept to production.

This is an opportunity to work on one of the most important infrastructure challenges in AI: building the compute foundation required to train and serve increasingly capable frontier models.

Key Responsibilities

  • Help build, scale, and operate OpenAI’s global compute infrastructure.

  • Solve complex problems across software, hardware, manufacturing supply chain, and data center systems.

  • Improve the reliability, performance, efficiency, and scalability of critical infrastructure.

  • Partner with cross-functional teams to bring new compute capacity online quickly and reliably.

  • Identify bottlenecks across technical, operational, and physical systems, and develop practical solutions.

  • Build tools, processes, systems, or infrastructure that improve execution at scale.

  • Contribute to the long-term architecture and operational maturity of OpenAI’s compute footprint.

Qualifications

  • Have experience building, scaling, or operating complex technical systems.

  • Enjoy working on ambiguous, high-impact problems where the path forward is not always defined.

  • Are comfortable collaborating across disciplines, including software, hardware, operations, and physical infrastructure.

  • Have strong technical judgment and a bias toward execution.

  • Care deeply about reliability, speed, safety, and operational excellence.

  • Are excited by the challenge of building infrastructure at unprecedented scale.

  • Want your work to directly support the development and deployment of frontier AI.

Preferred Skills

  • Have experience with AI infrastructure, high-performance computing, distributed systems, GPU clusters, or cloud-scale platforms.

  • Have worked on hardware systems, manufacturing, supply chain, data center development, or large capital infrastructure projects.

  • Have domain expertise in civil, controls, mechanical, hardware, electrical, thermal, power, networking, or facilities engineering.

  • Have helped bring new technical platforms, data centers, factories, or large-scale systems from concept to production.

  • Have experience operating in fast-moving environments where technical depth and execution speed both matter.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. 

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.

OpenAI Global Applicant Privacy Policy

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

What did you think of this job?

Comments 0

Want to leave a comment?
Sign in or create your account in seconds to join the discussion.
Loading comments...

Discover Other Areas

Understand the scope of work, key skills, and tools used in different career areas.

About Web Master

The Web Master is the professional responsible for maintaining, securing, and ensuring the technical performance of websites and web applications. They manage servers, hosting infrastructure, uptime monitoring, and ensure everything runs fast and reliably.

Key skills include server management (Apache, Nginx), hosting (AWS, Google Cloud, Azure), CDN (Cloudflare), SSL, DNS, web security (WAF, firewall), performance (Core Web Vitals, cache, compression), and versioning (Git, CI/CD). Knowledge of Docker, WordPress, cPanel, and monitoring (Sentry, New Relic) is a differentiator.

Web Masters in technology companies are highly valued, especially those who master DevOps, SRE, and can guarantee uptime and performance at scale. The field offers opportunities from junior webmaster to SRE and infrastructure engineer, with a focus on reliability, security, and speed.

About Technical Support

Technical Support is essential to ensure customer satisfaction and retention. Support professionals resolve technical issues, document solutions, and identify patterns that can lead to product improvements.

Key skills include troubleshooting, customer service, technical documentation, ITIL knowledge, and ticketing tools (Zendesk, Freshdesk, Intercom).

Technical support has evolved from a reactive to a proactive function, with high-level professionals working in Customer Engineering and Support Engineering.

About QA and Testing

QA and Software Testing are fundamental to ensure the quality and reliability of applications. QA professionals ensure that the delivered product meets requirements and is free of critical defects.

Key skills include manual and automated testing, Selenium, Cypress, Playwright, Postman, JMeter, and CI/CD pipeline knowledge. Performance and security testing are differentiators.

With the adoption of DevOps and continuous deployment, the demand for automation QAs and SDETs continues to grow.

About Infrastructure and DevOps

Infrastructure and DevOps are responsible for creating, maintaining, and optimizing IT environments that support applications at scale. This area is fundamental for system reliability and performance.

Key technologies include AWS, GCP, Azure, Docker, Kubernetes, Terraform, Ansible, CI/CD (GitHub Actions, GitLab CI, Jenkins), and monitoring (Datadog, Grafana, Prometheus).

DevOps engineers and SREs are highly sought-after professionals, with salaries among the highest in the technology sector.

About Design

The Design field, especially UX/UI and Product Design, has experienced significant growth in recent years. With accelerated business digitization, the demand for professionals who can create intuitive and pleasant digital experiences has never been higher.

Key skills include Figma, Sketch, Adobe XD, user research, design thinking, prototyping, and system design. Product designers are increasingly valued for their direct impact on business results.

Remote work has opened doors for Brazilian designers to work for global companies, with competitive salaries in dollars and euros.

Career Guides

Technology Career Guide

Planning, skills, interviews, and professional growth in IT, Data Science, DevOps, and Product.

Read full guide →

Design Career Guide

UX/UI, Graphic Design, Product Design. Portfolio, tools, interviews, and growth in the Design field.

Read full guide →

Marketing Career Guide

SEO, Paid Media, Growth, Content Marketing. Certifications, tools, and strategies to grow in Digital Marketing.

Read full guide →

Finance Career Guide

Financial market, investments, corporate finance, certifications, and strategies to grow in the financial field.

Read full guide →

Communication Career Guide

Journalism, PR, Corporate Communication, Content Marketing, and Multimedia Production.

Read full guide →

Administration Career Guide

Business Management, HR, Logistics, Consulting, Project Management, and Entrepreneurship.

Read full guide →

Data Career Guide

Data Science, Data Engineering, BI, Machine Learning, and AI. From training to the job market.

Read full guide →

Product Career Guide

Product Management, Product Ownership, Agile, Scrum, and OKRs. From strategy to execution.

Read full guide →

Tech & Remote Glossary

Stop getting lost in interviews and job descriptions

The job market, especially within tech and global companies, has developed its own dialect. Not understanding these acronyms can make you lose valuable opportunities or poorly negotiate your contract. To end this problem, we created the Definitive Glossary for the Remote Professional.

🏢 Work Models & Routine

Async (Asynchronous Work)
A communication model where responses don't need to be immediate. Instead of back-to-back meetings, the team relies on well-structured documents, threads, and messages. It's the gold standard for global companies spanning multiple time zones.
Sync (Synchronous Work)
The opposite of Async. It requires the team to be online and available at the same time for meetings, live chats, and real-time collaboration.
Daily / Stand-up
A quick daily meeting (usually 15 minutes) common in Agile (Scrum) methodologies. The team answers three questions: What did I do yesterday? What will I do today? Are there any blockers?
All-Hands / Town Hall
A company-wide meeting involving all employees. Usually led by the founders (C-Levels) to present results, new goals, and answer team questions.
1:1 (One-on-One)
A recurring individual meeting between a professional and their direct manager. It is used for career alignment, feedback, and problem-solving, not just for project status updates.

💰 Contracts, Benefits & Compensation

PTO (Paid Time Off)
Instead of strict, categorized leave policies, modern US companies usually offer a flexible pool of days (e.g., 20 days of PTO, or even "Unlimited PTO") that you can use for vacations, sick days, or personal matters, while receiving your regular compensation.
Equity / Stock Options
Company ownership. The startup offers you the right to buy shares at a heavily discounted strike price in the future. If the company grows, goes public, or is acquired, these shares can be highly lucrative.
Vesting (Vesting Schedule)
The rule that controls your Equity. It usually lasts 4 years. You don't get all the shares on day one; you "earn" them gradually as you stay with the company. A "1-year Cliff" means you must stay for at least one year to receive your first batch of shares.
RSUs (Restricted Stock Units)
Unlike Stock Options (where you have the right to buy the stock), RSUs are actual shares the company grants you as a bonus or part of your compensation package, following a strict Vesting schedule.
Independent Contractor (1099 / B2B)
The most common international hiring model for global talent working for US companies. You act as a service provider (business-to-business), receiving the gross salary (often six-figure compensation) without standard local payroll tax deductions at the source.

🤖 Recruitment & Hiring Process

ATS (Applicant Tracking System)
The "robot" that reads your resume. Software like Greenhouse, Ashby, and Workday are used to filter candidates by keywords before a human even looks at the document. (Pro tip: this is why your resume must be clean, semantic, and have the right keywords).
JD (Job Description)
The document that lists the responsibilities, technical requirements, and benefits of the open position.
Cultural Fit
The interview stage that evaluates if your core values, communication style, and worldview align with the company's culture. This is the ultimate test of your Soft Skills.
Onboarding
The integration process. It's the period of your first few weeks at the company, where you get your access credentials, learn about the culture, study the internal documentation, and understand how the product works.

🚀 How to use this to your advantage?

The secret isn't just knowing what these acronyms mean, but using them actively. If during an interview for a premium tech role you ask, "How does your PTO policy and Vesting schedule work?", the recruiter will immediately perceive you as a high-level professional, familiar with the global market standards.

The remote job market requires preparation. And having the right vocabulary is the first big step to securing six-figure proposals and standing out among thousands of applicants.

Expert Tip

The Back-End Development Market

The Back-End Development Market: Barriers, Opportunities, and the Path to the Top

Behind every brilliant application, revolutionary artificial intelligence, or successful fintech, there is an invisible and robust ecosystem. Welcome to the Back-End universe.

The Modern Back-End Paradox: Did AI Steal the Jobs?

With the rise of tools like GitHub Copilot and Cursor, many junior developers wonder if the Back-End career is threatened. The short answer is: no. In fact, it has evolved.

Artificial Intelligence has made writing basic "CRUD" (Create, Read, Update, Delete) code trivial. However, the market no longer pays six-figure salaries for writing repetitive code. The global market is actively hunting for Software Engineers—professionals who understand architecture, resilience, latency, and scalability. The Back-End didn't die; the bar was simply raised.

Barriers to Entry: What Separates Juniors from Seniors

Entering Back-End development today requires overcoming technical barriers that go far beyond mastering a programming language (like Java, C#, Go, or Python). Key barriers include:

  • System Design: Knowing how to design a system that supports 100 users is easy. Designing one that handles 1 million requests per second requires deep knowledge of load balancing, caching (Redis/Memcached), and message queues (RabbitMQ/Kafka).
  • Data Complexity: The debate is no longer just "SQL vs. NoSQL". It is about data modeling, replication, sharding, and how to avoid database bottlenecks in distributed systems.
  • Security: With data breaches costing millions, companies require developers to master security practices from day one. Not knowing the vulnerabilities listed by the OWASP Top 10 is a dealbreaker for premium remote roles.
  • DevOps and Cloud Culture: The modern Back-End developer must understand containerization (Docker), orchestration (Kubernetes), and cloud infrastructure (AWS, GCP, Azure).

Golden Opportunities: Where is the Money?

For those who overcome these barriers, the market is a blue ocean of opportunities, especially for US/Global Remote work.

  • Migration to Microservices and Serverless: Corporate giants continue to dismantle legacy monoliths. Professionals who understand the patterns described by Martin Fowler are highly sought after.
  • High-Performance Languages: While traditional languages maintain their corporate strength, the use of Go (Golang) and Rust has skyrocketed for systems requiring massive concurrency and low memory footprint (Green Computing).
  • AI Infrastructure: AI models don't run in a vacuum. There is a massive demand for Back-End engineers proficient in Python and C++ to build data pipelines (MLOps) and the APIs that serve these models in real time.

Success Stories: Architectural Decisions That Changed the Game

True Back-End engineering shines when solving impossible problems. Let's look at real-market examples:

The Discord Case (Migration to Rust): Discord faced latency spikes in its core Read States service, originally written in Go. Because Go's Garbage Collector caused critical millisecond freezes, the team rewrote the service in Rust, completely eliminating latency spikes and supporting trillions of messages with absurd efficiency. This proved the value of choosing the right tool for performance limits.

The Netflix Case (Pioneering Microservices): Netflix transformed a monolithic system that broke under pressure into an architecture of thousands of independently managed microservices. They pioneered the concept of Chaos Engineering, purposely shutting down production servers to ensure their Back-End was resilient to failure.

Practical Tips: How to Land Premium Global Jobs

  1. Master the Fundamentals: Before learning the trendy framework of the month, study data structures, algorithms, and time complexity (Big-O Notation). This is exactly what will be tested in high-level technical interviews (Whiteboard interviews).
  2. Build a Problem-Focused Portfolio: A GitHub repository with a "To-Do List" won't impress anyone. Build an API that handles asynchronous image processing, create a scalable URL shortener, or engineer a messaging system using WebSockets and Redis.
  3. Study Real Cloud Architecture: Get solutions-based certifications (like AWS Certified Developer or Solutions Architect). These serve as a "Seal of Approval" to bypass strict Applicant Tracking Systems (ATS) like Greenhouse or Ashby.
  4. Flawless Technical Communication: According to the Stack Overflow Developer Survey, the highest-paying roles require asynchronous global collaboration. Your code documentation, commit messages, and PR reviews must be pristine and professional.

Verdict: Is a Career in Back-End Worth It?

Absolutely. If you are an analytical person who loves solving complex puzzles and cares about the security and efficiency of things no one sees, the Back-End is your place.

It is a career virtually immune to visual fads. While Front-End libraries change every few years, the fundamentals of relational databases, networks, and operating systems remain the same. It is a rock-solid foundation for a highly lucrative and globalized career.