← Back to jobs

Operating Systems Engineer, On-Device Inference | Consumer Devices

openai

San Francisco
Uncategorized Systems Analyst

Job Score

80 pts
On-site model (+70) Systems Analyst (+10)

About the Team

OpenAI Consumer Devices is building the next generation of products that bring powerful AI into people’s everyday lives. Guided by OpenAI’s mission to ensure AGI benefits all of humanity, our team combines world-class researchers, engineers, designers, and operators who care deeply about creating useful, intuitive, and responsible technology. You’ll have the opportunity to work alongside exceptional people on ambitious, zero-to-one challenges at the intersection of hardware, software, and AI. This is a chance to help define an entirely new category of products—and shape how people experience AI in the future.

Our team works across silicon, embedded systems, operating systems, and cloud services to build reliable consumer devices and the novel platforms required to support them. We partner closely with research to bring advanced AI capabilities into the physical world.

About the Role

As an Operating Systems Engineer focused on on-device inference, you will design, develop, and ship the OS stack that makes advanced AI capabilities reliable, responsive, and energy efficient on consumer devices.

Your work will span OS services and frameworks, inference runtime integration, model fitting, scheduling, and performance and power management. You’ll partner with research to adapt models to device constraints, make design decisions across the stack, and carry solutions from early exploration through integration and production.

In this role, you will:

  • Build the inference platform: Design and implement maintainable OS services, frameworks, and clear interfaces for inference execution, model loading and lifecycle, and resource management.

  • Fit models to device constraints: Partner with researchers on quantization, runtime integration, and memory optimization to meet memory, compute, and energy budgets while evaluating model quality and product behavior.

  • Coordinate system resources: Develop scheduling and resource policies that balance inference with other device activity, preserving responsiveness within latency, memory, battery, and thermal constraints.

  • Advance performance and power management: Develop and validate execution strategies that adapt to workload needs, available resources, and changing device conditions.

  • Debug across the stack: Use tracing, profiling, and structured debugging to investigate correctness, concurrency, performance, and reliability issues across models, inference runtimes, and OS components.

  • Measure and validate improvements: Build diagnostic tools, instrumentation, representative device workloads, and automated tests to demonstrate repeatable performance and energy gains on physical devices, catch regressions, and validate sustained use.

  • Bring capabilities to production: Collaborate with research, hardware, firmware, platform, and product engineering teams to turn emerging model capabilities into maintainable systems and reliable shipped features.

Minimum Qualifications

  • Substantial hands-on experience designing, developing, and debugging operating system components, system services, or performance-critical platform software.

  • Proficiency in C++ for systems development, including concurrent programming, memory ownership, and resource lifetime management.

  • Hands-on experience integrating or optimizing inference runtimes or machine learning workloads in resource-constrained environments.

  • Strong understanding of scheduling, memory management, and how workloads compete for shared system resources.

  • Experience diagnosing complex system behavior and delivering performance or power improvements supported by repeatable measurements.

  • Ability to work across disciplines, translate research and product needs into system requirements, and explain technical decisions and tradeoffs clearly.

You might thrive in this role if you:

  • Have delivered end-to-end on-device inference in shipped products, from model adaptation and runtime integration through OS support and production debugging.

  • Have adapted models for deployment through quantization, compression, or related techniques, balancing quality, execution cost, and memory requirements.

  • Have optimized inference across CPUs, GPUs, or neural accelerators, accounting for data movement, synchronization, and execution placement.

  • Are proficient in Rust for systems programming.

  • Take ownership of complete solutions and follow problems across model, runtime, and operating system boundaries.

  • Challenge assumptions, develop new approaches when existing techniques fall short, and use experiments and measurements to guide decisions.

  • Work constructively across disciplines and make careful tradeoffs among model quality, performance, power, reliability, security, and maintainability.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. 

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.

OpenAI Global Applicant Privacy Policy

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

What did you think of this job?

Comments 0

Want to leave a comment?
Sign in or create your account in seconds to join the discussion.
Loading comments...

About Systems Analyst

The Systems Analyst is the professional responsible for analyzing, designing, and implementing technology solutions that meet business needs. They act as a bridge between business areas and the development team, ensuring that systems deliver real value to the organization.

Key skills include requirements gathering and analysis, process modeling (BPMN), data modeling, technical and functional documentation, system integration (APIs, microservices), and knowledge of ERPs and CRMs. Tools like Jira, Confluence, Visio, and project management platforms are essential.

Systems Analysts in technology companies are highly valued, especially those who master agile requirements analysis (user stories, backlog), system integration, and solution architecture. The field offers opportunities from junior analyst to solution architect, with a focus on efficiency, quality, and technological innovation.

Discover Other Areas

Understand the scope of work, key skills, and tools used in different career areas.

About Backend

The Backend area is responsible for all server logic, APIs, databases, and infrastructure that support web and mobile applications. Backend professionals ensure that systems are scalable, secure, and performant.

Key skills include languages like PHP, Java, Python, Ruby, Go, and Node.js, frameworks like Laravel, Spring Boot, Django, and Express, databases (MySQL, PostgreSQL, MongoDB, Redis), software architecture (clean architecture, DDD, microservices), and API security (OAuth, JWT).

Backend developers in technology companies are highly valued, especially those who master microservices architecture, cloud computing, and high-scale performance. The field offers opportunities from junior developer to software architect, with a focus on scalability, security, and efficiency.

About Project Manager

The Project Manager is the professional responsible for planning, executing, and controlling projects end-to-end, ensuring they are delivered on time, within budget, and with the expected quality. With the growing complexity of businesses, project management professionals are fundamental to organizational success.

Key skills include planning and scheduling, scope, cost, risk, quality, and resource management, stakeholder communication, cross-functional team leadership, and use of agile and traditional methodologies. Certifications like PMP, PRINCE2, and Six Sigma are important differentiators.

Project Managers in technology companies are highly valued, especially those who master agile methodologies (Scrum, Kanban), tools like Jira and MS Project, and can deliver complex projects efficiently. The field offers opportunities from project analyst to head of PMO, with a focus on execution, governance, and business value.

About Web3

The Web3 area represents the new phase of the decentralized internet, built on blockchain technology. Web3 professionals create decentralized applications (dApps), interact with smart contracts, manage digital assets (cryptocurrencies and NFTs), and utilize DeFi (Decentralized Finance) protocols, revolutionizing how data, ownership, and finance are managed online.

About Traffic Manager

The Traffic Manager is the professional responsible for planning, executing, and optimizing paid media campaigns across various digital platforms. With the competitiveness of the digital market, paid traffic professionals are essential for generating qualified leads and maximizing return on advertising investment.

Key skills include campaign management on Google Ads, Meta Ads, LinkedIn Ads, and TikTok Ads, media planning, metrics analysis (ROAS, CPA, CPC, CTR), A/B testing, remarketing, and landing page creation. Tools like Google Analytics, Google Tag Manager, Hotjar, and automation platforms are essential.

Traffic managers in technology companies are highly valued, especially those who master performance marketing, conversion funnel optimization, and scaling strategies. The field offers opportunities from media analyst to head of performance, with a focus on growth, budget efficiency, and return on investment.

About Ecommerce Manager

The Ecommerce Manager is the professional responsible for the entire strategic and operational management of online stores and marketplaces. They lead teams, define pricing, promotion, and catalog strategies, and monitor online sales performance across multiple platforms.

Key skills include catalog management, dynamic pricing, seasonal campaigns (Black Friday, Cyber Monday), marketplace management (Amazon, Mercado Livre, Shopee, Magalu), paid traffic, CRO, and team management. Knowledge of Shopify, VTEX, WooCommerce, Google Ads, Meta Ads, and performance metrics is a differentiator.

Ecommerce Managers in technology companies are highly valued, especially those who master multi-marketplace management, checkout optimization, and mobile commerce strategies. The field offers opportunities from ecommerce manager to head of ecommerce, with a focus on revenue, customer experience, and growth.

Career Guides

Technology Career Guide

Planning, skills, interviews, and professional growth in IT, Data Science, DevOps, and Product.

Read full guide →

Design Career Guide

UX/UI, Graphic Design, Product Design. Portfolio, tools, interviews, and growth in the Design field.

Read full guide →

Marketing Career Guide

SEO, Paid Media, Growth, Content Marketing. Certifications, tools, and strategies to grow in Digital Marketing.

Read full guide →

Finance Career Guide

Financial market, investments, corporate finance, certifications, and strategies to grow in the financial field.

Read full guide →

Communication Career Guide

Journalism, PR, Corporate Communication, Content Marketing, and Multimedia Production.

Read full guide →

Administration Career Guide

Business Management, HR, Logistics, Consulting, Project Management, and Entrepreneurship.

Read full guide →

Data Career Guide

Data Science, Data Engineering, BI, Machine Learning, and AI. From training to the job market.

Read full guide →

Product Career Guide

Product Management, Product Ownership, Agile, Scrum, and OKRs. From strategy to execution.

Read full guide →

Tech & Remote Glossary

Stop getting lost in interviews and job descriptions

The job market, especially within tech and global companies, has developed its own dialect. Not understanding these acronyms can make you lose valuable opportunities or poorly negotiate your contract. To end this problem, we created the Definitive Glossary for the Remote Professional.

🏢 Work Models & Routine

Async (Asynchronous Work)
A communication model where responses don't need to be immediate. Instead of back-to-back meetings, the team relies on well-structured documents, threads, and messages. It's the gold standard for global companies spanning multiple time zones.
Sync (Synchronous Work)
The opposite of Async. It requires the team to be online and available at the same time for meetings, live chats, and real-time collaboration.
Daily / Stand-up
A quick daily meeting (usually 15 minutes) common in Agile (Scrum) methodologies. The team answers three questions: What did I do yesterday? What will I do today? Are there any blockers?
All-Hands / Town Hall
A company-wide meeting involving all employees. Usually led by the founders (C-Levels) to present results, new goals, and answer team questions.
1:1 (One-on-One)
A recurring individual meeting between a professional and their direct manager. It is used for career alignment, feedback, and problem-solving, not just for project status updates.

💰 Contracts, Benefits & Compensation

PTO (Paid Time Off)
Instead of strict, categorized leave policies, modern US companies usually offer a flexible pool of days (e.g., 20 days of PTO, or even "Unlimited PTO") that you can use for vacations, sick days, or personal matters, while receiving your regular compensation.
Equity / Stock Options
Company ownership. The startup offers you the right to buy shares at a heavily discounted strike price in the future. If the company grows, goes public, or is acquired, these shares can be highly lucrative.
Vesting (Vesting Schedule)
The rule that controls your Equity. It usually lasts 4 years. You don't get all the shares on day one; you "earn" them gradually as you stay with the company. A "1-year Cliff" means you must stay for at least one year to receive your first batch of shares.
RSUs (Restricted Stock Units)
Unlike Stock Options (where you have the right to buy the stock), RSUs are actual shares the company grants you as a bonus or part of your compensation package, following a strict Vesting schedule.
Independent Contractor (1099 / B2B)
The most common international hiring model for global talent working for US companies. You act as a service provider (business-to-business), receiving the gross salary (often six-figure compensation) without standard local payroll tax deductions at the source.

🤖 Recruitment & Hiring Process

ATS (Applicant Tracking System)
The "robot" that reads your resume. Software like Greenhouse, Ashby, and Workday are used to filter candidates by keywords before a human even looks at the document. (Pro tip: this is why your resume must be clean, semantic, and have the right keywords).
JD (Job Description)
The document that lists the responsibilities, technical requirements, and benefits of the open position.
Cultural Fit
The interview stage that evaluates if your core values, communication style, and worldview align with the company's culture. This is the ultimate test of your Soft Skills.
Onboarding
The integration process. It's the period of your first few weeks at the company, where you get your access credentials, learn about the culture, study the internal documentation, and understand how the product works.

🚀 How to use this to your advantage?

The secret isn't just knowing what these acronyms mean, but using them actively. If during an interview for a premium tech role you ask, "How does your PTO policy and Vesting schedule work?", the recruiter will immediately perceive you as a high-level professional, familiar with the global market standards.

The remote job market requires preparation. And having the right vocabulary is the first big step to securing six-figure proposals and standing out among thousands of applicants.

Expert Tip

The Back-End Development Market

The Back-End Development Market: Barriers, Opportunities, and the Path to the Top

Behind every brilliant application, revolutionary artificial intelligence, or successful fintech, there is an invisible and robust ecosystem. Welcome to the Back-End universe.

The Modern Back-End Paradox: Did AI Steal the Jobs?

With the rise of tools like GitHub Copilot and Cursor, many junior developers wonder if the Back-End career is threatened. The short answer is: no. In fact, it has evolved.

Artificial Intelligence has made writing basic "CRUD" (Create, Read, Update, Delete) code trivial. However, the market no longer pays six-figure salaries for writing repetitive code. The global market is actively hunting for Software Engineers—professionals who understand architecture, resilience, latency, and scalability. The Back-End didn't die; the bar was simply raised.

Barriers to Entry: What Separates Juniors from Seniors

Entering Back-End development today requires overcoming technical barriers that go far beyond mastering a programming language (like Java, C#, Go, or Python). Key barriers include:

  • System Design: Knowing how to design a system that supports 100 users is easy. Designing one that handles 1 million requests per second requires deep knowledge of load balancing, caching (Redis/Memcached), and message queues (RabbitMQ/Kafka).
  • Data Complexity: The debate is no longer just "SQL vs. NoSQL". It is about data modeling, replication, sharding, and how to avoid database bottlenecks in distributed systems.
  • Security: With data breaches costing millions, companies require developers to master security practices from day one. Not knowing the vulnerabilities listed by the OWASP Top 10 is a dealbreaker for premium remote roles.
  • DevOps and Cloud Culture: The modern Back-End developer must understand containerization (Docker), orchestration (Kubernetes), and cloud infrastructure (AWS, GCP, Azure).

Golden Opportunities: Where is the Money?

For those who overcome these barriers, the market is a blue ocean of opportunities, especially for US/Global Remote work.

  • Migration to Microservices and Serverless: Corporate giants continue to dismantle legacy monoliths. Professionals who understand the patterns described by Martin Fowler are highly sought after.
  • High-Performance Languages: While traditional languages maintain their corporate strength, the use of Go (Golang) and Rust has skyrocketed for systems requiring massive concurrency and low memory footprint (Green Computing).
  • AI Infrastructure: AI models don't run in a vacuum. There is a massive demand for Back-End engineers proficient in Python and C++ to build data pipelines (MLOps) and the APIs that serve these models in real time.

Success Stories: Architectural Decisions That Changed the Game

True Back-End engineering shines when solving impossible problems. Let's look at real-market examples:

The Discord Case (Migration to Rust): Discord faced latency spikes in its core Read States service, originally written in Go. Because Go's Garbage Collector caused critical millisecond freezes, the team rewrote the service in Rust, completely eliminating latency spikes and supporting trillions of messages with absurd efficiency. This proved the value of choosing the right tool for performance limits.

The Netflix Case (Pioneering Microservices): Netflix transformed a monolithic system that broke under pressure into an architecture of thousands of independently managed microservices. They pioneered the concept of Chaos Engineering, purposely shutting down production servers to ensure their Back-End was resilient to failure.

Practical Tips: How to Land Premium Global Jobs

  1. Master the Fundamentals: Before learning the trendy framework of the month, study data structures, algorithms, and time complexity (Big-O Notation). This is exactly what will be tested in high-level technical interviews (Whiteboard interviews).
  2. Build a Problem-Focused Portfolio: A GitHub repository with a "To-Do List" won't impress anyone. Build an API that handles asynchronous image processing, create a scalable URL shortener, or engineer a messaging system using WebSockets and Redis.
  3. Study Real Cloud Architecture: Get solutions-based certifications (like AWS Certified Developer or Solutions Architect). These serve as a "Seal of Approval" to bypass strict Applicant Tracking Systems (ATS) like Greenhouse or Ashby.
  4. Flawless Technical Communication: According to the Stack Overflow Developer Survey, the highest-paying roles require asynchronous global collaboration. Your code documentation, commit messages, and PR reviews must be pristine and professional.

Verdict: Is a Career in Back-End Worth It?

Absolutely. If you are an analytical person who loves solving complex puzzles and cares about the security and efficiency of things no one sees, the Back-End is your place.

It is a career virtually immune to visual fads. While Front-End libraries change every few years, the fundamentals of relational databases, networks, and operating systems remain the same. It is a rock-solid foundation for a highly lucrative and globalized career.