← Back to jobs

Rma & Repair Lead

etched

OnSite San Jose
Advertising

Job Score

80 pts
On-site model (+70) Advertising (+10)

About Etched

Etched is building hardware for frontier intelligence. We co-design chips, racks, software, and manufacturing to deliver best-in-class throughput and latency across both prefill and decode workloads. Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is redefining the infrastructure layer for the fastest growing industry in history.

Job Summary

Etched is seeking an RMA and Repair Lead to build and run the end-to-end returns and repair operation for our AI hardware — from the moment a customer reports a failure through diagnosis, repair, and supplier recovery. You’ll be responsible for managing the repair capability whether it’s in-house or through outsourced partners. This role is deeply technical and hands-on: you'll own the repair, debug, and failure analysis requirements for our products, translate our manufacturing processes into repeatable repair processes, and feed what you learn back into design and production.

If you thrive in fast-paced environments, enjoy solving ambiguous problems, and want to shape the repair ecosystem at a rapidly scaling startup, this role is for you.

Key responsibilities

  • Own the end-to-end RMA process — from customer return authorization through triage, repair or replacement, reverse logistics, and supplier recovery/warranty claims

  • Develop and manage RMA workflows, including return authorization, repair/replacement, and reverse logistics

  • Build and manage Etched's repair operations, whether in-house or outsourced, capacity planning, and day-to-day management of in-house depots and/or outsourced repair partners

  • Own the technical definition of repair: repair strategy by product and level (rack, server, module, component), debug flows, failure analysis requirements, test coverage, and repair-vs-scrap criteria

  • Translate production and manufacturing processes into repair processes — repair travelers, work instructions, tooling and fixtures, test stations, calibration, and operator training and certification

  • Build the data layer for repair: turnaround time, failure rates by part and failure mode, cost of service, and repeat-failure tracking — and close the loop with design, quality, and manufacturing to drive corrective actions and improve product reliability

  • Design and implement metrics and dashboards (turnaround time, failure rates, cost of service) to monitor and improve operations

  • Define SLAs and escalation paths, and act as the technical point of contact for escalated customer issues.

You may be a good fit if you have (Must-have qualifications)

  • 8+ years of experience in Repair operations, RMA, or hardware support operations, including ownership of end-to-end returns flow

  • Direct experience standing up a repair operation — either building an in-house depot or selecting and managing an outsourced repair partner [ especially from scratch in a fast-paced environment

  • Strong hands-on technical depth in hardware debug, board- and system-level troubleshooting, and failure analysis.

  • Experience converting NPI/manufacturing process documentation into repair processes, work instructions, and test flows

  • Strong understanding of failure analysis (FA), root cause, and corrective action processes

  • Excellent communication skills and ability to work closely with customers, suppliers, and internal teams

Strong candidates may also have experience with (Nice-to-have qualifications)

  • Experience supporting server, networking, or AI hardware deployments at rack scale

  • Familiarity with PLM/ERP systems for tracking returns, parts, repairs and warranty claims

  • Contract manufacturer or ODM relationship management, including repair statements of work and pricing

  • Data-driven service analytics — failure rate analysis, Pareto/FMEA, cost optimization, predictive maintenance

  • Global reverse logistics, customs, and cross-border repair flows

  • ISO standards, quality systems, and regulatory requirements for hardware returns and repairs

Benefits

  • Medical, dental, and vision packages with generous premium coverage

    • $500 per month credit for waiving medical benefits

  • Housing subsidy of $2,500 per month for those living within walking distance of the office

  • Relocation support for those moving to San Jose (Santana Row)

  • Various wellness benefits covering fitness, mental health, and more

  • Daily lunch and dinner in our office

  • Unlimited compute budget subject to ROI justification

How we’re different

Etched believes in the Bitter Lesson. We are the first inference-focused frontier AI system, betting early on transformer and transformer-like architectures and on increasing model sizes. Our addressable market is the entirety of inference, unlike many of our competitors.

We are a fully in-person team in San Jose (Santana Row), and greatly value engineering skills. We do not have boundaries between engineering and research, and we expect all of our technical staff to contribute to both and work across disciplines as needed.


What did you think of this job?

Comments 0

Want to leave a comment?
Sign in or create your account in seconds to join the discussion.
Loading comments...

About Advertising

The Advertising area is aimed at the planning, creation, and delivery of communication campaigns to promote brands, products, ideas, or services. Professionals in the sector work in advertising agencies or in-house marketing departments in creative fields (art direction, copywriting), strategic planning, account management, and media buying.

Discover Other Areas

Understand the scope of work, key skills, and tools used in different career areas.

About Public Relations

The Public Relations (PR) area focuses on managing the reputation, image, and communication of an organization with its various stakeholders (such as clients, investors, employees, media, and the community). PR professionals develop corporate communication strategies, manage media relations (press relations), organize institutional events, and work in image crisis prevention and management.

About Content Manager

The Content Manager is the professional responsible for leading the entire content strategy, production, and management of an organization. They define the editorial strategy, coordinate writing teams, and ensure content aligns with business goals and brand identity.

Key skills include content strategy, editorial planning, content audit, buyer persona, customer journey, content ops, content governance, performance metrics (ROI, engagement, organic traffic), and team management. Knowledge of WordPress, Contentful, Notion, and analytics tools is a differentiator.

Content Managers in technology companies are highly valued, especially those who can align content with conversion funnels, lead multidisciplinary teams, and use data to optimize editorial strategy. The field offers opportunities from content manager to head of content, with a focus on strategy, quality, and scale.

About QA and Testing

QA and Software Testing are fundamental to ensure the quality and reliability of applications. QA professionals ensure that the delivered product meets requirements and is free of critical defects.

Key skills include manual and automated testing, Selenium, Cypress, Playwright, Postman, JMeter, and CI/CD pipeline knowledge. Performance and security testing are differentiators.

With the adoption of DevOps and continuous deployment, the demand for automation QAs and SDETs continues to grow.

About UX Design

The User Experience (UX) Design area focuses on optimizing the overall user experience when interacting with a product or service. UX professionals conduct user research (UX Research), map journeys, create wireframes, perform usability tests, and define navigation flows to ensure the product is intuitive, useful, and meets users' real needs.

About Frontend

The Frontend area is responsible for creating the visual interfaces that users interact with on websites and web applications. Frontend professionals combine technical skills with design to deliver intuitive, responsive, and accessible digital experiences.

Key skills include HTML, CSS, JavaScript/TypeScript, frameworks like React, Angular, and Vue, build tools (Webpack, Vite), CSS (Tailwind, Sass), testing (Jest, Cypress), and knowledge of web performance and accessibility (WCAG). Familiarity with design systems and reusable components is a differentiator.

Frontend developers in technology companies are highly valued, especially those who master React, Next.js, web performance, and accessibility. The field offers opportunities from junior developer to frontend architect, with a focus on user experience, performance, and code quality.

Career Guides

Technology Career Guide

Planning, skills, interviews, and professional growth in IT, Data Science, DevOps, and Product.

Read full guide →

Design Career Guide

UX/UI, Graphic Design, Product Design. Portfolio, tools, interviews, and growth in the Design field.

Read full guide →

Marketing Career Guide

SEO, Paid Media, Growth, Content Marketing. Certifications, tools, and strategies to grow in Digital Marketing.

Read full guide →

Finance Career Guide

Financial market, investments, corporate finance, certifications, and strategies to grow in the financial field.

Read full guide →

Communication Career Guide

Journalism, PR, Corporate Communication, Content Marketing, and Multimedia Production.

Read full guide →

Administration Career Guide

Business Management, HR, Logistics, Consulting, Project Management, and Entrepreneurship.

Read full guide →

Data Career Guide

Data Science, Data Engineering, BI, Machine Learning, and AI. From training to the job market.

Read full guide →

Product Career Guide

Product Management, Product Ownership, Agile, Scrum, and OKRs. From strategy to execution.

Read full guide →

Tech & Remote Glossary

Stop getting lost in interviews and job descriptions

The job market, especially within tech and global companies, has developed its own dialect. Not understanding these acronyms can make you lose valuable opportunities or poorly negotiate your contract. To end this problem, we created the Definitive Glossary for the Remote Professional.

🏢 Work Models & Routine

Async (Asynchronous Work)
A communication model where responses don't need to be immediate. Instead of back-to-back meetings, the team relies on well-structured documents, threads, and messages. It's the gold standard for global companies spanning multiple time zones.
Sync (Synchronous Work)
The opposite of Async. It requires the team to be online and available at the same time for meetings, live chats, and real-time collaboration.
Daily / Stand-up
A quick daily meeting (usually 15 minutes) common in Agile (Scrum) methodologies. The team answers three questions: What did I do yesterday? What will I do today? Are there any blockers?
All-Hands / Town Hall
A company-wide meeting involving all employees. Usually led by the founders (C-Levels) to present results, new goals, and answer team questions.
1:1 (One-on-One)
A recurring individual meeting between a professional and their direct manager. It is used for career alignment, feedback, and problem-solving, not just for project status updates.

💰 Contracts, Benefits & Compensation

PTO (Paid Time Off)
Instead of strict, categorized leave policies, modern US companies usually offer a flexible pool of days (e.g., 20 days of PTO, or even "Unlimited PTO") that you can use for vacations, sick days, or personal matters, while receiving your regular compensation.
Equity / Stock Options
Company ownership. The startup offers you the right to buy shares at a heavily discounted strike price in the future. If the company grows, goes public, or is acquired, these shares can be highly lucrative.
Vesting (Vesting Schedule)
The rule that controls your Equity. It usually lasts 4 years. You don't get all the shares on day one; you "earn" them gradually as you stay with the company. A "1-year Cliff" means you must stay for at least one year to receive your first batch of shares.
RSUs (Restricted Stock Units)
Unlike Stock Options (where you have the right to buy the stock), RSUs are actual shares the company grants you as a bonus or part of your compensation package, following a strict Vesting schedule.
Independent Contractor (1099 / B2B)
The most common international hiring model for global talent working for US companies. You act as a service provider (business-to-business), receiving the gross salary (often six-figure compensation) without standard local payroll tax deductions at the source.

🤖 Recruitment & Hiring Process

ATS (Applicant Tracking System)
The "robot" that reads your resume. Software like Greenhouse, Ashby, and Workday are used to filter candidates by keywords before a human even looks at the document. (Pro tip: this is why your resume must be clean, semantic, and have the right keywords).
JD (Job Description)
The document that lists the responsibilities, technical requirements, and benefits of the open position.
Cultural Fit
The interview stage that evaluates if your core values, communication style, and worldview align with the company's culture. This is the ultimate test of your Soft Skills.
Onboarding
The integration process. It's the period of your first few weeks at the company, where you get your access credentials, learn about the culture, study the internal documentation, and understand how the product works.

🚀 How to use this to your advantage?

The secret isn't just knowing what these acronyms mean, but using them actively. If during an interview for a premium tech role you ask, "How does your PTO policy and Vesting schedule work?", the recruiter will immediately perceive you as a high-level professional, familiar with the global market standards.

The remote job market requires preparation. And having the right vocabulary is the first big step to securing six-figure proposals and standing out among thousands of applicants.

Expert Tip

The Current State of the Cloud & DevOps Market

By Mondywork | The ultimate guide for tech professionals looking to level up their careers, land global remote roles, and master modern infrastructure.

The End of the "SysAdmin" and the Rise of Platform Engineering

If you've been tracking the most sought-after roles at startups and Fortune 500 companies, you've probably noticed a drastic shift. The Cloud Computing and DevOps market is no longer just about "keeping the app online." Today, companies demand extreme resilience, automation, and cost optimization (the famous FinOps).

DevOps culture has evolved. We are now in the era of Platform Engineering, where infrastructure teams build internal platforms to grant developers full autonomy (Self-Service). For those aiming for global remote roles with six-figure compensation, understanding this transition isn't just a plus—it's a strict requirement.

"Elite DevOps teams deploy code 973 times more frequently and boast a Mean Time To Recovery (MTTR) that is 6,570 times faster than low-performing teams."

— State of DevOps Report (DORA / Google Cloud)

What Do Recruiters and ATS Systems Really Look For?

Automated filters (ATS platforms like Greenhouse, Ashby, and Workday) are hardwired to scan for a very specific set of skills in Cloud/DevOps resumes. Simply listing 20 tools won't cut it anymore; you must demonstrate culture and real-world impact.

  • Infrastructure as Code (IaC): Forget clicking around the AWS console. Declarative tools like Terraform, Pulumi, and Ansible are the true industry standards.
  • Container Orchestration: Mastery of Docker and, above all, Kubernetes (K8s) is practically mandatory for premium roles.
  • CI/CD Culture: Building automated pipelines (GitHub Actions, GitLab CI, ArgoCD) that guarantee continuous and secure deployments.
  • Cloud Providers and Architecture: Proficiency in at least one major public cloud provider (AWS, Google Cloud, or Microsoft Azure), with a strong focus on Serverless, microservices, and system resilience.

How to Transition: The Roadmap to Success

Many Back-End developers, support analysts, and network professionals want to pivot to Cloud and DevOps. If that's your goal, follow this market-validated roadmap:

  1. Master the Basics (Linux and Networking): Modern DevOps runs on Linux. Understanding permissions, processes, bash scripting, and the fundamentals of TCP/IP, DNS, and Load Balancers is your foundation.
  2. Learn to Code: A DevOps engineer doesn't need to be a Senior Developer, but you must know how to automate tasks. Python and Go (Golang) are currently the most highly valued languages in this ecosystem.
  3. Are Certifications Worth It? Yes, especially to get past the initial HR screening. Certifications like the AWS Certified Solutions Architect – Associate or the CKA (Certified Kubernetes Administrator) carry massive weight in the global market.
  4. Build a Real Portfolio: Publish a project on GitHub where you provision infrastructure via Terraform, set up a CI/CD pipeline using GitHub Actions, and deploy a simple app to a Kubernetes cluster. Practical projects always beat theoretical resumes.

Essential References and Study Links

To stay up-to-date, we highly recommend following official organizations and annual market reports:


🚀 Want access to the best remote Cloud & DevOps jobs?

Join our mailing list and get the best jobs delivered straight to you, tailored to your areas of interest. Currently, we partner with over 200 companies that advertise their premium, remote-first roles with us. Sign up now and get the best opportunities in the global market right in your inbox!