(Eg0069) Senior Site Reliability Engineer (Sre) - Cassandra & Aws - Talent Connection
Nortal
Score da Vaga
100 pontosSenior Site Reliability Engineer (SRE) - Talent Connection
Real ownership from day one, not check-ins and micromanagement. That's what the Senior Site Reliability Engineer (SRE) role at Nortal looks like.
What working with us looks like
- Remote work with a LATAM team. Coffee breaks, tech talks, and games keep it human, even at a distance.
- No micromanagement. We hire for autonomy and expect you to use it.
- Support beyond the job. Our People Care team helps with time off, wellness, and anything in between.
- Our accounts team handles client relationships, so you focus on the work.
What you get
- Competitive USD salary.
- Full remote work, with coworking spaces across LATAM if you want to meet the team in person.
- Paid time off under your country's rules, at full salary.
- National holidays.
- Sick leave, no stress attached.
- A yearly refundable credit. Spend it on anything related to your health and well-being.
- A day off for your birthday.
About this search
Great talent doesn't wait for job postings, so we don't either. We're building a network of skilled professionals for roles that come up regularly with our clients.
Join our Future Talent network and you'll be one of the first people we reach when the right opening appears. We make around 160 hires a year, so it happens often.
The role
As a Senior Site Reliability Engineer, I help modernize and scale production platforms, most recently a customer-data caching platform that serves 20–25+ consuming applications at about 3,000 requests per second. I work across SRE, DevOps, engineering and architecture to improve resiliency, automation, observability, deployment practices and overall platform reliability. I work independently, set technical direction and do well in small engineering teams.
Your day-to-day:
- Design, build and operate highly available, fault-tolerant distributed systems and caching platforms.
- Evaluate the existing architecture and identify opportunities to scale the platform toward 6x current traffic while maintaining performance and resiliency.
- Improve caching strategies, including TTLs, refresh patterns, cache placement, performance and downstream dependency management.
- Help evolve the platform toward active-active resiliency and validate failure, recovery and capacity scenarios.
- Design and implement automated CI/CD pipelines, including rolling, blue/green or canary deployment strategies, automated checks and quality gates.
- Establish infrastructure, configuration, secrets and application deployment practices using an everything-as-code approach.
- Build performance and load-testing capabilities and establish meaningful performance gates for releases.
- Develop monitoring, alerting and observability for cache latency, throughput, availability, errors and other key reliability indicators.
- Define and improve SLIs, SLOs, error budgets and operational KPIs where appropriate.
- Automate operational tasks and reduce manual intervention through scripting, tooling and infrastructure automation.
- Troubleshoot production issues, perform root-cause analysis and drive reliability improvements through blameless post-incident reviews.
- Provide technical guidance and establish engineering guardrails for SRE and development teams.
- Collaborate with Java/Spring Boot engineers, architects, client FTEs and other technical teams to deliver platform improvements.
What we're looking for
- Bachelor's Degree in Computer Science, Engineering, or a related field.
- Strong experience in Site Reliability Engineering, DevOps, platform engineering or a similar production infrastructure role.
- Apache Cassandra experience, including data modeling, replication, consistency, tuning and multi-datacenter deployments.
- Experience operating and scaling distributed systems in production environments.
- Strong understanding of caching architectures, performance optimization and high-availability systems.
- Hands-on AWS experience and familiarity with cloud-native architecture.
- Experience designing and implementing automated CI/CD pipelines and deployment strategies for highly available systems.
- Experience with infrastructure/configuration/secrets as code and a strong preference for automated, repeatable processes.
- Strong understanding of observability, monitoring, alerting, performance metrics and production troubleshooting.
- Experience with performance testing, capacity planning and identifying system bottlenecks.
- Strong programming or scripting experience for automation and tooling; Java/Spring Boot experience strongly preferred.
- Ability to work independently, take ownership of ambiguous technical problems and move work forward with limited direction.
- Ability to provide technical leadership and establish practical engineering standards and guardrails within a small team.
- Advanced English Level is required for this role, as you will work with US clients. Effective communication in English is essential to deliver the best solutions to our clients and expand your horizons.
Strongly Preferred
- Experience with GraphQL and API gateway architectures.
- Kubernetes and containerized application experience.
- GitLab CI/CD experience.
- Terraform or other infrastructure-as-code tools.
- Prometheus, Grafana, Datadog or similar observability platforms.
- Kafka, Amazon MSK or other event-streaming technologies.
- Experience designing active-active or multi-region architectures.
- Experience with enterprise customer-data platforms or high-volume caching systems.
How the process works
- Apply and upload your resume.
- We review your profile and add it to our talent database.
- We reach out for an interview to learn about your experience and share more about what we're looking for.
- When a role opens that fits, we contact you to start the process. Because we already know you, it moves faster.
- No intermediaries. You deal with us directly.
About Nortal
For over 25 years, we've built the systems that global enterprises and public institutions run on. Today, that means using AI and data to help organizations make better decisions, not just build what they asked for.
We believe good solutions don't have to be complicated. That's why our clients trust us with the initiatives that matter most.
Real ownership, real impact, no micromanagement. If that's what you're after, apply now. 👇
By applying to this position, you authorize Nortal to collect, store, transfer, and process your personal data in accordance with our Privacy Policy. For more information, please review our Privacy Policy. (https://nortal.com/recruitment-and-hiring-privacy-statement/)
By applying to this position, you authorize Nortal to collect, store, transfer, and process your personal data in accordance with our Privacy Policy. For more information, please review our Privacy Policy.
Sobre a área de Backend
A área de Backend é responsável por toda a lógica do servidor, APIs, bancos de dados e infraestrutura que sustentam aplicações web e mobile. Profissionais de backend garantem que os sistemas sejam escaláveis, seguros e performáticos.
As principais habilidades incluem linguagens como PHP, Java, Python, Ruby, Go e Node.js, frameworks como Laravel, Spring Boot, Django e Express, bancos de dados (MySQL, PostgreSQL, MongoDB, Redis), arquitetura de software (clean architecture, DDD, microsserviços) e segurança de APIs (OAuth, JWT).
Desenvolvedores Backend em empresas de tecnologia são muito valorizados, especialmente aqueles que dominam arquitetura de microsserviços, cloud computing e performance de alta escala. A área oferece oportunidades desde desenvolvedor júnior até arquiteto de software, com foco em escalabilidade, segurança e eficiência.
Sobre a área de Fullstack
Desenvolvedores Fullstack são profissionais versáteis capazes de atuar tanto no frontend quanto no backend de aplicações web e mobile. Eles dominam múltiplas tecnologias e são capazes de construir produtos completos do início ao fim, desde a interface do usuário até a infraestrutura do servidor.
As principais habilidades incluem domínio de pelo menos uma stack completa (React/Vue/Angular + Node.js/PHP/Python/Java), banco de dados (SQL e NoSQL), APIs REST/GraphQL, versionamento com Git, CI/CD e conhecimento básico de infraestrutura (Docker, cloud). Arquitetura limpa, DDD e testes são diferenciadores importantes.
Desenvolvedores Fullstack são muito valorizados em startups e empresas que precisam de profissionais versáteis e autônomos. A área oferece oportunidades desde desenvolvedor júnior até arquiteto de software, com foco em entrega completa, visão holística do produto e capacidade de trabalhar em múltiplas camadas da aplicação.
Sobre a área de Web Master
O Web Master é o profissional responsável pela manutenção, segurança e performance técnica de sites e aplicações web. Ele gerencia servidores, infraestrutura de hospedagem, monitoramento de uptime e garante que tudo funcione com rapidez e disponibilidade.
As principais habilidades incluem gestão de servidores (Apache, Nginx), hospedagem (AWS, Google Cloud, Azure), CDN (Cloudflare), SSL, DNS, segurança web (WAF, firewall), performance (Core Web Vitals, cache, compressão) e versionamento (Git, CI/CD). Conhecimento de Docker, WordPress, cPanel e monitoramento (Sentry, New Relic) é diferencial.
Web Masters em empresas de tecnologia são muito valorizados, especialmente aqueles que dominam DevOps, SRE e conseguem garantir uptime e performance em escala. A área oferece oportunidades desde webmaster júnior até SRE e infrastructure engineer, com foco em confiabilidade, segurança e velocidade.
Sobre a área de Cloud Solutions
A área de Cloud Solutions é responsável por projetar, implementar e gerenciar infraestrutura e serviços em nuvem (AWS, Azure, GCP) para empresas. Profissionais de cloud arquitetam soluções escaláveis, seguras e otimizadas em custo, desde migrações de data centers até arquiteturas serverless e multi-cloud.
As principais habilidades incluem IaC (Terraform, CloudFormation), containers (Docker, Kubernetes), serverless (Lambda, Cloud Functions), bancos de dados gerenciados (RDS, DynamoDB, BigQuery), redes cloud (VPC, CDN, load balancer) e segurança (IAM, WAF, KMS). Conhecimento de FinOps, cloud governance e certificações AWS/Azure/GCP é diferencial.
Profissionais de Cloud Solutions em empresas de tecnologia são muito valorizados, especialmente aqueles que dominam arquiteturas multi-cloud, FinOps e conseguem otimizar custos enquanto mantêm performance e segurança. A área oferece oportunidades desde cloud engineer até cloud solutions architect, head of cloud e chief cloud architect.
Conheça Outras Áreas
Entenda melhor o escopo de atuação, habilidades necessárias e ferramentas utilizadas em diferentes áreas de carreira.
Sobre a área de Analista de Negócios
O Analista de Negócios (BA) é o profissional responsável por identificar problemas, oportunidades e soluções em processos organizacionais, atuando como ponte entre as áreas de negócio e a equipe de desenvolvimento de tecnologia. Ele realiza o levantamento e especificação de requisitos, mapeia fluxos de valor, desenha processos futuros e ajuda a garantir que as entregas de software estejam alinhadas com as metas estratégicas da empresa.
Sobre a área de Branding
Branding é a área responsável por construir, gerenciar e fortalecer a identidade e o valor de uma marca no mercado. Profissionais de branding criam estratégias que definem como a marca é percebida pelo público, desde o logotipo até a experiência completa do cliente.
As principais habilidades incluem estratégia de marca, identidade visual, brand guidelines, posicionamento, naming, brand voice, pesquisa de mercado, brand equity e gestão de marca. Conhecimento de design gráfico (Figma, Illustrator, Photoshop), storytelling e experiência de marca é diferencial.
Profissionais de Branding em empresas de tecnologia são muito valorizados, especialmente aqueles que dominam employer branding, digital branding e conseguem construir marcas fortes e memoráveis em mercados competitivos. A área oferece oportunidades desde designer de marca até diretor de brand, com foco em identidade, diferenciação e valor percebido.
Sobre a área de Produto
Product Management é uma das áreas mais strategicamente relevantes nas organizações de tecnologia. O Product Manager é responsável por definir a visão do produto, priorizar funcionalidades e coordenar equipes multidisciplinares para entregar valor ao usuário.
As habilidades essenciais incluem pensamento estratégico, análise de dados, comunicação, liderança e conhecimento técnico. Ferramentas como Jira, Confluence, Miro e analytics platforms são fundamentais no dia a dia.
Salários para PMs no Brasil variam de R$ 8.000 (júnior) a R$ 35.000+ (sênior em big techs), com oportunidades crescentes para trabalho remoto internacional.
Sobre a área de Engenheiro de Automação
O Engenheiro de Automação é o profissional responsável por projetar, desenvolver e implementar soluções que automatizam processos manuais e repetitivos em TI, infraestrutura, testes e operações. Ele combina conhecimento de programação com visão de DevOps e SRE para eliminar tarefas manuais e aumentar a eficiência operacional.
As principais habilidades incluem Infrastructure as Code (Terraform, Ansible, Pulumi), CI/CD (Jenkins, GitHub Actions, GitLab CI), automação de testes (Selenium, Cypress, Playwright), automação de rede (Netconf, SDN), RPA (UiPath, Power Automate) e scripting (Python, Bash, PowerShell). Conhecimento de Kubernetes, GitOps (ArgoCD, Flux) e plataforma de automação é diferencial.
Engenheiros de Automação em empresas de tecnologia são muito valorizados, especialmente aqueles que conseguem criar pipelines de deploy automatizados, infraestrutura auto-healing e plataformas internas de desenvolvimento (IDP). A área oferece oportunidades desde automation engineer júnior até automation architect e head of automation.
Sobre a área de Suporte Técnico
O Suporte Técnico é essencial para garantir a satisfação e retenção de clientes. Profissionais de suporte resolvem problemas técnicos, documentam soluções e identificam padrões que podem levar a melhorias no produto.
As principais habilidades incluem troubleshooting, atendimento ao cliente, documentação técnica, conhecimento de ITIL e ferramentas de ticketing (Zendesk, Freshdesk, Intercom).
O suporte técnico evoluiu de uma função reativa para proativa, com profissionais de alto nível atuando em Customer Engineering e Support Engineering.
Comentários 0