Regístrese para acceder a todas las funciones de nuestro servicio
  • Búsqueda de ofertas de trabajo
  • Favoritos
  • Crear CV
    Nuevo
  • Sueldos
  • Alertas de empleo

Senior Engineer, Insights & Reliability Engineering

Crunchyroll, LLC

About Crunchyroll

Founded by fans, Crunchyroll delivers the art and culture of anime to a passionate community. We super-serve over 100 million anime and manga fans across 200+ countries and territories, and help them connect with the stories and characters they crave. Whether that experience is online or in-person, streaming video, theatrical, games, merchandise, events and more, it’s powered by the anime content we all love.

Join our team, and help us shape the future of anime!

Crunchyroll, LLC is an independently operated joint venture between US-based Sony Pictures Entertainment, and Japan's Aniplex, a subsidiary of Sony Music Entertainment (Japan) Inc., both subsidiaries of Tokyo-based Sony Group Corporation.

About the role

The Insights & Reliability Engineering team builds data-intensive software systems that help Crunchyroll proactively detect, investigate, and improve reliability across user-facing experiences, including device QoE, payments, partner integrations, and core product journeys.

As a Senior Engineer, you will build, operate, and evolve tools, telemetry workflows, alerting systems, and operational insight platforms that turn high-volume product and platform data into actionable reliability signals. Your work will help engineering teams identify regressions, understand customer impact, improve signal quality, and drive systemic fixes before issues meaningfully affect users.

This is a software engineering role for someone who enjoys working close to data and production systems. The ideal candidate has strong coding fundamentals, SQL/data fluency, experience building reliable tools or data workflows, and the production instincts to solve ambiguous production problems across systems. As a senior engineer, you will help shape technical direction, define scalable operational practices, and influence cross-functional partners across engineering, product, data, and partner-facing teams.

What you will be doing

  • Build, operate, and evolve reliability tooling, telemetry workflows, alerting systems, dashboards, and automation that help teams detect and investigate customer-impacting issues.
  • Transform high-volume product, device, payment, partner, and platform telemetry into reliable operational signals for anomaly detection, alerting, debugging, and leadership visibility.
  • Use SQL, logs, metrics, dashboards, and product context to identify anomalies, baseline performance, validate hypotheses, and uncover patterns that drive reliability improvements.
  • Lead cross-team improvements to monitoring, alerting, signal quality, triage workflows, and operational automation, defining scalable patterns that reduce time to detection and resolution.
  • Participate in incident response, alert review, and production issue investigation for device, payment, and partner-related issues, with a focus on improving the systems and processes that make future issues easier to detect and resolve.
  • Work with internal service owners and external technical partners when reliability issues span device, payment, or partner ecosystems, with a focus on improving the software systems, telemetry, and ownership paths that make future issues easier to detect and resolve.
  • Collaborate with engineering, product, data, analytics, and partner teams to drive systemic fixes, postmortem learnings, and long-term reliability improvements.
  • Help set technical direction for reliability tooling and investigation workflows within your area, establishing patterns for telemetry quality, debugging, and operational readiness.
  • Communicate technical findings, customer impact, tradeoffs, and recommendations clearly to both technical and non-technical stakeholders.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, Data Science, or a related field, and/or 8+ years of practical engineering experience.
  • Strong coding skills, with Python preferred; experience with similar modern programming languages is also welcome if you can ramp up quickly in Python-based workflows.
  • Strong SQL skills and experience working with high-volume event, telemetry, product, or operational datasets.
  • Experience building, testing, deploying, and maintaining production-quality software, internal tools, data workflows, automation, or observability systems.
  • Experience with data-processing systems, analytics engineering workflows, reliability tooling, alerting systems, or operational insight platforms.
  • Ability to debug ambiguous production problems using data, logs, metrics, dashboards, code, system behavior, and product context.
  • Experience working across data, observability, and analytics platforms, such as Databricks/Spark for data processing, Airflow/dbt for workflow orchestration and modeling, Datadog/Grafana for monitoring, and Tableau/Mixpanel/Mux for product or experience analytics.
  • Experience defining scalable operational processes, improving signal quality, reducing alert noise, and turning repeated manual workflows into durable automation.
  • Experience in reliability engineering, SRE, incident response, support engineering, analytics engineering, data engineering, or operational tooling roles.
  • Comfortable working with consumer device, payment, partner integration, or distributed system domains.
  • Proven ability to influence stakeholders across Engineering, Product, Data, Analytics, and partner-facing teams.
  • Experience leading ambiguous, cross-system initiatives and aligning multiple teams around technical priorities, operational tradeoffs, and durable reliability improvements.
  • Experience using AI-assisted engineering tools, including agentic workflows where appropriate, to accelerate investigation, implementation, testing, or automation, with strong judgment around correctness, maintainability, security, and production safety.
  • Clear communicator, capable of translating technical findings, reliability risks, and customer impact for broad audiences.

You Might Be a Great Fit If You...

  • Enjoy building and evolving software, data workflows, and observability systems that make complex production behavior easier to understand and improve.
  • Are fluent in both code and data, and like using telemetry, SQL, logs, metrics, and product context to solve ambiguous problems.
  • Are obsessed with understanding root causes, not just patching symptoms.
  • Believe in turning repeated manual workflows into reliable tools, automation, and scalable operating practices.
  • Have a strong sense of ownership and pride in improving customer experience through better reliability systems.
  • Enjoy working behind the scenes to make things “just work” for millions of fans.

 

#LifeAtCrunchyroll #LI-Hybrid

About our Values

We want to be everything for someone rather than something for everyone and we do this by living and modeling our values in all that we do. We value

  • Courage. We believe that when we overcome fear, we enable our best selves.

  • Curiosity. We are curious, which is the gateway to empathy, inclusion, and understanding.

  • Kaizen. We have a growth mindset committed to constant forward progress.
  • Service. We serve our community with humility, enabling joy and belonging for others.

Our commitment to diversity and inclusion

Our mission of helping people belong reflects our commitment to diversity & inclusion. It's just the way we do business.

We are an equal opportunity employer and value diversity at Crunchyroll. Pursuant to applicable law, we do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Crunchyroll, LLC is an independently operated joint venture between US-based Sony Pictures Entertainment, and Japan's Aniplex, a subsidiary of Sony Music Entertainment (Japan) Inc., both subsidiaries of Tokyo-based Sony Group Corporation.

Questions about Crunchyroll’s hiring process? Please check out our Hiring FAQs: 

Please refer to our Candidate Privacy Policy for more information about how we process your personal information, and your data protection rights:

Please beware of recent scams to online job seekers. Those applying to our job openings will only be contacted directly from @crunchyroll.com email account.

Vacante publicada el 12 días atrás
Empleos similares que podrían interesarleBasado en la vacante Senior Engineer, Insights & Reliability Engineering en Ciudad de México
  •  ...Crunchyroll is seeking a Senior Engineer in Insights & Reliability Engineering to build and evolve tools, telemetry, alerting, and operational platforms across user experiences. You will help detect regressions, improve signal quality, and drive durable reliability improvements... 
    Senior
    Trabajo híbrido

    Ellation, Inc.

    Ciudad de México
    3 días atrás
  •  ...Qualcomm is looking for a Sr. IT Engineer in Mexico to support and operate enterprise Git platforms such as Gerrit, GitHub, and GitLab. This role demands strong experience in site reliability engineering, production operations, and automation. The ideal candidate will... 
    Senior

    Qualcomm

    Ciudad de México
    1 día atrás
  •  ...Qualcomm is seeking a Senior HPC Scheduler Operations Engineer to operate, scale, and optimize the IBM Spectrum LSF-based scheduling stack for large...  ...IT Hardware Infrastructure — EDA Compute, influencing reliability and performance across sites. Proven Linux administration... 
    Senior

    Qualcomm

    Ciudad de México
    2 días atrás
  •  ...Levi Strauss & Co. is seeking a Site Reliability Engineer for their Data & AI Platform Engineering team in Mexico City. You will monitor production systems, respond to incidents, and improve infrastructure automation on platforms including GCP. This role offers an exciting... 
    Senior

    Levi Strauss & Co.

    Ciudad de México
    15 horas atrás
  •  ...institutions and retailers across the globe. We are looking for Site Reliability Engineer (SRE) to join our team, with an initial focus on production...  ...and configuration management practices. Collaborate with senior engineers to identify opportunities for operational... 
    Senior

    NCR Atleos Corporation

    Ciudad de México
    3 días atrás
  •  ...Elsevier in Mexico City (Hybrid) is seeking a Senior MLOps Engineer to bridge data science and engineering for scalable AI platforms in research...  .../CD across AWS, Azure, and Databricks. The role emphasizes reliability, governance, and collaboration with product managers and... 
    Senior
    Trabajo híbrido

    LexisNexis Risk Solutions

    Ciudad de México
    4 días atrás
  •  ...EarnIn is seeking an experienced software engineer to join the DevelSoper Experience team. This remote role may be hybrid from the Mexico...  ...tooling to accelerate engineering delivery, improve CI/CD reliability, and reduce developer toil. You will work with Python/Go, LLM... 
    Senior
    Trabajar en la oficina
    Remoto
    Trabajo híbrido

    EarnIn

    Ciudad de México
    4 días atrás
  • $115,400 - $192,300 por año

    ## Senior MLOps EngineerApply: MEXICO-MEXICOCITY: Full time: Posted...  ...AI innovations into reliable, secure, and production-ready...  ...will bridge Data Science and Engineering to turn experimental NLP/IR/GenAI...  ...Engineering, Search and Recommendation Engines** Automate and orchestrate... 
    Senior
    Tiempo completo
    Remoto
    Inicio inmediato
    Trabajo híbrido
    Horario flexible

    LexisNexis Risk Solutions

    Ciudad de México
    15 horas atrás
  •  ...skills and become the heart of an innovative engine that is contributing to global impact...  ...diverse team. About the role: The Senior Analyst, Microsoft 365 Platform...  ...Your work will directly impact service reliability, user experience, operational efficiency... 
    Senior
    Tiempo completo
    Trabajar en la oficina
    Trabajo híbrido
    Horario flexible

    Takeda

    Santa Fe, Álvaro Obregón, D.F.
    2 días atrás
  •  ...Technology Group, Information Technology Group IT Engineering General Summary: Qualcomm is seeking a Senior Storage & Backup engineer with 3 to 5 years of experience in...  ...service availability, capacity, and reliability. ~ Efficiently handle performance issues... 
    Senior
    Trabajar en la oficina
    Fin de semana

    Qualcomm

    Ciudad de México
    4 días atrás
  •  ...Capital One National Association is seeking a Senior Manager SRE in Mexico City to lead the Site Reliability Engineering efforts. You'll shape the reliability vision and drive technical excellence across multiple teams, focusing on AI-driven solutions and operational... 
    Senior

    Capital One National Association

    Ciudad de México
    3 días atrás
  •  ...Audible is seeking a Senior Technical Program Manager to drive complex initiatives from concept to launch within the Support Engineering Program. You will lead cross-functional teams across...  ..., and deployment to deliver reliable global systems. You will partner... 
    Senior

    Amazon

    Ciudad de México
    2 días atrás
  •  ...Dana Canada Corp. seeks a senior Maintenance Director to lead preventive, predictive, and reliability programs across our manufacturing site. You will drive asset performance...  ..., and cost reduction while partnering with engineering, production, and quality teams. The ideal... 
    Senior

    Dana Canada Corp.

    Morelos, Venustiano Carranza, D.F.
    1 día atrás
  •  ...We are seeking a senior project leader to oversee strategic engineering, procurement, and construction (EPC) projects, including asset modernization, major maintenance, and operational improvement initiatives. The ideal candidate will bring experience working for EPC... 
    Senior
    Contratista
    Offshore

    Confidential

    Polanco, Miguel Hidalgo, D.F.
    1 día atrás
  •  ...Capital One seeks a Senior Manager SRE to lead the Site Reliability Engineering center in Mexico City. This role involves shaping technical vision, driving automation, and enhancing operational excellence across multiple teams. The ideal candidate will have extensive... 
    Senior

    Capital One

    Ciudad de México
    3 días atrás
  •  ...Takeda is seeking a Senior Analyst for Microsoft 365 Platform Engineering in Mexico City. You will deliver, configure, support, and continuously improve collaboration and productivity platforms within Takeda's digital workplace portfolio. You will design, implement,... 
    Senior

    Takeda

    Ciudad de México
    1 día atrás
  •  ...Zillow Group Inc. in Mexico (teleworker) seeks a Senior Analytics Engineer to translate complex business questions into actionable insights, building robust BI solutions and self-service reporting. You will lead the design, development, and maintenance of scalable... 
    Senior
    Remoto

    Zillow Group Inc.

    Ciudad de México
    2 días atrás
  •  ...Senior Systems Engineer IT Operations/Information Security·Full-Time·Mexico Build Your Future...  ...that delivers the accurate, real-time insights businesses need to make smarter decisions...  ..., working across teams to deliver reliable, scalable, and secure solutions that support... 
    Senior
    Tiempo completo
    Trabajo híbrido

    Hirebridge

    Ciudad de México
    15 horas atrás
  •  ...Delinea is seeking a senior backend engineer to enhance AI agent discovery across cloud platforms and endpoints. You will evaluate security posture and integrate agent data into the identity governance layer, enabling comprehensive protection of human and machine identities... 
    Senior

    Delinea

    Ciudad de México
    3 días atrás
  •  ...Takeda is seeking a Senior Analyst for Microsoft 365 Platform Engineering in Mexico City. You will deliver, configure, and enhance collaboration and productivity platforms within Takeda's digital workplace portfolio, focusing on automation, operations, and user experience... 
    Senior
    Trabajar en la oficina
    Trabajo híbrido

    Takeda

    Santa Fe, Álvaro Obregón, D.F.
    2 días atrás
  •  ...Salesforce is seeking a senior Platform/ML Infra Engineer to advance our AI-native engineering efforts. You will write and review infrastructure code with Claude Code, publish reusable AI tools, and design autonomous agents to speed engineering workflows. You will secure... 
    Senior

    Salesforce.com, inc.

    Ciudad de México
    4 días atrás
  •  ...Capital One seeks a Distinguished Engineer in Mexico City to lead high-impact engineering efforts, mentor teams, and shape architecture across platforms. You will drive adoption of modern patterns, collaborate with product partners, and influence technical direction at... 
    Senior

    Capital One Group

    Ciudad de México
    2 días atrás
  •  ...Knight Piésold, a specialized international engineering firm, seeks a Hydromechanical Engineering Leader to manage a group of water and mechanical engineers. This role focuses on hydromechanical analyses and modeling, including surface water management, water balance,... 
    Senior

    GeoWorld

    Ciudad de México
    4 días atrás
  •  ...Delinea is seeking a Senior Software Engineer to expand our AI platform integrations. You will build integrations to extend discovery, identity...  ..., Securing AI, and join a team that emphasizes security, reliability, and collaboration as you reuse shared identity components... 
    Senior

    Delinea

    Ciudad de México
    3 días atrás
  •  ...looking for a Full-Stack AI Engineer who builds production AI systems...  ...that keep AI systems reliable in production Optimize prompt...  ...Executive Interview - Meet with senior leadership to discuss role...  ...for the role while giving you insights into our company culture and... 
    Senior
    Práctica
    Tiempo completo
    Desde casa
    Remoto

    Hire Overseas

    Ciudad de México
    Hace 2 meses
  •  ...Posted Today: JR102790Sulzer is a leading engineering company with a proud heritage of...  ...roleWe are seeking a highly skilled **Senior Tendering Engineer** to act as a high-level technical...  ...tools (e.g., Salesforce/RFT) to ensure reliable sales forecasts and workload reporting... 
    Senior
    Tiempo completo
    Trabajar en la oficina

    Sulzer

    Cuautitlán Izcalli, Méx.
    2 días atrás
  •  ...Journeyfront, Inc. is seeking a Senior SDET Engineer to drive quality engineering for a multi phase commercial real estate modernization initiative. You will develop and maintain automated test coverage with Playwright while performing manual testing when needed. As... 
    Senior

    Journeyfront, Inc.

    Ciudad de México
    4 días atrás
  •  ...We are looking for a Senior SDET Engineer to drive quality engineering for a multi phase commercial real estate modernization initiative. In this role, you will develop and maintain automated test coverage while maintaining the flexibility to perform manual testing when... 
    Senior

    Journeyfront, Inc.

    Ciudad de México
    4 días atrás
  •  ...AgileEngine, LLC. in Ciudad de México is seeking a Senior InRiver PIM Engineer to maintain and customize PIM platforms within a global engineering team. You will design specifications, develop C#/.NET extensions, build integrations, and troubleshoot while collaborating... 
    Senior
    Remoto

    AgileEngine, LLC.

    Ciudad de México
    1 día atrás
  •  ...GBS Group is hiring a Senior Platform Engineer to design, deploy, and maintain internal tools, automation systems, and AI-driven workflows....  ...role requires ownership, API integrations, security, and reliability, with direct impact on business operations. The ideal candidate... 
    Senior
    Remoto

    GBS Group

    Ciudad de México
    2 días atrás

¿Desea recibir más vacantes?

Suscríbase y reciba vacantes similares a Senior Engineer, Insights & Reliability Engineering. ¡Sea el primero en aplicar!