Senior Engineer, Insights & Reliability Engineering
Crunchyroll, LLC
About Crunchyroll
Founded by fans, Crunchyroll delivers the art and culture of anime to a passionate community. We super-serve over 100 million anime and manga fans across 200+ countries and territories, and help them connect with the stories and characters they crave. Whether that experience is online or in-person, streaming video, theatrical, games, merchandise, events and more, it’s powered by the anime content we all love.
Join our team, and help us shape the future of anime!
Crunchyroll, LLC is an independently operated joint venture between US-based Sony Pictures Entertainment, and Japan's Aniplex, a subsidiary of Sony Music Entertainment (Japan) Inc., both subsidiaries of Tokyo-based Sony Group Corporation.
About the role
The Insights & Reliability Engineering team builds data-intensive software systems that help Crunchyroll proactively detect, investigate, and improve reliability across user-facing experiences, including device QoE, payments, partner integrations, and core product journeys.
As a Senior Engineer, you will build, operate, and evolve tools, telemetry workflows, alerting systems, and operational insight platforms that turn high-volume product and platform data into actionable reliability signals. Your work will help engineering teams identify regressions, understand customer impact, improve signal quality, and drive systemic fixes before issues meaningfully affect users.
This is a software engineering role for someone who enjoys working close to data and production systems. The ideal candidate has strong coding fundamentals, SQL/data fluency, experience building reliable tools or data workflows, and the production instincts to solve ambiguous production problems across systems. As a senior engineer, you will help shape technical direction, define scalable operational practices, and influence cross-functional partners across engineering, product, data, and partner-facing teams.
What you will be doing
- Build, operate, and evolve reliability tooling, telemetry workflows, alerting systems, dashboards, and automation that help teams detect and investigate customer-impacting issues.
- Transform high-volume product, device, payment, partner, and platform telemetry into reliable operational signals for anomaly detection, alerting, debugging, and leadership visibility.
- Use SQL, logs, metrics, dashboards, and product context to identify anomalies, baseline performance, validate hypotheses, and uncover patterns that drive reliability improvements.
- Lead cross-team improvements to monitoring, alerting, signal quality, triage workflows, and operational automation, defining scalable patterns that reduce time to detection and resolution.
- Participate in incident response, alert review, and production issue investigation for device, payment, and partner-related issues, with a focus on improving the systems and processes that make future issues easier to detect and resolve.
- Work with internal service owners and external technical partners when reliability issues span device, payment, or partner ecosystems, with a focus on improving the software systems, telemetry, and ownership paths that make future issues easier to detect and resolve.
- Collaborate with engineering, product, data, analytics, and partner teams to drive systemic fixes, postmortem learnings, and long-term reliability improvements.
- Help set technical direction for reliability tooling and investigation workflows within your area, establishing patterns for telemetry quality, debugging, and operational readiness.
- Communicate technical findings, customer impact, tradeoffs, and recommendations clearly to both technical and non-technical stakeholders.
Qualifications
- Bachelor’s degree in Computer Science, Engineering, Data Science, or a related field, and/or 8+ years of practical engineering experience.
- Strong coding skills, with Python preferred; experience with similar modern programming languages is also welcome if you can ramp up quickly in Python-based workflows.
- Strong SQL skills and experience working with high-volume event, telemetry, product, or operational datasets.
- Experience building, testing, deploying, and maintaining production-quality software, internal tools, data workflows, automation, or observability systems.
- Experience with data-processing systems, analytics engineering workflows, reliability tooling, alerting systems, or operational insight platforms.
- Ability to debug ambiguous production problems using data, logs, metrics, dashboards, code, system behavior, and product context.
- Experience working across data, observability, and analytics platforms, such as Databricks/Spark for data processing, Airflow/dbt for workflow orchestration and modeling, Datadog/Grafana for monitoring, and Tableau/Mixpanel/Mux for product or experience analytics.
- Experience defining scalable operational processes, improving signal quality, reducing alert noise, and turning repeated manual workflows into durable automation.
- Experience in reliability engineering, SRE, incident response, support engineering, analytics engineering, data engineering, or operational tooling roles.
- Comfortable working with consumer device, payment, partner integration, or distributed system domains.
- Proven ability to influence stakeholders across Engineering, Product, Data, Analytics, and partner-facing teams.
- Experience leading ambiguous, cross-system initiatives and aligning multiple teams around technical priorities, operational tradeoffs, and durable reliability improvements.
- Experience using AI-assisted engineering tools, including agentic workflows where appropriate, to accelerate investigation, implementation, testing, or automation, with strong judgment around correctness, maintainability, security, and production safety.
- Clear communicator, capable of translating technical findings, reliability risks, and customer impact for broad audiences.
You Might Be a Great Fit If You...
- Enjoy building and evolving software, data workflows, and observability systems that make complex production behavior easier to understand and improve.
- Are fluent in both code and data, and like using telemetry, SQL, logs, metrics, and product context to solve ambiguous problems.
- Are obsessed with understanding root causes, not just patching symptoms.
- Believe in turning repeated manual workflows into reliable tools, automation, and scalable operating practices.
- Have a strong sense of ownership and pride in improving customer experience through better reliability systems.
- Enjoy working behind the scenes to make things “just work” for millions of fans.
#LifeAtCrunchyroll #LI-Hybrid
About our Values
We want to be everything for someone rather than something for everyone and we do this by living and modeling our values in all that we do. We value
Courage. We believe that when we overcome fear, we enable our best selves.
Curiosity. We are curious, which is the gateway to empathy, inclusion, and understanding.
- Kaizen. We have a growth mindset committed to constant forward progress.
Service. We serve our community with humility, enabling joy and belonging for others.
Our commitment to diversity and inclusion
Our mission of helping people belong reflects our commitment to diversity & inclusion. It's just the way we do business.
We are an equal opportunity employer and value diversity at Crunchyroll. Pursuant to applicable law, we do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.
Crunchyroll, LLC is an independently operated joint venture between US-based Sony Pictures Entertainment, and Japan's Aniplex, a subsidiary of Sony Music Entertainment (Japan) Inc., both subsidiaries of Tokyo-based Sony Group Corporation.
Questions about Crunchyroll’s hiring process? Please check out our Hiring FAQs:
Please refer to our Candidate Privacy Policy for more information about how we process your personal information, and your data protection rights:
Please beware of recent scams to online job seekers. Those applying to our job openings will only be contacted directly from @crunchyroll.com email account.
- ...Crunchyroll is seeking a Senior Engineer in Insights & Reliability Engineering to build and evolve tools, telemetry, alerting, and operational platforms across user experiences. You will help detect regressions, improve signal quality, and drive durable reliability improvements...SeniorTrabajo híbrido
- ...Qualcomm is looking for a Sr. IT Engineer in Mexico to support and operate enterprise Git platforms such as Gerrit, GitHub, and GitLab. This role demands strong experience in site reliability engineering, production operations, and automation. The ideal candidate will...Senior
- ...Qualcomm is seeking a Senior HPC Scheduler Operations Engineer to operate, scale, and optimize the IBM Spectrum LSF-based scheduling stack for large... ...IT Hardware Infrastructure — EDA Compute, influencing reliability and performance across sites. Proven Linux administration...Senior
- ...Levi Strauss & Co. is seeking a Site Reliability Engineer for their Data & AI Platform Engineering team in Mexico City. You will monitor production systems, respond to incidents, and improve infrastructure automation on platforms including GCP. This role offers an exciting...Senior
- ...institutions and retailers across the globe. We are looking for Site Reliability Engineer (SRE) to join our team, with an initial focus on production... ...and configuration management practices. Collaborate with senior engineers to identify opportunities for operational...Senior
- ...Elsevier in Mexico City (Hybrid) is seeking a Senior MLOps Engineer to bridge data science and engineering for scalable AI platforms in research... .../CD across AWS, Azure, and Databricks. The role emphasizes reliability, governance, and collaboration with product managers and...SeniorTrabajo híbrido
- ...EarnIn is seeking an experienced software engineer to join the DevelSoper Experience team. This remote role may be hybrid from the Mexico... ...tooling to accelerate engineering delivery, improve CI/CD reliability, and reduce developer toil. You will work with Python/Go, LLM...SeniorTrabajar en la oficinaRemotoTrabajo híbrido
$115,400 - $192,300 por año
## Senior MLOps EngineerApply: MEXICO-MEXICOCITY: Full time: Posted... ...AI innovations into reliable, secure, and production-ready... ...will bridge Data Science and Engineering to turn experimental NLP/IR/GenAI... ...Engineering, Search and Recommendation Engines** Automate and orchestrate...SeniorTiempo completoRemotoInicio inmediatoTrabajo híbridoHorario flexible- ...skills and become the heart of an innovative engine that is contributing to global impact... ...diverse team. About the role: The Senior Analyst, Microsoft 365 Platform... ...Your work will directly impact service reliability, user experience, operational efficiency...SeniorTiempo completoTrabajar en la oficinaTrabajo híbridoHorario flexible
- ...Technology Group, Information Technology Group IT Engineering General Summary: Qualcomm is seeking a Senior Storage & Backup engineer with 3 to 5 years of experience in... ...service availability, capacity, and reliability. ~ Efficiently handle performance issues...SeniorTrabajar en la oficinaFin de semana
- ...Capital One National Association is seeking a Senior Manager SRE in Mexico City to lead the Site Reliability Engineering efforts. You'll shape the reliability vision and drive technical excellence across multiple teams, focusing on AI-driven solutions and operational...Senior
- ...Audible is seeking a Senior Technical Program Manager to drive complex initiatives from concept to launch within the Support Engineering Program. You will lead cross-functional teams across... ..., and deployment to deliver reliable global systems. You will partner...Senior
- ...Dana Canada Corp. seeks a senior Maintenance Director to lead preventive, predictive, and reliability programs across our manufacturing site. You will drive asset performance... ..., and cost reduction while partnering with engineering, production, and quality teams. The ideal...Senior
- ...We are seeking a senior project leader to oversee strategic engineering, procurement, and construction (EPC) projects, including asset modernization, major maintenance, and operational improvement initiatives. The ideal candidate will bring experience working for EPC...SeniorContratistaOffshore
- ...Capital One seeks a Senior Manager SRE to lead the Site Reliability Engineering center in Mexico City. This role involves shaping technical vision, driving automation, and enhancing operational excellence across multiple teams. The ideal candidate will have extensive...Senior
- ...Takeda is seeking a Senior Analyst for Microsoft 365 Platform Engineering in Mexico City. You will deliver, configure, support, and continuously improve collaboration and productivity platforms within Takeda's digital workplace portfolio. You will design, implement,...Senior
- ...Zillow Group Inc. in Mexico (teleworker) seeks a Senior Analytics Engineer to translate complex business questions into actionable insights, building robust BI solutions and self-service reporting. You will lead the design, development, and maintenance of scalable...SeniorRemoto
- ...Senior Systems Engineer IT Operations/Information Security·Full-Time·Mexico Build Your Future... ...that delivers the accurate, real-time insights businesses need to make smarter decisions... ..., working across teams to deliver reliable, scalable, and secure solutions that support...SeniorTiempo completoTrabajo híbrido
- ...Delinea is seeking a senior backend engineer to enhance AI agent discovery across cloud platforms and endpoints. You will evaluate security posture and integrate agent data into the identity governance layer, enabling comprehensive protection of human and machine identities...Senior
- ...Takeda is seeking a Senior Analyst for Microsoft 365 Platform Engineering in Mexico City. You will deliver, configure, and enhance collaboration and productivity platforms within Takeda's digital workplace portfolio, focusing on automation, operations, and user experience...SeniorTrabajar en la oficinaTrabajo híbrido
- ...Salesforce is seeking a senior Platform/ML Infra Engineer to advance our AI-native engineering efforts. You will write and review infrastructure code with Claude Code, publish reusable AI tools, and design autonomous agents to speed engineering workflows. You will secure...Senior
- ...Capital One seeks a Distinguished Engineer in Mexico City to lead high-impact engineering efforts, mentor teams, and shape architecture across platforms. You will drive adoption of modern patterns, collaborate with product partners, and influence technical direction at...Senior
- ...Knight Piésold, a specialized international engineering firm, seeks a Hydromechanical Engineering Leader to manage a group of water and mechanical engineers. This role focuses on hydromechanical analyses and modeling, including surface water management, water balance,...Senior
- ...Delinea is seeking a Senior Software Engineer to expand our AI platform integrations. You will build integrations to extend discovery, identity... ..., Securing AI, and join a team that emphasizes security, reliability, and collaboration as you reuse shared identity components...Senior
- ...looking for a Full-Stack AI Engineer who builds production AI systems... ...that keep AI systems reliable in production Optimize prompt... ...Executive Interview - Meet with senior leadership to discuss role... ...for the role while giving you insights into our company culture and...SeniorPrácticaTiempo completoDesde casaRemoto
- ...Posted Today: JR102790Sulzer is a leading engineering company with a proud heritage of... ...roleWe are seeking a highly skilled **Senior Tendering Engineer** to act as a high-level technical... ...tools (e.g., Salesforce/RFT) to ensure reliable sales forecasts and workload reporting...SeniorTiempo completoTrabajar en la oficina
- ...Journeyfront, Inc. is seeking a Senior SDET Engineer to drive quality engineering for a multi phase commercial real estate modernization initiative. You will develop and maintain automated test coverage with Playwright while performing manual testing when needed. As...Senior
- ...We are looking for a Senior SDET Engineer to drive quality engineering for a multi phase commercial real estate modernization initiative. In this role, you will develop and maintain automated test coverage while maintaining the flexibility to perform manual testing when...Senior
- ...AgileEngine, LLC. in Ciudad de México is seeking a Senior InRiver PIM Engineer to maintain and customize PIM platforms within a global engineering team. You will design specifications, develop C#/.NET extensions, build integrations, and troubleshoot while collaborating...SeniorRemoto
- ...GBS Group is hiring a Senior Platform Engineer to design, deploy, and maintain internal tools, automation systems, and AI-driven workflows.... ...role requires ownership, API integrations, security, and reliability, with direct impact on business operations. The ideal candidate...SeniorRemoto
¿Desea recibir más vacantes?
Suscríbase y reciba vacantes similares a Senior Engineer, Insights & Reliability Engineering. ¡Sea el primero en aplicar!
- ingeniero innovacion Ciudad de México
- ingeniero de medio ambiente Ciudad de México
- ingeniero en sitio Ciudad de México
- ingeniero aeropuerto Ciudad de México
- senior engineer Ciudad de México
- ingeniero sin experiencia Ciudad de México
- ingeniero pemex Ciudad de México
- ingeniero extranjero Ciudad de México
- ingeniero preventa Ciudad de México
- ingeniero nocturno Ciudad de México


