Software Engineer III, Media Delivery
San Francisco, United States · Hybrid · Full-time
- Posted 2w ago
- From Crunchyroll’s careers page
- Location
- San Francisco, United States
- Work mode
- Hybrid
- Type
- Full-time
- Level
- Senior
- Experience
- 5+ years
- Department
- Software Development
Opens the listing on boards.greenhouse.io
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.
About the role
About the role
As a Software Engineer III on the Media Delivery team, you will design, build, and operate highly reliable systems that stream video content to millions of Crunchyroll customers around the world.
Positioned at the intersection of software engineering, cloud infrastructure, networking, and delivery, you will build fault-tolerant media replication and delivery across multi-cloud origins and global CDNs.
You will drive maximizing availability, resilience, performance, security, and efficiency through automation and deep observability.
What You'll Do
- Architect and operate scalable content replication systems and multi-origin storage to ensure playback resilience.
- Build multi-CDN delivery infrastructure to intelligently route global traffic based on performance, availability, and cost.
- Design fault-tolerant AWS and GCP cloud architectures that remove single points of failure in the delivery path.
- Develop high-volume networking solutions between cloud origins, CDNs, and playback services.
- Automate infrastructure provisioning, deployment, and recovery using Infrastructure as Code (IaC).
- Implement end-to-end observability (metrics, tracing, logging, SLOs, and alerts) to rapidly resolve playback issues.
- Harden platform security by proactively scanning for vulnerabilities and building secure-by-default infrastructure.
- Enhance reliability using automated health detection, failover, and self-healing mechanisms.
- Own full end-to-end service lifecycles, including design, deployment, on-call incident response, and continuous optimization.
- Troubleshoot complex production delivery issues across software, infrastructure, networks, and vendor partners.
- Collaborate with external CDN and cloud providers to optimize performance and resolve delivery bottlenecks.
- Participate in design reviews and code reviews to maintain high engineering standards for scale and maintainability.
What Success Looks Like
In this role, you will own and drive:
- High playback uptime and minimal user impact during infrastructure or vendor failures.
- Robust multi-origin and multi-CDN automated failover capabilities.
- Rapid incident detection and resolution via Datadog observability and actionable alerts.
- A secure platform with minimized vulnerability risks across core services.
- Reduced operational toil through IaC automation and self-healing systems.
- Cost-effective infrastructure efficiency that maintains top-tier streaming quality.
- Strong operational readiness, comprehensive runbooks, and clear ownership.
About You
We encourage you to apply if you meet the following requirements:
- 5+ years of experience building and operating large-scale distributed systems.
- Proficiency in TypeScript (or another modern language) for application services and infrastructure automation.
- Production experience with AWS cloud infrastructure (GCP experience is a plus).
- Hands-on experience with Infrastructure as Code (IaC) tools and best practices.
- Experience operating high-volume CDN architectures (e.g., Fastly, Akamai, Google Media CDN, Cloudflare).
- Solid understanding of networking fundamentals (DNS, HTTP, TLS, caching, load balancing, routing).
- Background in designing high-throughput, low-latency delivery architectures across regions and external networks.
- Experience with Datadog (or similar tools) configuring metrics, logs, traces, SLOs, and alerts.
- Knowledge of cloud and application security principles integrated throughout design and operations.
- Proven track record of diagnosing and fixing complex production incidents across software and vendors.
- Strong ownership mindset with willingness to participate in on-call rotations and incident response.
- Clear technical communication and effective collaboration skills across cross-functional engineering teams.
Skills they ask for
Pick one to see other roles that ask for it.
About Crunchyroll
Anime streaming and fan experiencesCrunchyroll is an anime entertainment brand offering streaming content and experiences for anime fans worldwide.
See all 48 roles at CrunchyrollMore roles at Crunchyroll
See all 48- Senior Manager, Rights ManagementHyderabad · Senior · HybridBusiness Operations · Senior · HybridHyderabad, India18h
- Manager, Rights ManagementHyderabad · HybridBusiness Operations · HybridHyderabad, India18h
- Senior Manager, User Acquisition – Paid SocialLos Angeles · Senior · HybridMarketing · Senior · HybridLos Angeles, United States1d
- Senior Manager, Internal Communications, Programs & EventsLos Angeles · Senior · HybridCommunications and Public Affairs · Senior · HybridLos Angeles, United States1d
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.