Brightgrove logo
Українська
Senior Site Reliability Engineer

Senior Site Reliability Engineer

TV Streaming
Location
Remote, Medellin, Colombia
Area
DevOps/Cloud/Systems
Tech Level
Senior
Tech Stack
Kubernetes (K8, K8s), AWS, Golang
Refer a Friend

your info

REFERRAL'S INFO

0/4000

About the Client

Paramount creates entertainment experiences that drive conversation and culture around the world. Through television, film, digital media, live events, merchandise and solutions, our brands connect with diverse, young and the young at heart audiences in more than 180 countries.

Project details

Pluto TV , a Paramount company, is the leading free streaming television service in America, delivering 250+ live and original channels and thousands of on-demand movies in partnership with major TV networks, movie studios, publishers, and digital media companies. Pluto TV is available on all major mobile, web, and connected TV streaming devices and millions of viewers tune in each month to watch premium news, TV shows, movies, sports, lifestyle and trending digital series. Headquartered in West Hollywood, Pluto TV has offices in New York, Silicon Valley, Chicago, and Berlin. We have team members located all over the globe.

Your Team

We are a leading streaming platform delivering live and on-demand content to millions of users globally across web, mobile, and connected TV platforms. The Data Test Engineering (DTE) team protects the quality and integrity of analytics and data pipelines end-to-end — from analytics events generated on client applications through ingestion, transformation, data warehouses, and downstream reporting.

What's in it for you

  • Interview process that respects people and their time
  • Professional and open IT community
  • Internal meet-ups and resources for knowledge sharing
  • Time for recovery and relaxation
  • Bright online and offline events
  • Opportunity to become part of our internal volunteer community

Responsibilities

  • Analyze and improve system design to reduce failure modes and promote self-healing systems.
  • Establish and maintain robust systems that facilitate observability, encompassing logging, monitoring, distributed tracing, alerting, and offline test tools.
  • Work with development partners to shape the architecture, design, and implementation of new and existing systems to enhance their reliability, performance, efficiency, and scalability.
  • Ability to work both independently and as part of a geographically dispersed, yet integrated, team.
  • Collaborate with service engineers to establish Service Level Agreements (SLAs) and Service Level Objectives (SLOs) for backend services.
  • Identify indications or cues that demonstrate the effectiveness of an application and possess the knowledge to improve or repair its performance.
  • Assess options and suggest solutions when information is limited or unclear. This position requires a level of comfort and confidence in dealing with uncertain situations.
  • Work seamlessly within a team while effectively managing individual tasks.
  • Respond to emerging incidents, solve critical issues, and follow through with a plan for resolution or future mitigation.
  • Act as an SME on the Engineering Operations team, partnering with backend services teams and application teams to overcome challenges across all platforms where we stream our service.

Skills

  • 5+ years of experience in software development.
  • Degree in Computer Science or a related field, or equivalent work experience.
  • Solid engineering and coding skills, strong data structure knowledge, and the ability to write high-performance, production-quality code.
  • Experience building service-oriented APIs and cloud services.
  • Experience designing, implementing, and deploying microservices.
  • Extremely technical, hands-on server software experience.
  • Proficient in Golang and JavaScript, with the ability to quickly learn new languages.
  • Experience in Linux environments and a strong understanding of Linux fundamentals and internals, including file systems, modern memory management, threads and processes, and the user-kernel space divide.
  • Strong understanding of large-scale distributed systems in practice, including multi-tier architectures, application security, monitoring, and storage systems.
  • Working knowledge of the TCP/IP stack, internet routing, and load balancing.
  • Grit, drive, and a deep sense of ownership.
Recruiter Marina Yakushenko
Your personal recruiter
Marina Yakushenko

Apply Now

0/4000

sharing is caring & referral bonus