- Location
- Bengaluru, IN
- Posted
- 6d ago
About this role
<div class="content-intro"><h2 style="font-family: GothamBold,Helvetica,Arial,sans-serif; color: #662d91;">Teamwork makes the stream work.</h2> <p>&nbsp;</p> <h3 style="font-family: GothamBold,Helvetica,Arial,sans-serif;"><strong>Roku is changing how the world watches TV</strong></h3> <p>Roku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.</p> <p>From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.</p> <p>&nbsp;</p></div><h3 class="mt-6 mb-2 font-semibold text-xl" data-streamdown="heading-3"><span class="font-semibold" data-streamdown="strong">What does the team work on?</span></h3> <p class="text-neutral-900 dark:text-neutral-200 my-2.5 last:mb-0 first:mt-0 ">Roku runs one of the largest data lakes in the world. We store over 70 PB of data, run 10+M queries per month, scan over 100 PB of data per month. Big Data team is the one responsible for building, running, and supporting the platform that makes this possible. We provide all the tools needed to acquire, generate, process, monitor, validate and access the data in the lake for both streaming data and batch. We are also responsible for generating the foundational data. The systems we provide include Scribe, Kafka, Hive, Presto, Spark, Flink, Pinot, and others. The team is actively involved in the Open Source, and we are planning to increase our engagement over time.</p> <p class="text-neutral-900 dark:text-neutral-200 my-2.5 last:mb-0 first:mt-0 ">&nbsp;</p> <h3 class="mt-6 mb-2 font-semibold text-xl" data-streamdown="heading-3"><span class="font-semibold" data-streamdown="strong">What is the role?</span></h3> <p class="text-neutral-900 dark:text-neutral-200 my-2.5 last:mb-0 first:mt-0 ">Roku is in the process of modernizing its Big Data Platform. We are working on defining the new architecture to improve user experience, minimize the cost and increase efficiency. Are you interested in helping us build this state-of-the-art big data platform? Are you an expert with Big Data Technologies? Have you looked under the hood of these systems? Are you interested in Open Source? If you answered "Yes" to these questions, this role is for you!</p> <p class="text-neutral-900 dark:text-neutral-200 my-2.5 last:mb-0 first:mt-0 ">&nbsp;</p> <h3 class="mt-6 mb-2 font-semibold text-xl" data-streamdown="heading-3"><span class="font-semibold" data-streamdown="strong">How will I use AI at Roku?</span></h3> <p data-renderer-start-pos="14538" data-local-id="fd08136e2d75">At Roku, we don’t just use AI, we work with it. AI agents and smart tools help power drafts, analysis, and repetitive workflows, while our people bring direction, judgment, and accountability. We’re looking for curious, adaptable builders who can show how they’ve used AI or automation to move faster, raise the bar, and scale their impact.&nbsp;</p> <p data-renderer-start-pos="14881" data-local-id="195040e56241">We value your AI skills if have built fluency across the agentic engineering toolchain — coding harnesses like Claude Code or Cursor, <span data-highlighted="true" data-vc="highlighted-text"><span class="_kqswh2mm"><span class="_5pioz8co _189e1dm9 _1il9buyh _19lc184f _d0altlke" data-testid="definition-highlighter">MCP</span></span></span> servers, custom skills, or agent frameworks. And you can describe projects where you shipped real work with these tools. You know how to drive an agent, verify its output, and ramp on an unfamiliar codebase with an agent helping you.</p> <p class="text-neutral-900 dark:text-neutral-200 my-2.5 last:mb-0 first:mt-0 ">&nbsp;</p> <h3 class="mt-6 mb-2 font-semibold text-xl" data-streamdown="heading-3"><span class="font-semibold" data-streamdown="strong">What are the responsibilities of the role?</span></h3> <ul class="list-inside list-disc whitespace-normal [li_&amp;]:pl-6" data-streamdown="unordered-list"> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">You will be responsible for streamlining and tuning existing Big Data systems and pipelines and building new ones. Making sure the systems run efficiently and with minimal cost is a top priority</li> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">You will be making changes to the underlying systems and if an opportunity arises, you can contribute your work back into the open source</li> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">You will also be responsible for supporting internal customers and on-call services for the systems we host. Making sure we provided stable environment and great user experience is another top priority for the team</li> </ul> <h3 class="mt-6 mb-2 font-semibold text-xl" data-streamdown="heading-3">&nbsp;</h3> <h3 class="mt-6 mb-2 font-semibold text-xl" data-streamdown="heading-3"><span class="font-semibold" data-streamdown="strong">What experience would help someone be successful in this role at Roku?</span></h3> <ul class="list-inside list-disc whitespace-normal [li_&amp;]:pl-6" data-streamdown="unordered-list"> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">7+ years of production experience building big data platforms based upon Spark, Trino or equivalent</li> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">Strong programming expertise in Java, Scala, Kotlin or another JVM language</li> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">A robust grasp of distributed systems concepts, algorithms, and data structures</li> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">Strong familiarity with the Apache Hadoop ecosystem: Spark, Kafka, Hive/Iceberg/Delta Lake, Presto/Trino, Pinot, etc</li> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">Experience working with at least 3 of the technologies/tools mentioned here: Big Data / Hadoop, Kafka, Spark, Trino, Flink, Airflow, Druid, Hive, Iceberg, Delta Lake, Pinot, Storm etc</li> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">Extensive hands-on experience with public cloud AWS or GCP</li> <li class="py-1 [&amp;>p]:inline" data-streamdown="list-item">BS/MS degree in CS or equivalent</li> <li class="py-1 [&amp;>p]:inline&qu
If this role is no longer available, it will disappear from Praxy automatically.