Description
<div class="content-intro"><div class="c-message_kit__blocks c-message_kit__blocks--rich_text">
<div class="c-message__message_blocks c-message__message_blocks--rich_text" data-qa="message-text">
<div class="p-block_kit_renderer" data-qa="block-kit-renderer">
<div class="p-block_kit_renderer__block_wrapper p-block_kit_renderer__block_wrapper--first">
<div class="p-rich_text_block">
<div class="p-rich_text_section">Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the topics they care most about. With 100,000+ active communities and approximately 126 million daily active unique visitors, Reddit is one of the internet’s largest sources of information. For more information, visit <a class="c-link" href="http://www.redditinc.com/" target="_blank" data-stringify-link="http://redditinc.com" data-sk="tooltip_parent">www.redditinc.com</a>.</div>
</div>
</div>
</div>
</div>
</div></div><p>Reddit's Transport and Site Defense teams build and operate foundational infrastructure behind one of the largest sites in the world. These teams own how requests enter Reddit, how services discover and talk to each other, how traffic moves across clusters and clouds, and how Reddit protects itself from DDoS attacks, scrapers, bots, and other undesired traffic.</p>
<p>We are looking for an experienced Engineering Manager to lead and grow the Site Defense teams. This is a high-impact role for a manager who can guide senior engineers, make strong technical and organizational tradeoffs, and help Reddit scale critical traffic and defense platforms while keeping the site available.</p>
<h4><strong>How You'll Have Impact: </strong></h4>
<ul>
<li>Complete Reddit's migration from Fastly to Cloudflare, including high-traffic properties such as reddit.com, media endpoints, Ads endpoints, e.reddit.com/gql-fed, devvit.reddit.com, and long-tail CDN traffic.</li>
<li>Make Portal Transporter the predominant service discovery platform for Reddit, moving toward broad gRPC/xDS adoption and reducing fragmented service discovery.</li>
<li>Migrate high-volume Thrift traffic to gRPC through TGRPC while preserving reliability, customer trust, and migration velocity.</li>
<li>Reduce data transfer and cross-cluster networking costs through smarter routing, service discovery, and infrastructure standardization.</li>
<li>Expand Site Defense coverage across post-CDN traffic, including TAC, RLS, TRS, Recaptcha, edge challenges, device attestation, IP reputation, and bot context.</li>
<li>Deliver the Site Defense Platform and ML-at-ingress direction in a way that improves defense effectiveness while controlling false positives and preserving user experience.</li>
<li>Operate merged Transport and Site Defense on-call responsibilities while protecting team focus, improving runbooks, cross-training engineers, and clarifying ownership boundaries with Site Experience SRE.</li>
<li>Maintain high availability and low latency while performing large migrations on systems that sit directly in Reddit's request path.</li>
</ul>
<p><strong>What you'll do: </strong></p>
<ul>
<li>Lead, hire, onboard, and develop a high-caliber team of infrastructure engineers across Transport and Site Defense.</li>
<li>Partner with senior ICs to define technical strategy for global traffic, edge/CDN, service discovery, gRPC datapath, rate limiting, site defense, and bot detection infrastructure.</li>
<li>Drive multi-quarter execution on high-risk platform migrations, including CDN migration, inter-cluster load balancer adoption, Thrift to GRPC migration, and Site Defense Platform expansion.</li>
<li>Establish clear priority frameworks that balance KTLO, roadmap commitments, incidents, business-critical migrations, and cross-functional asks.</li>
<li>Build and maintain a healthy operational model for critical infrastructure, including on-call quality, SLOs, dashboards, incident response, postmortems, and proactive maintenance.</li>
<li>Collaborate deeply with partner teams across Consumer, Ads, Media Platform, Graph QL, SEO, Web Platform, Info Sec, Compute, ML Platform, Storage, DevEx, Safety/Trust, CIAM, and Site Experience SRE.</li>
<li>Coach engineers through technical design, stakeholder alignment, project planning, execution risk, communication, and career growth.</li>
<li>Create a culture of metrics-driven quality, operational transparency, crisp ownership, and pragmatic technical ambition.</li>
</ul>
<h4><strong>Who You Might Be:</strong></h4>
<ul>
<li>3+ years of people-management experience for software, infrastructure, platform, SRE, or security/abuse-defense engineering teams.</li>
<li>5+ years of experience building or operating internet-scale infrastructure, distributed systems, networking, cloud platforms, service mesh, edge/CDN, or reliability-critical systems.</li>
<li>Strong technical judgment in areas such as Kubernetes, Envoy, Contour, xDS, gRPC, CDNs, DNS, load balancing, cloud networking, rate limiting, service discovery, traffic routing, and observability.</li>
<li>Experience managing operationally critical services, including: on-call, incident response, reliability review, SLOs, and production-change safety.</li>
<li>Proven ability to lead large migrations or platform adoption efforts across many customer teams.</li>
<li>Ability to partner with senior/staff engineers and translate between technical detail, user impact, team capacity, and executive-level risk.</li>
<li>Strong communication, prioritization, and stakeholder-management skills.</li>
<li>High empathy for engineers and internal customers, with the judgment to protect team focus while still serving the business.</li>
</ul>
<p><strong>Nice To Have</strong></p>
<ul>
<li>Experience with edge security, bot defense, DDoS mitigation, WAF, traffic reputation, Recaptcha/challenge systems, abuse detection, or fraud/risk infrastructure.</li>
<li>Experience with Cloudflare, Fastly, Akamai, or similar CDN/edge providers.</li>
<li>Experience with ML-enabled platform systems, online scoring, request annotation, signal pipelines, or false-positive-sensitive enforcement systems.</li>
<li>Experience with cost optimization for large-scale cloud networking or service-to-service infrastructure.</li>
<li>Hands-on software engineering experience in Go, Rust, Python, Java, C++, or similar production languages.</li>
</ul>
<p><strong>Benefits:</strong></p>
<ul>
<li>Comprehensive Healthcare Benefits and Income Replacement Programs</li>
<li>401k with Employer Match</li>
<li>Global Benefit programs that fit your lifestyle, from workspace to professional development to caregiving support</li>
<li>Family Planning Support</li>
<li>Gender-Affirming Care</li>
<li>Mental Health & Coaching Benefits</li>
<li>Flexible Vacation & Paid Volunteer Time Off</li>
<li>Generous Paid Parental Leave </li>
</ul>
<p>#LI-Remote</p><div class="content-pay-transparency"><div class
TalentyGo is an aggregator of job postings from public sources. Always verify information directly with the company. Applications go through the original company website; TalentyGo does not manage hiring processes.