Overview
Build and lead Anthropic’s caching infrastructure as a managed service, including a managed Redis fleet and CDC-driven cache invalidation.
What you'll do
- Drive the technical direction for caching infrastructure used across Product and Research.
- Design, build, and operate a managed Redis fleet to scale for Claude’s ecosystem.
- Build client libraries and developer abstractions to make correct caching the default.
- Design and operate CDC-driven cache invalidation to keep cached data consistent.
- Architect caching solutions across GCP, AWS, first-party deployments, and other environments.
- Optimize latency, hit rates, reliability, and cost efficiency on critical paths.
- Develop observability and tooling to understand and debug cache behavior.
What you'll need
- Significant experience building and operating production distributed systems.
- Deep knowledge of caching architectures, including invalidation, consistency tradeoffs, and failure modes.
- Experience operating Redis, Memcached, or similar in-memory data stores in production.
- Proficiency in at least one systems programming language (Go, Rust, Java, C++, or Python at scale).
- Track record leading large, complex infrastructure projects as an engineer or tech lead.
- Ability to balance moving quickly with production reliability needs.
- Strong technical leadership and cross-functional collaboration skills.
Details
- Location-based hybrid policy: staff are expected to be in one of the company offices at least 25% of the time.
- Visa sponsorship is available; the company sponsors visas and makes reasonable efforts if an offer is made.
Read the full description and apply on the company’s own careers page.