Sr. Data Architect, Large Scale Distributed Systems
Proofpoint Inc
About Us:Proofpoint is a global leader in human- and agent-centric cybersecurity. We protect how people, data, and AI agents connect across email, cloud, and collaboration tools. Over 80 of the Fortune 100, 10,000 large enterprises, and millions of smaller organizations trust Proofpoint to stop threats, prevent data loss, and build resilience across their people and AI workflows. Our mission is simple: safeguard the digital world and empower people to work securely and confidently. Join us in our pursuit to defend data and protect people.How We Work:At Proofpoint you’ll be part of a global team that breaks barriers to redefine cybersecurity guided by our BRAVE core values: Bold in how we dream and innovateResponsive to feedback, challenges and opportunitiesAccountable for results and best in class outcomesVisionary in future focused problem-solvingExceptional in execution and impactRole OverviewWe are seeking an experienced Senior Architect to lead the design and evolution of enterprise-scale distributed systems supporting 50M+ connected sensors and high-volume event processing pipelines.This role is critical to building and operating mission-critical backend platforms that process millions of events per second across both synchronous and asynchronous architectures, with stringent requirements for scalability, reliability, security, and performance.The ideal candidate brings a proven track record of architecting and scaling production-grade systems at extreme scale, along with the ability to drive technical strategy, governance, and cross-functional alignment in a complex enterprise environment.Key ResponsibilitiesArchitecture & System DesignDefine and lead the architecture of large-scale distributed systems capable of ingesting and processing high-velocity data streams from 50M+ sensorsDesign resilient systems across synchronous (API-driven) and asynchronous (event-driven, streaming) paradigmsEstablish architectural standards for scalability, fault tolerance, and performance optimizationData Platform EngineeringArchitect real-time and batch data pipelines for high-throughput ingestion, transformation, and storageDrive design decisions across streaming, processing, and storage layers to ensure optimal performance and cost efficiencyEnable support for time-series, event-driven, and analytical workloadsTechnology Strategy & GovernanceDefine and enforce enterprise architecture principles, standards, and best practicesEvaluate and guide adoption of modern data technologies, including:Distributed messaging systems (e.g., Kafka, Pulsar)Scalable data stores (e.g., Cassandra, DynamoDB, Bigtable, ClickHouse, Elasticsearch)Stream and batch processing frameworks (e.g., Flink, Spark, Beam)Ensure alignment with security, compliance, and data governance requirementsScalability, Reliability & ObservabilityEstablish and operationalize SLAs, SLOs, and error budgetsDesign for high availability, multi-region resilience, and disaster recoveryImplement enterprise-grade observability frameworks (monitoring, logging, tracing)Leadership & CollaborationPartner with engineering, product, security, and data teams to align architecture with organizational objectivesProvide technical leadership, mentorship, and architectural oversight across multiple teamsLead design reviews and ensure adherence to architectural standardsRequired QualificationsExperience10+ years of experience in distributed systems and backend architectureDemonstrated success in scaling systems to:50M+ connected devices/sensors, orComparable high-scale environments (e.g., IoT, telecom, fintech, ad-tech, infrastructure platforms)Proven experience with high-throughput event-driven architectures in production environmentsTechnical ExpertiseDeep understanding of distributed systems concepts, including:CAP theorem, consistency models, and trade-offsPartitioning, replication, and sharding strategiesEvent delivery semantics (at-least-once, exactly-once, idempotency)Strong experience with:Streaming and messaging systems (Kafka, Pulsar, or equivalent)Real-time and batch processing frameworksScalable NoSQL and analytical data storesSystem Design & EngineeringExpertise in designing:Low-latency, high-throughput APIsEvent-driven and asynchronous processing systemsMulti-region, highly available architecturesStrong programming proficiency in one or more of: Go, Java, Scala, or RustPreferred QualificationsExperience with large-scale IoT or telemetry platformsFamiliarity with edge-to-cloud architecturesExperience operating in multi-cloud or hybrid environmentsKnowledge of enterprise security frameworks, data governance, and compliance (e.g., SOC2, ISO, GDPR)Exposure to AI/ML data pipelines and large-scale analytics platformsSuccess MetricsArchitecture supports billions of daily events with consistent performance and reliabilitySystems demonstrate horizontal scalability and fault isolationClear separation and optimization of real-time vs batch workloadsStrong adherence to enterprise architecture and governance standardsWhy Proofpoint?At Proofpoint, we believe that an exceptional career experience includes a comprehensive compensation and benefits package. Here are just a few reasons you’ll love working with us:Competitive compensationComprehensive benefitsCareer success on your termsFlexible work environmentAnnual wellness and community outreach daysAlways on recognition for your contributionsGlobal collaboration and networking opportunitiesOur Culture:Our culture is rooted in values that inspire belonging, empower purpose and drive success-every day, for everyone.We encourage applications from individuals of all backgrounds, experiences, and perspectives. If you need accommodation during the application or interview process, please reach out to View email address on click.appcast.io to ApplyInterested? Submit your application along with any supporting information- we can’t wait to hear from you!Consistent with Proofpoint values and applicable law, we provide the following information to promote pay transparency and equity. Our compensation reflects the cost of labor across several U.S. geographic markets, and we pay differently based on those defined markets as set out below. Pay within these ranges varies and depends on job-related knowledge, skills, and experience. The actual offer will be based on the individual candidate. The range provided may represent a candidate range and may not reflect the full range for an individual tenured employee. This role may be eligible for variable compensation and/or equity. We offer a competitive benefits package, including flexible time off, a comprehensive well-being program with two paid Wellbeing Days and two paid Volunteer Days per year, plus a three-week Work from Anywhere option.Base Pay Ranges:SF Bay Area, New York City Metro Area:Base Pay Range: 254,000.00 - 349,250.00 USDCalifornia (excludes SF Bay Area), Colorado, Connecticut, Illinois, Washington DC Metro, Maryland, Massachusetts, New Jersey, Texas, Washington, Virginia, and Alaska:Base Pay Range: 208,800.00 - 287,100.00 USDAll other cities and states excluding those listed above:Base Pay Range: 187,000.00 - 257,180.00 USDSummaryLocation: Sunnyvale, CA; Athens, Greece - Remote; IndiaType: Full time
- ...Mountain View is seeking a Staff / Principal ML Training Systems Engineer to lead the performance of large-scale multimodal training systems. This role involves... ...candidate will have proven experience in boosting distributed training performance and hands-on skills with...Senior
$272k - $431.25k
...models across multi-node distributed environments. Built in Rust... ...feel like a single system at datacenter scale. As large language models rapidly outgrow... ...-scale LLM inference.Architect and implement deep integrations... ..., RDMA/NVLink-based data planes, or KV-cache/CDN-like...SuggestedFull timeLocal areaRemote work- NVIDIA in Santa Clara, CA seeks a highly motivated systems architect to lead rack-scale factory and data center deployment flows for next-generation products. You will design end-to-end factory workflows, collaborate with data center architects, ODMs and OEMs, and mentor...Senior
- ...leading AI infrastructure company in California seeks a Member of Technical Staff — Training to design and optimize large-scale distributed training systems for frontier AI models. Candidates should have 5+ years of experience in ML systems and be proficient in Python...Suggested
- ...engineer to work on next‑generation technologies powering billions of users. You will explore information retrieval, distributed computing, large‑scale system design, and AI across multiple Google Ads teams, with opportunities to switch projects as you grow....Senior
$184k - $287.5k
...intelligence.NVIDIA is looking for a Datacenter System Architect to help define & design products for AI,... ...execute systems from concept through large production deploymentsMonitor... ...interaction between server architecture and at-scale datacenter deploymentHands on experience...SeniorFull time$272k - $431.25k
...as the technical focal point for rack-scale system SW/FW, working with CSP engineering teams... ...software, platform firmware, or large-scale distributed systems engineering. BS or MS in Computer... ...forefront of technological advancement.NVIDIA data center systems, such as DGX and HGX,...Full timeRemote workShift work$272k - $431.25k
...NVIDIA, as a Principal Rack Scale Systems Infrastructure Engineer, you... ...raising the engineering bar for large-scale networked systems,... ...architecture, system software, distributed systems, infrastructure control... ...workflows.Strong understanding of data center networking...Full timeRemote workShift work- ...computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of... ...are excited by the challenge of distributed training of large models on a large number of... ...of training generative AI at scale.THE PERSON:The ideal candidate...
- ...leading AI infrastructure company in California is seeking a Member of Technical Staff — Inference to design and optimize large-scale AI inference systems. The role demands 5+ years in systems engineering and expertise in large-scale inference systems. Successful...SeniorFlexible hours
$160.36k - $240.54k
...clear path to AVs at commercial scale, empowering a safer, richer,... ...to build/scale Nuro's large-scale computing infrastructure in the cloud/data center. This system is the foundation of many critical... ...building and developing large-scale distributed applications (e.g. Kubernetes...SeniorFull time$200k - $400k
...designs and operates ultra-scale GPU supercomputing systems to train next-generation foundation... ...communication performance, distributed reliability, and cross-layer optimization for large-scale training workloads.... ...across thousands of GPUs • Architect fault-tolerant distributed...SeniorVisa sponsorship$184k - $287.5k
...lives. We’re searching for a Senior Systems Software Engineer with deep expertise in distributed systems, Kubernetes, containers,... ...own hard technical problems at large scale and help shape how AI... ...accelerated runtime stack (control and data planes), including NVIDIA components...SeniorFull timeRemote work- ...building next-generation generalist robots and is seeking a Staff / Principal ML Training Systems Engineer to own training systems performance end-to-end. You will optimize large-scale multimodal training, define parallelism strategies, and drive efficiency across GPUs,...
- Applied Intuition, Inc. is seeking a performance engineer to accelerate large-scale ML workloads in the data center. You will own profiling, optimization, and cost-efficiency for distributed training and large offline inferences. You will work across accelerators, ML frameworks...
$174k - $252k
Write and test product or system development code in Java, C++... ...efficient, next-generation distributed caching and database... ...technologies for global Ads data storage and scaling systems.Review code developed... ...experience with developing large-scale infrastructure, distributed...Senior- ...View, California. This role offers the opportunity to provide technical leadership for advanced recommendation systems and AI initiatives, guiding large-scale projects across teams. Ideal candidates should have over 8 years of experience in machine learning and recommendation...
$153.2k - $234.1k
...from breakthrough hardware and battery systems to intuitive design, intelligent... ...future of transportation on a global scale.Role Overview:Are you passionate about... ...bring:3+ years of experience building large-scale distributed systems/applications or advanced ML Applications...SeniorFull timeWork at officeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours- ...with all of their business systems through natural language... ..., backed by the global scale of ServiceNow and the... ...responsibilities including distributed training and inference pipeline for large language models(LLM),... ...machine learning team, data infrastructure team and...SeniorWork at officeRemote workFlexible hours
- ...is seeking a seasoned software/solutions architect to mentor teams and lead the architecture of highly scalable distributed systems. You will identify performance... ...security, and engineering excellence across large-scale data plane platforms, with opportunities for...Senior
- ...firm in California seeks a Senior IC to advance AI/ML infrastructure for large models. This role involves collaborating with global teams to enhance simulation realism and designing distributed systems for ML lifecycles. Candidates should have over 5 years of software...Senior
$150k
...alongside world-class researchers, data scientists, and engineers,... ...pioneers. The Role The Distributed ML Engineer will play a role... ...develop new and cutting-edge systems. The ideal candidate will... ...coding, debug methodologies, and large-scale machine learning experience....Full timeWork experience placementVisa sponsorship- Google DeepMind in Mountain View, CA is seeking a Senior Staff Research Scientist to advance recommendation systems and large-scale AI research. You will design experiments, prototype architectures, and contribute to product and infrastructure improvements. The role emphasizes...Senior
$320k
NVIDIA data center systems, such as DGX and HGX, have become core to NVIDIA... ...for a strong technical architect to own the end-to-end architecture... ...direct authority in large-scale, collaborative environments... ...storage architectures and distributed parallel processing paradigmsNVIDIA...Full timeShift work- ...capability focused on frontier models. This role owns the infrastructure for large-scale training, RL experiments, and production-grade workflows. You will work at the intersection of distributed systems, GPU performance, and ML framework integration. The role requires...
- ...join the Build Environments and Tools team, focusing on scalable build infrastructure for large-scale C/C++ builds. You will design, develop, and maintain build orchestration systems, optimize build speeds, and mentor peers while partnering with engineering leads to...Senior
$174.9k - $261.3k
...understand the world!The Data Labeling Engineering... ...reliable training data at scale. Our tools and platform... ..., and direct impact on systems that unblock the next... ...experience building robust distributed platforms and... ...platformsor tools used by large labeling workforces (e....SeniorFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours- ...computing environment. We are seeking an HPC Systems Administrator who will steward the... ...nodes, high-density GPU racks, and petabyte-scale storage. The role requires strong Linux system... ...expertise, plus physical rigour for data center work and the ability to lift up to...Senior
- ...change how billions of users connect, explore, and interact with information and one another. Our products handle information at massive scale and extend beyond web search, with opportunities to switch teams as you grow. We’re looking for engineers who bring fresh ideas...Senior
- Bright Vision Technologies is seeking a Machine Learning Data Engineer to build and operate large-scale data systems powering AI training and evaluation pipelines. The role combines deep data engineering expertise with a strong understanding of AI workloads, focusing on...Remote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr. Data Architect, Large Scale Distributed Systems. Be the first to apply!
- data center architect Sunnyvale, CA
- database designer Sunnyvale, CA
- data architect Sunnyvale, CA
- senior associate architect Sunnyvale, CA
- senior dynamics crm developer Sunnyvale, CA
- senior application security Sunnyvale, CA
- senior account director Sunnyvale, CA
- senior plumbing designer Sunnyvale, CA
- senior ux designer Sunnyvale, CA
- senior cloud data engineer Sunnyvale, CA


