Skip to content
All jobs

Senior/Lead Software Engineer, Site Reliability (Agentforce Operations)

  • Salesforce
  • 3 Locations, United States of America
  • Full time

To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts. Job Category Software Engineering Job Details About Salesforce Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here, ambition meets action. Tech meets trust. And innovation isn’t a buzzword — it’s a way of life. The world of work as we know it is changing and we're looking for Trailblazers who are passionate about bettering business and the world through AI, driving innovation, and keeping Salesforce's core values at the heart of it all. Ready to level-up your career at the company leading workforce transformation in the agentic era? You’re in the right place! Agentforce is the future of AI, and you are the future of Salesforce. The Experience Global supply chains still rely on slow, manual processes—email, spreadsheets, and fragmented data. The economy moves fast but supply chains don't, creating an inefficiency that affects the $13T of goods shipped annually and is one of the largest untapped opportunities in modern enterprise. Agentforce Operations is reimagining the supply chain with an AI-powered platform for designing, automating, and running end-to-end business processes, with seamless collaboration through familiar channels like email. For Salesforce, this represents a massive growth opportunity in the back office, with innovations that flow into the front office. Customers are clamoring for more, rapidly expanding their use cases as we enter an exhilarating growth phase. As one user put it: "I've been waiting for this for 20 years." The Teams Agentforce Operations & Missionforce Team As a Site Reliability Engineer for Missionforce Operations, you will be a senior technical contributor who helps shape the infrastructure, process automation and engineering practices behind the product. You will partner closely with product engineers, sector experts, and forward-deployed engineers to deliver the most reliable systems for mission-critical public-sector workflows. Missionforce Operations is Agentforce Operations’ public-sector counterpart: rebuilt from the ground up to bring core capabilities to some of the world’s most sensitive operating environments. In this role, you will design and ship infrastructure, automation, and cost-efficient platform capabilities that enable customers to run Missionforce Operations reliably and securely in customer-owned, isolated public-cloud environments, where requirements for trust, security, and data classification are uncompromising. Super Agent Team As a Site Reliability Engineer for the Agentforce Operations Super Agent team, your mandate will be to radically optimize the scalability, resilience, and performance of our product and platform across the entire stack to help usher in this next phase of growth. The ideal candidate would be able to develop a deep technical understanding of our architecture and then deliver the end-to-end implementation of optimizations that would enable us to achieve greater scalability while simultaneously improving our reliability. Agentforce Operations manages complex, infinitely configurable states with equally complex state transitions and interfaces to teams across the organization. That means you'd be tackling a wide array of challenges that would include optimization of our existing data model, streamlining of our asynchronous operations patterns, profiling and rightsizing our infrastructure, and developing both the tools and observability to run experiments and empirically validate results. What You'll Actually Be Doing Agentforce Operations & Missionforce Team Scaling & Reliability: Own the reliability roadmap for major product areas, evolving startup-speed architectures into highly available, globally scalable systems for mission-critical customer-operated cloud environments. Collaborative Leadership: Partner with engineers to refine our infrastructure strategy, contributing senior-level perspectives on system design, capacity planning, bottleneck identification, and security. Infrastructure as Code: Define, maintain and evolve our tools for deployments, focusing on making deployments as “push-button” and robust as possible. AI Operations (AIOps): Support the scaling and deployment of our AI/ML infrastructure, ensuring the compute, deployment and platform capabilities required to run AI features reliably and cost-effectively. Production Excellence: Design and build the internal systems and processes for testing and releasing the product, along with the tooling and automation that enable customers and partners to deploy, upgrade, diagnose, and run Missionforce Operations successfully with minimal direct engineering involvement. Security Engineering: Design, review, and strengthen identity and access controls, networking, workload isolation, secret management, deployment security, and policy enforcement to ensure the product is secure by default. AI-First Workflow: Lean into the future of engineering by using AI tools to automate routine operational tasks and accelerate infrastructure delivery. Super Agent Diagnose and prioritize architectural problems across our distributed backend systems, contributing technical perspective to help the team invest in the right foundational work at the right time Design and implement solutions to hard system problems including application performance optimization, infrastructure management, and data model improvements. Raise and maintain our engineering quality bar through pragmatic standards, guardrails, and team practices that prevent regressions and keep our product healthy and performant. Collaborate with peers across customer, product, and engineering to deeply understand customer needs, make pragmatic tradeoffs that maximize impact, and align on engineering reality Mentor engineers as a senior peer and force multiplier, sharing expertise across backend architecture and system design while helping shape stronger engineering judgment, more effective collaboration, and a higher-performing team You're Our Person If... Agentforce Operations & Missionforce Team Proven Production Experience: Experience supporting mission-critical production systems and helping high-growth products navigate the technical debt and architectural changes required to scale reliably. Technical Breadth: Strong proficiency in Kubernetes, Terraform/OpenTofu, and AWS/GCP/Azure. Coding Mastery: Ability to write and review production-level code in Golang, TypeScript, or Python—you view automation as a software engineering problem. Systems Expert: Deep understanding of distributed systems, including how to debug complex interactions between microservices, databases, and AI agents. Low-Ego Collaboration: Experience working within a senior team of Principal engineers, capable of both leading specific initiatives and supporting the broader group’s technical vision. Mentorship Mindset: Enthusiasm for learning and growing as an engineer, and for helping your peers do the same. AI Fluency: A demonstrated ability to use modern AI development tools to move faster and build more reliable systems. Drive: Ability to work independently and collaboratively in a fast-paced startup environment. Excellent Communicator: Strong written and verbal communication skills, with the ability to communicate effectively with people from varied technical backgrounds. Super Agent 5+ years of experience in SRE, Production Engineering, or Backend Engineering with a heavy focus on operations and infrastructure. Mastery of Golang, GraphQL, and PSQL, with demonstrated experience delivering high-performance optimizations. Experience with RDS, Redis/Elasticache, and EKS. Deep expertise in multiple areas of backend engineering such as data modeling, db performance tuning, state management, API design, transactionality, concurrency, memory management, fault tolerance, and scaling, with the ability to pick up new technologies quickly Repeated ownership of foundational system work in complex distributed systems, from diagnosing architectural problems to designing solutions to driving adoption across a team. Excellent writing and speaking skills and ability to lead real-time technical-design discussions Demonstrated ownership and accountability, with a consistent track record of high standards, continuous improvement, and low-ego collaboration Even Better If... Agentforce Operations & Missionforce Team Advanced Degree in Computer Science or equivalent practical experience. Compliance and Regulated Industries: Experience building products for regulated industries, particularly the public sector or environments with strong security, compliance, and data-sovereignty requirements. Familiarity with developing for classified or limited-connectivity environments, including Department of Defense Impact Levels such as IL6. Advanced knowledge of microservice orchestration and durability patterns, including hands-on experience with Temporal for workflow reliability and service mesh for secure, observable service-to-service communication in high-growth SaaS environments. Deep knowledge of networking, security, and identity management within major cloud providers. Experience building AI products for supply chain, logistics, manufacturing, or operational workflows. Super Agent B.S. in Computer Science (M.S. preferred) Exposure to Temporal, Istio, or Typesense Experience with CI/CD systems, specifically Jenkins and Spinnaker. Exposure to the supply chain, logistics, or manufacturing industry. (We will teach you otherwise!) Familiarity with the Salesforce platform Experience with workflow engines Unleash Your Potential When you join Salesforce, you’ll be limitless in all areas of your life. Our benefits and resources support you to find balance and be your best, and our AI agents accelerate your impact so you can do your best. Together, we’ll bring the power of Agentforce to organizations of all sizes and deliver amazing experiences that customers love. Apply today to not only shape the future — but to redefine what’s possible — for yourself, for AI, and the world. Accommodations If you need a reasonable accommodation during the application or the recruiting process, please submit a request via this Accommodations Request Form. Please note that Salesforce uses artificial intelligence (AI) tools to help our recruiters assess and evaluate candidates’ resumes and qualifications throughout the recruiting process. Humans will always make any candidate selection and hiring decisions. Please see our Candidate Privacy Statement for more information about how we use your personal data and your rights, including with regard to use of AI tools and opt out options. Posting Statement Salesforce is an equal opportunity employer and maintains a policy of non-discrimination with all employees and applicants for employment. What does that mean exactly? It means that at Salesforce, we believe in equality for all. And we believe we can lead the path to equality in part by creating a workplace that’s inclusive, and free from discrimination. Know your rights: workplace discrimination is illegal. Any employee or potential employee will be assessed on the basis of merit, competence and qualifications – without regard to race, religion, color, national origin, sex, sexual orientation, gender expression or identity, transgender status, age, disability, veteran or marital status, political viewpoint, or other classifications protected by law. This policy applies to current and prospective employees, no matter where they are in their Salesforce employment journey. It also applies to recruiting, hiring, job assignment, compensation, promotion, benefits, training, assessment of job performance, discipline, termination, and everything in between. Recruiting, hiring, and promotion decisions at Salesforce are fair and based on merit. The same goes for compensation, benefits, promotions, trans