Search jobs now Find the right job type for you Create a job alert Explore how we help job seekers Contract talent Permanent talent Learn how we work with you Executive search Finance and Accounting Technology Marketing and Creative Legal Administrative and Customer Support Technology Risk, Audit and Compliance Finance and Accounting Digital, Marketing and Customer Experience Legal Operations Human Resources 2026 Salary Guide Demand for Skilled Talent Report Job Market Outlook Press Room Tech insights Labor market overview AI in recruiting Navigating the AI era Staffing for small businesses Cost of a bad hire Browse jobs Find your next hire Our locations

Add your latest resume to match with open positions.

6 results for Site Reliability Engineer in San Francisco, CA

Site Reliability Engineer (SRE)
  • San Francisco, CA
  • onsite
  • Permanent / Full Time
  • 170000 - 180000 USD / Yearly
  • We are looking for an experienced Site Reliability Engineer to strengthen the reliability, scalability, and operational maturity of our platform in San Francisco, California. This role will focus on improving service health, refining observability, and partnering with engineering teams to build systems that perform consistently under real-world demand. The ideal candidate brings deep production experience, a strong automation mindset, and a practical approach to incident response and continuous improvement.<br><br>Responsibilities:<br>• Establish measurable reliability standards for critical services by creating and maintaining service indicators, objectives, and error budget practices.<br>• Take ownership of production stability by monitoring uptime, latency, and availability, and driving improvements that reduce operational risk.<br>• Lead live incident response efforts, coordinate troubleshooting during outages, and ensure issues are resolved efficiently and thoroughly.<br>• Run blameless post-incident reviews, document findings clearly, and track corrective actions through completion.<br>• Design and enhance observability across logs, metrics, and distributed tracing using tools such as Datadog, CloudWatch, Grafana, OpenTelemetry, and Sentry.<br>• Improve alert quality and dashboard design so engineering teams can quickly identify meaningful system issues without unnecessary noise.<br>• Evaluate system behavior under load, uncover performance constraints, and recommend changes that improve scalability and resource efficiency.<br>• Build automation and internal tooling that streamline operational work, strengthen deployment safety, and support incident management, debugging, and capacity planning.<br>• Contribute to infrastructure and delivery workflows across AWS, Terraform, Ansible, Linux, and GitHub Actions with a focus on dependable releases and resilient systems.<br>• Partner with security and compliance stakeholders to support operational standards, audit readiness, and the integration of monitoring into broader engineering practices.
  • 2026-06-22T00:00:00Z
Sr DevOps Engineer
  • San Francisco, CA
  • onsite
  • Temporary / Contract
  • 75 - 85 USD / Hourly
  • <p>We are looking for a Sr. DevOps Engineer to help strengthen and scale a cloud-based platform in San Francisco, California. This opportunity is ideal for a senior engineer who can improve infrastructure performance, streamline delivery processes, and support production systems in a hybrid work environment. The role partners closely with engineering teams to build dependable deployment practices, enhance platform visibility, and maintain stable operations across modern cloud services. </p><p><br></p><p> This is a contract to permanent opportunity and requires 2 days a week on site. </p><p><br></p><p> Responsibilities: </p><p>• Develop, manage, and optimize cloud infrastructure that supports the company&#39;s live SaaS environment. </p><p>• Create and maintain delivery pipelines for application services and machine learning workloads to enable efficient releases. </p><p>• Oversee deployments across serverless components and container-based applications while ensuring consistency and reliability.</p><p>• Administer and support a PostgreSQL data environment, including performance, availability, and operational upkeep. </p><p>• Strengthen system dependability by improving monitoring, alerting, logging, and overall platform observability. </p><p>• Automate provisioning and release processes through infrastructure-as-code practices using tools such as Terraform. </p><p>• Collaborate with backend and machine learning teams to address infrastructure needs and enable smooth deployment workflows. </p><p>• Participate in production support activities, including incident response and an on-call rotation. </p><p>• Contribute to operational improvements that reduce manual effort and increase platform stability over time.</p>
  • 2026-07-13T00:00:00Z
DevOps Engineer
  • San Francisco, CA
  • onsite
  • Permanent / Full Time
  • 170000 - 180000 USD / Yearly
  • <p>We are looking for a Sr.DevOps Engineer to strengthen our engineering organization in San Francisco, California. In this role, you will improve the foundation that supports application delivery, platform reliability, and developer productivity across cloud-based systems. You will work closely with engineering partners to create scalable infrastructure, streamline release processes, and build dependable operational practices that support continued growth.</p><p><br></p><p>Responsibilities:</p><p>• Architect and maintain cloud infrastructure using infrastructure-as-code practices that promote consistency, scalability, and operational simplicity.</p><p>• Create and refine automated delivery pipelines to support dependable, efficient releases across multiple services and environments.</p><p>• Establish repeatable deployment approaches for serverless applications, container-based platforms, and orchestrated workflows.</p><p>• Build internal automation and developer tools that reduce manual effort and help backend and machine learning teams work more efficiently.</p><p>• Enhance development, testing, and staging environments to improve workflow speed and reduce friction during software delivery.</p><p>• Define strong monitoring and observability practices across logs, metrics, and tracing to improve system visibility and troubleshooting.</p><p>• Investigate reliability, latency, and performance issues in distributed environments and partner with engineers to implement lasting improvements.</p><p>• Strengthen operational readiness by supporting incident response, contributing to post-incident reviews, and turning findings into better tooling and processes.</p><p>• Integrate security and compliance controls into infrastructure and deployment workflows, including access management, patching, and vulnerability remediation.</p>
  • 2026-06-22T00:00:00Z
Software Engineer II
  • San Ramon, CA
  • remote
  • Permanent / Full Time
  • 85000 - 124000 USD / Yearly
  • <p>We are looking for a talented and motivated <strong>Software Engineer II</strong> to join our innovative IT team. As a Software Engineer II, you will be responsible for designing, developing, and maintaining software applications using modern technologies. You will work on both front-end and back-end development, ensuring seamless integration and optimal performance. Your role will involve collaborating with cross-functional teams to deliver high-quality software solutions that meet business requirements.</p><p><br></p><p>This role offers the opportunity to work on cutting-edge software solutions, improve system efficiency, and contribute to strategic IT initiatives. If you are passionate about software engineering and thrive in a dynamic environment, we’d love to hear from you!</p><p><br></p><p><strong>What You&#39;ll Do</strong></p><ul><li>Develop and maintain web applications using modern front-end frameworks such as React or Angular.</li><li>Design and implement RESTful APIs and microservices using C#.</li><li>Work with cloud platforms like Azure or AWS to deploy and manage applications.</li><li>Implement CI/CD pipelines using tools like Azure DevOps or GitHub Actions.</li><li>Collaborate with UX/UI designers to create intuitive and responsive user interfaces.</li><li>Write unit and integration tests to ensure code quality and reliability.</li><li>Participate in code reviews and provide constructive feedback to peers.</li><li>Troubleshoot and resolve software defects and production issues.</li><li>Design and contribute to an AI agent platform, enabling intelligent automation, workflow orchestration, and integration with enterprise systems.</li></ul><p><br></p>
  • 2026-07-06T00:00:00Z
AI Engineer (AI Data Privacy, Security & Governance)
  • San Ramon, CA
  • remote
  • Permanent / Full Time
  • 85000 - 124000 USD / Yearly
  • <p>obert Half is seeking a Software Engineer II – AI Engineer to analyze, design, program, debug, test, implement, and support the privacy, security, and governance controls that protect generative AI technologies. This role safeguards GenAI-enabled applications, including LLM-powered workflows, RAG pipelines, plugins, skills, and autonomous agents, by embedding data protection, security guardrails, and responsible AI governance across the development lifecycle. </p><p> </p><p>This role supports SDLC documentation across all phases, with a focus on data privacy, security guardrails, governance, compliance, and risk management. It also works with users to define requirements and support applications in production. </p><p><strong>What You’ll Do</strong> </p><ul><li>Design and implement data privacy controls, including PII/PHI detection, redaction, and data minimization. </li><li>Build security guardrails, including prompt-injection defense, jailbreak prevention, and output filtering. </li><li>Establish governance for plugins, skills, and agents, including registration, approval, and lifecycle management. </li><li>Review and vet third-party plugins, skills, and agent tools for security, privacy, and compliance risks. </li><li>Define and enforce access controls, authentication, and least-privilege permissions for AI components. </li><li>Implement guardrails for autonomous agents, including action scoping, tool-use restrictions, and human-in-the-loop approval. </li><li>Apply data classification, retention, and residency policies across GenAI data flows. </li><li>Monitor AI systems for policy violations, data leakage, and anomalous plugin, skill, or agent behavior. </li><li>Maintain audit trails and logging for plugin, skill, and agent activity. </li><li>Support compliance with regulations and frameworks, including GDPR, CCPA, and the EU AI Act. </li><li>Provide Level II production support for security, privacy, and governance incidents. </li><li>Support incident response, including containment, escalation, and remediation. </li></ul>
  • 2026-07-10T00:00:00Z
Software Engineer II – AI Engineer (w/ skills in Deployment)
  • San Ramon, CA
  • remote
  • Permanent / Full Time
  • 85000 - 124000 USD / Yearly
  • <p>Robert Half is seeking a Software Engineer II – AI Engineer who will analyze, design, program, debug, test, implement, deploy, and support software enhancements and new applications using Generative AI technologies. This role contributes to the development and production deployment of GenAI-enabled applications, including LLM-powered workflows, RAG pipelines, and AI-driven user experiences. </p><p> </p><p>This role supports SDLC documentation across all phases, with a focus on deployment, evaluation, observability, safety, and monitoring. It also interacts with users to define requirements and support applications in production. </p><p><strong>What You’ll Do</strong> </p><ul><li>Develop and modify application modules, including GenAI components. </li><li>Build prompt workflows, retrieval layers, APIs, and cloud services. </li><li>Troubleshoot production issues, including latency, hallucinations, and errors. </li><li>Provide Level II production support for deployed systems. </li><li>Design components, including LLM integrations and RAG pipelines. </li><li>Implement CI/CD pipelines, containerization, and release processes. </li><li>Develop RAG pipelines with embeddings, chunking, and vector search. </li><li>Apply prompt engineering techniques, including few-shot prompting and structured outputs. </li><li>Evaluate models for accuracy, relevance, and hallucination risk. </li><li>Implement safety guardrails, including PII protection and prompt-injection defense. </li><li>Execute testing, including unit, integration, and GenAI evaluation testing. </li><li>Monitor production systems for latency, cost, usage, and errors. </li><li>Support incident management with fallback and recovery strategies. </li></ul><p><br></p>
  • 2026-07-10T00:00:00Z