Skip to content
← Back to job listings

Operations Analyst

Hippocratic AI · Menlo Park, CA, United States

External listingfull-time22 days ago

About The Role

Join Hippocratic AI as an Operations Analyst, where you will play a critical role in ensuring the continuous operation of our production systems, integrations, and customer/partner environments. You will monitor system alerts, perform proactive maintenance, resolve common operational issues, and triage advanced issues to the appropriate teams. Your work will be essential in minimizing customer and partner downtime, maintaining trust, and ensuring our AI agents and supporting systems operate smoothly at all times.

  • Monitor system alerts, integrations, and operational reports; perform proactive maintenance; resolve common operational issues.
  • Coordinate incident response activities, track progress to resolution, and ensure clear internal handoffs during escalations.
  • Build and maintain scripts and automation to monitor system health, validate integrations, and generate customer- or partner-specific reports.
  • Familiarity with alerting and monitoring tools (e.g., Datadog, New Relic, CloudWatch, Prometheus, Grafana, PagerDuty, Opsgenie, or similar)
  • Bachelor’s degree in Computer Science, Health Informatics, Information Systems, or a related field
  • 2+ years of experience in operations, site reliability, NOC, technical support, or production monitoring roles
  • Strong understanding of incident management processes, escalation procedures, and SLA-driven operations
  • Bachelor’s degree in Information Systems, Computer Science, Operations, Engineering, or a related field (or equivalent practical experience)
  • Strong sense of ownership, reliability, and attention to detail
  • Hands-on experience monitoring production systems, integrations, APIs, or data pipelines in a 24×7 environment
  • Experience writing scripts or automation using tools/languages such as Python, Bash, SQL, or similar
  • Excellent organizational skills with the ability to manage multiple alerts, issues, and priorities simultaneously
  • Ability to troubleshoot common system, integration, and data-flow issues using logs, metrics, and dashboards
  • Clear written and verbal communication skills, especially during high-pressure incidents
  • Experience supporting cloud-based platforms (AWS, Azure, or Google Cloud)
  • Familiarity with REST APIs, webhooks, message queues, or integration workflows
  • Experience in healthcare, regulated environments, or HIPAA-compliant systems
  • Exposure to CI/CD pipelines, deployment monitoring, or change management processes
  • Background in Site Reliability Engineering (SRE), DevOps, or production support for SaaS platforms
  • Experience creating customer-facing operational or SLA reports
  • Experience supporting AI/ML platforms, data pipelines, or real-time systems

This is an external listing. JobSpring does not represent or verify the employer. Report this listing