Accela

Accela

Accela is the leading provider of productivity and civic engagement software solutions for local, state, and federal government agencies. Their platform offers cloud-based solutions for permitting, licensing, code enforcement, and service request manag...

Internet Software & Services
251-1K
Founded 1999
$234M raised

Description

  • Operate, maintain, and continuously improve Accela's production cloud environments.
  • Support platform modernization efforts, including containerization, cloud-native technologies, and automation.
  • Monitor platform health, availability, performance, and capacity using observability and monitoring tools.
  • Participate in incident response, troubleshoot production issues, and contribute to root cause analysis.
  • Develop and maintain automation, tooling, and scripts to improve reliability, scalability, deployment efficiency, and operational effectiveness.
  • Support the implementation and monitoring of SLOs, SLAs, and operational metrics.
  • Partner with Development, DevOps, Database Engineering, and Security teams to resolve reliability, performance, and scalability issues.
  • Assist with platform deployments, operational readiness reviews, and change management activities.
  • Contribute to observability through monitoring, logging, metrics collection, and distributed tracing.
  • Support compliance-related operational activities for SOC 2, HIPAA, FedRAMP, StateRAMP, and PCI-DSS environments.
  • Participate in post-incident reviews and corrective or preventive actions to improve platform stability.

Requirements

  • 4+ years of experience in Site Reliability Engineering, Cloud Operations, Systems Engineering, DevOps, Software Engineering, or a related technical discipline.
  • Experience supporting cloud-based SaaS environments, preferably within Microsoft Azure.
  • Experience with Kubernetes and containerized application environments.
  • Working knowledge of scripting and automation using Python, PowerShell, Bash, or similar languages.
  • Experience troubleshooting distributed systems across application, infrastructure, networking, and operating system layers.
  • Familiarity with monitoring, logging, metrics, and observability platforms.
  • Strong analytical and problem-solving skills with a structured approach to troubleshooting and root cause analysis.
  • Experience working within Incident, Problem, and Change Management processes.
  • Strong written and verbal communication skills and the ability to work effectively with cross-functional teams.
  • Experience using Git and GitHub-based workflows.
  • Experience with Infrastructure-as-Code tools such as Terraform (desired).
  • Experience with configuration management tools such as Ansible (desired).
  • Experience supporting SaaS platforms operating under compliance frameworks such as SOC 2, HIPAA, FedRAMP, StateRAMP, or PCI-DSS (desired).
  • Experience with GitOps deployment methodologies, including Argo CD or Flux (desired).
  • Familiarity with OpenTelemetry and modern observability practices (desired).
  • Exposure to cloud cost optimization and FinOps concepts (desired).
  • Linux systems administration experience (desired).
  • Experience using AI-assisted engineering tools such as GitHub Copilot, Claude Code, or similar technologies (desired).
  • Up to 10% travel for team meetings, training, and business needs.

Benefits

  • Annual base salary range of $125,000 to $145,000.
  • Eligible for an annual discretionary bonus target.
  • Flexible time off.
  • Comprehensive medical, dental, and vision plans.
  • Family planning benefits.
  • 401(k) retirement savings plan with company match.
  • Health savings account with company contributions.
  • Flexible spending account.
  • Life, accident, and disability coverage.
  • Business travel insurance.
  • Employee assistance programs and other well-being benefits.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its document workflow platform reliable, resilient, and available for customers.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
16 hours, 23 minutes ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its document workflow platform reliable through incident management, observability, production support, and resilience work across services.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
16 hours, 23 minutes ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to keep its document workflow platform highly available and resilient while supporting production operations and reliability improvements.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
1 day, 15 hours ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its production document workflow platform reliable, resilient, and low-downtime for customers.

AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
1 day, 16 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers