What are the responsibilities and job description for the Platform Support Engineer position at Render Networks?
About the Role
Reporting to the Lead DevOps/SRE Engineer, the Platform Support Engineer is responsible for supporting the day-to-day operation, reliability and supportability of Render's cloud platform and engineering systems.
This is an operationally focused role within the DevOps/SRE team, providing hands-on support across production systems, customer-impacting issues, internal engineering platforms and employee technology enablement.
You will work closely with Product Engineering, Customer Success and internal teams to investigate issues, support incident response, improve operational processes and ensure Render's SaaS platform remains secure, reliable and easy to operate.
As part of the DevOps/SRE team, you will participate in a 24x7 on-call roster to support production incidents affecting our SaaS platform. You'll play a key role in restoring services quickly, communicating effectively during incidents and helping drive long-term improvements that enhance platform reliability and operational excellence.
The role is ideal for an engineer who enjoys troubleshooting complex technical problems, automating repetitive tasks, improving operational processes and developing deeper expertise in cloud infrastructure, reliability engineering and SaaS operations.
What You'll Do
- Provide operational support for Render's production SaaS platform, assisting with investigation, troubleshooting and resolution of platform issues.
- Participate in a 24x7 on-call roster to respond to production incidents and customer-impacting outages, ensuring timely service restoration.
- Act as the primary operational contact for production support and customer-impacting issues during US business hours, coordinating with engineering teams globally when escalation is required.
- Participate in incident response activities, including triage, investigation, communication and post-incident follow-up.
- Monitor platform health and respond to alerts, operational events and customer-impacting issues.
- Support customer escalations by investigating technical issues across application, infrastructure and cloud environments.
- Assist with the operation and maintenance of Render's AWS cloud environments, including serverless workloads, databases, networking and security services.
- Support Infrastructure as Code practices using Cloudformation & Terraform, contributing to infrastructure changes and operational improvements.
- Maintain and improve operational tooling, scripts and automation to reduce manual effort and improve reliability.
- Support CI/CD operations, including troubleshooting Buildkite pipelines and assisting engineering teams with software delivery issues.
- Assist with observability and monitoring practices using tools such as Datadog.
- Support operational activities for platform services including Confluent Cloud and Databricks.
- Contribute to documentation, runbooks, knowledge sharing and continuous improvement initiatives.
- Support onboarding and offboarding of US employees, including access management, account provisioning and IT operational tasks.
- Assist with security and compliance activities supporting ISO 27001 controls and operational requirements.
- Work with engineering teams to improve developer experience, operational maturity and platform reliability.
What You'll Bring
- 2 years of experience in DevOps, Platform Engineering, Cloud Engineering, Systems Administration or a similar technical operations role.
- Experience supporting production software systems and troubleshooting technical issues in a SaaS or cloud environment.
- Willingness to participate in a 24x7 on-call support rotation and respond to production incidents outside normal business hours when rostered.
- Hands-on experience with cloud platforms, preferably AWS.
- Understanding of Infrastructure as Code concepts and experience with tools such as Terraform.
- Experience working with Linux-based systems and cloud-native technologies.
- Familiarity with CI/CD pipelines and software delivery practices.
- Experience with monitoring, logging and observability tools.
- Ability to investigate issues across multiple layers of a technology stack, from applications through infrastructure.
- Strong troubleshooting mindset and ability to work through ambiguous technical problems.
- Good understanding of networking fundamentals, security principles and identity/access management.
- Ability to participate in incident response and production support activities.
- Strong communication skills with the ability to work collaboratively with engineering, product and customer-facing teams.
- A willingness to learn, automate and continuously improve operational processes.
Nice To Have
- Experience with AWS services such as ECS, Lambda, API Gateway, RDS, S3, CloudFront or IAM.
- Experience with Terraform or other Infrastructure as Code tools.
- Experience supporting event-driven systems or data platforms.
- Experience with Datadog or similar observability platforms.
- Experience with Python, Bash or another scripting language for automation.
- Experience with IT administration, identity management or employee onboarding workflows.
- Experience working in environments with security or compliance requirements such as ISO 27001.
Salary : $85,000 - $100,000