Demo

Site Reliability Engineer -Jersey City, NJ & Dallas, TX

StradIT
Jersey, NJ Remote Full Time
POSTED ON 8/27/2026
AVAILABLE BEFORE 10/27/2026

Role: Site Reliability Engineer

Experience: Min 10 Years

Locations: Jersey City, NJ & Dallas, TX

Work mode: Hybrid

Employment: W2

The Application Support Engineering role advances Site Reliability Engineering (SRE) practices for applications running in production. The role scope includes Application Support Engineers, SDETs, Software Engineers, and SREs focused on improving reliability, observability, recovery, automation, and operational readiness. 

This role brings production support expertise and engineering discipline earlier in the lifecycle to influence design, validate reliability requirements, reduce production risk, and drive measurable operational improvement. 

Your Primary Responsibilities 

  • Participate in design reviews, sprint zero, and delivery planning to define and validate reliability requirements, including resiliency, observability, fault tolerance, performance, scalability, holiday and special-day processing, and disaster recovery. 
  • Collaborate with Major Release Management to ensure each release meets SRE standards for observability, resiliency, and reliability requirements, support readiness, and knowledge base coverage. 
  • Define and improve monitoring, observability, dashboards, telemetry coverage, and alert strategy to strengthen outage detection, reduce noise, improve signal quality, and accelerate incident response. 
  • Assist in major incident response and root cause analysis by identifying observability gaps, improving telemetry and knowledge articles, and driving actions that reduce repeat incidents. 
  • Drive automation, intelligent tooling, and AI-assisted remediation to reduce manual toil, improve consistency, accelerate recovery, and scale operational support. 
  • Serve as the operational readiness authority before production releases by validating reliability requirements, assessing support readiness, surfacing production risks, and confirming release supportability. 
  • Lead capacity, performance, workload trend, and resiliency analysis to ensure applications scale reliably under normal, peak, and stress conditions. 
  • Establish and track reliability metrics such as availability, incident volume, MTTx, alert quality, automation coverage, reliability requirement compliance, change failure rate, and repeat incident reduction. 
  • Participate in application reliability governance and service reviews by presenting incident trends, compliance metrics, operational risks, improvement actions, and readiness gaps. 
  • Prepare executive reporting on reliability posture, release readiness, observability maturity, alert quality, incident trends, automation progress, risks, and improvement outcomes. 
  • Promote SRE practices through mentoring, standards adoption, best-practice sharing, and approved AI tools that improve knowledge, observability, performance, security, and maintainability. 

Qualifications 

  • Minimum of 10  years of related technical experience across application support engineering, software engineering, site reliability engineering, production support, or application operations. 
  • Bachelor’s degree preferred or equivalent practical experience. 
  • Experience supporting business-critical applications in production environments. 
  • SRE, observability, automation, or ITIL certifications are a plus. 

Talents Needed for Success 

  • Proven experience in one or more in-scope roles including Application Support Engineer, SDET, Software Engineer, or SRE, with responsibility for improving reliability practices, application validation, observability coverage, automation frameworks, and reliability standards. 
  • Strong understanding of monitoring and observability platforms, including dashboard design, alert tuning, telemetry coverage, log analysis, metrics, traces, and event correlation. 
  • Programming or scripting proficiency in one or more languages such as Python, Java, Go, PowerShell, or similar for automation, tooling, and operational efficiency. 
  • Familiarity with distributed applications, middleware, messaging, batch processing, real-time processing, and production application behavior in high-availability environments. 
  • Experience in financial services, capital markets, regulated environments, or other high-availability operational settings. 
  • Demonstrated participation in disaster recovery, performance testing, resiliency testing, release readiness, incident response, and root cause analysis. 
  • Knowledge of AI concepts, data platforms, anomaly detection, incident correlation, and intelligent automation use cases. 
  • Strong collaboration skills across application support, application development, release management, risk, security, business, and vendor stakeholders. 
  • Ability to translate production support insights into actionable engineering improvements that reduce risk, improve stability, and enhance customer experience. 

Salary.com Estimation for Site Reliability Engineer -Jersey City, NJ & Dallas, TX in Jersey, NJ
$97,787 to $115,750
If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Site Reliability Engineer -Jersey City, NJ & Dallas, TX?

Sign up to receive alerts about other jobs on the Site Reliability Engineer -Jersey City, NJ & Dallas, TX career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$92,877 - $110,401
Income Estimation: 
$120,933 - $155,034
Income Estimation: 
$114,618 - $136,401
Income Estimation: 
$92,877 - $110,401
Income Estimation: 
$120,933 - $155,034
Income Estimation: 
$114,618 - $136,401
Employees: Get a Salary Increase
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at StradIT

  • StradIT Tampa, FL
  • ob Description: • Reviews processes and end users’ feedback in order to identify opportunities to adapt leading practices and/or more efficient/effective p... more
  • 3 Days Ago

  • StradIT Tampa, FL
  • Note : We need only W2 candidates User Experience Design Conduct user research, stakeholder interviews, and usability assessments. Analyze user journeys, w... more
  • 4 Days Ago

  • StradIT Jersey, NJ
  • Job Role: AI Gateway Engineer Locations: Jersey City NJ, Dallas TX & Tampa FL Work mode: Hybrid Experience: 8 to 12 Years Employment: W2 We are seeking an ... more
  • 6 Days Ago

  • StradIT Tampa, FL
  • Job Role: AI Gateway Engineer Locations: Jersey City NJ, Dallas TX & Tampa FL Work mode: Hybrid Experience: 8 to 12 Years Employment: W2 We are seeking an ... more
  • 6 Days Ago


Not the job you're looking for? Here are some other Site Reliability Engineer -Jersey City, NJ & Dallas, TX jobs in the Jersey, NJ area that may be a better fit.

  • Ampcus Inc Jersey, NJ
  • Job Position: SRE DevOps Engineer (Infrastructure) Location: Jersey City, NJ (Onsite) Type: Contract to hire (W2 Only) Infrastructure Automation (Terraform... more
  • 3 Days Ago

  • StradIT Jersey, NJ
  • Key Responsibilities User Experience Design Conduct user research, stakeholder interviews, and usability assessments. Analyze user journeys, workflows, and... more
  • 13 Days Ago

AI Assistant is available now!

Feel free to start your new journey!