Application Engineer View role →

NRnP Technology

Northern (KY)

Hybrid

USD 90,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NRnP Technology is seeking an experienced Production Support Engineer to join our 24/7 operations team in the United States. You will act as the primary contact for incidents, perform triage, root cause analysis, and code-level fixes, and collaborate across engineering, QA, and infra to minimize business impact.

You will design automation to reduce manual tasks, build observability dashboards, participate in post-incident reviews, and help with capacity planning.

Qualifications

  • 3+ years of experience in application support, production support, or a similar engineering role.
  • Strong understanding of software engineering principles and ability to read and debug code.
  • Experience with monitoring and observability tools.
  • Familiarity with cloud platforms and containerized environments.
  • Experience building automation scripts using Python or similar languages.
  • Familiarity with CI/CD pipelines and DevOps practices.
  • Strong problem-solving and analytical skills.
  • Excellent verbal and written communication skills.

Responsibilities

  • Act as the primary point of contact for production issues, alerts, and user-reported incidents.
  • Log, categorize, and prioritize tickets in accordance with defined SLAs.
  • Ensure timely resolution and minimize business impact.
  • Perform root cause analysis for production issues across application and infrastructure layers.
  • Collaborate with cross-functional teams to resolve complex technical problems.
  • Use logs, metrics, and traces to diagnose issues efficiently.
  • Review and debug application code to identify and resolve defects.
  • Implement code-level fixes and coordinate with development teams for larger changes.
  • Participate in code reviews to ensure quality and maintainability.
  • Design and build automation scripts and tools to reduce manual operational tasks.
  • Leverage AI-assisted tools to streamline troubleshooting and incident response.
  • Continuously identify opportunities to improve operational workflows.
  • Build and maintain dashboards, alerts, and monitoring frameworks.
  • Enhance system observability to proactively detect and prevent issues.
  • Ensure monitoring coverage aligns with business-critical services.
  • Apply SRE principles to improve system reliability, availability, and performance.
  • Participate in post-incident reviews and drive corrective actions.
  • Contribute to capacity planning and performance tuning efforts.
  • Escalate unresolved issues to appropriate teams while maintaining ownership of resolution.
  • Work closely with development, QA, and infrastructure teams to prevent recurring issues.
  • Foster strong cross-team relationships to support rapid issue resolution.
  • Create and maintain runbooks, troubleshooting guides, and knowledge base articles.
  • Document incident details, root causes, and resolutions for future reference.
  • Ensure documentation remains current and accessible to the team.
  • Provide clear and timely updates to stakeholders during incidents.
  • Communicate technical issues effectively to both technical and non-technical audiences.
  • Collaborate with global teams across different time zones.

Skills

Application support
Observability
Coding skills
Problem solving
Communication

Education

Bachelor’s degree in Computer Science, Engineering, or related field

Tools

Docker
Kubernetes
Python

Job description

Key Responsibilities

Incident Management

  • Act as the primary point of contact for production issues, alerts, and user-reported incidents
  • Log, categorize, and prioritize tickets in accordance with defined SLAs
  • Ensure timely resolution and minimize business impact

Triage and Troubleshooting

  • Perform root cause analysis for production issues across application and infrastructure layers
  • Collaborate with cross-functional teams to resolve complex technical problems
  • Use logs, metrics, and traces to diagnose issues efficiently

Code-Level Analysis and Fixes

  • Review and debug application code to identify and resolve defects
  • Implement code-level fixes and coordinate with development teams for larger changes
  • Participate in code reviews to ensure quality and maintainability

Automation and Operational Efficiency

  • Design and build automation scripts and tools to reduce manual operational tasks
  • Leverage AI-assisted tools to streamline troubleshooting and incident response
  • Continuously identify opportunities to improve operational workflows

Observability and Monitoring

  • Build and maintain dashboards, alerts, and monitoring frameworks
  • Enhance system observability to proactively detect and prevent issues
  • Ensure monitoring coverage aligns with business-critical services

SRE and Continuous Improvement

  • Apply SRE principles to improve system reliability, availability, and performance
  • Participate in post-incident reviews and drive corrective actions
  • Contribute to capacity planning and performance tuning efforts

Escalation and Collaboration

  • Escalate unresolved issues to appropriate teams while maintaining ownership of resolution
  • Work closely with development, QA, and infrastructure teams to prevent recurring issues
  • Foster strong cross-team relationships to support rapid issue resolution

Documentation and Knowledge Management

  • Create and maintain runbooks, troubleshooting guides, and knowledge base articles
  • Document incident details, root causes, and resolutions for future reference
  • Ensure documentation remains current and accessible to the team

Communication

  • Provide clear and timely updates to stakeholders during incidents
  • Communicate technical issues effectively to both technical and non-technical audiences
  • Collaborate with global teams across different time zones
Required Qualifications

Experience

  • 3+ years of experience in application support, production support, or a similar engineering role

Technical Skills

  • Strong understanding of software engineering principles and ability to read and debug code
  • Experience with monitoring and observability tools
  • Familiarity with cloud platforms and containerized environments

Automation and DevOps

  • Experience building automation scripts using Python or similar languages
  • Familiarity with CI/CD pipelines and DevOps practices

Soft Skills

  • Strong problem-solving and analytical skills
  • Excellent verbal and written communication skills
  • Ability to work effectively under pressure during incidents

Education

  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience

Preferred Qualifications

  • Experience with AI-assisted operational tools
  • Exposure to Site Reliability Engineering (SRE) practices
  • Experience working in a 24/7 production support environment

Full-time employment only. No C2C (corp-to-corp) arrangements.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr Application Support Engineer
Sr Application Support Engineer

IntePros • Pittsburgh

On-site
USD 80,000 - 110,000
Site-Reliability Engineer, Application Operations
Site-Reliability Engineer, Application Operations

Scorpion Therapeutics • Irving (TX)

On-site
USD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

System One • Pittsburgh

On-site
USD 140,000 - 190,000
I Site Reliability Engineer Incident IQ Alpharetta, Georgia, US
I Site Reliability Engineer Incident IQ Alpharetta, Georgia, US

Artha Nexgen • Alpharetta (GA), Northern (KY)

Hybrid
USD 120,000 - 160,000
Systems Engineer
Systems Engineer

Compunnel, Inc. • Westlake (OH)

Hybrid
USD 100,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

SCIGON • Naperville (IL)

Hybrid
USD 110,000 - 170,000
Application Support Engineer
Application Support Engineer

Programmers.io • Sunnyvale (CA)

On-site
USD 120,000 - 150,000
IT CONSULTANT SR
IT CONSULTANT SR

First Horizon Corp. • Memphis (TN), Northern (KY)

On-site
USD 120,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

Jobtailor • New Jersey

On-site
USD 120,000 - 180,000
Senior Application Support Engineer
Senior Application Support Engineer

Insight Global • United States

On-site
USD 90,000 - 120,000