Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
EviSmart Philippines is seeking a Production Support & Incident Response Lead to own incident response and platform continuity during our night operations. You will be the lead incident response person on shift, investigating first, restoring operations, and coordinating with Engineering and DevOps to minimize disruption.
The role supports 2,000+ dental labs on a single platform, requiring calm under pressure, clear communication, and accountability for end-to-end incident management.
We ship before we're 100% certain. We write things down because we have two offices and memory is lossy. We debate loudly and move without resentment. We treat the customer's real problem as more important than an elegant internal process. If you've spent time waiting for permission to try something obvious — you'll notice the difference here immediately.
On-site Full-Time Night Shift
We ship before we're 100% certain. We write things down because we have two offices and memory is lossy. We debate loudly and move without resentment. We treat the customer's real problem as more important than an elegant internal process. If you've spent time waiting for permission to try something obvious — you'll notice the difference here immediately.
When something goes wrong in production overnight, the Production Support & Incident Response Lead is the person leading the response.
EviSmart is looking for a Production Support & Incident Response Lead to own incident response and platform continuity during our night operations.
This is not a role where you simply monitor dashboards, create a ticket, and wait for Engineering.
You will be the lead incident response person on shift. You are expected to investigate first, understand what is happening, determine the safest way to restore operations, bring in the right technical people when necessary, and remain accountable for the incident until the platform is stable.
Our platform supports 2,000+ dental labs, so an issue in production can quickly become a real business problem for our customers. The goal is simple: keep cases moving and minimize disruption.
The role's ownership of monitoring, incident command, proactive customer communication, post-mortems and backup development is explicit in the operating playbook.
This is not a traditional Service Delivery Manager or ITIL governance position. It is also not a pure DevOps or Software Engineering role. You don't need to be the person who writes the permanent code fix for every problem. But you do need enough technical depth to investigate intelligently before asking someone else to solve it.
If your normal incident process is: Alert → Create ticket → Escalate → Wait ... then this probably isn't the right role.
We're looking for someone whose instinct is closer to: Detect → Investigate → Isolate → Restore or Work Around → Escalate Intelligently → Command Through Resolution → Prevent Recurrence
It's 2 AM. A production issue is preventing a customer from processing cases. The permanent fix requires an engineer who isn't immediately available.
We're looking for someone who doesn't stop at "I\'ll elevate it."
We want someone who starts asking: What actually broken? What's the business impact? What changed? What can I verify myself? Can I safely restore the previous working state? Is there another way to keep the customer's operation moving? Who genuinely needs to be involved? What can we do now instead of waiting until morning?
That's the mindset we're hiring for.
Your first 90 days are designed to progressively prove that we can trust you with the night.
First 30 days: Learn the platform and prove you can troubleshoot real issues and identify the correct workaround without being walked through every step.
By 60 days: Independently monitor the platform, recognize warning signals and own selected production tickets through resolution.
By 90 days: Independently command night incidents from detection through restoration, handle more complex issues, proactively identify problems and effectively delegate to your trained backup.
Ultimately, success means >99% platform health during night operations, incidents declared quickly, complete post-mortems, no dropped night-to-day handoffs, and a backup capable of independently handling night triage.
Deep DevOps expertise is not required. What matters is that you can investigate intelligently, understand what you're seeing, take the safest action available at your level, and know when specialist intervention is genuinely necessary.
This is a permanent night-shift role.
This is also an Individual Contributor role, not a traditional people-management position. You will act as the functional point person during night coverage and will help develop a designated backup, but you will not be joining to manage a large team.
The responsibility is significant because you are the person we need to trust when the daytime team isn't around.
If you're already strong in Application or Production Support and you're looking for an opportunity where you're given more ownership, more decision-making authority and the chance to become the person trusted to lead production incidents, we'd like to hear from you.
EviSmart Philippines evismart.com