Application Administrator in New York at Jobgether
Explore Related Opportunities
Job Description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an Application Administrator based in the United States.
This role supports a mission-critical Emergency Routing Service platform that enables reliable emergency communications for public safety organizations.
You will help maintain the availability, stability, and performance of production applications operating in a high-availability environment.
The position combines application administration, incident response, monitoring, troubleshooting, automation, and continuous improvement.
You will work closely with infrastructure, network, database, development, and customer-facing teams to resolve complex operational issues.
The role offers the opportunity to improve observability, automate repetitive processes, and strengthen platform resilience.
You will also contribute to upgrades, migrations, disaster recovery testing, and business continuity initiatives.
Success requires a technically curious problem-solver who can respond calmly and effectively when critical services are at risk.
- Monitor and maintain the health, performance, stability, and availability of production applications and supporting services.
- Manage IT service requests and tickets from multiple teams through an IT service management system, communicating clearly with requestors throughout resolution.
- Respond to system alerts, incidents, and service requests according to established operational procedures and escalation practices.
- Perform application maintenance, upgrades, patching, configuration management, and production deployment support.
- Validate application functionality after changes and maintain accurate system documentation, operational runbooks, and support procedures.
- Identify opportunities to automate recurring operational activities and enable support teams to self-service routine tasks.
- Troubleshoot and resolve application, server, database, and integration issues.
- Participate in major incident response and an on-call rotation covering nights and weekends for time-sensitive production events.
- Perform root cause analysis and help implement corrective and preventive actions to improve platform reliability.
- Collaborate with engineering teams to resolve defects, strengthen resiliency, and improve service delivery.
- Use monitoring and observability platforms to identify service degradation before it affects customers.
- Develop and enhance dashboards, alerts, operational reports, and log-analysis processes.
- Support system upgrades, migrations, modernization initiatives, disaster recovery testing, and business continuity activities.
- Recommend improvements that increase availability, operational efficiency, supportability, and readiness for new products and services.
- Communicate incident causes, mitigation steps, and operational updates effectively to customer-facing teams, leadership, and other stakeholders.
Requirements:
- Bachelor’s degree in Information Technology, Computer Science, or a related field, or equivalent professional experience.
- 2+ years of experience supporting enterprise applications in a production environment.
- Experience administering Linux and/or Windows server environments across physical and cloud-based infrastructure.
- Working knowledge of SQL and database troubleshooting concepts.
- Working knowledge of SIP and SBC technologies, with an understanding of telecommunications or application integration concepts.
- Experience with monitoring, alerting, observability, and log-analysis platforms.
- Familiarity with incident, problem, and change management processes.
- Strong analytical, troubleshooting, critical-thinking, and problem-solving abilities.
- Excellent written and verbal communication skills, with the ability to explain technical issues and resolutions to both technical and non-technical audiences.
- Ability to remain effective in a fast-paced operational environment and respond appropriately to high-priority incidents.
- Willingness and ability to participate in an on-call rotation that includes nights and weekends.
- Experience in 24x7 public safety, emergency communications, telecommunications, SaaS, or another high-availability environment is preferred.
- Experience communicating with customer-facing teams and executive leadership is a plus.
- Experience administering Java application servers such as Payara, GlassFish, or JBoss; F5 BIG-IP LTM; VMware/vCenter; or AWS-hosted web applications is beneficial.
- Knowledge of SIP/VoIP, Asterisk, Oracle platforms, and call-routing concepts is advantageous.
- Experience with Splunk, New Relic, or comparable monitoring and observability platforms is a plus.
- Experience with Ansible, Docker, Bash, Python, PowerShell, or similar automation technologies is desirable.
Benefits:
- Competitive compensation, with starting salary based on experience.
- Remote work opportunity within the United States.
- Medical, dental, and vision insurance.
- Life and disability coverage.
- Paid time off.
- 401(k) retirement savings plan.
- Paid parental leave.
- Access to extensive personal and professional training resources.
- Employee discount programs.
- Critical illness and hospital indemnity coverage.
- Legal support benefits.
- Pet insurance.
- Identity theft protection.
- Employee Assistance Program with free mental health resources and support.
- Opportunities for professional development and career growth.
- Collaborative environment supporting critical public safety and emergency communications services.