Application Engineer 3 (24×7) 5 openings available
Full-Time
OPS
Posted 2 weeks ago
Job Features
| Job Category | Application Engineer |
| Disclaimer | TS/SCI Full Scope Poly Required. |
| Requirements | Description: •As part of the Secure the Enterprise initiative, develop capabilities to shift from the current manual system security evaluation and authorization process to a new model that emphasizes automation, streamlined processes and approvals, continuous monitoring and assessment, and network data gathering across the entire life cycle of a project. Essential Duties and Responsibilities: This is a blended position across all teams in ANDS. As such, the successful candidate will have skills crossing from Software Development, Systems Engineering, Systems Administration, and Operations Support. •Ensure data reliability and accuracy by monitoring and maintaining application systems. •ontribute to operational success through continuous performance monitoring and proactive troubleshooting, to potentially include changes to system configuration, changes to the code base, etc. •Maintain system health by tracking metrics, logs, dashboards, alerts, and application status to prevent downtime and ensure optimal application performance. •Support watchfloor operations by identifying service degradation, failed dataflows, system errors, and application issues, and escalating as needed. •Perform basic Linux system administration tasks, including checking service status, starting, stopping, and restarting services, reviewing logs, validating disk, memory, and CPU usage, and supporting server reboots. •Support cloud-based operations by using the AWS Console or AWS CLI to check instance health, review system status, restart or reboot servers, and assist with basic operational troubleshooting. •Work in a rotating shift schedule, 6AM-6PM / 6PM-6AM, to provide 24/7 application support on our watch floor. Required: Prospective candidates must have proficiency in one or more of the following skillsets / technologies: •Monitoring / Watchfloor Operations: Grafana, Kibana, Splunk, CloudWatch, or similar monitoring tools; dashboard monitoring, alert triage, log review, incident escalation, and shift turnover support. •Linux / Systems Administration: Comfortable working in a Linux environment; ability to check system health, start, stop, and restart services, review logs, and reboot servers safely. •Cloud / AWS Operations: C2E, HCI, AWS, Lambda, AWS Console, or AWS CLI experience supporting EC2 instances, service health checks, and basic operational troubleshooting. •Dataflow / Application Support: Apache NiFi, Cribl, Kafka, Logstash, or similar data pipeline / dataflow tools; experience monitoring application health, failed jobs, queues, data movement, and ingestion issues. •DevOps: Terraform, Ansible, GitLab Pipelines, Docker. |