Job Title: Application Support / SRE Engineer
Experience: 2 4 Years,
Work Auth: All visas accepted (No H1 and No Fake Profiles)
Work Location: Chicago, IL (Hybrid)
We are looking for an experienced Application Support / SRE Engineer with strong expertise in AWS, Snowflake, SQL, Python, application monitoring, API troubleshooting, and production support. The ideal candidate will be responsible for supporting cloud-based applications, resolving production issues, improving system availability and performance, and developing automation and monitoring capabilities.
Bachelor's or Master's degree with 2 4 years of relevant experience in Application Support, SRE, Production Support, or a related area.
Candidates without a college degree may be considered with relevant technical certifications and 6+ years of applicable experience.
Minimum 2+ years of application support experience for cloud-based applications.
Strong experience in Application Support / Production Support.
Hands-on knowledge of relational databases, SQL querying, and database reporting.
Experience working with Snowflake for SQL queries, reporting, and troubleshooting.
Strong working knowledge of AWS services, including:
EC2
S3
VPC
Route 53
RDS
DynamoDB
Lambda
CloudWatch
IAM
Certificate Manager
ELB
EBS
ECS
CloudFront / WAF
SQS
SNS
SES
CloudFormation
Experience troubleshooting UI, API, data-flow, and integration issues.
Experience troubleshooting REST APIs and JSON-based integrations.
Good understanding of high-availability architecture.
Experience with monitoring and observability tools such as AppDynamics, Grafana, ThousandEyes, or similar platforms.
Experience with Python, PowerShell, SQL, and JSON.
Ability to develop scripts and automation solutions for operational and support activities.
Understanding of Site Reliability Engineering (SRE) principles, including application availability, reliability, and performance improvement.
Familiarity with data management, data engineering, or data operations.
Working knowledge of Azure DevOps (ADO) pipelines / CI-CD frameworks.
Understanding of auto-scaling concepts and cloud-based scalability features.
Triage, analyze, and resolve application support tickets within defined SLA timelines.
Troubleshoot and resolve critical production and technical issues.
Provide application support during off-hours or weekends when required for critical incidents.
Proactively identify production issues and coordinate resolution with appropriate technical teams.
Document incidents, root causes, corrective actions, and resolution steps.
Monitor application availability, reliability, and performance.
Develop scripts and automation tools to identify, prevent, and resolve recurring application issues.
Build and enhance monitoring, alerting, and operational dashboards.
Troubleshoot application issues across UI, APIs, databases, cloud infrastructure, and data flows.
Support AWS-hosted applications and Snowflake-based data environments.
Work closely with engineering, infrastructure, database, DevOps, and business teams to resolve production issues.
Participate in incident management, problem management, and production-support activities.
Support dealer/customer communication when required regarding application-related issues.
Previous experience working in an SRE or Production Support environment.
Strong troubleshooting and root-cause-analysis skills.
Experience supporting highly available, business-critical cloud applications.
Knowledge of automation, monitoring, alerting, and operational excellence practices.
Strong communication and stakeholder-management skills.