> Markdown version of [/jobs/ext/2730524-production-support-engineer](https://www.wearedevelopers.com/jobs/ext/2730524-production-support-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Production Support Engineer - **Company:** VLINK INC - **Location:** Bloomfield, CT, United States - **Experience:** Experienced - **Salary:** $128,960.0 - **Contract:** Temporary contract - **Skills:** Query Performance, Java (Programming Language), Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Amazon Cloudfront, Amazon S3, Application Integration Architecture, Application Performance Management, Interactive Voice Response, Software as a Service, Cloud Computing, Computer Programming, Databases, Data Validation, Data Synchronization, DevOps, Design of User Interfaces, Human-Computer Interaction, Python (Programming Language), Knowledge Management, Log Analysis, MongoDB, Query Optimization, Redis, Software Tools, Cloud Services, Runbook, Server Administration, Software Deployment, Software Engineering, Systems Integration, Amazon Connect, Scripting, Enterprise Software Applications, Cloud Platform System, ReactJS, System Availability, Prompt Engineering, Spring-boot, Software Troubleshooting, Caching, Generative AI, Indexer, Backend, Material UI, Information Technology, Performance Monitor, Software Coding, Restful APIs, Splunk, Dynatrace, Microservices - **Published:** September 5, 2026 - **Apply:** https://www.careerbuilder.com/job-details/production-support-engineer-bloomfield-ct--0edb541d-de7b-4e6d-ba6a-87d1cb2123fa ## About the Role * At least 6 8 years of overall IT experience, including at least 4 years of L2 production support for enterprise and business critical applications. * Having coding knowledge for supporting applications developed using React, Python, REST APIs, microservices, MongoDB, Redis Cache, and AWS services. * Experience troubleshooting end to end production issues across the UI, backend services, APIs, integrations, databases, caching layers, and cloud infrastructure. * Experience designing and building an enterprise alerting framework for proactive monitoring, alert correlation, notification, escalation, and incident prevention. * Experience supporting Doctor Tools and clinician facing applications, including application availability, integrations, workflow issues, and production performance. * Strong experience managing incident and change queues, including ticket prioritization, assignment, SLA tracking, technical analysis, change validation, stakeholder communication, and timely closure. * Experience supporting MongoDB, including connectivity, query performance, indexing, data validation, and production troubleshooting. * Working knowledge of Redis Cache, including availability, memory utilization, key expiration, synchronization, connectivity, and performance issues. * Practical experience using AI tools, Generative AI, prompt engineering, and RAG for incident analysis, troubleshooting, operational automation, incident summarization, and knowledge management. * Experience supporting, migrating, and operating AI enabled IVR and contact center solutions, including Kore.ai, Sierra AI for Health Services, Amazon Connect Outbound, Doctor Tools, and related AI tools, covering integrations, monitoring, incident resolution, production stability, and continuous improvement. * Experience supporting AWS hosted applications involving Amazon S3, CloudFront, application deployments, monitoring, and related AWS services. * Strong experience using Dynatrace and Splunk for log analysis, distributed tracing, dashboarding, alert investigation, performance monitoring, and root cause analysis. * Proven experience managing major incidents, incident bridges, problem management, root cause analysis, corrective actions, and preventive measures. * Experience developing Python automation scripts for health checks, log analysis, alert enrichment, operational reporting, and repetitive support activities. * Experience supporting production releases, deployment validation, post release monitoring, rollback coordination, runbooks, SOPs, and knowledge documentation. * Ability to participate in rotational on call support and communicate effectively with business stakeholders, client teams, development teams, infrastructure teams, and leadership. * Education: Bachelor's degree in computer science or related field., Amazon CloudFront, Amazon Simple Storage Service (S3), Amazon Web Services (AWS), Analysis Skills, Application Hosting, Application Integration, Application Programming Interface (API), Artificial Intelligence (AI), Automation, Caching, Call Centers, Candidate Screening, Change Management, Cloud Applications, Cloud Computing, Communication Skills, Computer Programming, Computer Science, Continuous Improvement, Corrective Action, Data Quality, Database Technology, DevOps, Diversity, Documentation, Enterprise Applications, Establish Priorities, Healthcare, High Availability, High Reliability, Identify Issues, Knowledge Management, Leadership, Memory Hardware, Microservices, MongoDB, On Call, Operational Strategy, Operational Support, Performance Analysis, Problem Solving Skills, Production Management, Production Support, Python Programming/Scripting Language, Query Optimization, React.js, Redis, Reporting Dashboards, Risk Analysis, Root Cause Analysis, Scripting (Scripting Languages), Service Level Agreement (SLA), Software Administration, Software Development, Software Engineering, Splunk, Standard Operating Procedures (SOP), Technical Analysis, Technical Support, Time Management, User Interface/Experience (UI/UX), Voice Response Systems ## Description Our Client is seeking a Production Support Engineer who will provide production support for enterprise applications and cloud-based platforms, ensuring high availability, reliability, performance, and stability. The role requires strong troubleshooting and coding skills across React, Java/Spring Boot, Python, APIs, microservices, databases, caching technologies, and AWS. The engineer will work closely with development, infrastructure, DevOps, and business teams to resolve production issues and continuously improve application support and operational efficiency. Your future duties and responsibilities: * Provide L2/L3 production support for enterprise applications developed using React, Python, APIs, microservices, MongoDB, Redis, and AWS services. * Monitor application availability, performance, and reliability, and proactively identify production risks and service degradation. * Having Coding knowledge on React UI, Java/Spring Boot backend service, integration, database, cache, and application performance issues. * Support MongoDB operations, including connectivity, data validation, query optimization, indexing, and performance troubleshooting. * Monitor and resolve Redis Cache issues related to availability, memory utilization, key expiration, data synchronization, and application connectivity. * Support AWS hosted applications and deployments involving Amazon S3, CloudFront, React UI components, and related cloud services. * Use Dynatrace and Splunk for log analysis, distributed tracing, alert investigation, performance monitoring, dashboarding, and production diagnostics. * Manage production incidents by performing impact assessment, participating in incident bridges, coordinating resolution, completing root cause analysis, and implementing preventive actions. * Develop Python based automation for application health checks, log analysis, alert enrichment, operational reporting, and repetitive support activities. * Apply Generative AI, prompt engineering, RAG, and AI assisted tools to accelerate incident analysis, generate incident summaries, support troubleshooting, and improve knowledge management. * Support application releases by completing readiness checks, validating deployments, monitoring post release performance, and coordinating rollback or remediation activities when required. * Communicate incident status, risks, technical findings, and recovery progress clearly to business stakeholders, development teams, infrastructure teams, and leadership. * Participate in rotational on call support and continuously improve application stability, support efficiency, monitoring coverage, and incident prevention. ## Related Videos - [Answering the Million Dollar Question: Why did I Break Production?](https://www.wearedevelopers.com/videos/1171-answering-the-million-dollar-question-why-did-i-break-production) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)