Required Skills: Java, Spring Boot, React.js, REST APIs, Microservices ,SRE, DevOps , CI/CD, Splunk, New Relic, Dynatrace , AWS, Azure, GCP, Docker, Kubernetes, OpenShift , SQL, databases , Production troubleshooting & RCA , Environment management , Release and deployment support
Job Description
Full Stack Engineer with strong SRE expertise
- Design, develop, and maintain scalable web applications using Java, Spring Boot, REST APIs, and React.
- Build and enhance frontend components using React, JavaScript/TypeScript, HTML, CSS, and modern UI frameworks.
- Develop backend services, APIs, integrations, and reusable components.
- Work with relational and NoSQL databases for application development and troubleshooting.
- Participate in code reviews, design discussions, performance tuning, and production-readiness reviews.
- Ensure applications follow secure coding, logging, monitoring, and performance best practices.
Environment Management
- Own non-production environment health end-to-end, including availability, readiness, and stability.
- Maintain an environment usage calendar to prevent team conflicts and scheduling collisions.
- Enforce environment readiness checklists before every sprint cycle, QA phase, integration test phase, and release cycle.
- Coordinate environment refreshes, deployments, configuration updates, data setup, and access readiness.
- Ensure lower environments are aligned with intended production-like configurations where applicable.
Triage & Issue Management
- Triage, classify, and route all non-production issues based on defined severity, priority, and SLA targets.
- Lead daily blocker calls to resolve non-prod issues before they impact sprint or release timelines.
- Ensure every non-prod defect is logged in GAB3 with correct environment labels, sprint links, impact details, and owners.
- Drive timely resolution by coordinating with application, infrastructure, database, QA, DevOps, and vendor teams.
- Track recurring issues and ensure long-term corrective actions are captured.
Monitoring & Observability
- Maintain and enhance Splunk, New Relic, or equivalent observability dashboards for:
- Error rates
- Slow transactions
- API failures
- Health checks
- Infrastructure/resource utilization
- Deployment status
- Correlate logs across services using trace IDs, session IDs, correlation IDs, and request IDs to accelerate QA and defect debugging.
- Monitor CI/CD pipeline runs and immediately flag failed, partial, or inconsistent deployments.
- Define and maintain non-prod alerting thresholds, dashboards, and operational reports.
Release & Change Support
- Validate non-production environment stability before approving promotion to production.
- Support release readiness reviews, deployment validations, smoke testing, and rollback planning.
- Rehearse rollback procedures in non-prod, including:
- Application rollback
- Script rollback
- Database reversals
- Configuration resets
- Feature flag reversals
- Track feature flag states across environments to ensure non-prod mirrors intended production behavior.
- Partner with release managers, Scrum Masters, product owners, QA leads, and engineering teams to manage release risks.
Integration & Dependency Management
- Monitor third-party sandbox and integration availability to prevent test execution blockers.
- Manage third-party stubs, mock services, test data, certificates, and endpoint availability.
- Coordinate proactively for certificate renewals, credentials, tokens, firewall rules, and integration dependencies.
- Support FHIR/API-based integrations and troubleshoot connectivity, authentication, payload, and contract issues.
- Ensure dependent services and test environments are available for planned sprint and release testing.
Runbooks & Knowledge Management
- Maintain non-production operational runbooks for:
- Application restarts
- Cache clears
- Data reseeds
- Pipeline retriggers
- Deployment validations
- Smoke tests
- Rollback steps
- Maintain a living known-issues registry to reduce repeat escalations.
- Document environment dependencies, support contacts, escalation paths, and common troubleshooting steps.
- Ensure support documentation is current, accurate, and usable by development, QA, and operations teams.
Stakeholder Communication
- Deliver daily non-prod blocker summaries to Scrum Masters, Release Managers, Product Owners, QA Leads, and Engineering Managers.
- Own escalation paths when non-prod issues threaten sprint goals or release timelines.
- Provide clear status updates on environment availability, incident progress, risk items, and ETA for fixes.
- Facilitate cross-functional discussions to remove blockers and improve environment stability.
Continuous Improvement & Automation
- Identify recurring non-prod failure patterns and drive root cause fixes instead of temporary workarounds.
- Convert operational pain points into actionable GAB3 backlog items.
- Automate environment health checks, deployment validations, test data resets, smoke tests, and reporting.
- Improve CI/CD pipeline reliability, observability, and deployment traceability.
- Recommend process improvements to increase development, QA, and release efficiency.
Required Skills
Backend Development
- Strong hands-on experience with Java 8/11/17, Spring Boot, Spring MVC, Spring Security, and RESTful APIs.
- Experience designing and consuming microservices.
- Strong understanding of API design, exception handling, logging, authentication, and authorization.
- Experience with JUnit, Mockito, integration testing, and code quality tools.
Frontend Development
- Strong experience with React.js.
- Good knowledge of JavaScript, TypeScript, HTML5, CSS3, Redux/Context API, and frontend build tools.
- Experience integrating frontend applications with REST APIs.
- Ability to debug browser, API, and UI performance issues.
SRE / DevOps / Environment Support
- Strong experience in SRE, DevOps, environment management, or production support.
- Hands-on experience with CI/CD tools such as Jenkins, GitLab CI, GitHub Actions, Azure DevOps, or similar.
- Experience monitoring applications using Splunk, New Relic, Dynatrace, AppDynamics, or equivalent tools.
- Strong troubleshooting skills across application logs, APIs, databases, infrastructure, and deployment pipelines.
- Experience with incident management, root cause analysis, SLA management, and operational reporting.
Cloud / Infrastructure
- Experience with cloud platforms such as AWS, Azure, or GCP.
- Good understanding of containers and orchestration tools such as Docker and Kubernetes/OpenShift.
- Experience with config management, secrets management, certificates, and environment-specific deployment configurations.
Database
- Experience with relational databases such as Oracle, PostgreSQL, MySQL, or SQL Server.
- Ability to write and troubleshoot SQL queries.
- Experience with data refreshes, reseeding, DB scripts, and rollback validations.
Preferred Skills
- Experience with healthcare integrations, FHIR APIs, Epic/Cerner sandboxes, or similar third-party integrations.
- Experience with feature flag tools such as LaunchDarkly or similar.
- Experience with API testing tools such as Postman, SoapUI, Swagger, or REST Assured.
- Experience with infrastructure-as-code tools such as Terraform, Ansible, or CloudFormation.
- Experience working in Agile/Scrum delivery models.
- Familiarity with ITIL practices, change management, and release governance.
Tools & Technologies
- Languages: Java, JavaScript, TypeScript
- Backend: Spring Boot, REST APIs, Microservices
- Frontend: React.js, HTML, CSS, Redux/Context API
- Databases: Oracle, PostgreSQL, MySQL, SQL Server
- CI/CD: Jenkins, GitLab CI, GitHub Actions, Azure DevOps
- Monitoring: Splunk, New Relic, Dynatrace, AppDynamics
- Cloud/Infra: AWS/Azure/GCP, Docker, Kubernetes/OpenShift
- Testing: JUnit, Mockito, Postman, Selenium/Cypress preferred
- Tracking: GAB3/Jira or equivalent Agile tracking tools
- Version Control: Git, Bitbucket/GitHub/GitLab
Qualifications
-
Bachelor’s degree in Computer Science, Engineering, Information Technology, or related field.
-
10+ years of experience in software engineering, full stack development, SRE, DevOps, or environment support.
-
Strong hands-on experience in Java, Spring Boot, React, CI/CD, monitoring, troubleshooting, and release support.
-
Excellent analytical, problem-solving, and communication skills.
-
Ability to work with cross-functional teams in a fast-paced Agile environment.