Cloud Solutions Engineer
Tyler Technologies
Plano, TX, United States
18 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$93,547.0 - $152,629.0
Working hours
Regular working hours
Job source
Tech stack
Query Performance
.NET Framework
Microsoft Windows
Agile Methodology
Amazon Web Services
Antivirus Softwares
Application Release Automation
JIRA
Microsoft Azure
Bash Shell
C Sharp (Programming Language)
Cloud Computing
+45 more
Configuration Management
Code Review
Information Systems
Computer Networks
Computer Engineering
Continuous Integration
Database Theory
Software Debugging
Linux
DevOps
Web Development
Microsoft Dynamics CRM
Monitoring of Systems
Internet Information Services (IIS)
Issue Tracking Systems
Python (Programming Language)
PostgreSQL
Linux System Administration
Microsoft SQL Server
Visual Basic .NET (Programming Language)
Windows Servers
Performance Tuning
Windows PowerShell
Reliability Engineering
Site Reliability Engineering Practices
Cloud Services
Software Engineering
Transact-SQL
Web Applications
Datadog
Data Logging
Scripting
Database Performance Analyzer
Cloud Platform System
Grafana
Indexer
Infrastructure Automation Frameworks
Information Technology
SolarWinds (Software)
Deployment Automation
Cloudwatch
Terraform
Pagerduty
Golang
Programming Languages
Job description
- Design, build, and maintain highly available, scalable, secure, and cost-effective cloud infrastructure and supporting services.
- Implement, maintain, and continuously improve monitoring, alerting, logging, and observability practices to detect reliability risks early and ensure alerts are actionable.
- Proactively assess production systems for operational risk, performance bottlenecks, resiliency gaps, and recurring failure patterns; recommend and implement preventative improvements.
- Define, track, and improve service reliability indicators, objectives, and error budgets in partnership with engineering, product, and operations teams to balance reliability, customer impact, and delivery priorities.
- Partner with engineering, product support, cloud operations, and infrastructure teams to strengthen incident response readiness, improve communication, and drive timely resolution of production issues.
- Lead root cause analysis and cross-functional incident retrospectives, translating findings into preventative actions that reduce recurring incidents and improve system resiliency.
- Develop automation, operational tooling, deployment support, and infrastructure provisioning practices to reduce manual effort, improve consistency, and minimize operational toil.
- Conduct regular performance tuning, capacity analysis, reliability reviews, and system health assessments to support SLA commitments and customer expectations.
- Drive continuous improvement initiatives that improve reliability engineering practices, operational maturity, and productivity across teams.
- Participate in a structured on-call rotation during primary business hours and contribute to incident management best practices.
- Stay current with industry trends, best practices, and emerging technologies in site reliability engineering, DevOps, cloud operations, automation, and observability.
Requirements
- BS/BA degree in Computer Science, Computer Engineering, Information Systems, or a related field, or equivalent practical experience.
- 3+ years of experience in Site Reliability Engineering, Cloud Operations, DevOps, Software Engineering, or a related technical role supporting cloud-based production systems.
- 5+ years of hands-on software engineering experience, including application design, code review, testing, debugging, release processes, and supporting production software systems.
- Proficiency with cloud computing platforms such as AWS, Azure, or GCP, including experience with infrastructure as code tools such as Terraform.
- Hands-on experience with monitoring, logging, alerting, and observability tools such as Datadog, AWS CloudWatch, SolarWinds Database Performance Analyzer, or similar platforms.
- Experience participating in or improving incident response processes, including use of tools such as PagerDuty, JSM Operations, or similar incident management platforms.
- Experience conducting root cause analysis, identifying systemic risks, and implementing preventative measures to reduce production incidents.
- Familiarity with SRE practices such as service level indicators, service level objectives, error budgets, toil reduction, blameless post-incident reviews, and reliability-focused engineering.
- Expertise in scripting or programming languages such as PowerShell, Python, Bash, C#, Go, or .NET for automation and tooling.
- Strong understanding of database concepts and administration, including T-SQL scripting, indexing, query performance, and performance tuning for MS SQL Server and/or PostgreSQL.
- Solid understanding of networking concepts, security best practices, and system administration in Windows and Linux environments.
- Experience with CI/CD practices, deployment automation, configuration management, and modern DevOps methodologies.
- Strong analytical and problem-solving skills, with the ability to troubleshoot complex issues and drive timely, sustainable resolution.
- Clear communicator who thrives in collaborative, cross-functional environments and can explain technical risks, tradeoffs, and recommendations to varied audiences.
- Demonstrated curiosity, ownership, and a continuous improvement mindset focused on reliability, prevention, and operational excellence.
Required Skills and Technologies
- AWS or comparable cloud platform experience
- MS SQL Server and/or PostgreSQL
- Windows Server OS and Linux OS
- PowerShell and at least one additional scripting or programming language such as Python, Bash, C#, or .NET
- IIS and web application hosting concepts
- Monitoring, alerting, observability, and incident management practices
- Infrastructure as Code and automation practices
Preferred Skills and Technologies
- Knowledge of .NET Framework and languages such as C# or VB.NET
- Experience with monitoring and observability tools such as Datadog, AWS CloudWatch, and SolarWinds Database Performance Analyzer
- Experience with PagerDuty, JSM Operations, or similar incident management platforms
- Knowledge of Agile development practices and tools such as JIRA
- Knowledge of web development practices and technologies
- Experience with ticketing systems such as Microsoft CRM
- Advanced knowledge of AWS hosting technologies
- Advanced knowledge of DevOps practices, release automation, and operational readiness
- Experience with security and endpoint protection tools such as CrowdStrike Anti-Virus
Benefits & conditions
Salary will generally fall between $93,547 - $152,629 before adjustment for geographic differences. Recruiter can confirm if position is incentive eligible.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on diversityjobs.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
AJ
Austin Joy
over 4 years ago
LM
Luis Minvielle
Is Software Engineering Over-Saturated?
over 2 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
LM
Luis Minvielle
The 12 Best Jobs for Software Engineers
over 2 years ago
LM
Luis Minvielle
The Most Popular IT Jobs on the Market
over 2 years ago
LM
Luis Minvielle
Fully Remote Software Engineer Jobs
about 2 years ago