Just postedUrgently hiring Use left and right arrow keys to navigate
Provided by the employer
Verified Pay check_circle $48.48-$53.87 per hour
Hours Full-time, Part-time
Location Austin, Texas

Compare Pay

Verified Pay check_circleProvided by the employer
This job pays about average compared to similar jobs in your area.

$30.16

$51.17

$72.92


About this job

Overview

Placement Type:

Temporary

Salary:

$48.48-53.87 Hourly

Start Date:

Oct 8, 2026

The hiring company is a leader in the financial services industry, dedicated to empowering individuals and institutions to achieve their financial goals. They are at the forefront of innovation, constantly evolving their platforms and services to provide unparalleled value and security to their clients. This is an opportunity to join a dynamic team that is passionate about leveraging technology to drive business success and enhance the client experience.

We are seeking a highly motivated and experienced Site Reliability Engineer to join a pivotal team, partnering with our client to elevate the reliability and operational excellence of critical production systems. In this dynamic contract role, you will be instrumental in shaping the future of our infrastructure, driving automation, and ensuring seamless operations across both on-premises and cloud platforms. Your expertise will directly impact the stability, performance, and scalability of systems that serve millions, making a tangible difference in the daily lives of our clients. If you thrive on solving complex challenges, have an automation-first mindset, and are eager to contribute to a high-impact environment, this is your chance to shine.

Key Responsibilities

  • Develop Python-based automation solutions to reduce manual operational effort and enhance efficiency.
  • Automate infrastructure management across Linux, Windows, Kubernetes, cloud platforms, and cloud-native environments.
  • Integrate various tools and platforms through APIs and client libraries to streamline workflows.
  • Assist in implementing robust infrastructure automation using industry-standard technologies.
  • Support CI/CD automation and deployment reliability initiatives to ensure smooth and consistent releases.
  • Monitor and maintain production systems to consistently meet and exceed reliability and availability objectives.
  • Actively participate in incident response, thorough troubleshooting, and root cause analysis activities to prevent recurrence.
  • Develop proactive automation and operational improvements to eliminate recurring issues and improve system resilience.
  • Support disaster recovery, failover testing, and other operational readiness activities to ensure business continuity.
  • Perform in-depth performance analysis and regular system health reviews to optimize system behavior.
  • Build and maintain comprehensive dashboards, alerts, and monitoring solutions using Splunk, Grafana, Prometheus, or similar tools.
  • Improve visibility into application and infrastructure health through enhanced metrics, logs, and traces.
  • Investigate alerts thoroughly and identify opportunities to reduce noise and improve detection accuracy.
  • Explore AI/ML-driven operational improvements such as anomaly detection, intelligent alerting, and log analytics.
  • Assist in developing automation solutions that leverage AI to significantly improve operational efficiency.
  • Participate in evaluating emerging AIOps capabilities and cutting-edge observability technologies.

Required Qualifications

  • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
  • 3 to 5 years of hands-on experience in Site Reliability Engineering, Production Engineering, DevOps, Systems Engineering, or Platform Engineering, with a strong focus on production operations.
  • Must have supported production systems at scale, demonstrating a deep understanding of operational challenges in large environments.
  • Operations ownership, incident response, and reliability engineering experience are essential.
  • Strong programming skills in Python for automation and tooling development; Python must be a primary skill, not just basic scripting.
  • Demonstrated experience building automation tools, scripts, frameworks, or operational solutions, with the ability to provide examples of personally developed automation.
  • Experience supporting Kubernetes and cloud platforms (GCP, AWS, or Azure).
  • Familiarity with infrastructure automation and configuration management tools.
  • Experience with monitoring and observability platforms such as Splunk, Grafana, Prometheus, Datadog, or similar.
  • Understanding of Linux systems, networking, and distributed applications.
  • Strong analytical, troubleshooting, and problem-solving skills, with experience in critical production incident management and Root Cause Analysis (RCA).
  • Ability to work effectively in fast-paced, mission-critical environments, reducing operational toil through automation.

Preferred Qualifications

  • Experience with Terraform, Ansible, or other Infrastructure as Code solutions.
  • Exposure to OpenTelemetry and modern observability practices.
  • Experience with CI/CD pipelines and deployment automation.
  • Knowledge of AI/ML, AIOps, or intelligent operational tooling.
  • Experience supporting highly available production systems in regulated or enterprise environments.

About Aquent Talent

Aquent Talent connects the best talent in marketing, creative, and design with the world's biggest brands.

Our eligible talent get access to amazing benefits like subsidized health, vision, and dental plans, paid sick leave, and retirement plans with a match.

Aquent is an equal-opportunity employer. We evaluate qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, veteran status, and other legally protected characteristics. We're about creating an inclusive environment-one where different backgrounds, experiences, and perspectives are valued, and everyone can contribute, grow their careers, and thrive.


Nearby locations

Posting ID: 1297057443 Posted: 2026-09-16 Job Title: Site Reliability Engineer