Major Incident Manager Job at SonarSource, Austin, TX

  • SonarSource
  • Austin, TX

Job Description

Austin, TexasMission Control – IT Ops /Employee / Full-Time /On-siteWho is Sonar?Sonar is driving the future of agent-centric software development. As the leader in AI code verification and governance, we solve a critical problem: ensuring that software generated by AI-assisted developers or autonomous agents is reliable, secure, and maintainable.Integrating seamlessly with Claude Code, Codex, Cursor, GitHub Copilot, Gemini, and Devin, we help over 75% of the Fortune 100 build trusted, reliable, compliant software. Customers who use Sonar are 44% less likely to report an outage due to AI-generated code.We believe code verification is the critical missing link in the Agent-Centric Development Cycle (AC/DC). Industry giants like Nvidia, ServiceNow,Booking.com, Goldman Sachs, AstraZeneca, and Ford Motor Company count on us to provide independent, explainable, consistent review and governance of their AI-generated code via products like:SonarQube: The world’s leading AI code review and verification platform.SonarQube Foundation Agent: Currently topping the leaderboards for agentic software repair.SonarSweep & Sonar Context Augmentation: Providing the enterprise-grade context and constraints agents need to be truly effective.Our team operates across global hubs in Austin, Bochum, Dubai, Geneva, London, Singapore, Tokyo, and Washington D.C. We move with a mindset we call CODE:Committed to our customers and community.Obsessed with quality.Deliberate in our decisions.Effective as one team.With over $400M in revenue and profitable, fast-paced growth, we are building the backbone of the AI software revolution. If you’re hungry to have an impact, want to build at a fast pace, and ready to work at the forefront of AI, we want to hear from you.Position descriptionWe are still at the beginning of our growth journey, so we are putting new processes, technologies, and tools in place on a continuous basis. Your role is a pivotal engineering contributor to the tooling and services to automate and enhance the software development lifecycle, empowering our fellow SonarSourcers to deliver with speed, confidence, and security. You would be a member of a team that delivers solutions across all of our 5 offices: Austin (Texas, US), Geneva (Switzerland), Bochum (Germany) and Singapore.As a Major Incident Manager, you use and create automation tools to monitor and observe production infrastructure services both on premises and in the cloud. You are allergic to repetitive tasks, preferring to maximize automation and reliability. You are expert in change management, infrastructure management, system support, and configuration management. What you will doSystem Health Monitoring, Alert Triaging, and Error Budget Management: Dedicate time to monitoring critical security infrastructure (e.g., identity platforms, firewalls, compliance systems) and core infrastructure components. Focus on using and maintaining dashboards tied to Service Level Objectives (SLOs), triaging high-severity alerts, and analyzing the current Error Budget burn rate to guide prioritization for the rest of the day.Infrastructure as Code (IaC) and Policy as Code Development: Spend the largest portion of time writing, reviewing, and testing code (e.g., Python, Go, Terraform, or proprietary tools) to automate the deployment, configuration, and security hardening of infrastructure. This involves treating infrastructure and security policies as software to ensure consistency and prevent configuration drift.Toil Elimination and Automation of Operational Tasks: Identify, scope, and implement automated solutions for manual, repetitive, and time-consuming tasks (toil) related to security patching, compliance checks, certificate rotations, or infrastructure maintenance. The goal is to continuously reduce the operational workload for the team.Security Pipeline and Observability Maintenance: Maintain and enhance the DevSecOps security tools integrated into the CI/CD pipelines (e.g., static analysis, vulnerability scanning, security configuration checks). Ensure the end-to-end logging, metrics, and tracing (observability) systems for both infrastructure and security tools are robust, accurate, and provide immediate diagnostic capability during incidents.Incident Response Engineering and Post-Mortem Action: Participate in the on-call rotation and actively engage in engineering solutions derived from post-mortems. This means turning incident root causes into preventative measures implemented via code, improving runbooks into automated actions, and reducing Mean Time To Resolution (MTTR) for future incidents.Experience and qualificationsDeep IaC Expertise: Professional experience provisioning and managing complex infrastructure using tools like Terraform or CloudFormation (AWS), or similar tools like Ansible or Puppet for configuration management.Cloud/Platform Experience: Hands-on experience with a major cloud provider (AWS, GCP, Azure) or managing large-scale internal/private cloud infrastructure.SLO/SLI Implementation: Practical experience defining, measuring, and reporting on Service Level Indicators (SLIs) and Service Level Objectives (SLOs) for critical services.Logging/Metrics/Tracing Stacks: Proven experience with modern observability platforms (e.g., Prometheus/Grafana, ELK/EFK stack, proprietary systems, or vendor solutions like Datadog/Splunk) for proactive issue identification.Networking: Strong understanding of core networking concepts (TCP/IP, DNS, Load Balancing, Firewalls, Proxies) sufficient to debug complex service connectivity and latency issues.Automation of Security Controls: Experience implementing security best practices via code, such as automated vulnerability scanning, configuration hardening, secret management (e.g., HashiCorp Vault), and key rotation.Identity and Access Management (IAM): Practical experience managing large-scale IAM systems (e.g., implementing least-privilege policies, single sign-on).Incident Management: Experience running or significantly contributing to post-incident reviews (post-mortems) and prioritizing resulting engineering work (error budget management).In-office cultureWe're intentional about this. We believe the best teams are built in the room together. Three anchor days — Mondays, Tuesdays, and Thursdays — create the collaboration rhythm that makes a hub office worth having. Candidates need to be genuinely based in the location the role is posted — if that's not where you are today, we're happy to support relocation for the right person.We value diversity, equity, and inclusionAt Sonar, we believe that our diversity is our strength. We are a global company that values and respects different backgrounds, perspectives, and cultures. We are committed to fostering a diverse and inclusive work environment where everyone feels valued and empowered to contribute their best. We are proud to be an equal opportunity employer and welcome all qualified applicants, regardless of race, color, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status.If you need any accommodation, please reach out to us at [email protected]. All offers of employment at Sonar are contingent upon the results of a comprehensive background check and reference verification conducted before the start date. Applications that are submitted through agencies or third party recruiters will not be considered. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Job Tags

Full time, Work at office, Immediate start, Relocation

Similar Jobs

Proficient Auto Transport

CDL-A Auto Haulers Needed: Earn $70,000-$120,000+ per year, Regional and OTR Routes Available, Home Job at Proficient Auto Transport

 ...earnings potential (percentage pay)* Consistent home time depending on lane (local, regional...  ...to load and secure vehicles* Willing to work at least five (5) days per week within...  ...999 and based in Bound Brook, New Jersey, Delta Auto Transport built its reputation on speed... 

Insomnia Cookies

Bike Delivery Courier Job at Insomnia Cookies

 ...Bike Delivery Courier Insomnia Cookies is one of the fastest growing, late-night, sweet indulgence companies in the country, and at the present time, we are actively interviewing Bike Delivery Courier for our Charlotte Downtown store! As a Bike Courier, you are our... 

AAA Life Insurance Company

Senior Email Marketing Specialist Job at AAA Life Insurance Company

 ...perspective to influence, innovate, motivate, and thrive. Responsibilities What You'll Do AAA Life is seeking a Senior Email Marketing Specialist that will be responsible for leading the development and optimization of AAA Life's acquisition email campaigns from... 

James C. Standring, DDS

Office Manager Job at James C. Standring, DDS

 ...Be the Welcoming Face of James C. Standring, DDS Full-Time Office Manager Needed! At James C. Standring, DDS, we are committed to providing exceptional dental care in a warm, patient-focused environment. Our practice is built on compassion, integrity, and a dedication... 

Flextronics

IT Director, SOX Job at Flextronics

 ...infrastructure, delivering end-to-end power and thermal management solutions for AI data centers and other mission-critical applications.The IT Director, SOX Compliance is responsible for leading the global IT Sarbanes-Oxley (SOX) compliance program and ensuring the effectiveness...