Job description
Note The job is a remote job and is open to candidates in USA. Stage 4 Solutions is hiring a Site Reliability Engineer for a large, global B2B high-tech company. The role focuses on deploying software for cloud and SaaS customers, managing infrastructure reliability, responding to incidents, and providing advanced technical support and guidance to customers. Responsibilities Deploy software for Cloud Prem and SAAS customers Respond to and diagnose system incidents in a timely and efficient manner, minimizing downtime and impact on users Collaborate with other engineers to establish root causes and implement effective resolutions. Continuously improve incident response processes and documentation for future occurrences. Proactively monitor and maintain the health and performance of our infrastructure and services Perform routine administrative tasks such as system configuration, user management, and data backups. Identify and implement operational improvements to ensure ongoing system reliability and efficiency. Develop and implement scripts and automated solutions to streamline operational tasks and reduce manual workload Participate in the on-call rotation to address critical incidents outside of regular business hours Ensure effective handoff between on-call engineers and document post-incident information for future reference Document processes for support and create, maintain and execute run-books for identified situations Provide tier 2/3 technical support to customers experiencing platform issues or requiring advanced troubleshooting Work directly with customer technical teams to resolve complex deployment, configuration, and integration challenges Conduct technical onboarding sessions and provide guidance on best practices for customer implementations Collaborate with customer success teams to ensure smooth customer experiences and rapid issue resolution Create and maintain customer-facing technical documentation, troubleshooting guides, and knowledge base articles Escalate customer feedback and feature requests to product and engineering teams Participate in customer calls and technical discussions to provide expert-level platform guidance Track and analyze customer support metrics to identify trends and areas for improvement Skills 3+ years of experience in Site Reliability Engineering 2+ years of experience working with cloud platforms and cloud automation tools, especially in AWS Strong experience with Kubernetes, Helm, Linux, AWS networking(VPC), and Terraform Experience with the GitOps model for deployment Familiarity with distributed version control Experience with monitoring and alerting tools (e.g., Prometheus, Grafana) Understanding of software configuration best practices Ability to wear multiple hats in a fast-paced environment Hands-on, can do attitude and a bias for action Comfortable working across time zones to support a global customer base Excellent communication skills with the ability to explain technical concepts to both technical and non-technical audiences Strong customer service orientation with patience and empathy when working with frustrated customers BS degree in Computer Science or related field Bazel and CueLang experience a plus Benefits Health benefits 401K Remote role in the US Company Overview Stage 4 Solutions is a management consulting firm that provides marketing solutions services. It was founded in 2001, and is headquartered in Saratoga, California, USA, with a workforce of 51-200 employees. Its website is https//www.stage4solutions.com.
Prepare your application
Use the description above to assess your fit for RemoteNow. This checklist is guidance from Resumize, rather than additional requirements from the employer. Imported listings can be shortened or change after collection, so confirm the complete description and current availability before sending an application.
Match requirements to evidence
Identify the responsibilities and skills the employer explicitly names. For each important requirement, choose one example from your work, education, training, or projects. Explain what you did, how you did it, and what changed as a result. Include a measurable outcome when you can support it. If you lack a requirement, describe related experience accurately instead of adding an unfamiliar skill to your resume.
Put the most relevant examples near the top of your resume. Keep job titles, dates, qualifications, and contact details consistent across your application materials. Use the employer’s terminology when it describes your experience accurately; a copied list of keywords does not explain your contribution. A short, specific bullet is easier to review than a paragraph that mixes several unrelated duties.
Check the working arrangement
Confirm the location, schedule, employment type, and any eligibility requirements on the employer’s site. A remote label does not establish worldwide eligibility or flexible hours. If compensation, benefits, equipment, travel, or contract duration are missing from this listing, write down your questions for the recruiter. Do not assume details from another opening at the same company apply to this role.
Review before submitting
Follow the employer’s requested file format and application instructions. Open your exported resume and check that its text is selectable, its sections read in order, and its links work. Proofread names, dates, and contact information. Share confidential work samples only when you have permission, and use an anonymized example when necessary.
Submit through a trusted employer or recruiting destination. If the original listing has disappeared, search the employer’s careers page for the title rather than assuming the vacancy remains open. Save the application date, job reference, and the version of your resume you sent. Those details help you prepare for a later conversation and avoid submitting conflicting information through several job boards.
Review your resume for writing and readability feedback, or browse other openings if this opportunity is unavailable.