Senior Site Reliability Engineer

Senior Site Reliability Engineer

27 Sep 2024
Michigan, Lansing, 48901 Lansing USA

Senior Site Reliability Engineer

At Ford Motor Company, we believe freedom of movement drives human progress. We also believe in providing you with the freedom to define and realize your dreams. With our incredible plans for the future of mobility, we have a wide variety of opportunities for you to accelerate your career potential as you help us define tomorrow’s transportation.Site Reliability Engineering at Ford Motor Company plays a critical role in maintaining and improving the reliability, scalability, and performance of our services. As a Site Reliability Engineer, you will work closely with our development teams to build and maintain large-scale, distributed systems and ensure our products meet our high standards for availability and user experience.

Write, configure, and deploy code that improves service reliability for existing or new systems; set standard for others with respect to code quality

Provide helpful and actionable feedback and review for code or production changes

Drive repair/optimization of complex systems with consideration towards a wide range of contributing factors

Lead debugging, troubleshooting, and analysis of service architecture and design

Participate in on-call rotation

Write documentation: design, system analysis, runbooks, playbooks. Provide design feedback and uplevel design skills of others.

Implement and manage monitoring solutions using Dynatrace, Splunk, and OpenTelemetry to ensure visibility and proactive issue detection across our platforms.

Work within GCP infrastructure, optimizing performance, and cost, and scaling resources to meet demand.

Collaborate with development teams to enhance system reliability and performance, applying a platform engineering mindset to system administration tasks.

Develop and maintain automated solutions for operational aspects such as on-call monitoring, performance tuning, and disaster recovery.

Troubleshoot and resolve issues in our dev, test, and production environments.

Participate in postmortem analysis and create preventative measures for future incidents.

Bachelor’s degree in Computer Science, Engineering, or equivalent experience.

3+ years of experience as an SRE, DevOps Engineer, or in a similar role.

Strong experience with monitoring and observability tools, particularly Dynatrace and OpenTelemetry.

Proficient with cloud services, with a strong preference for Google Cloud Platform (GCP) experience.

Solid programming skills in Java, with a good understanding of software development best practices.

Experience managing and optimizing PostgreSQL databases.

Familiarity with front-end development frameworks, particularly React.

Ability to debug, optimize code, and automate routine tasks.

Strong problem-solving skills and the ability to work under pressure in a fast-paced environment.

Excellent verbal and written communication skills.

You may not check every box, or your experience may look a little different from what we've outlined, but if you think you can bring value to Ford Motor Company, we encourage you to apply!As an established global company, we offer the benefit of choice. You can choose what your Ford future will look like: will your story span the globe, or keep you close to home? Will your career be a deep dive into what you love, or a series of new teams and new skills? Will you be a leader, a changemaker, a technical expert, a culture builder…or all of the above? No matter what you choose, we offer a work life that works for you, including: Immediate medical, dental, and prescription drug coverage Flexible family care, parental leave, new parent ramp-up programs, subsidized back-up child care and more Vehicle discount program for employees and family members, and management leases Tuition assistance Established and active employee resource groups Paid time off for individual and team community service A generous schedule of paid holidays, including the week between Christmas and New Year’s Day Paid time off and the option to purchase additional vacation time.For a detailed look at our benefits, click here: Benefit Summary (https://urldefense.com/v3/https:/fordcareers.co/GSR-HTHD;NLtwI-RPugbI9wg0dJn!DsiF5SEXdZDVZwdZgnRezGpR7r347NCo7gj6QzTiwtjge5qTcqMFhCCblgcCy1VjavkencOcNa-g$)Visa sponsorship is available for this position.SOUTHEAST MI RESIDENTS: Please note, this job is posted as remote unless the selected candidates lives within 50 miles of Dearborn, MI. We request the candidate to be onsite 1-2 days a week.Candidates for positions with Ford Motor Company must be legally authorized to work in the United States. Verification of employment eligibility will be required at the time of hire.We are an Equal Opportunity Employer committed to a culturally diverse workforce. All qualified applicants will receive consideration for employment without regard to race, religion, color, age, sex, national origin, sexual orientation, gender identity, disability status or protected veteran status. In the United States, If you need a reasonable accommodation for the online application process due to a disability, please call 1-888-336-0660.# LI-RemoteRequisition ID : 34488

Related jobs

  • Cribl does differently. What does that mean? It means we are a serious company that doesn\'t take itself too seriously; and we\'re looking for people who love to get stuff done, and laugh a bit along the way. We\'re growing rapidly - looking for collaborative, curious, and motivated team members who are passionate about putting customers first. As a remote-first company we believe in empowering our employees to do their best work, wherever they are. As the data engine for IT and Security many of the biggest names in the most demanding industries trust Cribl to solve their most pressing data needs. Ready to do the best work of your career? Join the herd and unlock your opportunity. Why you\'ll love this role: Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission to unlock the value of all observability data. Cribl provides users a new level of observability, intelligence and control over their real-time data. You will join a team of technical engineers who are committed to shipping only high-quality software and enjoying all the goat gifs the internet has to offer. This role is remote and you will be part of the engineering organization where you will contribute in our efforts to envision, create, deploy, test, and ship Cribl products. Not often do you get to be part of something that is fundamentally changing a technology. But here at Cribl we are building the next generation of software that puts our customers in full control of their observability data. If this is something that interests you, and you want to be truly at the center of the wheel helping make this work better every day. Then this opportunity might be something you have been waiting for to be a part of making a real impact. We are looking for Cloud Site Reliability Engineers and Developers at all levels at Cribl, who enjoy being in the thick of it. Fixing things at the operational side should always be the last resort, so our SRE engineers are involved from conception to design to development and all the way through production and beyond. You provide your creative input into all things Cloud, Scaling, Reliability, High Availability and much more. If reliability is your passion, and you have always had strong opinions on how to make things better and have the desire to build consensus around ideas. Then let\'s talk! As An Active Member Of Our Team, You Will Engage with teams and improve service delivery and reliability across their entire lifecycle Measure and monitor all production systems with an eye towards availability, latency and overall system health Seek out the cause of errors and instability in our production cloud services and drive teams towards better operational excellence Engage with product and platform teams to improve and evolve systems by lobbying for changes that improve reliability, resilience, and observability Help Identify and drive down toil with creative innovation and automation On-call responsibilities If You Got It, We Want It Extensive experience with enterprise scale continuous delivery environments 5+ years of experience with a DevOps or SRE job title Development with JavaScript/Node.js/TypeScript in a Linux/Mac environment Experience with Configuration Management Tools like Terraform (preferred) or Puppet, Chef, Ansible Experience with sustainable incident response in a blameless environment Knowledge of cloud platforms (prefer AWS) and container + orchestration technologies Experience with APM and Observability and related tools such as, New Relic, Splunk, CloudWatch, Prometheus, Grafana/Kibana, Sentry etc. Background in Linux Systems Engineering Experience with Incident response related tools for instance, PagerDuty, FireHydrant, Blameless etc. Comfortable with a high level of autonomy and working with a distributed team Preferred Qualifications Knowledge of Cloud and application security Strong knowledge of cloud design patterns for scale, data management, resiliency,

  • Job Description

  • Job Description

  • Job Description

  • Job Description

  • Job Description

  • Job Number 24162788

Job Details

Jocancy Online Job Portal by jobSearchi.