Powered by pgvector · cosine kNN
About the role
At Google, we have a vision of empowerment and equitable opportunity for all Aboriginal and Torres Strait Islander peoples and commit to building reconciliation through Google’s technology, platforms and people and we welcome Indigenous applicants. Please see our Reconciliation Action Plan for more information.
In most instances, this position requires in-person interviews as part of the hiring process.Minimum qualifications:
Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages.3 years of experience with managing people or teams.3 years of experience with leading projects.3 years of experience in designing, analyzing, and troubleshooting distributed systems.
Preferred qualifications:
Master's degree in Computer Science or Engineering.Experience in problem solving and analyzing distributed systems.Experience with mobile development, application deployment.Experience with algorithms, data structures, analysis and software design or in Unix/Linux systems, Internet Protocol (IP) networking, performance and application issues.Ability to perform technical analysis across code, networking, operating systems, and storage, while maintaining the cognitive and verbal agility to manage discussions with executive leadership.Ability to manage the strategy while providing technical guidance to the team, enabling them to execute and deliver products on time and within budget.
About the jobSite Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.
Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation. On the SRE team, you’ll have the opportunity to manage the complex challenges of scale which are unique to Google, while using your expertise in coding, algorithms, complexity analysis and large-scale system design.
SRE's culture of intellectual curiosity, problem solving and openness is key to its success. Our organization brings together people with a wide variety of backgrounds, experiences and perspectives. We encourage them to collaborate, think big and take risks in a blame-free environment. We promote self-direction to work on meaningful projects, while we also strive to create an environment that provides the support and mentorship needed to learn and grow.
To learn more: check out our books on Site Reliability Engineering or read a career profile about why a Software Engineer chose to join SRE.
Android Site Reliability Engineering (SRE) manages the mission-critical infrastructure powering the global Android ecosystem of devices. The mission is to bridge the reliability gap between mobile and web platforms, building user trust through availability. We empower Product and Development teams to scale securely and seamlessly.Behind everything our users see online is the architecture built by the Technical Infrastructure team to keep it running. From developing and maintaining our data centers to building the next generation of Google platforms, we make Google's product portfolio possible. We're proud to be our engineers' engineers and love voiding warranties by taking things apart so we can rebuild them. We keep our networks up and running, ensuring our users have the best and fastest experience possible.
Responsibilities
Lead software and systems engineers through planning, technical execution, and quality delivery. Grow engineering talent through mentoring and coaching strategies.Manage end-to-end availability and performance for mission-critical services while building automation to prevent recurrence.Collaborate with distributed partner teams and manage international on-call rotations.Align with Product and Developer teams to define and deliver Service Level Objectives (SLOs) that ensure reliability.Drive projects by leveraging existing frameworks and providing leadership in changing environments.
Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form .
Your match
See how you fit
Scored against this job in seconds
Your account
Sign in to apply
Your profile and your match for this job appear right here.
By continuing you agree to our Terms and Privacy Policy.
Next step
Apply for this role
via careers.google.com — opens their site
Applications via careers.google.com
Your job hunt, handled
Ask about any role and get a straight answer on your fit. Then stop searching: new matches land in your WhatsApp the moment they’re listed.
Free for jobseekers
Commonwealth Bank
Join us and lead our Site Reliability Engineering team to define and deliver world class observability and reliability tooling and practices at Australia’s largest bank. Position the core Site Reliability team to supp…
Universal Music Australia Pty Limited
Job Summary We are UMG, the Universal Music Group. We are the world’s leading music company. In everything we do, we are committed to artistry, innovation and entrepreneurship. We own and operate a broad array of busi…
IMPROVE - 3Rs concepts to improve the quality of biomedical science (CA21139)
🚀 WE'RE HIRING💼 SITE RELIABILITY ENGINEER (SRE)📍 AUSTRALIA🕒 FULL-TIME✨ ENSURE UPTIME. AUTOMATE OPERATIONS. BUILD SCALABLE SYSTEMS.We are looking for a talented and proactive Site Reliability Engineer (SRE) to join…
Leidos
We’re a ‘Family Friendly’ certified workplace – we understand the often many and varied roles our team members need to play within their own unique family setting and actively support them. Our team feel Leidos is a g…
Arista Networks
Company Description Arista Networks is an industry leader in data-driven, client-to-cloud networking for large data center, campus and routing environments. What sets us apart is our relentless pursuit of innovation. …
McMillan Shakespeare (MMSG)
This is a great opportunity to establish and lead Site Reliability Engineering at Macmillan Shakespeare. Here’s how you will make a difference in this role… You'll start as a hands-on technical leader, designing and b…