Job Information

Oracle Fusion Apps-As-A-Service (FAaaS) : SRE Role in BENGALURU, India

Job Description

Solve complex problems related to infrastructure cloud services and build automation to prevent problem recurrence. Design, write, and deploy software to improve the availability, scalability, and efficiency of Oracle products and services. Design and develop designs, architectures, standards, and methods for large-scale distributed systems. Facilitate service capacity planning and demand forecasting, software performance analysis, and system tuning.

Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible for the design and delivery of the mission critical stack, with focus on security, resiliency, scale, and performance. Authority for end-to-end performance and operability. Partner with development teams in defining and implementing improvements in service architecture. Articulate technical characteristics of services and technology areas and guide Development Teams to engineer and add premier capabilities to the Oracle Cloud service portfolio. Understand and communicate the scale, capacity, security, performance attributes, and requirements of the service and technology stack. Demonstrate clear understanding of automation and orchestration principles. Act as ultimate escalation point for complex or critical issues that have not yet been documented as Standard Operating Procedures (SOPs). Utilize a deep understanding of service topology and their dependencies required to troubleshoot issues and define mitigations. Understand and explain the affect of product architecture decisions on distributed systems. Professional curiosity and a desire to a develop deep understanding of services and technologies.

A BS or MS in Computer Science, or equivalent. Provides strategic and comprehensive complex business solutions to knowledge of server hardware and software configuration, networking, standard internet services, scripting languages, cloud computing patterns, technology security and compliance. Experience running large scale customer facing web services. Provides strategic and comprehensive complex business solutions to understanding of load balancing technologies and experience with development in programming languages, databases and big data stores, and container technologies. Work involves defining and documenting technical architecture of complex and highly scalable products. A minimum of 12+ years experience of running large scale customer facing web services.


Cloud Engineering SRE

Fusion Applications-As-A-Service (FAaaS)– Principal Software Engineer

Location : Bangalore / Hyderabad

At Oracle Cloud Infrastructure (OCI) , we build the future of the cloud for Enterprises as a diverse team of fellow creators and inventors. We act with the speed and attitude of a start-up, with the scale and customer-focus of the leading enterprise software company in the world.

Values are OCI’s foundation and how we deliver excellence. We strive for equity, inclusion, and respect for all. We are committed to the greater good in our products and our actions. We are constantly learning and taking opportunities to grow our careers and ourselves. We challenge each other to stretch beyond our past to build our future.

You are the builder here. You will be part of a team of really smart, motivated, and diverse people and given the autonomy and support to do your best work. It is a dynamic and flexible workplace where you’ll belong and be encouraged.

We offer unique opportunities for smart, hands-on engineers with the expertise and passion to solve difficult problems in distributed highly available services and virtualized infrastructure. At every level, our engineers have a significant technical and business impact designing and building innovative new systems to power our customer’s business-critical applications.

Fusion Applications-As-A-Service (FAaaS) is an exciting new team working on at the intersection of infrastructure and applications where we are leveraging OCI to transform some of the largest SaaS properties in the industry. We are making the existing Oracle's SaaS portfolios, including the multi-billion dollar revenue producing Fusion applications, become first class Oracle cloud citizens. To support the vision we are building a platform that manages end-to-end lifecycle, from provisioning to upgrade to terminate; and provides a self-service cloud experience to the customer for our SaaS product by leveraging the underlying OCI platform and services.

As a Principal Software Engineer on the FAaaS SRE team you will be responsible for the Security, Availability, Performance, Compliance, Cost/COGS, Change management, Monitoring, Emergency response and capacity planning . You will build innovative automated solutions and tools to help debug and resolve problems in production and prevent them from recurring. Further, you will proactively seek out system weaknesses and find ways to fix them before they cause production issues using monitoring data, watching trends, metrics and events leveraging a plethora of internal tooling at OCI.

About You:

  • You are an experienced cloud engineer with a proven track record of delivering high-scale, high-impact solutions

  • You are obsessed with the customer, always exceeding expectations

  • You have excellent communication skills. You can clearly explain complex technical concepts

  • You are a disciplined engineer who understands the importance of high standards, never satisfied with mediocrity and constantly striving for excellence

  • You are comfortable with ambiguity in a chaotic and fluid environment

  • You are passionate about technology and are not afraid to defend your opinions or position with peers/superiors

  • You work through metrics and logs to ensure that potential service inhibitors are identified before they are detected

  • You avoid logging into servers directly and prefer using automation and aggregation to manage them

  • You will work closely with Product Engineering teams to proactively identify and fix potential service issues

  • You will handle communication and notification on major site issues to the company and executive management team

  • You will document resolution run books and standard operating procedures

  • You will be the technical mentor and coach for a team of SRE’s and guide them on day-to-day engineering tasks while working on individual deliverables in collaboration with engineering leadership

Minimum Qualifications

  • 12- 15 years of experience developing and managing scalable, cloud-native distributed systems

  • Engineering Degree (BE/BTech/BS) or higher in Computer Science as educational Qualification

  • Ability to work in a collaborative, cross-functional team environment

  • Strong grasp of Computer Science concepts (data structures, algorithms, and programming paradigms)

  • Proficient in at Java/C++, Python and shell scripting tools

  • Proficient in one or more hyperscaler service providers Oracle Cloud Infrastructure (preferred), Azure, AWS, GCP

  • Experience with container orchestration like Kubernetes/Docker Swarm/Mesos, experience working on Helm Charts, etc.

  • Experience in day to day operations and general engineering efforts for scalability, availability and security

  • Experience in working with development teams to design scalable, robust systems using cloud architecture

  • Build automation using industry tools (like GitHub/Bitbucket, TeamCity/Hudson, Maven/Gradle, Terraform etc) to deploy services.

  • Experience with components of modern infrastructure like service discovery, secret storage, containerization, software-defined networking, etc.

  • Experience with production operations and best practices for putting quality code in production and troubleshoot issues when they arise

  • Able to effectively communicate technical ideas verbally and in writing (technical proposals, design specs, architecture diagrams and presentations)

  • Working in an Agile team implementing Scrum practices

  • Availability for 24 hour on-call rotation for service related issues

Preferred Qualifications

  • Experience in a fast-paced start-up environment

  • Experience building/managing control plane/data plane solutions for cloud native companies

  • Experience in diagnosing, troubleshooting and resolving performance issues in complex environments

  • Deep understanding of Unix-like operating systems

  • Prior experience in managing production services deployed on Oracle database, Fusion Middleware and Fusion Apps stack will be considered a big plus

  • Production experience with Cloud and ML technologies

  • Data science and machine learning knowledge would be helpful but not required

About Us

Innovation starts with inclusion at Oracle. We are committed to creating a workplace where all kinds of people can be themselves and do their best work. It’s when everyone’s voice is heard and valued, that we are inspired to go beyond what’s been done before. That’s why we need people with diverse backgrounds, beliefs, and abilities to help us create the future, and are proud to be an affirmative-action equal opportunity employer.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans status, age, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.