Description
At Apple, we don’t just build products — we craft experiences that enrich people’s daily lives. The Application Production Support (APS) & Site Reliability Engineering (SRE) team is responsible for ensuring that mission-critical enterprise systems and customer-facing cloud services operate with high availability, resilience, and performance at massive scale. We are seeking a motivated Junior Application Support & Site Reliability Engineer with 1 to 2 years of experience to join our fast-paced, agile engineering team. In this role, you will keep mission-critical applications running reliably across multiple environments—triaging production incidents and debugging application, database, core system, and performance bottlenecks through to root cause and permanent remediation. You will partner cross-functionally with development, QA, database, and infrastructure teams to resolve large-scale production disruptions, eliminate operational toil through automation, and manage platform health using modern observability and emerging GenAI technologies. This is an exceptional opportunity for an engineer with a strong foundation in Linux, relational databases, and programming who wants to build a deep, high-impact career in production and reliability engineering at Apple.
Job Description :-
| Company: | Apple |
| Job Role: | Junior Application Support & Site Reliability Engineer |
| Batches: | 2021-2025 |
| Degree: | Bachelor’s degree |
| Experience: | 1 to 2 years |
| Location: | Bangalore, India |
| CTC/Salary: | ₹8-16 LPA (Expected) |
Responsibilities
- High-Availability Operations: Monitor the health, performance, and capacity of mission-critical applications and services across multi-region environments.
- Incident Triage & Response: Act as a first-line responder to production alerts, outages, and system anomalies, participating in on-call rotations and rapid-response workflows.
- End-to-End Troubleshooting: Investigate and debug complex application, core system, network, and performance bottlenecks through to resolution.
- Database Diagnostics: Perform database-level troubleshooting, query analysis, data verification, and performance tuning across PostgreSQL, Oracle DB, and MySQL environments.
- Root-Cause Analysis (RCA): Partner with development, QA, database, and infrastructure teams to identify underlying root causes and implement permanent fixes.
- Observability & Alerting: Build and refine telemetry dashboards, log aggregations, and alerting rules using tools such as Splunk, Prometheus, Grafana, and Dynatrace to eliminate blind spots.
- Service Level Tracking: Track and uphold key reliability metrics (SLIs/SLOs) to minimize Time to Detect (TTD) and Time to Mitigate (TTM).
- Toil Reduction & Automation: Develop robust scripts and tools (Python, Bash) to automate repetitive operational tasks, health checks, and routine maintenance.
- Release & CI/CD Enablement: Support production deployments, CI/CD pipeline runs (Jenkins, GitHub Actions), build validations (Maven, Gradle), and canary/smoke testing.
- Change & Configuration Governance: Execute production change requests (CRs) in accordance with enterprise change management and risk assessment standards.
- Continuous Improvement & Modern Tooling: Leverage modern engineering practices and GenAI-assisted workflows to accelerate log analysis, incident documentation, and operational efficiency.
Minimum Qualifications
- Experience: 1 to 2 years of combined work experience in Application Production Support (APS), DevOps, Systems Engineering, or Site Reliability Engineering (SRE).
- Linux Fundamentals & System Engineering: Solid understanding of Linux/Unix internals (filesystems, process management, networking, permissions, systemd) and core system engineering concepts (servers, services, resource management).
- Application & Build Ecosystem: Programming/debugging exposure in Java and Python, along with familiarity with Java-based build tools (Maven, Gradle).
- Database Troubleshooting: Strong debugging and querying skills with relational databases such as PostgreSQL, Oracle DB, and MySQL.
- Observability & Monitoring: Hands-on exposure to observability and logging platforms (e.g., Splunk, Grafana, Prometheus, Dynatrace).
- Incident & Change Management: Proven track record in incident management, triage, and adhering to standard change/configuration management processes.
- CI/CD Foundations: Basic understanding of CI/CD concepts and continuous deployment tools (e.g., Jenkins, GitHub Actions).
- Modern Tech Exposure: Exposure to GenAI solutions, AI models, and modern AI-assisted developer/support workflows.
- Bachelor’s degree or foreign equivalent in Computer Science, Computer Engineering, Information Technology or related field
Preferred Qualifications
- AI-Assisted Operations: Demonstrated ability to effectively leverage AI/GenAI tools to accelerate log analysis, script generation, and issue investigation.
- Containers & Orchestration: Hands-on experience with Docker containers and foundational knowledge of Kubernetes concepts.
- Cloud Infrastructure: Familiarity with public or private cloud platforms (AWS, GCP, or Azure).
- Infrastructure as Code (IaC): Basic awareness of configuration automation tools (e.g., Ansible, Terraform).
Apply Through This Link: Click Here
Join our Telegram group: Click here
Follow us on Instagram: Click here
Follow us on WhatsApp: Click here
