About TCS
Tata Consultancy Services (TCS) is a global leader in IT services, consulting, and business solutions. With a presence in over 55 countries and a workforce of more than 600,000 associates worldwide, TCS partners with leading organizations to drive innovation, accelerate digital transformation, and achieve business growth through cutting-edge technologies, including Artificial Intelligence (AI), Cloud Computing, Data Analytics, Cybersecurity, and Software Engineering.
Entry-Level Site Reliability Engineer (SRE)
Required Skills
0-2 years of experience, including relevant internship, academic project, or hands-on experience in Site Reliability Engineering, DevOps, Cloud Technologies, Production Support, or Infrastructure Operations.
Basic understanding of Linux administration, system operations, and troubleshooting techniques.
Familiarity with AWS cloud services, including EC2, S3, RDS, IAM, and VPC.
Knowledge of monitoring, logging, observability, and system health management concepts.
Basic understanding of incident management processes and troubleshooting methodologies.
Familiarity with Shell scripting and/or Python programming.
Exposure to relational database systems such as PostgreSQL or MySQL.
Knowledge of Git and version control best practices.
Understanding of CI/CD principles and tools such as Jenkins.
Strong analytical, problem-solving, communication, and collaboration skills.
Roles & Responsibilities
Assist in maintaining highly available, scalable, and reliable production environments.
Monitor applications, services, and infrastructure to proactively identify issues and support incident resolution activities.
Support Linux server administration, troubleshooting, routine maintenance, and operational tasks.
Assist in provisioning, managing, and supporting AWS cloud resources and services.
Participate in Root Cause Analysis (RCA) activities and contribute to problem resolution initiatives.
Develop and maintain automation scripts to improve operational efficiency and reduce manual effort.
Support software deployment processes and CI/CD pipeline operations.
Collaborate with development, QA, cloud, and support teams to enhance system reliability, stability, and performance.
Assist with database monitoring, performance tracking, and basic database administration tasks.
Adhere to best practices related to automation, security, observability, reliability, and operational excellence.
Participate in production support activities and on-call rotations as required.