Description
Job Description - IT Administrator
Job Summary
Location: Hyderabad
Employment Type: Full-Time
We are looking for an experienced IT Administrator – Cloud & Infrastructure to manage, maintain, secure, and optimize enterprise IT infrastructure and cloud environments. The ideal candidate should have hands-on experience in AWS infrastructure, Linux/server administration, cloud operations, automation, containerization, monitoring, networking, and infrastructure management. The role will involve ensuring high availability and reliability of infrastructure, supporting application environments, automating operational activities, implementing security controls, monitoring system performance, and collaborating with development and IT teams for smooth deployments and production operations.
Key Skills
- Cloud Platform: AWS – EC2, S3, EBS, ELB/ALB/NLB, Auto Scaling, IAM, VPC, RDS
- Server & Infrastructure Administration: Apache, Tomcat, Linux/Unix environments
- Containerization: Docker, Docker Compose, Kubernetes, AWS ECS, AWS EKS
- Infrastructure as Code: Terraform
- Configuration Management: Ansible
- CI/CD: Jenkins, GitHub Actions
- Monitoring & Logging: Prometheus, Grafana, Amazon CloudWatch, Datadog
- Scripting & Automation: Bash, Shell Scripting, Python
- Networking & Security: VPC, IAM, Security Groups, iptables, SSL/TLS, VPN
- Security & Secrets Management: AWS Secrets Manager, SSM Parameter Store, AWS Config
- Version Control: Git, GitHub
- Tools: AWS CLI, kubectl, Terraform CLI, SonarQube, Nexus, YAML, JSON
Roles & Responsibilities
- Administer and maintain AWS cloud infrastructure and enterprise IT environments, including EC2, S3, EBS, ELB, Auto Scaling, IAM, and VPC.
- Ensure infrastructure availability, reliability, scalability, and performance while supporting application migration from on-premise environments to AWS.
- Manage and support Apache and Tomcat application servers, including provisioning, configuration, troubleshooting, and maintenance.
- Ensure consistent server configurations across development, staging, and production environments using Ansible.
- Provision and manage cloud infrastructure using Terraform and maintain reusable Infrastructure-as-Code modules.
- Manage development, staging, and production environments, including resource optimization, capacity management, and cloud cost optimization.
- Deploy and manage containerized applications using Docker and AWS ECS.
- Support Kubernetes/EKS environments, ensuring high availability, scalability, fault tolerance, and efficient container operations.
- Monitor container and cluster health and troubleshoot infrastructure-related issues to support zero-downtime deployments.
- Automate infrastructure and administrative tasks using Bash, Shell Scripting, Python, AWS Lambda, and EventBridge.
- Develop automation for scheduled infrastructure operations and resource management to reduce manual intervention.
- Monitor infrastructure, applications, and cloud resources using Prometheus, Grafana, CloudWatch, and Datadog.
- Analyze system alerts, logs, and performance metrics to identify and resolve infrastructure and production issues.
- Implement monitoring and recovery mechanisms to improve system availability and reduce incident response time.
- Implement and maintain AWS IAM roles, policies, Security Groups, and VPC security controls.
- Manage SSL/TLS configurations and network security controls while supporting infrastructure governance, compliance, and auditing requirements.
- Secure application credentials and sensitive information using AWS Secrets Manager and SSM Parameter Store.
- Manage and support CI/CD pipelines using Jenkins and GitHub Actions across development, staging, and production environments.
- Coordinate application deployments, troubleshoot deployment failures, and support rollback procedures.
- Collaborate with development teams to ensure smooth, reliable, and timely application releases.
- Troubleshoot connectivity, access, server, application, network, and infrastructure-related issues.
- Work with VPCs, load balancers, VPNs, Security Groups, and network-level configurations.
- Perform root-cause analysis of infrastructure incidents and implement preventive measures.
- Maintain infrastructure configurations, operational procedures, deployment documentation, and technical runbooks.
- Collaborate with application developers, DevOps engineers, security teams, and other IT stakeholders.
- Participate in production support, infrastructure upgrades, and continuous improvement initiatives.
Preferred Qualifications
- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
- 4+ years of experience in IT Infrastructure, Cloud Administration, System Administration, or DevOps.
- Experience with Linux/server administration, networking, security, and automation.
- Good troubleshooting and problem-solving skills.
- Ability to work in production environments and handle infrastructure incidents.
- AWS Certified Solutions Architect – Associate or equivalent certification is an added advantage.
Key Competencies
Cloud Administration | IT Infrastructure | AWS | Server Administration | Linux | Networking | Security | Infrastructure Automation | Monitoring | Incident Management | Docker | Kubernetes | Terraform | Ansible | CI/CD | System Reliability