Senior System Software Engineer - Github
NVIDIA
India, PunePosted 2 months ago
Skill Required
Software EngineerSoftware DeveloperGitCI/CDSystem DesignMachine LearningDeep LearningElasticsearchEngineeringR ProgrammingKubernetesData StructuresAndroidMySQLLinuxCloudJavaCIFulltime
Key highlights
- 5+ years of proven experience
- BS/MS in Computer Science, Computer Engineering, or equivalent experience
- Work on a large‑scale private cloud system handling millions of jobs daily
- Opportunity to develop with Java, Kubernetes, and Machine Learning technologies
- Collaboration across NVIDIA divisions (Graphics, Mobile, AI, Autonomous Vehicles)
- Equal opportunity employer with accommodations for disabilities
Role overview
NVIDIA is seeking a senior software developer to contribute to a large‑scale private cloud system hosted on GitHub and Kubernetes that provides Continuous Integration services for multiple NVIDIA teams. The role involves building scalable cloud solutions, tackling infrastructure challenges, and collaborating across diverse groups such as Graphics, Mobile, Deep Learning, AI, and Autonomous Vehicles.
Responsibilities
- Build creative, scalable cloud solutions to handle millions of jobs and thousands of systems
- Tackle challenging problems in infrastructure such as job scheduling, resource management, and automated recovery
- Develop complete solutions including Metrics, Alert, and Storage Services
- Dig into data, analyze it extensively, and apply deep learning algorithms/machine learning to improve system performance and predictability
- Contribute to our GitHub-based CI workflow to streamline and optimize processes
Requirements
- Strong object-oriented programming background, with a preference for Java
- Proven experience in developing large-scale cloud infrastructure applications
- Knowledge of various technologies including Kubernetes and Message brokers
- Experience with relational databases like MySQL, and NoSQL databases such as Elasticsearch
- Ability to work effectively with various teams across different time zones
- BS/MS in Computer Science, Computer Engineering, or equivalent experience
- 5+ years of proven experience
Nice to have
- Real-world experience with distributed systems, containers, and Kubernetes API
- Proficiency in computer algorithms with the capability to select the most suitable algorithms for complex problems
- Skill in breaking down complex problems into manageable sub-problems and reusing solutions effectively
- Experience in crafting, implementing, and deploying major infrastructure features across multiple servers with incremental rollout
- Proficiency in Machine Learning and Data Analytics, and their application in Infrastructure as well as the ability to build simple systems that operate efficiently with minimal support
Additional details
- As a team, we work with various groups within NVIDIA such as Graphics Processors, Mobile Processors, Deep Learning, Artificial Intelligence, and Autonomous Vehicles to meet their diverse infrastructure needs.
- These cloud services will be scaled to run on thousands of servers and complete millions of automated jobs per day.
- We host a heterogeneous mix of machines with various operating systems (Windows/Linux/Android) and multiple hardware platforms (x86/ARM) featuring both NVIDIA GPUs and Tegra Processors.
- NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward‑thinking and dedicated people in the world working for us.
- We are an equal opportunity employer and value diversity at our company. We do not discriminate based on race, religion, color, national origin, sex, gender, gender expression, sexual orientation, age, marital status, veteran status, or disability status. We provide reasonable accommodations to individuals with disabilities. These accommodations help with the application process, essential job functions, and other employment benefits. Please contact us to request accommodation.