Walmart Global Tech logo

Principal, Software Engineer

Walmart Global Tech · Posted Oct 8

Software development, cloud infrastructure, and retail technology services for Walmart

Bangalore, Karnataka, IndiaFull-timeOnsiteLead/Staff12+ years₹80.0L–₹1.3Cr yearly76 applicants
Information TechnologyRetail TechnologySoftware DevelopmentCloud ComputingArtificial IntelligencePublic Company
Full time

Get a personal compatibility score

Add a resume for personal matches

About the role

GDF (Global Data Foundation) is part of Walmart's AI & Data organization, building the foundational data platforms that power AI, analytics, and decision-making across every Walmart banner and geography, from supply chain and merchandising to customer experience and enterprise AI agents. The newly created platform team is chartered with building next-generation data infrastructure at Walmart scale: reliable, high-throughput data pipelines, distributed processing systems, and platform capabilities that downstream teams depend on daily. This Principal Software Engineer role will design, build, and operate distributed platform services and APIs, contribute to Spark-based data engineering work, build AI/GenAI and agentic platform capabilities, and set technical direction across the broader GDF platform.

What you will do

  • Design, build, and operate distributed platform services and APIs using Java and Python, applied against Kubernetes or equivalent cloud-native infrastructure.
  • Apply domain-oriented design, system design, and low-level design (LLD) rigor to ambiguous platform problems - evaluating dependencies, trade-offs, failure modes, scalability, cost, and long-term durability before writing a line of code.
  • Flex into data engineering work when needed: contribute directly to Spark-based data pipelines and processing jobs, working comfortably in GDF's data engineering codebase alongside platform work.
  • Build AI, GenAI, and agentic-powered platform capabilities: intelligent developer tooling, workflow automation, agentic services, and AI-assisted operational tooling (monitoring, incident triage, capacity planning).
  • Own platform reliability, scalability, and observability - service-level objectives, distributed tracing, alerting, and incident response for services you build, with a strong Operational Excellence / Engineering Excellence discipline.
  • Guide architecture and system design decisions across microservices, event-driven systems, and enterprise integrations, with an explicit eye toward how platform choices affect downstream data engineering workloads.
  • Partner with data engineers, data scientists, and ML engineers to ensure platform services expose the right primitives for data pipelines, feature stores, and model-serving paths.
  • Establish engineering standards through architecture reviews, design reviews, code reviews, and production readiness reviews for a team building process from scratch.
  • As a Principal engineer, set technical direction and best practices across multiple teams and the broader GDF platform without needing a management title - mentor other engineers and influence GDF's broader platform engineering standards.
  • Continuously evaluate emerging platform, cloud, and AI/agentic technologies, adopting them pragmatically where they improve reliability, productivity, or business value.

Skills used in this role

JavaPythonKubernetesGoC++ScalaSparkHadoopHDFSHiveKafkaMicroservicesEvent-Driven ArchitectureCI/CDContainerizationInfrastructure AutomationGenAILLMEmbeddingsVector SearchMLOpsObservabilityDistributed TracingDistributed SystemsAPI Design

What the employer is looking for

  • Option 1: Bachelor's degree in computer science, computer engineering, computer information systems, software engineering, or related area and 5 years' experience in software engineering or related area.
  • Option 2: 7 years' experience in software engineering or related area.
  • Genuinely hands-on coding: 12+ years of professional software engineering experience building and operating large-scale distributed platforms; strong proficiency in Java and Python is required - comfortable writing, reviewing, and debugging production code, not just directing others. Additional experience with Go, C++, or Scala is a plus.
  • Operational Excellence (OE) and Engineering Excellence (EE) mindset is a required skill, not a nice-to-have: SLO/SLA ownership, incident management and blameless post-incident reviews, proactive reliability and quality metrics, and continuous improvement built into how you ship, not bolted on afterward.
  • Bachelor's or Master's degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
  • Deep experience with microservices architectures, distributed systems, event-driven patterns, API design, service contracts, and enterprise system integration - backed by strong system design, domain-oriented design (DDD), and low-level design (LLD) skills, with demonstrated ownership of production code and the judgment to evaluate trade-offs, failure modes, scalability, and cost.
  • Deep understanding of cloud-native engineering practices - CI/CD, containerization, Kubernetes or equivalent orchestration, infrastructure automation, and production operations.
  • Strong exposure to Big Data / data engineering technologies - Spark required at a working level, with familiarity in Hadoop/HDFS, Hive, or Kafka - enough to flex into data engineering work confidently, not just integrate with it from a distance.
  • Full exposure to AI and Agentic technologies: hands-on experience integrating model APIs and GenAI/LLM tooling, building or operating agentic workflows (tool-calling agents, multi-step autonomous systems), working with embeddings/vector search, and applying MLOps practices (feature stores, model lifecycle, retraining pipelines) - not just conceptual awareness.
  • Strong knowledge of reliability engineering - observability, distributed tracing, logging, metrics, alerting, incident response, capacity planning, and fault tolerance for large-scale systems.
  • Strong grounding in security, privacy, and enterprise engineering practices, including authentication, authorization, data protection, and secure API design.
  • Proven ability to translate ambiguous technical/business problems into clear execution plans, particularly across a newly formed platform organization without fully established process.
  • Strong ownership mindset and technical judgment; comfortable operating with high autonomy in a 0-to-1 environment.

Benefits and support

  • Incentive awards for performance
  • Maternity and parental leave
  • PTO
  • Health benefits

About Walmart Global Tech

Walmart Global Tech is the technological powerhouse and engineering arm of Walmart Inc., responsible for building and managing foundational technologies, cloud infrastructure, and software systems. The organization powers global retail operations across Walmart U.S., Sam's Club, and Walmart International, delivering customer experiences and enterprise services worldwide. Operating across multiple global hubs, it unites engineering, data science, product design, and artificial intelligence to digitally transform retail.

Industry
Information Technology
Company size
25000+ employees
Founded
2011
Location
Bentonville, Arkansas, USA
Funding stage
Public Company

Leadership

SK
Suresh Kumar

Executive Vice President, Global Chief Technology Officer and Chief Development Officer

DD
Daniel Danker

Executive Vice President, AI Acceleration, Product and Design

SR
Sanjay Radhakrishnan

Senior Vice President and Tech Operating Partner, Walmart International