Find jobs
Pricing
Free Resume Checker
Sign in
Sign up
FIND JOBSPRICINGFREE RESUME CHECKERSIGN INSIGN UP
Jobs
/Software Engineer jobs
Walmart

Software Engineer III– AI Systems

SalaryNo pay range posted
LocationBentonville, AR; Bellevue, WA; (USA) Crossman Excellence Building CA SUNNYVALE Home Office
Work modeon-site
Typefull-time
SenioritySenior
Experience4+ yrs
DepartmentEngineering
Company size10,000+ people
First seenOct 8, 2026 · 3d ago
Verified live1d ago
Washington and California laws require most employers to post a pay range.
At a glanceSummarised by Seekless from the posting.
Must have10
Bachelor’s/Master’s in CS, Engineering, or equivalent industry experience
4+ years building production backend or platform services (preferably in AI/ML contexts)
Proficiency in Python (primary), plus one of Go/Java/C++ for performance services
Experience with distributed frameworks: Ray, Spark, or Dask
Experience with accelerated compute: RAPIDS (cuDF/cuML/cuGraph) and GPU-aware programming concepts
Proficiency in service frameworks: FastAPI/Flask (Python), K8s (Kubernetes) and containerization (Docker)
Strong foundations in data structures/algorithms, concurrency, networking, and systems design
Pragmatic problem solver with a bias for measurable outcomes (latency, throughput, reliability)
Excellent communicator able to translate between research goals and production constraints
Drives clarity in ambiguous problem spaces; mentors others and uplifts engineering standards
Nice to have6
Production experience with agent frameworks (e.g., LangGraph-style planners, tool-use patterns, retrieval and memory components)
Experience with vector databases (e.g., FAISS, Milvus, pgvector, Pinecone) and feature stores
Familiarity with LLM and embedding services, prompt/tooling patterns, and evaluation harnesses
Hands-on with Kubernetes, autoscaling (HPA/KEDA), and GPU scheduling/operators
Experience with profiling tools: PyTorch profiler, Nsight, line-profiler, Ray dashboard
Experience with vLLM, Triton Inference Server, ONNX Runtime, or TensorRT for high‑throughput inference
Skills
Python
Go
Java
C++
Ray
Spark
Dask
RAPIDS
cuDF
cuML
cuGraph
FastAPI
Flask
Kubernetes
Docker
LangGraph
FAISS
Milvus
pgvector
Pinecone
PyTorch
vLLM
Triton Inference Server
ONNX Runtime
TensorRT
Position Summary…
We’re seeking a Software Engineer to design and build AI-first systems with a focus on agentic AI, high performance data/compute frameworks, and scalable, production-grade services. You’ll work across model-driven features and platform layers—integrating LLMs/agents, orchestrating pipelines with Ray, accelerating data science workloads with RAPIDS, and delivering robust APIs and services that power high-impact AI applications at scale.
The ideal candidate blends strong software engineering fundamentals with practical ML systems exposure and a passion for performance, reliability, and developer experience.
What you’ll do…
Key Responsibilities AI Systems & Agentic Workflows
•
Build agentic AI services (planning, tool use, retrieval, feedback loops) and integrate them with internal systems and APIs.
•
Implement orchestration, memory, tooling, evaluation, and guardrails for agentic workflows.
•
Collaborate with DS/MLE partners to productionize models (LLMs, GNNs, embedding services) behind stable APIs and SDKs.
Accelerated Compute & Data Pipelines
•
Develop GPU‑accelerated pipelines using RAPIDS (cuDF/cuML/cuGraph) and optimize end‑to‑end performance.
•
Use Ray (or similar) for distributed compute, batch/stream processing, and scalable workflow orchestration.
•
Profile and optimize bottlenecks across CPU/GPU, memory, and I/O layers; implement caching, vectorization, and async patterns.
Service & Platform Engineering
•
Design and maintain reliable microservices for training/inference, vector indexing, and real-time decisioning.
•
Implement observability (tracing/metrics/logging), fault tolerance, auto-scaling, and cost-aware execution.
•
Create internal SDKs/CLIs to streamline developer workflows, testing, and reproducibility.
Quality, Security & MLOps Integration
•
Establish CI/CD for AI services (unit/integration/e2e tests, canaries, blue/green, rollback).
•
Integrate with feature stores, vector databases, artifact registries, and model catalogs.
•
Enforce security, privacy, and compliance (data minimization, PII handling, governance, auditability).
Collaboration & Influence
•
Partner with product, platform, and DS/MLE teams to align requirements, SLAs, and success metrics.
•
Document systems thoroughly; contribute to design reviews and engineering best practices.
•
Mentor peers on AI systems patterns, distributed compute, and performance engineering.
Minimum Qualifications
•
Bachelor’s/Master’s in CS, Engineering, or equivalent industry experience.
•
4+ years building production backend or platform services (preferably in AI/ML contexts).
•
Proficiency in:
•
Languages: Python (primary), plus one of Go/Java/C++ for performance services.
•
Distributed frameworks: Ray, Spark, or Dask.
•
Accelerated compute: RAPIDS (cuDF/cuML/cuGraph) and GPU-aware programming concepts (streams, memory).
•
Service frameworks: FastAPI/Flask (Python), K8s (Kubernetes) and containerization (Docker).
•
Strong foundations in data structures/algorithms, concurrency, networking, and systems design.
Preferred Qualifications
•
Production experience with agent frameworks (e.g., LangGraph-style planners, tool-use patterns, retrieval and memory components).
•
Experience with vector databases (e.g., FAISS, Milvus, pgvector, Pinecone) and feature stores.
•
Familiarity with LLM and embedding services, prompt/tooling patterns, and evaluation harnesses.
•
Hands-on with Kubernetes, autoscaling (HPA/KEDA), and GPU scheduling/operators.
•
Performance profiling: PyTorch profiler, Nsight, line-profiler, Ray dashboard.
•
Experience with vLLM, Triton Inference Server, ONNX Runtime, or TensorRT for high‑throughput inference.
Soft Skills & Leadership
•
Pragmatic problem solver with a bias for measurable outcomes (latency, throughput, reliability).
•
Excellent communicator able to translate between research goals and production constraints.
•
Drives clarity in ambiguous problem spaces; mentors others and uplifts engineering standards.
About Walmart Global Tech
Imagine working in an environment where one line of code can make life easier for hundreds of millions of people. That’s what we do at Walmart Global Tech. We’re a team of software engineers, data scientists, cybersecurity expert’s and service professionals within the world’s leading retailer who make an epic impact and are at the forefront of the next retail disruption. People are why we innovate, and people power our innovations. We are people-led and tech-empowered. We train our team in the skillsets of the future and bring in experts like you to help us grow. We have roles for those chasing their first opportunity as well as those looking for the opportunity that will define their career. Here, you can kickstart a great career in tech, gain new skills and experience for virtually every industry, or leverage your expertise to innovate at scale, impact millions and reimagine the future of retail. Walmart’s culture is a competitive advantage, and it’s fostered by being together. Working together in person allows us to collaborate, align quickly and innovate with greater speed. We use our campuses to create purposeful connection rooted in deepening understanding and investing in the development of our associates.
More roles at Walmart
(CAN) Grocery Associate(CAN) Produce Stocker(CAN) OMNI Customer Fulfillment Associate(CAN) Stock Unloader Associate(CAN) Overnight Associate
Seekless
Seek less: one search across companies' own career pages, instead of a dozen job boards.
Product
Find jobsCompaniesPricingFree Resume CheckerBlog
Legal
AboutPrivacyTermsCookies
© 2026 SEEKLESS. ALL RIGHTS RESERVED.BUILT WITH CARE