Back to jobs
New

Senior Software Engineer, AI Memory Systems & Resource Management (Architecture)

Houston, TX

About HP

At HP, you’ll have a chance to create tools, technology, and solutions that reshape the way the world works in the future. If you’re looking to join a company that allows you to connect with a network of professionals eager to support you in doing your best work, we want to talk to you. Our legendary culture guides every employee toward success—fostering collaboration and driving innovation.

Operating in over 170 countries, HP is always creating new services, products, and capabilities giving you more opportunities to advance your career. Here, innovation is the key to professional development and career mobility.

When you join HP, you’re joining a company that believes every voice matters and that we all deserve a seat at the table. From the boardroom to factory floor, we create a culture where everyone is respected and where people can be themselves. You will be part of a global laboratory where different perspectives and experiences will help you solve problems in new ways. This is where you can build a long and wide-ranging career.

Job Summary

The Commercial Systems Software Engineering organization is developing HP's Memory Management platform, a foundational component of the Agentic AI Software Stack designed to optimize memory utilization, model execution, and resource orchestration for next-generation AI PCs.

The Senior Software Engineer will define the architecture and technical strategy for AI memory management across heterogeneous compute environments, including CPUs, GPUs, NPUs, storage, and cloud resources. The role will focus on enabling efficient execution of large language models, multi-agent workloads, multimodal AI applications, and enterprise AI scenarios through intelligent memory allocation, model residency management, KV cache optimization, workload prioritization, and Quality of Service (QoS) controls.

This position requires deep expertise in operating systems, memory architectures, AI runtime systems, workload scheduling, performance optimization, and large-scale software architecture. The architect will work closely with silicon vendors, operating system partners, AI framework teams, and product organizations to develop innovative technologies that maximize AI performance while maintaining system responsiveness, security, and power efficiency.

Responsibilities

  • Define and drive the end-to-end architecture for HP's Memory Management platform.

  • Architect systems for intelligent memory allocation, prioritization, and optimization across AI workloads running on client devices.

  • Lead design of model residency management capabilities to enable efficient deployment and execution of large AI models on resource-constrained systems.

  • Define policies for KV cache management, context window optimization, model sharing, and memory multiplexing across multiple AI agents.

  • Develop policy engines that dynamically balance performance, memory consumption, power usage, and user responsiveness.

  • Architect workload QoS and resource arbitration mechanisms across CPU, GPU, NPU, memory, storage, and network resources.

  • Define strategies for local, hybrid, and cloud model execution based on system state, workload requirements, and resource availability.

  • Lead development of telemetry and analytics systems that monitor memory pressure, model utilization, resource contention, and workload efficiency.

  • Partner with operating system, firmware, silicon, and AI framework teams to optimize memory behavior across the full platform stack.

  • Establish technical roadmaps, engineering standards, and architectural guidance across multiple software teams.

  • Mentor senior engineers and architects while influencing company-wide AI platform strategy.

  • Engage with ecosystem partners including Microsoft, Intel, AMD, Qualcomm, NVIDIA, and ISVs to align memory optimization technologies and industry standards.

Education & Experience Recommended

  • Four-year or Graduate Degree in Computer Science, Computer Engineering, Electrical Engineering, Software Engineering, or related technical field, or equivalent experience.

  • 12+ years of industry experience in operating systems, systems software, platform architecture, AI infrastructure, memory management, or performance engineering.

  • Proven track record delivering scalable platform technologies, runtime systems, or enterprise software architectures.

  • Experience building software that operates close to hardware and operating system layers.

Core Technical Expertise

Memory Systems Architecture

  • Deep expertise in operating system memory management concepts including virtual memory, paging, allocation strategies, memory pressure management, caching, NUMA architectures, and resource isolation.

  • Advanced understanding of DRAM architectures, memory controllers, UMA designs, storage hierarchies, and memory subsystems.

  • Experience optimizing memory behavior for high-performance applications and distributed systems.

AI Memory Optimization

  • Expertise in AI model memory management including:

    • Model residency management

    • Memory budgeting

    • Model quantization

    • Dynamic model loading and unloading

    • Context window optimization

    • Hybrid local/cloud model execution

  • Experience reducing AI memory footprints while maintaining model accuracy and performance.

LLM Runtime Systems

  • Strong knowledge of:

    • KV cache optimization

    • Context management

    • Token generation pipelines

    • Model serving architectures

    • Multi-model execution frameworks

    • Agentic AI runtimes

  • Experience supporting concurrent AI agents sharing compute and memory resources.

Resource Governance & Scheduling

  • Expertise in workload scheduling, admission control, resource allocation, and QoS management.

  • Experience coordinating workloads across heterogeneous compute engines including CPU, GPU, NPU, and cloud services.

  • Understanding of power-aware and thermal-aware resource management techniques.

Performance Engineering

  • Expert-level understanding of performance profiling, bottleneck analysis, workload characterization, and observability systems.

  • Experience building telemetry platforms that generate insights from large-scale performance data.

  • Ability to model and predict system behavior under varying workload conditions.

AI Infrastructure & Platform Skills

  • Experience with AI runtimes such as ONNX Runtime, OpenVINO, DirectML, TensorRT, CUDA, ROCm, Qualcomm AI SDKs, or equivalent technologies.

  • Experience optimizing AI inferencing pipelines on client computing platforms.

  • Familiarity with model compression techniques including quantization, pruning, distillation, and low-rank adaptation methods.

  • Knowledge of agentic AI frameworks and orchestration systems.

Programming & Development Skills

  • Expert proficiency in C++ and Python.

  • Strong systems programming experience.

  • Experience developing operating system services, runtime systems, middleware, or platform software.

  • Familiarity with Windows internals and client platform architecture.

  • Experience building large-scale analytics, telemetry, and automation solutions.

Leadership & Collaboration Skills

  • Demonstrated technical leadership driving highly complex platform initiatives.

  • Ability to align diverse engineering organizations behind a common architectural vision.

  • Strong cross-functional collaboration skills spanning hardware, software, firmware, product, and partner organizations.

  • Exceptional communication skills with the ability to influence senior leadership and external partners.

  • Proven mentoring and talent development experience.

Preferred Qualifications

  • Experience working on operating systems, memory managers, hypervisors, container runtimes, or distributed resource management systems.

  • Prior work on AI PCs, edge AI systems, client computing platforms, or performance-critical infrastructure.

  • Contributions to memory optimization, AI infrastructure, operating systems, or runtime technologies through patents, publications, or open-source projects.

Impact & Scope

This role will define how future HP AI PCs efficiently execute increasingly large and complex AI workloads. The architect will establish the foundational technologies that enable multi-agent systems, local AI inferencing, context persistence, and intelligent resource sharing while ensuring responsive user experiences. The work directly influences HP's AI PC differentiation strategy, ecosystem partnerships, and long-term software platform roadmap.

Complexity

Provides company-wide architectural leadership for one of the most technically challenging areas of the Agentic AI Software Stack. Solves highly complex problems involving memory management, AI model execution, workload orchestration, heterogeneous computing, operating systems, and performance optimization across hardware and software boundaries. The role is expected to create industry-leading innovations in AI memory efficiency and resource governance that become core differentiators for HP AI PCs.

Salary: $154,400 - $227,750 

 

Compensation & Benefits (Full-Time Employees)

The salary range for this role is listed above. Final salary offered is based upon multiple factors including individual job-related qualifications, education, experience, knowledge and skills.

At HP, we offer a competitive and comprehensive benefits package, including:

  • Health insurance
  • Dental insurance
  • Vision insurance
  • Long term/short term disability insurance
  • Employee assistance program
  • Flexible spending account
  • Life insurance
  • Generous time off policies, including; 
    • 4-12 weeks fully paid parental leave based on tenure
    • 11 paid holidays
    • Additional flexible paid vacation and sick leave (US benefits overview)

Why join HP?

When you join HP, you’re investing in your future—and so are we. You have priorities beyond work. That’s why we offer flexibility, support, and benefits that help you shape the life you want. 

Equal Opportunity Employer (EEO) Statement

HP, Inc. provides equal employment opportunity to all employees and prospective employees, without regard to race, color, religion, sex, national origin, ancestry, citizenship, sexual orientation, age, disability, or status as a protected veteran, marital status, familial status, physical or mental disability, medical condition, pregnancy, genetic predisposition or carrier status, uniformed service status, political affiliation or any other characteristic protected by applicable national, federal, state, and local law(s).

Please be assured that you will not be subject to any adverse treatment if you choose to disclose the information requested. This information is provided voluntarily. The information obtained will be kept in strict confidence.

Apply for this job

*

indicates a required field

Phone
Resume/CV

Accepted file types: pdf, doc, docx, txt, rtf

Cover Letter

Accepted file types: pdf, doc, docx, txt, rtf


Select...
Select...