3h ago
New
USD 154400-227750/yr

Senior Software Engineer, AI Memory Systems & Resource Management (Architecture)

United StatesUnited States·Houstonsenior
OtherSoftware Engineer Ai
2 views0 saves0 applied

Quick Summary

Overview

About HP At HP, you’ll have a chance to create tools, technology, and solutions that reshape the way the world works in the future.

Technical Tools
OtherSoftware Engineer Ai

At HP, you’ll have a chance to create tools, technology, and solutions that reshape the way the world works in the future. If you’re looking to join a company that allows you to connect with a network of professionals eager to support you in doing your best work, we want to talk to you. Our legendary culture guides every employee toward success—fostering collaboration and driving innovation.

Operating in over 170 countries, HP is always creating new services, products, and capabilities giving you more opportunities to advance your career. Here, innovation is the key to professional development and career mobility.

When you join HP, you’re joining a company that believes every voice matters and that we all deserve a seat at the table. From the boardroom to factory floor, we create a culture where everyone is respected and where people can be themselves. You will be part of a global laboratory where different perspectives and experiences will help you solve problems in new ways. This is where you can build a long and wide-ranging career.

The Commercial Systems Software Engineering organization is developing HP's Memory Management platform, a foundational component of the Agentic AI Software Stack designed to optimize memory utilization, model execution, and resource orchestration for next-generation AI PCs.

The Senior Software Engineer will define the architecture and technical strategy for AI memory management across heterogeneous compute environments, including CPUs, GPUs, NPUs, storage, and cloud resources. The role will focus on enabling efficient execution of large language models, multi-agent workloads, multimodal AI applications, and enterprise AI scenarios through intelligent memory allocation, model residency management, KV cache optimization, workload prioritization, and Quality of Service (QoS) controls.

This position requires deep expertise in operating systems, memory architectures, AI runtime systems, workload scheduling, performance optimization, and large-scale software architecture. The architect will work closely with silicon vendors, operating system partners, AI framework teams, and product organizations to develop innovative technologies that maximize AI performance while maintaining system responsiveness, security, and power efficiency.

Responsibilities

~1 min read
  • →

    Define and drive the end-to-end architecture for HP's Memory Management platform.

  • →

    Architect systems for intelligent memory allocation, prioritization, and optimization across AI workloads running on client devices.

  • →

    Lead design of model residency management capabilities to enable efficient deployment and execution of large AI models on resource-constrained systems.

  • →

    Define policies for KV cache management, context window optimization, model sharing, and memory multiplexing across multiple AI agents.

  • →

    Develop policy engines that dynamically balance performance, memory consumption, power usage, and user responsiveness.

  • →

    Architect workload QoS and resource arbitration mechanisms across CPU, GPU, NPU, memory, storage, and network resources.

  • →

    Define strategies for local, hybrid, and cloud model execution based on system state, workload requirements, and resource availability.

  • →

    Lead development of telemetry and analytics systems that monitor memory pressure, model utilization, resource contention, and workload efficiency.

  • →

    Partner with operating system, firmware, silicon, and AI framework teams to optimize memory behavior across the full platform stack.

  • →

    Establish technical roadmaps, engineering standards, and architectural guidance across multiple software teams.

  • →

    Mentor senior engineers and architects while influencing company-wide AI platform strategy.

  • →

    Engage with ecosystem partners including Microsoft, Intel, AMD, Qualcomm, NVIDIA, and ISVs to align memory optimization technologies and industry standards.

  • Four-year or Graduate Degree in Computer Science, Computer Engineering, Electrical Engineering, Software Engineering, or related technical field, or equivalent experience.

  • 12+ years of industry experience in operating systems, systems software, platform architecture, AI infrastructure, memory management, or performance engineering.

  • Proven track record delivering scalable platform technologies, runtime systems, or enterprise software architectures.

  • Experience building software that operates close to hardware and operating system layers.

  • Deep expertise in operating system memory management concepts including virtual memory, paging, allocation strategies, memory pressure management, caching, NUMA architectures, and resource isolation.

  • Advanced understanding of DRAM architectures, memory controllers, UMA designs, storage hierarchies, and memory subsystems.

  • Experience optimizing memory behavior for high-performance applications and distributed systems.

  • Expertise in AI model memory management including:

    • Model residency management

    • Memory budgeting

    • Model quantization

    • Dynamic model loading and unloading

    • Context window optimization

    • Hybrid local/cloud model execution

  • Experience reducing AI memory footprints while maintaining model accuracy and performance.

  • Strong knowledge of:

    • KV cache optimization

    • Context management

    • Token generation pipelines

    • Model serving architectures

    • Multi-model execution frameworks

    • Agentic AI runtimes

  • Experience supporting concurrent AI agents sharing compute and memory resources.

  • Expertise in workload scheduling, admission control, resource allocation, and QoS management.

  • Experience coordinating workloads across heterogeneous compute engines including CPU, GPU, NPU, and cloud services.

  • Understanding of power-aware and thermal-aware resource management techniques.

  • Expert-level understanding of performance profiling, bottleneck analysis, workload characterization, and observability systems.

  • Experience building telemetry platforms that generate insights from large-scale performance data.

  • Ability to model and predict system behavior under varying workload conditions.

  • Experience with AI runtimes such as ONNX Runtime, OpenVINO, DirectML, TensorRT, CUDA, ROCm, Qualcomm AI SDKs, or equivalent technologies.

  • Experience optimizing AI inferencing pipelines on client computing platforms.

  • Familiarity with model compression techniques including quantization, pruning, distillation, and low-rank adaptation methods.

  • Knowledge of agentic AI frameworks and orchestration systems.

  • Expert proficiency in C++ and Python.

  • Strong systems programming experience.

  • Experience developing operating system services, runtime systems, middleware, or platform software.

  • Familiarity with Windows internals and client platform architecture.

  • Experience building large-scale analytics, telemetry, and automation solutions.

  • Demonstrated technical leadership driving highly complex platform initiatives.

  • Ability to align diverse engineering organizations behind a common architectural vision.

  • Strong cross-functional collaboration skills spanning hardware, software, firmware, product, and partner organizations.

  • Exceptional communication skills with the ability to influence senior leadership and external partners.

  • Proven mentoring and talent development experience.

Requirements

~1 min read
  • Experience working on operating systems, memory managers, hypervisors, container runtimes, or distributed resource management systems.

  • Prior work on AI PCs, edge AI systems, client computing platforms, or performance-critical infrastructure.

  • Contributions to memory optimization, AI infrastructure, operating systems, or runtime technologies through patents, publications, or open-source projects.

This role will define how future HP AI PCs efficiently execute increasingly large and complex AI workloads. The architect will establish the foundational technologies that enable multi-agent systems, local AI inferencing, context persistence, and intelligent resource sharing while ensuring responsive user experiences. The work directly influences HP's AI PC differentiation strategy, ecosystem partnerships, and long-term software platform roadmap.

Provides company-wide architectural leadership for one of the most technically challenging areas of the Agentic AI Software Stack. Solves highly complex problems involving memory management, AI model execution, workload orchestration, heterogeneous computing, operating systems, and performance optimization across hardware and software boundaries. The role is expected to create industry-leading innovations in AI memory efficiency and resource governance that become core differentiators for HP AI PCs.

Salary: $154,400 - $227,750 

 

What We Offer

~1 min read

The salary range for this role is listed above. Final salary offered is based upon multiple factors including individual job-related qualifications, education, experience, knowledge and skills.

At HP, we offer a competitive and comprehensive benefits package, including:

✓Health insurance
✓Dental insurance
✓Vision insurance
✓Long term/short term disability insurance
✓Employee assistance program
✓Flexible spending account
✓Life insurance
✓Generous time off policies, including;  4-12 weeks fully paid parental leave based on tenure
✓11 paid holidays
✓Additional flexible paid vacation and sick leave (US benefits overview)

HP, Inc. provides equal employment opportunity to all employees and prospective employees, without regard to race, color, religion, sex, national origin, ancestry, citizenship, sexual orientation, age, disability, or status as a protected veteran, marital status, familial status, physical or mental disability, medical condition, pregnancy, genetic predisposition or carrier status, uniformed service status, political affiliation or any other characteristic protected by applicable national, federal, state, and local law(s).

Please be assured that you will not be subject to any adverse treatment if you choose to disclose the information requested. This information is provided voluntarily. The information obtained will be kept in strict confidence.

Location & Eligibility

Where is the job
Houston, United States
On-site at the office
Who can apply
US

Listing Details

Posted
October 8, 2026
First seen
October 8, 2026
Last seen
October 8, 2026

Posting Health

Days active
0
Repost count
0
Trust Level
60%
Scored at
October 8, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

Senior Software Engineer, AI Memory Systems & Resource Management (Architecture)USD 154400-227750