Hiring.Camp

AI Model Optimization Architect - Cork, Ireland

Qualcomm

·

Yesterday

Location
Cork, CO,IE, IE
Type
Full-time
Department
Engineering
Education
PhD
Source
Eightfold

Description

##

Company:

QT Technologies Ireland Limited

## Job Area:

Engineering Group, Engineering Group > Software Engineering

General Summary:

General Summary:

Qualcomm is leveraging its strengths in compute, connectivity, and AI acceleration to play a central role in the evolution of Cloud AI. The Qualcomm Cloud AI team develops hardware and software platforms enabling efficient inference of large-scale foundation models.

We are seeking a Staff Engineer – AI Model Optimization Architect to lead end-to-end model transformation and optimization for LLMs, VLMs, diffusion, and multimodal models on Qualcomm inference accelerators. This role works closely with compiler, performance, and accuracy teams to translate models into accelerator efficient execution while balancing throughput, latency, memory, and quality. The scope spans Day0 enablement through production deployment, with a strong emphasis on scaling optimizations to future architectures.

Key Responsibilities

· Architect and deliver model optimization strategies that transform PyTorch models for efficient inference on Qualcomm accelerators.

· Drive graph capture and deployment using PyTorch, ONNX, and torch.compile, including model rewrites and graph-level transformations.

· Design and implement fusion kernels using DSL based approaches (e.g., Triton), enabling fused operations and performance critical algorithmic rewrites.

· Partner deeply with compiler, performance, and accuracy teams to co-design lowering strategies, kernel fusion, layout decisions, and runtime integration.

· Profile and optimize LLM/VLM/diffusion inference for throughput and latency across batch sizes, sequence lengths, and serving modes.

· Own transformer specific optimizations including KVcache management, decoding behavior, and long context performance.

· Enable and optimize continuous batching (dynamic/iteration-level scheduling), understanding its impact on memory, scheduling, and tail latency.

· Architect and scale distributed inference strategies (e.g., sharding and parallelism) across multi-core and multi-device systems.

· Establish reusable approaches to scale model optimizations to new hardware architectures, creating robust patterns and tooling.

· Debug complex performance or stability issues to root cause and drive production ready solutions.

Required Qualifications

· Expert level expertise in PyTorch and inference focused model optimization; strong Python engineering skills.

· Hands on experience with torch.compile / TorchDynamo or related graph capture and compilation workflows.

· Deep understanding of transformer architectures, attention mechanisms, MoEs, and performance trade-offs.

· Practical experience with KVcache behavior, serving time optimizations, and memory/performance tradeoffs.

· Strong foundation in computer architecture, ML accelerators, and distributed systems.

· Proven ability to lead cross-functional technical efforts and influence design decisions.

· MS in Computer Science, Machine Learning, Computer Engineering, or Electrical Engineering, or equivalent experience.

Preferred / Bonus Qualifications

· Experience developing fusion kernels using Triton or similar DSLs, and collaborating with ML compiler teams.

· Familiarity with LLM serving stacks and continuous batching systems.

· Background in numerical methods, performance/accuracy trade-off analysis, or evaluation frameworks.

Where you will be working

Cork has a proud reputation as Ireland's second largest economic engine and is now one of the Top 20 location choices in Europe with 39,000 people being employed by over 170 overseas companies.

There's a growing diversity in the region with people from many nationalities relocating to Cork, relishing the opportunity to work and live in a location that offers an excellent quality of life.

A gateway to Europe, Cork airport provides access to almost 50 international destinations including transatlantic air routes.

Equal Opportunities

We are an Equal Opportunity employer; all qualified applicants will receive consideration for employment without regard to race, colour, religion, sexual orientation, gender identity, national origin, disability, veteran status, or any protected classification.

What's on Offer

Apart from working in an open, relaxed and collaborative space, you will enjoy:

  • Salary, stock and performance related bonus
  • Maternity/Paternity Leave
  • Employee stock purchase scheme
  • Matching pension scheme
  • Education Assistance
  • Relocation and immigration support (if needed)
  • Life, Medical, Income and Travel Insurance
  • Subsidised memberships for physical and mental well-being
  • Bicycle purchase scheme
  • Employee run clubs, including, running, football, chess, badminton + many more

Minimum Qualifications:

  • Bachelor's degree in Engineering, Information Systems, Computer Science, or related field and 6+ years of Software Engineering or related work experience.

OR

Master's degree in Engineering, Information Systems, Computer Science, or related field and 5+ years of Software Engineering or related work experience.

OR

PhD in Engineering, Information Systems, Computer Science, or related field and 4+ years of Software Engineering or related work experience.

  • 3+ years of work experience with Programming Language such as C, C++, Java, Python, etc.

*References to a particular number of years experience are for indicative purposes only. Applications from candidates with equivalent experience will be considered, provided that the candidate can demonstrate an ability to fulfill the principal duties of the role and possesses the required competencies.

Qualcomm is an equal opportunity employer. If you are an individual with a disability and need an accommodation during the application/hiring process, rest assured that Qualcomm is committed to providing an accessible process. You may e-mail [email protected] or call Qualcomm's toll-free number found here. Upon request, Qualcomm will provide reasonable accommodations to support individuals with disabilities to be able participate in the hiring process. Qualcomm is also committed to making our workplace accessible for individuals with disabilities. (Keep in mind that this email address is used to provide reasonable accommodations for individuals with disabilities. We will not respond here to requests for updates on applications or resume inquiries).

Qualcomm expects its employees to abide by all applicable policies and procedures, including but not limited to security and other requirements regarding protection of Company confidential information and other confidential and/or proprietary information, to the extent those requirements are permissible under applicable law.

To all Staffing and Recruiting Agencies: Our Careers Site is only for individuals seeking a job at Qualcomm. Staffing and recruiting agencies and individuals being represented by an agency are not authorized to use this site or to submit profiles, applications or resumes, and any such submissions will be considered unsolicited. Qualcomm does not accept unsolicited resumes or applications from agencies. Please do not forward resumes to our jobs alias, Qualcomm employees or any other company location. Qualcomm is not responsible for any fees related to unsolicited resumes/applications.

If you would like more information about this role, please contact Qualcomm Careers.

Skills

PythonJavaMachine LearningPyTorch

Similar Jobs

8

Lead AI Data Trainer & AI Model Optimization Engineer

Purestorage · Prague, Czech Republic · Hybrid, Onsite

1 month ago

AI Data Trainer & AI Model Optimization Engineer

Purestorage · Prague, Czech Republic · Hybrid, Onsite

1 month ago

Sr Software Engineer, AI Tools – On-Device Generative AI Model Optimization

Qualcomm · San Diego, CA,US, US · Onsite

2 months ago

Senior Quantization Engineer - Edge AI Model Optimization

Job Listings · Hyderabad, India

2 months ago

Edge AI/Model Optimization Engineer

NextGen Federal Systems · Aberdeen, Maryland · Hybrid

3 months ago

AI Inference Engineer - Model Optimization & Deployment

Zoox · Foster City, CA +2 · Hybrid

4 months ago

AI Frameworks Engineer - Model and Kernel Optimization

Intel · CHN - Minhang, China · Onsite

1 week ago

AI Frameworks Engineer - Model and Kernel Optimization

Intel · CHN - Minhang, China · Onsite

1 week ago