OPSWAT

Senior AI Engineer

OPSWAT · Ho Chi Minh City, Ho Chi Minh City, Vietnam

Posted 7 days ago

or apply directly on OPSWAT's site. We never take the application ourselves.

Is this posting real?

This role has been open
7 days
OPSWAT's roles stay open a median of 45 days
Reposted
No
Salary listed
No
1% of OPSWAT's roles list one
Ghost-job risk at OPSWAT
high
56 stale, 5 reposted of 89 open
Hiring momentum
118 roles opened in the last 90 days
↑ up vs. the prior 90 days
Last confirmed on the employer's board
2026-09-17

Measured from postings appearing on and disappearing from OPSWAT's own greenhouse board since 2026-08-03. Full hiring picture for OPSWAT.

About this role

As a Senior AI Engineer at OPSWAT, you will lead the deployment and optimization of Large Language Models (LLMs) for high-performance inference on GPU hardware. Your responsibilities include designing local LLM serving environments, optimizing model performance through quantization and compression techniques, and managing production serving infrastructure. You will collaborate with various teams to ensure efficient and cost-effective AI service delivery.

Our read on this posting3.2out of 5
benefits
1/5
freshness
5/5
career value
5/5
role clarity
5/5
pay transparency
0/5

Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of OPSWAT as an employer.

What you need

  • Bachelor's degree in Computer Science, Software Engineering, or a related field
  • 5+ years of experience in software engineering, focusing on ML infrastructure, LLM inference, or model optimization
  • Hands-on experience deploying and serving LLMs on GPU hardware in production
  • Strong understanding of quantization and model compression techniques
  • Experience with high throughput inference engines
  • Solid understanding of GPU architecture, CUDA, and memory bandwidth in LLM inference

Nice to have

  • Experience writing or tuning custom CUDA / Triton kernels
  • Experience with multi GPU and distributed inference
  • Certifications in AWS or Azure architecture
  • Experience with CI/CD and cloud production deployment

Worth weighing

  • No salary listed
  • Focus on GPU hardware and LLMs may limit exposure to other AI technologies
  • Fast-paced environment may require adaptability to changing priorities

Summarised from OPSWAT's posting. Read the full original.

Listed by OPSWAT on their greenhouse job board, last confirmed open on 2026-09-17. PitchMeAI is not the employer.

More roles at OPSWAT

All 89 open roles at OPSWAT