or apply directly on OPSWAT's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 7 days OPSWAT's roles stay open a median of 45 days
- Reposted
- No
- Salary listed
- No 1% of OPSWAT's roles list one
- Ghost-job risk at OPSWAT
- high 56 stale, 5 reposted of 89 open
- Hiring momentum
- 118 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-09-17
Measured from postings appearing on and disappearing from OPSWAT's own greenhouse board since 2026-08-03. Full hiring picture for OPSWAT.
About this role
As a Senior AI Engineer at OPSWAT, you will lead the deployment and optimization of Large Language Models (LLMs) for high-performance inference on GPU hardware. Your responsibilities include designing local LLM serving environments, optimizing model performance through quantization and compression techniques, and managing production serving infrastructure. You will collaborate with various teams to ensure efficient and cost-effective AI service delivery.
- benefits
- 1/5
- freshness
- 5/5
- career value
- 5/5
- role clarity
- 5/5
- pay transparency
- 0/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of OPSWAT as an employer.
What you need
- Bachelor's degree in Computer Science, Software Engineering, or a related field
- 5+ years of experience in software engineering, focusing on ML infrastructure, LLM inference, or model optimization
- Hands-on experience deploying and serving LLMs on GPU hardware in production
- Strong understanding of quantization and model compression techniques
- Experience with high throughput inference engines
- Solid understanding of GPU architecture, CUDA, and memory bandwidth in LLM inference
Nice to have
- Experience writing or tuning custom CUDA / Triton kernels
- Experience with multi GPU and distributed inference
- Certifications in AWS or Azure architecture
- Experience with CI/CD and cloud production deployment
Worth weighing
- No salary listed
- Focus on GPU hardware and LLMs may limit exposure to other AI technologies
- Fast-paced environment may require adaptability to changing priorities
Summarised from OPSWAT's posting. Read the full original.
Listed by OPSWAT on their greenhouse job board, last confirmed open on 2026-09-17. PitchMeAI is not the employer.
More roles at OPSWAT
- Senior Software EngineerHo Chi Minh City, Ho Chi Minh City, Vietnam
- Joint Professional Services – Technical Support Engineer in Bangalore, IndiaBengaluru, Karnataka, India
- AI-Native Software Engineer - MD CoreHo Chi Minh City, Ho Chi Minh City, Vietnam
- AI-Native Software EngineerHo Chi Minh City, Ho Chi Minh City, Vietnam
- AI Security Automation EngineerHo Chi Minh City, Ho Chi Minh City, Vietnam
- AI Security EngineerHo Chi Minh City, Ho Chi Minh City, Vietnam
- Machine Learning EngineerHo Chi Minh City, Ho Chi Minh City, Vietnam
- Director of Software Engineering - My OPSWAT PortalHo Chi Minh City, Ho Chi Minh City, Vietnam