SSE - Optimization Engineer
MulticoreWare Pvt Ltd · IT Services & Consulting
- Ramapuram, India
- On-site
- Posted about a month ago
- Software Engineering
- Full time
About the job
We are looking
for a Senior Software Engineer to develop and optimize deep learning models,
including CNNs, LLMs, and MoE, for efficient inference across CPU, GPU,
hardware accelerators, and edge devices. The role focuses on quantization,
model compression, high-performance kernel implementation, transformer
optimization, and production deployment.
Responsibilities:
LLMs, MoE) for efficient inference across CPU, GPU, hardware accelerators, and
edge devices.
(PTQ, QAT, GPTQ, AWQ) from scratch.
pruning, decomposition, and distillation.
INT4, FP8) using C++ for high performance.
implementations.
for real-world deployment.
KV-cache, PEFT (LoRA/QLoRA), and MoE models.
across different hardware platforms.
teams to deliver optimized solutions.
Requirements
Education:
BE/BTech/MS/MTech in Computer Science or a related field.
Technical Skills (Must haves):
(PTQ, QAT, GPTQ, AWQ).
compression, and inference optimization.
optimization techniques from scratch.
LLM architectures.
deployment pipelines.
optimization skills.
Need to have (Can be bridged):
No
additional bridged skills were specified.
Good to have (Not essential):
techniques (LoRA, QLoRA).
MLIR.
across GPU, NPU, and edge devices.
open-source contributions.
Preferred Qualifications (Optional):
No
additional preferred qualifications were specified.