Quantization jobs
Explore 11 current openings. Compare roles, companies, locations, and work options, then open any job to view the full details.
Current openings
Showing 20 jobsSenior Quantization Engineer
NXP SemiconductorsEdge AI Model Optimization team at NXP focuses on advancing on‑device intelligence by optimizing CNN, LLM and VLM models for the Ara2 NPU family. The Senior Quantization Engineer researches state‑of‑the‑art quantization …
Hyderabad,IN,IndiaOn-siteFull-timeActiveSenior Quantization Engineer
NXP SemiconductorsQuantization Engineering Team at NXP focuses on optimizing AI models for edge devices. The Senior Quantization Engineer designs and implements quantization strategies, validates code against NXP standards, and …
Hyderabad,IndiaOn-siteFull-timeActiveSenior Quantization Engineer
NXP SemiconductorsEdge AI Model Optimization team at NXP focuses on quantizing CNNs, LLMs, and VLMs for the Ara2 NPU family. The role involves researching state‑of‑the‑art quantization, prototyping on NXP hardware, and delivering …
Hyderabad,IN,IndiaFull-timeActiveData Scientist
Commonwealth Bank of AustraliaData Scientist at Commonwealth Bank of Australia leading Gen AI and multi-agent solutions for finance. Architect and build large language models, multi-modal generative models, and agentic frameworks to drive innovation …
Bangalore, IndiaOn-siteFull-timeActiveAI Research Engineer
Tether Operations LimitedAI research team at Tether focuses on model compression for multimodal AI. The role drives low‑bit quantization, knowledge distillation, and pruning to enable LLMs and VLMs on edge devices. Work is fully remote with a …
Bangalore, Karnataka, IndiaRemoteFull-timeActiveComputer Vision Data Scientist
CodvoDesign, develop, and optimize computer vision models for Object Detection, Image & Instance Segmentation, Multi-Object Tracking, Human Pose Estimation. Build and train deep learning models using frameworks such as …
Pune, Maharashtra, IndiaOn-siteFull-timeActiveSr Software Engineer III - AI
Elevance HealthSr Software Engineer III – AI at Carelon Global Solutions India. Lead AI/ML engineering for healthcare solutions. Design scalable data pipelines, build LLM inference architectures, optimize models with quantization and …
Bengaluru, KA, IndiaOn-siteFull-timeActiveSenior Software Engineer
Elevance HealthSenior Software Engineer at Carelon Global Solutions India, part of Elevance Health, leading AI and digital platform initiatives. Focus on machine learning, NLP, and large language models to transform healthcare data. …
Bengaluru, KA, IndiaOn-siteFull-timeActiveSenior Software Engineer Modelzoo
NXP SemiconductorsModelzoo team at NXP Semiconductors builds AI inference solutions for specialized hardware. The Senior Software Engineer implements model optimizations, focusing on quantization and performance tuning, and ports deep …
Hyderabad,IndiaFull-timeActiveSoftware Engineer
NXP SemiconductorsModelZoo team at NXP Semiconductors builds and optimizes deep learning models for specialized AI accelerator hardware. The role implements model porting, performance optimization and quantization to improve inference …
Hyderabad,IN,IndiaFull-timeActiveMachine Learning Engineer, Inference Optimization
jobgetherMachine Learning Engineering team focuses on building high‑performance AI systems for production. The role optimizes inference pipelines, implements quantization, KV‑cache and batching techniques, and improves …
IndiaRemoteFull-timeActiveGen AI Intern
AirbusGen AI Intern at Airbus India Private Limited focuses on advancing language model capabilities through Retrieval-Augmented Generation (RAG) and related techniques. Responsibilities include researching and implementing …
Bangalore, IndiaOn-siteInternshipActiveInference Optimization Architect, Speech AI
NvidiaNVIDIA seeks an Inference Optimization Architect for its Speech AI team to accelerate and scale conversational AI models. The role focuses on reducing inference latency, improving throughput, and optimizing resource …
Pune, Maharashtra, IndiaOn-siteFull-timeActiveSenior Data Scientist
S&P GlobalSenior Data Scientist at S&P Global Capital IQ Solutions Data Science team. Responsible for designing, developing, evaluating, and deploying advanced NLP, generative AI, and large language models to power the Capital IQ …
Gurugram, IndiaOn-siteFull-timeActiveStaff Machine Learning Engineer
Automation AnywhereAutomation Anywhere seeks a Staff Machine Learning Engineer to develop and optimize NLP, Computer Vision, and GenAI models, architect scalable ML pipelines, drive large-scale ML infrastructure, implement MLOps best …
Bengaluru, IndiaOn-siteFull-timeActiveAI Engineer
CaterpillarJoin Caterpillar’s Cat Digital team as an AI Engineer to build the next‑generation Manufacturing & Supply Digital Platform. Leverage NVIDIA Omniverse, OpenUSD, and generative AI models (transformers, diffusion, GPT, …
Chennai, Tamil NaduOn-siteFull-timeActiveInference Optimization Architect, Speech AI
NvidiaJoin NVIDIA’s Speech AI Engineering team as an Inference Optimization Architect. Accelerate and scale speech AI models by reducing inference latency, boosting throughput, and optimizing resource use across GPU …
Pune, IndiaOn-siteFull-timeActiveMachine Learning Engineer
AdobeAdobe’s Digital Video & Audio group builds creative tools like Premiere, After Effects, Audition. Seeking a Machine Learning Engineer to transform research into production‑ready AI features for millions of creators. …
Bangalore, IndiaOn-siteFull-timeActiveSenior Associate AI/ML Engineer
PricewaterhouseCoopersPwC data & analytics engineering team builds robust data solutions and AI/ML systems to transform raw data into actionable insights. Senior Associate AI/ML Engineer will design, optimize, and deploy advanced AI models, …
Bengaluru, Karnataka, IndiaHybridFull-timeActiveMachine Learning Engineer II
DriveCamMachine Learning Engineer II at Lytx India will join the Applied Machine Learning Team in Bangalore to develop and deploy deep learning and computer vision models that monitor driver behavior and environments, enhancing …
Bangalore, IndiaOn-siteFull-timeActive