Develop, train, and optimize machine learning and deep learning models. * Build scalable AI/ML ... Exposure to distributed training, model compression, quantization, and inference optimization ...
New
Develop, train, and optimize machine learning and deep learning models. * Build scalable AI/ML ... Exposure to distributed training, model compression, quantization, and inference optimization ...
New
Develop, train, and optimize machine learning and deep learning models. * Build scalable AI/ML ... Exposure to distributed training, model compression, quantization, and inference optimization ...
New
Develop, train, and optimize machine learning and deep learning models. * Build scalable AI/ML ... Exposure to distributed training, model compression, quantization, and inference optimization ...
New
Develop, train, and optimize machine learning and deep learning models. * Build scalable AI/ML ... Exposure to distributed training, model compression, quantization, and inference optimization ...
New
New York, NY · On-site
$122K - $143K/yr
The position We are looking for our lead deep learning engineer to spearhead the development of our ... Optimize models for embedded deployment using quantization, pruning, TensorRT, and NVIDIA Triton
New York, NY · On-site
$122K - $143K/yr
The position We are looking for our lead deep learning engineer to spearhead the development of our ... Optimize models for embedded deployment using quantization, pruning, TensorRT, and NVIDIA Triton
$117K - $138K/yr
The position We are looking for our lead deep learning engineer to spearhead the development of our ... Optimize models for embedded deployment using quantization, pruning, TensorRT, and NVIDIA Triton
$117K - $138K/yr
The position We are looking for our lead deep learning engineer to spearhead the development of our ... Optimize models for embedded deployment using quantization, pruning, TensorRT, and NVIDIA Triton
New York, NY · On-site
$122K - $143K/yr
The position We are looking for our lead deep learning engineer to spearhead the development of our ... Optimize models for embedded deployment using quantization, pruning, TensorRT, and NVIDIA Triton
Quick apply
New York, NY · On-site
$122K - $143K/yr
The position We are looking for our lead deep learning engineer to spearhead the development of our ... Optimize models for embedded deployment using quantization, pruning, TensorRT, and NVIDIA Triton
Manhattan, NY · On-site
$180 - $260/hr
Implement quantization techniques and deploy large language models (LLMs) to maximize efficiency ... Deep knowledge and passion for data science fundamentals, training and deploying models
New
Manhattan, NY · On-site
$180 - $260/hr
Implement quantization techniques and deploy large language models (LLMs) to maximize efficiency ... Deep knowledge and passion for data science fundamentals, training and deploying models
New
Manhattan, NY · On-site
$112K - $148K/yr
Implement quantization techniques and deploy large language models (LLMs) to maximize efficiency ... Deep knowledge and passion for data science fundamentals, training and deploying models
Manhattan, NY · On-site
$112K - $148K/yr
Implement quantization techniques and deploy large language models (LLMs) to maximize efficiency ... Deep knowledge and passion for data science fundamentals, training and deploying models
Manhattan, NY · On-site
$164K - $260K/yr
Implement quantization techniques and deploy large language models (LLMs) to maximize efficiency ... Deep knowledge and passion for data science fundamentals, training and deploying models
Manhattan, NY · On-site
$164K - $260K/yr
Implement quantization techniques and deploy large language models (LLMs) to maximize efficiency ... Deep knowledge and passion for data science fundamentals, training and deploying models
Manhattan, NY · On-site
$112K - $148K/yr
Implement quantization techniques and deploy large language models (LLMs) to maximize efficiency ... Deep knowledge and passion for data science fundamentals, training and deploying models
Manhattan, NY · On-site
$112K - $148K/yr
Implement quantization techniques and deploy large language models (LLMs) to maximize efficiency ... Deep knowledge and passion for data science fundamentals, training and deploying models
New York, NY · On-site
$118K - $161K/yr
Strong foundation in machine learning, deep learning, and optimization. * Excellent software ... Model compression techniques including quantization (FP8, NVFP4, INT4), pruning, knowledge ...
New
New York, NY · On-site
$118K - $161K/yr
Strong foundation in machine learning, deep learning, and optimization. * Excellent software ... Model compression techniques including quantization (FP8, NVFP4, INT4), pruning, knowledge ...
New
Jersey City, NJ · On-site +1
$76K - $102K/yr
Apply techniques like quantization, distillation, and pruning to optimize LLM models for efficient ... Deep understanding of LLM architectures (e.g., Transformers), training techniques, and inference ...
Jersey City, NJ · On-site +1
$76K - $102K/yr
Apply techniques like quantization, distillation, and pruning to optimize LLM models for efficient ... Deep understanding of LLM architectures (e.g., Transformers), training techniques, and inference ...
New York, NY · On-site
$175K - $275K/yr
Implement techniques such as distillation, quantization, and pruning to aggressively accelerate ... Strong experience in deep learning systems and infrastructure * Expertise in PyTorch, CUDA, Triton ...
New York, NY · On-site
$175K - $275K/yr
Implement techniques such as distillation, quantization, and pruning to aggressively accelerate ... Strong experience in deep learning systems and infrastructure * Expertise in PyTorch, CUDA, Triton ...
Manhattan, NY · Hybrid
$115K - $157K/yr
Tune vector quantization strategies (PQ, SQ, Binary Quantization) to reduce memory footprint and ... Deep knowledge of vector embedding generation, storage and retrieval, with preference for hands-on ...
Manhattan, NY · Hybrid
$115K - $157K/yr
Tune vector quantization strategies (PQ, SQ, Binary Quantization) to reduce memory footprint and ... Deep knowledge of vector embedding generation, storage and retrieval, with preference for hands-on ...
Manhattan, NY · On-site
$115K - $157K/yr
Tune vector quantization strategies (PQ, SQ, Binary Quantization) to reduce memory footprint and ... Deep knowledge of vector embedding generation, storage and retrieval, with preference for hands-on ...
Manhattan, NY · On-site
$115K - $157K/yr
Tune vector quantization strategies (PQ, SQ, Binary Quantization) to reduce memory footprint and ... Deep knowledge of vector embedding generation, storage and retrieval, with preference for hands-on ...
$175K - $250K/yr
Prior experience in the domains of LLMs, foundation models, or large-scale deep learning systems, with a complete understanding of modern training, fine-tuning, quantization, and model evaluation.
$175K - $250K/yr
Prior experience in the domains of LLMs, foundation models, or large-scale deep learning systems, with a complete understanding of modern training, fine-tuning, quantization, and model evaluation.
New York, NY · On-site
$175K - $250K/yr
Prior experience in the domains of LLMs, foundation models, or large-scale deep learning systems, with a complete understanding of modern training, fine-tuning, quantization, and model evaluation.
New York, NY · On-site
$175K - $250K/yr
Prior experience in the domains of LLMs, foundation models, or large-scale deep learning systems, with a complete understanding of modern training, fine-tuning, quantization, and model evaluation.
New York, NY · On-site
$175K - $275K/yr
Implement techniques such as distillation, quantization, and pruning to aggressively accelerate ... Strong experience in deep learning systems and infrastructure * Expertise in PyTorch, CUDA, Triton ...
Quick apply
New York, NY · On-site
$175K - $275K/yr
Implement techniques such as distillation, quantization, and pruning to aggressively accelerate ... Strong experience in deep learning systems and infrastructure * Expertise in PyTorch, CUDA, Triton ...
... quantization, compression, and resource-efficient AI, to drive performance improvements and ... Research experience in machine learning, deep learning, natural language processing, and/or ...
... quantization, compression, and resource-efficient AI, to drive performance improvements and ... Research experience in machine learning, deep learning, natural language processing, and/or ...
New York, NY · On-site
$180K - $360K/yr
Deep dive into underlying codebases of TensorRT, PyTorch, TensorRT-LLM, vllm, sglang, CUDA, and ... Familiarity with LLM optimization techniques (e.g., quantization, speculative decoding, continuous ...
New York, NY · On-site
$180K - $360K/yr
Deep dive into underlying codebases of TensorRT, PyTorch, TensorRT-LLM, vllm, sglang, CUDA, and ... Familiarity with LLM optimization techniques (e.g., quantization, speculative decoding, continuous ...
$196K - $253K/yr
Strong foundation in machine learning algorithms, including deep learning architectures (e.g ... quantization, pruning, and knowledge distillation. * Experience with model interpretability ...
$196K - $253K/yr
Strong foundation in machine learning algorithms, including deep learning architectures (e.g ... quantization, pruning, and knowledge distillation. * Experience with model interpretability ...
$22.3K is the 25th percentile. Wages below this are outliers.
$11.2K - $23.2K
27% of jobs
$23.2K - $35.1K
0% of jobs
$35.1K - $47.1K
0% of jobs
$47.1K - $59.1K
0% of jobs
$59.1K - $71K
0% of jobs
The median wage is $82K / yr.
$71K - $83K
25% of jobs
$83K - $95K
18% of jobs
$103.5K is the 75th percentile. Wages above this are outliers.
$95K - $106.9K
7% of jobs
$106.9K - $118.9K
2% of jobs
$118.9K - $130.8K
0% of jobs
$130.8K - $142.8K
21% of jobs
$11.2K
$85.6K
$142.8K
| Aspect | Deep Learning Quantization | Machine Learning Engineer |
|---|---|---|
| Required Credentials | Advanced degrees in AI, Computer Science, or related fields; knowledge of neural networks | Bachelor's or Master's in CS, Data Science, or related fields; programming skills |
| Work Environment | Research labs, AI development teams, hardware optimization settings | Software development teams, data-driven projects, product-focused environments |
| Industry Usage | AI hardware optimization, model deployment, edge computing | Model development, data analysis, software solutions across industries |
Deep Learning Quantization focuses on reducing model size and improving inference speed through techniques like weight and activation quantization, often in hardware or embedded systems. Machine Learning Engineers develop, implement, and optimize machine learning models for various applications. While both roles require knowledge of AI and programming, Deep Learning Quantization is more specialized in model optimization techniques, whereas Machine Learning Engineers work broadly on model development and deployment.
Other
Medical, Dental, Vision, Life, Retirement, PTO
Posted 3 days ago
New
7.9
Based on 341 frontline employees who took The Breakroom Quiz
7th of 26 rated airlines
Choosing Capgemini means choosing a company where you will be empowered to shape your career in the way you'd like, where you'll be supported and inspired bya collaborative community of colleagues around the world, and where you'll be able to reimagine what's possible. Join us and help the world's leading organizationsunlock the value of technology and build a more sustainable, more inclusive world.
We are seeking a highly skilled GenAI & AI/ML Framework Specialist to design, develop, and scale next-generation Artificial Intelligence solutions. This role will focus on building advanced machine learning and Generative AI applications, optimizing open-source AI frameworks, and integrating AI-powered capabilities into enterprise platforms.
The ideal candidate will possess deep expertise in AI/ML frameworks, Large Language Models (LLMs), MLOps, cloud platforms, and modern data infrastructure. You will work closely with product, engineering, and data science teams to deliver high-performance, production-ready AI solutions that drive business value.
The base compensation range for this role in the posted location is: 85786- 105237
Capgemini provides compensation range information in accordance with applicable national, state, provincial, and local pay transparency laws. The base compensation range listed for this position reflects the minimum and maximum target compensation Capgemini, in good faith, believes it may pay for the role at the time of this posting. This range may be subject to change as permitted by law.
The actual compensation offered to any candidate may fall outside of the posted range and will be determined based on multiple factors legally permitted in the applicable jurisdiction.
These may include, but are not limited to: Geographic location, Education and qualifications, Certifications and licenses, Relevant experience and skills, Seniority and performance, Market and business consideration, Internal pay equity.
It is not typical for candidates to be hired at or near the top of the posted compensation range.
In addition to base salary, this role may be eligible for additional compensation such as variable incentives, bonuses, or commissions, depending on the position and applicable laws.
Capgemini offers a comprehensive, non-negotiable benefits package to all regular, full-time employees. In the U.S. and Canada, available benefits are determined by local policy and eligibility and may include:
Important Notice: Compensation (including bonuses, commissions, or other forms of incentive pay) is not considered earned, vested, or payable until it becomes due under the terms of applicable plans or agreements and is subject to Capgemini's discretion, consistent with applicable laws. The Company reserves the right to amend or withdraw compensation programs at any time, within the limits of applicable legislation.
Disclaimers
Capgemini is an Equal Opportunity Employer encouraging inclusion in the workplace. Capgemini also participates in the Partnership Accreditation in Indigenous Relations (PAIR) program which supports meaningful engagement with Indigenous communities across Canada by promoting fairness, accessibility, inclusion and respect. We value the rich cultural heritage and contributions of Indigenous Peoples and actively work to create a welcoming and respectful environment. All qualified applicants will receive consideration for employment without regard to race, national origin, gender identity/expression, age, religion, disability, sexual orientation, genetics, veteran status, marital status or any other characteristic protected by law.
This is a general description of the Duties, Responsibilities and Qualifications required for this position. Physical, mental, sensory or environmental demands may be referenced in an attempt to communicate the manner in which this position traditionally is performed. Whenever necessary to provide individuals with disabilities an equal employment opportunity, Capgemini will consider reasonable accommodations that might involve varying job requirements and/or changing the way this job is performed, provided that such accommodation does not pose an undue hardship. Capgemini is committed to providing reasonable accommodation during our recruitment process. If you need assistance or accommodation, please reach out to your recruiting contact.
Please be aware that Capgemini may capture your image (video or screenshot) during the interview process and that image may be used for verification, including during the hiring and onboarding process.
Click the following link for more information on your rights as an Applicant in the United States. http://www.capgemini.com/resources/equal-employment-opportunity-is-the-law
Capgemini is a global business and technology transformation partner, helping organizations to accelerate their dual transition to a digital and sustainable world, while creating tangible impact for enterprises and society. It is a responsible and diverse group of 340,000 team members in more than 50 countries. With its strong over 55-year heritage, Capgemini is trusted by its clients to unlock the value of technology to address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, generative AI, cloud and data, combined with its deep industry expertise and partner ecosystem.
Get the full story on Breakroom
Sourced by ZipRecruiter
United Airlines is embarking on an exciting journey to become the best airline in aviation history. Our purpose, "Connecting People, Uniting the World," extends beyond transportation, emphasizing our commitment to uplift and create opportunities in the places we serve. With a global presence and diverse workforce, we value inclusivity and are dedicated to hiring tens of thousands of individuals across various roles. Our comprehensive benefits package, including perks like space available travel, parental leave, and 401k, aims to support your well-being and growth.
Aviation
10,000+ Employees
Chicago, IL, US
1926