Job Summary:
Nace AI is a company focused on machine learning solutions, and they are seeking a Machine Learning Engineer to translate cutting-edge research into scalable, production-ready solutions. The role involves designing and maintaining ML systems, collaborating with cross-functional teams, and enhancing existing models with the latest advancements in machine learning.
Responsibilities:
โข Design, build, and maintain end-to-end ML systems, including synthetic data pipelines, model training, debugging, and performance evaluation.
โข Fine-tune large language models (LLMs) and implement meta-learning methods to enhance model generalization and efficiency.
โข Improve existing Nace.AI models by incorporating advancements from recent ML research.
Qualifications:
Required:
โข Hands-on experience training and fine-tuning large language models (LLMs) and vision-language models (VLMs), including practical work with pre-training, instruction tuning, and alignment techniques (GRPO,RLHF/DPO/PPO).
โข Hands-on Experience with Deep Learning Models, especially Transformers.
โข Ability to translate cutting-edge research from papers into clean, production-ready code (Paper to Code).
โข Proven experience scaling inference infrastructure for LLMs/VLMs, including expertise in model serving frameworks like vLLM, TGI.
โข Proficient in Python with a strong track record of building substantial projects.
โข Solid foundation in computer science fundamentals (data structures, algorithms, design patterns).
โข BS degree in CS or related technical field.
โข Solid Experience with ML frameworks and libraries (PyTorch, TensorFlow).
โข Self-starter comfortable working in a fast-paced, dynamic environment.
Preferred:
โข MS/PhD in CS or related technical field.
โข Familiarity with data processing stacks such as Spark and Airflow.
โข Experience with multi-node GPU training.
โข Contributor to open-source ML projects.
โข Deep knowledge in Linear Programming.
โข Experience with advanced NLP and Multimodal post-training experience (e.g., model distillation, quantization, deployment optimization).
โข Experienced in inference time optimization, deep understanding of LLM serving optimizations for LLMs/VLMs.
โข Hands on experience with quantization techniques (AWQ, GPTQ, FP8/GGUF).
Company:
Enterprise AI product & research company, building long-horizon reasoning models and agents. Founded in 2024, the company is headquartered in Palo Alto, USA, with a team of 11-50 employees. The company is currently Early Stage.