Engineering Group, Engineering Group > Software Engineering
General Summary
Job Overview:
The Qualcomm Cloud AI team is developing hardware and software for Machine Learning solutions spanning the data center, edge, infrastructure, automotive market. We are seeking ambitious, bright, and innovative engineers with experience in machine learning framework development. Job activities span the whole product life cycle from early design to commercial deployment. The environment is fast-paced and requires cross-functional interaction daily so positive communication, planning and execution skills are a must.
We are seeking a highly skilled and motivated Language Model Engineer to join our team. The primary role of the engineer will be to train Large Language Models (LLMs) from scratch and fine-tune existing LLMs on various datasets using state-of-the-art techniques.
Responsibilities
Model architecture optimizations : optimize latest LLM and GenAI model architectures for NPUs,
which involves reimplementing basic building blocks of models for NPUs
Model Training and Fine-tuning: Fine-tune pre-trained models on specific tasks or datasets to improve performance. Implement state-of-the-art LLM training techniques such as Reinforcement Learning from Human Feedback (RLHF), ZeRO (Zero Redundancy Optimizer), Speculative Sampling, and other speculative techniques.
Data Management: Handle large datasets effectively. Ensure data quality and integrity. Implement data cleaning and preprocessing techniques. Hands-on with EDA is a plus.
Model Evaluation: Evaluate model performance using appropriate metrics. Understand the trade-offs between different evaluation metrics.
LLM metrics: Sound understanding of various LLM metrics like MMLU, Rouge, BLEU, Perplexity etc.
AWQ: Understanding of Quantization is a plus. Knowledge on QAT will be a plus.
Research and Development: Stay up to date with the latest research in NLP and LLMs. Implem