Senior AI Research Scientist (New Delhi)

Senior AI Research Scientist (New Delhi)

27 Aug
|
Mindfire
|
New Delhi

27 Aug

Mindfire

New Delhi

Frontier Model Architecture Research

- Design, implement, and evaluate new neural-network architectures for language, reasoning, multimodal generation, and agentic workloads.

- Research the strengths and limitations of Transformers, SSMs, diffusion models, recurrent architectures, mixture-of-experts models, and hybrid designs.

- Develop architectures that combine attention, state-space mechanisms, recurrence, memory, sparse routing, retrieval, and modular components where appropriate.

- Explore efficient long-context methods, external and recurrent memory, adaptive computation, speculative execution, and improved reasoning techniques.

- Investigate diffusion and flow-based approaches for text, image, audio, video, structured data, and multimodal generation.

- Translate promising research papers and mathematical concepts into reproducible prototypes and production-relevant experiments.

Synthetic Data and Model-Generated Training

- Design scalable pipelines for creating high-quality synthetic examples, reasoning traces, preferences, critiques, simulations, and task-specific training data.

- Develop teacher-student, self-training, rejection-sampling, curriculum-learning, process-supervision, and knowledge-distillation approaches.

- Evaluate alignment and post-training methods such as supervised fine-tuning, preference optimization, reinforcement learning, and AI-generated feedback.

- Build filtering, scoring, deduplication, provenance, contamination-detection, and quality-control systems for synthetic datasets.

- Study and reduce the risks of feedback loops, bias amplification, hallucinations, overfitting, reward hacking, and model collapse caused by poorly controlled synthetic data.

- Establish methods for combining synthetic, public, licensed, customer-authorized, and human-generated data while maintaining privacy and traceability.

Distributed Inference and Neural Systems

- Develop algorithms that partition, route, and execute neural-network workloads across heterogeneous devices and infrastructure.

- Research tensor, pipeline, expert, sequence, and context parallelism for inference in resource-constrained and geographically distributed environments.

- Design efficient methods for model sharding, layer placement, distributed KV-cache management,



cache-aware scheduling, and dynamic workload migration.

- Improve inference across mixed hardware, including data-center GPUs, consumer GPUs, CPUs, NPUs, Apple Silicon, integrated graphics, edge devices, and browser-based runtimes where appropriate.

- Develop routing and scheduling strategies that account for memory capacity, bandwidth, latency, thermal limits, energy use, device availability, privacy rules, and workload priority.

- Create resilient inference methods that tolerate node loss, unreliable connectivity, changing resource availability, and partial system failure.

- Explore decentralized or collaborative neural networks in which multiple devices jointly execute models without requiring all data or model components to reside in one location.

- Advance compression and acceleration methods, including quantization, sparsity, pruning, distillation, speculative decoding, optimized kernels, and hardware-aware model design.

Research Evaluation and Scientific Rigor

- Form clear hypotheses, define baselines, design ablation studies, and run statistically sound experiments.

- Create evaluation suites covering accuracy, reasoning, robustness, safety, privacy, latency, throughput, memory, energy use, resilience, and cost.

- Identify where conventional benchmarks fail to predict real-world performance and develop task-relevant evaluations.

- Analyze quality-versus-efficiency tradeoffs and produce evidence that guides architecture and product decisions.

- Maintain reproducible research code, experiment records, model cards, dataset documentation, and technical reports.

- Monitor relevant research and clearly communicate which developments are promising, immature, or unsuitable for systems.

Technical Leadership and Product Collaboration

- Help define the companys research roadmap, technical strategy, and standards for scientific quality.





- Collaborate with engineering teams to move successful research from prototype to reliable production systems.

- Work with product leaders to connect research goals to customer needs, deployment constraints, and measurable business or social outcomes.

- Mentor researchers and engineers, review experimental designs, and raise the technical quality of the broader team.

- Contribute to patents, peer-reviewed publications, open research, technical demonstrations, grant proposals, and strategic partnerships when appropriate.

- Explain complex research clearly to technical teams, customers, partners, investors, and nontechnical stakeholders.

- Experience with distributed training or inference systems and parallel-computing frameworks.

- Hands-on work with model parallelism, expert parallelism, distributed KV caches, inference schedulers, or heterogeneous clusters.

- Experience with architectures such as Transformers, selective SSMs, Mamba-style systems, mixture-of-experts models, diffusion models, or hybrid combinations.

- Knowledge of inference engines and optimization stacks such as vLLM, SGLang, TensorRT-LLM, DeepSpeed, Ray, Triton, CUDA, ROCm, MLX, ONNX Runtime, WebGPU, or similar technologies.

- Experience with quantization, low-rank adaptation, agile adapters, pruning, sparsity, speculative decoding, or custom kernels.

- Experience creating or governing synthetic datasets at scale, including quality scoring, safety filtering, provenance, and contamination controls.

- Familiarity with alignment and post-training techniques, including preference optimization, reinforcement learning, AI feedback, red teaming, and safety evaluation.

- Experience with multimodal models spanning language, images, audio, video, documents, or structured data.

- Knowledge of privacy-preserving, federated, decentralized, on-device, or edge machine learning.

- A record of leading ambiguous research projects and helping others turn exploratory work into reliable systems.

Disclaimer: This has been sourced from a public domain and may have been modified by Naukri.com to improve clarity for our users. We encourage job seekers to verify all details directly with the employer via their official channels before applying.

📌 Senior AI Research Scientist (New Delhi)
🏢 Mindfire
📍 New Delhi

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior ai research scientist (new delhi) / new delhi

Subscribe to this job alert:

Get the latest job offers by email for: senior ai research scientist (new delhi) / new delhi