Python Developer – LLM Post-Training (Remote | San Francisco) (India)

Python Developer – LLM Post-Training (Remote | San Francisco) (India)

07 Sep
|
Parsewave
|
India

07 Sep

Parsewave

India

We are a San Francisco-based AI infrastructure company working with leading frontier AI labs to build post-training data and evaluation infrastructure for foundation models. We are hiring a Python Developer to create high-quality datasets, reinforcement learning environments, and benchmarking pipelines used to improve and evaluate state-of-the-art LLMs. This is a remote role with flexible working hours.

Responsibilities

* Create and curate datasets for LLM post-training (SFT, RLHF, RL, preference optimization).

* Build and maintain RL environments for agent evaluation.

* Develop Python tooling for dataset generation, validation, and transformation.

* Evaluate models on custom benchmarks and testing pipelines.

* Collaborate with research and engineering teams to deliver client-specific post-training datasets.

* Work with terminal-first development workflows and cloud infrastructure.

Required Skills

* Strong Python programming skills.

* Understanding of LLM fundamentals and post-training concepts (SFT, RLHF, RL).

* Experience working with structured data (JSON, CSV, YAML).

* Git, Linux/Unix command line, and solid software engineering fundamentals.

Good to Have





Experience with RAG, agentic AI systems, Hugging Face Transformers, LoRA/PEFT, LangChain or LlamaIndex, vector databases (FAISS, Qdrant, Milvus, Pinecone, Weaviate, ChromaDB), Docker, AWS/GCP, FastAPI/Flask, Bash, CLI tooling, model evaluation frameworks, benchmarking, and AI infrastructure.

Compensation

Base Salary: USD $1,250/month

Equity: ESOP/Equity package included.

Performance Bonuses: Up to USD $4,000/month (in addition to base salary).

Location

Remote (Worldwide)

Work Hours

Adaptable, remote-first, asynchronous work environment.

How to Apply

Apply here: https://tally.so/r/wLReJG

Please complete the application form and submit the required details. Only shortlisted candidates will be contacted.

Skills:- Python, Large Language Models (LLM), Fine-tuning LLMs, PEFT (Parameter-Efficient Fine-Tuning), Amazon Web Services (AWS), Google Cloud Platform (GCP), Docker, Retrieval Augmented Generation (RAG), Vector database, FastAPI, Agentic AI, Hugging Face Transformers, LoRA / QLoRA, LangChain, LangGraph, LlamaIndex and Bash

📌 Python Developer – LLM Post-Training (Remote | San Francisco) (India)
🏢 Parsewave
📍 India

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: python developer – llm post-training (remote | san francisco) (india) / india

Subscribe to this job alert:

Get the latest job offers by email for: python developer – llm post-training (remote | san francisco) (india) / india