Prompt Engineer (Founding Team) (Hyderabad)

Prompt Engineer (Founding Team) (Hyderabad)

19 Sep
|
Gradientflo Labs
|
Hyderabad

19 Sep

Gradientflo Labs

Hyderabad

Role Overview

You will be the language architect of Vibecoderz. The TutorAgent and its family of sub-agents only think as clearly as the prompts that guide them — and you will design those “thought patterns.” As the founding Prompt Engineer, you’ll build a prompt library and evaluation system that ensures Vibecoderz produces consistent, safe, and high-quality artifacts: slides, quizzes, code snippets, and runnable mini-apps.

This is a role for a builder of mental scaffolding. You’ll be deeply involved in orchestrating system prompts, tool schemas, chaining strategies, and eval frameworks that make our AI both reliable and delightful. You’ll collaborate with AI Engineers on orchestration, Backend Engineers on schema integration, and PM on defining product outcomes.

You’ll use Linear for execution, Notion for prompt specs and experiments, and GitHub for version control, making every iteration observable, testable, and reproducible.

Key Responsibilities

1. Prompt Library Ownership - Design, version, and maintain system prompts for TutorAgent, PlannerAgent, CodeAgent, QuizAgent, and ContentAgent.

2. Prompt Chaining & Orchestration - Implement advanced strategies: reflection, planning, few-shot, self-consistency, and tool-use flows.

3. Schema Definition - Collaborate with BE and AI engineers to define JSON schemas and tool contracts that LLMs must respect.

4. Evaluation Pipelines - Use LangSmith or Langfuse to build eval frameworks that test accuracy, latency, drift, and hallucination rates.

5. Performance Optimization - Continuously refine prompts to minimize cost and latency while maximizing reliability and user trust.

6. Hybrid mode Router Support - Adapt prompt strategies for Gemini Flash vs. Pro vs. Claude depending on use case (speed vs. depth).

7. Guardrails & Safety - Embed safety layers in prompts to prevent unsafe, biased,



or irrelevant outputs.

8. RAG Integration - Build prompt strategies for retrieval-augmented generation (GitHub docs, MDN, StackOverflow).

9. Documentation & Experimentation - Maintain prompt playbooks in Notion, documenting rationale, variations, and outcomes for reproducibility.

10. Cross-Team Collaboration - Work with PM to translate learning flows into prompt chains, AI engineers for orchestration, and QA for regression testing.

11. Problem Solving - Debug prompt failures, identify root causes (model vs. schema vs. orchestration), and propose iterative solutions.

Success Metrics

90 Days (Probation):

- Deliver prompt library v1 for TutorAgent, PlannerAgent, and QuizAgent.

- Build baseline LangSmith eval pipeline with at least 20 golden prompts.

- Achieve <15% hallucination rate in core flows.

12 Months:

- Maintain <5% hallucination rate across all core workflows.

- Prompt system powering 10+ specialized agents in production.

- Automated regression evals integrated into CI/CD pipeline.

- AI outputs powering >50% of artifacts with consistent schema compliance.

Must-Haves

- 10+ years in NLP, prompt engineering, or applied AI.

- Expertise in LLM prompting strategies and evaluation frameworks.

- Strong grasp of JSON schema enforcement, tool-use, and prompt chaining.

- Proven ability to debug and refine AI outputs in production products.

- Product company background with evidence of shipping AI-first systems.

Nice-to-Haves





- Contributions to LangChain, CrewAI, or LangSmith open source.

- Research background in prompt optimization, LLM evals, or safety.

- Prior work on developer-facing AI tutor or EdTech products.

- Startup/founding engineer experience.

Tech Stack Visibility

- Prompting & Orchestration: LangSmith, Langfuse, Google ADK

- Models: Gemini Pro, Gemini Flash, Gemini Vision, Gemini Live API, Claude 3 Opus

- Data: Firestore, Redis, Neo4j (schema enforcement targets)

- Infra: Cloud Run, Pub/Sub for agent comms

- CI/CD: GitHub Actions with prompt regression evals

- Tools: Linear (execution), Notion (prompt specs), GitHub (prompt versions)

Assessment

Objective: Validate ability to design, evaluate, and optimize a prompt system powering multi-agent tutoring.

Challenge (Candidate PoC):

1. Design system prompts for:

- TutorAgent: teaching “React Hooks.”

- QuizAgent: generating 5 MCQs with answers + explanations.

- CodeAgent: generating runnable JS snippets.

- Implement prompt chaining strategy for:

- Outline → Lesson → Quiz → Mini-App.

- Build a LangSmith eval pipeline:

- Test against 20 golden prompts.

- Measure accuracy, latency, hallucination rate, schema compliance.

- Optimize:

- Compare Gemini Flash vs. Pro routing for cost/latency.

- Document tradeoffs and improvements.

Deliverables:

- Prompt library (YAML/JSON).

- Eval report with metrics table (baseline vs. optimized).

- GitHub repo with eval scripts + LangSmith integration.

- Notion doc summarizing prompt strategies.

- 5-min Loom walkthrough of the workflow.

Evaluation Criteria:

- Prompt Design & Schema Compliance (30%)

- Prompt Chaining & Orchestration Strategy (20%)

- Evaluation Pipeline & Metrics (20%)

- Performance Optimization (15%)

- Documentation & Reproducibility (15%)

📌 Prompt Engineer (Founding Team) (Hyderabad)
🏢 Gradientflo Labs
📍 Hyderabad

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: prompt engineer (founding team) (hyderabad) / hyderabad