Staff AI Engineer

PW (PhysicsWallah) · Delhi

  • Experience7–12 yrs
  • SalaryNot disclosed
  • Work modeonsite
  • Posted28 Sept 2026

About PW (PhysicsWallah)

PW (PhysicsWallah) is hiring in Delhi in education. This role looks for around 7+ years of experience.

Skills

  • Generative AI
  • Large Language Models
  • Retrieval-Augmented Generation
  • LLM Fine-tuning
  • LoRA
  • QLoRA
  • PEFT
  • RLHF
  • Quantization
  • Distillation
  • LLM-as-a-Judge
  • vLLM
  • Text Generation Inference
  • SGLang
  • LangChain
  • LangGraph
  • LlamaIndex
  • Pinecone
  • Weaviate
  • pgvector
  • Hugging Face Transformers
  • PyTorch
  • Flash Attention
  • AI Agents
  • Multi-Agent Systems
  • Autonomous AI workflows
  • MLflow
  • Langfuse
  • Weights & Biases
  • LLM latency optimization
  • LLM inference optimization
  • GPU utilization
  • Cost optimization
  • Distributed systems
  • AWS
  • Azure
  • GCP

The role

A generative AI engineer at an education technology company designs production systems using Generative AI, Large Language Models, and Retrieval-Augmented Generation, and builds AI Agents and Multi-Agent Systems. The role also applies PyTorch and MLflow to optimize, evaluate, and observe scalable AI applications.

Full job description

Job Title: AI Staff Data ScientistLocation: Noida / BangaloreExperience: 7–12+ YearsEmployment Type: Full-time

About PhysicsWallah:PhysicsWallah (PW) is one of India's leading education technology companies, empowering millions of learners through innovative digital learning solutions. We are building next-generation AI platforms to transform education through personalized learning, intelligent tutoring, content generation, and AI-powered learning experiences.As an AI Staff Data Scientist, you will play a key technical leadership role in shaping our Generative AI and Large Language Model (LLM) initiatives. You will work closely with Product, Engineering, and AI teams to build scalable, production-grade AI systems that power experiences for millions of students.If you're passionate about solving complex AI challenges and building state-of-the-art LLM applications at scale, we'd love to hear from you.

Roles & Responsibilities:Design, develop, and deploy production-grade Generative AI and LLM-powered applications.Architect and implement scalable Retrieval-Augmented Generation (RAG) pipelines for knowledge-intensive AI systems.Fine-tune foundation models using techniques such as LoRA, QLoRA, PEFT, and RLHF.Optimize LLM inference for latency, throughput, scalability, and infrastructure cost.Build intelligent Agentic AI and multi-agent systems capable of autonomous reasoning and task execution.Design scalable vector search architectures using modern vector databases.Develop robust evaluation pipelines, including LLM-as-a-Judge methodologies for automated quality assessment.Build and maintain end-to-end ML workflows, experimentation pipelines, and model observability.Collaborate with Product Managers, Engineering teams, and Data Scientists to translate business problems into AI-powered solutions.Drive architectural decisions, establish AI best practices, and mentor engineers across teams.Stay up to date with the latest advancements in Generative AI, LLM optimization, and model serving technologies.

Skills & Qualifications:Required Skills:Large Language Models & Generative AIStrong experience designing and deploying production-grade LLM applications.Hands-on expertise with:vLLM or Text Generation Inference (TGI) or SGLangRetrieval-Augmented Generation (RAG)LLM Fine-tuningLoRA, QLoRA, or PEFTReinforcement Learning from Human Feedback (RLHF)Quantization and/or DistillationLLM-as-a-Judge evaluation frameworks

AI FrameworksExperience with one or more of the following:LangChainLangGraphLlamaIndex

Vector DatabasesHands-on experience with one or more:PineconeWeaviatepgvector

Deep Learning & Model OptimizationHugging Face TransformersPyTorchFlash AttentionTransformer architectures

Agentic AIExperience building:AI AgentsMulti-Agent SystemsAutonomous AI workflows

MLOps & Experiment TrackingExperience with:MLflowLangfuseWeights & Biases (W&B)

Performance OptimizationStrong understanding of:LLM latency optimizationLLM inference optimizationGPU utilizationCost optimization for production AI systems

QualificationsBachelor's or Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, or a related field.7–12+ years of experience in Machine Learning, Data Science, AI Engineering, or related domains.Proven experience building and deploying AI solutions in production environments.Strong understanding of distributed systems, scalable architectures, and cloud platforms (AWS, Azure, or GCP).Excellent problem-solving, communication, and stakeholder management skills.Experience mentoring engineers and driving technical initiatives is highly preferred.

Why Join PhysicsWallah?Build AI products that impact millions of learners across India.Work on cutting-edge technologies in Generative AI, LLMs, Agentic AI, and RAG.Solve challenging real-world problems with large-scale AI systems.Collaborate with high-performing Product, Engineering, and AI teams.Enjoy a fast-paced, ownership-driven culture where innovation is encouraged.Lead strategic AI initiatives and influence the future of AI-powered education.Accelerate your career by working on some of the most impactful AI use cases in the EdTech ecosystem.