AI / ML Engineer
- Engineering
- Remote (India)
- Full-time
- 3+ years
- Python
- LLMs
- RAG Pipelines
- Cloud AI Services
Build the AI layer inside our client products — retrieval pipelines, LLM integrations, evaluation harnesses, and agents that actually hold up in production. This is applied AI with real users and real constraints, not research for its own sake. You will own intelligent features from prototype to production and keep them reliable once they are there.
Model Development & Integration
Integrate and orchestrate LLM APIs (Claude, OpenAI, open-weight models) with proper guardrails.
Apply modern techniques across NLP, generative AI, and where relevant, computer vision.
Fine-tune prompts, context strategies, and model selection for accuracy, cost, and latency.
Retrieval & Data Pipelines
Design and implement RAG pipelines, embeddings, and vector-search systems for client products.
Build ingestion and preprocessing flows that turn messy, unstructured data into usable context.
Ensure data quality, privacy, and compliance across everything that feeds a model.
Evaluation & Optimization
Build evaluation suites so AI features have measurable quality, not vibes.
Run experiments and A/B tests to compare approaches and prove improvements.
Monitor production models for latency, cost, and quality drift, and act before users notice.
Collaboration & Research
Collaborate with full-stack engineers to ship AI features behind clean, stable APIs.
Translate ambiguous business goals into practical, deliverable AI architectures.
Stay current on the AI landscape and advise clients on what is real versus hype.
What We're Looking For
3+ years in software or ML engineering with strong Python skills.
Hands-on production experience with LLM APIs and at least one RAG or agent system.
Understanding of embeddings, vector databases, and prompt engineering fundamentals.
Experience with cloud AI services (AWS Bedrock, Azure OpenAI, or GCP Vertex).
Ability to reason about latency, token cost, and failure modes of AI systems.
Nice to Have
Experience fine-tuning open-weight models or building evaluation frameworks.
Familiarity with LangChain, LlamaIndex, or similar orchestration frameworks — and their limits.
Published side projects, blog posts, or open-source contributions in the AI space.
Don't tick every box? Apply anyway — we hire for trajectory and attitude, not checklists.
- 01
Apply
Send us your resume and a few lines about the work you are proudest of.
- 02
Intro Call
A relaxed 30-minute conversation about your goals, experience, and the role.
- 03
Skills Round
A practical technical or portfolio discussion — real problems, no trick puzzles.
- 04
Offer & Onboarding
A clear offer, a structured first month, and a buddy to help you settle in.
Apply for this role
Under five minutes. We review every submission personally.
- Department
- Engineering
- Location
- Remote (India)
- Job Type
- Full-time
- Experience
- 3+ years
Prefer email? support@neonaitech.com
What You Get
Health Insurance Coverage
Flexible & Remote Work
Generous Paid Time Off
Learning & Certification Budget
Performance Bonuses
Fast-Track Career Growth
Team Retreats & Celebrations
Latest Tools & Hardware
