FELLOWSHIP & PPO TRACK AI Engineering 100% REMOTE (INDIA)

Synthetic Reasoning & Evaluation Intern

Build benchmark datasets, evaluate chain-of-thought accuracy, and verify mathematical logic deduction traces for frontier AI models.

Compensation Paid Fellowship — ₹15,000 – ₹24,000/month + Full-Time PPO Option
Location & Work Model Remote (India)
Experience Requirement Freshers / Final Year
Hiring Benchmark 5-Day Code Audit
SECTION 01

Role Overview & Operational Scope

Evaluation is the engine of AI progress. As a Reasoning & Evaluation Intern, you will verify whether agents genuinely deduce correct solutions or merely hallucinate plausible-looking steps, ensuring our reasoning engines meet enterprise accuracy bars.

PythonMathematical LogicReasoning BenchmarksPrompting TechniquesData AnalysisPandasGitHub
SECTION 02

Key Responsibilities & Production Deliverables

  • Assemble and annotate high-difficulty benchmark problems spanning formal logic, math, and software architecture.
  • Execute automated test runs across frontier models and analyze reasoning trace failure points.
  • Implement scoring heuristics that verify intermediate reasoning lemmas.
  • Visualize accuracy trends and token efficiency metrics for engineering leadership.
  • Support senior researchers in prompt optimization and error categorization.
SECTION 03

Mandatory Foundational Knowledge

  • Strong foundation in discrete mathematics, logic, or computer science theory.
  • Familiarity with Python for data manipulation and test automation.
  • Understanding of reasoning techniques: Chain-of-Thought, Self-Consistency, and verification.
SECTION 04

Mandatory Practical Skills & Architecture

  • Proficiency with Python, Pandas, and JSON data structures.
  • Sharp analytical eye capable of spotting subtle logic flaws in model outputs.
  • Disciplined data management and documentation skills.
SECTION 05

Problem Solving, Execution Rigor & Curiosity

  • A love for logic puzzles, competitive programming, or mathematical problem solving.
  • Reluctance to accept unsubstantiated assertions without empirical verification.
  • High ambition to develop deep expertise in machine learning evaluation.
SECTION 06 · PRACTICAL EVALUATION BENCHMARK

5-Day Live Technical Evaluation Milestone

5-Day Practical Milestone: Construct an automated evaluation suite that tests a model on 30 multi-step logic problems, scoring intermediate steps and computing verification precision (strictly 5 working days).

Institutional Hiring Protocol: Candidates who pass initial resume screening are invited to a live, practical evaluation milestone spanning strictly not more than 5 working days. Verifiable completion and code audit by your assigned senior engineering mentor is the sole prerequisite for official corporate offer letter issuance.

SECTION 07

Compensation, Total Rewards & Advancement

  • Paid stipend (₹15,000–₹24,000/month) with performance bonuses.
  • Direct PPO track into Full-Time Synthetic Reasoning Architect.
  • Deep practical immersion in frontier AI reasoning architectures.
  • Flexible remote working conditions.
SECTION 08 · DIRECT INQUIRIES & CATCH-ALL

Dedicated Inquiries Inbox for This Role

Have questions regarding architecture scope or wish to share private research repos directly? Messages sent to this address route straight to the engineering leads reviewing this opening.

synthetic-reasoning-evaluation-i-careers@cehpoint.co.in Open Mail Client →

Related Engineering Appointments

FELLOWSHIP

AI Personality & Persona Research Intern

Remote (India) · Paid Fellowship — ₹15,000 – ₹25,000/month + Full-Time PPO (₹12–18 LPA)
View Specification →
FELLOWSHIP

Agentic Skill & Tool Synthesis Intern

Remote (India) · Paid Fellowship — ₹14,00,000 CTC PPO Track (Stipend: ₹15,000–₹22,000/month)
View Specification →
FELLOWSHIP

AI Memory Systems & Vector DB Intern

Remote (India) · Paid Fellowship — ₹14,000 – ₹22,000/month + PPO (₹10–16 LPA)
View Specification →