Gen AI Researcher – Reasoning & Benchmarks (Artificial Intelligence Researcher) at Recrew AI
Listed by DevShelfHub
About the Role
Frontier reasoning AI is becoming the next battleground for enterprise AI—and Recrew AI, a research-first company incubated at IISc, is building at the cutting edge. Recrew AI is hiring a Gen AI Researcher – Reasoning & Benchmarks in Bangalore to design and build evaluation infrastructure for AI systems at the frontier of reasoning and planning.
This role sits at the core of the company's research agenda, working directly with founders and IISc faculty in a lean, research-driven environment.
What you'll do
- The Gen AI Researcher will design and build reasoning benchmarks to rigorously evaluate AI systems.
- Design, build, and validate reasoning benchmarks spanning science and engineering domains, ensuring rigor and reproducibility
- Develop evaluation pipelines using eval harnesses and benchmark suites (e.g., SWE-bench) to systematically assess AI system performance
- Build and iterate on agents capable of tackling complex reasoning and planning tasks, guided by benchmark results
- Contribute to the company's core research agenda — document methodologies, findings, and experimental results to publication-ready standards
- Collaborate closely with IISc faculty, and internal research, agent, and data teams to align benchmark design with real-world network planning problems
- Engage with the open-source research community to stay current with and contribute to frontier developments in reasoning and evaluation
What we're looking for
- BTech from a top-tier institute (IITs/IISc/BITS or equivalent) with 2+ years of relevant AI/ML research experience, OR MTech/PhD from a top-tier institute
- Solid foundations in probability, statistics, and machine learning
- Hands-on experience with evaluation harnesses and benchmark suites (e.g., SWE-bench or comparable frameworks)
- Strong coding and software engineering skills for reproducible research pipelines
Skills & Technologies
Required
Benefits & Perks
- Health Insurance
- Flexible Leave Policy
- Learning Budget
- EPF / NPS
Why This Role is Good for Experienced Professionals
- This role offers ownership over a core research function at a seed-stage company building frontier reasoning AI.
- Design, build, and validate reasoning benchmarks spanning science and engineering domains
- Develop evaluation pipelines using eval harnesses and benchmark suites (e.g., SWE-bench)
- Build and iterate on agents for complex reasoning and planning tasks
- Contribute to core research agenda with publication-ready documentation
- Collaborate directly with IISc faculty and founding team
- Direct collaboration with a founding team with $100M+ raised and frontier AI publications
- Opportunity to publish and contribute to the broader AI research community
- Lean, research-first culture with high autonomy
- Cross-disciplinary exposure across agents, data, and network AI
- Shape the company's technical direction through your research output
About Recrew AI
- Industry
- Technology
- Company Size
- 500+
- Website
- recrewai.com
Skip the queue
Contact the recruiter directly at Recrew AI and follow up personally — most applicants never do.
Browse Recruiter DatabaseFree trial · 40 contacts · ₹0
Job Details
- Type
- Full time
- Level
- Mid-level
- Experience
- 2–6 years
- Location
- Bangalore, India
- Work Mode
- Hybrid
- Category
- AI / ML
Share this Job
Recruiter Database
Want to stand out? Talk to the recruiter directly.
Most applicants never hear back. Our Recruiter Database gives you direct email access to the people making hiring decisions — so you can follow up personally and actually get noticed.
- Recruiter name, company & direct email address
- Sourced from active job postings across top companies
- Free trial — 40 contacts, no credit card needed