NextHire Consulting Logo

NextHire Consulting

Forbes Advisor : Data Research Engineer role - AI/ML

Posted 8 Hours Ago
Be an Early Applicant
Remote
Hiring Remotely in India
Mid level
Remote
Hiring Remotely in India
Mid level
Design and implement AI/LLM-driven data extraction and ETL pipelines, build web crawling and API integrations, implement RAG systems, mitigate LLM hallucinations, evaluate third-party tools, and collaborate with content and analytics teams to deliver high-quality, scalable data for downstream analysis.
The summary above was generated by AI
Job Title:

Data Research Engineer – Data Extraction Team

Experience:

4+ Years

Location:

Chennai or Remote (India)

About Forbes Advisor

Forbes Advisor is a new initiative under the Forbes Marketplace umbrella that provides journalist- and expert-written insights, news, and reviews on personal finance, health, business, and everyday life decisions.

Our mission is to help readers turn their aspirations into reality by arming them with trusted advice and data-driven insights, enabling them to make confident decisions and focus on what matters most.

The Marketplace team brings decades of industry experience across geographies and functions including Content, SEO, Business Intelligence, Finance, HR, Marketing, Production, Technology, and Sales, with expertise in diverse sectors such as consumer credit, banking, insurance, small business, education, real estate, and travel.

About the Data Extraction Team

The Data Extraction Team plays a crucial role in designing, implementing, and maintaining advanced web scraping frameworks and data pipelines. The team develops methodologies to gather precise, high-quality data from a wide range of digital sources and ensures its seamless integration into internal systems through ETL (Extract, Transform, Load) processes.

This team also explores the integration of Artificial Intelligence (AI) and Large Language Models (LLMs) to automate and enhance data extraction, processing, and analysis workflows.

Role Overview

The Data Research Engineer will help shape how Forbes Advisor leverages AI and LLM technologies to streamline data operations, optimize research workflows, and build intelligent, scalable data systems.

This is a forward-looking role that combines elements of Data Engineering and AI/LLM Engineering. The ideal candidate is a creative problem-solver who proactively explores emerging technologies and identifies innovative ways to harness their potential for business impact.

Key Responsibilities
  • Develop and implement methods to leverage AI and LLMs (Large Language Models) for process automation and data research efficiency.

  • Proactively identify new AI/LLM-based solutions to streamline operations and improve data workflows.

  • Act as a visionary for AI/LLM adoption, anticipating future technological developments and preparing the team to capitalize on them early.

  • Assist in acquiring and integrating data from multiple sources, including web crawling, APIs, and other data pipelines.

  • Design and optimize ETL workflows to ensure high-quality data availability for downstream analysis.

  • Explore and evaluate third-party tools for modernizing legacy data systems and enhancing scalability.

  • Collaborate cross-functionally with content, research, and analytics teams to understand and fulfill data requirements.

  • Ensure timely delivery of project milestones in a fast-paced, dynamic environment.

  • Support and collaborate with fellow engineers on the Data Research and Extraction Team.

  • Utilize online technical resources effectively (e.g., StackOverflow, ChatGPT, Bard) while understanding their limitations.

Required Skills & Experience
  • Bachelor’s degree in Computer Science, Data Science, Engineering, or a related field (advanced degree is a plus).

  • Minimum 4+ years of experience in Data Engineering, AI/ML Engineering, or related fields.

  • Strong proficiency in Python for data manipulation, automation, and API integration.

  • Experience in AI/ML engineering and data extraction workflows.

  • Proficiency in implementing Retrieval-Augmented Generation (RAG) pipelines using tools such as ChromaDB or Pinecone.

  • Experience with agentic AI platforms (e.g., CrewAI, LangChain) for modular and autonomous task execution.

  • Hands-on experience working with LLMs, including prompt engineering and mitigation of model “hallucinations.”

  • Familiarity with machine learning frameworks such as TensorFlow or PyTorch.

  • Exposure to NLP frameworks (spaCy, NLTK, Hugging Face, etc.).

  • Understanding of SQL and data querying (a plus).

  • Familiarity with web crawling techniques and API integration (a plus).

  • Experience using version control tools such as Git for collaborative development.

  • Strong problem-solving, analytical, and critical-thinking skills.

  • Excellent communication and teamwork abilities.

  • Ability to thrive in a high-growth environment with shifting priorities.

  • Experience managing or mentoring AI/LLM-focused teams is a plus.

  • Familiarity with Agile development methodologies (a plus).

Perks & Benefits
  • Day off on the 3rd Friday of every month (one long weekend each month)

  • Monthly Wellness Reimbursement Program to promote health and well-being

  • Monthly Office Commutation Reimbursement Program

  • Paid Paternity and Maternity Leave

Summary Classification

This position blends Data Engineering and AI/LLM Engineering expertise, focusing on data acquisition, pipeline automation, and intelligent AI-driven workflows.
Ideal for candidates with 4+ years of experience in building scalable data systems, integrating AI/LLM technologies, and driving innovation in data operations.

Similar Jobs

4 Hours Ago
Easy Apply
In-Office or Remote
Easy Apply
Senior level
Senior level
Cloud • Information Technology • Security • Software
Lead quality efforts for major product areas, define test strategies, write and maintain complex automated tests, enable developers to shift testing left, participate in incident reviews and root-cause analysis, mentor junior QEs, and ensure releases meet reliability, performance, and security standards while partnering with architects and development leads.
Top Skills: Ci/CdJavaScriptPlaywrightPytestPythonTypescript
4 Hours Ago
Remote or Hybrid
India
Expert/Leader
Expert/Leader
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Lead clinical trial disclosure strategy and operations for Pfizer-sponsored interventional trials. Ensure timely, compliant posting of protocols, SAPs, CSRs, and clinical summaries per EMA Policy 70 and other regulations. Manage vendors, develop processes and technical solutions, represent Medical Writing on governance committees, and maintain regulatory knowledge and best practices to drive quality and consistency across disclosures.
9 Hours Ago
Easy Apply
Remote
India
Easy Apply
Senior level
Senior level
Artificial Intelligence • Fintech • Hardware • Information Technology • Sales • Software • Transportation
Own and roadmap fleet management product experiences for dispatchers, operators, and drivers. Translate customer pain into web and mobile workflows, define metrics, run betas, work with data scientists to build AI features from sensor data, and collaborate cross-functionally while gathering field feedback.
Top Skills: AIAutomotive SystemsIndustrial IotSQLTelematics

What you need to know about the Pune Tech Scene

Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account