Algo8 AI
201301 / Global
201301 / Global
We're are Hiring @Algo8AI
Role:
Data Engineer — Data Platform & Pipelines
Lake, canonical model and data quality
Location: Noida | Hybrid (4 days/week in office)
Experience: 5–7 years
Employment: Permanent | Algo8 Payroll
Reports to: Tech Lead
Joining: Early joiners / short notice preferred
You will build and own the spine of the programme — the on-premises data platform, the canonical model that everything lands into, and the quality framework that determines whether the foundation can be trusted.
About Algo8
Algo8 is an industrial and manufacturing AI firm. We work inside plants and R&D centres — with engineers, scientists and plant teams — building the data foundations and AI capabilities that enable industrial organisations to make better decisions. Our work is hands-on and close to the client.
The Engagement
You will join a multi-year programme with the R&D centre of excellence of a large Indian manufacturer.
The centre holds four decades of engineering knowledge — design records, formulations, laboratory results, test data, reports and studies — spread across enterprise systems, departmental databases, local servers and document archives.
The programme recovers that knowledge into a single governed foundation, connects it to manufacturing context, and then validates where engineering intelligence can be built on top of it.
This is a data engineering programme first. The AI comes later, and only where the evidence supports it.
The solution is on-premises.
Because of the proprietary nature of new product development, this foundation is being built inside the client's own infrastructure.
If your experience is entirely on managed cloud services, this will not be the right fit.
What You Will Own
The platform
Stand up and operate the on-premises data lake and processing layer — storage, compute, orchestration and access control.
The canonical model
Work with engineers and business analysts to design a shared data model across engineering and manufacturing domains, then keep it coherent as new sources land.
Ingestion pipelines
Build the pipelines that bring recovered and synchronised data into the foundation, with lineage from source to consumption.
Data quality
Define and implement quality thresholds per source. Each estate has to pass an acceptance gate before it is declared complete — you own the evidence behind that.
Enabling everyone else
The integration and document engineers land data into your platform. Analytics and validation work builds on top of it. Making both easy is part of the job.
The Expected Stack
The architecture is being finalised with the client. We expect to work with an on-premises stack along these lines, or close equivalents:
We are more interested in whether you have run this class of system on your own infrastructure than in which specific tools you have used.
What We Need From You
Useful, Not Essential
How You Will Work
Discovery is part of the job — expect to map what is actually there before you design against it.
Who This Role Is Not For
We would rather be direct than waste your time. This role is probably not a fit if:
Working at Algo8
Applying
We need this team in place early.
Please tell us your notice period when you apply — candidates who can join immediately or at short notice will be prioritised.
Our process is a CV review, a short written assignment, and interviews.
If you have the experience to build and own an on-premises data platform from the ground up, we'd like to hear from you.
You can share your CVs at: [email protected]
Kanpur / Global
Uttar Pradesh / Global
201301 / Global
Kanpur / Global
Uttar Pradesh / Global
Kanpur / Global