Talentmate
India
27th September 2026
2609-17024-17
The hardest problem in our system isn't retrieval. It's deciding when two mentions are the same person.
A full name in one document. An abbreviation in another. A transliteration in a third script. An honorific. A description with no name at all. Everything above this depends on getting it right. Getting it wrong produces something worse than an error: a confident, well-sourced picture of a person who doesn't exist.
We're hiring one person to own it.
Why this work
India has been caught by surprise before. Not because the warning didn't exist, but because it existed somewhere in the system and nobody put it together in time. People died because of that gap. We're closing it.
We're building the intelligence system that helps India see the next attack coming before it happens, and helps it win the next war before the first shot is fired. Not with better sensors; India already collects enough. With the ability to actually use what it collects, at the speed the threat moves.
This doesn't get built by a foreign company, and it doesn't get built for a demo. It gets built by people who decided this mattered enough to build it here, for real, before it's needed. If we do this right, the payoff is a warning that gets acted on in time, and a war that's already won in preparation before it's fought at all.
What you'd work on
Why it's hard
At this scale all-pairs comparison is out, and standard candidate generation assumes comparable surface forms, the exact assumption that breaks first here. Rare-event base rates are punishing, which makes overall accuracy meaningless and puts all the weight on the top of a ranked list.
Who this is for
Given an ambiguous problem, you decompose it, name the tradeoffs, and design, rather than looking for a tutorial. You understand embeddings and retrieval mechanically. You've worked on real, messy data, and you know a system that's right 95% of the time can be useless if the other 5% is catastrophic.
Direct experience with entity resolution, knowledge graphs, or multilingual NLP is a strong signal, not a prerequisite. Linguistics, IR, database theory and formal methods backgrounds all interest us. How you think matters more than what you've built.
Not this role: RAG over a document store, or maintaining an ontology someone else designed.
| Role Level: | Not Applicable | Work Type: | Full-Time |
|---|---|---|---|
| Country: | India | City: | Bengaluru ,Karnataka |
| Company Website: | auricai.in | Job Function: | Data Science & AI |
| Company Industry/ Sector: |
Other | ||
Searching, interviewing and hiring are all part of the professional life. The TALENTMATE Portal idea is to fill and help professionals doing one of them by bringing together the requisites under One Roof. Whether you're hunting for your Next Job Opportunity or Looking for Potential Employers, we're here to lend you a Helping Hand.
Disclaimer: talentmate.com is only a platform to bring jobseekers & employers together. Applicants are advised to research the bonafides of the prospective employer independently. We do NOT endorse any requests for money payments and strictly advice against sharing personal or bank related information. We also recommend you visit Security Advice for more information. If you suspect any fraud or malpractice, email us at abuse@talentmate.com.