15 Aug
|
Auric AI Labs
|
Bengaluru
15 Aug
Auric AI Labs
Bengaluru
The hardest problem in our system isn't retrieval. It's deciding when two mentions are the same person.
A full name in one document. An abbreviation in another. A transliteration in a third script. An honorific. A description with no name at all. Everything above this depends on getting it right. Getting it wrong produces something worse than an error: a confident, well-sourced picture of a person who doesn't exist.
We're hiring one person to own it.
Why this work
India has been caught by surprise before. Not because the warning didn't exist, but because it existed somewhere in the system and nobody put it together in time. People died because of that gap. We're closing it.
We're building the intelligence system that helps India see the next attack coming before it happens, and helps it win the next war before the first shot is fired. Not with better sensors; India already collects enough. With the ability to actually use what it collects, at the speed the threat moves.
This doesn't get built by a foreign company, and it doesn't get built for a demo.
It gets built by people who decided this mattered enough to build it here, for real, before it's needed. If we do this right, the payoff is a warning that gets acted on in time, and a war that's already won in preparation before it's fought at all.
What you'd work on
- Identity across scripts:
A name in Devanagari, Roman and a third script has no canonical form. Edit distance assumes a shared alphabet. Phonetic methods were built around English. Multilingual embeddings will cheerfully merge two distinct people who share a name component, the exact failure that ends you. What replaces the standard toolkit is open.
- Decisions under asymmetric cost:
A wrong merge is far more expensive than a missed one, and most of the published literature optimises a metric that assumes the two costs are equal. You'd set thresholds against that asymmetry with very few labels.
- Extraction from data that resists it:
Mixed-language
📌 AI Research Engineer: Entity Resolution (Bengaluru)
🏢 Auric AI Labs
📍 Bengaluru