Uploaded October 2023 | Updated September 2026, 2 weeks ago
This talk walks through the process of transforming an unstructured database of over 30 million entities into a text-searchable resource. Instead of using an existing solution, we opted to build our search in-house using ngrams, serverless computing and object storage. We will deep-dive concepts of indexing, (n)grams and tf-idf. And explore the insights from the successes, challenges and limitations of this technical project.
Akshata (@iamaatoh) is an engineer with 8+ years of building software products across the ed tech, media, and fintech sectors, working in research, multinationals and startups. Previously an architect (for buildings!) and a self-taught dev, she enjoys traveling, painting and geeking out with fellow techies.
Visit https://geekcamp.sg for more information about GeekcampSG
This talk walks through the process of transforming an unstructured database of over 30 million entities into a text-searchable resource. Instead of using an existing solution, we opted to build our search in-house using ngrams, serverless computing and object storage. We will deep-dive concepts of indexing, (n)grams and tf-idf. And explore the insights from the successes, challenges and limitations of this technical project.
Akshata (@iamaatoh) is an engineer with 8+ years of building software products across the ed tech, media, and fintech sectors, working in research, multinationals and startups. Previously an architect (for buildings!) and a self-taught dev, she enjoys traveling, painting and geeking out with fellow techies.
Visit https://geekcamp.sg for more information about GeekcampSG










