Mostly about the unglamorous half of building with language models:
deciding whether the output actually got better, and what to do when
it did not.
I hand-labelled 149 examples before I trusted my own scorer
Everyone can generate. Far fewer can tell you whether the output got better. Notes on building a small evaluation set by hand, calibrating a scorer against it, and what the disagreements taught me.
evaluationllmragagentsengineering
Read the post →
Bridging Cultures: My Journey from India to Japan
A personal reflection on growing up in India, living and working in Japan, and learning to navigate two cultures that shaped who I am.
Read the post →
Autonomous Driving: Concepts, Levels, and System Architecture
An overview of autonomous driving concepts, SAE J3016 automation levels, and core system components used in modern self-driving vehicles.
Install NDDS (NVIDIA Deep Learning Dataset Synthesizer)
Step-by-step guide to installing NVIDIA Deep Learning Dataset Synthesizer (NDDS) with Unreal Engine 4.22 for generating synthetic datasets for deep learning.