Can Jev Help Voice Bots Stop Taking Wrong Actions On Misheard Words?
The problem nobody likes admitting Here's the thing about voice bots: they're really good at sounding sure of themselves even when they have no idea what just ...
11 entries found in this collection
The problem nobody likes admitting Here's the thing about voice bots: they're really good at sounding sure of themselves even when they have no idea what just ...
Everyone talks about "AI inference" like it's one problem. It isn't. Serving text and serving speech are as different as catering a wedding and running a dosa s...
Is My RAG System Over-Engineered? I Tested It. If you're building with Large Language Models, you've probably heard of Retrieval-Augmented Generation, or RAG. ...
I Benchmarked 6 Vector Databases for RAG — Here’s What Surprised Me Most If you’ve ever worked with LLMs, you’ve probably hit the same wall I did — choosing t...
The Papers That Built ABCs of LLMs From sequence-to-sequence models to transformers that changed everything --- Timeline of Innovation | Date | Pa...
CNNs ,RNNs and LSTM <img src="https://raw.githubusercontent.com/somiljain7/somiljain7.github.io/main/images/foundationpapers.png"> | Paper ...
Welcome to the wonderfully wacky world of AI image generation, where noise isn’t just unwanted static—it’s the secret sauce that transforms raw data into breath...
I've been looking into getting started with using transformers for speech. I've been doing some reading and attended a talk where I learned about using Hubert f...
====== I was checking out few repositories for Language translation and came across set of following keywords which got me more interested towards checking thes...
A Deep Dive into Automatic Speech Recognition Technology ====== ASR, or automatic speech recognition, is a technology that aims to convert spoken utterances int...
Getting started with Whisper by OPENAI ====== Whisper is a cutting-edge speech recognition model developed by OpenAI in October 2022. Its primary purpose is to ...