April 14, 2024, 1:23 p.m. | /u/Puzzleheaded_Bee5489

Machine Learning www.reddit.com

I'm working on a user authentication project using **voice** i.e, **voice authentication.** I was researching on what are the different aspects of a given audio/speech which I can make use of to identify a particular person, one of the most commonly used things are the [MFCC](https://www.kaggle.com/code/ilyamich/mfcc-implementation-and-tutorial) features, which are extracted using any standard audio processing library like Librosa.

Now, in recent times we have Embeddings which essential capture the information in the form of vectors, it could be audio, video, …

audio embeddings features form image information machinelearning mean text the information the way vectors video

Senior Machine Learning Engineer

@ GPTZero | Toronto, Canada

ML/AI Engineer / NLP Expert - Custom LLM Development (x/f/m)

@ HelloBetter | Remote

Doctoral Researcher (m/f/div) in Automated Processing of Bioimages

@ Leibniz Institute for Natural Product Research and Infection Biology (Leibniz-HKI) | Jena

Seeking Developers and Engineers for AI T-Shirt Generator Project

@ Chevon Hicks | Remote

Principal Data Architect - Azure & Big Data

@ MGM Resorts International | Home Office - US, NV

GN SONG MT Market Research Data Analyst 11

@ Accenture | Bengaluru, BDC7A