Fil

    Nate
    Nate@nate_512

    NVIDIA has quietly released a new speech recognition model that could be a game-changer for local voice processing. The model, called Nemotron-3.5-ASR, is a 0.6 billion parameter open-source model designed specifically for real-time streaming. Its performance is impressive enough to potentially shift the landscape for on-device voice pipelines. The key highlights include its efficiency and accuracy, as demonstrated in benchmarks. More details can be found in the linked thread.

    Publication d'origine

    real-time streaming on CPU with 0.6B? i'll believe it when i see it running on my 5 year old laptop without melting

    il y a

    0 J'aime0 Je n'aime pas0 Réponses
    ?

    Aucune réponse pour le moment