Skip to content
View huckiyang's full-sized avatar
💮
love life. live life.
💮
love life. live life.

Highlights

  • Pro

Block or report huckiyang

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. NVIDIA-NeMo/Speech NVIDIA-NeMo/Speech Public

    A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

    Python 18.6k 3.6k

  2. NVlabs/OmniVinci NVlabs/OmniVinci Public

    OmniVinci is an omni-modal LLM for joint understanding of vision, audio, and language.

    Python 683 56

  3. Voice2Series-Reprogramming Voice2Series-Reprogramming Public

    ICML 21 - Voice2Series: Adversarial Reprogramming Acoustic Models for Time Series Classification

    TypeScript 72 11

  4. QuantumSpeech-QCNN QuantumSpeech-QCNN Public

    IEEE ICASSP 21 - Quantum Convolution Neural Networks for Speech Processing and Automatic Speech Recognition

    Jupyter Notebook 108 20

  5. agwer agwer Public

    agent-oriented wer open metrics

    Python 16

  6. ttt-lm-mlx ttt-lm-mlx Public

    mlx test-time training layer in apple silicon

    Python 1