Skip to content
View DaoyuanLi2816's full-sized avatar
  • Greater Seattle Area

Highlights

  • Pro

Block or report DaoyuanLi2816

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
DaoyuanLi2816/README.md

Pinned Loading

  1. huggingface/peft huggingface/peft Public

    🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.

    Python 21.5k 2.4k

  2. deepseek-ai/DeepSpec deepseek-ai/DeepSpec Public

    DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms

    Python 6.9k 646

  3. can-i-finetune-this can-i-finetune-this Public

    Estimate whether a Hugging Face model fits and fine-tunes on your local GPU.

    Python 792 107

  4. pairjudge pairjudge Public

    Pairwise LLM judges (A/B/tie): budget-aware multi-turn packing, position-bias correction, pseudo-label distillation. Generalized from the 4th-place (gold) solution to Kaggle LMSYS Chatbot Arena.

    Python 169 12

  5. mini-verl mini-verl Public

    Auditable one-GPU alignment and distillation runtime with shared-backbone training and a fail-closed verl artifact bridge.

    Python 59 13

  6. tracedistill tracedistill Public

    Distill teacher chains-of-thought into a LoRA adapter via a strict boxed-answer format contract + two-phase Train→Nudge (silver-medal NVIDIA Nemotron reasoning recipe, as a tested library).

    Python 43 5