Skip to content

Navigation Menu

Sign in
Appearance settings
Sign up
Appearance settings
View DaoyuanLi2816's full-sized avatar
  • Greater Seattle Area

Highlights

  • Pro

Block or report DaoyuanLi2816

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
DaoyuanLi2816/README.md

Pinned Loading

  1. huggingface/peft huggingface/peft Public

    🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.

    Python 21.6k 2.5k

  2. deepseek-ai/DeepSpec deepseek-ai/DeepSpec Public

    DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms

    Python 7k 659

  3. can-i-finetune-this can-i-finetune-this Public

    Estimate whether a Hugging Face model fits and fine-tunes on your local GPU.

    Python 792 107

  4. mini-verl mini-verl Public

    Run a documented subset of verl-style OPD on one consumer GPU—typed config, Parquet prompts, and PEFT scale-out artifacts.

    Python 191 50

  5. pairjudge pairjudge Public

    Pairwise LLM judges (A/B/tie): budget-aware multi-turn packing, position-bias correction, pseudo-label distillation. Generalized from the 4th-place (gold) solution to Kaggle LMSYS Chatbot Arena.

    Python 170 12

  6. tracedistill tracedistill Public

    Distill teacher chains-of-thought into a LoRA adapter via a strict boxed-answer format contract + two-phase Train→Nudge (silver-medal NVIDIA Nemotron reasoning recipe, as a tested library).

    Python 85 17

Morty Proxy This is a proxified and sanitized view of the page, visit original site.