1 14 10

Le Yu

vanillaOVO

https://yule-buaa.github.io/

yule-BUAA

AI & ML interests

None yet

Recent Activity

upvoted a paper 11 days ago

CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings

upvoted a paper 23 days ago

Qwen2.5 Technical Report

upvoted a paper 3 months ago

A Unified View of Delta Parameter Editing in Post-Trained Large-Scale Models

View all activity

Organizations

None yet

vanillaOVO's activity

upvoted a paper 11 days ago

CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings

Paper • 2501.01257 • Published 12 days ago • 46

upvoted a paper 23 days ago

Qwen2.5 Technical Report

Paper • 2412.15115 • Published 25 days ago • 339

upvoted a paper 3 months ago

A Unified View of Delta Parameter Editing in Post-Trained Large-Scale Models

Paper • 2410.13841 • Published Oct 17, 2024 • 16

upvoted a collection 6 months ago

Llama 3.1

Collection

This collection hosts the transformers and original repos of the Llama 3.1, Llama Guard 3 and Prompt Guard models • 11 items • Updated Dec 6, 2024 • 639

upvoted an article 6 months ago

Article

Llama 3.1 - 405B, 70B & 8B with multilinguality and long context

Jul 23, 2024

• 226

upvoted a collection 7 months ago

Qwen2

Collection

Qwen2 language models, including pretrained and instruction-tuned models of 5 sizes, including 0.5B, 1.5B, 7B, 57B-A14B, and 72B. • 39 items • Updated Nov 28, 2024 • 354

upvoted 2 articles 9 months ago

Article

Fine-tune Llama 3 with ORPO

•

Apr 22, 2024

• 230

Article

Merge Large Language Models with mergekit

•

Jan 9, 2024

• 89

upvoted a paper 10 months ago

DoRA: Weight-Decomposed Low-Rank Adaptation

Paper • 2402.09353 • Published Feb 14, 2024 • 26

upvoted 2 papers 11 months ago

Resolving Interference When Merging Models

Paper • 2306.01708 • Published Jun 2, 2023 • 13

Neural Network Diffusion

Paper • 2402.13144 • Published Feb 20, 2024 • 95

upvoted a collection 12 months ago

Papers about model merging

Collection

referenced in the mergekit repo: https://github.com/cg123/mergekit • 4 items • Updated Feb 13, 2024 • 14

upvoted 2 papers 12 months ago

Large Language Models are Superpositions of All Characters: Attaining Arbitrary Role-play via Self-Alignment

Paper • 2401.12474 • Published Jan 23, 2024 • 35

Language Models are Super Mario: Absorbing Abilities from Homologous Models as a Free Lunch

Paper • 2311.03099 • Published Nov 6, 2023 • 28