Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

The author claimed that the models he modified with this layer repetition method topped the huggingface open llm leaderboard in his first post: https://dnhkng.github.io/posts/rys/

Do you remember the names of the previous experiments done on this? Would love to take a look.



Just learned about it the other day from this thread from Feb, 2024: https://old.reddit.com/r/LocalLLaMA/comments/1aqrd7t/i_made_...

Has some interesting github links.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: