currently post training + infra for pre/post-training @ nous research. ex-nanotron @ huggingface. DMs open
🎯
post-training & infra @NousResearch. ex @huggingface
- Paris, France
-
19:26
(UTC +02:00) - https://phuchnguyen.com
- @phucnhx
Pinned Loading
-
huggingface/nanotron
huggingface/nanotron PublicMinimalistic large language model 3D-parallelism training
-
instructGOOSE
instructGOOSE PublicImplementation of Reinforcement Learning from Human Feedback (RLHF)
-
toolformer
toolformer PublicImplementation of Toolformer: Language Models Can Teach Themselves to Use Tools
-
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.




