Welcome to open-posttraining-system Discussions! #1
shaheennabi
announced in
Announcements
Replies: 0 comments
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Uh oh!
There was an error while loading. Please reload this page.
👋 Welcome to Open Post-Training System Discussions
Welcome to the community!
Open Post-Training System is an open-source effort focused on understanding, implementing, and advancing the post-training stack for reasoning models. The project covers topics such as supervised reasoning fine-tuning, preference learning, reward modeling, reinforcement learning for reasoning, inference-time scaling, evaluation, and the infrastructure needed to support these systems.
This Discussions space is where contributors, learners, researchers, and engineers can connect, ask questions, share ideas, and collaborate.
What can you use Discussions for?
❓ Ask Questions
Stuck on an implementation? Confused by a concept? Need help understanding a paper, algorithm, or training pipeline? Ask here.
💡 Share Ideas
Have an idea for improving the project, adding new notebooks, implementing a paper, or expanding the evaluation framework? We'd love to hear it.
📚 Discuss Research
Talk about reasoning models, RLHF, RLVR, preference optimization, reward models, inference-time compute, evaluation methodologies, and related research directions.
🛠️ Collaborate
Interested in contributing code, documentation, experiments, benchmarks, or educational content? Start a discussion and introduce yourself.
🚀 Showcase Your Work
Built something using the project? Trained a model? Reproduced an experiment? Share your results and lessons learned with the community.
Community Values
We aim to build a collaborative, respectful, and technically rigorous community.
Please:
Introduce Yourself
If you're new here, we'd love to meet you.
Tell us:
Whether you're just getting started or already working on frontier AI systems, you're welcome here.
Let's build open reasoning systems together.
All reactions