Skip to content
View cameron-chen's full-sized avatar

Highlights

  • Pro

Block or report cameron-chen

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. sail-sg/oat sail-sg/oat Public

    🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.

    Python 669 63

  2. sail-sg/dice sail-sg/dice Public

    Official implementation of Bootstrapping Language Models via DPO Implicit Rewards

    Python 49 3

  3. sail-sg/understand-r1-zero sail-sg/understand-r1-zero Public

    Understanding R1-Zero-Like Training: A Critical Perspective

    Python 1.3k 63

  4. axon-rl/gem axon-rl/gem Public

    A Gym for Agentic LLMs

    Python 506 34