Skip to content
@RUCBM

RUCBM

Popular repositories Loading

  1. G-OPD G-OPD Public

    Official repository for the paper "Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation"

    Python 296 13

  2. GUICourse GUICourse Public

    GUICourse: From General Vision Langauge Models to Versatile GUI Agents

    Python 144 7

  3. DeepCritic DeepCritic Public

    Official repository for paper "DeepCritic: Deliberate Critique with Large Language Models"

    Python 41

  4. LaSeR LaSeR Public

    [ICLR 2026] Official repository for the paper "LaSeR: Reinforcement Learning with Last-Token Self-Rewarding"

    Python 36 3

  5. AgentProcessBench AgentProcessBench Public

    Python 34 2

  6. AtomMem AtomMem Public

    Python 30

Repositories

Showing 10 of 19 repositories
  • TIPS Public

    Inducing Process Supervision from Outcome-Only Reinforcement Learning

    RUCBM/TIPS's past year of commit activity
    Python 0 Apache-2.0 0 0 0 Updated Aug 27, 2026
  • CURE Public

    Official repository for the ACL 2026 paper CURE: Critique-Driven Unified Reinforcement Learning for Test-Time Self-Improvement

    RUCBM/CURE's past year of commit activity
    Python 7 Apache-2.0 0 0 0 Updated Jul 6, 2026
  • TTExplore Public

    Test-Time Deep Thinking to Explore Implicit Rules

    RUCBM/TTExplore's past year of commit activity
    SAS 0 0 0 0 Updated Jun 24, 2026
  • ExpInternalization Public

    Official code for "Rethinking Continual Experience Internalization for Self-Evolving LLM Agents"

    RUCBM/ExpInternalization's past year of commit activity
    Python 10 0 1 0 Updated Jun 16, 2026
  • RUCBM/AgentProcessBench-Homepage's past year of commit activity
    CSS 0 0 0 0 Updated May 29, 2026
  • G-OPD Public

    Official repository for the paper "Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation"

    RUCBM/G-OPD's past year of commit activity
    Python 296 MIT 13 1 0 Updated May 28, 2026
  • DelTA Public

    Code for Paper 'DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards'

    RUCBM/DelTA's past year of commit activity
    Python 18 0 2 0 Updated May 21, 2026
  • AtomMem Public
    RUCBM/AtomMem's past year of commit activity
    Python 30 0 1 0 Updated Mar 31, 2026
  • RUCBM/AgentProcessBench's past year of commit activity
    Python 34 2 0 0 Updated Mar 17, 2026
  • DARC Public
    RUCBM/DARC's past year of commit activity
    Python 13 0 1 0 Updated Mar 8, 2026

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Top languages

Python CSS SAS

Most used topics

Loading…