Skip to content
@THUDM

THUKEG

ChatGLM, GLM-4, CogVLM, CodeGeeX, CogView, ImageReward, CogVideoX | CogDL, GraphMAE, AMiner | Zhipu.ai (Z.ai) & Knowledge Engineering Group (KEG)

Pinned Loading

  1. GLM GLM Public

    GLM (General Language Model)

    Python 3.7k 371

  2. slime slime Public

    slime is an LLM post-training framework for RL Scaling.

    Python 8.4k 1.2k

  3. P-tuning-v2 P-tuning-v2 Public

    An optimized deep prompt tuning strategy comparable to fine-tuning across scales and tasks

    Python 2.1k 212

  4. ReST-MCTS ReST-MCTS Public

    ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search (NeurIPS 2024)

    Python 712 52

  5. T1 T1 Public

    RL Scaling and Test-Time Scaling (ICML'25)

    116 1

  6. AgentRL AgentRL Public

    Scaling Agentic Reinforcement Learning with a Multi-Turn, Multi-Task Framework

    Python 352 29

Repositories

Showing 10 of 130 repositories
  • Awesome-Parameter-Efficient-Fine-Tuning-for-Foundation-Models Public

    Parameter-Efficient Fine-Tuning for Foundation Models

    THUDM/Awesome-Parameter-Efficient-Fine-Tuning-for-Foundation-Models's past year of commit activity
    110 4 0 0 Updated Sep 8, 2026
  • slime Public

    slime is an LLM post-training framework for RL Scaling.

    THUDM/slime's past year of commit activity
    Python 8,441 Apache-2.0 1,239 224 264 Updated Sep 3, 2026
  • INFTY Public

    INFTY Engine: An Optimization Toolkit to Support Continual AI

    THUDM/INFTY's past year of commit activity
    Python 575 MIT 13 6 1 Updated Sep 1, 2026
  • SCALE-CUA Public

    Open-source framework for computer use agents: VeriGen verifiable task synthesis, online RL training (AgentRL), and OSWorld/ScienceBoard evaluation.

    THUDM/SCALE-CUA's past year of commit activity
    Python 57 2 1 0 Updated Aug 3, 2026
  • KARL Public
    THUDM/KARL's past year of commit activity
    Python 4 1 0 0 Updated Jul 7, 2026
  • CodeRM-NT Public

    [Findings of ACL 2026] CodeRM-NT: Reward Model for Code RL without Unit Tests

    THUDM/CodeRM-NT's past year of commit activity
    Python 0 0 0 0 Updated Jul 1, 2026
  • DeepDive Public

    DeepDive: Advancing Deep Search Agents with Knowledge Graphs and Multi-Turn RL

    THUDM/DeepDive's past year of commit activity
    Python 346 38 1 0 Updated Jun 17, 2026
  • ReST-RL Public

    Reinforcing LLM Reasoning through Self-Training and Value-Guided Decoding

    THUDM/ReST-RL's past year of commit activity
    Python 19 MIT 0 0 0 Updated May 6, 2026
  • CaRR Public

    This repository contains the code and data for the paper "Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents with Citation-Aware Rubric Rewards".

    THUDM/CaRR's past year of commit activity
    Python 74 MIT 9 0 0 Updated Apr 8, 2026
  • IndexCache Public

    IndexCache: Accelerating Sparse Attention via Cross-Layer Index Reuse

    THUDM/IndexCache's past year of commit activity
    141 11 7 1 Updated Mar 14, 2026