Skip to content
View kkli08's full-sized avatar
🌳
Focusing
🌳
Focusing

Organizations

@CMPUT301F21T24

Block or report kkli08

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
kkli08/README.md

Hi, I'm Ke Li

LLM post-training · RLVR · GPU and systems engineering

I'm an engineer at XPENG Robotics, where I work on reinforcement learning with verifiable rewards (RLVR) for large language models. I hold an M.Eng. in Electrical and Computer Engineering from the University of Toronto and a B.A.Sc. in Computer Science from the University of Alberta.

I enjoy turning model-training ideas into reliable systems, with a particular interest in efficient post-training, inference, and the low-level software that makes them work.

Current focus

  • RL post-training
  • High-performance systems with C++, CUDA, and Rust

Elsewhere

Pinned Loading

  1. VeloxDB VeloxDB Public

    Persistent Key-Value Storage Database Library.

    C++ 1

  2. AReaL AReaL Public

    Forked from areal-project/AReaL

    The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

    Python