Skip to content
View yangjianxin1's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report yangjianxin1

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

DeepSeek Harness: Everything is a Plugin.

TypeScript 217,475 25,728 Updated Sep 9, 2026

Framework for evaluating and improving agents

Python 5,073 1,758 Updated Sep 10, 2026

🐧 Harness for RSI. Let AI Build AI. Multi-Agent Auto-Dev Platform. Everything is Transparent.

TypeScript 2,049 216 Updated Sep 10, 2026

Qwen-AgentWorld: Language World Models for General Agents

Python 1,000 92 Updated Jul 20, 2026

Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning

Python 451 49 Updated May 28, 2026

AI agents running research on single-GPU nanochat training automatically

Python 95,495 13,412 Updated Mar 26, 2026

🦉 OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation

Python 20,127 2,299 Updated Aug 27, 2026

"OpenHarness: Open Agent Harness with a Built-in Personal Agent--Ohmo!"

Python 15,692 2,553 Updated Jun 4, 2026

A benchmark for LLMs on complicated tasks in the terminal

Python 2,573 568 Updated Jul 11, 2026

The open source coding agent.

TypeScript 206,179 26,944 Updated Sep 10, 2026

Official code of "StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs".

Python 77 5 Updated Jun 23, 2025

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

Python 2,294 220 Updated Jun 9, 2026

Pytorch Implementation of "Multi-Level Optimal Transport for Universal Cross-Tokenizer Knowledge Distillation on Language Models", AAAI 2025

Python 38 4 Updated Feb 4, 2026

Pytorch Implementation of "Sinkhorn Distance Minimization for Knowledge Distillation", COLING 2024 and TNNLS 2024

Python 129 2 Updated Apr 27, 2025

Unleashing the Power of Reinforcement Learning for Math and Code Reasoners

Python 738 44 Updated Jun 6, 2025

Super-Efficient RLHF Training of LLMs with Parameter Reallocation

Python 336 22 Updated Apr 24, 2025

Official Repo for Open-Reasoner-Zero

Python 2,099 120 Updated Jun 2, 2025

Scalable RL solution for advanced reasoning of language models

Python 1,874 115 Updated Mar 18, 2025

Speech recognition

C 1,506 223 Updated Sep 4, 2026

An MCP-based chatbot | 一个基于MCP的聊天机器人

C++ 29,771 6,904 Updated Sep 9, 2026

Efficient Triton Kernels for LLM Training

Python 6,607 597 Updated Sep 10, 2026

O1 Replication Journey

2,001 61 Updated Jan 14, 2025
Python 1,340 54 Updated Nov 21, 2024

[ICLR 2024]EMO: Earth Mover Distance Optimization for Auto-Regressive Language Modeling(https://arxiv.org/abs/2310.04691)

Python 129 13 Updated Mar 7, 2024

POT : Python Optimal Transport

Python 2,841 561 Updated Sep 9, 2026

OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Python 1,855 131 Updated Jan 17, 2025

[ICLR2023] PLOT: Prompt Learning with Optimal Transport for Vision-Language Models

Python 177 14 Updated Dec 14, 2023

Implementation of Sinkhorn algorithms in Torch.

Python 4 1 Updated Aug 16, 2024

code for paper "BiLD: Bi-directional Logits Difference Loss for Large Language Model Distillation"

Python 11 1 Updated Dec 17, 2024
Next