- Princeton, NJ, USA
-
19:02
(UTC -04:00) - https://orcid.org/0009-0006-3667-0665
- in/heyu-guo
Stars
[ICRA 2026] VITRA: Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos
A curated list of state-of-the-art research in embodied AI, focusing on vision-language-action (VLA) models, vision-language navigation (VLN), and related multimodal learning approaches.
RFSec-ToolKit is a collection of Radio Frequency Communication Protocol Hacktools.无线通信协议相关的工具集,可借助SDR硬件+相关工具对无线通信进行研究。Collect with ♥ by HackSmith
Train transformer language models with reinforcement learning.
Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging
A simulation framework for RLHF and alternatives. Develop your RLHF method without collecting human data.
Reference implementation for DPO (Direct Preference Optimization)
主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题
A collection of 3D reconstruction papers in the deep learning era.
Simulation codes for "Channel Estimation and Equalization for CP-OFDM-based OTFS in Fractional Doppler Channels"
Examples and guides for using the Gemini API
Course review material for the Jan 2024 course "Intro to Version Control with Git and GitHub"
SSD: Single Shot MultiBox Detector | a PyTorch Tutorial to Object Detection
Image augmentation for machine learning experiments.
Multi-object tracking, Multi-camera, Mouse group, Deep learning, Faster R-CNN, Tracklets fusion
Image annotation with Python. Supports polygon, rectangle, circle, line, point, and AI-assisted annotation.