Picture for Yifei Cao

Yifei Cao

Predicting Large Language Model Capabilities on Closed-Book QA Tasks Using Only Information Available Prior to Training

Add code
Feb 06, 2025
Viaarxiv icon

TranStable: Towards Robust Pixel-level Online Video Stabilization by Jointing Transformer and CNN

Add code
Jan 25, 2025
Viaarxiv icon

Mutual Information as Intrinsic Reward of Reinforcement Learning Agents for On-demand Ride Pooling

Add code
Jan 07, 2024
Figure 1 for Mutual Information as Intrinsic Reward of Reinforcement Learning Agents for On-demand Ride Pooling
Figure 2 for Mutual Information as Intrinsic Reward of Reinforcement Learning Agents for On-demand Ride Pooling
Figure 3 for Mutual Information as Intrinsic Reward of Reinforcement Learning Agents for On-demand Ride Pooling
Figure 4 for Mutual Information as Intrinsic Reward of Reinforcement Learning Agents for On-demand Ride Pooling
Viaarxiv icon