葵 渡辺
jacobbaker2023
AI & ML interests
None yet
Recent Activity
upvoted a paper 3 days ago
CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization upvoted a paper 3 days ago
StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents upvoted a paper 4 days ago
Progress Reward Modeling for Robotic Learning: A Comprehensive SurveyOrganizations
None yet