Demonstration venue · read-only. Every page can be browsed; the buttons that would change it are switched off. Create an account to run TaxoReview on your own data.

Hana Hoshino

dblp:301/9464 · DBLP profile ↗
← Back
2ranked-venue papers
1as first author
2since 2021 · last 2023
—ORCID · none

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 2 · 1 first-author · 2 since 2021Systems, architecture and hardware · 1 · 1 first-author · 1 since 2021Graphics, computer vision, multimedia, augmented reality and games · 1 · 1 since 2021

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Artificial intelligence
2 papers
Autonomous driving · 64% Reinforcement learning · 36%

Topics — the 4 heaviest of 4, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Robotics › Autonomous driving
interaction modeling
0.712023
Joint Metrics Matter: A Better Standard for Trajectory Forecasting · ICCV 2023
Robotics › Autonomous driving
trajectory prediction
0.712023
Joint Metrics Matter: A Better Standard for Trajectory Forecasting · ICCV 2023
Machine learning › Reinforcement learning › imitation learning
inverse reinforcement learning
0.612022
OPIRL: Sample Efficient Off-Policy Inverse Reinforcement Learning via Distribution Matching · ICRA 2022
Machine learning › Reinforcement learning
reward learning
0.212022
OPIRL: Sample Efficient Off-Policy Inverse Reinforcement Learning via Distribution Matching · ICRA 2022

Methods — techniques the papers use, named apart from their topics

joint loss function · 0.7off-policy learning · 0.6distribution matching · 0.6
YearPublicationVenuePosition
2023 Joint Metrics Matter: A Better Standard for Trajectory Forecasting
abstract
Multi-modal trajectory forecasting methods commonly evaluate using single-agent metrics (marginal metrics), such as minimum Average Displacement Error (ADE) and Final Displacement Error (FDE), which fail to capture joint performance of multiple interacting agents. Only focusing on marginal metrics can lead to unnatural predictions, such as colliding trajectories or diverging trajectories for people who are clearly walking together as a group. Consequently, methods optimized for marginal metrics lead to overly-optimistic estimations of performance, which is detrimental to progress in trajectory forecasting research. In response to the limitations of marginal metrics, we present the first comprehensive evaluation of state-of-the-art (SOTA) trajectory forecasting methods with respect to multi-agent metrics (joint metrics): JADE, JFDE, and collision rate. We demonstrate the importance of joint metrics as opposed to marginal metrics with quantitative evidence and qualitative examples drawn from the ETH / UCY and Stanford Drone datasets. We introduce a new loss function incorporating joint metrics that, when applied to a SOTA trajectory forecasting method, achieves a 7% improvement in JADE / JFDE on the ETH / UCY datasets with respect to the previous SOTA. Our results also indicate that optimizing for joint metrics naturally leads to an improvement in interaction modeling, as evidenced by a 16% decrease in mean collision rate on the ETH / UCY datasets with respect to the previous SOTA. Code is available at github.com/ericaweng/joint-metrics-matter.
Erica Weng, Hana Hoshino, Deva Ramanan, Kris Makoto Kitani
ICCV2
2022 OPIRL: Sample Efficient Off-Policy Inverse Reinforcement Learning via Distribution Matching
abstract
Inverse Reinforcement Learning (IRL) is attractive in scenarios where reward engineering can be tedious. However, prior IRL algorithms use on-policy transitions, which require intensive sampling from the current policy for stable and optimal performance. This limits IRL applications in the real world, where environment interactions can become highly expensive. To tackle this problem, we present Off-Policy Inverse Reinforcement Learning (OPIRL), which (1) adopts off-policy data distribution instead of on-policy and enables significant reduction of the number of interactions with the environment, (2) learns a reward function that is transferable with high generalization capabilities on changing dynamics, and (3) leverages mode-covering behavior for faster convergence. We demonstrate that our method is considerably more sample efficient and generalizes to novel environments through the experiments. Our method achieves better or comparable results on policy performance baselines with significantly fewer interactions. Furthermore, we empirically show that the recovered reward function generalizes to different tasks where prior arts are prone to fail.
Hana Hoshino, Kei Ota, Asako Kanezaki, Rio Yokota
ICRA1