Skip to content
STIMSMITH

Proximal Policy Optimization

Technique
First seen 6/14/2026
Last seen 9/1/2026
Evidence 11 chunks

NEIGHBORHOOD

3 nodes · 4 edges
graph · Proximal Policy Optimization · depth=1

RELATIONSHIPS

5 connections
Basic Block Agent ← uses 100% 3e
The Basic Block Agent uses Proximal Policy Optimization to learn its policy.
GenHuzz ← uses 100% 2e
GenHuzz uses PPO to bridge the fuzzer-DUT interaction in the HGRL framework.
Hardware-Guided Reinforcement Learning ← uses 100% 2e
The HGRL framework uses PPO to optimize the fuzzing policy.
HiFuzz ← uses 1e
HiFuzz uses Proximal Policy Optimization for the Basic Block Agent's policy updates.
Proximal Policy Optimization Algorithms ← introduces 1e
The Proximal Policy Optimization Algorithms paper introduces the PPO technique used by HiFuzz.