genpark-gae-generalized-advantage-estimation-skill
mcp
Warn
Health Warn
- No license — Repository has no license file
- Description — Repository has a description
- Active repo — Last push 0 days ago
- Low visibility — Only 7 GitHub stars
Code Pass
- Code scan — Scanned 6 files during light audit, no dangerous patterns found
Permissions Pass
- Permissions — No dangerous permissions requested
No AI report is available for this listing yet.
Generalized Advantage Estimation (GAE-Lambda) exponentially weighted temporal difference calculator
README.md
genpark-gae-generalized-advantage-estimation-skill
Agent Skill implementing Generalized Advantage Estimation (GAE-$\lambda$) balancing bias and variance across temporal difference steps for Actor-Critic algorithms.
Architectural Overview
flowchart TD
Rewards["Rewards r_t"] & Values["State Values V(s_t)"] --> Delta["TD Residuals: delta_t = r_t + gamma * V(s_{t+1}) - V(s_t)"]
Delta --> Recurse["Backward Accumulator: A_t = delta_t + (gamma * lambda) * A_{t+1}"]
Recurse --> Advantages["Generalized Advantages A^{GAE}"]
Advantages & Values --> Returns["Empirical Target Value Returns: R_t = A_t + V(s_t)"]
Reviews (0)
Sign in to leave a review.
Leave a reviewNo results found