fix(reflect): support continuous reward scores in failure filtering
not r.get("hard") treats non-zero floats as success.
Add explicit float threshold check (< 1e-9).
Backward compatible with binary hard=0/1.
This commit is contained in:
+588
-588
File diff suppressed because it is too large
Load Diff
Reference in New Issue
Block a user