SkillOpt v0.1.0: initial release
- Skill optimization framework with training loop analogy - 11 benchmarks, 4 model backends (Azure OpenAI, Claude, Codex, Qwen) - WebUI for browser-based training control - Pluggable architecture for extending benchmarks and backends
This commit is contained in:
@@ -0,0 +1,17 @@
|
||||
"""ReflACT Gradient -- trajectory analysis and patch generation.
|
||||
|
||||
Analogous to gradient computation in neural network training: analyzes
|
||||
minibatch rollout trajectories to produce skill-edit patches (the "gradient"
|
||||
that drives skill updates).
|
||||
|
||||
Modules
|
||||
-------
|
||||
- reflect: minibatch trajectory analysis (gradient computation)
|
||||
- aggregate: hierarchical patch merging (gradient aggregation)
|
||||
- deep_probe: diagnostic probe generation (gradient probing)
|
||||
"""
|
||||
from skillopt.gradient.reflect import ( # noqa: F401
|
||||
run_minibatch_reflect,
|
||||
)
|
||||
from skillopt.gradient.aggregate import merge_patches # noqa: F401
|
||||
from skillopt.gradient.deep_probe import generate_deep_probe_instruction # noqa: F401
|
||||
Reference in New Issue
Block a user