chimera bench
Run the continuous-evolution benchmark on a demo task set. Requires a key.
Options
--limitintdefault:0Limit number of demo tasks (0 = all).
--model, -mstrOverride the model slug.
--fusebooleanUse the fusion engine as the solver.
--chainbooleanRun the stateful chained benchmark (error propagation).
--hardbooleanUse the hard suite (traps / propagating chain).
--roundsintdefault:1Re-run the suite N times; report stagnation + cost trend across rounds.