still folding · tools
Still folding.
Two initial tools are on PyPI, and the leaderboard and benchmarks are coming soon.
Evaluate generated regular expressions: semantic equivalence, correctness, and ReDoS safety.
Inspired by autoresearch, an agent-driven experiment loop: propose a change, run it time-boxed, keep it only if the metric improves.
Runs regexbench across models and publishes the
numbers: scores, methodology, and a re-run command.
Reach out:
info@plicara.ai