Skip to content
#

experiment-comparison

Here is 1 public repository matching this topic...

TraceOS standardizes AI experiments into reproducible, searchable, and comparable assets. One command runs experiments, generates reports, and produces structured analysis: capability vectors, failure taxonomy, and recommendations. Every run is tracked, traceable, and comparable. Built on ABC-130K (amazon-far/abc). Apache 2.0.

  • Updated Jul 3, 2026
  • Python

Add this topic to your repo

To associate your repository with the experiment-comparison topic, visit your repo's landing page and select "manage topics."

Learn more