feat(harbor-manager): unified benchmark normalization and reporting
- Implemented a unified benchmark normalization layer to handle metrics and artifacts across harbor, edit, and snapcompact benchmarks. - Integrated the TypeScript edit benchmark directly into the manager, migrating logic and deprecating the standalone package. - Updated the API and database schema to support standardized benchmark configurations, metrics, and trace-based reporting. - Enhanced the UI to visualize comparative performance metrics, including pass rate, cost, and latency deltas for benchmark runs.
This commit is contained in:
@@ -149,7 +149,6 @@
|
||||
"ci:release:publish": "bun scripts/ci-release-publish.ts",
|
||||
"ci:release:publish-native-leaf": "bun scripts/ci-release-publish.ts --native-leaf",
|
||||
"bench:gen-fixtures": "bun --cwd=packages/typescript-edit-benchmark run src/generate.ts --typescript-dir /tmp/typescript-source --count-per-type 8",
|
||||
"bench:edit": "bun --cwd=packages/typescript-edit-benchmark run start",
|
||||
"stats:sync": "python3 scripts/session-stats/sync.py",
|
||||
"stats:tools": "python3 scripts/session-stats/analyze.py tools",
|
||||
"stats:edits": "python3 scripts/session-stats/analyze.py edits",
|
||||
|
||||
Reference in New Issue
Block a user