fix(benchmarks): count unfenced code, ASCII-safe output, refresh llama3.2 results (#67)
Fixes the local benchmark LOC counter (counted only fenced code, scored bare output 0), makes summary output ASCII-safe (a Unicode arrow crashed the script on Windows cp1252), gitignores generated artifacts, and refreshes the llama3.2 writeup with n=5 data showing the LOC effect is within the noise floor. Follow-up to #63. Verified live. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
This commit is contained in:
co-authored by
Claude Opus 4.8
parent
386f95734a
commit
2e6a93765a
@@ -9,6 +9,10 @@ node_modules/
|
||||
# promptfoo eval artifacts
|
||||
.promptfoo/
|
||||
benchmarks/output*
|
||||
benchmarks/benchmark-local-results.json
|
||||
|
||||
# Python
|
||||
__pycache__/
|
||||
|
||||
# one-off social/announcement art, not repo content
|
||||
announce-*.png
|
||||
|
||||
Reference in New Issue
Block a user