NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYCRC7JTP80FGCX1YHDVW5S9
Run Details
Model:
Claude Opus 5 (Reasoning)
Test:
N-Length Sentences
Scenario:
Write sentences with 10 words each
Duration:
19.0s
Cost:
$0.038400
Provider:
Amazon Bedrock
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 10 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Cold rain hammered the greenhouse glass like fingers demanding entry. Inspector Vale counted the orchids before she counted the corpses. There was only one of each, which somehow felt insufficient. The dead man lay among the ferns, wearing evening clothes. His shoes were spotless, though the garden path was mud. Someone had carried him here, or he had never walked. Vale crouched, ignoring the water soaking through her wool coat. A single orchid petal rested inside the victim's open mouth. Death had come near midnight, judging by the stiffened fingers. Behind her, the young constable made a small unprofessional sound. Vale did not turn around; she had expected that much. Every house this large hides something rotting beneath the roses. The family waited inside, drinking sherry, rehearsing their alibis quietly. She would give them nine more minutes of that peace.
created_at
"2026-07-25T13:45:31.49Z"
7/25/2026, 1:45:31 PM
model
"anthropic/claude-opus-5-20260723"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
5844
moderation_latency
(null)
generation_time
18836
tokens_prompt
46
tokens_completion
335
native_tokens_prompt
60
native_tokens_completion
1524
native_tokens_completion_images
(null)
native_tokens_reasoning
116
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"end_turn"
service_tier
"default"
usage
0.0384
router
(null)
provider_responses
0
endpoint_id
"76cb4608-f48c-483d-8da8-9957fb44244e"
id
"msg_011CdNsq7LfpDGiMAHd7dwVx"
is_byok
false
latency
2862
model_permaslug
"anthropic/claude-opus-5-20260723"
provider_name
"Amazon Bedrock"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1784987131-1Ip2XkT8w4nobJClQCw7"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1784987131-LnyJB8Jf5gIuxb12i4IJ"
upstream_id
"msg_011CdNsq7LfpDGiMAHd7dwVx"
total_cost
0.0384
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Amazon Bedrock"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
14
wordCounts
0
10
1
10
2
10
3
10
4
10
5
10
6
10
7
10
8
10
9
10
10
10
11
10
12
10
13
10