NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYCQPRDPMYWD930J4Q48RXZB
Run Details
Model:
Claude Opus 5 (Reasoning)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
32.2s
Cost:
$0.076350
Provider:
Amazon Bedrock
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
The rain had been falling on Thornbury Hall for three days, and the gravel drive shone like wet polished coal. Inspector Wren arrived at seven o'clock, hat brim dripping, and found the front door standing wide open to the storm. Nobody had thought to close it, which told him more about the household than any written statement could possibly convey. In the library, Lord Ashcombe lay face down across his desk, one hand still gripping an expensive black fountain pen. The wound at the back of his skull was small and neat, entirely inconsistent with the fall the butler described. Five guests waited in the drawing room, arranged around a dying fire like reluctant witnesses at their own private trial. Not one of them asked what had happened, and Wren noticed that particular silence more than anything else that evening. Outside, the clocks struck eight, and somewhere above them a floorboard creaked beneath the careful footsteps of a sixth guest.
created_at
"2026-07-25T13:33:47.936Z"
7/25/2026, 1:33:47 PM
model
"anthropic/claude-opus-5-20260723"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
4078
moderation_latency
(null)
generation_time
31899
tokens_prompt
46
tokens_completion
664
native_tokens_prompt
60
native_tokens_completion
3042
native_tokens_completion_images
(null)
native_tokens_reasoning
426
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"end_turn"
service_tier
"default"
usage
0.07635
router
(null)
provider_responses
0
endpoint_id
"76cb4608-f48c-483d-8da8-9957fb44244e"
id
"msg_011CdNrwCjuaSuELqK2D28Qd"
is_byok
false
latency
1443
model_permaslug
"anthropic/claude-opus-5-20260723"
provider_name
"Amazon Bedrock"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1784986428-DgtksYjGSIFOELdci2dB"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1784986428-pZhtJLLvATQdYujphC7t"
upstream_id
"msg_011CdNrwCjuaSuELqK2D28Qd"
total_cost
0.07635
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Amazon Bedrock"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20