NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-29-01M3P2FWJHAF8F86X3GJ8MC28V
Run Details
Model:
Claude Sonnet 5.5 (Reasoning)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
15.6s
Cost:
$0.019654
Provider:
Claude Platform on AWS
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Rain hammered the windows of Blackwood Manor the night Lord Edmund Harrow was found dead in his locked oak study. Inspector Mara Quill arrived at midnight, her umbrella broken, her patience thinner than the fog curling across the gravel drive. The butler, pale and trembling, led her through the great hall where portraits of dead Harrows watched her every step. Nobody had touched the study door, he insisted, until the family broke it down together just before eleven o'clock tonight. Inside, Lord Harrow slumped across his desk, one hand still clutching a silver letter opener buried deep beneath his ribs. Mara knelt beside the body and noticed that the window was latched from the inside, exactly like the heavy door. Somebody in this house had killed a man inside a sealed room, and she fully intended to learn precisely who. Behind her, the grandfather clock struck twelve, and somewhere upstairs a woman began to scream, high and long and terrible.
created_at
"2026-09-29T07:53:56.054Z"
9/29/2026, 7:53:56 AM
model
"anthropic/claude-sonnet-5.5-20260928"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3824
moderation_latency
(null)
generation_time
15596
tokens_prompt
46
tokens_completion
548
native_tokens_prompt
62
native_tokens_completion
1953
native_tokens_completion_images
(null)
native_tokens_reasoning
1626
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"end_turn"
service_tier
"default"
usage
0.019654
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1790668436-ZiyQO3sWOaayZGIBFnOA"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790668436-AM0wQLrM5bcheSJg6tiF"
upstream_id
"msg_011CfXNQTE9mLrNs9JLRgcS1"
provider_responses
0
endpoint_id
"99aaad94-923b-4fc1-b763-271ed5486f7a"
id
"msg_011CfXNQTE9mLrNs9JLRgcS1"
is_byok
false
latency
650
model_permaslug
"anthropic/claude-sonnet-5.5-20260928"
provider_name
"Claude Platform on AWS"
status
200
total_cost
0.019654
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Claude Platform on AWS"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20