NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-29-01M3P2FY624KABAW9Y737VQQ5D
Run Details
Model:
Claude Sonnet 5.5 (Reasoning)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
16.4s
Cost:
$0.022404
Provider:
Claude Platform on AWS
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Rain hammered the windows of Blackthorn Manor the night Lord Edmund Harrow was found dead in his locked private study. Inspector Mara Quill arrived at midnight, shaking water from her coat and studying the crowd gathered in the marble hall. Five guests stood beneath the chandelier, each wearing the same pale expression of shock that Mara had learned to distrust. The butler, Mr. Pemberton, explained that the study door had been bolted from inside, with the only key still missing. Mara climbed the staircase, counting every step, because she believed murderers often left their first mistakes closest to the crime. Inside the study, Lord Harrow slumped across his mahogany desk, one hand still clutching a torn scrap of blue silk. The window was latched, the fire had burned to embers, and a single wine glass sat untouched beside the body. Mara knelt, sniffed the glass, and caught the bitter almond scent that told her this death was never an accident. Behind her, the grandfather clock chimed once, and every guest downstairs fell suddenly silent as though waiting for a confession. Mara stood, brushed off her knees, and whispered that somebody in this house had just become a very careful liar.
created_at
"2026-09-29T07:53:57.702Z"
9/29/2026, 7:53:57 AM
model
"anthropic/claude-sonnet-5.5-20260928"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3841
moderation_latency
(null)
generation_time
16385
tokens_prompt
46
tokens_completion
675
native_tokens_prompt
62
native_tokens_completion
2228
native_tokens_completion_images
(null)
native_tokens_reasoning
1833
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"end_turn"
service_tier
"default"
usage
0.022404
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1790668437-YiilszeDqudZvhoKZbcm"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790668437-s9juspDJXywCxZWx4I28"
upstream_id
"msg_011CfXNQaELMDk6HJA4uJNYg"
provider_responses
0
endpoint_id
"99aaad94-923b-4fc1-b763-271ed5486f7a"
id
"msg_011CfXNQaELMDk6HJA4uJNYg"
is_byok
false
latency
614
model_permaslug
"anthropic/claude-sonnet-5.5-20260928"
provider_name
"Claude Platform on AWS"
status
200
total_cost
0.022404
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Claude Platform on AWS"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
10
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20
8
20
9
20