NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37PRC57MK4AQ3M09W78N6KY
Run Details
Model:
Claude Opus 5.5 (Reasoning)
Test:
N-Length Sentences
Scenario:
Write sentences with 10 words each
Duration:
20.2s
Cost:
$0.040488
Provider:
Claude Platform on AWS
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 10 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
The storm arrived at Blackwood Manor just after midnight struck. Rain hammered the tall windows like fists demanding to enter. Inside, the guests gathered nervously around the library's dying fire. Lord Ashford had not appeared for dinner, which was unusual. His niece Clara kept glancing toward the locked study door. The butler, Graves, insisted his master simply wished for solitude. Nobody believed him, but nobody dared to say so aloud. Then a single gunshot echoed through the dark, silent house. Clara screamed, and the doctor rushed toward the study immediately. Graves fumbled with his keys, his trembling hands betraying him. When the heavy door finally swung open, everyone froze completely. Lord Ashford lay slumped across his desk, a pistol nearby. The window was bolted, and his key lay beside him. Yet the doctor noticed something strange about the fresh wound. There were no powder burns, and the blood was cold. Lord Ashford had been dead long before that gunshot rang. Someone in this house was lying, and Clara knew it.
created_at
"2026-09-23T17:59:29.208Z"
9/23/2026, 5:59:29 PM
model
"anthropic/claude-opus-5.5-20260921"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
4737
moderation_latency
(null)
generation_time
20142
tokens_prompt
46
tokens_completion
595
native_tokens_prompt
62
native_tokens_completion
2012
native_tokens_completion_images
(null)
native_tokens_reasoning
1654
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"end_turn"
service_tier
"default"
usage
0.040488
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790186369-o64uDzYomkvTNg8dgDcy"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790186369-VL8FxHhV4odzKvbFA3AH"
upstream_id
"msg_011CfLojEdGBTxKgZBaM5zr8"
provider_responses
0
endpoint_id
"3a2388bc-3740-4e64-a1f5-4e301726b6b9"
id
"msg_011CfLojEdGBTxKgZBaM5zr8"
is_byok
false
latency
1345
model_permaslug
"anthropic/claude-opus-5.5-20260921"
provider_name
"Claude Platform on AWS"
status
200
total_cost
0.040488
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Claude Platform on AWS"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
17
wordCounts
0
10
1
10
2
10
3
10
4
10
5
10
6
10
7
10
8
10
9
10
10
10
11
10
12
10
13
10
14
10
15
10
16
10