NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-05-01M46B3BJFZACCFVFRMY22HMW9
Run Details
Model:
GPT-6.1 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
23.1s
Cost:
$0.009742
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the storm extinguished every window in Blackthorn House, except the library, where someone had recently lit a candle. Inspector Mara Vale arrived eleven minutes later, carrying a soaked invitation and wearing the expression reserved for particularly unwelcome reunions. The butler admitted her without speaking, although his trembling hands suggested he had already exhausted every useful word that evening. Beyond him, six guests stood around Lord Blackthorn, whose body occupied the hearthrug with an almost theatrical disregard for dignity. His dinner jacket was immaculate, his throat was bruised, and a silver key protruded between his lips like a joke. Mara recognized the key immediately, but she kept her face still while the dead man's daughter began explaining the impossible. She had locked the library after dinner, she said, and nobody possessed another key, except her father now lying there. Under the candle, a sealed envelope bore Mara's name, written in the handwriting of the husband she had buried yesterday.
created_at
"2026-10-05T15:32:13.524Z"
10/5/2026, 3:32:13 PM
model
"openai/gpt-6.1-sol-20260929"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
21206
moderation_latency
(null)
generation_time
23101
tokens_prompt
46
tokens_completion
263
native_tokens_prompt
41
native_tokens_completion
966
native_tokens_completion_images
0
native_tokens_reasoning
766
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.009742
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791214333-W3pV5YFuyo1lERtkETwI"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791214333-O086gFvDio1B451ryLIk"
upstream_id
"resp_0fb0738c5f1554e5016ac3c2fd9fb887d1b3f3d8a9767e5667"
provider_responses
0
endpoint_id
"dfb9d5c3-63f9-4263-8b78-4fa97dcb3d5e"
id
"resp_0fb0738c5f1554e5016ac3c2fd9fb887d1b3f3d8a9767e5667"
is_byok
false
latency
320
model_permaslug
"openai/gpt-6.1-sol-20260929"
provider_name
"OpenAI"
status
200
total_cost
0.009742
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20