NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37J1E6Y19FE6212EV9273H4
Run Details
Model:
GPT-6 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
25.5s
Cost:
$0.019202
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the village clock struck thirteen, and Detective Mara Voss found blood seeping beneath the locked oak library door. The librarian had vanished, leaving one wet footprint, a shattered teacup, and a note addressed to whoever arrived there first. Mara unfolded the paper with gloved fingers, expecting a confession, but found only tomorrow's date written in thin red ink. Behind her, Constable Bell whispered that the library had been empty when he checked its windows just ten minutes earlier. Then someone knocked from inside the locked room, three slow taps followed by the unmistakable scrape of dragging heavy furniture. Bell reached for his keys, but Mara stopped him when a voice behind the door called her forgotten childhood name. No one in the village knew that name, not even Bell, who had served beside her for seven long years. Mara turned the key, and the door swung inward before she touched it, releasing the sharp smell of fresh rain. Inside, the librarian lay beneath the desk, while Mara's own reflection smiled quietly from the tall, darkened window behind him.
created_at
"2026-09-23T16:37:03.336Z"
9/23/2026, 4:37:03 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
23678
moderation_latency
(null)
generation_time
25407
tokens_prompt
46
tokens_completion
273
native_tokens_prompt
41
native_tokens_completion
1912
native_tokens_completion_images
0
native_tokens_reasoning
1690
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.019202
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790181423-P817rQxZAZAg777ToK5x"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790181423-ybTJscYSPNsKgMtVEn67"
upstream_id
"resp_0dca31e91ebc0651016ab4002f77fc87d191f2b4ab298daaab"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_0dca31e91ebc0651016ab4002f77fc87d191f2b4ab298daaab"
is_byok
false
latency
1692
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.019202
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
9
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20
8
20