NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HKBMN53DAR77YWP086EE4
Run Details
Model:
GPT-6 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
31.4s
Cost:
$0.022892
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the lighthouse keeper found a woman's body seated at his kitchen table, still wearing her white wedding veil. The strange part was that her husband had drowned three winters earlier, according to every surviving official record in town. She held a damp envelope addressed to Detective Mara Vale, whose name nobody in the isolated village should have known. But Mara had arrived that afternoon, summoned by a telegram warning her that someone would die before the tide turned. The keeper swore he had locked the kitchen before sunset, when the chair had stood empty beside an untouched supper. Outside, waves struck the rocks with enough force to shake the windows, yet the woman's veil somehow remained perfectly dry. On the table lay three silver coins, each stamped with tomorrow's date and marked by the same unmistakable bloody fingerprint. Mara lifted the envelope, and somewhere overhead, a door slammed in the lantern room that had been sealed for decades. From behind the sealed door came a woman's voice, whispering that the murderer was standing directly behind Mara herself now.
created_at
"2026-09-23T16:29:21.949Z"
9/23/2026, 4:29:21 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
5242
moderation_latency
(null)
generation_time
31362
tokens_prompt
46
tokens_completion
711
native_tokens_prompt
41
native_tokens_completion
2281
native_tokens_completion_images
0
native_tokens_reasoning
2070
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.022892
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790180961-f3NNwRTMs3WNcAzIWwYQ"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790180961-SvARFp7TJyZy6b2VOQtA"
upstream_id
"resp_0e001d8c3e601885016ab3fe620fe087d1b3eb6b1f367ab0dd"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_0e001d8c3e601885016ab3fe620fe087d1b3eb6b1f367ab0dd"
is_byok
false
latency
523
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.022892
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
9
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20
8
20