NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HF5V8KS1MFS5CFK1D9M51
Run Details
Model:
GPT-6 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
22.2s
Cost:
$0.015842
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the bell above Blackthorn Manor rang thirteen times, though its rope had been cut three bitter winters earlier. Detective Mara Voss arrived beneath a moonless sky, carrying a letter that smelled faintly of smoke and expensive violet perfume. Inside, twelve dinner guests sat motionless around a table set for thirteen, their untouched soup filmed with cooling yellow grease. The thirteenth place belonged to Lord Blackthorn, whose body lay upstairs behind a heavy oak door locked from the inside. His widow insisted nobody had climbed the stairs, but mud darkened the hem of her otherwise immaculate white silk gown. Mara noticed the footprints stopped outside the study, where someone had scratched a single word into the polished oak floorboards. The word was tomorrow, and beneath it gleamed a fresh drop of blood that had not yet fully turned brown. From upstairs came a soft click, followed by the unmistakable sound of a key sliding slowly into another hidden lock.
created_at
"2026-09-23T16:27:04.943Z"
9/23/2026, 4:27:04 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
7068
moderation_latency
(null)
generation_time
22205
tokens_prompt
46
tokens_completion
638
native_tokens_prompt
41
native_tokens_completion
1576
native_tokens_completion_images
0
native_tokens_reasoning
1380
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.015842
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790180824-ywyy0megrjTDASxIwl6Y"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790180824-fDHDDn935mKlZXeEG9nr"
upstream_id
"resp_01079afb6edb1bc4016ab3fdd90a2887d1bf978209bd79be94"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_01079afb6edb1bc4016ab3fdd90a2887d1bf978209bd79be94"
is_byok
false
latency
1217
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.015842
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20