NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HQ93SHWGNSW0D8B92V909
Run Details
Model:
GPT-6 Sol
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
5.2s
Cost:
$0.001282
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
92.13%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, Inspector Vale found the conservatory locked, its glass walls fogged from within despite the cold November rain outside. Inside, beneath the dead orchids, a man lay face down, one hand curled around a fresh white chess knight. The victim was Judge Harrow, whose dinner guests had spent the evening insisting he had no enemies among them whatsoever. Vale counted six guests in the hallway, then noticed seven damp umbrellas dripping onto the marble floor behind their feet. Mrs Harrow said her husband never played chess, but someone upstairs had just moved a piece across a wooden board.
created_at
"2026-09-23T16:31:30.436Z"
9/23/2026, 4:31:30 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3595
moderation_latency
(null)
generation_time
5155
tokens_prompt
46
tokens_completion
151
native_tokens_prompt
41
native_tokens_completion
120
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.001282
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790181090-hY2O6xyZAPdJKhx4wPqP"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790181090-gCCXZupz1Sfvmdtlefak"
upstream_id
"resp_058b0542ad54eeb4016ab3fee2936087d180253a26338892ae"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_058b0542ad54eeb4016ab3fee2936087d180253a26338892ae"
is_byok
false
latency
1786
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.001282
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
92.13%
Matches word count
n/a
neededClean
false
sentences
5
wordCounts
0
20
1
19
2
20
3
20
4
20