NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HS236CBDHZAYDNY9DE9CJ
Run Details
Model:
GPT-6 Sol
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
5.0s
Cost:
$0.001762
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
94.38%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, Detective Mara Vale arrived at Blackthorn House, where every clock had stopped and the front door stood open. Inside, the guests gathered around a dinner table set for eight, though nine chairs scraped against the polished floorboards. The ninth chair held a folded coat, a silver pocket watch, and a note bearing Mara's name in red ink. Beyond the dining room, someone began playing the piano, slowly enough for Mara to hear a man's footsteps between notes. She followed the music upstairs and found Lord Blackthorn beneath his portrait, dead with a black ribbon around his throat. His youngest daughter stood nearby, holding a candle, and asked Mara why she had taken so long to come tonight. Mara looked at the note again; she had received the invitation only that morning, three days after its writer died.
created_at
"2026-09-23T16:32:28.784Z"
9/23/2026, 4:32:28 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3143
moderation_latency
(null)
generation_time
4999
tokens_prompt
46
tokens_completion
211
native_tokens_prompt
41
native_tokens_completion
168
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.001762
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790181148-XzIKAmnRRi0tZHU12ZhV"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790181148-yBrLVo4uofokbI830kVU"
upstream_id
"resp_016b13e900c7cea4016ab3ff1ced5887d1ae1c6ef25c7de05b"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_016b13e900c7cea4016ab3ff1ced5887d1ae1c6ef25c7de05b"
is_byok
false
latency
1826
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.001762
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
94.38%
Matches word count
n/a
neededClean
false
sentences
7
wordCounts
0
20
1
19
2
20
3
20
4
20
5
20
6
20